Chapter 4
The Form of the Evolution Equation
With the diffraction of electrons we saw that a wave is in some way associated with the electron. If there is a wave, there must be an equation that describes it: this is the question Schrödinger asked himself. That equation describes the time evolution of the system, that is, how the state changes in time. In this card we shall look for the general form of the evolution equation.
Probabilistic interpretation.
A wave associated with a particle can be imagined in more than one way. Here we assume that the wave gives the probability of finding the particle at a point in space. We then describe the wave with a complex number at each point of space. In other areas of physics complex numbers are a computational convenience for dealing with waves, and it is therefore natural to use them for our own purpose as well, which is to find the equation of time evolution of what will be a wave.
According to this probabilistic interpretation we postulate that a material particle, from the point of view of Quantum Mechanics, is a physical system on which position measurements can be made. When we perform such a measurement, the result is in general not certain, but random. So we have a probability distribution for the variable x. We postulate the principle of the complex function: the probabilities are obtained as the squared moduli of certain complex numbers, called probability amplitudes. So we have a certain complex function and the probability is given by .
The function is a distribution of probability amplitudes and characterises the state of the particle; this function can be seen as a vector of infinitely many complex numbers, and we will represent it with the ket-vector symbol .
When the state of the particle evolves, the function ψ changes with time, so we obtain a function that depends on space and time , which we can represent with a time-dependent ket .
We postulate the principle of linear superposition: two states can be combined, each with a complex coefficient, and the combination is still a state; the state evolves continuously in time and, under evolution, each term evolves as it would on its own, with the same coefficient, and the subsequent state is the sum of the evolved terms. It follows that the time evolution is described by a linear law . Here is the vector associated with the function representing the initial state; is the vector associated with the function , and is the time-evolution matrix.
At this point the problem arises of determining the time-evolution matrix . We will solve this problem later; for now we want to anticipate the following qualitative result: the final equation we will obtain will closely resemble the wave equation, and the solutions will have the appearance of wave packets. So the wave that accompanies a material particle, of which we spoke in the previous section, is nothing other than the vector of probability amplitudes associated with the position variable.
For clarity we recap the important points we have introduced:
A material particle is characterised by the position variable, that is, by a coordinate which we will briefly denote by the symbol x.
The state of a material particle at an instant t is represented by the distribution of probability amplitudes of the variable x, which we denote by the ket symbol .
The evolution of the state is described by the following equation:
where is a matrix of ∞×∞ components.
Time-evolution equation in differential form.
The time-evolution equation we have written allows a finite jump between the instants t0 and t. However, it is more convenient to consider an infinitesimal time jump . In this case we have the equation
subtracting from both sides and dividing by dt we obtain
To simplify the right-hand side we can write the matrix with a first-order approximation
where is the derivative
Substituting this first-order approximation we have
So in the end we obtain the equation
The problem of determining has become the problem of determining the derivative .
Before proceeding, we must dwell on some mathematical topics.
Algebra of operators.
Vectors and matrices of infinite dimension
In this card we have introduced vectors and matrices of infinite dimension. In finite-dimensional cases a vector is represented by an n-tuple of components ; in infinite-dimensional cases, instead, we can represent a vector by a function , where the variable x takes the role of the indices. Analogously, an infinite-dimensional matrix is represented by a function of two variables .
The product of a matrix and a vector, for example, can be written by means of an integral
In an analogous way one can write other types of products between matrices or between vectors.
Adjoint matrix and Hermitian, anti-Hermitian and unitary matrices.
Suppose we have a ket vector given by the product of a matrix A and another vector
and suppose we have to determine the conjugate bra . Consider for example the two-dimensional case
So the bra is obtained by conjugating and transposing the ket
In the end we can write
where the matrix is obtained by conjugating and transposing the matrix A.
By definition we will say that the matrix is the adjoint of the matrix A, and we will denote it with the symbol
For example, for a two-dimensional matrix
On the basis of this definition we can write that the conjugate bra of the ket is .
By definition, if a matrix is equal to its adjoint , then it is said to be Hermitian; if instead the matrix is equal to its adjoint with the sign changed , then it is said to be anti-Hermitian. A matrix such that , where I is the identity matrix, is said to be unitary.
The operation of taking the adjoint matrix, in the field of matrices, takes the role of the operation of taking the complex conjugate in the field of complex numbers. So Hermitian matrices take the role of purely real numbers, while anti-Hermitian matrices take the role of purely imaginary numbers. Unitary matrices, finally, take the role of numbers of unit modulus.
Hermitian, anti-Hermitian and unitary matrices have very interesting properties that we will study later, and they are of considerable importance in Quantum Mechanics.
Relation between matrices and operators.
In the case of finite-dimensional vectors we know, from the study of linear algebra, that any linear operator can be written as a matrix product , where R is a particular matrix associated with the operator in question.
This property remains valid also in the case of infinite-dimensional vectors, but it is much more delicate from the mathematical point of view.
Consider for example the derivative operator, which we will denote by D. The operator D transforms a vector associated with the function into the vector associated with the function
We ask: what is the matrix, that is, the function of two variables, that represents the derivative operator? The function we seek is the derivative of the Dirac δ
indeed, performing the matrix product and integrating by parts, we have
The boundary term vanishes because δ vanishes at infinity; in the remaining integral the δ selects the value of the derivative at the point x:
The operator is anti-Hermitian, indeed
that is, by conjugating and transposing the matrix associated with D one obtains a matrix of opposite sign .
In general one prefers to use the operator , which is Hermitian, indeed
From now on we will speak indifferently of operators or of matrices. A derivative operator will in general be represented by the derivative operation rather than by the associated matrix; in any case it is important to know that to any linear operator there is always associated a definite matrix, whether finite- or infinite-dimensional.
Functions of operators, or functions of matrices.
If we take an operator A and apply it twice we obtain the operator ; in the same way we can obtain , , etc.
If we take the inverse operator and apply it twice we define the operator ; in the same way we obtain , , etc.
If we consider a polynomial
from it we can build the operator
In general, from a function that can be expanded in a power series
we can build the operator
(one sets , where I is the identity operator)
Let us now return to the study of the time-evolution equation.
Conservation of the scalar product.
Finally, we postulate the principle of conservation of the total: as long as the system is not observed, the sum of the squared moduli over all positions does not change during the evolution. From this principle, together with linearity, there follows an important property of the time-evolution process: during this process the scalar products between different kets remain constant. Now we will see what this fact implies for the matrix A that appears in the equation
Consider two kets and ; let us compute the evolution of these kets at the instant
By the property of conservation of the scalar product, it must be
substituting the formulas found for and we have
dividing by dt and taking the limit as dt→0 we have
This equation is valid for every and , so it must be
So the matrix A is anti-Hermitian. We replace the matrix A with the matrix , where is the imaginary unit.
The matrix H is called the Hamiltonian and is Hermitian, indeed
The evolution equation written in terms of the matrix H appears thus
This equation is called the Schrödinger equation.