The Hamilton-Jacobi equation

Jordan Bell
April 16, 2014

1 Example of free particle in one dimension

Define L⁢(q,v)=m2⁢v2, where m is a nonzero constant. Fixing two times t0<t1, we define the action for a path γ in ℝ by

S⁢(γ) = ∫t0t1L⁢(γ⁢(t),γ˙⁢(t))⁢𝑑t
= m2⁢∫t0t1(γ˙⁢(t))2⁢𝑑t.

Suppose that γ is satisfies the Euler-Lagrange equation for the Lagrangian L. That is,

dd⁢t⁢(∂⁡L∂⁡v⁢(γ⁢(t),γ˙⁢(t)))-∂⁡L∂⁡q⁢(γ⁢(t),γ˙⁢(t))=0,

and here ∂⁡L∂⁡v=m⁢v and ∂⁡L∂⁡q=0, so

dd⁢t⁢(m⁢γ˙⁢(t))=0,

i.e.

m⁢γ¨⁢(t)=0,

so γ¨⁢(t)=0, and hence γ⁢(t)=a⁢t+b for some constants a,b. If we are given the conditions γ⁢(t0)=q0 and γ⁢(t1)=q1, then a solution of the Euler-Lagrange equation that satisfies these conditions must be

γ⁢(t)=q0+q1-q0t1-t0⁢(t-t0).

The action of this path is

S⁢(γ)=m2⁢∫t0t1(q1-q0t1-t0)2⁢𝑑t=m2⁢(q1-q0)2t1-t0.

The dimensions of the right hand side are

kg⁢m2⁢s-1=kg⁢m⁢s-2⁢m⁢s=N⁢m⁢s=J⁢s,

which are indeed the dimensions that action ought to have. If instead of talking about action that is a function of paths we talk about action that is a function of the end point of a motion and the time at which the motion ends, taking the time and location at which the motion starts as fixed, then

S⁢(q,t)=m2⁢(q-q0)2t-t0,

for which

∂⁡S∂⁡q=m⁢q-q0t-t0

and

∂⁡S∂⁡t=-m2⁢(q-q0)2(t-t0)2.

These satisfy

∂⁡S∂⁡t+12⁢m⁢(∂⁡S∂⁡q)2=0.

If we write

H⁢(q,p)=pm⁢∂⁡L∂⁡v⁢(q,pm)-L⁢(q,pm)=pm⋅m⋅pm-m2⁢(pm)2=p22⁢m.

Then,

∂⁡S∂⁡t⁢(q,t)+H⁢(q,∂⁡S∂⁡q⁢(q,t))=0.

2 Motivation for the Hamilton-Jacobi equation

Suppose that γ is a path that satisfies the Euler-Lagrange equation for some Lagrangian L. If we perturb the path to start at the same position at time t0 but to end at q+δ⁢q instead of at q, then the perturbed path is s↦γ⁢(s)+(δ⁢γ)⁢(s), where (δ⁢γ)⁢(0)=0 and (δ⁢γ)⁢(t)=δ⁢q. Then, first doing a Taylor approximation in which we drop all powers of δ⁢γ or δ⁢γ˙ higher than the first and then using the Euler-Lagrange equation, and using Einstein summation notation,

S⁢(γ+δ⁢γ)-S⁢(γ) = ∫t0tL⁢(γ+δ⁢γ,γ˙+δ⁢γ˙)-L⁢(γ,γ˙)⁢d⁢s
= ∫t0t∂⁡L∂⁡qi⁢(γ,γ˙)⁢δ⁢γi+∂⁡L∂⁡vi⁢(γ,γ˙)⁢δ⁢γi˙⁢d⁢s
= ∫t0t(dd⁢t⁢∂⁡L∂⁡vi⁢(γ,γ˙))⁢δ⁢γi+∂⁡L∂⁡vi⁢(γ,γ˙)⁢δ⁢γi˙⁢d⁢s
= ∫t0tdd⁢t⁢(∂⁡L∂⁡vi⁢(γ,γ˙)⁢δ⁢γi)⁢𝑑s
= ∂⁡L∂⁡vi⁢(γ⁢(t),γ˙⁢(t))⁢(δ⁢γi)⁢(t)-∂⁡L∂⁡vi⁢(γ⁢(t0),γ˙⁢(t0))⁢(δ⁢γi)⁢(t0)
= ∂⁡L∂⁡vi⁢(γ⁢(t),γ˙⁢(t))⁢δ⁢q.

We have not been precise about what we mean by perturbing a path, but what we have obtained suggests that if we think of S as a function of the endpoint of a path and the time at which the path ends rather than as a function of a path itself, we have

∂⁡S∂⁡qi=∂⁡L∂⁡vi⁢(γ⁢(t),γ˙⁢(t))=pi⁢(t).

If on the other hand we fix the point q at which a path ends and change the time at which it arrives at this point from t to t+δ⁢t, then, doing a Taylor expansion and dropping all powers of δ⁢t higher than the first,

γ⁢(t+δ⁢t)-γ⁢(t)=γ˙⁢(t)⁢δ⁢t.

But γ⁢(t+δ⁢t)=q, so

γ⁢(t)=q-γ˙⁢(t)⁢δ⁢t.

Then

δ⁢S=L⁢δ⁢t-∂⁡L∂⁡vi⁢vi⁢δ⁢t.

Defining H=∂⁡L∂⁡vi⁢vi-L, what we have done suggests that

∂⁡S∂⁡t=-H.

Then, using ∂⁡S∂⁡qi=pi, with which H⁢(q,p)=H⁢(q,∂⁡S∂⁡q), and using ∂⁡S∂⁡t=-H⁢(q,p), we have

∂⁡S∂⁡t+H⁢(q,∂⁡S∂⁡q)=0.

We call this equation the Hamilton-Jacobi equation.

To precisely sort out where the Hamilton-Jacobi equation comes from and what it means, the only place I can imagine that does an adequate job is Abraham and Marsden.11 1 Abraham and Marsden, Foundations of Mechanics, second ed. Certainly there are other sources that present this more precisely than I have presented it, but it is almost universal to thoughtlessly confound the variables on which H or S depends with paths; that is, to write things like ∂⁡H∂⁡q and also to think of q not as a point but rather as a path which for each time goes through a particular point, in which case one has no certain way of knowing whether d⁢qd⁢t=0, as is the case for the derivative of any fixed point, or to say that q is a path and that d⁢qd⁢t is a tangent vector at the point q⁢(t) on the path. If one plainly states that what one has said is only suggestive of how symbols work together then one does not need to apologize for the absence of precision, but there is a foul area between suggestive symbol manipulation and actual precision in which one tricks oneself into believing that one has given a precise presentation, and this is the path followed by some presentations of the Hamilton-Jacobi equation.

3 Harmonic oscillator

Let

H=p22⁢m+12⁢m⁢ω2⁢q2.

The Hamilton-Jacobi equation for this Hamiltonian is

∂⁡S∂⁡t+12⁢m⁢(∂⁡S∂⁡q)2+12⁢m⁢ω2⁢q2=0.

We set S⁢(q,t)=S0⁢(q)-E⁢t; for this to make sense presumes that there is in fact a constant E and a function S0 so that S⁢(q,t)-S0⁢(q) depends just on t. With this, ∂⁡S∂⁡q=∂⁡S0∂⁡q and ∂⁡S∂⁡t=-E, and so the Hamilton-Jacobi equation becomes

-E+12⁢m⁢(∂⁡S0∂⁡q)2+12⁢m⁢ω2⁢q2=0,

or

12⁢m⁢(∂⁡S0∂⁡q)2+12⁢m⁢ω2⁢q2=E.

Supposing that S0 is nonnegative we get

∂⁡S0∂⁡q⁢(q)=2⁢m⁢E-m2⁢ω2⁢q2,

a primitive of which is

S0 = ∫2⁢m⁢E-m2⁢ω2⁢q2⁢𝑑q
= Eω⁢(arcsin⁡m⁢ω⁢q2⁢m⁢E+m⁢ω⁢q2⁢m⁢E⁢1-(m⁢ω⁢q2⁢m⁢E)2)

for which

∂⁡S0∂⁡E=1ω⁢arcsin⁡m⁢ω⁢q2⁢m⁢E.

But ∂⁡S0∂⁡E=t (using the expression involving S and E⁢t), so ω⁢t=arcsin⁡m⁢ω⁢q2⁢m⁢E, hence

sin⁡(ω⁢t)=m⁢ω⁢q2⁢E,

and therefore

q=2⁢Em⁢ω2⁢sin⁡(ω⁢t).

As well,

p=∂⁡S∂⁡q=∂⁡S0∂⁡q=2⁢m⁢E-m2⁢ω2⁢q2,

and using the above expression for q this becomes

p=2⁢m⁢E-m2⁢ω2⁢2⁢Em⁢ω2⁢sin2⁡(ω⁢t)=2⁢m⁢E-2⁢m⁢E⁢sin2⁡(ω⁢t).

We have thus written q and p as functions of E and t. Since for a particular trajectory the energy is fixed, on a particular trajectory the position and momentum have thus been expressed as functions of t.

4 Schrödinger equation

Write

i⁢ℏ⁢∂⁡ψ∂⁡t=H⁢ψ,

called the Schrödinger equation. If H⁢(q,p)=p22⁢m+V⁢(q), and p=-i⁢ℏ⁢∇, then p2=-ℏ2⁢Δ and the Schrödinger equation is

i⁢ℏ⁢∂⁡ψ∂⁡t=-ℏ22⁢m⁢Δ⁢ψ+V⁢(q)⁢ψ.

Supposing that there is a solution ψ of the form ψ=ei⁢Sℏ, we get

i⁢ℏ⁢ei⁢Sℏ⁢iℏ⁢∂⁡S∂⁡t = -ℏ22⁢m⁢∂∂⁡q⁢(ei⁢Sℏ⁢iℏ⁢∂⁡S∂⁡q)+V⁢(q)⁢ei⁢Sℏ
= -ℏ22⁢m⁢(ei⁢Sℏ⁢(iℏ⁢∂⁡S∂⁡q)2+ei⁢Sℏ⁢iℏ⁢∂2⁡S∂⁡q2)+V⁢(q)⁢ei⁢Sℏ.

Diving both sides by ei⁢Sℏ gives

-∂⁡S∂⁡t=12⁢m⁢(∂⁡S∂⁡q)2-i⁢ℏ2⁢m⁢∂2⁡S∂⁡q2+V⁢(q).

Taking ℏ→0 yields the equation

-∂⁡S∂⁡t=12⁢m⁢(∂⁡S∂⁡q)2+V⁢(q)=H⁢(q,∂⁡S∂⁡q),

which is the Hamilton-Jacobi equation.

The above derivation of the Hamilton-Jacobi equation from the Schrödinger equation is suggestive symbol manipulation. Rather than stating that we assume ψ=ei⁢Sℏ, I could have written that we “make an Ansatz”; “ein Ansatz” means “an approach”, and is used to mean a guess which may work out. Of course, if there is some problem for which one knows there is a unique solution and we find an explicit solution starting from some unjustified assumption, then we don’t need to have justified the assumption because we can explicitly check that what we have found is a solution. But this is the only situation where there is precision to “making an Ansatz”. Otherwise, to talk about an Ansatz is a sophisticated sounding way of saying “we make an assumption”, and after making this assumption we have no guarantee that anything we end up with need make any sense. This does not mean that it is useless to make unjustified assumptions; but it is deceitful to smuggle Ansätze into the realm of proved things, and confuses those who later would rely on one’s work.