Skip to article frontmatterSkip to article content
Site not loading correctly?

This may be due to an incorrect BASE_URL configuration. See the MyST Documentation for reference.

15.1: Introduction to Advanced Hamiltonian Mechanics

This study of classical mechanics has involved climbing a vast mountain of knowledge, while the pathway to the top has led us to elegant and beautiful theories that underlie much of modern physics. Being so close to the summit provides the opportunity to take a few extra steps in order to provide a glimpse of applications to physics at the summit. These are described in chapters 151815 − 18.

Hamilton’s development of Hamiltonian mechanics in 1834 is the crowning achievement for applying variational principles to classical mechanics. A fundamental advantage of Hamiltonian mechanics is that it uses the conjugate coordinates q,p\mathbf{q}, \mathbf{p}, plus time tt, which is a considerable advantage in most branches of physics and engineering. Compared to Lagrangian mechanics, Hamiltonian mechanics has a significantly broader arsenal of powerful techniques that can be exploited to obtain an analytical solution of the integrals of the motion for complicated systems. In addition, Hamiltonian dynamics provides a means of determining the unknown variables for which the solution assumes a soluble form, and is ideal for study of the fundamental underlying physics in applications to fields such as quantum or statistical physics. As a consequence, Hamiltonian mechanics has become the preeminent variational approach used in modern physics. This chapter introduces the following four techniques in Hamiltonian mechanics:

  1. the elegant Poisson bracket representation of Hamiltonian mechanics, which played a pivotal role in the development of quantum theory;

  2. the powerful Hamilton-Jacobi theory coupled with Jacobi’s development of canonical transformation theory;

  3. action-angle variable theory; and

  4. canonical perturbation theory.

Prior to further development of the theory of Hamiltonian mechanics, it is useful to summarize the major formula relevant to Hamiltonian mechanics that have been presented in chapters 7, 8, and 9.

Action functional SS:

As discussed in chapter 9.2, Hamiltonian mechanics is built upon Hamilton’s action functional

S(q,p,t)=t1t2L(q,q˙,t)dtS( \mathbf{ q}, \mathbf{ p},t) = \int^{t_2}_{t_1} L( \mathbf{ q}, \mathbf{\dot{q}},t)dt

Hamilton’s Principle of least action states that

δS(q,p,t)=δt1t2L(q,q˙,t)dt=0\delta S( \mathbf{ q}, \mathbf{ p},t) = \delta \int^{t_2}_{t_1} L( \mathbf{ q}, \mathbf{\dot{q}},t)dt = 0

Generalized momentum pp:

In chapter 7.2, the generalized (canonical) momentum was defined in terms of the Lagrangian LL to be

piL(q,q˙,t)q˙ip_i \equiv \frac{\partial L(\mathbf{q}, \mathbf{\dot{q}},t)}{ \partial \dot{q}_i}

Chapter 9.2 defined the generalized momentum in terms of the action functional SS to be

pj=S(q,p,t)qjp_j = \frac{\partial S(\mathbf{q}, \mathbf{p},t)}{\partial q_j}

Generalized energy h(q,q˙,t)h(\mathbf{q}, \dot{q},t ):

Jacobi’s Generalized Energy h(q,q˙,t)h(\mathbf{q}, \dot{q},t ) was defined in equation (7.7.6)(7.7.6) as

h(q,q˙,t)j(q˙jL(q,q˙,t)q˙j)L(q,q˙,t)h(\mathbf{q}, \mathbf{\dot{q}},t ) \equiv \sum_j \left( \dot{q}_j \frac{\partial L(\mathbf{q}, \mathbf{\dot{q}}, t)}{ \partial \dot{q}_j} \right) − L(\mathbf{q}, \mathbf{\dot{q}}, t)

Hamiltonian function:

The Hamiltonian H(q,p,t)H(\mathbf{q},\mathbf{p},t) was defined in terms of the generalized energy h(q,q˙,t)h(\mathbf{q}, \mathbf{\dot{q}},t ) plus the generalized momentum. That is

H(q,p,t)h(q,q˙,t)=jpjq˙jL(q,q˙,t)=pq˙L(q,q˙,t)H(\mathbf{q},\mathbf{p},t) \equiv h(\mathbf{q}, \mathbf{\dot{q}},t ) = \sum_j p_j \dot{q}_j − L(\mathbf{q}, \mathbf{\dot{q}}, t) = \mathbf{p} \cdot \mathbf{\dot{q}}−L(\mathbf{q}, \mathbf{\dot{q}}, t)

where q,p\mathbf{q}, \mathbf{p} correspond to nn-dimensional vectors, e.g. q(q1,q2,...,qn)\mathbf{q} \equiv (q_1, q_2, ..., q_n) and the scalar product pq˙=ipiq˙i\mathbf{p}\cdot\mathbf{\dot{q}} = \sum_i p_i \dot{q}_i. Chapter 8.2 used a Legendre transformation to derive this relation between the Hamiltonian and Lagrangian functions. Note that whereas the Lagrangian L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) is expressed in terms of the coordinates q\mathbf{q}, plus conjugate velocities q˙\mathbf{\dot{q}}, the Hamiltonian H(q,p,t)H (\mathbf{q}, \mathbf{p}, t) is expressed in terms of the coordinates q\mathbf{ q} plus their conjugate momenta p\mathbf{ p}. For scleronomic systems, plus assuming the standard Lagrangian, then equations (7.9.4)(7.9.4) and (7.6.13)(7.6.13) give that the Hamiltonian simplifies to equal the total mechanical energy, that is, H=T+UH = T + U.

Generalized energy theorem:

The equations of motion lead to the generalized energy theorem which states that the time dependence of the Hamiltonian is related to the time dependence of the Lagrangian.

dH(q,p,t)dt=jq˙j[QjEXC+k=1mλkgkqj(q,t)]L(q,q˙,t)t\frac{dH (\mathbf{q},\mathbf{p},t)}{ dt} = \sum_j \dot{q}_j \left[ Q^{EXC}_j + \sum^m_{k=1} \lambda_k \frac{\partial g_k}{ \partial q_j} (\mathbf{q}, t) \right] − \frac{\partial L(\mathbf{q}, \mathbf{\dot{q}}, t)}{ \partial t}

Note that if all the generalized non-potential forces and Lagrange multiplier terms are zero, and if the Lagrangian is not an explicit function of time, then the Hamiltonian is a constant of motion.

Hamilton’s equations of motion:

Chapter 8.3 showed that a Legendre transform plus the Lagrange-Euler equations led to Hamilton’s equations of motion. Hamilton derived these equations of motion directly from the action functional, as shown in chapter 9.2.

q˙j=H(q,p,t)pj\dot{q}_j = \frac{\partial H(\mathbf{q},\mathbf{p},t)}{ \partial p_j}
p˙j=Hqj(q,p,t)+[k=1mλkgkqj+QjEXC]\dot{p}_j = −\frac{\partial H}{ \partial q_j} (\mathbf{q}, \mathbf{p},t) + \left[ \sum^{m}_{k=1} \lambda_k \frac{\partial g_k}{ \partial q_j} + Q^{EXC}_j \right]
H(q,p,t)t=L(q,q˙,t)t\frac{\partial H(\mathbf{q},\mathbf{p},t) }{\partial t} = −\frac{\partial L(\mathbf{q}, \mathbf{\dot{q}}, t)}{ \partial t}

Note the symmetry of Hamilton’s two canonical equations. The canonical variables pk,qkp_k,q_k are treated as independent canonical variables. Lagrange was the first to derive the canonical equations but he did not recognize them as a basic set of equations of motion. Hamilton derived the canonical equations of motion from his fundamental variational principle and made them the basis for a far-reaching theory of dynamics. Hamilton’s equations give 2s2s first-order differential equations for pk,qkp_k,q_k for each of the ss degrees of freedom. Lagrange’s equations give ss second-order differential equations for the variables qk,q˙kq_k,\dot{q}_k.

Hamilton-Jacobi equation:

Hamilton used Hamilton’s Principle to derive the Hamilton-Jacobi equation (9.2.17)(9.2.17).

St+H(q,p,t)=0\frac{\partial S }{\partial t} + H(\mathbf{q}, \mathbf{p},t)=0

The solution of Hamilton’s equations is trivial if the Hamiltonian is a constant of motion, or when a set of generalized coordinates can be identified for which all the coordinates qiq_i are constant, or are cyclic (also called ignorable coordinates). Jacobi developed the mathematical framework of canonical transformations required to exploit the Hamilton-Jacobi equation.

15.2: Poisson bracket Representation of Hamiltonian Mechanics

Poisson Brackets

Poisson brackets were developed by Poisson, who was a student of Lagrange. Hamilton’s canonical equations of motion describe the time evolution of the canonical variables (q,p)(q,p) in phase space. Jacobi showed that the framework of Hamiltonian mechanics can be restated in terms of the elegant and powerful Poisson bracket formalism. The Poisson bracket representation of Hamiltonian mechanics provides a direct link between classical mechanics and quantum mechanics.

The Poisson bracket of any two continuous functions of generalized coordinates F(p,q)F(p, q) and G(p,q)G(p, q), is defined to be

{F,G}qpi(FqiGpiFpiGqi)(15.12)\{F,G\}_{qp} \equiv \sum_i \left(\frac{\partial F}{\partial q_i} \frac{\partial G} {\partial p_i} − \frac{\partial F} {\partial p_i }\frac{\partial G} {\partial q_i} \right) \tag{15.12}

Note that the above definition of the Poisson bracket, written using the common brace notation, leads to the following identity, antisymmetry, linearity, Leibniz rules, and Jacobi Identity.

{F,F}=0{F,G}={G,F}{G,F+Y}={G,F}+{G,Y}{G,FY}={G,F}Y+F{G,Y}0={F,{G,Y}}+{G,{Y,F}}+{Y{F,G}}\begin{align} \{F, F\} &= 0 \\[4pt] \{F,G\} &= − \{G, F\} \\[4pt] \{G, F + Y \} &= \{G, F\}+\{G, Y \} \\[4pt] \{G, F Y \} &= \{G, F\} Y + F \{G, Y \} \\[4pt] 0 &= \{F, \{G, Y \}\} + \{G, \{Y,F\}\} + \{Y \{F,G\}\} \tag{15.17} \end{align}

where GG, HH, and YY are functions of the canonical variables plus time. Jacobi’s identity; 15.17 states that the sum of the cyclic permutation of the double Poisson brackets of three functions is zero. Jacobi’s identity plays a useful role in Hamiltonian mechanics as will be shown.

Fundamental Poisson Brackets

The Poisson brackets of the canonical variables themselves are called the fundamental Poisson brackets. They are

{qk,ql}qp=i(qkqiqlpiqkpiqlqi)=i(δki00δli)=0\{q_k, q_l\}_{qp} = \sum_i \left(\frac{\partial q_k}{ \partial q_i} \frac{\partial q_l}{\partial p_i} − \frac{\partial q_k}{ \partial p_i} \frac{\partial q_l}{ \partial q_i} \right) = \sum_i (\delta_{ki} \cdot 0 − 0 \cdot \delta_{li}) = 0
{pk,pl}qp=i(pkqiplpipkpiplqi)=i(0δliδki0)=0\{p_k, p_l\}_{qp} = \sum_i \left(\frac{\partial p_k }{\partial q_i} \frac{\partial p_l}{ \partial p_i} − \frac{\partial p_k }{\partial p_i} \frac{\partial p_l}{ \partial q_i } \right) = \sum_i (0 \cdot \delta_{li} − \delta_{ki} \cdot 0) = 0
{qk,pl}qp=i(qkqiplpiqkpiplqi)=i(δkiδli00)=δkl\{q_k, p_l\}_{qp} = \sum_i \left(\frac{\partial q_k}{ \partial q_i} \frac{\partial p_l}{ \partial p_i} − \frac{\partial q_k}{ \partial p_i} \frac{\partial p_l}{ \partial q_i} \right) = \sum_i (\delta_{ki} \cdot \delta_{li} − 0 \cdot 0) = \delta_{kl}

In summary, the fundamental Poisson brackets equal

{qk,ql}qp=0\{q_k, q_l\}_{qp} = 0
{pk,pl}qp=0\{p_k, p_l\}_{qp} = 0
{qk,pl}qp={pl,qk}qp=δkl\{q_k, p_l\}_{qp} = − \{p_l, q_k\}_{qp} = \delta_{kl}

Note that the Poisson bracket is antisymmetric under interchange in pp and qq. It is interesting that the only non-zero fundamental Poisson bracket is for conjugate variables where k=lk = l, that is

{qk,pk}pq=1\{q_k, p_k\}_{pq} = 1

Poisson bracket invariance to canonical transformations

The Poisson brackets are invariant under a canonical transformation from one set of canonical variables (qk,pk)(q_k, p_k) to a new set of canonical variables (Qk,Pk)(Q_k, P_k) where QkQk(q,p)Q_k \rightarrow Q_k(\mathbf{q}, \mathbf{p}) and PkPk(q,p)P_k \rightarrow P_k(\mathbf{q}, \mathbf{p}). This is shown by transforming Equation 15.12 to the new variables by the following derivation

{F,G}qp=j(FqjGpjFpjGqj)=jk(Fqj(GQkQkpj+GPkPkpj)Fpj(GQkQkqj+GPkPkqj))\begin{align} \{F,G\}_{qp} & = \sum_{j} \left( \frac{\partial F}{ \partial q_j} \frac{\partial G} {\partial p_j} − \frac{\partial F} {\partial p_j} \frac{\partial G} {\partial q_j} \right) \tag{15.25} \\[4pt] & = \sum_{jk} \left( \frac{\partial F}{ \partial q_j }\left( \frac{\partial G}{ \partial Q_k }\frac{\partial Q_k }{\partial p_j} + \frac{\partial G} {\partial P_k }\frac{\partial P_k}{ \partial p_j} \right) − \frac{\partial F}{ \partial p_j} \left( \frac{\partial G} {\partial Q_k} \frac{\partial Q_k}{ \partial q_j} + \frac{\partial G}{ \partial P_k }\frac{\partial P_k} {\partial q_j} \right)\right) \tag{15.26}\end{align}

The terms can be rearranged to give

{F,G}qp=k(GQk{F,Qk}qp+GPk{F,Pk}qp)(15.27)\{F,G\}_{qp} = \sum_k \left( \frac{\partial G}{ \partial Q_k} \{F, Q_k\}_{qp} + \frac{\partial G}{ \partial P_k} \{F, P_k\}_{qp}\right) \tag{15.27}

Let F=QkF = Q_k and replace GG by FF, and use the fact that the fundamental Poisson brackets {Qk,Qj}qp=0\{Q_k, Q_j \}_{qp} = 0 and {Qk,Pj}qp=δjk\{Q_k, P_j \}_{qp} = \delta_{jk}, then Equation 15.25 reduces to

{Qk,F}qp=j(FQj{Qk,Qj}+FPj{Qk,Pj})=jFPjδjk\{Q_k, F\}_{qp} = \sum_j \left( \frac{\partial F}{ \partial Q_j} \{Q_k, Q_j \} + \frac{\partial F} {\partial P_j } \{Q_k, P_j \} \right) = \sum_j \frac{\partial F}{ \partial P_j} \delta_{jk}

That is

{F,Qk}=FPk(15.29)\{F, Q_k\} = − \frac{\partial F} {\partial P_k} \tag{15.29}

Similarly

{Pk,F}qp=j(FQj{Pk,Qj}qp+FPj{Pk,Pj}qp)\{P_k, F\}_{qp} = \sum_j \left( \frac{\partial F}{ \partial Q_j} \{P_k, Q_j \}_{qp} + \frac{\partial F}{ \partial P_j} \{P_k, P_j \}_{qp} \right)

leading to

{F,Pk}qp=FQk(15.31)\{F, P_k\}_{qp} = \frac{\partial F} {\partial Q_k} \tag{15.31}

Substituting equations 15.29 and 15.31 into Equation 15.27 gives

{F,G}qp=k(FQkGPkFPkGQk)={F,G}QP\{F,G\}_{qp} = \sum_k \left( \frac{\partial F} {\partial Q_k} \frac{\partial G}{ \partial P_k} − \frac{\partial F} {\partial P_k} \frac{\partial G} {\partial Q_k} \right) = \{F,G\}_{QP}

Thus the canonical variable subscripts (q,p)(q,p) and (Q,P)(Q,P) can be ignored since the Poisson bracket is invariant to any canonical transformation of canonical variables. The counter argument is that if the Poisson bracket is independent of the transformation, then the transformation is canonical.

Correspondence of the Commutator and the Poisson Bracket

In classical mechanics there is a formal correspondence between the Poisson bracket and the commutator. This can be shown by deriving the Poisson Bracket of four functions taken in two pairs. The derivation requires deriving the two possible Poisson Brackets involving three functions.

{F1F2,G}=j[(F1qjF2+F1F2qj)Gpj(F1pjF2+F1F2pj)Gqj]={F1,G}F2+F1{F2,G}\begin{align} \{F_1F_2, G\} & = \sum_j \left[ \left(\frac{\partial F_1}{ \partial q_j } F_2 + F_1 \frac{\partial F_2}{ \partial q_j} \right) \frac{\partial G} {\partial p_j} − \left(\frac{\partial F_1}{ \partial p_j} F_2 + F_1 \frac{\partial F_2}{ \partial p_j} \right) \frac{\partial G}{ \partial q_j} \right] \\[4pt] &= \{F_1, G\} F_2 + F_1 \{F_2, G\} \tag{15.33} \end{align}
{F,G1G2}={F,G1}G2+G1{F,G2}(15.34)\{F,G_1G_2\} = \{F,G_1\} G_2 + G_1 \{F,G_2\} \tag{15.34}

These two Poisson Brackets for three functions can be used to derive the Poisson Bracket of four functions, taken in pairs. This can be accomplished two ways using either Equation 15.33 or 15.34.

{F1F2,G1G2}={F1,G1G2}F2+F1{F2,G1G2}=[{F1,G1}G2+G1{F1,G2}]F2+F1[{F2,G1}G2+G1{F2,G2}]={F1,G1}G2F2+G1{F1,G2}F2+F1{F2,G1}G2+F1G1{F2,G2}(15.35)\{F_1F_2, G_1G_2\} = \{F_1, G_1G_2\} F_2 + F_1 \{F_2, G_1G_2\} \\ = [ \{F_1, G_1\} G_2 + G_1 \{F_1, G_2\} ] F_2 + F_1 [\{F_2, G_1\} G_2 + G_1 \{F_2, G_2\}] \\ = \{F_1, G_1\} G_2F_2 + G_1 \{F_1, G_2\} F_2 + F_1 \{F_2, G_1\} G_2 + F_1G_1 \{F_2, G_2\} \tag{15.35}

The alternative approach gives

{F1F2,G1G2}={F1F2,G1}G2+G1{F1F2,G2}={F1,G1}F2G2+F1{F2,G1}G2+G1{F1,G2}F2+G1F1{F2,G2}(15.36)\{F_1F_2, G_1G_2\} = \{F_1F_2,G_1\} G_2 + G_1 \{F_1F_2, G_2\} \\ = \{F_1, G_1\} F_2G_2 + F_1 \{F_2, G_1\} G_2 + G_1 \{F_1, G_2\} F_2 + G_1F_1 \{F_2, G_2\} \tag{15.36}

These two alternate derivations give different relations for the same Poisson Bracket. Equating the alternative equations 15.35 and 15.36 gives that

{F1,G1}(F2G2G2F2)=(F1G1G1F1){F2,G2}\{F_1, G_1\} (F_2G_2 − G_2F_2) = (F_1G_1 − G_1F_1) \{F_2, G_2\} \nonumber

This can be factored into separate relations, the left-hand side for body 1, and the right-hand side for body 2.

(F1G1G1F1){F1,G1}=(F2G2G2F2){F2,G2}=λ\frac{(F_1G_1 − G_1F_1)}{ \{F_1, G_1\}} = \frac{(F_2G_2 − G_2F_2)}{ \{F_2, G_2\} } = \lambda

Since the left-hand ratio holds for F1,G1F_1, G_1 independent of F2,G2F_2, G_2, and vise versa, then they must equal a constant λ\lambda that does not depend on F1,G1F_1, G_1, does not depend on F2,G2F_2, G_2, and λ\lambda must commute with (F1G1G1F1)(F_1G_1 − G_1F_1). That is, λ\lambda must be a constant number independent of these variables.

(F1G1G1F1)=λ{F1,G1}λi(F1qiG1piF1piG1qi)(15.38)(F_1G_1 − G_1F_1) = \lambda \{F_1, G_1\} \equiv \lambda \sum_i \left(\frac{\partial F_1}{ \partial q_i} \frac{\partial G_1 }{\partial p_i} − \frac{\partial F_1 }{\partial p_i }\frac{\partial G_1}{ \partial q_i} \right) \tag{15.38}

Equation 15.38 is an especially important result which states that to within a multiplicative constant numberλ\lambda, there is a one-to-one correspondence between the Poisson Bracket and the commutator of two independent functions. An important implication is that if two functions, FiGkF_iG_k have a Poisson Bracket that is zero, then the commutator of the two functions also must be zero, that is, FiF_i and GkG_k commute.

Consider the special case where the variables F1F_1 and G1G_1 correspond to the fundamental canonical variables, (qk,pl)(q_k, p_l). Then the commutators of the fundamental canonical variables are given by

qkplplqk=λ{qk,pl}=λδklq_kp_l − p_lq_k = \lambda \{q_k, p_l\} = \lambda\delta_{kl}
qkqlqlqk=λ{qk,ql}=0q_kq_l − q_lq_k = \lambda \{q_k, q_l\} = 0
pkplplpk=λ{pk,pl}=0p_kp_l − p_lp_k = \lambda \{p_k, p_l\} = 0

In 1925, Paul Dirac, a 23-year old graduate student at Bristol, recognized that the formal correspondence between the Poisson bracket in classical mechanics, and the corresponding commutator, provides a logical and consistent way to bridge the chasm between the Hamiltonian formulation of classical mechanics, and quantum mechanics. He realized that making the assumption that the constant λi\lambda \equiv i\hbar, leads to Heisenberg’s fundamental commutation relations in quantum mechanics, as is discussed in chapter 18.3.1. Assuming that λi\lambda \equiv i\hbar provides a logical and consistent way that builds quantization directly into classical mechanics, rather than using ad-hoc, case-dependent, hypotheses as was used by the older quantum theory of Bohr.

Observables in Hamiltonian mechanics

Poisson brackets, and the corresponding commutation relations, are especially useful for elucidating which observables are constants of motion, and whether any two observables can be measured simultaneously and exactly. The properties of any observable are determined by the following two criteria.

Time dependence:

The total time differential of a function G(qi,pi,t)G (q_i, p_i, t) is defined by

dGdt=Gt+i(Gqiq˙i+Gpip˙i)\frac{dG}{ dt} = \frac{\partial G}{ \partial t} +\sum_i \left(\frac{\partial G} {\partial q_i} \dot{q}_i + \frac{\partial G}{ \partial p_i} \dot{p}_i \right)

Hamilton’s canonical equations give that

q˙i=Hpi\dot{q}_i = \frac{\partial H}{ \partial p_i}
p˙i=Hqi\dot{p}_i = −\frac{\partial H}{ \partial q_i }

Substituting these in the above relation gives

dGdt=Gt+i(GqiHpiGpiHqi)\frac{dG}{ dt} = \frac{\partial G} {\partial t} +\sum_i \left(\frac{\partial G} {\partial q_i} \frac{\partial H}{ \partial p_i} − \frac{\partial G}{ \partial p_i} \frac{\partial H}{ \partial q_i} \right) \nonumber

that is

dGdt=Gt+{G,H}(15.45)\frac{dG }{dt} = \frac{\partial G}{ \partial t} + \{G, H\} \tag{15.45}

This important equation states that the total time derivative of any function G(q,p,t)G(q, p, t) can be expressed in terms of the partial time derivative plus the Poisson bracket of G(q,p,t)G(q, p, t) with the Hamiltonian.

Any observable G(p,q,t)G(p, q, t) will be a constant of motion if dGdt=0\frac{dG}{ dt} = 0, and thus Equation 15.45 gives

Gt+{G,H}=0(If G is a constant of motion)\frac{\partial G} {\partial t} + \{G, H\} = 0 \tag{If G is a constant of motion}

That is, it is a constant of motion when

Gt={H,G}\frac{\partial G}{ \partial t} = \{H, G\}

Moreover, this can be extended further to the statement that if the constant of motion GG is not explicitly time dependentthen

{G,H}=0\{G, H\} = 0

The Poisson bracket with the Hamiltonian is zero for a constant of motion GG that is not explicitly time dependent. Often it is more useful to turn this statement around with the statement that if {G,H}=0\{G, H\} = 0, and Gt=0\frac{\partial G} {\partial t} = 0, then dGdt=0\frac{dG}{dt} = 0, implying that GG is a constant of motion.

Independence

Consider two observables F(p,q,t)F(p, q, t) and G(p,q,t)G(p, q, t). The independence of these two observables is determined by the Poisson bracket

{F,G}={G,F}\{F,G\} = − \{G, F\}

If this Poisson bracket is zero, that is, if the two observables F(p,q,t)F(p, q, t) and G(p,q,t)G(p, q, t) commute, then their values are independent and can be measured independently. However, if the Poisson bracket {F,G}0\{F,G\} \neq 0, that is F(p,q,t)F(p, q, t) and G(p,q,t)G(p, q, t) do not commute, then FF and GG are correlated since interchanging the order of the Poisson bracket changes the sign which implies that the measured value for FF depends on whether GG is simultaneously measured.

A useful property of Poisson brackets is that if FF and GG both are constants of motion, then the double Poisson bracket {H,{F,G}}=0\{H, \{F,G\}\} = 0. This can be proved using Jacobi’s identity

{F,{G,H}}+{G,{H,F}}+{H,{F,G}}=0(15.49)\{F, \{G, H\}\} + \{G, \{H, F\}\} + \{H, \{F,G\}\} = 0 \tag{15.49}

If {G,H}=0\{G, H\} = 0 and {F,H}=0\{F,H\} = 0, then {H,{F,G}}=0\{H, \{F,G\}\} = 0, that is, the Poisson bracket {F,G}\{F,G\} commutes with HH. Note that if FF and GG do not depend explicitly on time, that is Ft=Gt=0\frac{\partial F}{ \partial t} = \frac{\partial G}{ \partial t} = 0, then combining equations 15.45 and 15.49 leads to Poisson’s Theorem that relates the total time derivatives.

ddt{F,G}={dFdt,G}+{F,dGdt}\frac{d}{ dt} \{F,G\} = \left\{ \frac{dF}{ dt} , G\right\} + \left\{ F, \frac{dG}{ dt} \right\}

This implies that if FF and GG are invariants, that is dFdt=dGdt=0\frac{dF}{ dt} = \frac{dG}{ dt} = 0, then the Poisson bracket {F,G}\{F,G\} is an invariant if FF and GG are not explicitly time dependent.

Hamilton’s equations of motion

An especially important application of Poisson brackets is that Hamilton’s canonical equations of motion can be expressed directly in the Poisson bracket form. The Poisson bracket representation of Hamiltonian mechanics has important implications to quantum mechanics as will be described in chapter 18.

In Equation 15.45 assume that GG is a fundamental coordinate, that is, Gqk,G \equiv q_k,. Since qkq_k is not explicitly time dependent, then

dqkdt=qkt+{qk,H}=0+i(qkqiHpiqkpiHqi)=i(δikHpi0Hqi)=Hpk\begin{align} \frac{dq_k}{ dt} &= \frac{\partial q_k}{ \partial t} + \{q_k, H\} \tag{15.51} \\[4pt] &= 0+\sum_i \left(\frac{\partial q_k }{\partial q_i} \frac{\partial H }{\partial p_i} − \frac{\partial q_k}{ \partial p_i} \frac{\partial H }{\partial q_i }\right) \nonumber \\[4pt] &= \sum_i \left( \delta_{ik} \frac{\partial H}{ \partial p_i} − 0 \cdot \frac{\partial H}{ \partial q_i} \right) \nonumber \\[4pt] &= \frac{\partial H}{ \partial p_k} \tag{15.52}\end{align}

That is

q˙k={qk,H}=Hpk\dot{q}_k = \{q_k, H\} = \frac{\partial H}{ \partial p_k}

Similarly consider the fundamental canonical momentum GpkG \equiv p_k. Since it is not explicitly time dependent, then

dpkdt=pkt+{pk,H}=0+i(qkqiHpiqkpiHqi)=i(0HpiδikHqi)=Hqk\begin{align} \frac{dp_k}{ dt} &= \frac{\partial p_k}{ \partial t} + \{p_k, H\} \tag{15.54} \\[4pt] &= 0+\sum_i \left(\frac{\partial q_k }{\partial q_i} \frac{\partial H }{\partial p_i} − \frac{\partial q_k}{ \partial p_i} \frac{\partial H }{\partial q_i }\right) \nonumber \\[4pt] &= \sum_i \left( 0 \frac{\partial H}{ \partial p_i} − \delta_{ik} \cdot \frac{\partial H}{ \partial q_i} \right) \nonumber \\[4pt] &= \frac{\partial H}{ \partial q_k} \tag{15.55}\end{align}

That is

p˙k={pk,H}=Hqk\dot{p}_k = \{p_k, H\} = \frac{\partial H}{ \partial q_k}

Thus, it is seen that the Poisson bracket form of the equations of motion includes the Hamilton equations of motion. That is,

q˙k={qk,H}=Hpk(15.57)\dot{q}_k = \{q_k, H\} = \frac{\partial H}{ \partial p_k} \tag{15.57}
p˙k={pk,H}=Hqk(15.58)\dot{p}_k = \{p_k, H\} = −\frac{\partial H}{ \partial q_k} \tag{15.58}

The above shows that the full structure of Hamilton’s equations of motion can be expressed directly in terms of Poisson brackets.

The elegant formulation of Poisson brackets has the same form in all canonical coordinates as the Hamiltonian formulation. However, the normal Hamilton canonical equations in classical mechanics assume implicitly that one can specify the exact position and momentum of a particle simultaneously at any point in time which is applicable only to classical mechanics variables that are continuous functions of the coordinates, and not to quantized systems. The important feature of the Poisson Bracket representation of Hamilton’s equations is that it generalizes Hamilton’s equations into a form 15.57, 15.58 where the Poisson bracket is equally consistent with both classical and quantum mechanics in that it allows for non-commuting canonical variables and Heisenberg’s Uncertainty Principle. Thus the generalization of Hamilton’s equations, via use of the Poisson brackets, provides one of the most powerful analytic tools applicable to both classical and quantal dynamics. It played a pivotal role in derivation of quantum theory as described in chapter 18.

Liouville’s Theorem

Liouvilles Theorem illustrates an application of Poisson Brackets to Hamiltonian phase space that has important implications for statistical physics. The trajectory of a single particle in phase space is completely determined by the equations of motion if the initial conditions are known. However, many-body systems have so many degrees of freedom it becomes impractical to solve all the equations of motion of the many bodies. An example is a statistical ensemble in a gas, a plasma, or a beam of particles. Usually it is not possible to specify the exact point in phase space for such complicated systems. However, it is possible to define an ensemble of points in phase space that encompasses all possible trajectories for the complicated system. That is, the statistical distribution of particles in phase space can be specified.

Infinitessimal element of area in phase space

Figure 15.2.1:Infinitessimal element of area in phase space

Consider a density ρ\rho of representative points in (q,p)(\mathbf{q}, \mathbf{p}) phase space. The number NN of systems in the volume element dvdv is

N=ρdvN = \rho dv

where it is assumed that the infinitessimal volume element dv=dq1,dq2....dqs,dp1,dp2....dpsdv = dq_1, dq_2....dq_s,dp_1, dp_2....dp_s contains many possible systems so that ρ\rho can be considered a continuous distribution. For the conjugate variables (qi,pi)(q_i, p_i) shown in Figure 15.2.1, the number of representative points moving across the left-hand edge into the area per unit time is

ρq˙idpi\rho \dot{q}_i dp_i

The number of representative points flowing out of the area along the right-hand edge is

[ρq˙i+qi(ρq˙i)dqi]dpi\left[ \rho \dot{q}_i + \frac{\partial}{ \partial q_i} (\rho \dot{q}_i) dq_i \right] dp_i

Hence the net increase in ρ\rho in the infinitessimal rectangular element dqidpidq_idp_i due to flow in the horizontal direction is

qi(ρq˙i)dqidpi− \frac{\partial}{ \partial q_i} (\rho \dot{q}_i) dq_idp_i

Similarly, the net gain due to flow in the vertical direction is

pi(ρp˙i)dpidqi− \frac{\partial}{ \partial p_i} (\rho \dot{p}_i) dp_idq_i

Thus the total increase in the element dqidpidq_idp_i per unit time is therefore

[qi(ρq˙i)+pi(ρp˙i)]dpidqi− \left[ \frac{\partial}{ \partial q_i } (\rho \dot{q}_i) + \frac{\partial}{ \partial p_i} (\rho \dot{p}_i) \right] dp_idq_i

Assume that the total number of points must be conserved, then the total increase in the number of points inside the element dqidpidq_idp_i must equal the net changes in ρ\rho on the infinitessimal surface element per unit time. That is

(ρt)dqidpi\left(\frac{\partial \rho}{ \partial t} \right) dq_idp_i

Thus summing over all possible values of ii gives

ρt+i[qi(ρq˙i)+pi(ρp˙i)]=0\frac{\partial \rho }{\partial t} + \sum_i \left[ \frac{\partial}{ \partial q_i} (\rho \dot{q}_i) + \frac{\partial}{ \partial p_i} (\rho \dot{p}_i) \right] = 0

or

ρt+i[q˙iρqi+p˙iρpi]+ρi[p˙ipi+q˙iqi]=0\frac{\partial \rho}{ \partial t} +\sum_i \left[ \dot{q}_i \frac{\partial \rho}{ \partial q_i } + \dot{p}_i \frac{\partial \rho}{ \partial p_i} \right] + \rho \sum_i \left[ \frac{\partial \dot{p}_i}{ \partial p_i} + \frac{\partial \dot{q}_i}{ \partial q_i } \right] = 0

Inserting Hamilton’s canonical equations into both brackets and differentiating the last bracket results in

ρt+i[HpiρqiHqiρpi]+ρi[2Hpiqi2Hpiqi]=0\frac{\partial \rho}{ \partial t} +\sum_i \left[ \frac{\partial H}{ \partial p_i} \frac{\partial \rho }{\partial q_i} − \frac{\partial H}{ \partial q_i} \frac{\partial \rho}{ \partial p_i } \right] + \rho \sum_i \left[\frac{ \partial^2 H}{ \partial p_i\partial q_i} − \frac{\partial^2H}{ \partial p_i\partial q_i} \right] = 0

The two terms in the last bracket cancel and thus

ρt+i[HpiρqiHqiρpi]=ρt+{ρ,H}=0\frac{\partial \rho }{\partial t} +\sum_i \left[ \frac{\partial H }{\partial p_i} \frac{\partial \rho} { \partial q_i} − \frac{\partial H}{ \partial q_i} \frac{\partial \rho}{ \partial p_i} \right] = \frac{\partial \rho}{ \partial t} + \{\rho , H\} = 0

However, this just equals dρdt\frac{d\rho}{ dt}, therefore

dρdt=ρt+{ρ,H}=0(15.70)\frac{d\rho }{dt} = \frac{\partial \rho}{ \partial t} + \{\rho , H\} = 0 \tag{15.70}

This is called Liouville’s theorem which states that the rate of change of density of representative points vanishes, that is, the density of points is a constant in the Hamiltonian phase space along a specific trajectory. Liouville’s theorem means that the system acts like an incompressible fluid that moves such as to occupy an equal volume in phase space at every instant, even though the shape of the phase-space volume may change, that is, the phase-space density of the fluid remains constant. Equation 15.70 is another illustration of the basic Poisson bracket relation 15.45 and the usefulness of Poisson brackets in physics.

Liouville’s theorem is crucially important to statistical mechanics of ensembles where the exact knowledge of the system is unknown, only statistical averages are known. An example is in focussing of beams of charged particles by beam handling systems. At a focus of the beam, the transverse width in xx is minimized, while the width in pxp_x is largest since the beam is converging to the focus, whereas a parallel beam has maximum width xx and minimum spreading width pxp_x. However, the product xpxxp_x remains constant throughout the focussing system. For a two dimensional beam, this applies equally for the yy and pyp_y coordinates, etc. It is obvious that the final beam quality for any beam transport system is ultimately limited by the emittance of the source of the beam, that is, the initial area of the phase space distribution. Note that Liouville’s theorem only applies to Hamiltonian qipiq_i − p_i phase space, not to xx˙x − \dot{x} Lagrangian state space. As a consequence, Hamiltonian dynamics, rather than Lagrange dynamics, is used to discuss ensembles in statistical physics.

Note that Liouville’s theorem is applicable only for conservative systems, that is, where Hamilton’s equations of motion apply. For dissipative systems the phase space volume shrinks with time rather than being a constant of the motion.

15.3: Canonical Transformations in Hamiltonian Mechanics

Hamiltonian mechanics is an especially elegant and powerful way to derive the equations of motion for complicated systems. Unfortunately, integrating the equations of motion to derive a solution can be a challenge. Hamilton recognized this difficulty, so he proposed using generating functions to make canonical transformations which transform the equations into a known soluble form. Jacobi, a contemporary mathematician, recognized the importance of Hamilton’s pioneering developments in Hamiltonian mechanics, and therefore he developed a sophisticated mathematical framework for exploiting the generating function formalism in order to make the canonical transformations required to solve Hamilton’s equations of motion.

In the Lagrange formulation, transforming coordinates (qi,q˙i)(q_i, \dot{q}_i) to cyclic generalized coordinates (Qi,Q˙i)(Q_i, \dot{Q}_i), simplifies finding the Euler-Lagrange equations of motion. For the Hamiltonian formulation, the concept of coordinate transformations is extended to include simultaneous canonical transformation of both the spatial coordinates qiq_i and the conjugate momenta pip_i from (qi,pi)(q_i, p_i) to (Qi,Pi)(Q_i, P_i), where both of the canonical variables are treated equally in the transformation. Compared to Lagrangian mechanics, Hamiltonian mechanics has twice as many variables which is an asset, rather than a liability, since it widens the realm of possible canonical transformations.

Hamiltonian mechanics has the advantage that generating functions can be exploited to make canonical transformations to find solutions, which avoids having to use direct integration. Canonical transformations are the foundation of Hamiltonian mechanics; they underlie Hamilton-Jacobi theory and action-angle variable theory, both of which are powerful means for exploiting Hamiltonian mechanics to solve problems in physics and engineering. The concept underlying canonical transformations is that, if the equations of motion are simplified by using a new set of generalized variables (Q,P)(\mathbf{Q},\mathbf{P}), compared to using the original set of variables (q,p)(\mathbf{q},\mathbf{p}), then an advantage has been gained. The solution, expressed in terms of the generalized variables (Q,P)(\mathbf{Q},\mathbf{P}), can be transformed back to express the solution in terms of the original coordinates, (q,p)(\mathbf{q},\mathbf{p}).

Only a specialized subset of transformations will be considered, namely canonical transformations that preserve the canonical form of Hamilton’s equations of motion. That is, given that the original set of variables (qi,pi)(q_i, p_i) satisfy Hamilton’s equations

q˙=H(q,p,t)pp˙=H(q,p,t)q(15.71)\mathbf{\dot{q}} = \frac{\partial H (\mathbf{q},\mathbf{p}, t)}{ \partial \mathbf{p}} \quad − \mathbf{\dot{p}} = \frac{\partial H (\mathbf{q},\mathbf{p}, t)}{ \partial \mathbf{q}} \tag{15.71}

for some Hamiltonian H(q,p,t)H(\mathbf{q},\mathbf{p}, t), then the transformation to coordinates Qi(qk,pk,t),Pi(qk,pk,t)Q_i(q_k,p_k, t), P_i (q_k, p_k, t) is canonical if, and only if, there exists a function H(Q,P,t)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) such that the P\mathbf{ P} and Q\mathbf{ Q} are still governed by Hamilton’s equations. That is,

Q˙=H(Q,P,t)PP˙=H(Q,P,t)Q(15.72)\mathbf{\dot{Q}} = \frac{\partial\mathcal{H}(\mathbf{Q},\mathbf{P}, t)}{ \partial \mathbf{P}} \quad − \mathbf{\dot{P}} = \frac{\partial\mathcal{H}(\mathbf{Q},\mathbf{P}, t) }{\partial \mathbf{Q}} \tag{15.72}

where H(Q,P,t)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) plays the role of the Hamiltonian for the new variables. Note that H(Q,P,t)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) may be very different from the old Hamiltonian H(q,p,t)H(\mathbf{q},\mathbf{p}, t). The invariance of the Poisson bracket to canonical transformations, chapter 15.2, provides a powerful test that the transformation is canonical.

Hamilton’s Principle of least action, discussed in chapter 9, states that

δS=δt1t2L(q,q˙,t)dt=δt1t2[pq˙H(q,p,t)]dt=0(15.73)\delta S = \delta \int^{t_2}_{t_1} L(\mathbf{q}, \mathbf{\dot{q}}, t)dt = \delta \int^{t_2}_{t_1} [\mathbf{p} \cdot \mathbf{\dot{q}} − H(\mathbf{q},\mathbf{p}, t)] dt = 0 \tag{15.73}

Similarly, applying Hamilton’s Principle of least action to the new Lagrangian L(Q,Q˙,t)\mathcal{L}(\mathbf{Q}, \mathbf{\dot{Q}} , t) gives

δS=δt1t2L(Q,Q˙,t)dt=δt1t2[PQ˙H(Q,P,t)]dt=0(15.74)\delta S = \delta \int^{t_2}_{t_1} \mathcal{L}(\mathbf{Q}, \mathbf{\dot{Q}} , t)dt = \delta \int^{t_2}_{t_1} \left[ \mathbf{P} \cdot \mathbf{\dot{Q}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) \right] dt = 0 \tag{15.74}

The discussion of gauge-invariant Lagrangians, chapter 9.3, showed that LL and L\mathcal{L} can be related by the total time derivative of a generating function FF where

dFdt=LL(15.75)\frac{dF}{ dt} = \mathcal{L} − L \tag{15.75}

The generating function FF can be any well-behaved function with continuous second derivatives of both the old and new canonical variables p\mathbf{p}, q\mathbf{q}, P\mathbf{P}, Q\mathbf{Q} and tt. Thus the integrands of 15.73 and 15.74 are related by

pq˙H(q,p,t)=λ[PQ˙H(Q,P,t)]+dFdt(15.76)\mathbf{p} \cdot \mathbf{\dot{q}} − H(\mathbf{q},\mathbf{p}, t) = \lambda \left[ \mathbf{P} \cdot \mathbf{\dot{Q}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) \right] + \frac{dF}{dt} \tag{15.76}

where λ\lambda is a possible scale transformation. A scale transformation, such as changing units, is trivial, and will be assumed to be absorbed into the coordinates, making λ=1\lambda = 1. Assuming that λ1\lambda \neq 1 is called an extended canonical transformation.

Generating functions

The generating function FF has to be chosen such that the transformation from the initial variables (q,p)( \mathbf{q},\mathbf{p}) to the final variables (Q,P)(\mathbf{Q},\mathbf{P}) is a canonical transformation. The chosen generating function contributes to 15.76 only if it is a function of the old plus new variables. The four possible types of generating functions of the first kind, are F1(q,Q,t)F_1(\mathbf{q}, \mathbf{Q}, t), F2(q,P,t)F_2(\mathbf{q},\mathbf{P}, t), F3(p,Q,t)F_3(\mathbf{p}, \mathbf{Q}, t), and F4(p,P,t)F_4(\mathbf{p}, \mathbf{P}, t). These four generating functions lead to relatively simple canonical transformations, are shown below.

Type 1: F=F1(q,Q,t)F = F_1(\mathbf{q}, \mathbf{Q},t):

The total time derivative of the generating function F=F1(q,Q,t)F = F_1(\mathbf{q}, \mathbf{Q},t) is given by

dF(q,Q,t)dt=[F1(q,Q,t)qq˙+F1(q,Q,t)QQ˙]+F1(q,Q,t)t(15.77)\frac{dF(\mathbf{q}, \mathbf{Q},t)}{ dt} = \left[ \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{q}} \cdot \mathbf{\dot{q}} + \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{Q}} \cdot \mathbf{\dot{Q}} \right] + \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t) }{\partial t} \tag{15.77}

Insert Equation 15.77 into Equation 15.76, and assume that the trivial scale factor λ=1\lambda = 1, then

[pF1(q,Q,t)q]q˙H(q,p,t)=[P+F1(q,Q,t)Q]Q˙H(Q,P,t)+F1(q,Q,t)t\left[ \mathbf{p} − \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{q}} \right] \cdot \mathbf{\dot{q}} − H(\mathbf{q},\mathbf{p}, t) = \left[ \mathbf{P} + \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{Q}} \right] \cdot \mathbf{\dot{Q}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) + \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t) }{\partial t} \nonumber

Assume that the generating function F1F_1 determines the canonical variables p\mathbf{p} and P\mathbf{P} to be

p=F1(q,Q,t)qP=F1(q,Q,t)Q(15.78)\mathbf{p} = \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{q}} \qquad \mathbf{P} = −\frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial \mathbf{Q}} \tag{15.78}

then the terms in each square bracket cancel, leading to the required canonical transformation

H(Q,P,t)=H(q,p,t)+F1(q,Q,t)t(15.79)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial F_1(\mathbf{q}, \mathbf{Q},t)}{ \partial t} \tag{15.79}

Type 2: F=F2(q,P,t)QPF = F_2(\mathbf{q},\mathbf{P},t) − \mathbf{Q} \cdot \mathbf{P}:

The total time derivative of the generating function F=F2(q,P,t)QPF = F_2(\mathbf{q},\mathbf{P},t)−\mathbf{Q} \cdot \mathbf{P} is given by

dFdt=[F2(q,P,t)qq˙+F2(q,P,t)Pp˙PQ˙P˙Q]+F2(q,P,t)t(15.80)\frac{dF}{ dt} = \left[ \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial \mathbf{q}} \cdot \mathbf{\dot{q}} + \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial \mathbf{P}} \cdot \mathbf{\dot{p}} − \mathbf{P} \cdot \mathbf{\dot{Q}} − \mathbf{\dot{P}} \cdot \mathbf{Q} \right] + \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial t} \tag{15.80}

Insert this into Equation 15.76, and assume that the trivial scale factor λ=1\lambda = 1, then

(pF2(q,P,t)q)q˙H(q,p,t)=PQ˙PQ˙+[F2(q,P,t)PQ]P˙H(Q,P,t)+F2(q,P,t)t\left( \mathbf{p} − \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial \mathbf{q}} \right) \cdot \mathbf{\dot{q}} − H(\mathbf{q},\mathbf{p}, t) = \mathbf{P} \cdot \mathbf{\dot{Q}} − \mathbf{P} \cdot \mathbf{\dot{Q}} + \left[ \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial \mathbf{P}} − \mathbf{Q} \right] \cdot \mathbf{\dot{P}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) + \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial t } \nonumber

Assume that the generating function F2F_2 determines the canonical variables p\mathbf{p} and Q\mathbf{Q} to be

p=F2(q,P,t)qQ=F2(q,P,t)P(15.81)\mathbf{p} = \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial \mathbf{q}} \qquad \mathbf{Q} = \frac{\partial F_2(\mathbf{q},\mathbf{P},t) }{\partial \mathbf{P}} \tag{15.81}

then the terms in brackets cancel, leading to the required transformation

H(Q,P,t)=H(q,p,t)+F2(q,P,t)t(15.82)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial F_2(\mathbf{q},\mathbf{P},t)}{ \partial t} \tag{15.82}

Type 3: F=F3(p,Q,t)+qpF = F_3(\mathbf{p}, \mathbf{Q},t) + \mathbf{q} \cdot \mathbf{p}:

The total time derivative of the generating function F=F3(p,Q,t)+qpF = F_3(\mathbf{p}, \mathbf{Q},t) + \mathbf{q} \cdot \mathbf{p} is given by

dFdt=[F3(p,Q,t)pp˙+F3(p,Q,t)QQ˙+q˙p+qp˙]+F3(p,Q,t)t(15.83)\frac{dF }{dt} = \left[ \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial \mathbf{p}} \cdot \mathbf{\dot{p}} + \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial\mathbf{ Q}} \cdot \mathbf{\dot{Q}} + \mathbf{\dot{q}} \cdot \mathbf{p} + \mathbf{q} \cdot \mathbf{\dot{p}} \right] + \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial t} \tag{15.83}

Insert this into Equation 15.76, and assume that the trivial scale factor λ=1\lambda = 1, then

[q+F3(p,Q,t)p]p˙H(q,p,t)=[P+F3(p,Q,t)Q]Q˙H(Q,P,t)+F3(p,Q,t)t− \left[ \mathbf{q}+ \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial \mathbf{p}} \right] \cdot \mathbf{\dot{p}} − H(\mathbf{q},\mathbf{p}, t) = \left[ \mathbf{P}+ \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial \mathbf{Q}} \right] \cdot \mathbf{\dot{Q}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) + \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial t} \nonumber

Assume that the generating function F3F_3 determines the canonical variables q\mathbf{q} and P\mathbf{P} to be

q=F3(p,Q,t)pP=F3(p,Q,t)Q(15.84)\mathbf{q} = −\frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial \mathbf{p}} \qquad \mathbf{P} = −\frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial \mathbf{Q}} \tag{15.84}

then the terms in brackets cancel, leading to the required transformation

H(Q,P,t)=H(q,p,t)+F3(p,Q,t)t(15.85)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial F_3(\mathbf{p}, \mathbf{Q},t)}{ \partial t} \tag{15.85}

Type 4: F=F4(p,P,t)+qpQPF = F_4(\mathbf{p}, \mathbf{P},t) + \mathbf{q} \cdot \mathbf{p} − \mathbf{Q} \cdot \mathbf{P}:

The total time derivative of the generating function F=F4(p,P,t)+qpQPF = F_4(\mathbf{p}, \mathbf{P},t) + \mathbf{q} \cdot \mathbf{p} − \mathbf{Q} \cdot \mathbf{P} is given by

dFdt=[F4(p,P,t)pp˙+F4(p,P,t)Pp˙+q˙p+qp˙Q˙PQP˙]+F4(p,P,t)t(15.86)\frac{dF }{dt} = \left[ \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial \mathbf{p}} \cdot \mathbf{\dot{p}} + \frac{\partial F_4(\mathbf{p}, \mathbf{P},t) }{\partial \mathbf{P}} \cdot \mathbf{\dot{p}} + \mathbf{\dot{q}} \cdot \mathbf{p} + \mathbf{q} \cdot \mathbf{\dot{p}} − \mathbf{\dot{Q}} \cdot \mathbf{P} − \mathbf{Q} \cdot \mathbf{\dot{P}} \right] + \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial t}\tag{15.86}

Insert this into Equation 15.76, and assume that the trivial scale factor λ=1\lambda = 1, then

[q+F4(p,P,t)p]p˙H(q,p,t)=[F4(p,P,t)PQ]P˙H(Q,P,t)+F4(p,P,t)t− \left[ \mathbf{q}+ \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial \mathbf{p}} \right] \cdot \mathbf{\dot{p}} − H(\mathbf{q},\mathbf{p}, t) = \left[ \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial \mathbf{P} } − \mathbf{Q} \right] \cdot \mathbf{\dot{P}} − \mathcal{H}(\mathbf{Q},\mathbf{P}, t) + \frac{\partial F_4(\mathbf{p}, \mathbf{P},t) }{\partial t} \nonumber

Assume that the generating function F4F_4 determines the canonical variables q\mathbf{q} and Q\mathbf{Q} to be

q=F4(p,P,t)pQ=F4(p,P,t)P(15.87)\mathbf{q} = −\frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial \mathbf{p}} \qquad \mathbf{Q} = \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial \mathbf{P}} \tag{15.87}

then the terms in brackets cancel, leading to the required transformation

H(Q,P,t)=H(q,p,t)+F4(p,P,t)t(15.88)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial F_4(\mathbf{p}, \mathbf{P},t)}{ \partial t} \tag{15.88}

Note that the last three generating functions require the inclusion of additional bilinear products of qq, pp, QQ, PP in order for the terms to cancel to give the required result. The addition of the bilinear terms, ensures that the resultant generating function FF is the same using any of the four generating functions F1F_1, F2F_2, F3F_3, F4F_4. Frequently the F2(q,P,t)F_2(\mathbf{q},\mathbf{P}, t) generating function is the most convenient. The four possible generating functions of the first kind, given above, are related by Legendre transformations. A canonical transformation does not have to conform to only one of the four generating functions FkF_k for all the degrees of freedom, they can be a mixture of different flavors for the different degrees of freedom. The properties of the generating functions are summarized in table 15.3.1.

Generating functionGenerating function derivativesTrivial special examples
F=F1(q,Q,t)F = F_1(\mathbf{q}, \mathbf{Q}, t)pi=F1qiPi=F1Qip_i = \frac{\partial F_1}{ \partial q_i } \quad P_i = −\frac{\partial F_1}{ \partial Q_i}F1=qiQiQi=piPi=qiF_1 = q_iQ_i \quad Q_i = p_i \quad P_i = −q_i
F=F2(q,P,t)QPF = F_2(\mathbf{q},\mathbf{P}, t) − \mathbf{Q} \cdot \mathbf{P}pi=F2qiQi=F2Pip_i = \frac{\partial F_2}{ \partial q_i} \quad Q_i = \frac{\partial F_2}{ \partial P_i}F2=qiPiQi=qiPi=piF_2 = q_iP_i \quad Q_i = q_i \quad P_i = p_i
F=F3(p,Q,t)+qpF = F_3(\mathbf{p}, \mathbf{Q},t) + \mathbf{q} \cdot \mathbf{p}qi=F3piPi=F3Qiq_i = −\frac{\partial F_3}{ \partial p_i} \quad P_i = −\frac{\partial F_3}{ \partial Q_i}F3=piQiQi=qiPi=piF_3 = p_iQ_i \quad Q_i = −q_i \quad P_i = −p_i
F=F4(p,P,t)+qpQPF = F_4(\mathbf{p},\mathbf{P},t) + \mathbf{q} \cdot \mathbf{p} − \mathbf{Q} \cdot \mathbf{P}qi=F4piQi=F4Piq_i = −\frac{\partial F_4}{ \partial p_i } \quad Q_i = \frac{\partial F_4}{ \partial P_i}F4=piPiQi=piPi=qiF_4 = p_iP_i \quad Q_i = p_i \quad P_i = −q_i

The partial derivatives of the generating functions FiF_i determine the corresponding conjugate variables not explicitly included in the generating function FiF_i. Note that, for the first trivial example F1=qiQiF_1 = q_iQ_i, the old momenta become the new coordinates, Qi=piQ_i = p_i, and vice versa, Pi=qiP_i = −q_i. This illustrates that it is better to name them “conjugate variables” rather than “momenta” and “coordinates”.

In summary, Jacobi has developed a mathematical framework for finding the generating function FF required to make a canonical transformation to a new Hamiltonian H(Q,P,t)\mathcal{H}(\mathbf{Q},\mathbf{P}, t), that has a known solution. That is,

H(Q,P,t)=H(q,p,t)+Ft(15.89)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial F}{ \partial t} \tag{15.89}

When H(Q,P,t)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) is a constant, then a solution has been obtained. The inverse transformation for this solution Q(t),P(t)q(t),p(t)\mathbf{Q}(t), \mathbf{P}(t) \rightarrow \mathbf{q}(t), \mathbf{p}(t) now can be used to express the final solution in terms of the original variables of the system.

Note the special case when H(Q,P,t)=0\mathcal{H}(\mathbf{Q},\mathbf{P}, t)=0, then Equation 15.89 has been reduced to the Hamilton-Jacobi relation 15.11

H(q,p,t)+St=0(15.11)H(\mathbf{q},\mathbf{p}, t) + \frac{\partial S}{\partial t} = 0 \tag{15.11}

In this case, the generating function FF determines the action functional SS required to solve the Hamilton-Jacobi equation (15.4.23)(15.4.23)). Since Equation 15.89 has transformed the Hamiltonian H(q,p,t)H(Q,P,t)H(\mathbf{q},\mathbf{p}, t) \rightarrow \mathcal{H}(\mathbf{Q},\mathbf{P}, t), for which H(Q,P,t)=0\mathcal{H}(\mathbf{Q},\mathbf{P}, t)=0, then the solution Q(t),P(t)\mathbf{Q}(t), \mathbf{P}(t) for the Hamiltonian H(Q,P,t)=0\mathcal{H}(\mathbf{Q},\mathbf{P}, t)=0 is obtained easily. This approach underlies Hamilton-Jacobi theory presented in chapter 15.4.

Applications of Canonical Transformations

The canonical transformation procedure may appear unnecessarily complicated for solving the examples given in this book, but it is essential for solving the complicated systems that occur in nature. For example, canonical transformations can be used to transform time-dependent, (non-autonomous) Hamiltonians to time-independent, (autonomous) Hamiltonians for which the solutions are known. Example 15.6.2 describes such a system. Canonical transformations provide a remarkably powerful approach for solving the equations of motion in Hamiltonian mechanics, especially when using the Hamilton-Jacobi approach discussed in chapter 15.4.

15.4: Hamilton-Jacobi Theory

Hamilton used the Principle of Least Action to derive the Hamilton-Jacobi relation (chapter 15.3)

H(q,p,t)+St=0(15.11)H(\mathbf{q},\mathbf{p}, t) + \frac{\partial S}{\partial t} = 0 \tag{15.11}

where q,p\mathbf{q}, \mathbf{p} refer to the 1in1 \leq i \leq n variables qi,piq_i, p_i and S(qj(t1),t1,qj(t2),t2)S(q_j (t_1), t_1, q_j (t_2), t_2) is the action functional. Integration of this first-order partial differential equation is non trivial which is a major handicap for practical exploitation of the Hamilton-Jacobi equation. This stimulated Jacobi to develop the mathematical framework for canonical transformation that are required to solve the Hamilton-Jacobi equation. Jacobi’s approach is to exploit generating functions for making a canonical transformation to a new Hamiltonian H(Q,P,t)\mathcal{H}(\mathbf{Q}, \mathbf{P}, t) that equals zero.

H(Q,P,t)=H(q,p,t)+St=0(15.90)\mathcal{H}(\mathbf{Q},\mathbf{P}, t) = H(\mathbf{q},\mathbf{p}, t) + \frac{\partial S}{\partial t} = 0 \tag{15.90}

The generating function for solving the Hamilton-Jacobi equation then equals the action functional SS.

The Hamilton-Jacobi theory is based on selecting a canonical transformation to new coordinates (Q,P,t)(Q, P, t) all of which are either constant, or the QiQ_i are cyclic, which implies that the corresponding momenta PiP_i are constants. In either case, a solution to the equations of motion is obtained. A remarkable feature of Hamilton-Jacobi theory is that the canonical transformation is completely characterized by a single generating function, SS. The canonical equations likewise are characterized by a single Hamiltonian function, HH. Moreover, the generating function SS, and Hamiltonian function HH, are linked together by Equation 15.11. The underlying goal of Hamilton-Jacobi theory is to transform the Hamiltonian to a known form such that the canonical equations become directly integrable. Since this transformation depends on a single scalar function, the problem is reduced to solving a single partial differential equation.

Time-dependent Hamiltonian

Jacobi’s complete integral S(qi,Pi,t)S(q_i, P_i, t)

The principle underlying Jacobi’s approach to Hamilton-Jacobi theory is to provide a recipe for finding the generating function F=SF = S needed to transform the Hamiltonian H(q,p,t)H(\mathbf{q}, \mathbf{p}, t) to the new Hamiltonian H(Q,P,t)\mathcal{H}(\mathbf{Q}, \mathbf{P}, t) using Equation 15.90. When the derivatives of the transformed Hamiltonian H(Q,P,t)\mathcal{H}(\mathbf{Q}, \mathbf{P}, t) are zero, then the equations of motion become

Q˙i=HPi=0(15.91)\dot{Q}_i = \frac{\partial \mathcal{H}}{ \partial P_i} = 0 \tag{15.91}
P˙i=HQi=0(15.92)\dot{P}_i = − \frac{\partial \mathcal{H}}{ \partial Q_i } = 0 \tag{15.92}

and thus QiQ_i and PiP_i are constants of motion. The new Hamiltonian H\mathcal{H} must be related to the original Hamiltonian HH by a canonical transformation for which

H(Q,P,t)=H(q,p,t)+St(15.93)\mathcal{H}(\mathbf{Q}, \mathbf{P}, t) = H(\mathbf{q}, \mathbf{p}, t) + \frac{\partial S}{ \partial t} \tag{15.93}

Equations 15.91 and 15.92 are automatically satisfied if the new Hamiltonian H=0\mathcal{H} = 0 since then Equation 15.93 gives that the generating function SS satisfies Equation 15.90.

Any of the four types of generating function can be used. Jacobi chose the type 2 generating function as being the most useful for many practical cases, that is, S(qi,Pi,t)S(q_i, P_i, t) which is called Jacobi’s complete integral.

For generating functions F1F_1 and F2F_2 the generalized momenta are derived from the action by the derivative

pi=Sqi(15.4)p_i = \frac{\partial S}{ \partial q_i} \tag{15.4}

Use this generalized momentum to replace pip_i in the Hamiltonian HH, given in Equation 15.93, leads to the Hamilton-Jacobi equation expressed in terms of the action SS.

H(q1,...qn;Sq1,...,Sqn;t)+St=0(15.94)H(q_1, ...q_n; \frac{\partial S}{ \partial q_1 }, ..., \frac{\partial S}{ \partial q_n} ;t) + \frac{\partial S}{ \partial t} = 0 \tag{15.94}

The Hamilton-Jacobi equation, 15.94, can be written more compactly using tensors q\mathbf{q} and S\boldsymbol{\nabla}S to designate (q1,..qn)(q_1, ..q_n) and Sq1,...,Sqn\frac{\partial S}{ \partial q_1 }, ..., \frac{\partial S}{ \partial q_n} respectively. That is

H(q,S,t)+St=0(15.95)H(\mathbf{q}, \boldsymbol{\nabla}S, t) + \frac{\partial S}{\partial t} = 0 \tag{15.95}

Equation 15.95 is a first-order partial differential equation in n+1n + 1 variables which are the old spatial coordinates qiq_i plus time tt. The new momenta PiP_i have not been specified except that they are constants since H=0\mathcal{H} = 0.

Assume the existence of a solution of 15.95 of the form S(qi,Pi,t)=S(q1,..qn;α1,..αn+1;t)S(q_i, P_i, t) = S(q_1, ..q_n; \alpha_1, ..\alpha_{n+1};t) where the generalized momenta Pi=α1,α2,....αP_i = \alpha_1, \alpha_2, ....\alpha plus tt are the n+1n + 1 independent constants of integration in the transformed frame. One constant of integration is irrelevant to the solution since only partial derivatives of S(qi,Pi,t)S(q_i, P_i, t) with respect to qiq_i and tt are involved. Thus, if SS is a solution of the first-order partial differential equation, then so is S+αS + \alpha where α\alpha is a constant. Thus it can be assumed that one of the n+1n + 1 constants of integration is just an additive constant which can be ignored leading effectively to a solution

S(qi,Pi,t)=S(q1,.....qn;α1,.....αn;t)(15.96)S(q_i, P_i, t) = S(q_1, .....q_n;\alpha_1, .....\alpha_n;t) \tag{15.96}

where none of the nn independent constants are solely additive. Such generating function solutions are called complete solutions of the first-order partial differential equations since all constants of integration are known.

It is possible to assume that the nn generalized momenta, PiP_i are constants αi\alpha_i, where the αi\alpha_i are the constants. This allows the generalized momentum to be written as

pi=S(q,α,t)qi(15.97)p_i = \frac{\partial S(\mathbf{q}, \boldsymbol{\alpha}, t)}{ \partial q_i } \tag{15.97}

Similarly, Hamilton’s equations of motion give the conjugate coordinate Q=β\mathbf{Q} = \boldsymbol{\beta}, where βi\beta_i are constants. That is

Qi=βi=S(q,α,t)αi(15.98)Q_i = \beta_i = \frac{\partial S(\mathbf{q}, \boldsymbol{\alpha}, t)}{ \partial \alpha_i} \tag{15.98}

The above procedure has determined the complete set of 2n2n constants (Q=β,P=α)(\mathbf{Q} = \boldsymbol{\beta}, \mathbf{P} = \boldsymbol{\alpha}). It is possible to invert the canonical transformation to express the above solution, which is expressed in terms of Qi=βiQ_i = \beta_i and Pi=αiP_i = \alpha_i, back to the original coordinates, that is, qj=qj(α,β,t)q_j = q_j (\alpha , \beta , t) and momenta pj=pj(α,β,t)p_j = p_j (\alpha , \beta , t) which is the required solution.

Hamilton’s principle function SH(qi,t;qoto)S_H(\mathbf{q}_i, t; \mathbf{q}_o t_o)

Hamilton’s approach to solving the Hamilton-Jacobi Equation 15.95 is to seek a canonical transformation from variables (p,q)(\mathbf{p}, \mathbf{q}) at time tt, to a new set of constant quantities, which may be the initial values (q0,p0)(\mathbf{q}_0, \mathbf{p}_0) at time t=0t = 0. Hamilton’s principle function SH(qi,t;qoto)S_H(q_i, t; q_ot_o) is the generating function for this canonical transformation from the variables (q,p)(\mathbf{q}, \mathbf{p}) at time t to the initial variables (q0,p0)(\mathbf{q}_0, \mathbf{p}_0) at time t0t_0. Hamilton’s principle function SH(qi,t;qoto)S_H(q_i, t; q_ot_o) is directly related to Jacobi’s complete integral S(qi,Pi,t)S(q_i, P_i, t).

Note that SHS_H is the generating function of a canonical transformation from the present time (q,p,t)(\mathbf{q}, \mathbf{p}, t) variables to the initial (q0,p0,t0)(\mathbf{q}_0, \mathbf{p}_0, t_0), whereas Jacobi’s SS is the generating function of a canonical transformation from the present (q,p,t)(\mathbf{q},\mathbf{p}, t) variables to the constant variables (Q=β,P=α)(\mathbf{Q} = \boldsymbol{\beta}, \mathbf{P} = \boldsymbol{\alpha}). For the Hamilton approach, the canonical transformation can be accomplished in two steps using SS by first transforming from (q,p,t)(\mathbf{q}, \mathbf{p}, t) at time tt, to (β,α)(\boldsymbol{\beta}, \boldsymbol{\alpha}), then transforming from (β,α)(\boldsymbol{\beta}, \boldsymbol{\alpha}) to (q0,p0,t0)(\mathbf{q}_0,\mathbf{p}_0, t_0). That is, this two-step process corresponds to

SH(q,t;qoto)=S(q,α,t)S(q0,α,t0)(15.99)S_H(\mathbf{q}, t; \mathbf{q}_ot_o) = S(\mathbf{q}, \boldsymbol{\alpha}, t) − S(\mathbf{q}_0, \boldsymbol{\alpha}, t_0) \tag{15.99}

Hamilton’s principle function SH(q,t;qoto)S_H(\mathbf{q}, t; \mathbf{q}_ot_o) is related to Jacobi’s complete integral S(q,α,t)S(\mathbf{q}, \boldsymbol{\alpha}, t), and it will not be discussed further in this book.

Time-independent Hamiltonian

Frequently the Hamiltonian does not explicitly depend on time. For the standard Lagrangian with time-independent constraints and transformation, then H(q,p,t)=EH (\mathbf{q}, \mathbf{p},t) = E which is the total energy. For this case, the Hamilton-Jacobi equation simplifies to give

St=H(q,p,t)=E(α)(15.100)\frac{\partial S}{ \partial t} = −H( \mathbf{ q}, \mathbf{ p}, t) = −E (\boldsymbol{\alpha}) \tag{15.100}

The integration of the time dependence is trivial, and thus the action integral for a time-independent Hamiltonian equals

S(q,α,t)=W(q,α)E(α)t(15.101)S(\mathbf{q}, \boldsymbol{\alpha},t) = W (\mathbf{q}, \boldsymbol{\alpha}) − E (\boldsymbol{\alpha})t \tag{15.101}

That is, the action integral has separated into a time independent term W(q,α)W (\mathbf{q}, \boldsymbol{\alpha}) which is called Hamilton’s characteristic function plus a time-dependent term E(α)t−E (\boldsymbol{\alpha})t. Thus using equations 15.97, 15.101 gives that the generalized momentum is

pi=W(q,α)qi(15.102)p_i = \frac{\partial W(\mathbf{q}, \boldsymbol{\alpha})}{ \partial q_i} \tag{15.102}

The physical significance of Hamilton’s characteristic function W(q,α)W (\mathbf{q}, \boldsymbol{\alpha}) can be understood by taking the total time derivative

dWdt=iW(q,α)qiq˙i=ipiq˙i\frac{dW}{ dt} = \sum_i \frac{\partial W(\mathbf{q}, \boldsymbol{\alpha})}{ \partial q_i} \dot{q}_i = \sum_i p_i\dot{q}_i \nonumber

Taking the time integral then gives

W(q,α)=piq˙idt=pidqi(15.103)W (\mathbf{q}, \boldsymbol{\alpha}) = \int \sum p_i\dot{q}_i dt =\int \sum p_idq_i \tag{15.103}

Note that this equals the abbreviated action described in chapter 9.2.3, that is W(q,α)=S0(q,α)W(\mathbf{q}, \boldsymbol{\alpha}) = S_0(\mathbf{q}, \boldsymbol{\alpha}).

Inserting the action S(q,α)S (\mathbf{q}, \boldsymbol{\alpha}) into the Hamilton-Jacobi equation (15.2.1)(15.2.1) gives

H(q;W(q,α)q)=E(α)(15.104)H(\mathbf{q}; \frac{\partial W(\mathbf{q}, \boldsymbol{\alpha})}{ \partial \mathbf{q}} ) = E (\boldsymbol{\alpha}) \tag{15.104}

This is called the time-independent Hamilton-Jacobi equation. Usually it is convenient to have EE equal the total energy. However, sometimes it is more convenient to exclude the kthk^{th} energy E(αk)E(\alpha_k) in the set, in which case E=E(α1,α2,...αk1)E = E(\alpha_1, \alpha_2, ...\alpha_k−1); the Routhian exploits this feature.

The equations of the canonical transformation expressed in terms of W(q,α)W (\mathbf{q}, \boldsymbol{\alpha}) are

pi=W(q,α)qiβi+E(α)αit=W(q,α)αi(15.105)p_i = \frac{\partial W(\mathbf{q}, \boldsymbol{\alpha}) }{\partial q_i } \quad \beta_i + \frac{\partial E(\boldsymbol{\alpha}) }{\partial \alpha_i} t = \frac{\partial W(\mathbf{q}, \boldsymbol{\alpha})}{ \partial \alpha_i} \tag{15.105}

These equations show that Hamilton’s characteristic function W(q,α)W (\mathbf{q}, \boldsymbol{\alpha}) is itself the generating function of a time-independent canonical transformation from the old variables (q,p)(q, p) to a set of new variables

Qi=βi+E(α)αitPi=αi(15.106)Q_i = \beta_i + \frac{\partial E(\boldsymbol{\alpha})}{ \partial \alpha_i } t \quad P_i = \alpha_i \tag{15.106}

Table 15.4.1 summarizes the time-dependent and time-independent forms of the Hamilton-Jacobi equation.

Hamiltonian

Time dependent H(q,p,t)H(q, p, t)

Time independent H(q,p)H(q, p)

Transformed Hamiltonian

H=0\mathcal{H}= 0

H\mathcal{H} is cyclic

Canonical transformed variables

All QiPiQ_iP_i are constants of motion

All PiP_i are constants of motion

Transformed equations of motion

Q˙i=HPi=0\dot{Q}_i = \frac{\partial \mathcal{H}}{ \partial P_i} = 0, therefore Qi=βiQ_i = \beta_i P˙i=HQi=0\dot{P}_i = − \frac{\partial \mathcal{H}}{ \partial Q_i} = 0, therefore Pi=αiP_i = \alpha_i

Q˙i=HPi=vi\dot{Q}_i = \frac{\partial \mathcal{H}}{ \partial P_i} = v_i, therefore Qi=vit+βiQ_i = v_i t + \beta_i P˙i=HQi=0\dot{P}_i = − \frac{\partial \mathcal{H}} {\partial Q_i} = 0, therefore Pi=αiP_i = \alpha_i

Generating function

Jacobi’s complete integral S(q,P,t)S(\mathbf{q}, \mathbf{P}, t)

Characteristic Function W(q,P)W(\mathbf{q}, \mathbf{P})

Hamilton-Jacobi equation

H(q1,...qn;Sq1,...,Sqn;t)+St=0H(q_1, ...q_n; \frac{\partial S}{ \partial q_1 }, ..., \frac{\partial S} {\partial q_n} ;t)+\frac{\partial S}{ \partial t} = 0

H(q1,...qn;Wq1,...,Wqn)=EH(q_1, ...q_n; \frac{\partial W }{\partial q_1} , ..., \frac{\partial W}{ \partial q_n} ) = E

Transformation equations

pi=Sqip_i= \frac{\partial S}{ \partial q_i} Qi=Sαi=βiQ_i= \frac{\partial S}{ \partial \alpha_i} = \beta_i

pi=Wqip_i=\frac{\partial W}{ \partial q_i} Qi=Wαi=vit+βiQ_i=\frac{\partial W}{ \partial \alpha_i} = v_i t + \beta_i

Separation of variables

Exploitation of the Hamilton-Jacobi theory requires finding a suitable action function SS. When the Hamiltonian is time independent, then Equation 15.101 shows that the time dependence of the action integral separates out from the dependence on the spatial variables. For many systems, the Hamilton’s characteristic function W(q,P)W(\mathbf{q}, \mathbf{P}) separates into a simple sum of terms each of which is a function of a single variable. That is,

W(q,α)=W1(q1)+W2(q2)+Wn(qn)(15.107)W(\mathbf{q}, \boldsymbol{\alpha}) = W_1(q_1) + W_2(q_2) + \cdots \cdot \cdot W_n(q_n) \tag{15.107}

where each function in the summation on the right depends only on a single variable. Then Equation 15.100 reduces to

H(q1,...qn;Wq1,...,Wqn)=E(15.108)H(q_1, ...q_n; \frac{\partial W }{\partial q_1} , ...,\frac{ \partial W}{ \partial q_n} ) = E \tag{15.108}

where EE is the constant denoting the total energy.

Hamilton’s characteristic function W(q,P)W( \mathbf{ q}, \mathbf{ P}) can be used with equations 15.101, 15.102, 15.91, 15.92, and 15.93 to derive

pi=W(q,α)qiQi=W(q,α)Pi(15.109)p_i = \frac{\partial W( \mathbf{ q}, \boldsymbol{\alpha}) }{\partial q_i} \quad Q_i = \frac{\partial W( \mathbf{ q}, \boldsymbol{\alpha}) }{\partial P_i} \tag{15.109}
Q˙i=HPi=0P˙i=HQi=0(15.110)\dot{Q}_i = \frac{\partial \mathcal{H}}{ \partial P_i} = 0 \quad \dot{P}_i = \frac{\partial \mathcal{H}}{ \partial Q_i} = 0 \tag{15.110}
H=H+St=HE=0(15.111)\mathcal{H} = H + \frac{\partial S}{\partial t} = H − E = 0 \tag{15.111}

which has reduced the problem to a simple sum of one-dimensional first-order differential equations.

If the ithi^{th} variable is cyclic, then the Hamiltonian is not a function of qiq_i and the ithi^{th} term in Hamilton’s characteristic function equals Wi=αiqiW_i = \alpha_iq_i which separates out from the summation in Equation 15.107. That is, all cyclic variables can be factored out of W(q,α)W( \mathbf{ q}, \boldsymbol{\alpha}) which greatly simplifies solution of the Hamilton-Jacobi equation. As a consequence, the ability of the Hamilton-Jacobi method to make a canonical transformation to separate the system into many cyclic or independent variables, which can be solved trivially, is a remarkably powerful way for solving the equations of motion in Hamiltonian mechanics.

Visual representation of the action function SS.

Surfaces of constant action integral S (dashed lines) and the corresponding particle momenta (solid lines) with arrows showing the direction.

Figure 15.4.1:Surfaces of constant action integral S (dashed lines) and the corresponding particle momenta (solid lines) with arrows showing the direction.

The important role of the action integral SS can be illuminated by considering the case of a single point mass mm moving in a time independent potential U(r)U(r). Then the action reduces to

S(q,α,t)=W(q,α)Et(15.112)S(q, \alpha , t) = W(q, \alpha ) − Et \tag{15.112}

Let q1=x,q2=y,q3=z,p1=px,p2=py,p3=pzq_1 = x, q_2 = y, q_3 = z, p_1 = p_x, p_2 = p_y, p_3 = p_z. The momentum components are given by

pi=W(q,α)qi(15.113)p_i = \frac{\partial W(q, \alpha ) }{\partial q_i} \tag{15.113}

which corresponds to

p=W=S(15.114)\mathbf{p} = \boldsymbol{\nabla}W = \boldsymbol{\nabla}S \tag{15.114}

That is, the time-independent Hamilton-Jacobi equation is

12mW2+U(r)=E(15.115)\frac{1}{2m} |\boldsymbol{\nabla}W|^2 + U(r) = E \tag{15.115}

This implies that the particle momentum is given by the gradient of Hamilton’s characteristic function and is perpendicular to surfaces of constant WW as illustrated in Figure 15.4.1. The constant WW surfaces are time dependent as given by Equation 15.101. Thus, if at time t=0t = 0 the equi-action surface S0(q,t)=W0(q,Pi)=0S_0(q, t) = W_0(q, P_i)=0, then at t=1t = 1 the same surface S0(q,t)=0S_0(q, t)=0 now coincides with the S0(q,t)=ES_0(q, t) = E surface etc. That is, the equi-action surfaces move through space separately from the motion of the single point mass.

The above pictorial representation is analogous to the situation for motion of a wavefront for electromagnetic waves in optics, or matter waves in quantum physics where the wave equation separates into the form ϕ=ϕ0eiS=ϕ0ei(krωt)\phi = \phi_0 e^{\frac{ iS}{ \hbar }} = \phi_0 e^{i(\mathbf{k} \cdot \mathbf{r}−\omega t)}. Hamilton’s goal was to create a unified theory for optics that was equally applicable to particle motion in classical mechanics. Thus the optical-mechanical analogy of the Hamilton-Jacobi theory has culminated in a universal theory that describes wave-particle duality; this was a Holy Grail of classical mechanics since Newton’s time. It played an important role in development of the Schrödinger representation of quantum mechanics.

Advantages of Hamilton-Jacobi theory

Initially, only a few scientists, like Jacobi, recognized the advantages of Hamiltonian mechanics. In 1843 Jacobi made some brilliant mathematical developments in Hamilton-Jacobi theory that greatly enhanced exploitation of Hamiltonian mechanics. Hamilton-Jacobi theory now serves as a foundation for contemporary physics, such as quantum and statistical mechanics. A major advantage of Hamilton-Jacobi theory, compared to other formulations of analytic mechanics, is that it provides a single, first-order partial differential equation for the action SS, which is a function of the nn generalized coordinates q\mathbf{q} and time tt. The generalized momenta no longer appear explicitly in the Hamiltonian in equations 15.94, 15.95. Note that the generalized momentum do not explicitly appear in the equivalent Euler-Lagrange equations of Lagrangian mechanics, but these comprise a system of nn second-order, partial differential equations for the time evolution of the generalized coordinate q\mathbf{q}. Hamilton’s equations of motion are a system of 2n2n first-order equations for the time evolution of the generalized coordinates and their conjugate momenta.

An important advantage of the Hamilton-Jacobi theory is that it provides a formulation of classical mechanics in which motion of a particle can be represented by a wave. In this sense, the Hamilton-Jacobi equation fulfilled a long-held goal of theoretical physics, that dates back to Johann Bernoulli, of finding an analogy between the propagation of light and the motion of a particle. This goal motivated Hamilton to develop Hamiltonian mechanics. A consequence of this wave-particle analogy is that the Hamilton-Jacobi formalism featured prominently in the derivation of the Schrödinger equation during the development of quantum-wave mechanics.

15.5: Action-angle Variables

Canonical transformation

Systems possessing periodic solutions are a ubiquitous feature in physics. The periodic motion can be either an oscillation, for which the trajectory in phase space is a closed loop (libration), or rolling (rotational) motion as discussed in chapter 3.4. For many problems involving periodic motion, the interest often lies in the frequencies of motion rather than the detailed shape of the trajectories in phase space. The action-angle variable approach uses a canonical transformation to action and angle variables which provide a powerful, and elegant method to exploit Hamiltonian mechanics. In particular, it can determine the frequencies of periodic motion without having to calculate the exact trajectories for the motion. This method was introduced by the French astronomer Ch. E. Delaunay(1816 − 1872) for applications to orbits in celestial mechanics, but it has equally important applications beyond celestial mechanics such as to bound solutions of the atom in quantum mechanics.

The action-angle method replaces the momenta in the Hamilton-Jacobi procedure by the action phase integral for the closed loop (libration) trajectory in phase space defined by

Jipidqi(15.116)J_i \equiv \oint p_idq_i \tag{15.116}

where for each cyclic variable the integral is taken over one complete period of oscillation. The cyclic variable IiI_i is called the action variable where

Ii12πJi=12πpidqi(15.117)I_i \equiv \frac{1}{ 2\pi} J_i = \frac{1}{ 2\pi} \oint p_idq_i \tag{15.117}

The canonical variable to the action variable I\mathbf{I} is the angle variable ϕ\boldsymbol{\phi}. Note that the name “action variable” is used to differentiate I\mathbf{I} from the action functional S=LdtS = \int Ldt which has the same units; i.e. angular momentum.

The general principle underlying the use of action-angle variables is illustrated by considering one body, of mass mm, subject to a one-dimensional bound conservative potential energy U(q)U(q). The Hamiltonian is given by

H(p,q)=p22m+U(q)(15.118)H(p,q) = \frac{p^2}{ 2m} + U(q) \tag{15.118}

This bound system has a (q,p)(q,p) phase space contour for each energy H=EH = E.

p(q,E)=±2m(EU(q))(15.119)p(q,E) = \pm \sqrt{2m(E − U(q))} \tag{15.119}

For an oscillatory system the two-valued momentum of Equation 15.119 is non-trivial to handle. By contrast, the area JpdqJ \equiv \oint pdq of the closed loop in phase space is a single-valued scalar quantity that depends on EE and U(q)U(q). Moreover, Liouville’s theorem states that the area of the closed contour in phase space JpdqJ \equiv \oint pdq is invariant to canonical transformations. These facts suggest the use of a new pair of conjugate variables, (ϕ,I)(\phi , I), where I(E)I(E) uniquely labels the trajectory, and corresponding area, of a closed loop in phase space for each value of EE, and the single-valued function ϕ\phi is a corresponding angle that specifies the exact point along the phase-space contour as illustrated in Fig 15.5.1.

For simplicity consider the linear harmonic oscillator where

U(q)=12mω2q2(15.120)U(q) = \frac{1}{ 2} m\omega^2q^2 \tag{15.120}

Then the Hamiltonian, 15.118 equals

H(p,q)=p22m+12mω2q2(15.121)H(p,q) = \frac{p^2 }{2m} + \frac{1}{ 2} m\omega^2q^2 \tag{15.121}

Hamilton’s equations of motion give that

p˙=Hq=mω2q(15.122)\dot{p} = −\frac{\partial H}{ \partial q} = −m\omega^2q \tag{15.122}
q˙=Hp=pm(15.123)\dot{q} = \frac{\partial H}{ \partial p} = \frac{p}{ m} \tag{15.123}

The solution of equations 15.122 and 15.123 is of the form

q=Ccos(ω(tt0))(15.124)q = C \cos(\omega (t − t_0)) \tag{15.124}
p=mωCsinω(tt0)(15.125)p = −m\omega C \sin \omega (t − t_0) \tag{15.125}

where CC, and t0t_0 are integration constants. For the harmonic oscillator, equations 15.124 and 15.125 correspond to the usual elliptical contours in phase space, as illustrated in Figure 15.5.1.

The potential energy V (q), (upper) and corresponding phase space (p,q) (middle) for the harmonic oscillator at four equally spaced total energies E. The corresponding action-angles (I \phi) resulting from a canonical transformation of this system are shown in the lower plot.

Figure 15.5.1:The potential energy V(q)V (q), (upper) and corresponding phase space (p,q)(p,q) (middle) for the harmonic oscillator at four equally spaced total energies EE. The corresponding action-angles (Iϕ)(I \phi) resulting from a canonical transformation of this system are shown in the lower plot.

The action-angle canonical transformation involves making the transform

(q,p)(ϕ,I)(15.126)(q,p) \rightarrow (\phi , I) \tag{15.126}

where II is defined by Equation 15.117 and the angle ϕ\phi being the corresponding canonical angle. The logical approach to this canonical transformation for the harmonic oscillator is to define qq and pp in terms of ϕ\phi and II

q=2Imωcosϕ(15.127)q = \sqrt{\frac{ 2I}{ m\omega}} \cos \phi \tag{15.127}
p=2mIωsinϕ(15.128)p = \sqrt{2mI\omega } \sin \phi \tag{15.128}

Note that the Poisson bracket is unity

[q,p](ϕ,I)=1[q, p]_{(\phi , I)} = 1 \nonumber

which implies that the above transformation is canonical, and thus the phase space area I(E)12πpdqI(E) \equiv \frac{1}{ 2\pi} \oint pdq is conserved.

For this canonical transformation the transformed Hamiltonian H(ϕ,I)\mathcal{H} (\phi , I) is

H(ϕ,I)=12m(2mωI)sin2ϕ+12mω22Imωcos2ϕ=ωI(15.129)\mathcal{H} (\phi , I) = \frac{1}{ 2m } (2m\omega I) \sin^2 \phi + \frac{1}{ 2 }m\omega^2 \frac{2I}{ m\omega } \cos^2 \phi = \omega I \tag{15.129}

Note that this Hamiltonian is a constant that is independent of the angle ϕ\phi, and thus Hamilton’s equations of motion give

I˙=H(ϕ,I)ϕ=0(15.130)\dot{ I} = −\frac{\partial \mathcal{H} (\phi , I)}{ \partial \phi} = 0 \tag{15.130}
ϕ˙=H(ϕ,I)I=ω(15.131)\dot{\phi} = \frac{\partial \mathcal{H} (\phi , I) }{\partial I} = \omega \tag{15.131}

Thus we have mapped the harmonic oscillator to new coordinates (ϕ,I)(\phi , I) where

I=H(ϕ,I)ω=Eω(15.132)I = \frac{\mathcal{H} (\phi , I)}{ \omega} = \frac{E }{\omega} \tag{15.132}
ϕ=ω(tt0)(15.133)\phi = \omega (t − t_0) \tag{15.133}

That is, the phase space has been mapped from ellipses, with area proportional to EE in the (q,p)(q,p) phase space, to a cylindrical (ϕ,I)(\phi , I) phase space where I=EωI = \frac{E}{\omega} are constant values that are independent of the angle, while ϕ\phi increases linearly with time. Thus the variables (q,p)(q,p) are periodic with modulus Δϕ=2π\Delta\phi = 2\pi.

q(ϕ+2π,I)=q(ϕ,I)(15.134)q(\phi + 2\pi, I) = q (\phi , I) \tag{15.134}
p(ϕ+2π,I)=p(ϕ,I)(15.135)p(\phi + 2\pi, I) = p (\phi , I) \tag{15.135}

The period τ\tau of the periodic oscillatory motion is given simply by Δϕ=2π=ωτ\Delta\phi = 2\pi = \omega \tau which is the well known result for the harmonic oscillator. Note that the action-angle variable canonical transformation has determined the frequency of the periodic motion without solving the detailed trajectory of the motion.

The above example of the harmonic oscillator has shown that, for integrable periodic systems, it is possible to identify a canonical transformation to (ϕ,I)(\phi , I) such that the Hamiltonian is independent of the angle ϕ\phi which specifies the instantaneous location on the constant energy contour II. If the phase space contour is a separatrix, then it divides phase space into invariant regions containing phase-space contours with differing behavior. The action-angle variables are not useful for separatrix contours. For rolling motion, the system rotates with continuously increasing, or decreasing angle, and there is no natural boundary for the action angle variable since the phase space trajectory is continuous and not closed. However, the action-angle approach still is valid if the motion involves periodic as well as rolling motion.

The example of the one-dimensional, one-body, harmonic oscillator can be expanded to the more general case for many bodies in three dimensions. This is illustrated by considering multiple periodic systems for which the Hamiltonian is conservative and where the equations of the canonical transformation are separable. The generalized momenta then can be written as

pi=Wi(qi;α1,α2,..αn)qi(15.136)p_i = \frac{\partial W_i (q_i; \alpha_1, \alpha_2, ..\alpha_n)}{ \partial q_i} \tag{15.136}

for which each pip_i is a function of qiq_i and the nn integration constants αj\alpha_j

pi=pi(qi,α1,α2,..αn)(15.137)p_i = p_i (q_i, \alpha_1, \alpha_2, ..\alpha_n) \tag{15.137}

The momentum pi(qi,α1,α2,..αn)p_i (q_i, \alpha_1, \alpha_2, ..\alpha_n) represents the trajectory of the system in the (qi,pi)(q_i, p_i) phase space that is characterized by Hamilton’s characteristic function W(q,J)W(q,J). Combining equations 15.116, 15.136 gives

JiWi(qi;α1,α2,..αn)qidqi(15.138)J_i \equiv \oint \frac{\partial W_i (q_i; \alpha_1, \alpha_2, ..\alpha_n)}{ \partial q_i} dq_i \tag{15.138}

Since qiq_i is merely a variable of integration, each active action variable JiJ_i is a function of the nn constants of integration in the Hamilton-Jacobi equation. Because of the independence of the separable-variable pairs (qi,pi)(q_i, p_i), the JiJ_i form nn independent functions of the αi\alpha_i, and hence are suitable for use as a new set of constant momenta. Thus the characteristic function WW can be written as

W(q1,...qn;J1,...Jn)=jWj(qj;J1,...Jn)(15.139)W (q_1, ...q_n; J_1, ...J_n) = \sum_j W_j (q_j ; J_1, ...J_n) \tag{15.139}

while the Hamiltonian is only a function of the momenta H(J1,....Jn)H (J_1, .... J_n)

The generalized coordinate, conjugate to JJ, is known as the angle variable ϕi\phi_i which is defined by the transformation equation

ϕi=WJi=j=1nWj(qj;J1,...Jn)Ji(15.140)\phi_i = \frac{\partial W }{\partial J_i} = \sum^n_{j=1} \frac{\partial W_j (q_j ; J_1, ...J_n)}{ \partial J_i} \tag{15.140}

The corresponding equation of motion for ϕ\phi is given by

ϕ˙i=H(J)Ji=2πωi(J1,...Jn)(15.141)\dot{\phi}_i = \frac{\partial H(J) }{\partial J_i} = 2\pi\omega_i(J_1, ...J_n) \tag{15.141}

where ωi(J)\omega_i(J) are constant functions of the action variables JjJ_j with a solution

ϕi=2πωit+βi(15.142)\phi_i = 2\pi\omega_it + \beta_i \tag{15.142}

that is, they are linear functions of time. The constants ωi\omega_i can be identified with the frequencies of the multiple periodic motions.

The action-angle variables appear to be no different than a particular set of transformed coordinates. Their merit appears when the physical interpretation is assigned to ωi\omega_i. Consider the change δϕi\delta \phi_i as the qjq_j are changed infinitesimally

δϕi=jϕiqjqj=j2WJiqjqj(15.143)\delta \phi_i = \sum_j \frac{\partial \phi_i}{ \partial q_j} \partial q_j = \sum_j \frac{\partial^2 W }{\partial J_i\partial q_j } \partial q_j \tag{15.143}

The derivative with respect to qiq_i vanishes except for the WjW_j component of WW. Thus Equation 15.143 reduces to

δϕi=Jijpj(qj,J)dqj(15.144)\delta \phi_i = \frac{\partial}{ \partial J_i} \sum_j p_j (q_j , J) dq_j \tag{15.144}

Therefore, the total change in ϕ\phi, as the system goes through one complete cycle is

Δϕi=jJipj(qj,J)dqj=2πδij(15.145)\Delta\phi_i = \sum_j \frac{\partial}{ \partial J_i} \oint p_j (q_j , J) dq_j = 2\pi\delta_{ij} \tag{15.145}

where Ji\frac{\partial }{ \partial J_i} is outside the integral since the JiJ_i are constants for cyclic motion. Thus Δϕi=2π=ωiτi\Delta\phi_i = 2\pi = \omega_i\tau_i where τi\tau_i is the period for one cycle of oscillation, where the angular frequency ωi\omega_i is given by

ωi2π=νi=1τi(15.146)\frac{\omega_i}{ 2\pi} = \nu_i = \frac{1}{ \tau_i} \tag{15.146}

Thus the frequency ν\nu associated with the periodic motion is the reciprocal of the period τ\tau. The secret here is that the derivative of HH with respect to the action variable JJ given by Equation 15.141 directly determines the frequency of the periodic motion without the need to solve the complete equations of motion. Note that multiple periodic motion can be represented by a Fourier expansion of the form

qk=j1=j2=...jn=aj1,..,jnke2πi(j1ω1+j2ω2+j3ω3+..+jnωn)(15.147)q_k = \sum^{\infty}_{j_1=−\infty} \sum^{\infty}_{j_2=−\infty} ... \sum^{\infty}_{j_n=−\infty} a^k_{j_1,..,j_n} e^{2\pi i(j_1\omega_1+ j_2\omega_2+ j_3\omega_3+..+ j_n\omega_n)} \tag{15.147}

Although the action-angle approach to Hamilton-Jacobi theory does not produce complete equations of motion, it does provide the frequency decomposition that often is the physics of interest. The reason that the powerful action-angle variable approach has been introduced here is that it is used extensively in celestial mechanics. The action-angle concept also played a key role in the development of quantum mechanics, in that Sommerfeld recognized that Bohr’s ad hoc assumption that angular momentum is quantized, could be expressed in terms of quantization of the angle variable as is mentioned in chapter 18.

Adiabatic invariance of the action variables

When the Hamiltonian depends on time it can be quite difficult to solve for the motion because it is difficult to find constants of motion for time-dependent systems. However, if the time dependence is sufficiently slow, that is, if the motion is adiabatic, then there exist dynamical variables that are almost constant which can be used to solve for the motion. In particular, such approximate constants are the familiar action-angle integrals. The adiabatic invariance of the action variables played an important role in the development of quantum mechanics during the 1911 Solvay Conference. This was a time when physicists were grappling with the concepts of quantum mechanics. Einstein used the following classical mechanics example of adiabatic invariance, applied to the simple pendulum, in order to illustrate the concept of adiabatic invariance of the action. This example demonstrates the power of using action-angle variables.

15.6: Canonical Perturbation Theory

Most examples in classical mechanics discussed so far have been capable of exact solutions. In real life, the majority of problems cannot be solved exactly. For example, in celestial mechanics the two-body Kepler problem can be solved exactly, but solution of the three-body problem is intractable. Typical systems in celestial mechanics are never as simple as the two-body Kepler system because of the influence of additional bodies. Fortunately in most cases the influence of additional bodies is sufficiently small to allow use of perturbation theory. That is, the restricted three-body approximation can be employed for which the system is reduced to considering it as an exactly solvable two-body problem, subject to a small perturbation to this solvable two-body system. Note that even though the change in the Hamiltonian due to the perturbing term may be small, the impact on the motion can be especially large near a resonance.

Consider the Hamiltonian, subject to a time-dependent perturbation, is written as

H(q,p,t)=H0(q,p,t)+ΔH(q,p,t)H(q, p, t) = H_0(q, p, t) + \Delta H(q, p, t) \nonumber

where H0(q,p,t)H_0(q, p, t) designates the unperturbed Hamiltonian and ΔH(q,p,t)\Delta H(q, p, t) designates the perturbing term. For the unperturbed system the Hamilton-Jacobi equation is given by

H(Qi,Pi,t)=H0(q1,...qn;Sq1...,Sqn;t)+St=0(15.90)\mathcal{H}(Q_i, P_i, t) = H_0(q_1, ...q_n; \frac{\partial S}{\partial q_1} ..., \frac{\partial S}{\partial q_n };t) + \frac{\partial S}{\partial t} = 0 \tag{15.90}

where S(qi,Pi,t)S(q_i, P_i, t) is the generating function for the canonical transformation (q,p)(Q,P)(q, p) \rightarrow (Q, P). The perturbed S(qi,Pi,t)S(q_i, P_i, t) remains a canonical transformation, but the transformed Hamiltonian H(Qi,Pi,t)0\mathcal{H}(Q_i, P_i, t) \neq 0. That is,

H(Qi,Pi,t)=H0+ΔH(q,p,t)+St=ΔH(q,p,t)(15.148)\mathcal{H}(Q_i, P_i, t) = H_0 + \Delta H(q, p, t) + \frac{\partial S}{\partial t} = \Delta H(q, p, t) \tag{15.148}

The equations of motion satisfied by the transformed variables now are

Q˙i=ΔHPiP˙i=ΔHQi(15.149)\dot{Q}_i = \frac{\partial \Delta H}{ \partial P_i} \tag{15.149} \\ \dot{P}_i = \frac{\partial \Delta H }{\partial Q_i}

These equations remain as difficult to solve as the full Hamiltonian. However, the perturbation technique assumes that ΔH\Delta H is small, and that one can neglect the change of (Qi,Pi)(Q_i, P_i) over the perturbing interval. Therefore, to a first approximation, the unperturbed values of ΔHPi\frac{\partial \Delta H }{\partial P_i} and ΔHQi\frac{\partial \Delta H}{ \partial Q_i } can be used in equations 15.149. A detailed explanation of canonical perturbation theory is presented in chapter 12 of Goldstein[Go50].

15.7: Symplectic Representation

The Hamilton’s first-order equations of motion are symmetric if the generalized and constraint force terms, in equation (15.1.9)(15.1.9), are excluded.

q˙=Hpp˙=Hq\mathbf{\dot{q}} = \frac{\partial H}{ \partial \mathbf{p}} \quad − \mathbf{\dot{p}} = \frac{\partial H}{ \partial \mathbf{q}} \nonumber

This stimulated attempts to treat the canonical variables (q,p)(\mathbf{q}, \mathbf{p}) in a symmetric form using group theory. Some graduate textbooks in classical mechanics have adopted use of symplectic symmetry in order to unify the presentation of Hamiltonian mechanics. For a system of nn degrees of freedom, a column matrix η\boldsymbol{\eta} is constructed that has 2n2n elements where

ηj=qjηn+j=pjjn(15.150)\eta_j = q_j \quad \eta_{n+j} = p_j \quad j \leq n \tag{15.150}

Therefore the column matrix

(Hη)j=Hqj(Hη)n+j=Hpjjn(15.151)\left(\frac{\partial H}{ \partial \boldsymbol{\eta}} \right)_j = \frac{\partial H }{\partial q_j} \quad \left(\frac{\partial H}{ \partial \boldsymbol{\eta}} \right)_{n+j} = \frac{\partial H }{\partial p_j} \quad j \leq n \tag{15.151}

The symplectic matrix J\mathbf{J} is defined as being a 2n2n by 2n2n skew-symmetric, orthogonal matrix that is broken into four n×nn \times n null or unit matrices according to the scheme

J=([0]+[1][1][0])(15.152)\mathbf{J} = \begin{pmatrix} [\mathbf{0}] & +[\mathbf{1}] \\ − [\mathbf{1}] & [\mathbf{0}] \end{pmatrix} \tag{15.152}

where [0][\mathbf{0}] is the nn-dimension null matrix, for which all elements are zero. Also [1][\mathbf{1}] is the nn-dimensional unit matrix, for which the diagonal matrix elements are unity and all off-diagonal matrix elements are zero. The J\mathbf{J} matrix accounts for the opposite signs used in the equations for q˙\mathbf{\dot{q}} and p˙\mathbf{\dot{p}}. The symplectic representation allows the Hamilton’s equations of motion to be written in the compact form

η˙=JHη(15.153)\boldsymbol{\dot{\eta}} = \mathbf{J}\frac{\partial H }{\partial \boldsymbol{\eta}} \tag{15.153}

This textbook does not use the elegant symplectic representation since this representation ignores the important generalized forces and Lagrange multiplier forces.

15.8: Comparison of the Lagrangian and Hamiltonian Formulations

Common features

The discussion of Lagrangian and Hamiltonian dynamics has illustrated the power of such algebraic formulations. Both approaches are based on application of variational principles to scalar energy which gives the freedom to concentrate solely on active forces and to ignore internal forces. Both methods can handle manybody systems and exploit canonical transformations, which are impractical or impossible using the vectorial Newtonian mechanics. These algebraic approaches simplify the calculation of the motion for constrained systems by representing the vector force fields, as well as the corresponding equations of motion, in terms of either the Lagrangian function L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) or the action functional S(q,p,t)S(\mathbf{q},\mathbf{p},t) which are related by the definite integral

S(q,p,t)=t1t2L(q,q˙,t)dt(15.1)S(\mathbf{q},\mathbf{p},t) = \int^{t_2}_{t_1} L(\mathbf{q}, \mathbf{\dot{q}}, t)dt \tag{15.1}

The Lagrangian function L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t), and the action functional S(q,p,t)S(\mathbf{q},\mathbf{p},t), are scalar functions under rotation, but they determine the vector force fields and the corresponding equations of motion. Thus the use of rotationally-invariant functions L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) and S(q,p,t)S(\mathbf{q},\mathbf{p},t) provide a simple representation of the vector force fields. This is analogous to the use of scalar potential fields ϕ(q,t)\phi (\mathbf{q}, t) to represent the electrostatic and gravitational vector force fields. Like scalar potential fields, Lagrangian and Hamiltonian mechanics represents the observables as derivatives of L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) and S(q,p,t)S(\mathbf{q},\mathbf{p},t), and the absolute values of L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) and S(q,p,t)S(\mathbf{q},\mathbf{p},t) are undefined; only differences in L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}}, t) and S(q,p,t)S(\mathbf{q},\mathbf{p},t) are observable. For example, the generalized momenta are given by the derivatives piLq˙ip_i \equiv \frac{\partial L}{ \partial \dot{q}_i} and pj=Sqjp_j = \frac{\partial S }{\partial q_j }. The physical significance of the least action S(q,α,t)S(\mathbf{q}, \boldsymbol{\alpha},t) is illustrated when the canonically transformed momenta P=α\mathbf{P} = \boldsymbol{\alpha} is a constant. Then the generalized momenta and the Hamilton-Jacobi equation, imply that the total time derivative of the action equals

dSdt=Sqiq˙i+St=piqiH=L(15.154)\frac{dS}{ dt} = \frac{\partial S}{ \partial q_i} \dot{q}_i + \frac{\partial S}{ \partial t} = p_iq_i − H = L \tag{15.154}

The indefinite integral of this equation reproduces the definite integral 15.1 to within an arbitrary constant, i.e.

S(q,p)=L(q,q˙,t)dt+ constant(15.155)S(\mathbf{q}, \mathbf{p}) = \int L(\mathbf{q}, \mathbf{\dot{q}}, t)dt + \text{ constant} \tag{15.155}

Lagrangian Formulation

Consider a system with nn independent generalized coordinates, plus mm constraint forces that are not required to be known. The Lagrangian approach can reduce the system to a minimal system of s=nms = n − m independent generalized coordinates leading to s=nms = n - m second-order differential equations. By comparison, the Newtonian approach uses n+mn + m unknowns. Alternatively, the Lagrange multipliers approach allows determination of the holonomic constraint forces resulting in s=n+ms = n + m second order equations to determine s=n+ms = n + m unknowns. The Lagrangian potential function is limited to conservative forces, but generalized forces can be used to handle non-conservative and non-holonomic forces. The advantage of the Lagrange equations of motion is that they can deal with any type of force, conservative or non-conservative, and they directly determine q,q˙q, \dot{q} rather than q,pq,p which then requires relating pp to q˙\dot{q}. The Lagrange approach is superior to the Hamiltonian approach if a numerical solution is required for typical undergraduate problems in classical mechanics. However, Hamiltonian mechanics has a clear advantage for addressing more profound and philosophical questions in physics.

Hamiltonian Formulation

For a system with nn independent generalized coordinates, and mm constraint forces, the Hamiltonian approach determines 2n2n first-order differential equations. In contrast to Lagrangian mechanics, where the Lagrangian is a function of the coordinates and their velocities, the Hamiltonian uses the variables q\mathbf{q} and p\mathbf{p}, rather than velocity. The Hamiltonian has twice as many independent variables as the Lagrangian which is a great advantage, not a disadvantage, since it broadens the realm of possible transformations that can be used to simplify the solutions. Hamiltonian mechanics uses the conjugate coordinates q,p\mathbf{q},\mathbf{p}, corresponding to phase space. This is an advantage in most branches of physics and engineering. Compared to Lagrangian mechanics, Hamiltonian mechanics has a significantly broader arsenal of powerful techniques that can be exploited to obtain an analytical solution of the integrals of the motion for complicated systems. These techniques include, the Poisson bracket formulation, canonical transformations, the Hamilton-Jacobi approach, the action-angle variables, and canonical perturbation theory. In addition, Hamiltonian dynamics provides a means of determining the unknown variables for which the solution assumes a soluble form, and it is ideal for study of the fundamental underlying physics in applications to other fields such as quantum or statistical physics. However, the Hamiltonian approach endemically assumes that the system is conservative putting it at a disadvantage with respect to the Lagrangian approach. The appealing symmetry of the Hamiltonian equations, plus their ability to utilize canonical transformations, makes it the formalism of choice for examination of system dynamics. For example, Hamilton-Jacobi theory, action-angle variables and canonical perturbation theory are used extensively to solve complicated multibody orbit perturbations in celestial mechanics by finding a canonical transformation that transforms the perturbed Hamiltonian to a solved unperturbed Hamiltonian.

The Hamiltonian formalism features prominently in quantum mechanics since there are well established rules for transforming the classical coordinates and momenta into linear operators used in quantum mechanics. The variables q,q˙\mathbf{q}, \mathbf{\dot{q}} used in Lagrangian mechanics do not have simple analogs in quantum physics. As a consequence, the Poisson bracket formulation, and action-angle variables of Hamiltonian mechanics played a key role in development of matrix mechanics by Heisenberg, Born, and Dirac, while the Hamilton-Jacobi formulation played a key role in development of Schrödinger’s wave mechanics. Similarly, Hamiltonian mechanics is the preeminent variational approached used in statistical mechanics.

15.E: Advanced Hamiltonian Mechanics (Exercises)

  1. Poisson brackets are a powerful means of elucidating when observables are constant of motion and whether two observables can be simultaneously measured with unlimited precision. Consider a spherically symmetric Hamiltonian

    H=12m(pr2+pθ2r2+pϕ2r2sin2θ)+U(r)H = \frac{1}{2m} \left( p^2_r + \frac{p^{2}_{\theta}}{r^2} + \frac{p^2_{\phi}}{r^2 \sin^2 \theta} \right) + U(r) \nonumber

    for a mass mm where U(rU(r is a central potential. Use the Poisson bracket plus the time dependence to determine the following:

  2. Does pϕp_{\phi} commute with HH and is it a constant of motion?

  3. Does pθ2+pϕ2sin2θp^2_{\theta} + \frac{p^2_{\phi}}{ \sin^2 \theta } commute with HH and is it a constant of motion?

  4. Does prp_r commute with HH and is it a constant of motion?

  5. Does pϕp_{\phi} commute with pθp_{\theta} and what does the result imply?

  6. Consider the Poisson brackets for angular momentum LL

  7. Show {Li,rj}=ϵijkrk\{L_i, r_j \} = \epsilon_{ijk}r_k, where the Levi-Cevita tensor is,

    ϵijk={+1if ijk are cyclically permuted1if ijk are anti-cyclically permuted0if i=j or i=k or j=k\epsilon_{ijk} = \begin{cases} +1 & \mbox{if } ijk \mbox{ are cyclically permuted}\\ −1 & \mbox{if } ijk \mbox{ are anti-cyclically permuted} \\ 0 & \mbox{if } i = j \mbox{ or } i = k \mbox{ or } j = k \end{cases} \nonumber
  8. Show {Li,pj}=ϵijkpk\{L_i, p_j \} = \epsilon_{ijk}p_{k}.

  9. Show {Li,Lj}=ϵijkLk\{L_i, L_j \} = \epsilon_{ijk}L_k. The following identity may be useful: ϵijkϵilm=δjlδkmδjmδkl\epsilon_{ijk}\epsilon_{ilm} = \delta_{jl}\delta_{km} − \delta_{jm}\delta_{kl }.

  10. Show {Li,L2}=0\{L_i, L^2 \} = 0.

  11. Consider the Hamiltonian of a two-dimensional harmonic oscillator,

    H=p22m+12m(ω12r12+ω22r22)H = \frac{\mathbf{p}^2 }{2m} + \frac{1 }{2 }m ( \omega^2_1r^2_1 + \omega^2_2r^2_2 ) \nonumber

    What condition is satisfied if L2L^2 a conserved quantity?

  12. Consider the motion of a particle of mass mm in an isotropic harmonic oscillator potential U=12kr2U = \frac{1}{ 2} kr^2 and take the orbital plane to be the xyx − y plane. The Hamiltonian is then

    HS0=12m(px2+py2)+12k(x2+y2)H \equiv S_0 = \frac{1}{2m}(p^2_x + p^2_y) +\frac{1}{2}k(x^2 + y^2) \nonumber

Introduce the three quantities

S1=12m(px2py2)+12k(x2y2)S_1 = \frac{1}{2m}(p^2_x − p^2_y) +\frac{1}{2}k(x^2 − y^2) \nonumber
S2=1mpxpy+kxyS_2 = \frac{1}{ m} p_{x}p_{y} + kxy \nonumber
S3=ω(xpyypx)S_3 = \omega (xp_{y} − yp_{x}) \nonumber

with ω=km\omega = \sqrt{\frac{k}{m}}. Use Poisson brackets to solve the following:

  1. Show that {S0,Si}=0\{S_0, S_i\}=0 for i=1,2,3i = 1, 2, 3 proving that (S1,S2,S3)(S_1, S_2, S_3) are constants of motion.

  2. Show that

    {S1,S2}=2ωS3\{S_1, S_2\}=2\omega S_3 \nonumber
{S2,S3}=2ωS1\{S_2, S_3\}=2\omega S_1 \nonumber
{S3,S1}=2ωS2\{S_3, S_1\}=2\omega S_2 \nonumber

so that (2ω)1(S1,S2,S3)(2\omega )^{ −1 } (S_1, S_2, S_3) have the same Poisson bracket relations as the components of a 3-dimensional angular momentum.

S02=S12+S22+S32S^2_0 = S^2_1 + S^2_2 + S^2_3 \nonumber
  1. Assume that the transformation equations between the two sets of coordinates (q,p)(q, p) and (Q,P)(Q, P) are

Q=ln(1+q12cosp)Q = \ln (1 + q^{\frac{1}{2}} \cos p) \nonumber
P=2(1+q12cosp)q12sinp)P = 2(1 + q^{\frac{1}{2}} \cos p)q^{\frac{1}{2}} \sin p) \nonumber
  1. Assuming that q,pq, p are canonical variables, i.e. [q,p]=1[q, p]=1, show directly from the above transformation equations that Q,PQ, P are canonical variables.

  2. Show that the generating function that generates this transformation between the two sets of canonical variables is

    F3=[eQ1]2tanpF_3 = −[e^Q − 1]^2 \tan p \nonumber
  3. Consider a bound two-body system comprising a mass mm in an orbit at a distance rr from a mass MM. The attractive central force binding the two-body system is

F=kr2r^\mathbf{F} = \frac{k}{r^2}\mathbf{\hat{r}} \nonumber

where kk is negative. Use Poisson brackets to prove that the eccentricity vector A=p×L+μkr^A = p\times L+\mu k\hat{r} is a conserved quantity.

  1. Consider the case of a single mass m where the Hamiltonian H=12p2H =\frac{1}{2}p^2.

  2. Use the generating function S(q,P,t)S(q, P, t) to solve the Hamilton-Jacobi equation with the canonical transformation q=q(Q,P)q = q(Q, P) and p=p(Q,P)p = p(Q, P) and determine the equations relating the (q,p)(q, p) variables to the transformed coordinate and momentum (Q,P)(Q, P).

  3. If there is a perturbing Hamiltonian ΔH=12q2\Delta H =\frac{1}{2}q^2, then PP will not be constant. Express the transformed Hamiltonian HH (using the transformation given above in terms of PP, QQ, and tt). Solve for Q(t)Q(t) and P(t)P(t) and show that the perturbed solution q[Q(t),P(t)]q[Q(t), P(t)], p[Q(t),P(t)]p[Q(t), P(t)] is simple harmonic.

15.S: Advanced Hamiltonian mechanics (Summary)

This chapter has gone beyond what is normally covered in an undergraduate course in classical mechanics, in order to illustrate the power of the remarkable arsenal of methods available for solution of the equations of motion using Hamiltonian mechanics. This has included the Poisson bracket representation of Hamiltonian formulation of mechanics, canonical transformations, Hamilton-Jacobi theory, action-angle variables, and canonical perturbation theory. The purpose was to illustrate the power of variational principles in Hamiltonian mechanics and how they relate to fields such as quantum mechanics and astronomy. The following are the key points made in this chapter.

Poisson brackets:

The elegant and powerful Poisson bracket formalism of Hamiltonian mechanics was introduced. The Poisson bracket of any two continuous functions of generalized coordinates F(p,q)F(p,q) and G(p,q)G(p,q), is defined to be

{F,G}pqi(FqiGpiFpiGqi)\{F, G\}_{pq} \equiv \sum_i \left( \frac{\partial F}{\partial q_i} \frac{\partial G}{\partial p_i} − \frac{\partial F}{\partial p_i} \frac{\partial G}{\partial q_i}\right)

The fundamental Poisson brackets equal

{qk,ql}=0\{q_k, q_l\}=0
{pk,pl}=0\{p_k, p_l\}=0
{qk,pl}={pl,qk}=δkl\{q_k, p_l\} = − \{p_l, q_k\} = \delta_{kl}

The Poisson bracket is invariant to a canonical transformation from (q,p)(q, p) to (Q,P)(Q, P). That is

{F,G}qp=k(FQkGPkFPkGQk)={F,G}QP\{F, G\}_{qp} = \sum_k \left( \frac{\partial F}{\partial Q_k} \frac{\partial G}{\partial P_k } − \frac{\partial F}{\partial P_k }\frac{\partial G}{\partial Q_k} \right) = \{F, G\}_{QP}

There is a one-to-one correspondence between the commutator and Poisson Bracket of two independent functions,

(F1G1G1F1)=λ{F1,G1}(F_1G_1 − G_1F_1) = \lambda \{F_1, G_1\}

where λ\lambda is an independent constant. In particular F1G1F_1G_1 commute of the Poisson Bracket {F1,G1}=0\{F_1, G_1\}=0.

Poisson Bracket representation of Hamiltonian mechanics:

It has been shown that the Poisson bracket formalism contains the Hamiltonian equations of motion and is invariant to canonical transformations. Also this formalism extends Hamilton’s canonical equations to non-commuting canonical variables. Hamilton’s equations of motion can be expressed directly in terms of the Poisson brackets

q˙k={qk,H}=Hpk\dot{q}_k = \{q_k, H\} = \frac{\partial H }{\partial p_k}
p˙k={pk,H}=Hqk\dot{p}_k = \{p_k, H\} = −\frac{\partial H}{ \partial q_k }

An important result is that the total time derivative of any operator is given by

dGdt=Gt+{G,H}\frac{dG}{dt} = \frac{\partial G}{\partial t} + \{G, H\}

Poisson brackets provide a powerful means of determining which observables are time independent and whether different observables can be measured simultaneously with unlimited precision. It was shown that the Poisson bracket is invariant to canonical transformations, which is a valuable feature for Hamiltonian mechanics. Poisson brackets were used to prove Liouville’s theorem which plays an important role in the use of Hamiltonian phase space in statistical mechanics. The Poisson bracket is equally applicable to continuous solutions in classical mechanics as well as discrete solutions in quantized systems.

Canonical transformations:

A transformation between a canonical set of variables (q,p)(q,p) with Hamiltonian H(q,p,t)H(q,p, t) to another set of canonical variable (Q,P)(Q,P) with Hamiltonian H(Q,P,t)\mathcal{H}(Q,P, t) can be achieved using a generating functions FF such that

H(Q,P,t)=H(q,p,t)+Ft\mathcal{H}(Q,P, t) = H(q,p, t) + \frac{\partial F}{\partial t}

Possible generating functions are summarized in the following table.

Generating functionGenerating function derivativesTrivial special case
F=F1(q,Q,t)F = F_1 (\mathbf{q}, \mathbf{Q}, t)pi=F1qiPi=F1Qip_i = \frac{\partial F_1}{\partial q_i} \quad P_i = -\frac{\partial F_1}{\partial Q_i}F1=qiQiQi=piPi=qiF_1 = q_iQ_i \quad Q_i = p_i \quad P_i = -q_i
F=F2(q,P,t)QPF = F_2 (\mathbf{q}, \mathbf{P}, t) - \mathbf{Q} \cdot \mathbf{P}pi=F2qiQi=F2Pip_i = \frac{\partial F_2}{\partial q_i} \quad Q_i = \frac{\partial F_2}{\partial P_i}F2=qiPiQi=qiPi=piF_2 = q_iP_i \quad Q_i = q_i \quad P_i = p_i
F=F3(p,Q,t)+qpF = F_3 (\mathbf{p}, \mathbf{Q}, t) + \mathbf{q} \cdot \mathbf{p}qi=F3piPi=F3Qiq_i = -\frac{\partial F_3}{\partial p_i} \quad P_i = -\frac{\partial F_3}{\partial Q_i}F3=piQiQi=qiPi=piF_3 = p_iQ_i \quad Q_i = -q_i \quad P_i = -p_i
F=F4(p,P,t)+qpQPF = F_4 (\mathbf{p}, \mathbf{P}, t) + \mathbf{q} \cdot \mathbf{p} - \mathbf{Q} \cdot \mathbf{P}qi=F4piQi=F4Piq_i = -\frac{\partial F_4}{\partial p_i} \quad Q_i = \frac{\partial F_4}{\partial P_i}F1=piPiQi=piPi=qiF_1 = p_iP_i \quad Q_i = p_i \quad P_i = -q_i

If the canonical transformation makes H(Q,P,t)=0\mathcal{H}(Q,P, t)=0 then the conjugate variables (Q,P)(Q,P) are constants of motion. Similarly if H(Q,P,t)\mathcal{H}(Q,P, t) is a cyclic function then the corresponding PP are constants of motion.

Hamilton-Jacobi theory:

Hamilton-Jacobi theory determines the generating function required to perform canonical transformations that leads to a powerful method for obtaining the equations of motion for a system. The Hamilton-Jacobi theory uses the action function SF2S \equiv F_2 as a generating function, and the canonical momentum is given by

pi=Sqip_i = \frac{\partial S}{ \partial q_i}

This can be used to replace pip_i in the Hamiltonian HH leading to the Hamilton-Jacobi equation

H(q;Sq;t)+St=0H(q; \frac{\partial S}{ \partial q} ;t) + \frac{\partial S}{ \partial t} = 0

Solutions of the Hamilton-Jacobi equation were obtained by separation of variables. The close optical-mechanical analogy of the Hamilton-Jacobi theory is an important advantage of this formalism that led to it playing a pivotal role in the development of wave mechanics by Schrödinger.

Action-angle variables:

The action-angle variables exploits a canonical transformation from (q,p)(ϕ,I)(q,p) \rightarrow (\phi , I) where

Ii12πJi=12πpidqiI_i \equiv \frac{1}{ 2\pi} J_i = \frac{1}{ 2\pi} \oint p_i dq_i

For periodic motion the phase-space trajectory is closed with area given by JJ and this area is conserved for the above canonical transformation. For a conserved Hamiltonian the action variable II is independent of the angle variable ϕ\phi. The time dependence of the angle variable ϕ\phi directly determines the frequency of the periodic motion without recourse to calculation of the detailed trajectory of the periodic motion.

Canonical perturbation theory:

Canonical perturbation theory is a valuable method of handling multibody interactions. The adiabatic invariance of the action-angle variables provides a powerful approach for exploiting canonical perturbation theory.

Comparison of Lagrangian and Hamiltonian formulations:

The remarkable power, and intellectual beauty, provided by use of variational principles to exploit the underlying principles of natural economy in nature, has had a long and rich history. It has led to profound developments in many branches of theoretical physics. However, it is noted that although the above algebraic formulations of classical mechanics have been used for over two centuries, the important limitations of these algebraic formulations to non-linear systems remain a challenge that still is being addressed.

It has been shown that the Lagrangian and Hamiltonian formulations represent the vector force fields, and the corresponding equations of motion, in terms of the Lagrangian function L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}},t), or the action functional S(q,p,t)S(\mathbf{q},\mathbf{p},t), which are scalars under rotation. The Lagrangian function L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}},t) is related to the action functional S(q,p,t)S(\mathbf{q},\mathbf{p},t) by

S(q,p,t)=t1t2L(q,q˙,t)dt(15.1)S(\mathbf{q},\mathbf{p},t) = \int^{t_2}_{t_1}L(\mathbf{q}, \mathbf{\dot{q}},t) dt\tag{15.1}

These functions are analogous to electric potential, in that the observables are derived by taking derivatives of the Lagrangian function L(q,q˙,t)L(\mathbf{q}, \mathbf{\dot{q}},t) or the action functional S(q,p,t)S(\mathbf{q},\mathbf{p},t). The Lagrangian formulation is more convenient for deriving the equations of motion for simple mechanical systems. The Hamiltonian formulation has a greater arsenal of techniques for solving complicated problems plus it uses the canonical variables (qi,pi)(q_i, p_i) which are the variables of choice for applications to quantum mechanics and statistical mechanics.