9. Hamilton’s Action Principle¶
9.1: Introduction to Hamilton’s Action Principle¶
Hamilton’s principle of stationary action was introduced in two papers published by Hamilton in 1834 and 1835. Hamilton’s Action Principle provides the foundation for building Lagrangian mechanics that had been pioneered 46 years earlier. Hamilton’s Principle now underlies theoretical physics and many other disciplines in mathematics and economics. In 1834 Hamilton was seeking a theory of optics when he developed both his principle of stationary action, plus the field of Hamiltonian mechanics.
Hamilton’s Action Principle is based on defining the action functional[1] for generalized coordinates which are expressed by the vector and their corresponding velocity vector .
The scalar action is a functional of the Lagrangian , integrated between an initial time and final time . In principle, higher order time derivatives of the generalized coordinates could be included, but most systems in classical mechanics are described adequately by including only the generalized coordinates, plus their velocities. The definition of the action functional allows for more general Lagrangians than the standard Lagrangian that has been used throughout chapters . Hamilton stated that the actual trajectory of a mechanical system is that given by requiring that the action functional is stationary with respect to change of the variables. The action functional is stationary when the variational principle can be written in terms of a virtual infinitessimal displacement, to be
Typically the stationary point corresponds to a minimum of the action functional. Applying variational calculus to the action functional leads to the same Lagrange equations of motion for systems as the equations derived using d’Alembert’s Principle, if the additional generalized force terms, , are omitted in the corresponding equations of motion.
These are used to derive the equations of motion, which then are solved for an assumed set of initial conditions. Prior to Hamilton’s Action Principle, Lagrange developed Lagrangian mechanics based on d’Alembert’s Principle in contrast to Newtonian equations of motion which are defined in terms of Newton’s Laws of Motion.
9.2: Hamilton’s Principle of Stationary Action¶
Hamilton’s crowning achievement was his use of the general form of Hamilton’s principle of stationary action , equation , to derive both Lagrangian mechanics, and Hamiltonian mechanics. Consider the action for the extremum path of a system in configuration space, that is, along path for coordinates at initial time to at a final time as shown in Figure 9.2.1.

Figure 9.2.1:Extremum path A, plus the neighboring path B, shown in configuration space.
Then the action is given by
As used in chapter a family of neighboring paths is defined by adding an infinitessimal fraction of a continuous, well-behaved neighboring function where for the extremum path. That is,
In contrast to the variational case discussed when deriving Lagrangian mechanics, the variational path used here does not assume that the functions vanish at the end points. Assume that the neighboring path has an action where
Expanding the integrand of in Equation 9.5 gives that, relative to the extremum path , the incremental change in action is
The second term in the integral can be integrated by parts since leading to
Note that Equation 9.7 includes contributions from the entire path of the integral as well as the variations at the ends of the curve and the terms. Equation 9.7 leads to the following two pioneering principles of least action in variational mechanics that were developed by Hamilton.
Stationary-action principle in Lagrangian mechanics¶
Derivation of Lagrangian mechanics in chapter 6 was based on the extremum path for neighboring paths between two given locations and that the system occupies at the initial and final times and respectively. For this special case, where the end points do not vary, that is, when , and , then the least action for the stationary path 9.8 reduces to
For independent generalized coordinates , the integrand in brackets vanishes leading to the Euler-Lagrange equations. Conversely, if the Euler-Lagrange equations in 9.8 are satisfied, then, that is, the path is stationary. This leads to the statement that the path in configuration space between two configurations and that the system occupies at times and respectively, is that for which the action is stationary. This is a statement of Hamilton’s Principle.
Stationary-action principle in Hamiltonian mechanics¶
Hamilton used the general variation of the least-action path to derive the basic equations of Hamiltonian mechanics. For the general path, the integral term in Equation 9.7 vanishes because the Euler-Lagrange equations are obeyed for the stationary path. Thus the only remaining non-zero contributions are due to the end point terms, which can be written by defining the total variation of each end point to be
where and are evaluated at and . Then Equation 9.7 reduces to
Since the generalized momentum , then Equation 9.10 can be expressed in terms of the Hamiltonian and generalized momentum as
Equation 9.11 contains Hamilton’s Principle of Least-action. Equation 9.12 gives an alternative relation of the generalized momentum that is expressed in terms of the action functional . Note that equations 9.11 and 9.12 were derived directly without invoking reference to the Lagrangian.
Integrating the action , Equation 9.10, between the end points gives the action for the path between and , that is, to be
The stationary path is obtained by using the variational principle
The integrand, in this modified Hamilton’s principle, can be used in the Euler-Lagrange equations for to give
Similarly, the other Euler-Lagrange equations give
Thus Hamilton’s principle of least-action leads to Hamilton’s equations of motion, that is equations 9.15 and 9.16.
The total time derivative of the action , which is a function of the coordinates and time, is
But the total time derivative of Equation 9.14 equals
Combining equations 9.17 and 9.18 gives the Hamilton-Jacobi equation which is discussed in chapter 15.4.
In summary, Hamilton’s principle of least action leads directly to Hamilton’s equations of motion 9.15, 9.16 plus the Hamilton-Jacobi Equation 9.19. Note that the above discussion has derived both Hamilton’s Principle 9.8, and Hamilton’s equations of motion 9.15, 9.16, directly from Hamilton’s variational concept of stationary action, , without explicitly invoking the Lagrangian.
Abbreviated action¶
Hamilton’s Action Principle determines completely the path of the motion and the position on the path as a function of time. If the Lagrangian and the Hamiltonian are time independent, that is, conservative, then and Equation 9.13 equals
The term in Equation 9.20, is called the abbreviated action which is defined as
The abbreviated action can be simplified assuming use of the standard Lagrangian with a velocity-independent potential , then equation 8.1.4 gives.
Abbreviated action provides for use of a simplified form of the principle of least action that is based on the kinetic energy, and not potential energy. For conservative systems it determines the path of the motion, but not the time dependence of the motion. Consider virtual motions where the path satisfies energy conservation, and where the end points are held fixed, that is but allow for a variation in the final time. Then using the Hamilton-Jacobi equation, 9.19
However, Equation 9.21 gives that
Therefore
That is, the abbreviated action has a minimum with respect to all paths that satisfy the conservation of energy which can be written as
Equation 9.26 is called the Maupertuis’ least-action principle which he proposed in 1744 based on Fermat’s Principle in optics. Credit for the formulation of least action commonly is given to Maupertuis; however, the Maupertuis principle is similar to the use of least action applied to the “vis viva”, as was proposed by Leibniz four decades earlier. Maupertuis used teleological arguments , rather than scientific rigor, because of his limited mathematical capabilities. In 1744 Euler provided a scientifically rigorous argument, presented above, that underlies the Maupertuis principle. Euler derived the correct variational relation for the abbreviated action to be
Hamilton’s use of the principle of least action to derive both Lagrangian and Hamiltonian mechanics is a remarkable accomplishment. It underlies both Lagrangian and Hamiltonian mechanics and confirmed the conjecture of Maupertuis.
Hamilton’s Principle applied using initial boundary conditions¶
Galley[Gal13] identified a subtle inconsistency in the applications of Hamilton’s Principle of Stationary Action to both Lagrangian and Hamiltonian mechanics. The inconsistency involves the fact that Hamilton’s Principle is defined as the action integral between the initial time and the final time as boundary conditions, that is, it is assumed to be time symmetric. However, most applications in Lagrangian and Hamiltonian mechanics assume that the action integral is evaluated based on the initial values as the boundary conditions, rather than the initial and final times . That is, typical applications require use of a time-asymmetric version of Hamilton’s principle. Galley proposed a framework for transforming Hamilton’s Principle to a time-asymmetric form in order to handle problems where the boundary conditions are based on using only the initial values at the initial time , rather than the initial plus final times that is assumed in the time-symmetric definition of the action in Hamilton’s Principle.

Figure 9.2.2:The left schematic shows paths between the initial and final times for conservative mechanics. The solid line designates the path for which the action is stationary, while the dashed lines represent the varied paths. The right schematic shows the paths applied to the doubled degrees of freedom with two initial boundary conditions, that is, and plus assuming that both paths are identical at their intersection and that they intersect at the same final time, that is, .
The following describes the framework proposed by Galley for transforming Hamilton’s Principle to a time-asymmetric form. Let and designate sets of generalized coordinates, plus their velocities, where and are the fundamental variables assumed in the definition of the Lagrangian used by Hamilton’s Principle. As illustrated schematically in Figure 9.2.2, Galley proposed doubling the number of degrees of freedom for the system considered, that is, let and . In addition he defines two identical variational paths 1 and where path 2 is the time reverse of path1. That is, path 1 starts at the initial time , and ends at , whereas path 2 starts at and ends at . That is, he assumes that and specify the two paths in the space of the doubled degrees of freedom that are identical, and that they intersect at the final time . The arrows shown on the paths in Figure 9.2.2 designate the assumed direction of the time integration along these paths.
For the doubled system of degrees of freedom, the total action for the sum of the two paths is given by the time integral of the doubled variables, which can be written as
The above relation assumes that the doubled variables and are decoupled from each other. More generally one can assume that the two sets of variables are coupled by some arbitrary function . Then the action can be written as
The effective Lagrangian for this doubled system then can be defined as
and the action can be written as
The coupling term for the doubled system of degrees of freedom must satisfy the following two properties.
(a) If it can be expressed as the difference of two scalar potentials, , then it can be absorbed into the potential term for each of the doubled variables in the Lagrangian. This implies that and there is no reason to double the number of degrees of freedom because the system is conservative. Thus describes generalized forces that are not derivable from potential energy, that is, conservative.
(b) A second property of the coupling term is that it must be antisymmetric under interchange of the arbitrary labels . That is,
Therefore the antisymmetric function vanishes when .
The variational condition requires that the action has a well defined stationary point for the doubled system. This is achieved by parametrizing both coordinate paths as
where are the coordinates for which the action is stationary, and where are arbitrary functions of time denoting virtual displacements of the paths. The doubled system has two independent paths connecting the two initial boundary conditions at , and it requires that these paths intersect at . The variational system for the two intersecting paths requires specifying four conditions, two per path. Two of the four conditions are determined by requiring that at the initial boundary conditions satisfies that . The remaining two conditions are derived by requiring that the variation of the action satisfies
The canonical momenta conjugate to the doubled coordinates are defined using the nonconservative Lagrangian to be
where the superscript designates the solution based on the initial conditions. Note that the conjugate momentum while the term is part of the total momentum due to the nonconservative interaction. Similarly the momentum for the second path is
The last term in Equation 9.34, that is, the term results from integration by parts, which will vanish if
The equality condition at the intersection of the two paths at requires that
Therefore equations 9.37 and 9.38 imply that
Therefore equations 9.38 and 9.39 constitute the equality condition that must be satisfied when the two paths intersect at . The equality condition ensures that the boundary term for integration by parts in Equation 9.34 will vanish for arbitrary variations provided that the two unspecified paths agree at the final time . Similarly the conjugate momenta must agree, but otherwise are unspecified. As a consequence, the equality condition ensures that the variational principle is consistent with the final state at not being specified. That is, the equations of motion are only specified by the initial boundary conditions of the time-asymmetric action for the doubled system.
More physics insight is provided by using a more convenient parametrization of the coordinates in terms of their average and difference. That is, let
Then the physical limit is
That is, the average history is the relevant physical history, while the difference coordinate simply vanishes. For these coordinates, the nonconservative Lagrangian is and the equality conditions reduce to
which implies that the physically relevant average quantities are not specified at the final time in order to have a well-defined variational principle.
The canonical momenta are given by
The equations of motion can be written as.
Equation 9.46 is identically zero for the subscript, while, in the physical limit (PL), the negative subscript gives that
Substituting for the Lagrangian gives that
where is a generalized nonconservative force derived from .
Note that Equation 9.46 can be derived equally well by taking the direct functional derivative with respect to , that is,
The above time-asymmetric formalism applies Hamilton’s action principle to systems that involve initial boundary conditions while the second path corresponds to the final boundary conditions. This framework, proposed recently by Galley, provides a remarkable advance for the handling of nonconservative action in Lagrangian and Hamiltonian mechanics.[2] This formalism directly incorporates the variational principle for initial boundary conditions and causal dynamics that are usually required for applications of Lagrangian and Hamiltonian mechanics. Currently, there is limited exploitation of this new formalism because there has been insufficient time for it to become well known, for full recognition of its importance, and for the development and publication of applications. Chapter 10 discusses an application of this formalism to nonconservative systems in classical mechanics.
9.3: Lagrangian¶
Standard Lagrangian¶
Lagrangian mechanics, as introduced in chapter was based on the concepts of kinetic energy and potential energy. d’Alembert’s principle of virtual work was used to derive Lagrangian mechanics in chapter 6 and this led to the definition of the standard Lagrangian. That is, the standard Lagrangian was defined in chapter 6.2 to be the difference between the kinetic and potential energies.
Hamilton extended Lagrangian mechanics by defining Hamilton’s Principle, equation , which states that a dynamical system follows a path for which the action functional is stationary, that is, the time integral of the Lagrangian. Chapter 6 showed that using the standard Lagrangian for defining the action functional leads to the Euler-Lagrange variational equations
The Lagrange multiplier terms handle the holonomic constraint forces and handles the remaining excluded generalized forces. Chapters showed that the use of the standard Lagrangian, with the Euler-Lagrange equations 9.51, provides a remarkably powerful and flexible way to derive second-order equations of motion for dynamical systems in classical mechanics.
Note that the Euler-Lagrange equations, expressed solely in terms of the standard Lagrangian 9.51, that is, excluding the terms, are valid only under the following conditions:
The forces acting on the system, apart from any forces of constraint, must be derivable from scalar potentials.
The equations of constraint must be relations that connect the coordinates of the particles and may be functions of time, that is, the constraints are holonomic.
The terms extend the range of validity of using the standard Lagrangian in the Lagrange-Euler equations by introducing constraint and omitted forces explicitly.
Chapters exploited Lagrangian mechanics based on use of the standard definition of the Lagrangian. The present chapter will show that the powerful Lagrangian formulation, using the standard Lagrangian, can be extended to include alternative non-standard Lagrangians that may be applied to dynamical systems where use of the standard definition of the Lagrangian is inapplicable. If these non-standard Lagrangians satisfy Hamilton’s Action Principle, , then they can be used with the Euler-Lagrange equations to generate the correct equations of motion, even though the Lagrangian may not have the simple relation to the kinetic and potential energies adopted by the standard Lagrangian. Currently, the development and exploitation of non-standard Lagrangians is an active field of Lagrangian mechanics.
Gauge invariance of the standard Lagrangian¶
Note that the standard Lagrangian is not unique in that there is a continuous spectrum of equivalent standard Lagrangians that all lead to identical equations of motion. This is because the Lagrangian is a scalar quantity that is invariant with respect to coordinate transformations. The following transformations change the standard Lagrangian, but leave the equations of motion unchanged.
The Lagrangian is indefinite with respect to addition of a constant to the scalar potential which cancels out when the derivatives in the Euler-Lagrange differential equations are applied.
The Lagrangian is indefinite with respect to addition of a constant kinetic energy.
The Lagrangian is indefinite with respect to addition of a total time derivative of the form for any differentiable function of the generalized coordinates plus time, that has continuous second derivatives.
This last statement can be proved by considering a transformation between two related standard Lagrangians of the form
This leads to a standard Lagrangian that has the same equations of motion as as is shown by substituting Equation 9.52 into the Euler-Lagrange equations. That is,
Thus even though the related Lagrangians and are different, they are completely equivalent in that they generate identical equations of motion.
There is an unlimited range of equivalent standard Lagrangians that all lead to the same equations of motion and satisfy the requirements of the Lagrangian. That is, there is no unique choice among the wide range of equivalent standard Lagrangians expressed in terms of generalized coordinates. This discussion is an example of gauge invariance in physics.
Modern theories in physics describe reality in terms of potential fields. Gauge invariance, which also is called gauge symmetry, is a property of field theory for which different underlying fields lead to identical observable quantities. Well-known examples are the static electric potential field and the gravitational potential field where any arbitrary constant can be added to these scalar potentials with zero impact on the observed static electric field or the observed gravitational field. Gauge theories constrain the laws of physics in that the impact of gauge transformations must cancel out when expressed in terms of the observables. Gauge symmetry plays a crucial role in both classical and quantal manifestations of field theory, e.g. it is the basis of the Standard Model of electroweak and strong interactions.
Equivalent Lagrangians are a clear manifestation of gauge invariance as illustrated by equations 9.52, 9.53 which show that adding any total time derivative of a scalar function to the Lagrangian has no observable consequences on the equations of motion. That is, although addition of the total time derivative of the scalar function changes the value of the Lagrangian, it does not change the equations of motion for the observables derived using equivalent standard Lagrangians.
For Lagrangian formulations of classical mechanics, the gauge invariance is readily apparent by direct inspection of the Lagrangian.
Non-standard Lagrangians¶
The definition of the standard Lagrangian was based on d’Alembert’s differential variational principle. The flexibility and power of Lagrangian mechanics can be extended to a broader range of dynamical systems by employing an extended definition of the Lagrangian that is based on Hamilton’s Principle, equation . Note that Hamilton’s Principle was introduced 46 years after development of the standard formulation of Lagrangian mechanics. Hamilton’s Principle provides a general definition of the Lagrangian that applies to standard Lagrangians, which are expressed as the difference between the kinetic and potential energies, as well as to non-standard Lagrangians where there may be no clear separation into kinetic and potential energy terms. These non-standard Lagrangians can be used with the Euler-Lagrange equations to generate the correct equations of motion, even though they may have no relation to the kinetic and potential energies. The extended definition of the Lagrangian based on Hamilton’s action functional can be exploited for developing non-standard definitions of the Lagrangian that may be applied to dynamical systems where use of the standard definition is inapplicable. Non-standard Lagrangians can be equally as useful as the standard Lagrangian for deriving equations of motion for a system. Secondly, non-standard Lagrangians, that have no energy interpretation, are available for deriving the equations of motion for many nonconservative systems. Thirdly, Lagrangians are useful irrespective of how they were derived. For example, they can be used to derive conservation laws or the equations of motion. Coordinate transformations of the Lagrangian is much simpler than that required for transforming the equations of motion. The relativistic Lagrangian defined in chapter 17.6 is a well-known example of a non-standard Lagrangian.
Inverse variational calculus¶
Non-standard Lagrangians and Hamiltonians are not based on the concept of kinetic and potential energies. Therefore, development of non-standard Lagrangians and Hamiltonians require an alternative approach that ensures that they satisfy Hamilton’s Principle, equation , which underlies the Lagrangian and Hamiltonian formulations. One useful alternative approach is to derive the Lagrangian or Hamiltonian via an inverse variational process based on the assumption that the equations of motion are known. Helmholtz developed the field of inverse variational calculus which plays an important role in development of non-standard Lagrangians. An example of this approach is use of the well-known Lorentz force as the basis for deriving a corresponding Lagrangian to handle systems involving electromagnetic forces. Inverse variational calculus is a branch of mathematics that is beyond the scope of this textbook. The Douglas theorem states that, if the three Helmholtz conditions are satisfied, then there exists a Lagrangian that, when used with the Euler-Lagrange differential equations, leads to the given set of equations of motion. Thus, it will be assumed that the inverse variational calculus technique can be used to derive a Lagrangian from known equations of motion.
9.4: Application of Hamilton’s Action Principle to Mechanics¶
Knowledge of the equations of motion is required to predict the response of a system to any set of initial conditions. Hamilton’s action principle, that is built into Lagrangian and Hamiltonian mechanics, coupled with the availability of a wide arsenal of variational principles and techniques, provides a remarkably powerful and broad approach to deriving the equations of motions required to determine the system response.
As mentioned in the Prologue, derivation of the equations of motion for any system, based on Hamilton’s Action Principle, separates naturally into a hierarchical set of three stages that differ in both sophistication and understanding, as described below.
Action stage: The primary “action stage” employs Hamilton’s Action functional, to derive the Lagrangian and Hamiltonian functionals. This action stage provides the most fundamental and sophisticated level of understanding. It involves specifying all the active degrees of freedom, as well as the interactions involved. Symmetries incorporated at this primary action stage can simplify subsequent use of the Hamiltonian and Lagrangian functionals.
Hamiltonian/Lagrangian stage: The “Hamiltonian/Lagrangian stage” uses the Lagrangian or Hamiltonian functionals, that were derived at the action stage, in order to derive the equations of motion for the system of interest. Symmetries, not already incorporated at the primary action stage, may be included at this secondary stage.
Equations of motion stage: The “equations-of-motion stage” uses the derived equations of motion to solve for the motion of the system subject to a given set of initial boundary conditions. Nonconservative forces, such as dissipative forces, that were not included at the primary and secondary stages, may be added at the equations of motion stage.
Lagrange omitted the action stage when he used d’Alembert’s Principle to derive Lagrangian mechanics. The Newtonian mechanics approach omits both the primary “action” stage, as well as the secondary “Hamiltonian/Lagrangian” stage, since Newton’s Laws of Motion directly specify the “equations-of-motion stage”. Thus these do not exploit the considerable advantages provided by the use of the action, the Lagrangian, and the Hamiltonian. Newtonian mechanics requires that all the active forces be included when deriving the equations of motion, which involves dealing with vector quantities. In Newtonian mechanics, symmetries must be incorporated directly at the equations of motion stage, which is more difficult than when done at the primary “action” stage, or the secondary “Lagrangian/Hamiltonian” stage. The “action” and “Hamiltonian/Lagrangian” stages allow for use of the powerful arsenal of mathematical techniques that have been developed for applying variational principles.
There are considerable advantages to deriving the equations of motion based on Hamilton’s Principle, rather than derive them using Newtonian mechanics. It is significantly easier to use variational principles to handle the scalar functionals, action, Lagrangian, and Hamiltonian, rather than starting at the equationsof-motion stage. For example, utilizing all three stages of algebraic mechanics facilitates accommodating extra degrees of freedom, symmetries, and interactions. The symmetries identified by Noether’s theorem are more easily recognized during the primary “action” and secondary “Hamiltonian/Lagrangian” stages rather than at the subsequent “equations of motion” stage. Approximations made at the “action” stage are easier to implement than at the “equations-of-motion” stage. Constrained motion is much more easily handled at the primary “action”, or secondary “Hamilton/Lagrangian” stages, than at the equations-of-motion stage. An important advantage of using Hamilton’s Action Principle, is that there is a close relationship between action in classical and quantal mechanics, as discussed in chapters 15 and 18. Algebraic principles, that underly analytical mechanics, naturally encompass applications to many branches of modern physics, such as relativistic mechanics, fluid motion, and field theory.
In summary, the use of the single fundamental invariant quantity, action, as described above, provides a powerful and elegant framework, that was developed first for classical mechanics, but now is exploited in a wide range of science, engineering, and economics. An important feature of using the algebraic approach to classical mechanics is the tremendous arsenal of powerful mathematical techniques that have been developed for use of variational calculus applied to Lagrangian and Hamiltonian mechanics. Some of these variational techniques were presented in chapters 6, 7, 8, and 9, while others will be introduced in chapter 15.
9.S: Hamilton’s Action Principle (Summary)¶
The Hamilton’s 1834 publication, introducing both Hamilton’s Principle of Stationary Action and Hamiltonian mechanics, marked the crowning achievements for the development of variational principles in classical mechanics. A fundamental advantage of Hamiltonian mechanics is that it uses the conjugate coordinates , , plus time , which is a considerable advantage in most branches of physics and engineering. Compared to Lagrangian mechanics, Hamiltonian mechanics has a significantly broader arsenal of powerful techniques that can be exploited to obtain an analytical solution of the integrals of the motion for complicated systems, as described in chapter 15. In addition, Hamiltonian dynamics provides a means of determining the unknown variables for which the solution assumes a soluble form, and is ideal for study of the fundamental underlying physics in applications to fields such as quantum or statistical physics. As a consequence, Hamiltonian mechanics has become the preeminent variational approach used in modern physics.
This chapter has introduced and discussed Hamilton’s Principle of Stationary Action, which underlies the elegant and remarkably powerful Lagrangian and Hamiltonian representations of algebraic mechanics. The basic concepts employed in algebraic mechanics are summarized below.
Hamilton’s Action Principle¶
As discussed in chapter 9.2, Hamiltonian mechanics is built upon Hamilton’s action functional
Hamilton’s Principle of least action states that
Generalized momentum ¶
In chapter 7.2, the generalized (canonical) momentum was defined in terms of the Lagrangian to be
Chapter 9.2.2 defined the generalized momentum in terms of the action functional to be
Generalized energy ¶
Jacobi’s Generalized Energy was defined in Equation 7.37 as
Hamiltonian function ¶
The Hamiltonian was defined in terms of the generalized energy plus the generalized momentum. That is
where , correspond to -dimensional vectors, e.g. and the scalar product . Chapter 8.2 used a Legendre transformation to derive this relation between the Hamiltonian and Lagrangian functions. Note that whereas the Lagrangian is expressed in terms of the coordinates , plus conjugate velocities , the Hamiltonian is expressed in terms of the coordinates plus their conjugate momenta . For scleronomic systems, using the standard Lagrangian, in equations and , shows that the Hamiltonian simplifies to be equal to the total mechanical energy, that is, .
Generalized energy theorem¶
The equations of motion lead to the generalized energy theorem which states that the time dependence of the Hamiltonian is related to the time dependence of the Lagrangian.
Note that if all the generalized non-potential forces and Lagrange multiplier terms are zero, and if the Lagrangian is not an explicit function of time, then the Hamiltonian is a constant of motion.
Lagrange equations of motion¶
Equation 6.60 gives that the Lagrange equations of motion are
where .
Hamilton’s equations of motion¶
Chapter 8.3 showed that a Legendre transform, plus the Lagrange-Euler equations, lead to Hamilton’s equations of motion. Hamilton derived these equations of motion directly from the action functional, as shown in chapter 9.2.
Note the symmetry of Hamilton’s two canonical equations. The canonical variables , are treated as independent canonical variables. Lagrange was the first to derive the canonical equations but he did not recognize them as a basic set of equations of motion. Hamilton derived the canonical equations of motion from his fundamental variational principle and made them the basis for a far-reaching theory of dynamics. Hamilton’s equations give first-order differential equations for , for each of the degrees of freedom. Lagrange’s equations give second-order differential equations for the variables , .
Hamilton-Jacobi equation¶
Hamilton used Hamilton’s Principle plus Equation 9.19 to derive the Hamilton-Jacobi equation.
The solution of Hamilton’s equations is trivial if the Hamiltonian is a constant of motion, or when a set of generalized coordinate can be identified for which all the coordinates are constant, or are cyclic (also called ignorable coordinates). Jacobi developed the mathematical framework of canonical transformation required to exploit the Hamilton-Jacobi equation.
Hamilton’s Principle applied using initial boundary conditions¶
The definition of Hamilton’s Principle assumes integration between the initial time and final time . A recent development has extended applications of Hamilton’s Principle to apply to systems that are defined in terms of only the initial boundary conditions. This method doubles the number of degrees of freedom and uses a coupling Lagrangian between the corresponding and doubled degrees of freedom
and where is a generalized nonconservative force derived from .
Standard Lagrangians¶
Derivation of Lagrangian mechanics, using d’Alembert’s principle of virtual work, assumed that the Lagrangian is defined by Equation 9.52
This was used in equation to derive the action in terms of the fundamental Lagrangian defined by Equation 9.52. The assumption that the action is the fundamental property inverts this procedure and now equation is used to derived the Lagrangian. That is, the assumption that Hamilton’s Principle is the foundation of algebraic mechanics defines the Lagrangian in terms of the fundamental action .
Non-standard Lagrangians¶
The flexibility and power of Lagrangian mechanics can be extended to a broader range of dynamical systems by employing an extended definition of the Lagrangian that assumes that the action is the fundamental property, and then the Lagrangian is defined in terms of Hamilton’s variational action principle using Equation 9.2. It was illustrated that the inverse variational calculus formalism can be used to identify non-standard Lagrangians that generate the required equations of motion. These nonstandard Lagrangians can be very different from the standard Lagrangian and do not separate into kinetic and potential energy components. These alternative Lagrangians can be used to handle dissipative systems which are beyond the range of validity when using standard Lagrangians. That is, it was shown that several very different Lagrangians and Hamiltonians can be equivalent for generating useful equations of motion of a system. Currently the use of non-standard Lagrangians is a narrow, but active, frontier of classical mechanics with important applications to relativistic mechanics.
Gauge invariance of the standard Lagrangian¶
It was shown that there is a continuum of equivalent standard Lagrangians that lead to the same set of equations of motion for a system. This feature is related to gauge invariance in mechanics. The following transformations change the standard Lagrangian, but leave the equations of motion unchanged.
The Lagrangian is indefinite with respect to addition of a constant to the scalar potential which cancels out when the derivatives in the Euler-Lagrange differential equations are applied.
Similarly the Lagrangian is indefinite with respect to addition of a constant kinetic energy.
The Lagrangian is indefinite with respect to addition of a total time derivative of the form for any differentiable function of the generalized coordinates, plus time, that has continuous second derivatives.
Application of Hamilton’s Action Principle to mechanics¶
The derivation of the equations of motion for any system can be separated into a hierarchical set of three stages in both sophistication and understanding. Variational principles are employed during the primary “action” stage and secondary “Hamilton/Lagrangian” stage to derive the required equations of motion, which then are solved during the third “equations-of-motion stage”. Hamilton’s Action Principle, is a scalar function that is the basis for deriving the Lagrangian and Hamiltonian functions. The primary “action stage” uses Hamilton’s Action functional, to derive the Lagrangian and Hamiltonian functionals that are based on Hamilton’s action functional and provide the most fundamental and sophisticated level of understanding. The second “Hamiltonian/Lagrangian stage” involves using the Lagrangian and Hamiltonian functionals to derive the equations of motion. The third “equations-of-motion stage” uses the derived equations of motion to solve for the motion subject to a given set of initial boundary conditions. The Newtonian mechanics approach bypasses the primary “action” stage, as well as the secondary “Hamiltonian/Lagrangian” stage. That is, Newtonian mechanics starts at the third “equations-of-motion” stage, which does not allow exploiting the considerable advantages provided by use of action, the Lagrangian, and the Hamiltonian. Newtonian mechanics requires that all the active forces be included when deriving the equations of motion, which involves dealing with vector quantities. This is in contrast to the action, Lagrangian, and Hamiltonian which are scalar functionals. Both the primary “action” stage, and the secondary “Lagrangian/Hamiltonian” stage, exploit the powerful arsenal of mathematical techniques that have been developed for exploiting variational principles.
The term “action functional” was named “Hamilton’s Principal Function” in older texts. The name usually is abbreviated to “action” in modern mechanics.
This topic goes beyond the planned scope of this book. It is recommended that the reader refer to the work of Galley, Tsang, and Stein[Gal13, Gal14] for further discussion plus examples of applying this formalism to nonconservative systems in classical mechanics, electromagnetic radiation, RLC circuits, fluid dynamics, and field theory.