A new perspective in linear Cauchy Elasticity: variational minimum principles for statics, dynamics, and heterogeneous materials
Abstract
A variational minimum principle for linear elastodynamics of a possibly heterogeneous material without a stored energy function is developed. It involves a change of variables to dual fields, and results in a degenerate elliptic Euler-Lagrange system, even when the primal elastodynamics is hyperbolic. Uniqueness assertions for the dual dynamic and static problems and implications of the degenerate ellipticity are sketched. Some implications pertaining to heterogeneous materials and ones with indefinite elastic moduli are discussed.
1 Introduction
This work describes an application to Linear Cauchy Elasticity of a technique [1, 3, 2] for associating a family of variational minimum (i.e., convex) principles to a system of equations with boundary and initial conditions when appropriate. The given ‘primal’ equations may not have a variational structure in terms of the primal fields or variables they are originally posed in, e.g., elasticity without a strain energy function. The procedure involves an adapted change of variables to a new set of ‘dual’ fields in terms of which a functional can be posed whose Euler-Lagrange equations become the primal set of equations, interpreted through the change of variables mapping involved. This approach has been further developed and demonstrated through computation in [8, 9, Ach6, 15, 10, 16], with rigorous results developed in [15, AGS24, AG25, 10, 17, ASZ24]. Here, we work out the details of the scheme in the context of the governing system being that of linear elastodynamics of a generally heterogeneous material possibly without a strain energy function. As is well understood, accounting for initial conditions of an initial value problem is problematic in a variational principle - including Hamilton’s celebrated principle - see, e.g., [5] for a discussion; Hamilton’s principle also has a problem with causality due to the need for an acausal final-time boundary condition on position (cf. [19], [4, 7]). The way out of this difficulty in linear problems has been to resort to working in the Laplace transform domain [5, 19], or at ‘fixed frequency’ [11], i.e., for elastodynamic fields satisfying a special ansatz for its time-dependent part, assumed separable from its space-dependent part; also see [7] in the context of Hamiltonian particle mechanics using an adaptation of the Schwinger-Keldysh formalism which involves a doubling of variables and a backward-in-time path. The approach adopted here works in real space-time domains without involving convolutions, assumptions of separability, or a doubling of variables in which the original system is posed, and works as well for nonlinear and dissipative time-dependent equations as shown in [9, 16, 17, 15].
The work [15] develops the dual approach for nonlinear elasticity, focusing primarily on elastostatics; the form of the dual functional is developed for quadratic nonlinearity in stress. The emphasis in this work is to explicitly work out the case of linear elasticity in full detail, which is also of interest to a much broader class of researchers interested in its theoretical and/or practical aspects. As examples, the specialization to linear elastodynamics greatly facilitates the explorations of the type provided in Sections 3 and 4. For comparison to other variational techniques for linear elastodynamics in the literature, to our knowledge this work provides the first convex variational principle for linear elastodynamics (with a strain energy function and without) in the time domain without approximation, recovering field equations, boundary conditions, and initial conditions without any loss of causality. In situations when a stored energy function exists, the recovery of initial conditions, a causal formulation, and the convexity of the variational principle for the whole range of positive-definite to indefinite stored energy function are some of the main contributions of this work.
An outline of the paper is as follows: in Sec. 2 the variational principle for elastodynamics is developed. In Sec. 3 a uniqueness result for the Euler-Lagrange system of the variational principle is studied; some discussion of its ellipticity properties (in space-time domains) is also provided. In Sec. 4 implications of the scheme for heterogeneous/composite materials are discussed. Some concluding remarks are recorded in Sec. 5.
2 A convex variational principle for linear elastodynamics
Given a 3-dimensional domain with boundary the closure of a union of disjoint sets, and an interval of time , consider the system of linear elastodynamics whose given elastic moduli and density fields, can vary in space. has minor symmetries and may or may not have major symmetries; we will use the notation . The fields , will denote the displacement and velocity, and and the symmetric and antisymmetric parts of the displacement gradient. When using direct notation a “” will denote the inner product of two vectors and a “” that of two second-order tensors. For fourth-order tensors, , and . For this work, it will be convenient to work with a non-dimensional representation of the equations of linear elasticity. Using as a physical scale for stress11 1 E.g., the Young’s modulus for the material., as some physical scale for length22 2 E.g., dimension of a body, a feature-size like grain size, a wave length of imposed disturbance, or a ratio of an elastic modulus, say , and the magnitude of the body force density at some point or its average over a physically relevant region in an infinite homogeneous medium., and as a scale for time33 3 E.g., the time for an elastic wave to traverse the length given by where is a constant mass density scale; or a reciprocal frequency of harmonic loading., the equations of linear elastodynamics can be written as
| (1) | ||||
where
Here, the body force density, the displacement and traction boundary conditions, and the displacement and velocity initial conditions. Henceforth, we work with the system (1) dropping all tildes for convenience:
| (2) | ||||
Using the notation , and where are arbitrarily specified functions referred to as base states (in space-time domains), we introduce an auxiliary potential
| (3) | ||||
for the system of linear elastodynamics, where are nondimensional fields, with freedom to be designed as needed. Due to the quadratic form through which it enters, is assumed to have major symmetry without loss of generality.
Consider now a pre-dual functional
| (4) | ||||
with
| (5) | ||||
The fields are referred to as dual fields. The values (right hand-sides) of the dual (Dirichlet) space-time boundary conditions (5) can be arbitrarily specified, without loss of generality (note that in such cases no extra terms are, nevertheless, included in the pre-dual functional which is free to design.)
We refer to the first space-time bulk integrand in (4) as the Lagrangian
| (6) |
Defining the DtP (dual-to-primal) map as the solution of
for in terms of so that
| (7) |
we have
| (8a) | ||||
| (8b) | ||||
| (8c) | ||||
| (8d) | ||||
where square brackets around a pair of indices represents antisymmetrization e.g., and ordinary brackets, symmetrization . Also, for any second-order tensor, say , will represent its symmetric and antisymmetric parts, respectively.
Now define the dual functional, , by substituting for in the pre-dual functional :
| (9) | ||||
with the essential space-time boundary conditions (5). Then, noting (7) and that the Lagrangian (6) is affine in by construction, it is easily seen that the Euler-Lagrange equations of the functional is the governing set of equations and side conditions of linear Cauchy elastodynamics (2) with the replacement , regardless of the choices of the specified functions
subject to and merely invertible on the space of symmetric second-orders tensors.
Let us now assume that and is positive-definite on the space of symmetric second order tensors and note that
Then, due to the lack of any differential constraints on in the definition of ,
Therefore,
and because is affine in by construction, and therefore convex in , defining our problem statement as finding the infimum of , i.e.,
corresponds to a minimization problem for a convex functional, solving formally the problem of linear Cauchy Elastodynamics (statics is a special case) through the use of the DtP map, including for heterogeneous materials.
To obtain the explicit form of the dual functional, the (dual) Lagrangian may be expressed, using the DtP map (8), as
Thus,
| (11) | ||||
with its fields satisfying the following essential boundary conditions on the space-time domain :
| (12) | ||||
For positive-definite and , the semi-definiteness of the second variation of the representation (11) of the dual functional furnishes another (direct) proof of its convexity, regardless of the symmetry or positivity of or .
The DtP map for the problem is given by
| (13) | ||||
The system of linear elastodynamics along with its initial and boundary conditions, expressed in terms of the dual fields, is
| (14) | ||||
The following observations are in order:
- 1.
When the base state, , is a solution to the system (2), is a solution to the dual system (14) and an (absolute) minimizer of the dual functional, (11). This is an important property of the dual scheme for both theoretical and practical purposes [17], as it motivates the reasoning that if a base state is ‘close’ to a solution of (2), then it may be possible to obtain a solution of (2) by using as a good guess to the corresponding dual problem (14).
- 2.
- 3.
When solutions of the primal system (2) are unique, any solution to the dual system (14), generated from, e.g., different choices of boundary conditions (5) or , must generate the unique primal solution through the use of the DtP map. Examples of this fact are provided in [8] in the context of the heat equation. However, if the primal system is known to have non-unique solutions, the auxiliary potential, parametrized by , can be used to obtain such solutions through the dual scheme in a ‘stable’ manner - e.g., even when is indefinite. In other words, the use of the auxiliary potential, , acts as a selection criterion for solutions of the primal problem.
- 4.
Given that the primal problem (2) is an initial value-problem a natural question is whether prescribing final-time Dirichlet boundary conditions on the dual fields interferes in any way with causality of the primal solutions obtained from the above dual approach. Such interference does not arise because the DtP map (8) involves time derivatives of the dual fields, and prescribing the values of dual fields at final-time allows these time derivatives (at time ) to adjust so as to recover the correct primal solution. This feature has been explained and demonstrated in [2, Sec. 7], [8, 9, 16], with an exact solution for the first-order initial value problem worked out by the dual scheme in [16, Sec. 3.1].
- 5.
Due to the availability of the functional , a gradient descent algorithm in a fake time-like variable, can be formulated to obtain solutions to the primal system (2), see [17, Sec. 3]. The fact that is convex also guarantees that the norm of its gradient is non-increasing along a gradient flow - this is of great importance in a solution scheme that employs more than one functional (based on choices of base states) while keeping the gradient continuous at such switches, with the final goal of reaching a vanishing norm of the gradient (a weak solution to the primal equations is then obtained). Such ideas can be useful for iterative schemes for linear elasticity of heterogeneous materials.
- 6.
The point of view adopted here is that all physics is contained in the primal system and the variational scheme is simply a mathematical device to solve that system. Thus, the dual fields are unlikely to have much physical significance - e.g., they admit acausal boundary conditions. Also, in nonlinear problems the use of a whole sequence of dual functionals, parametrized by base states, is employed to obtain a single primal solution; physically important quantities are not usually found in such abundance.
- 7.
From a practical standpoint, having to solve a boundary-value-problem on a large space-time domain as might be the case to probe long-time behavior can be prohibitive. Solution of the dual system (14) can be divided into space-time domain slices, with an arbitrary, finite, sub-division of the time interval . In each such space-time slice the dual problem can be solved, the primal solution recovered at the final time of the slice and used to define the initial conditions for the next slice. This strategy for solving the dual problem has been demonstrated in the solution of Euler’s ODE system for motion of a rigid body about a fixed point (with and without damping), the heat equation, and the linear transport equation in [8], and in the context of the (in)viscid Burgers equation in [9].
3 Formal uniqueness for the dual system (14) and its degenerate ellipticity
For practical purposes related to the use of the dual system (14), it is reassuring to have an assertion of uniqueness of solutions for any specific set of dual Dirichlet boundary conditions (5) and base states , at least when the elastic modulus is positive definite and has major symmetry. We provide such a sufficient condition in this Section. In the static case, uniqueness holds simply with positive definiteness without an assumption of major symmetry and we sketch that proof as well.
Consider two solutions of system (14) for and let their difference be denoted as . Then
Also, the difference fields satisfy the following space-time boundary conditions:
| (16a) | ||||
| (16b) | ||||
| (16c) | ||||
| (16d) | ||||
| (16e) | ||||
| (16f) | ||||
| (16g) | ||||
| (16h) | ||||
On forming products of the three sets of equations in (15) with , respectively, and using the space-time boundary conditions (16a)-(16d), one obtains,
and due to the positive-definiteness of on the space of symmetric tensors and ,
| (17a) | ||||
| (17b) | ||||
| (17c) | ||||
| (17d) | ||||
hold on . Then (17) (all four equations together) gives
| (18) |
which further gives, after forming a scalar product with and integrating over , when has major symmetry,
| (19) |
Now, from (16g) one has that
| (20) |
Assuming that (using (17d) as motivation)
| (21) |
holds, and combining with (16h) gives
| (22) |
Assuming as well that (using (17c) as motivation),
| (23) |
(20)-(22)-(23) imply that the boundary term in (19) vanishes so that
| (24) |
for all . Assuming (using (17b) as motivation)
| (25) |
and noting (16e) and (16f), the first integral in (24) vanishes and when and is positive-definite (major symmetry has been assumed), we have that
| (26a) | ||||
| (26b) | ||||
This implies, from (17c)-(17d) (and an assumption of continuous extension to the boundary) that
which combined with (17a) and (16e) gives
Of course, (26a) along with (16f) implies that
The (formal) argument above for dual elastodynamics may be summarized as follows:
Consider a class of functions , , , such that the difference of any two functions in the class satisfy the conditions (16), (21), (23), (25) with the replacement . Furthermore, assume that is continuous on .
When and is positive-definite with major symmetry, there can be at most one solution in this class of functions to the system (14).
Uniqueness for dual Elastostatics: In the problem of dual elastostatics, the dual fields are redundant, and the difference solution satisfies
with the boundary conditions
| (28a) | ||||
| (28b) | ||||
| (28c) | ||||
Then, following the previous arguments for the dynamic case, one has
| (29a) | ||||
| (29b) | ||||
| (29c) | ||||
which further implies
| (30) |
Forming the scalar product with , integrating by parts and using (28b) gives
| (31) |
and assuming (29b)-(29c) also hold on , and combining with (28c), the boundary term in (31) vanishes.
Then, assuming to be positive-definite (but not necessarily symmetric, a difference from what was used in the proof of uniqueness for the dynamic case), one has on and combining with (28b), uniqueness for holds. But then, from (29b)-(29c), uniqueness for holds as well.
The (formal) argument above for dual elastostatics may be summarized as follows:
Consider a class of functions , , , such that the difference of any two functions in the class satisfy the conditions i) (29b) and (29c) on and ii) (28), with the replacement .
When is positive-definite (major symmetry is not required), there can be at most one solution in this class of functions to the system
Returning to the full dual elastodynamic problem, it is to be noted that while uniqueness in the absence of strict convexity might seem surprising, degenerate elliptic equations are known to have uniqueness in many circumstances, see, e.g., [13] (here we deal with a system of equations).
Purely from the standpoint of the operator containing the highest order derivatives of the dual system (14) given by
with the corresponding quadratic form
| (33) |
the system (14) is at worst degenerate-elliptic since the quadratic form (33) is positive semi-definite, regardless of whether is positive-definite and is positive (assuming is positive definite and , which are free to choose). This property is true of the full quadratic form in (11) as well.
Thus, it is worthy of note that the degenerate ellipticity of the dual problem holds even when the primal elastodynamic problem is hyperbolic or ill-posed hyperbolic with no continuous dependence on initial data when is indefinite. Initial demonstrations of success in solving such problems in the context of the linear transport equation, the (in)viscid Burgers equation, and elastodynamics of a bar with non-convex strain energy have been provided in [8, 9, 15].
We end this Section by exploring the following question: even strictly convex linear elastodynamics admits stress-waves induced by boundary/initial conditions (or body forces), which travel at the linear elastic wave speed of the material. This naturally must involve a propagating strain discontinuity as well. How can a ‘close-to’ elliptic PDE formulation recover such intrinsically ‘hyperbolic’ behavior?
Consider the dual ‘action’ (11) (for and positive-definite) and one asks about solutions that have vanishing bulk ‘dual energy/action.’ This necessarily requires the system (17) to hold (where the ∗ fields now represents the 0-dual-action solution under consideration), which further implies that (18) holds. But this directly implies, because of its hyperbolic nature, that admits rank-one discontinuities in the usual way across the characteristic surfaces of the physical linear elastodynamic equation and, through the DtP map (8), in the physical strain field produced by the dual scheme, regardless of the chosen , , . On the other hand, in the static situation there can be no discontinuities in (for +ve-definite) by the usual properties of elliptic equations. Thus, we conclude that the degenerate ellipticity of the dual system (14) is crucial for recovering physically mandated behavior as per classical linear elasticity, and ellipticity of the dual system for elastodynamics must fail. These are all manifestations of a convex minimum principle which is not strictly convex (cf., [3, Sec. 3]).
The above arguments lead to a natural conjecture: that the dual formulation, due to its degenerate ellipticity selects only ‘good’ solutions, i.e. ones with fairly mild discontinuities, for elastodynamics with indefinite tensor.
4 Heterogeneous materials
For application to heterogeneous materials, i.e., vary in space, it is useful to write the system (14) in the form (we focus only on the governing equations)
In the above, in the first equation we have added a term to both sides of the equation; for the third equation we have done the same using .
With the goal of converting the left-hand-side operator of (34) to a constant coefficient one (the ‘comparison medium’), let be a positive-definite, spatially constant, compliance tensor with major and minor symmetries which is otherwise arbitrary and a spatially constant density field. Define
and assume (and ) are invertible on the space of second order tensors. One then makes the choice
| (35) |
resulting in (34) taking the form
| (36a) | ||||
| (36b) | ||||
| (36c) | ||||
It is worthy of note that obtaining a constant-coefficient highest order operator on the l.h.s. of (36) does not come at the expense of introducing a highest order operator on the r.h.s. of the equation, as is effectively the case in the conventional method used to solve the linear elastic problem for heterogeneous materials [6, 19, 12] (we will refer to this as the ‘established formulation’). There, the addition arises from the (negative) divergence of the stress polarization tensor in the static case:
To elaborate a bit further, consider the elastostatic case and ignore the base states for the moment in (36). On defining the polarization tensor here to be
(36b) appears to have a very similar structure to the established formulation for solving the problem for a given stress polarization field, as described in the above references. Here, the polarization depends on the tensor field which, very roughly speaking, from (36c) would seem to have one less spatial derivative than , the latter being the analogous situation in the case in the established formulation. Thus, it is conceivable that solving for and here, using a natural adaptation of the Moulinec-Suquet [12] iterative scheme for solving such integral equations, and through these fields the solution for the physical strain field through the DtP map (8c), has better properties. A similar argument would also be applicable to the problem of heterogeneous elastodynamics.
We note the following features, potentially important for possible practical application:
- 1.
- 2.
In either approach, changes of base states in an iterative scheme [17, Sec. 3], based on progressive iterates for the primal fields, can be invoked to enhance convergence to a solution.
5 Concluding remarks
A scheme for generating variational minimum principles for Cauchy elastodynamics has been described as an application of a broader program to ease the solution/approximation of difficult (linear and nonlinear) PDE problems [3]. This is achieved through a conversion of the question to a convex variational one which allows the definition of a weaker notion of solutions (Variational dual solutions [15, 17, ASZ24]) to PDE than weak solutions, thus allowing for the use of tools from the Calculus of Variations and Convex Analysis in PDE. Such solutions have the consistency that when they can be proven to be sufficiently regular, they define genuine weak solutions of the primal PDE problem. The development lends itself to natural computational implementation for practical purposes.
Several positive aspects of the development in the context of Linear Elasticity have been pointed out. An unpalatable feature is that the transformation to an at most first-order differential system, due to the requirement that the pre-dual Lagrangian have no derivatives on primal variables appearing in it, increases the number of field degrees of freedom. The number of degrees of freedom in classical linear elastodynamics is 6 (we count velocity/momentum as a separate field as that is the optimal setting, especially when it comes to questions of heterogeneous materials, see. e.g., [18]). In contrast, the present formulation needs a minimum of 12 (6 for when (2)3 is posed in symmetrized form, and three each for displacement, and velocity). Thus, while providing conceptual benefits and exposing unexpected connections, the present approach is sub-optimal for standard problems of linear elasticity from a practical point of view. However, for non-standard or difficult problems, e.g., a convex variational formulation of ‘Odd Elasticity’ [14], problems with indefinite elastic moduli (when the dynamic Cauchy problem changes type and is ill-posed due to a lack of continuous dependence w.r.t. initial data), problems with non-unique solutions, or for heterogeneous materials with high contrast, the methodology can be potentially beneficial. ‘Metamaterials’ are included in this class of problems.
While the availability of a convex variational principle for a general PDE system is a definite advantage, the considerations related to degenerate ellipticity at the end of Sec. 3 make it clear that weak coercivity of the convex dual functional, i.e., roughly speaking, the growth of the functional to as a relevant norm of its argument function tends to , is not to be expected. Thus, one of the three criteria for use of the ‘big hammer’ of the Calculus of Variations (cf., [20] for the Generalized Weierstrass Theorem44 4 Strictly speaking, one needs the theorem for extended real-valued functionals in the present case.) to prove existence of minimizers is not likely to be satisfied in many situations55 5 The affineness of the pre-dual in gives weak lower-semicontinuity, and one works in a closed, convex subset of a Hilbert space., when working within the present dual formulation. The formulation of Vorotnikov [V22], extending the pioneering work of Brenier [CMP18], is to be noted in this regard, but in any case, the gradient flow scheme proposed in [17] is perhaps the most pragmatic strategy to work with in efforts to obtain weak solutions of the original primal PDE through the dual scheme. That strategy is also immediately transferable to a computational algorithm for approximate solutions.
Finally, a speculation from a non-expert in the theory of homogenization: it seems like the natural transformation of the elastostatic(dynamic) problem of an arbitrary linear elastic composite to a space-time boundary value problem with constant-coefficient highest-order operator (36), along with a (formal) uniqueness guarantee in the usually encountered situations (major symmetry and +ve definiteness for all phases), may be expected to provide some advantages for mathematical homogenization schemes for such materials.
References
- [1] (2023) A dual variational principle for nonlinear dislocation dynamics. Journal of Elasticity 154 (1), pp. 383–395. Cited by: §1.
- [2] (2023) Variational principle for nonlinear PDE systems via duality. Quarterly of Applied Mathematics LXXXI, pp. 127–140. External Links: Link Cited by: §1, item 4.
- [3] (2025) A hidden convexity in continuum mechanics, with application to classical, continuous-time, rate-(in)dependent plasticity. Mathematics and Mechanics of Solids 30(3), pp. 701–719. External Links: Link Cited by: §1, §3, §5.
- [4] (2013) Classical mechanics of nonconservative systems. Physical Review Letters 110 (17), pp. 174301. Cited by: §1.
- [5] (1964) Variational principles for linear initial-value problems. Quarterly of Applied Mathematics XXII (3), pp. 252–256. Cited by: §1.
- [6] (1962) A variational approach to the theory of the elastic behaviour of polycrystals. Journal of the Mechanics and Physics of Solids 10 (4), pp. 343–352. Cited by: §4.
- [7] (2026) Hamilton revised: the action principle for initial value problems. arXiv e-prints. External Links: 2603.02350 Cited by: §1.
- [8] (2024) Hidden convexity in the heat, linear transport, and Euler’s rigid body equations: A computational approach. Quarterly of Applied Mathematics LXXXII, pp. 673–703. Cited by: §1, item 3, item 4, item 7, §3.
- [9] (2025) Inviscid Burgers as a degenerate elliptic problem. Quarterly of Applied Mathematics LXXXIII, pp. 315–360. Cited by: §1, item 4, item 7, §3.
- [10] (2025) Traveling wave profiles for a semi-discrete Burgers equation. Physica D, pp. 134961. External Links: Link Cited by: §1, §1.
- [11] (2009) Minimization variational principles for acoustics, elastodynamics and electromagnetism in lossy inhomogeneous bodies at fixed frequency. Proceedings of the Royal Society A: Mathematical, Physical and Engineering Sciences 465 (2102), pp. 367–396. Cited by: §1.
- [12] (1994) A fast numerical method for computing the linear and nonlinear mechanical properties of composites. Comptes Rendus de l’Académie des sciences. II 318, pp. 1417–1423. Cited by: item 1, §4, §4.
- [13] (2009) Uniqueness of solutions to degenerate elliptic problems with unbounded coefficients. Annales de l’Institut Henri Poincaré C 26 (5), pp. 2001–2024. Cited by: §3.
- [14] (2020) Odd elasticity. Nature Physics 16 (4), pp. 475–480. Cited by: §5.
- [15] (2024) A hidden convexity of nonlinear elasticity. Journal of Elasticity 156 (3), pp. 975–1014. Cited by: §1, §1, §1, §3, §5.
- [16] (2025) Variational formulation based on duality to solve partial differential equations: use of B-splines and machine learning approximants. Computer Methods in Applied Mechanics and Engineering 441, pp. 117909. Cited by: §1, item 4.
- [17] (2025) On the variational dual formulation of the nash system and an adaptive convex gradient-flow approach to nonlinear pdes. arXiv e-prints. External Links: 2512.12878 Cited by: §1, §1, item 1, item 5, item 1, item 2, §5, §5.
- [18] (1980) Polarization approach to the scattering of elastic waves—i. scattering by a single inclusion. Journal of the Mechanics and Physics of Solids 28 (5-6), pp. 287–305. Cited by: §5.
- [19] (1981) Variational principles for dynamic problems for inhomogeneous elastic media. Wave Motion 3 (1), pp. 1–11. Cited by: §1, §4.
- [20] (1995) Applied functional analysis: main principles and their applications. Vol. 109, Springer. Cited by: §5.