Carleman linearization
Linearization technique for nonlinear differential systems
In mathematics, Carleman linearization (also called Carleman embedding) is a method that formally represents a finite-dimensional nonlinear system of ordinary differential equations as an infinite-dimensional linear system. Truncating the resulting hierarchy at a finite order yields approximate solutions of the original nonlinear system, known as Carleman approximants.
The method was introduced by the Swedish mathematician Torsten Carleman in 1932. For polynomial systems, the construction successively introduces monomials of the state variables as new dependent variables. More generally, the additional variables may be chosen according to the nonlinear structure of the differential system. Carleman linearization may be regarded as a systematic extension of the usual local linearization of nonlinear differential equations, in which higher-order nonlinear terms are retained through an enlargement of the phase space.
Carleman linearization is related to composition operators, particularly the Koopman operator, which represents nonlinear dynamics through linear evolution on a space of observable functions. It has been applied in nonlinear control theory, state estimation, chemical kinetics, process modelling, mathematical biology, and the construction of finite-dimensional approximations to Koopman operators. The method has also been used in the development of quantum algorithms for nonlinear differential equations.
01Method
Consider an autonomous system of ordinary differential equations,
Carleman linearization consists of replacing nonlinear terms appearing on the right-hand side of the system with new dependent variables. Differentiating these variables generally produces further nonlinear terms, which are treated successively in the same way. The procedure therefore generates an infinite sequence of linear differential equations. A finite-dimensional approximation is obtained when the sequence is terminated according to a chosen truncation criterion. The solution of the truncated linear system is called a Carleman approximant.
For polynomial systems, the additional variables are monomials of the original state variables. The order of the approximation is then defined as the degree of the highest-order monomial retained. Other sets of dependent variables may be chosen when they are better adapted to the structure of the differential system.
After truncation at order , the original
-dimensional nonlinear system is replaced by a finite-dimensional linear system of dimension
. The value of
depends on the approximation order and on the nonlinear structure of the original equations. The approximation to the original solution is given by the components of the enlarged linear system corresponding to the original state variables.
Two forms of the construction can be distinguished according to the point in phase space around which the Carleman approximant is built: linearization around an equilibrium point, referred to as Type A, and linearization around the initial condition, referred to as Type B.
Type A: linearization around an equilibrium point
Let be an equilibrium point of the nonlinear system,
Introducing the shifted variable
places the equilibrium point at the origin. Since the constant term of the transformed differential system vanishes, truncation of the Carleman sequence produces a homogeneous linear system with constant coefficients,
where contains the retained variables and
is the corresponding
coefficient matrix. Its solution is
The approximate solution of the original nonlinear system is provided by the components of
associated with the original variables, after reversing the coordinate shift.
The first-order Type A approximant coincides with the usual local linearization obtained from the Jacobian matrix of at the equilibrium point. Higher-order approximants include additional nonlinear terms and may therefore be regarded as an extension of the local or qualitative analysis of nonlinear differential systems. For a fixed approximation order, their accuracy generally depends on the distance between the trajectory of interest and the equilibrium point around which the approximant was constructed.
Type B: linearization around the initial condition
Alternatively, the Carleman approximant may be constructed around the initial condition itself. Introducing
gives
Unless is an equilibrium point, the transformed system contains the nonzero constant term
. Consequently, truncation of the Carleman sequence produces a non-homogeneous linear system,
where is a constant vector determined by the initial condition and by the nonlinear differential system. The solution can be written as
If is invertible, this becomes
The approximation to is obtained by adding
to the components of
associated with the shifted original variables.
Type B approximants require the solution of a non-homogeneous linear system and generally involve a greater algebraic burden than Type A approximants. However, centring the approximation at the initial condition can substantially improve its local accuracy.
Local accuracy
Comparison of the derivatives at of the exact solution and the Carleman approximants reveals different local-accuracy patterns. In the one- and two-dimensional systems analysed, an approximant of order
reproduces the derivatives of the exact solution from order
through order
, independently of the point around which it is constructed.
In one dimension, let be the point around which the approximant
is constructed, and let
be the initial condition. The differences between the derivatives of the exact and approximate solutions have the following pattern:
| Order |
|
|---|---|
When the approximant is centred at the initial condition, , the differences in the second row also vanish. Thus, for the one-dimensional systems considered, a Type B approximant reproduces the derivatives of the exact solution through order
.
In two dimensions, let the approximants and
be constructed around
, and let
be the initial condition. For
, the corresponding pattern is
| Order |
|
|---|---|
When , all the displayed differences vanish. The additional agreement of the derivatives from order
through order
accounts for the greater local accuracy observed for Type B approximants of a given order.

02Examples
The following examples illustrate the construction of Type A and Type B Carleman approximants in one and two dimensions. Only one equilibrium point is considered in each Type A construction.
One-dimensional example: Riccati equation
Consider the Riccati equation
Its exact solution is
The equation has two equilibrium points, and
. The Type A approximant below is constructed only around
.
Type A approximant
Introduce the Carleman variables
Differentiation gives the infinite linear sequence
A second-order approximation is obtained by retaining and
and discarding the term
. The resulting homogeneous system is
Solving the linear system gives the second-order Type A approximant
Type B approximant
To construct the approximant around the initial condition, introduce
The differential equation becomes
Define
Retaining and
gives the second-order non-homogeneous linear system
Writing
the solution is
when is invertible. The second-order Type B approximant is therefore
where , and
are the eigenvalues of
. Although this expression is algebraically more involved than the Type A approximant, centring the construction at
generally improves its local accuracy.
The plot compares the exact solutions with second- and third-order Type A approximants constructed around the two equilibrium points and with a second-order Type B approximant constructed around each initial condition.
Two-dimensional example: center
Consider the two-dimensional nonlinear system
with initial condition
The system has an equilibrium at , which is a center and is stable in the sense of Lyapunov, but not asymptotically stable. It also has a saddle point at
. For
, its trajectories satisfy the first integral
where is constant along each trajectory.
Type A approximant
For the Type A construction, consider only the equilibrium point . At second order, introduce the enlarged set of variables
After discarding terms of degree greater than two, the Carleman system is
with initial condition
Thus,
where is the displayed coefficient matrix. The second-order Type A approximants
and
are respectively the third and first components of
.
Type B approximant
To centre the approximation at the initial condition , introduce
The transformed nonlinear system is
At first Carleman order, the nonlinear term is discarded, yielding
Let
The first-order Type B approximant is
when is invertible. The eigenvalues of the coefficient matrix are
Despite being only first order, the Type B approximation can have a numerical accuracy intermediate between the first- and second-order Type A approximants because it is centred directly on the trajectory's initial condition.
The phase portrait compares second-order Type A approximants constructed around the two critical points with first-order Type B approximants constructed around several initial conditions.
03Application to epidemiology
The Carleman approximants for the SIR model yield an explicit algebraic estimator of the time-dependent effective reproduction number from prevalence data. This application illustrates how an analytic approximation of a nonlinear system can be used not only to approximate its trajectories, but also to obtain closed-form expressions for quantities of practical interest.
Consider the SIR model
where ,
, and
are the susceptible, infectious, and removed population fractions, respectively. The transmission and removal rates are denoted by
and
. The effective reproduction number is
To account explicitly for the approximately exponential behaviour of the infectious population, the change of variable
transforms the first two SIR equations into the system
This formulation is referred to as the SYR system. Its Carleman sequence can be constructed using nonlinear terms of the form , with approximation order
.
Let
where is the value of the effective reproduction number at the time chosen as the origin. At second order, the Carleman approximant for the logarithmic prevalence is
Because this approximation is an explicit analytic function of the epidemiological parameters, it can be fitted directly to prevalence data. Assume that the removal rate is known and remains approximately constant within a short time interval, and that prevalence
has been reconstructed from incidence data. For observations indexed by
, define
The second-order approximation is fitted over a time window containing observations by minimizing the residual sum of squares
with respect to the parameters ,
, and
. The parameters are assumed to remain approximately constant within the selected time window. The resulting value of
estimates the effective reproduction number
at the central point of the interval.
Although the resulting variational equations are nonlinear, their structure allows them to be solved in closed form. Define the centred quantities
where
Minimization of the residual sum of squares then gives the algebraic estimator
Thus, once the prevalence and removal time scale have been specified, can be estimated from finite sums over the data. The expression has been tested using simulated epidemic waves with time-dependent transmission rates and SARS-CoV-2 incidence data from the Valencian Country, Spain, yielding estimates consistent with the corresponding reference values.

04Control-theoretic formulation
In control theory and state estimation, Carleman linearization can be applied to nonlinear systems with external inputs or disturbances. One such formulation considers the system
where is the state vector,
and
are analytic vector fields, and
denotes the
-th component of an external input or disturbance. Because the disturbances may depend explicitly on time, the system is not autonomous in general. When the functions
are treated as inputs, the lifted model obtained below is bilinear in the lifted state and the inputs.
Let be a nominal point. After introducing the shifted coordinate
, the nominal point is placed at the origin. For notational simplicity, the shifted variable is again denoted by
. The vector fields may then be approximated by finite Taylor expansions,
where is the Taylor truncation degree and
and
are the corresponding matricized Taylor coefficient tensors,
Here, denotes the
-th Kronecker power,
with the convention .
The time derivative of the -th Kronecker power is
Substitution of the Taylor expansions generates Kronecker powers of progressively higher degree. To obtain a finite-dimensional approximation, the hierarchy is truncated at a maximum lifted degree . Discarding terms whose total degree exceeds
gives, for
,
where
and similarly,
In these expressions, is the
-dimensional identity matrix,
denotes its
-th Kronecker power, and
.
Introducing the truncated lifted state
the system can be written in block-matrix form as
where and
are assembled from the blocks
and
, respectively. The vectors
and
contain the constant terms of the Taylor expansions, while
represents the residual produced by the Taylor and Carleman truncations.
If the disturbances are prescribed functions of time and the residual is neglected, the truncated model is a finite-dimensional affine linear time-varying system. When the disturbances are regarded as independent inputs, it is a bilinear state-input system. This formulation has been used, for example, in moving-horizon estimation of nonlinear processes.
05Convergence and limitations
Except in particular systems for which the Carleman sequence closes after finitely many steps, a finite-order Carleman system provides an approximation rather than an exact representation of the original nonlinear dynamics. The accuracy depends on the order of truncation, the point in phase space around which the approximant is constructed, the initial condition, and the time interval considered.
The dimension of the enlarged linear system grows rapidly with both the number of original variables and the approximation order. Consequently, the computation and analytic exponentiation of the coefficient matrix may become difficult at high orders. Type B approximants may offer better local accuracy, but require the solution of a non-homogeneous system and generally have more complicated algebraic expressions than Type A approximants.
Convergence and truncation-error estimates have been established for particular classes of nonlinear systems, but the applicable bounds and time intervals depend on assumptions about the vector field and the solution.
Because finite-order Carleman approximants are solutions of linear differential systems with constant coefficients, they consist of linear combinations of functions such as ,
, and
. They therefore do not reproduce finite singularities of an exact nonlinear solution in the complex time plane. They may also fail to preserve global dynamical structures such as closed orbits; at higher orders, repeated eigenvalues can generate secular terms that prevent an approximate trajectory from closing.
Sources and credits
This article is adapted from the Wikipedia article “Carleman linearization”, written by its contributors and licensed under CC BY-SA 4.0. Fathomly has changed the layout, removed citation markers, navigation and maintenance notices, and adjusted punctuation. This adapted version is shared under the same license. For references, see the original article.
Images, from Wikimedia Commons:
- Carleman approximants for the Riccati equation.svg by Actinopterigio, CC BY 4.0
- Carleman approximants for the center system.svg by Actinopterigio, CC BY 4.0
Fathomly is not affiliated with or endorsed by the Wikimedia Foundation. Spotted a problem? Tell us.