Characteristic Functions, Moments, Cumulants, and Generating Functionals
A characteristic function exists for every probability law and determines that law. A moment-generating function is more restrictive: it is useful only where the relevant exponential moment is finite. Whenever differentiation is justified, derivatives of the generator give moments, while derivatives of its normalized logarithm give cumulants. In several variables the same construction gives joint cumulants; with a finite Euclidean source, it also packages connected correlations and nonlinear response.
Those statements have different hypotheses, and keeping them separate is the main task of this page. A finite list of moments need not determine a law, a characteristic function can have no useful derivatives, and a formal Lorentzian path integral is not a probability measure. The logarithm is also only local unless zeros and branches have been controlled.
Required background. Probability Spaces, Random Variables, and Conditional Expectation supplies expectation, independence, integrability criteria, and dominated convergence.
Helpful background. Fourier Series, Fourier Transforms, and Plancherel Theory supplies the transform convention and inversion language. Readers needing a shorter review can use the statistical ensembles and probability repair or the Fourier, distributions, and Green functions repair.
Characteristic, moment, cumulant, and generating functions
Section titled “Characteristic, moment, cumulant, and generating functions”Let be an -valued random vector with law . We use the characteristic-function convention
The sign is explicit because the inverse transform below uses . Four related objects should not be conflated:
| Object | Definition | Domain | What it encodes |
|---|---|---|---|
| Characteristic function | Every real | The complete probability law | |
| Moment-generating function | Only the set where the expectation is finite | Moments and exponential tilts, under additional hypotheses | |
| Cumulant-generating function | Real points where is finite; a neighborhood of is an extra hypothesis | Joint cumulants and additive structure | |
| Generating functional | or | Depends on the field, test space, and source | Smeared moments or cumulants through directional derivatives |
The first row is automatic because . The other rows require domain, differentiability, normalization, and sometimes infinite-dimensional existence arguments.
Characteristic functions determine probability laws
Section titled “Characteristic functions determine probability laws”Several basic facts follow directly from the definition:
They require no moment assumption. Characteristic functions are uniformly continuous, since
by dominated convergence, with a bound independent of .
They also obey a positivity condition. For any points and coefficients ,
Thus every characteristic function is normalized, continuous, and positive definite. These properties are valuable diagnostics, although uniqueness of the underlying probability law requires the inversion result below.
Inversion and uniqueness
Section titled “Inversion and uniqueness”For a real-valued , the inversion theorem states that at continuity points of the distribution function,
The quotient at is interpreted by continuity. Without the continuity assumption, the right-hand side is
Thus endpoint atoms contribute half their mass. The formula proves the uniqueness result
This is the inversion theorem in Durrett’s Durrett 2019, § 3.3.1, pp. 128–30, PDF.
When , the law has the bounded continuous density
This is a sufficient condition, not a characterization of all distributions with densities.
Characteristic functions also connect this page to the preceding treatment of weak convergence. Lévy’s continuity theorem says that implies pointwise convergence . Conversely, a pointwise limit of characteristic functions is a weak limit when the limiting function is continuous at the origin. The last condition is essential: for ,
and the discontinuity records probability mass escaping to infinity. The two directions and the continuity-at-zero condition are stated in Durrett’s Durrett 2019, Theorem 3.3.17, p. 132, PDF. The full convergence framework belongs to Probabilistic Convergence, Laws of Large Numbers, and Central Limit Theorems.
Moments come from derivatives only under hypotheses
Section titled “Moments come from derivatives only under hypotheses”Let be a multi-index and . If
then dominated differentiation gives
and hence
The integrable dominating function is . This is a clean sufficient condition; it is not a license to infer every absolute moment from the mere existence of a derivative at zero. Durrett’s Durrett 2019, Theorem 3.3.18, p. 134, PDF gives the scalar dominated-differentiation statement.
The Cauchy law makes the boundary visible. Its characteristic function is
which exists for all but is not differentiable at the origin. The mean does not exist, and for every .
The domain of the moment-generating function
Section titled “The domain of the moment-generating function”Define
The set contains and is convex by Hölder’s inequality, but can lie on its boundary. If is finite on an open neighborhood of , then all absolute moments exist, is analytic on a possibly smaller neighborhood, and
In that case the local MGF determines the law. None of these conclusions follows from a formal Taylor series alone. McCullagh’s McCullagh 2017, § 2.2.1, pp. 29–30, PDF explicitly separates a convergent MGF from a finite-order or divergent formal moment expansion.
The lognormal law supplies two distinct warnings. Let with . Then
so every nonnegative integer moment is finite, but for every . Moreover, if is the lognormal density, then
defines a family of distinct densities with the same nonnegative integer moments. Indeed,
Thus “all moments exist,” “the MGF exists near zero,” and “the moments determine the law” are three different statements. Durrett gives the lognormal moment-indeterminacy construction in Durrett 2019, § 3.3.5, p. 140, PDF.
Cumulants are logarithmic coordinates
Section titled “Cumulants are logarithmic coordinates”Suppose first that is finite on a neighborhood of . Because for real , define
The joint cumulant tensor is
One can instead work with the characteristic function. Continuity and give a zero-free ball around the origin. On that ball choose the local logarithm satisfying . If , then the dominated-differentiation result applies through order , and for ,
If such log derivatives happen to exist without the corresponding absolute moments, they are derivative-defined coefficients; they should not be silently identified with the moment cumulants below.
The word local matters. For a symmetric Rademacher variable, , so the characteristic function has zeros and no single global logarithm.
For a scalar variable with raw moments , coefficient comparison in gives
Equivalently, if and ,
The third and fourth cumulants are not themselves the standardized skewness and excess kurtosis; those divide by appropriate powers of .
The partition formula and connected blocks
Section titled “The partition formula and connected blocks”Let be the set of partitions of . If every product appearing below is integrable, the joint cumulant has the finite-order definition
Möbius inversion on the partition lattice gives the inverse identity
This formula is the precise algebraic meaning of a full correlation being a sum of products of connected blocks. It can define a cumulant through a fixed finite order even when no neighborhood-valued MGF exists: expand the finite polynomial
and read the coefficient of in its formal logarithm. This is finite-order algebra, not a claim that an infinite moment series converges. McCullagh derives both partition identities and this finite-order interpretation in McCullagh 2017, § 2.3.4, p. 38, PDF.
Direct partition enumeration has combinatorial cost equal to the Bell number . For low orders this makes every subtraction transparent; for high orders, recursive differentiation or symbolic partition algorithms are less error-prone.
Independence, joint structure, and the Gaussian boundary
Section titled “Independence, joint structure, and the Gaussian boundary”For random vectors and , independence is equivalent to factorization of the joint characteristic function:
for all . The reverse implication uses uniqueness of characteristic functions. If two nonempty subfamilies of the arguments of a joint cumulant are independent, that mixed cumulant vanishes. One vanishing covariance, or any finite list of vanishing mixed cumulants, does not generally prove independence.
For independent scalar variables and ,
on their common local domain. Consequently,
Affine transformations obey
Thus, for iid variables with the required cumulants,
Bazant 2006, pp. 3–6 develops the same cumulant additivity and iid scaling from characteristic functions.
This scaling explains why higher standardized cumulants are suppressed in classical central-limit regimes. More precisely, if and , then
This scaling does not replace the hypotheses of a central limit theorem.
A Gaussian vector with mean and covariance has
Its first two cumulants are and , and every higher cumulant vanishes. Conversely, an actual cumulant-generating function that is quadratic on a neighborhood of the origin determines a Gaussian law. A finite list of zero higher cumulants does not: the distribution
has mean zero, variance one, , and , but is discrete. The full Gaussian characterization, Gaussian processes, random distributions, and Wick structure belong to Gaussian Vectors, Processes, Random Distributions, and Wick Structure.
A reliable calculation procedure
Section titled “A reliable calculation procedure”The method applies to probability laws, finite regulated Euclidean integrals, and characteristic or exponential functionals whose domains have been established.
- Choose the generator. Use a characteristic function when law determination or weak convergence is the goal. Use an MGF or Euclidean source generator only after checking exponential integrability.
- Normalize and state the domain. Verify or divide an unnormalized Euclidean source integral by its value at zero. Record the neighborhood in which it is finite and nonzero.
- Justify every derivative. Supply an integrable dominating function, analyticity from exponential integrability, or a finite-order algebraic definition. Stop at the highest justified order.
- Take the logarithm locally. Fix the branch by . Derivatives of the generator give full moments; derivatives of the logarithm give cumulants.
- Validate independently. Check normalization, conjugation symmetry or positive definiteness, covariance positivity, independence factorization, and at least one direct moment calculation.
The output is a law, a collection of moments or connected cumulants, or a source-response relation. Differentiation is usually inexpensive at low order; converting all order- moments and cumulants by partitions grows with the Bell number. The calculation must stop when the proposed source is outside the finite domain, the needed moment is absent, a logarithm crosses a zero, normalization fails, or a formal functional has not been given a mathematical definition.
Source tilting turns cumulants into responses
Section titled “Source tilting turns cumulants into responses”Let have an MGF finite on an open set. Write
For in the interior of the finite domain, define the tilted probability measure
Differentiating gives
Higher derivatives are the joint cumulants in the tilted measure. The first derivative is the sourced mean response; the Hessian is the linear response and is positive semidefinite. Therefore is convex wherever this differentiation is valid. Fithian, n.d., “Differential identities” derives the gradient-as-mean and Hessian-as-covariance formulas on the interior of the finite natural-parameter domain. Positivity of the Hessian is a useful independent check on signs and source conventions.
For a random distribution acting on test functions , the characteristic functional
exists for every admissible . The exponential functional still needs exponential integrability. When directional derivatives are justified,
The smeared variables are essential. A continuum field may be a distribution rather than a pointwise random variable, so unsmeared functional derivatives need a separate kernel or distributional argument.
Euclidean and Lorentzian conventions
Section titled “Euclidean and Lorentzian conventions”For a positive finite-dimensional Euclidean weight, this page uses
Then . At , derivatives of give unsourced full Euclidean correlations, and derivatives of give their cumulants. At nonzero , the same derivatives give correlations in the tilted measure.
The site’s Lorentzian convention is instead
Writing , the corresponding formal translation is
These factors belong to the displayed source convention; they must not be copied into the Euclidean probability formulas. Nor does the Lorentzian symbol by itself construct a probability measure. Vacuum normalization, ordering, and the developed physical interpretation are handled in Connected Correlators and Cumulants. For the displayed convention and the full-versus-connected derivative formulas, see Schwartz 2014, § 14.3, pp. 261–63 and § 34.1.2, pp. 737–39.
Worked application: a normalized two-mode Euclidean source
Section titled “Worked application: a normalized two-mode Euclidean source”Consider the finite regulated field with
The inequalities make positive definite, so the weight is integrable and positive. For a real source , define
Set . Completing the square gives
with
Translation invariance of the ordinary Lebesgue integral yields
After normalization,
This is the finite positive Euclidean counterpart of the free quadratic source completion in Schwartz 2014, § 14.3, pp. 261–263, with normalization and all signs stated locally here.
The source responses are therefore
At zero source,
and every cumulant of order three or higher vanishes because is quadratic.
The fourth derivative of , however, is not zero. For component indices ,
The three terms are exactly the three two-block partitions of four centered variables; the connected four-point cumulant is the fourth derivative of and vanishes. This calculation shows why the logarithm selects the one-block contribution without yet invoking a diagrammatic linked-cluster theorem.
There are several independent checks:
- , so the source object is normalized.
- , and , as a covariance matrix must be.
- The response equation follows both from completing the square and from differentiating .
- Directly differentiating four times reproduces the three pairings above.
The stop condition is equally explicit. As , one eigenvalue of approaches zero, the covariance diverges, and ceases to define a normalizable probability measure at the boundary. Interacting continuum fields require a regulator, existence arguments, and renormalization beyond this example.
Common pitfalls and stop rules
Section titled “Common pitfalls and stop rules”Calling every transform an MGF. A characteristic function always exists; an MGF may be finite only at zero or on a one-sided domain. State which exponential and which domain are being used.
Differentiating before checking integrability. A formal derivative of an expectation is not a moment theorem. Supply domination or local exponential integrability, and stop at the highest justified order.
Replacing an actual function by its formal moment series. The lognormal example has all integer moments but no two-sided MGF neighborhood, and its moments do not determine its law. Formal coefficient identities remain finite-order algebra only.
Taking a global logarithm without checking zeros. The branch normalized by always exists near the origin, but a characteristic function can vanish elsewhere. Never continue the cumulant logarithm through a zero without a separate branch analysis.
Equating zero covariance with independence. Independence kills every mixed cumulant that straddles independent groups, but a single vanishing mixed cumulant proves very little. Use factorization of the full joint characteristic function for an exact criterion.
Declaring Gaussianity from a few cumulants. Vanishing third and fourth cumulants is not enough. Use the full Gaussian characteristic function or a quadratic generator on a domain that determines the law.
Ignoring normalization or signature. Divide a Euclidean source integral by its zero-source value before reading it as an MGF. Translate the factors of when moving to the Lorentzian convention, and do not call an oscillatory path integral a probability measure.
Using point fields when only smeared fields exist. In an infinite-dimensional theory, establish the test-function space and the continuity of the functional before interpreting functional derivatives as correlation kernels.
Check your understanding
Section titled “Check your understanding”1. Find a local logarithm that cannot be global
Section titled “1. Find a local logarithm that cannot be global”Let take the values with equal probability. Compute its characteristic function, MGF, and first four cumulants. Why can the characteristic logarithm not be defined globally on the real line?
Solution
Direct averaging gives
Near zero,
Since ,
The characteristic function vanishes at , so a logarithm normalized at the origin cannot be continued as one finite branch through all real .
2. Track cumulants of an iid average
Section titled “2. Track cumulants of an iid average”Suppose are iid and have cumulants through order . Derive the cumulants of .
Solution
Independence first gives
Scaling by then gives
For the mean is unchanged, for the variance falls as , and higher unstandardized cumulants fall faster. The calculation uses independence and the existence of the stated cumulants.
3. Separate a centered four-point function
Section titled “3. Separate a centered four-point function”For centered variables , use the partition identity to express in terms of second and fourth cumulants.
Solution
Every partition containing a singleton contributes a factor . The surviving partitions are the one-block partition and the three pair partitions, so
In the Gaussian example the fourth cumulant is zero, leaving the three pairings. For a non-Gaussian law the one-block term need not vanish.
4. Translate the Lorentzian source factors
Section titled “4. Translate the Lorentzian source factors”Under the convention , compute the first two derivatives of and identify the connected two-point function.
Solution
At a general source,
Differentiating again gives
Therefore, at zero source,
The minus- is the case of ; it would be absent in the positive Euclidean convention used in the worked example.
Synthesis and continuations
Section titled “Synthesis and continuations”Use a characteristic function when an always-defined, law-determining transform is required. Use an MGF only on its finite domain, and take moments from derivatives only after justifying differentiation. Normalize before taking a logarithm: the derivatives of the original generator give full moments, while derivatives of the logarithm give the one-block cumulants. Independence turns products into sums and makes cumulants additive. A source then upgrades the same algebra into response theory, with the Hessian of the Euclidean log generator equal to a positive-semidefinite covariance.
Continue to Gaussian Vectors, Processes, Random Distributions, and Wick Structure for the full Gaussian and infinite-dimensional theory. Continue to Connected Correlators and Cumulants for vacuum normalization, physical source derivatives, and the linked-cluster interpretation in QFT.
References
Section titled “References”-
Martin Z. Bazant, “Lecture 2: Moments, Cumulants, and Scaling”, MIT OpenCourseWare 18.366, Fall 2006, pp. 3–6. This is the teaching source for the progression from characteristic functions to tensor moments, low-order cumulants, and additivity under independent sums.
-
Rick Durrett, Probability: Theory and Examples, fifth edition, PDF, Cambridge University Press, 2019. This is the structural probability source. § 3.3.1, pp. 125–31, supports the definition, basic properties, inversion, and uniqueness of characteristic functions; § 3.3.2, pp. 132–33, supports Lévy’s continuity theorem; § 3.3.3, pp. 134–36, supports moment differentiation; and § 3.3.5, pp. 140–43, supports the moment-problem and lognormal counterexamples.
-
William Fithian, “Exponential Families”, undated Statistics 210A course notes, University of California, Berkeley, accessed August 11, 2026. The sections “Exponential family structure” and “Differential identities” support the finite natural-parameter domain, exponential tilting, and the gradient-as-mean and Hessian-as-covariance identities.
-
Peter McCullagh, Tensor Methods in Statistics, Dover edition, 2017. This is the structural cumulant source. Chapter 2, §§ 2.2–2.7, pp. 29–44, supports the distinction between convergent and finite-order cumulant expansions, the moment–cumulant set-partition formulas, Gaussian cumulants, mixed cumulants across independent blocks, and additivity for independent sums.
-
Matthew D. Schwartz, Quantum Field Theory and the Standard Model, Cambridge University Press, 2014. Section 14.3, pp. 261–63, supports normalized source derivatives and the free quadratic generator; § 34.1.2, pp. 737–39, supports and the separation of full and connected Green functions.