The Fubini–Tonelli theorem
Statement
Let \((X,\mathcal{A},\mu)\) and \((Y,\mathcal{B},\nu)\) be \(\sigma\)-finite measure spaces, and let \(\mathcal{A}\otimes\mathcal{B}\) denote the product \(\sigma\)-algebra on \(X\times Y\), with \(\mu\otimes\nu\) the unique product measure on \((X\times Y,\mathcal{A}\otimes\mathcal{B})\) satisfying \((\mu\otimes\nu)(A\times B)=\mu(A)\nu(B)\) for all \(A\in\mathcal{A}\), \(B\in\mathcal{B}\). (Tonelli.) If \(f:X\times Y\to[0,\infty]\) is \(\mathcal{A}\otimes\mathcal{B}\)-measurable, then the maps \(x\mapsto\int_Y f(x,y)\,d\nu(y)\) and \(y\mapsto\int_X f(x,y)\,d\mu(x)\) are measurable (with values in \([0,\infty]\)), and \[\int_{X\times Y} f\,d(\mu\otimes\nu)=\int_X\left(\int_Y f(x,y)\,d\nu(y)\right)d\mu(x)=\int_Y\left(\int_X f(x,y)\,d\mu(x)\right)d\nu(y),\] all three quantities being equal in \([0,\infty]\) with no integrability hypothesis. (Fubini.) If instead \(f:X\times Y\to\mathbb{R}\) (or \(\mathbb{C}\)) is \(\mathcal{A}\otimes\mathcal{B}\)-measurable and \(\int_{X\times Y}|f|\,d(\mu\otimes\nu)\lt\infty\), then \(f(x,\cdot)\) is \(\nu\)-integrable for \(\mu\)-a.e. \(x\), \(f(\cdot,y)\) is \(\mu\)-integrable for \(\nu\)-a.e. \(y\), the a.e.-defined functions \(x\mapsto\int_Y f(x,y)\,d\nu(y)\) and \(y\mapsto\int_X f(x,y)\,d\mu(x)\) are integrable, and the same iterated-integral equality holds, now as an equality of finite real (or complex) numbers.
Why it matters
Iterated integration is the only computationally tractable way to evaluate most multivariable integrals — one reduces a genuinely two-dimensional problem to two successive one-dimensional problems. But an iterated integral is defined regardless of whether the double integral it purports to compute even exists as a well-defined quantity, and the two iterated integrals \(\int_X\int_Y\) and \(\int_Y\int_X\) can disagree with each other and with the double integral when the hypotheses fail. Fubini–Tonelli identifies exactly the circumstances — measurability of \(f\) on the product, plus either non-negativity (Tonelli) or absolute integrability (Fubini) — under which all three quantities coincide, licensing the routine swap of integration order used throughout analysis, probability, and PDE theory.
The theorem also underlies the entire machinery of independent random variables: the joint law of an independent pair is a product measure, and Fubini–Tonelli is precisely the statement that expectations of joint functionals may be computed by conditioning on one coordinate at a time.
Hypotheses
Proof
Result
Reading. On a product of \(\sigma\)-finite measure spaces, a jointly measurable function's double integral can be computed by integrating one variable at a time, in either order, and getting the same answer — provided either the function is non-negative (Tonelli, no other conditions needed) or its absolute value has finite double integral (Fubini). Tonelli is typically used first, on \(|f|\), to check the finiteness hypothesis that then licenses Fubini for \(f\) itself.
Scope. Applies to any pair of \(\sigma\)-finite measure spaces — Lebesgue measure on \(\mathbb{R}^n\), counting measure on a countable set (recovering rearrangement theorems for double series), or probability spaces (recovering independence and joint-expectation formulas). Does not by itself extend to non-\(\sigma\)-finite spaces or to integrands that are merely separately measurable in each variable rather than jointly \(\mathcal{A}\otimes\mathcal{B}\)-measurable.
Corollaries & converses
- Order of summation for double series. If \(a_{mn}\geq0\), or if \(\sum_{m,n}|a_{mn}|\lt\infty\), then \(\sum_m\sum_n a_{mn}=\sum_n\sum_m a_{mn}\) — the special case \(X=Y=\mathbb{N}\) with counting measure.
- Differentiation under the integral sign and interchange of integral and infinite sum are routinely justified by recasting the sum as an integral against counting measure and invoking Fubini–Tonelli jointly with dominated convergence.
- Independence and product expectation. If \(X,Y\) are independent random variables with finite \(\mathbb{E}|g(X)h(Y)|\), then \(\mathbb{E}[g(X)h(Y)]=\mathbb{E}[g(X)]\,\mathbb{E}[h(Y)]\), by applying Fubini to the product measure \(\mu_X\otimes\mu_Y\), which is the joint law precisely because of independence.
- Convolution. For \(f,g\in L^1(\mathbb{R}^n)\), \((f*g)(x)=\int f(x-y)g(y)\,dy\) is defined for a.e.\ \(x\), lies in \(L^1\), and \(\|f*g\|_1\leq\|f\|_1\|g\|_1\); this is Fubini–Tonelli applied to \(F(x,y)=f(x-y)g(y)\).
- Converse: does equality of the two iterated integrals imply the Fubini/Tonelli hypotheses hold? No. The two iterated integrals can coincide "by accident" even when \(f\) is not integrable, or (under CH, as in the Hypotheses section) even when \(f\) is not jointly measurable at all. Equality of iterated integrals is necessary but nowhere near sufficient for the theorem's conclusion, so it can never be used to retroactively certify the hypotheses.
Fails without
- Drop \(\sigma\)-finiteness: with \(\mu=\) Lebesgue on \([0,1]\) and \(\nu=\) counting measure on \([0,1]\) (not \(\sigma\)-finite), the indicator of the diagonal gives iterated integrals \(1\) and \(0\) respectively — see the Hypotheses section for the full computation.
- Drop joint measurability: a CH-constructed subset of \([0,1]^2\) with co-countable vertical sections and countable horizontal sections yields an \(f\) whose iterated integrals are \(1\) and \(0\), yet \(f\) is not \(\mathcal{A}\otimes\mathcal{B}\)-measurable, so it lies outside the theorem's scope entirely.
- Drop absolute integrability in the signed case: \(f(x,y)=\dfrac{x^2-y^2}{(x^2+y^2)^2}\) on \((0,1)\times(0,1)\) has \(\int_0^1\!\int_0^1 f\,dy\,dx=\tfrac{\pi}{4}\) but \(\int_0^1\!\int_0^1 f\,dx\,dy=-\tfrac{\pi}{4}\); indeed \(\int\!\int|f|=\infty\), so Fubini's hypothesis fails and the two orders genuinely disagree.
- Drop non-negativity without absolute integrability (Tonelli misapplied to a signed function): the row/column-alternating array \(a_{mn}\) of the Hypotheses section shows a signed, non-integrable \(f\) for which the two iterated sums are unequal, confirming Tonelli's conclusion cannot be invoked once \(f\) takes both signs and \(|f|\) is not integrable.
Common errors
- Swapping the order of integration first and only checking (or never checking) integrability afterwards — the standard fallacy is to compute both iterated integrals, find they agree, and conclude Fubini "applies"; agreement of iterated integrals does not certify the hypotheses (see Corollaries).
- Invoking "Fubini's theorem" for a manifestly non-negative integrand and then separately worrying about absolute convergence — for \(f\geq0\) one should cite Tonelli, which needs no integrability hypothesis at all; conflating the two names obscures which hypothesis is actually doing the work.
- Checking integrability of \(f\) itself (i.e., that \(\int f\) is finite) rather than of \(|f|\) — a function can have well-defined, finite, and even equal iterated integrals of \(f\) while \(\int\int|f|=\infty\), and Fubini's conclusion can still fail in such cases (the classical \(f(x,y)=(x^2-y^2)/(x^2+y^2)^2\)-type examples).
- Forgetting the a.e. qualifiers: assuming \(x\mapsto\int_Y f(x,y)\,d\nu(y)\) is defined for every \(x\), when Fubini only guarantees this for \(\mu\)-a.e. \(x\); on the exceptional null set the inner integral may not exist (both \(f^+\) and \(f^-\) sections may be non-integrable there).
- Applying the theorem to a function known only to be separately measurable in each variable (measurable in \(x\) for each fixed \(y\), and vice versa) rather than jointly \(\mathcal{A}\otimes\mathcal{B}\)-measurable; separate measurability does not imply joint measurability in general (it does under extra continuity assumptions, e.g. Carathéodory functions, but not unconditionally).
Discussion
The theorem is really two theorems bundled under one name for good pedagogical reason: Tonelli handles the "easy" non-negative case with no integrability hypothesis whatsoever — only measurability and \(\sigma\)-finiteness — because Monotone Convergence tolerates \(\infty\) gracefully, whereas Fubini's signed/complex case must additionally guard against the \(\infty-\infty\) indeterminacy that plagues subtraction of infinite quantities. The standard proof strategy, visible in the steps above, is the "climb the ladder of function classes" technique common throughout measure theory: establish a result for indicators, extend to simple functions by linearity, extend to non-negative measurable functions by monotone approximation, then extend to signed/complex functions by decomposition — precisely the same ladder used to define the Lebesgue integral itself.
Guido Fubini proved the integrable case in 1907 for Lebesgue measure on \(\mathbb{R}^n\); Leonida Tonelli gave the non-negative version in 1909, freeing the theorem from any a priori integrability assumption. The abstract measure-space formulation used today, built via Carathéodory extension of the product premeasure, is due to the general development of measure theory in the following decades and is the version needed for probability theory, where \(X\) and \(Y\) are typically infinite or continuous sample spaces.
In probability, the theorem is the analytic engine behind the definition of independence itself: two random variables are independent exactly when their joint law is the product of their marginal laws, and Fubini's theorem is what allows "integrate out one variable, then the other" computations of joint expectations, convolutions of independent sums, and conditional expectation formulas via disintegration. The \(\sigma\)-finiteness hypothesis is automatic for any probability space (total mass \(1\)), which is one reason probabilists rarely dwell on it, but it is essential and is exactly what fails for, e.g., certain non-\(\sigma\)-finite Palm measures in point-process theory.
A subtler point often elided in first courses: the product \(\sigma\)-algebra \(\mathcal{A}\otimes\mathcal{B}\) is in general strictly smaller than the Borel \(\sigma\)-algebra of the product topology, and also strictly smaller than the completed product \(\sigma\)-algebra used to build Lebesgue measure on \(\mathbb{R}^{m+n}\) from Lebesgue measure on \(\mathbb{R}^m\) and \(\mathbb{R}^n\) separately. The completed version of Fubini's theorem — needed to identify Lebesgue measure on \(\mathbb{R}^{m+n}\) with the completion of \(\mathcal{L}(\mathbb{R}^m)\otimes\mathcal{L}(\mathbb{R}^n)\) — requires an extra null-set bookkeeping step (sections of a \(\mu\otimes\nu\)-null set are null for a.e. fixed coordinate, but not literally every section is null), which is routine but is a genuine addition to the argument given here, not a free consequence of it.
Worked examples
Reading. Introducing the extra variable \(y\) via \(1/x=\int_0^\infty e^{-xy}dy\) converts a transcendental integral into a Gaussian-type double integral that becomes tractable after reversing the order — a standard "Feynman trick" whose legitimacy rests entirely on Fubini–Tonelli.
Reading. By symmetry \(P(X\lt Y)=P(Y\lt X)\) and \(P(X=Y)=0\) (a null set under a jointly continuous density, itself an instance of Tonelli applied to the diagonal), so the answer \(1/2\) is exactly what symmetry demands — Fubini–Tonelli is what makes the double integral over the "wedge" region \(\{x\lt y\}\) computable as a clean iterated integral in the first place.
Problems
- State precisely which of Tonelli's or Fubini's theorem applies to \(f(x,y)=xye^{-(x^2+y^2)}\) on \(\mathbb{R}^2\), and evaluate \(\int_{\mathbb{R}^2}f\,d(\mu\otimes\mu)\) (Lebesgue \(\times\) Lebesgue).
Solution
\(f\) changes sign (positive in quadrants 1 and 3, negative in 2 and 4), so Tonelli does not directly apply to \(f\); check \(\int\int|f|\): \(|f(x,y)|=|x||y|e^{-(x^2+y^2)}\geq0\), and by Tonelli, \(\int_{\mathbb{R}^2}|f|=\left(\int_{\mathbb{R}}|x|e^{-x^2}dx\right)^2=\left(2\int_0^\infty xe^{-x^2}dx\right)^2=(2\cdot\tfrac12)^2=1\lt\infty\). Since \(\int\int|f|\lt\infty\), Fubini applies to \(f\) itself: \(\int_{\mathbb{R}^2}f\,d(\mu\otimes\mu)=\left(\int_{\mathbb{R}}xe^{-x^2}dx\right)^2\). But \(\int_{\mathbb{R}}xe^{-x^2}dx=0\) (odd integrand over symmetric interval), so the answer is \(0\). - Let \(\mu=\nu=\) Lebesgue measure on \((0,1)\). Show that \(f(x,y)=\dfrac{1}{(x+y)^{3}}\) has \(\int_0^1\int_0^1 f\,dx\,dy=\infty\) directly, and explain why this is consistent with (not a counterexample to) Tonelli.
Solution
Compute \(\int_0^1\frac{dx}{(x+y)^3}=\left[-\tfrac12(x+y)^{-2}\right]_0^1=\tfrac12\left(\tfrac1{y^2}-\tfrac1{(1+y)^2}\right)\). Then \(\int_0^1\tfrac1{2y^2}dy=\infty\) since \(1/y^2\) is non-integrable near \(0\), while \(\int_0^1\tfrac{1}{2(1+y)^2}dy\) is finite; hence the outer integral diverges to \(+\infty\). Since \(f\geq0\) is jointly measurable and \(\mu,\nu\) are \(\sigma\)-finite, Tonelli guarantees the double integral and both iterated integrals are all equal in \([0,\infty]\) — here they are all equal to \(+\infty\). This is fully consistent with Tonelli: the theorem never claims finiteness, only equality (in the extended sense), and \(\infty=\infty=\infty\) is a valid instance of it. - Construct, using Fubini's theorem, a proof that if \(X\geq0\) is a random variable, then \(\mathbb{E}[X]=\int_0^\infty P(X\gt t)\,dt\) (the "layer-cake" / tail-integral formula).
Solution
Write \(X(\omega)=\int_0^\infty \mathbf{1}_{\{t\lt X(\omega)\}}\,dt\) (true for every \(\omega\), since the inner integral is just the length of \([0,X(\omega))\)). Then \(\mathbb{E}[X]=\int_\Omega\left(\int_0^\infty \mathbf{1}_{\{t\lt X(\omega)\}}\,dt\right)dP(\omega)\). The integrand \(g(\omega,t)=\mathbf{1}_{\{t\lt X(\omega)\}}\) is non-negative and jointly measurable (as \(\{(\omega,t):t\lt X(\omega)\}\) is measurable in the product \(\sigma\)-algebra \(\mathcal{F}\otimes\mathcal{B}((0,\infty))\), being the preimage of an open set under the jointly measurable map \((\omega,t)\mapsto X(\omega)-t\)), and \((\Omega,\mathcal{F},P)\) and \(((0,\infty),\mathcal{B},\text{Leb})\) are both \(\sigma\)-finite (indeed \(P\) is finite). By Tonelli, the order of integration may be swapped: \(\mathbb{E}[X]=\int_0^\infty\left(\int_\Omega \mathbf{1}_{\{t\lt X(\omega)\}}\,dP(\omega)\right)dt=\int_0^\infty P(X\gt t)\,dt\). - Give an example, distinct from the ones in this article, of a non-\(\sigma\)-finite measure space pairing for which the two iterated integrals of a non-negative jointly measurable function disagree, and identify exactly which step of the proof breaks down.
Solution
Take \(X=Y=[0,1]\), \(\mu=\) Lebesgue measure (\(\sigma\)-finite, indeed finite), \(\nu=\) counting measure on \([0,1]\) (not \(\sigma\)-finite, since \([0,1]\) is uncountable and every set of finite \(\nu\)-measure is finite/countable, so no countable union of finite-\(\nu\)-measure sets covers \([0,1]\)). Let \(f=\mathbf{1}_D\) where \(D=\{(x,x):x\in[0,1]\}\) is the diagonal, which is closed hence Borel, hence jointly measurable. For fixed \(x\), \(\int_Y f(x,y)\,d\nu(y)=\nu(\{x\})=1\), so \(\int_X\int_Y f\,d\nu\,d\mu=\int_0^1 1\,dx=1\). For fixed \(y\), \(\int_X f(x,y)\,d\mu(x)=\mu(\{y\})=0\), so \(\int_Y\int_X f\,d\mu\,d\nu=\sum_{y\in[0,1]}0\cdot(\text{as a }\nu\text{-integral of the zero function})=0\). The failure is precisely in Step 5 of the proof: the passage from Step 4 (valid only for finite \(\mu,\nu\)) to general \(\sigma\)-finite \(\mu,\nu\) requires writing \(Y\) as a countable increasing union \(Y_n\uparrow Y\) of finite-\(\nu\)-measure sets and invoking Monotone Convergence; with \(\nu\) counting measure on an uncountable set, no such exhaustion exists, so Step 5's Monotone Convergence argument cannot be run, and the section formula \((\mu\otimes\nu)(E)=\int_X\nu(E_x)\,d\mu(x)\) itself fails to hold (indeed \(\mu\otimes\nu\) is not even uniquely determined by its values on rectangles here). - Let \(f(x,y)=\dfrac{x^2-y^2}{(x^2+y^2)^2}\) on \((0,1)\times(0,1)\). Verify \(\int_0^1\int_0^1 f\,dy\,dx=\tfrac{\pi}{4}\) and \(\int_0^1\int_0^1 f\,dx\,dy=-\tfrac{\pi}{4}\), and confirm \(\int\int|f|=\infty\) so this is not a violation of Fubini's theorem.
Solution
Note \(\dfrac{\partial}{\partial y}\left(\dfrac{-y}{x^2+y^2}\right)=\dfrac{-(x^2+y^2)+y\cdot2y}{(x^2+y^2)^2}=\dfrac{y^2-x^2}{(x^2+y^2)^2}=-f(x,y)\), so \(\int_0^1 f(x,y)\,dy=\left[\dfrac{-y}{x^2+y^2}\right]_{y=0}^{1}\cdot(-1)=\dfrac{1}{x^2+1}\) (careful sign bookkeeping: \(\int_0^1 f\,dy=\left[\frac{y}{x^2+y^2}\right]_0^1=\frac{1}{x^2+1}\)). Then \(\int_0^1\frac{dx}{x^2+1}=\arctan(1)-\arctan(0)=\pi/4\). By the antisymmetry \(f(x,y)=-f(y,x)\), swapping the names of the variables shows \(\int_0^1\int_0^1 f\,dx\,dy=-\pi/4\) immediately, without repeating the computation. For the divergence of \(\int\int|f|\): along the ray \(y=x\), \(f(x,x)=0\), but near the origin \(|f(x,y)|\) is comparable to \(1/r^2\) in polar coordinates \(r^2=x^2+y^2\) (since \(f\) is homogeneous of degree \(-2\) and not identically zero on the unit circle), and \(\int_0^\epsilon \frac{1}{r^2}\cdot r\,dr\,d\theta=\int_0^\epsilon\frac{dr}{r}=\infty\); a full estimate confirms \(\iint_{(0,1)^2}|f|\,dx\,dy=\infty\). Since Fubini's absolute-integrability hypothesis fails, the unequal iterated integrals \(\pi/4\neq-\pi/4\) are fully consistent with (not a counterexample to) the theorem as stated.