ArticleslgStudy

mathematics

Mercer's theorem

Mercer's theorem is a mathematics topic covered in the lgStudy science library. This page brings together a partial reference excerpt, illustrations, worked examples, real-world applications and a short study plan, so you can understand Mercer's theorem rather than just read about it. In short: In mathematics, specifically functional analysis, Mercer's theorem is a representation of a symmetric positive-definite function on a square as a sum of a convergent sequence of product functions. This theorem, presented in (Mercer 1909), is one of the most notable results of the work of James Mercer (1883–1932).

Key takeaways

  • Mercer's theorem belongs to mathematics; place it in that map before memorising details.
  • Learn the definition first, then one example that makes the definition concrete.
  • Connect Mercer's theorem to a quantity you can measure, compute or draw — that is where exam questions come from.
  • Reproduce the core statement of Mercer's theorem from memory before moving on to harder problems.

Reference excerpt

In mathematics, specifically functional analysis, Mercer's theorem is a representation of a symmetric positive-definite function on a square as a sum of a convergent sequence of product functions. This theorem, presented in (Mercer 1909), is one of the most notable results of the work of James Mercer (1883–1932). It is an important theoretical tool in the theory of integral equations; it is used in the Hilbert space theory of stochastic processes, for example the Karhunen–Loève theorem; and it is also used in the reproducing kernel Hilbert space theory where it characterizes a symmetric positive-definite kernel as a reproducing kernel.

Introduction To explain Mercer's theorem, we first consider an important special case; see below for a more general formulation. A kernel, in this context, is a symmetric continuous function

K : [ a , b ] × [ a , b ] → R {\displaystyle K:[a,b]\times [a,b]\rightarrow \mathbb {R} }

where K ( x , y ) = K ( y , x ) {\displaystyle K(x,y)=K(y,x)} for all x , y ∈ [ a , b ] {\displaystyle x,y\in [a,b]} . K is said to be a positive-definite kernel if and only if

∑ i = 1 n ∑ j = 1 n K ( x i , x j ) c i c j ≥ 0 {\displaystyle \sum _{i=1}^{n}\sum _{j=1}^{n}K(x_{i},x_{j})c_{i}c_{j}\geq 0}

for all finite sequences of points x1, ..., xn of [a, b] and all choices of real numbers c1, ..., cn. Note that the term "positive-definite" is well-established in literature despite the weak inequality in the definition. The fundamental characterization of stationary positive-definite kernels (where K ( x , y ) = K ( x − y ) {\displaystyle K(x,y)=K(x-y)} ) is given by Bochner's theorem. It states that a continuous function K ( x − y ) {\displaystyle K(x-y)} is positive-definite if and only if it can be expressed as the Fourier transform of a finite non-negative measure μ {\displaystyle \mu } :

K ( x − y ) = ∫ − ∞ ∞ e i ( x − y ) ω d μ ( ω ) {\displaystyle K(x-y)=\int _{-\infty }^{\infty }e^{i(x-y)\omega }\,d\mu (\omega )}

This spectral representation reveals the connection between positive definiteness and harmonic analysis, providing a stronger and more direct characterization of positive definiteness than the abstract definition in terms of inequalities when the kernel is stationary, e.g., when it can be expressed as a 1-variable function of the distance between points rather than the 2-variable function of the positions of pairs of points. Associated to K is a linear operator (more specifically a Hilbert–Schmidt integral operator when the interval is compact) on functions defined by the integral

[ T K φ ] ( x ) = ∫ a b K ( x , s ) φ ( s ) d s . {\displaystyle [T_{K}\varphi ](x)=\int _{a}^{b}K(x,s)\varphi (s)\,ds.}

We assume φ {\displaystyle \varphi } can range through the space of real-valued square-integrable functions L2[a, b]; however, in many cases the associated reproducing kernel Hilbert space can be strictly larger than L2[a, b]. Since TK is a linear operator, the eigenvalues and eigenfunctions of TK exist. Theorem. Suppose K is a continuous symmetric positive-definite kernel. Then there is an orthonormal basis {ei}i of L2[a, b] consisting of eigenfunctions of TK such that the corresponding sequence of eigenvalues {λi}i is nonnegative. The eigenfunctions corresponding to non-zero eigenvalues are continuous on [a, b] and K has the representation

K ( s , t ) = ∑ j = 1 ∞ λ j e j ( s ) e j ( t ) {\displaystyle K(s,t)=\sum _{j=1}^{\infty }\lambda _{j}\,e_{j}(s)\,e_{j}(t)}

where the convergence is absolute and uniform.

Details We now explain in greater detail the structure of the proof of Mercer's theorem, particularly how it relates to spectral theory of compact operators.

… excerpt ends here. Continue reading the full article.

Worked examples

Example 1 — a first encounter with Mercer's theorem

Start with the simplest possible case. Write down what Mercer's theorem claims or describes in one sentence, then invent the smallest concrete situation in which that sentence is true. In mathematics, the smallest case is usually a single object, a single equation or a single measurement. Check that every symbol or term in your sentence has a meaning in that case.

Example 2 — changing one variable

Take the situation from Example 1 and change exactly one quantity: double it, halve it, or set it to zero. Predict what should happen to Mercer's theorem before you calculate. Comparing your prediction with the result is the fastest way to find out whether you understand the idea or only the words.

Example 3 — an exam-style question

Typical questions about Mercer's theorem ask you to (a) state it precisely, (b) apply it to given data, and (c) explain a limitation. Practise writing all three answers in under five minutes; the third part is what separates a full-mark answer from an average one.

Applications of Mercer's theorem

In research
Mercer's theorem appears in mathematics research whenever the underlying quantities have to be modelled precisely. Papers usually cite it as a starting assumption and then explore where it breaks down.
In technology and industry
Engineering practice reuses Mercer's theorem in design rules, simulations and safety margins. Knowing the idea lets you read a specification sheet and understand why the numbers look the way they do.
In the classroom
Mercer's theorem is common in secondary-school and first-year university syllabi. It links to neighbouring topics Theorems in functional analysis, so understanding it makes those chapters shorter.
In everyday life
Look for Mercer's theorem outside the textbook — in sport, cooking, traffic, electronics or the sky above you. An example you found yourself is remembered far longer than one you were given.

Affiliate

Preply — study more efficiently by working with a personal tutor. 50% off.

How to study Mercer's theorem in 20 minutes

  1. Read the reference excerpt below once, without taking notes.
  2. Close the page and write down what Mercer's theorem means in your own words.
  3. Compare your version with the excerpt and mark what you missed.
  4. Work through the three examples above with pen and paper.
  5. Explain Mercer's theorem out loud to somebody else — or to Teacher Smith in the lgStudy chat.

Frequently asked questions

What is Mercer's theorem in simple terms?

In mathematics, specifically functional analysis, Mercer's theorem is a representation of a symmetric positive-definite function on a square as a sum of a convergent sequence of product functions. This theorem, presented in (Mercer 1909), is one of the most notable results of the work of James Merc…

Why does Mercer's theorem matter?

Because it connects several mathematics ideas at once: it gives you a definition you can apply, a quantity you can calculate, and a way to check whether a result is plausible.

How should I study Mercer's theorem?

Read the excerpt, restate it from memory, then work through the examples and applications listed on this page. The five-step study plan above takes about twenty minutes.

What does this page cover?

It gives you a compact reference excerpt plus original lgStudy explanations, examples, applications and study material on Mercer's theorem.

Tags

  • Theorems in functional analysis

Keep exploring