Two conditions have to hold at once before a decay asymmetry can exist at all, and arranging for both is why this measurement took thirty years.
🎯 Why this matters
A non-zero ε′ moved CP violation out of the states and into the interaction itself. Mixing alone could have been an accident of how two particles happen to overlap; an asymmetry between decay amplitudes cannot be.§8.5 established CP violation in the mixing: the propagating states are not CP eigenstates, and the impurity is . This section asks the independent question — do the decay amplitudes themselves differ between a process and its CP conjugate?
The answer is yes, by a further factor of a thousand, and extracting it took from 1970 to 2002.
Two conditions, and both are necessary
Start with the general statement, because it governs §8.9 and §8.10 as well. Suppose a decay proceeds through two amplitudes. Each carries a weak phase, which flips sign under CP, and a strong phase weak phase vs strong phase weak phases change sign under CP; strong phases do not. A rate asymmetry needs both kinds of phase difference: it goes as sin(δ_{S1} − δ_{S2}) sin(φ_{W1} − φ_{W2}), so either one alone gives nothing. defined in §8.8 — open in glossary from final-state interactions, which does not:
Bettini p. 342. The only difference between the two lines is the sign of the weak phases. Everything follows from that.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
Subtract the two rates and almost everything cancels:
The master formula for a direct CP asymmetry. Two sines multiplied together — and the whole of experimental CP violation is the problem of making neither of them zero.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
A product of two sines. Set either to zero and there is no asymmetry, however large the other is. That is the section’s central claim, and it is worth playing with rather than reading:
A rate asymmetry needs two phase differences, not one
Both present. Flipping the weak phase reflects A₂ about the direction of A₁, and because the two are not collinear the reflected chain closes somewhere else. The rate difference is −4|A₁||A₂| sin(δ_S1−δ_S2) sin(δ_W1−δ_W2) = -1.6164, matching |A|² − |Ā|² = -1.6164.
⚙️ Engineer’s bridge — a phase is only observable against a reference, and CP has to move one and not the other
The reason both conditions are needed is not an accident of the algebra. It is the same reason a single phasor has no measurable phase at all.
An overall phase on an amplitude is unobservable — for any . To see a phase you need two amplitudes and their interference, which measures only the difference. That is the first condition: one amplitude gives you nothing.
Now ask what CP does. It flips the weak phases and leaves the strong ones. So the CP conjugate rate differs only if flipping the weak phases changes the relative angle between the two phasors — and if the two phasors are collinear to begin with, reflecting one about the other’s direction puts it back where it was. Hence the second condition.
An engineer will recognise the structure as a lock-in detection with a reference channel that must be in quadrature. A lock-in multiplies your signal by a reference and low-passes; if the signal and reference are exactly in phase the output is insensitive to small phase shifts, and if they are 90° apart it is maximally sensitive. The strong phase is that quadrature: it is nature’s reference channel, and is exactly the quadrature factor.
Where it breaks: the lock-in analogy assumes you supply the reference. Here nature supplies it, and you cannot calculate it. A strong phase you cannot compute sits directly in front of the quantity you want. is proportional to , and that number comes from pion scattering data, not from the CKM matrix. Direct CP violation is therefore never a clean CKM measurement in the way was (§8.6) — which is exactly why the chapter’s cleanest result came from interference rather than from decay.
The kaon case: two isospin amplitudes
For the two amplitudes are supplied by isospin. The two pions are in a spatially symmetric state, so their isospin wave function must be symmetric too: or , never 1. Call the corresponding weak amplitudes and . Then (Eq. 8.67)
The two physical decays written in terms of two isospin amplitudes. This is the change of basis that supplies the 'two paths to one final state' the master formula above demands.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
and the conjugate amplitudes are the same with — which is CPT, not an assumption about the decay.
Experimentally dominates: , the ΔI = 1/2 rule. That single number can be checked against the measured rate ratio, and doing so turns up something:
🔢 Worked example — the K_S rate ratio, from Eq. (8.67) and two numbers
Square the two amplitudes above, keeping the interference term:
With and , and multiplying by the phase-space ratio :
Now the measurement. The PDG branching ratios are and , so the ratio is
Agreement to 0.7 % — which simultaneously confirms the ΔI = 1/2 rule and the strong phase difference, neither of which was used to obtain the branching ratios. Note that pure alone would give only 1.97; the 14 % excess is the admixture.
Erratum — Eq. (8.68) prints 2.55 where 2.255 belongs
The book gives
Four independent checks say this is a transposed digit for 2.255:
- The PDG branching ratios the book cites give 2.2548. .
- The quoted uncertainty fits 2.255 and not 2.55. Propagating on each branching ratio gives — the printed . On a value of 2.55 that would be a 0.2 % measurement of nothing in particular.
- The book’s own Eq. (8.67) predicts 2.24, using the book’s own and . That is 0.7 % from 2.255 and 12 % from 2.55.
- 2.55 would require , contradicting the sentence immediately before the equation.
Nothing downstream uses the number, so no conclusion changes. But it is quoted precisely as evidence that the final state is nearly pure , and at 2.55 that evidence points the wrong way.
ε′, and why it is so hard
With the Wu–Yang convention ( real and positive), the two observable amplitude ratios η₊₋ and η₀₀ η₊₋ and η₀₀ the amplitude ratios A(K_L → ππ)/A(K_S → ππ) for the charged and neutral pion pairs. Equal to each other if there is no direct CP violation, which is why their difference is the observable. defined in §8.8 — open in glossary come out as, together with the direct-violation parameter ε′ ε′ the direct-CP-violating parameter of the kaon system. Only ever quoted as the ratio Re(ε′/ε) = (1.66 ± 0.23) × 10⁻³ — direct violation is a further thousand times smaller than indirect. defined in §8.8 — open in glossary ,
The two observable ratios, and the parameter that separates CP violation in the decay from CP violation in the mixing. All three conditions of the master formula are visible in the last expression.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
Read carefully — it is zero unless is both non-zero and non-real. Non-zero is the ΔI = 1/2 rule being imperfect; non-real is a weak phase. And the is the strong phase that makes the real part observable. All three conditions of the bridge, in one expression.
The trouble is the size. If CP violation lived only in the mixing then exactly, so the search is for a difference between two nearly equal numbers:
💡 What this really says — why the observable is a double ratio and not anything simpler
The natural thing to measure would be against — a decay and its conjugate, differing by . You cannot: by the time anything decays, the propagating states are and , not and .
So the observable has to be built from and rates. Take the ratio of to for each, then the ratio of those — the double ratio double ratio [Γ(K_L→π⁺π⁻)/Γ(K_S→π⁺π⁻)] ÷ [Γ(K_L→π⁰π⁰)/Γ(K_S→π⁰π⁰)] = 1 + 6 Re(ε′/ε). Constructed so that everything except direct CP violation cancels. defined in §8.8 — open in glossary :
The double ratio — the most systematics-driven observable in the book. Every element of its design exists so that something cancels.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
Every source of systematic error that is common to a pair cancels in that pair’s ratio, and everything common to the two pairs cancels again in the double ratio. Incident kaon flux, detector acceptance for a given topology, trigger efficiency — all gone.
That is why the experiments were built the way §8.8 describes: simultaneous detection of both final states so the flux cancels, and and beams present at the same time with matched energy spectra and matched decay distributions so the acceptance cancels. Every element of the design exists to make a cancellation exact.
The engineering instinct is the same one behind a Wheatstone bridge or a four-terminal resistance measurement: when the quantity you want is a small difference between two large ones, do not measure the two and subtract. Build a configuration in which everything you do not want cancels before the measurement, and read out only the residual.
Bettini p. 341. Four decay rates, arranged so that everything except direct CP violation cancels.
Every symbol, one at a time
Hover or tap a symbol above — it lights up in the equation and its meaning, units and type appear here.
from the measured double ratio to ε′
eta_ratio, eps = 0.9950, 2.232e-3
d = 1 - eta_ratio**2
print("the double ratio and what it gives:")
print(f" |eta_00 / eta_+-| = {eta_ratio} +- 0.0007")
print(f" 1 - |eta_00/eta_+-|^2 = {d:.3e}")
print(f" Re(eps'/eps) = (1/6) x {d:.3e} = {d/6:.3e}")
print( " book (8.78): (1.66 +- 0.23)e-03 -- exact agreement")
print("\nso how big is direct CP violation, absolutely?")
print(f" |eps| = {eps:.3e} (violation in the MIXING)")
print(f" Re(eps'/eps) = {d/6:.3e}")
print(f" |eps'| ~ {eps*d/6:.2e} (violation in the DECAY)")
print("\nthree levels, each about a thousand times smaller than the last:")
print( " CP-conserving amplitude 1")
print(f" mixing violation |eps| {eps:.1e}")
print(f" decay violation |eps'| {eps*d/6:.1e}") the double ratio and what it gives: |eta_00 / eta_+-| = 0.995 +- 0.0007 1 - |eta_00/eta_+-|^2 = 9.975e-03 Re(eps'/eps) = (1/6) x 9.975e-03 = 1.662e-03 book (8.78): (1.66 +- 0.23)e-03 -- exact agreement so how big is direct CP violation, absolutely? |eps| = 2.232e-03 (violation in the MIXING) Re(eps'/eps) = 1.662e-03 |eps'| ~ 3.71e-06 (violation in the DECAY) three levels, each about a thousand times smaller than the last: CP-conserving amplitude 1 mixing violation |eps| 2.2e-03 decay violation |eps'| 3.7e-06
Thirty years, and a standoff
The measurement is worth telling as history because it shows what a effect on top of a effect actually costs.
| experiment↕ | year↕ | result↕ | ↕ |
|---|---|---|---|
| NA31 (CERN) | 1993 | 2.30 ± 0.65 | 3.5σ from zero — a claim of discovery, on 428 000 decays |
| E731 (FNAL) | 1993 | −0.74 ± 0.56 | similar statistics, consistent with zero. The two disagreed with each other at 3.5σ |
| KTeV (FNAL) | 2003 | 2.07 ± 0.28 | decays and much better systematics |
| NA48 (CERN) | 2002 | 1.47 ± 0.22 | the same, independently — and the effect was settled |
the 1993 standoff, quantified
import numpy as np
m = [('NA31', 2.30, 0.65), ('E731', -0.74, 0.56), ('KTeV', 2.07, 0.28), ('NA48', 1.47, 0.22)]
print("NA31 vs E731, 1993:")
d, sd = 2.30 - (-0.74), np.hypot(0.65, 0.56)
print(f" 2.30 +- 0.65 against -0.74 +- 0.56")
print(f" difference {d:.2f} +- {sd:.3f} -> {d/sd:.1f} sigma apart")
print( " one claimed a 3.5 sigma discovery, the other was consistent with zero.")
print("\nKTeV vs NA48, the next generation:")
d2, sd2 = 2.07 - 1.47, np.hypot(0.28, 0.22)
print(f" 2.07 +- 0.28 against 1.47 +- 0.22")
print(f" difference {d2:.2f} +- {sd2:.3f} -> {d2/sd2:.1f} sigma apart")
print( ' the book says "the two values agree"; 1.7 sigma is marginal but fair.')
wsum = sum(v/s**2 for _, v, s in m); w = sum(1/s**2 for _, v, s in m)
avg, err = wsum/w, 1/np.sqrt(w)
chi2 = sum(((v-avg)/s)**2 for _, v, s in m)
print(f"\nnaive weighted average of all four: {avg:.2f} +- {err:.2f}")
print(f" chi^2 = {chi2:.1f} for 3 dof, so the four are NOT mutually consistent;")
print( " the PDG inflates the error for exactly this reason, reaching")
print( " book (8.78): 1.66 +- 0.23") NA31 vs E731, 1993: 2.30 +- 0.65 against -0.74 +- 0.56 difference 3.04 +- 0.858 -> 3.5 sigma apart one claimed a 3.5 sigma discovery, the other was consistent with zero. KTeV vs NA48, the next generation: 2.07 +- 0.28 against 1.47 +- 0.22 difference 0.60 +- 0.356 -> 1.7 sigma apart the book says "the two values agree"; 1.7 sigma is marginal but fair. naive weighted average of all four: 1.54 +- 0.16 chi^2 = 21.6 for 3 dof, so the four are NOT mutually consistent; the PDG inflates the error for exactly this reason, reaching book (8.78): 1.66 +- 0.23
Three experimental obstacles made it that hard, and all three are named in the book:
- has BR , and sits under , which is 200 times more frequent and gives the same photons plus two more. Losing two photons out of six turns the background into the signal.
- The two kaons have wildly different decay lengths. At 110 GeV, km against m — yet the two decay distributions inside the detector must be made as similar as possible, or the acceptance does not cancel.
- Everything must be simultaneous. Both final states at once so the flux cancels; both beams at once with matched spectra so the acceptance does.
Erratum — Eq. (8.81) is missing a factor of 2
The book prints
The coefficient is −4. It is pure algebra from Eq. (8.80): the cross term in already carries a factor 2, and
supplies another. .
Checked numerically over 20 000 random choices of the two magnitudes and four phases: with the maximum residual is ; with as printed it is — i.e. the printed relation simply fails.
Nothing in the argument depends on it. The point of Eq. (8.81) is the product of two sines, and that is right as printed; only the overall coefficient is wrong, and the equation is never used numerically.
Supplied — the table above has the four numbers and the disagreement only becomes obvious when they share an axis. Two experiments of comparable statistics, published the same year, disagreeing at 3.5σ: one claiming a discovery, the other consistent with no direct CP violation at all. Neither was wrong about its own data; the gap was systematics, which is what a 10⁻³ effect measured on top of a 10⁻³ effect costs. The next generation, with ~10⁷ π⁰π⁰ decays instead of 4×10⁵, agrees with itself to 1.7σ and excludes zero decisively — and note that the settled value sits closer to NA31 than to E731, so the discovery claim was right and the disagreement was still real.
🔑 If you remember only three things
-
It is a part in a thousand of a part in a thousand. ε′ sits on top of an effect that was itself barely visible, and the smallness is the whole difficulty.
-
Two experiments disagreed for a decade and neither was careless. That is what a double ratio’s systematics produce when the signal is this small.
-
The kaon happens to have the structure the question needs. Two amplitudes with different strong phases are required, and not every system offers them.
Where this goes next
CP violation in the decay is now established in the kaon at — direct violation being a thousand times smaller again than the indirect violation of §8.5.
The B mesons make it much easier. Their masses open hundreds of channels, so individual branching ratios are and asymmetries of several per cent are within reach of pairs — no double ratio required, because the rates of two charge-conjugate decays can simply be compared.
§8.9 applies all of this to charm, the only up-type system available, where CP violation was not seen until 2019 and the observable is a difference of two asymmetries — the double-ratio trick one level up. §8.10 then does the charged B, where there is no mixing at all and the payoff is the unitarity-triangle angle from tree diagrams alone.
✅ Check yourself — CP violation in the decay
0/6 answered · 0 correct
1.A rate asymmetry between a decay and its CP conjugate requires two phase differences. Why is one not enough?
2.ε′ ∝ Im A₂ · e^{i(δ₂−δ₀)}. What does each factor require?
3.Why is the observable a double ratio of four decay rates rather than a single comparison of a decay with its conjugate?
4.Eq. (8.68) gives Γ(K_S→π⁺π⁻)/Γ(K_S→π⁰π⁰) = 2.55 ± 0.005. What is wrong?
5.In 1993 NA31 reported (2.30 ± 0.65) × 10⁻³ and E731 reported (−0.74 ± 0.56) × 10⁻³. What should one have concluded?
6.Why is K_L → π⁰π⁰ so much harder than K_L → π⁺π⁻?