These notes start from one sharp question: a qubit’s state is parameterized by two complex numbers, so if you know , is determined? Following the logic of that question leads through an entire stretch of the single-qubit map: why phase is real physical information, how it takes physical form in polarized light, how qubits relate to hardware, the three properties of the Pauli matrices, rotation gates and circuit notation, and finally the formal mathematics of measurement. The first section reviews the prerequisites in one line apiece; everything after that is self-contained.
0. Prerequisites at a glance
Only the following facts are needed, each in one sentence:
- A complex number is an arrow: length times direction, . Here is the point on the unit circle at angle ; its modulus is always 1, so multiplying by it changes direction but never length.
- A qubit’s state: , where are complex numbers (called amplitudes), also writable as the column vector .
- The Born rule: measuring in the 0/1 basis gives 0 with probability and 1 with probability ; the normalization condition is .
- Bras and inner products: , obtained from the ket by transposing and conjugating (the combined operation is written , dagger); is a row times a column, and the result is one complex number.
- The standard superposition states: , orthogonal to each other.
- The Bloch sphere: the map of single-qubit states. , where the latitude governs probabilities and the longitude governs phase; two orthogonal states sit at antipodal points (directly opposite), not at right angles.
1. The seed question: knowing , can you determine ?
The answer splits in two. Normalization pins down the length:
But it cannot pin down the direction. There are infinitely many complex numbers of length , a full circle of them:
This free is the phase, and it is real, measurable physical information. A counterexample is the most convincing argument: fix (so is also fixed) and change only :
| State | Position on the Bloch sphere | |
|---|---|---|
| axis | ||
| axis | ||
| axis |
Same , same , yet different means entirely different points on the equator. and are even orthogonal to each other, two states that can be distinguished perfectly. Measure in another basis, or pass the state through a gate, and the outcomes differ completely.
In the two-arrows picture: and are each an arrow (length plus direction). Physically there is no absolute phase, just as there is no absolute position; saying ” points at ” means nothing without a reference, and the only reference in sight is the other arrow, . So exactly one quantity carries physical meaning: the angle between the two arrows.
- Rotate both together (global phase): the angle is unchanged, no experiment can detect it, it means nothing;
- Rotate only one (relative phase): the angle changes, and both interference and measurements in other bases can see it.
An analogy: tilt the whole clock by and the angle between the hour and minute hands is unchanged, so the time is unchanged; move only the minute hand by and the time has changed. The convention “take real” is hanging the clock straight: point at , and all phase information compresses into a single angle .
The bookkeeping of degrees of freedom now matches the Bloch sphere: two complex numbers carry 4 real parameters, normalization eats one, global phase eats another, and 2 remain, exactly . Knowing gives you only the latitude; the longitude is an independent second piece of information.
2. Where the phase hides: why
A common confusion: “I don’t see any phase in this formula.” Answer: the phase has been there all along, packed inside the complex numbers. Every complex number carries a direction by birth, so writing actually writes down four pieces of information:
The lengths govern probabilities; the directions are the phases. The notation is the packed form (easy to write and compute with); the Bloch form is the unpacked form (the redundant global phase already discarded, the one physical phase laid out in the open). They are two notations for the same thing.
Why is “a complex two-dimensional unit vector” the right container for a state? Because it is the least common multiple of three experimental facts:
- Quantum states superpose (diagonal polarization really exists), so states need an additive structure: linear combinations of basis states;
- Quantum states interfere (two paths can cancel), so coefficients must carry direction. Real numbers cannot cancel naturally; complex numbers do it exactly: ;
- Measurement yields probabilities, so the squared modulus of a coefficient serves as probability, is precisely “the vector has length 1”, and since quantum operations are all length-preserving rotations, conservation of probability comes for free.
Addable, direction-carrying, squared-modulus-as-probability: exactly one object satisfies all three at once, the unit-length complex two-dimensional vector.
3. The physical incarnation of phase: polarized light
3.1 What light is, what polarization is
Light is an electromagnetic wave: an electric field oscillating as it travels, with the oscillation perpendicular to the direction of travel. The rope analogy: the rope stretches away from you (the direction of travel), and your hand can shake it up and down, side to side, or at a slant; the wave runs forward either way, but the direction of shaking differs. That direction is the polarization: horizontal oscillation is , vertical is . Looking into the beam, the tip of the electric field traces a line segment back and forth, and the tilt of that line is the polarization angle.
3.2 45° polarization = vector decomposition = superposition demystified
What is the electric field of light oscillating at ? Middle-school vector decomposition:
that is, an equal horizontal component plus an equal vertical component. Translated into quantum notation (with ):
“Superposition state” sounds mystical, but for polarization it is just vector decomposition. The field is really and simultaneously composed of a horizontal oscillation and a vertical one, and the coefficient is not an abstract symbol but the length of a geometric projection. This also answers whether is “secretly already H or V”: neither. It is one definite vector pointing at , and asking “but is it really horizontal or vertical” is like asking “is northeast really east or north”. The question itself is malformed.
3.3 Arbitrary angles and Malus’s law
A photon polarized at physical angle :
Example: at polarization, and . Note that and are the same polarization (the field oscillates back and forth anyway), so the distinct linear polarizations live in .
A polarizer passes only the component along its transmission axis. Classical optics gives Malus’s law (1809): transmitted intensity . That and the Born rule’s are the same one: at the single-photon level a photon is indivisible, so it either passes whole (with probability ) or is absorbed whole, and classical light is the statistical average over myriad photons. Malus’s law is the classical shadow of the Born rule; it is just that in 1809 nobody knew photons were hiding in the denominator.
3.4 The “angle doubling” mystery, resolved physically
Compare with the Bloch formula : the physical polarization angle plays exactly the role of :
| Physical angle | State | Position on the Bloch sphere |
|---|---|---|
| north pole | ||
| equator, | ||
| south pole | ||
| equator, | ||
| again | back to the north pole |
Half a physical turn sweeps a full circle on the sphere: one-to-one, nothing repeated, nothing missed. The design decision “why ” acquires a visible incarnation in polarization: H and V differ physically by only , yet they are perfectly distinguishable (orthogonal), so the map must draw them as the two poles.
3.5 Circular polarization: the physical identity of
So far the field has oscillated along a line. If the H component and the V component fall out of step, offset by a quarter of a period, the field tip no longer traces a line but draws a circle: circular polarization. “Delayed by a quarter period” translates into complex language as multiplication by (since , a turn):
These are the two ends of the Bloch sphere’s axis (and the eigenvectors of Pauli ; see §5). The sphere is now completely filled by polarization: the – great circle holds all linear polarizations (real coefficients), the two ends of the axis are left and right circular polarization, and every remaining point is elliptical. Here ” means rotate by ” stops being a metaphor: it is literally a quarter-period delay, implemented in the lab by a quarter-wave plate.
3.6 An experiment you can do by hand: three polarizers
- Light passes an H polarizer, so everything that exits is ;
- Add a V polarizer after it: , total darkness, as expected;
- Insert a polarizer between them, and light comes through (about ).
More obstacles, more light? Because the middle sheet is a measurement: hits the sheet and passes with probability , and having passed, its state is projected onto ; then hits the V sheet, another . Total . This is the tabletop demonstration of “measuring in the wrong basis changes the state”. Quantum key distribution (the BB84 protocol) catches eavesdroppers with exactly this law of physics, and three lenses from polarized sunglasses let you watch it happen.
4. A qubit is not a photon: three layers of abstraction
A natural misconception: “qubit = photon.” The best correction starts with the classical question: what is a bit made of? In a CPU it is a voltage, on a hard drive the orientation of a magnetic domain, on an optical disc a pit, on a punched card a hole. The carriers could hardly differ more, and algorithms care about none of it. “Bit” is an abstract unit of information; the voltages are merely its physical implementations.
A qubit is the same story, quantum version. Its definition: any two-level quantum system. Whenever there are two mutually orthogonal (perfectly distinguishable) quantum states that can be prepared, manipulated, and measured, you have a qubit. are logical labels that deliberately say nothing about hardware. The correct picture has three layers:
- The logical layer: the qubit itself. , algorithms, protocols, and all the mathematics in these notes live here;
- The degree-of-freedom layer: the particular pair of orthogonal states chosen to carry the information — a photon’s polarization, an electron’s spin (), an atom’s energy levels (ground/excited), a superconducting circuit’s current states;
- The carrier layer: the physical entities, photons, electrons, atoms, circuits. One entity can offer several degrees of freedom, hence several potential qubits.
Strictly, even “a photon’s polarization is a qubit” is half a step off; the right sentence is “a photon’s polarization implements a qubit”, just as “the voltage on this wire implements a bit”. This layering is why quantum computing works as an engineering discipline: theorists design algorithms at the logical layer, experimentalists swap hardware at the carrier layer, and the interface contract is “two orthogonal states, unitary operations, measurement”.
Two practical footnotes. First, labels collide: a single photon has polarization, path, arrival time, photon number, several degrees of freedom, each encodable. In polarization encoding ; in photon-number encoding is the vacuum (no photon present) and means one photon. Identical symbols on paper, entirely different physics. The first thing to do with an experimental paper is find the authors’ definition of the encoding. Labels are cheap; always ask what physical states they name. Second, for a spin- particle the Bloch sphere almost stops being an abstract map: the arrow’s direction on the sphere really is the direction the spin magnetic moment points in laboratory space. That is also the context in which the sphere was invented (Felix Bloch, studying spins in nuclear magnetic resonance).
5. Pauli matrices: one set of matrices, three properties
The three protagonists:
They are gates: they act by matrix-times-vector. swaps the two components (, the quantum NOT); flips the sign of the second component; does both and throws in an . Their geometric identity: each is a rotation of the Bloch sphere about the , , or axis respectively.
5.1 Property one: self-inverse
“Doing it twice equals doing nothing”: rotating twice is a full turn, just like classical NOT. It looks trivial and is the hidden workhorse in several places: it forces the eigenvalues to satisfy (see 5.3); it collapses the rotation operator’s power series into and (see §6); and it will appear again in the proof of the uncertainty principle (see 5.4).
5.2 Property two: both Hermitian and unitary — two credentials
Two definitions, both built from :
- Unitary: . These matrices preserve length (, provable in two lines). Physical meaning: states must stay normalized and probabilities must keep summing to 1, so the legal quantum gates are exactly the unitary matrices — the license to be a gate.
- Hermitian: . These matrices have all-real eigenvalues. Physical meaning: measurement readouts must be real numbers, so observables are represented by Hermitian matrices — the license to be an instrument (details in §8).
The two credentials usually cannot be held at once (a rotation gate is unitary but generally not Hermitian, so it can only be a gate). The Paulis hold both, and the reason is property one: Hermitian () plus self-inverse () immediately gives . Consequence: the same matrix is a gate inside a circuit box and an instrument in the sentence “measure ”. Beginners trip hardest over “measuring a matrix”; it refers to the second identity.
5.3 Eigenvectors, and eigenvalues
Definition: if — the matrix leaves the direction unchanged and only multiplies by a number — then is an eigenvector of and the factor is its eigenvalue. Plainly: eigenvectors are the directions the matrix cannot move, and the eigenvalue is the only thing it can do in that direction (stretch or flip).
For the Paulis the factor takes only two values. Verify :
swaps components: the two components of are equal, so swapping changes nothing (); picks up an overall minus sign (). The complete dictionary:
| Pauli | Eigenvalue | Eigenvalue | Axis |
|---|---|---|---|
| (the poles) | |||
Why the factor can only be : self-inverse forces ; Hermitian forces real. The list is locked.
Why the eigenvectors sit exactly on the corresponding axis: a Pauli gate is a rotation about its axis. When a globe spins about its north–south axis, the only points that do not move are the two on the axis, the north and south poles. “Direction unchanged under the action” equals “position fixed under the rotation” equals being on the rotation axis. So each row’s eigenvectors are precisely the two endpoints of that axis, and an old rule gets confirmed in passing: the two ends of any diameter are a pair of orthogonal (perfectly distinguishable) states.
The two physical identities of (matching the matrix’s two roles):
- As a gate, is a phase. : on alone the is a global phase, no motion on the sphere; but inside a superposition, the two components pick up different factors ( versus ), which is a relative phase: , a turn along the equator, in agreement with ” = half-turn about the axis”.
- As an instrument, are the dial readings. The measurement rules (§8) say readouts must be eigenvalues, so “measuring ” outputs or : reading means the state collapsed to , reading means it collapsed to . Choosing rather than has a convenience: the expectation value expresses “which way the state leans” in a single number.
5.4 Property three: cyclic non-commutation — the algebraic fingerprint of quantum weirdness
The commutator is defined as
It measures how much the order of operations matters: means order is irrelevant (they commute); otherwise order counts. Ordinary numbers always commute; matrices generally do not. Compute one by hand:
The three relations rotate cyclically (advance the letters ; remember one, turn the crank twice):
The structure is identical to the vector cross product , and that is no coincidence: it is the algebra of three-dimensional rotations. And rotations are inherently non-commuting: take your phone, flip it about the horizontal axis then about the vertical axis, and compare with the reverse order; the final orientations differ. Pauli non-commutation is that everyday geometric fact in matrix form.
The main event: non-commutation = the uncertainty principle. The notebook version of the statement is that the observables which fail to commute are exactly the ones that cannot be known simultaneously, and it can be proven in five lines with tools you now fully possess.
Step one, translate “known simultaneously” into mathematics: “the state has a definite value of observable ” means the outcome of measuring is 100% certain, which means the state is an eigenvector of . So “definite and definite at once” means a common eigenvector exists.
Step two, prove none exists, by contradiction. Suppose a nonzero satisfies both and (with ). Then
(numbers commute, so the two terms cancel). But , hence , that is, . And property one says : is invertible, and an invertible matrix sends only the zero vector to zero, so , contradiction.
Conclusion: and share no eigenvector at all. Not “hard to know both at once”: there exists no state that is definite for both. All three properties appear within the five lines: the commutator supplies , self-inverseness makes invertible, and “definite value = eigenvector” rests on Hermiticity.
Step three, the physical face. is an eigenvector of (its value is definite), but measure : — maximal uncertainty. On the Bloch sphere it is obvious: knowing means the arrow lies on the axis, knowing means it lies on the axis, and one arrow cannot lie on two perpendicular axes at once — a seesaw whose two ends cannot both be up. One more layer deserves puncturing: this is not “the value exists but we are ignorant of it”. For a qubit in , the value is simply not defined (like asking whether northeast is really east or north); and this is experimentally distinguishable from classical ignorance — distinguishing exactly that is what Bell-inequality experiments do.
Step four, recognize an old acquaintance. Heisenberg’s starts from the same algebra: . The quantitative general form (the Robertson inequality):
The product of two uncertainties is bounded below by the commutator. Classical physics commutes everywhere, so everything can be known at once; all of quantum “weirdness” grows out of the single algebraic fact that matrix multiplication does not commute. A cryptographic application: BB84’s two encoding bases are precisely the eigenvectors of and , and means no common eigenvector, which means no measurement an eavesdropper can make reads both bases reliably — the five-line mathematical backbone of the protocol’s security.
6. Rotation gates: Euler’s formula in matrix form
Pauli gates are “violent” ( teleports the north pole to the south pole). To move continuously on the Bloch sphere, feed a Pauli into the matrix exponential:
where is the Pauli of the rotation axis (rotating about uses , and so on) and is the angle you want to turn on the sphere. The left side is the definition, the right side the computed closed form; the step between them is this section’s main course.
6.1 The matrix exponential: an old formula with a widened domain
The Taylor series uses only multiplication and addition, both of which matrices can do. So define directly:
6.2 The series splits itself into two piles
Substitute ; the -th term is . Two cycles run simultaneously: the powers of cycle with period four (), and the powers of cycle with period two ( enters: ). So every even term carries and every odd term carries , and the series groups into two piles:
The two brackets are exactly the Taylor series of and . Closed form obtained. The derivation parallels the derivation of Euler’s formula word for word: there splits the series, here does the same job:
| Euler’s formula | Rotation operator | |
|---|---|---|
| Engine that splits the series | ||
| Result | ||
| What rotates | an arrow in the complex plane | the state arrow on the Bloch sphere |
One observation runs deeper: — a square equal to , which is the matrix version of the defining property of (namely ). are three different “matrix versions of ”, one per rotation axis; the rotation formula is Euler’s formula replayed with a new engine.
6.3 Verification and use
Sanity checks at special angles: gives (no rotation) ✓; gives , which up to a global phase is the Pauli gate itself, so “Pauli = rotation about its axis” turns from slogan into corollary ✓; gives (see 6.4).
The explicit matrix of (running Euler’s formula backwards to fold the diagonal entries into exponentials):
(the last step factors out a global phase and discards it). adds to the relative phase and leaves the probabilities untouched — it is the knob that turns the angle between the two arrows; changing probabilities takes or . Verify one more:
North pole to equator ✓. Unitarity in one line: dagger turns into , so , hence — “rotating back” is the inverse.
Why : the sphere runs at double speed (third appearance in these notes). The matrix lives in state space, the sphere is a map, map angle = state-space angle × 2, so to turn the map by the matrix must contain . Why the minus sign: a convention fixing the positive rotation direction (right-hand rule), inherited from the solution of the Schrödinger equation, .
6.4 A full 360° turn is not “doing nothing”
Substitute : . The arrow on the sphere returns home, but the state has been multiplied by ; to truly return home mathematically takes , that is, .
An apparent contradiction arrives: isn’t a global phase, the thing that cannot be measured? The resolution is exactly §1’s principle: multiplying the whole state means nothing; multiplying one branch means everything. The trick: do not rotate the entire system, rotate only one branch of a superposition. Implement it with an interferometer — split one particle into a superposition of two paths and rotate the spin by on the lower path only:
The hangs on one branch only and becomes a relative phase of : at recombination the interference flips from constructive to destructive, the fringes invert, and the detectors see it. This is no thought experiment: in 1975 the groups of Rauch and of Werner each did it with neutron interferometers — a silicon crystal splits the neutron beam, a magnetic field precesses the spin on one arm, the fringes cycle with period in the rotation angle, and a turn lands exactly in antiphase. All spin- particles (electrons, protons, neutrons) behave this way; the technical term is spinor.
A shadow of it exists in daily life: the plate trick — hold a plate palm-up and rotate your arm , and your arm ends up twisted; continue another in the same direction and the arm untwists by itself, total to return to the start. The mathematical root is a group-theoretic fact: is the double cover of the rotation group — every rotation of the sphere corresponds to two matrices , and that is the shared ID card behind the mystery, the phenomenon, and the plate trick.
6.5 The hardware identity
In the lab this formula is not abstract notation: hit a superconducting qubit with a microwave pulse, or an ion with a laser, and the hardware is solving the Schrödinger evolution ; when the Hamiltonian is proportional to some Pauli, the evolution is exactly , and is proportional to the pulse duration. “Calibrating a gate” is, to a large extent, tuning the pulse length until hits the target angle.
7. Circuit diagrams and the unitary world
7.1 Circuit notation: time reads left to right, matrices right to left
Three rules for reading a circuit: each horizontal wire is one qubit (its time axis, not a spatial wire); each box is a gate; left to right = earlier to later. The trap is converting to matrices: applying matrices is function composition, the first gate hugs and later gates wrap around the outside, so a matrix string reads right to left:
Walk through it and see how fatal reading the wrong way is:
- Correct ( first): , then . Output .
- Reversed ( first): , then (as is the eigenvector of , it does not move). Output .
and are orthogonal — as wrong as wrong gets. The algebraic reason order matters: and do not commute. The mnemonic: the state is fed into the matrix string from the right, and whatever it hits first acts first. ( is the Hadamard gate, : the standard tool for creating superposition and realizing interference.)
7.2 The complete roster of gates = the unitary matrices
Fact: the legal single-qubit operations are exactly the unitary matrices. “Exactly” is a two-way promise. (⟹) They must be unitary: probability has to survive, and the unitaries are all of the length-preservers. (⟸) Unitary suffices: any unitary decomposes into three rotations (), and rotations are just pulses. The roster is neither more nor less.
7.3 Closed versus open: where unitarity applies
- Closed system: zero contact with the outside — the Schrödinger equation takes over and the evolution is always unitary;
- Open system: the environment gets involved — non-unitary things start happening, and measurement is the most extreme example (random, irreversible, superposition-destroying).
An ideal quantum computer stays closed between measurements, unitary throughout, with measurement as the single sanctioned exit. This also explains hardware’s enemy number one, decoherence: any “peek” by the environment is an uninvited little measurement, quietly collapsing your superposition.
7.4 Bonus one: every gate can be undone
Unitary implies invertible, and the antidote is dagger: ; undoing a stretch of circuit means running it backwards with every box daggered. Contrast the classical world: an AND gate takes two bits in and puts one bit out — seeing output 0, you cannot tell whether the input was 00, 01, or 10; information is destroyed, irreversibly. Quantum circuits have no destruction option (before measurement everything is unitary), so compiling classical logic into a quantum circuit requires reversibilization first (introducing reversible gates such as the Toffoli). This is not philosophical trivia; it is a design constraint you actually hit when writing circuits.
7.5 Bonus two: any point to any point on the sphere
Unitaries preserve inner products, hence the “angle” between any two states, hence they act on the Bloch sphere as rigid rotations, able to carry any point to any other point. This is the license on the first line of every algorithm: hardware natively prepares only , but any desired initial state is just ” plus a suitable “.
8. Measurement: the mathematical identity of the instrument
8.1 Observables: one Hermitian matrix per instrument
Postulate: every measurable physical quantity (an “observable”: energy, spin along some direction, polarization at some angle…) corresponds to a Hermitian matrix . Here is a placeholder for “whatever instrument”; you already know three concrete members: . The dictionary between matrix and instrument has exactly two entries:
- Eigenvalues = the list of possible readings on the dial (Hermiticity guarantees they are all real);
- Eigenvectors = the states for which each reading is certain, and also where the system lands after the measurement.
The three clauses of measuring : the reading is always one of ‘s eigenvalues; the state projects (collapses) onto the eigenvector of that reading; which reading occurs is random, with probabilities determined by the state (Born). Corollary: only states already sitting on an eigenvector give predictable outcomes; for every other state, no one can predict a single shot.
Readings are not always : the energy observable is the Hamiltonian matrix , whose eigenvalues are the energies of the levels (arbitrary real numbers) and whose eigenvectors are the levels themselves — the same clauses, word for word.
8.2 The key clarification: measurement is not computing
A misconception that almost everyone produces: “multiply the matrix into the state and get a classical bit.” No. Demonstrate why the “hard multiply” must be wrong: take (eigenvalues 2 and 5) and the state , and actually multiply:
Three sins: the squared length is , not a legal quantum state; no number was output; and matrix multiplication is fully deterministic, while real measurement is random. The real outcome: with probability 50% read 2 and the state becomes ; with probability 50% read 5 and the state becomes — bearing no resemblance to the product.
8.3 The correct picture: a codebook plus dice
The true identity of is a package of two things: an orthogonal basis (the eigenvectors) plus a reading label glued to each basis state (the eigenvalues). The actual procedure of measurement has four steps, and multiplication never appears:
- Consult the codebook: solve for ‘s eigenvectors and eigenvalues (done once, when the instrument was designed);
- Decompose: expand in the eigenbasis, ;
- Roll the dice: select the outcome by the Born rule, reading with probability ;
- Collapse: the state jumps to the winning .
The physical process is carried out by apparatus (beam splitters, magnets…); the matrix describes not the machinery but the interface specification — is the instrument’s API signature: the list of return values (eigenvalues) plus the set of deterministic inputs (eigenvectors). Implementation belongs to hardware; the signature belongs to the matrix. The spectral theorem guarantees that a Hermitian matrix is exactly equivalent to this data: — the matrix is the codebook compressed into a single object.
Matrix multiplication has three legitimate uses, all paper bookkeeping: making the codebook (solving , multiplication as a diagnostic tool); computing the average reading (in the example above, ✓ — a statistical tool, not the measurement itself); and testing compatibility (the commutator machine needs matrices to subtract).
8.4 The dual identity, side by side
| Used as a gate | Used as an observable | |
|---|---|---|
| Mathematical action | genuinely multiply: | consult codebook + Born dice |
| Determinism | fully deterministic | single-shot outcome random |
| Reversibility | reversible ( undoes) | irreversible (information burned) |
| Output | a new quantum state, no number | one classical number + the collapsed state |
Matrix multiplication belongs to the left column’s world (unitary evolution); measurement is the right column’s world, two independent rules in the axiom list, neither reducible to the other.
Finally, state the “Hermitian” credential in full: a qualified instrument needs real readings (Hermitian ⟹ all eigenvalues real), distinguishable outcomes (spectral theorem ⟹ eigenvectors of distinct eigenvalues are automatically orthogonal — was never a coincidence), and every state measurable (the eigenvectors form a complete basis, so every state has a page in the codebook). All three in one package: the Hermitian matrix is the mathematical name of the concept “measuring instrument”.
9. Projective measurement: the formal uniform
9.1 A new part: the outer product — brackets facing outward
You know the inner product : row times column, a number. Reverse the order and you get the outer product: column times row, a matrix:
The mnemonic: brackets facing inward, , close up into a number; brackets facing outward, , open out into a matrix.
9.2 The projector: casting a shadow
Write . Act on and use the sandwich reading — the on the right bites first and spits out a number:
The part is killed, the part kept: asks “how much does contain” and keeps only that. The geometric picture: push the vector down onto the axis and take its shadow.
Two properties, one line each. Hermitian: is real symmetric ✓. Idempotent () — “the shadow of a shadow is the shadow”:
9.3 Walking the formula chain
Read it segment by segment: the starting point is the actual law, — the squared length of the shadow; expanding by the definition of the norm ( with ) gives ; Hermitian plus idempotent collapses into ; and the sandwich splits into two numbers multiplied, and . The whole chain is the Born rule wearing the projector uniform; no new physics anywhere.
The post-measurement state falls out along the way: the shadow has length , not a legal state — renormalize (divide by the length) to get , and the unit-modulus factor in front is a global phase, discarded, leaving . Two old friends, renormalization and discarding global phase, take the stage together.
9.4 Spectral decomposition: the codebook is a stack of projectors
One projector per page, one reading label glued on — the simplest instance of the spectral decomposition mentioned in §8.3. The expectation value also turns transparent at once: , the probability-weighted average reading.
9.5 Why it is worth the trouble: it transplants unchanged to many qubits
On a single qubit this machinery is overkill ( can be read off by eye). The real payoff comes with multiple qubits: when measuring only the first qubit of an entangled pair, “the amplitude” is no longer a single number, yet the projector recipe carries over without changing a word:
(Here is the tensor product, the operation that joins subsystems into a composite system; reads “project the first qubit, do nothing to the second”.) “Probability , post-state normalized ” is the one version that upgrades losslessly, and the form worth memorizing.
9.6 True randomness
Three worked cases: gives 0 with probability 1 (already on the eigenvector); likewise; gives 50/50 — and nobody in the universe can tell you in advance which outcome this particular run will produce; only ensemble statistics are lawful. This is not the ignorance-style randomness of a coin hidden under a cup (the kind in a complexity class like BPP, where peeking at the random tape would in principle let you predict), but true randomness: there is no coin, the value simply was not defined before the measurement (Bell experiments ruled out the hidden coin). Two direct consequences: quantum random number generators are the best randomness sources because their randomness is backed by physical law rather than algorithmic disguise; and the quantum complexity class BQP is defined with bounded error (error at most , then amplified) precisely because quantum algorithms are born with random output — the “two-thirds, then amplify” machinery built for BPP works off the shelf.
10. A checklist of common misconceptions
- “Measurement means computing ” — wrong. The matrix is never multiplied into the state; measurement = decompose → roll dice → collapse, and is the codebook, not the action (the three-sins counterexample of §8.2).
- “Eigenvalue means the state changed into another state” — on its own it is only a global phase, motionless on the sphere; only inside a superposition does it become a consequential relative phase (§5.3).
- “Rotating returns you to the start” — off by a factor of ; only truly returns home. Rotate just one branch of a superposition and that shows up in interference fringes (§6.4, the neutron experiments).
- “Circuits and matrices read in the same direction” — circuits read left to right (time), matrices right to left (function composition); reading backwards can be wrong by as much as orthogonal ( versus ).
- “A qubit is a photon” — the qubit is a role; the photon (one of its degrees of freedom) is one of the actors; keep the three layers straight: logical, degree-of-freedom, carrier (§4).
- “Superposition is mysterious” — for polarization it is vector decomposition: a field really is composed of simultaneous horizontal and vertical oscillation (§3.2).
- “Quantum randomness is like classical randomness” — BPP’s randomness is an unflipped coin; quantum randomness is the coin not existing; Bell experiments are the evidence (§9.6).
- “Global phase can never be measured” — needs sharpening: acting on the entire system, it cannot; acting on only part of a superposition, it demotes itself to a relative phase and becomes measurable (§6.4).
11. Self-test (answers included)
Problems
- Verify by hand that .
- What does do to ? Which Pauli gate is it equivalent to?
- A photon linearly polarized at enters an H/V polarizing beam splitter. What are the probabilities at the two exits?
- Do and share a common eigenvector? Explain using the five-line proof.
- Define . For , compute and the post-measurement state.
- Why can a classical AND gate not be used directly as a quantum gate?
- What is the output of the circuit ? (Hint: multiply right to left, or walk it step by step.)
Answers
- ; ; subtracting gives ✓.
- ; and . North pole to south pole; up to a global phase it is the gate (which also sends to ; the difference from lies in phase details).
- , .
- No. If were an eigenvector of both, then ; but and makes invertible, so , contradiction — a state definite in is necessarily maximally uncertain in .
- ; the shadow renormalizes to .
- AND takes two bits to one bit, destroying information, hence irreversible; quantum gates must be unitary (reversible), so classical logic must first be rewritten with reversible gates such as the Toffoli.
- ; ; . Output . (In matrix language: — “a sandwiched between two ‘s becomes an ”, a classic example of basis change.)
12. Symbol quick reference (new in these notes)
| Symbol | Name | One-line meaning |
|---|---|---|
| commutator | how much order matters; is required for simultaneous definiteness | |
| Pauli matrices | rotations about the three axes; moonlighting as three instruments | |
| Hermitian conjugate | transpose + entrywise conjugate | |
| unitary | preserves length; the license to be a gate | |
| Hermitian | all-real eigenvalues; the license to be an instrument | |
| eigenvalue equation | : a direction the matrix cannot move; : reading / phase factor | |
| rotation gate | rotate about axis by ; ∝ pulse duration | |
| outer product / projector | brackets opening outward into a matrix; casts shadows | |
| idempotent | the shadow of a shadow is the shadow | |
| projective-measurement probability | squared shadow length; upgrades losslessly to many qubits | |
| spectral decomposition | the codebook: a stack of projectors, each with a reading label | |
| tensor product | joins subsystems into a composite system |
The whole article in one sentence: knowing pins down the length of but not the angle between the two arrows — that angle is real physical information, taking bodily form as an angle you can verify with polarizers, turned by the knob, and exposed by the of a full turn inside an interferometer; gates work by multiplying unitary matrices, instruments roll dice from a Hermitian codebook, and the projector’s “squared shadow length” formula is the one piece of luggage that travels to the many-qubit world unaltered.