Lorentz transformation Solutions
Sections
- Quick note:
- Lorentz transformation: general structure
- 1D boost
- Solution
- Rapidity and velocity
- Physical consequences
- Time dilation
- Length contraction (the wrong way to see the concept)
- Relativity of simultaneity
- Length contraction, done properly
- Boost of a moving object
- Boost of light
- Doppler effect
- Aberration
- 2D boost
Quick note:§
Remembers that in previous articles, we defined the lorentz transformation using
This immidieatly hint us to have in the calculation of invariant, as the definition above is already about invariant of no matter what it is it should obey the definition. Let's sandwich it with and note that .
Yeah, we see that indeed is invariant, write it in index notation it is being invariant. And it is
This invariant, we call it proper length squared, whos actual meaning will be clear later. Although from first principle this is not how we defined lorentz transformation. But since this invariant in universally true, we might as well treat it as an alternative definition of lorentz transformation. A transformation that keep this value invariant.
And note that we can form an invariant of another unit ( of time ) rather than lenght by dividing through . We will get
Again, we will see why is this the proper time later.
Invariant in general
Note that the derivation of invariant works for any four vectors, not only for four positions
Lorentz transformation: general structure§
Let's remind ourselves that we have a generator for the Lorentz transformation
(using a different convention from previous articles, but the structure stays the same). We have the generator
This is the general , the Lie algebra of — the proper, orthochronous Lorentz group, the part reachable continuously from the identity by exponentiating a generator like . (The full Lorentz group also contains parity and time reversal, sitting in disconnected pieces that no real can reach this way — we won't need them here.) We split into six basis generators, exactly matching the six degrees of freedom: three boosts, three rotations.
so that
with the meaning of and now immediate: one parametrizes boosts, the other rotations.
These six generators don't just sit next to each other — they talk to each other:
Keep that last relation in your pocket. We won't need it for a while, but it is not decorative.
A fully general six-parameter Lorentz transformation can be written down by exponentiating directly, but the closed form carries no physical insight on its own. So instead of solving the general case, we'll build up from the physically transparent one. Homogeneity and isotropy of space mean we only need to understand a boost with no rotation — so set — and by isotropy, any single boost direction is as good as any other. We start with one dimension.
1D boost§
Solution§
Let only be nonzero. Call it simply , and its generator simply , so that in one dimension
Notice
which turns the exponential into a familiar pair of series:
so that
Rapidity and velocity§
is just a parameter so far — let's find out what it physically means. Boost into the frame of an object moving as , and demand that in the new frame the object sits still: .
Set , so , giving
with , . Notice doesn't care about the sign of — the transformation matrix has the same structure whether you boost forward or backward, which is just isotropy of space showing up again. That also tells you : undoing a boost is the same as boosting the other way.
The parameter itself has a name — rapidity — and unlike velocity, rapidities for collinear boosts simply add. We won't need that fact yet, but it's worth knowing it's there.
Physical consequences§
Everything that follows — time dilation, simultaneity, length contraction — comes out of exactly the same two equations:
What changes from one effect to the next is only which pair of events we plug in, and which simultaneity condition we impose. Watch how much mileage we get out of the same formula.
Time dilation§
Take a clock sitting still in , so its worldline has . Then
Since is the clock's own rest frame, here is its proper time ( defined as the time measured in the rest frame ), so really
The moving observer watches the stationary clock accumulate less proper time than their own coordinate time ticks off — a moving clock runs slow. It's worth stressing that is not the universal statement of time dilation; it only holds because we chose . The frame-independent statement is the one with in it.
And let's we call we call and we can see it why here, since it is reference frame independant ( invariant ), we use the rest reference frame (). Then we find
Length contraction (the wrong way to see the concept)§
Now for length — and here we need to slow down, because length is a trickier thing to define than time. Time can be read off along a single worldline. Length can't: it needs the positions of two endpoints, at the same time, in whichever frame is doing the measuring.
If you note down where the front of a train is in the morning, and where its back is at night after it's crossed an entire country, that difference is not the train's length — you've just measured the gap between two unrelated events. So the measurement has to be simultaneous. Let's see what happens if we're careless about whose simultaneity we mean.
Take two endpoint events simultaneous in , so , and push them through the spatial transformation:
So the moving object comes out longer by a factor of . That should make you suspicious — length contraction is supposed to shrink things, not stretch them. Something in this calculation isn't measuring what we think it's measuring.
Relativity of simultaneity§
Here's the catch. We fixed — simultaneous in — but we never checked whether those same two events are simultaneous in . Let's check:
Since for two distinct endpoints, . The two events we used — simultaneous in — are not simultaneous in . That's exactly the mistake in the train story above: we measured one end and the other end at different times, just dressed up in Lorentz-transformation language instead of morning-and-night language. What we computed as "" was never the train's length in at all.
The same logic runs in reverse. Force instead — simultaneous in the moving frame — and see what that requires of :
So events simultaneous in are generally not simultaneous in either. Simultaneity isn't something the two frames agree on — it's frame-dependent, full stop, and it was hiding inside the Lorentz transformation the whole time.
Length contraction, done properly§
Armed with that, let's redo the measurement honestly. If we want the train's proper length, we need to measure both ends at the same time in the train's own rest frame — that means , not . Using the relation we just derived, , substitute into the spatial transformation:
And we write as the proper length and as the measured length. Where the proper length is deifned as the length measreud in the rest frame.
This is length contraction. The earlier factor of wasn't wrong arithmetic — it was the right arithmetic applied to the wrong pair of events. The deeper lesson: a Lorentz transformation doesn't just rescale numbers, it also reshuffles which events count as simultaneous, and length contraction is really a simultaneity effect wearing a geometry costume.
And since, now proper length is defined, let's see why the invariant is the proper length ( squared ). Well, since it is invariant ( it's value doesn't depend on the reference frame ), we can evaluate it's value in any reference frame that makes the meaning trivial. We chose the frame that makes , Which measure the length of the object in the rest frame, which by definition of proper length.
Boost of a moving object§
Now let's keep the velocity explicit instead of setting it to zero. Suppose an object moves at velocity in , so , and we boost by :
If , the denominator collapses to and we recover the Galilean , exactly as it should.
Boost of light§
Now push it to the extreme: set .
Light stays light speed no matter how you boost. That's not an extra rule bolted on afterward — it falls straight out of the same transformation we've been using all along.
Doppler effect§
A source at rest in emits with period — its proper period. Time dilation alone tells us the observed period lengthens to . But that's not the whole story: while the wave takes time to cross the widening gap, the observer (moving away, say) keeps retreating further, so each successive crest has a little farther to travel than the last. That extra travel time is , and folding it in gives
so the observed frequency is
This is the longitudinal Doppler effect — receding along the line of sight gives you both the time-dilation piece and the widening-gap piece together. If instead the motion is purely transverse — perpendicular to the line of sight, so the separation isn't changing — that second piece vanishes and only the dilation survives:
Aberration§
Doppler changes a wave's frequency. The same boost, applied to a different object, also changes the direction it appears to arrive from — that's aberration.
Let a photon travel in at angle to the -axis, so its momentum components are , , with . Photon momentum transforms as a four-vector, exactly the way did — we haven't shown why forms a four-vector yet, that's a later article, so take it on credit for now:
Divide the first by the second, and use :
A source moving toward you doesn't just blueshift — its light also gets funneled forward, toward , no matter which direction it was actually emitted in. Push and nearly everything a fast-moving source radiates ends up crammed into a narrow forward cone. That's the headlight effect, and it's the same we've had since the first page.
2D boost§
By isotropy, a boost pointed in some other single direction is no more interesting than the 1D case — the math gets uglier, the physics doesn't. The only reason 2D boosts deserve their own section is what happens when you chain two boosts in different, non-collinear directions. (Chain two boosts in the same direction and you just get one bigger boost — nothing new there.)
So look back at the generator, and consider composing two boosts:
For small , expand to leading order:
This is where that commutation relation we set aside earlier finally gets used. Recall , so , and substituting,
There it is: a term. Two pure boosts, composed, have produced a piece of a rotation.
(That's only the leading term. For finite rapidities, the exact rotation angle between two perpendicular boosts of rapidity and works out to , which reduces to our result once both are small. We won't derive the finite case here, but it's worth knowing the infinitesimal calculation isn't hiding anything qualitatively different from the exact one.)
Here's the physical picture behind that algebra. Take a rod already moving along , and in its own instantaneous rest frame, give both of its ends a simultaneous sideways kick in — a "rigid" perpendicular push. Simultaneous in that frame, that is. That frame is reached from the lab by an -boost, and an -boost's relativity of simultaneity depends on -position — so two events that share a or coordinate but differ in , simultaneous in the boosted frame, are not simultaneous back in the lab. Since the rod's two ends are separated along — the same direction as the boost it's already carrying — the "simultaneous" kick lands on one end before the other, as clocked in the lab. By the time you look at the whole rod at a single lab instant, one end has had a head start accelerating sideways and the other hasn't caught up — so the rod appears to have tilted, not because anything physically twisted it, but because the two kicks were smeared apart in time. And this only shows up when the boost you add is genuinely non-collinear with the one already there — a further kick purely along , collinear with the first boost, never desynchronizes anything.
This composition of non-collinear boosts producing a rotation is called a Wigner rotation. Let an object's velocity change continuously — think of it as an unbroken chain of these infinitesimal non-collinear boosts — and the accumulated rotation of its comoving frame is what we call Thomas precession.
It's worth stating the group-theoretic version of what just happened: pure boosts do not close under composition. Compose two non-collinear ones and you land outside the set of pure boosts entirely, in something with a rotation baked in. Boosts alone aren't a subgroup of the Lorentz group — which is exactly why needed both and from the very first page. by itself doesn't close under commutation; kicks you straight back out into rotation territory.
So that commutation relation we asked you to remember, , was never just bookkeeping. At the infinitesimal level it's already saying: change the direction you're boosting in, and a rotation is unavoidable.
This isn't just a geometric curiosity. The textbook non-relativistic calculation of spin-orbit coupling in hydrogen — the electron's spin interacting with the effective magnetic field produced by its own orbit — comes out a factor of 2 too large compared to experiment, until you account for the fact that the electron's rest frame is itself undergoing exactly this kind of precession as it orbits the nucleus. Thomas precession supplies the missing factor of . A purely kinematic fact about composing boosts turns out to matter for the fine structure of every atom you're made of.
Discussion
No comments yet — yours could open the discussion.