Part 1 ended with a switch that electricity can flip, the relay, and with the two things it could not do: flip fast, and make a weak signal strong. Both of those are the vacuum tube's story. Go back to the microwave oven from the opening of Part 1. The magnetron behind its back wall is the great-grandchild of a device found by accident in 1883, inside a light bulb, by a man who patented it and then put it in a drawer. This part builds that device up a piece at a time, follows it through the first computers, and ends where the series really begins, with the transistor.

The vacuum tube: from light bulb to one-way valve

The device that made electronics possible was found by accident, inside a light bulb, and the quickest way to understand it is to build it up a piece at a time: first the bulb, then a plate, and, in the next section but one, the part that changed everything.

Start with the bulb itself. Thomas Edison's lamp of 1879 is a thin filament inside a glass envelope with the air pumped out. The filament is heated white-hot by the current through it, and the vacuum is there because a filament that hot would burn away in seconds if there were any oxygen to burn in. Around 1883 Edison noticed something odd. If he put a second, separate piece of metal inside the bulb, a current would flow across the empty space from the hot filament to that plate, but only when the plate was connected to the positive side of the battery, and never the other way round. He patented the effect and did nothing with it.

Here is what was happening. Heat the filament enough and it boils off electrons into the vacuum, a cloud of them hovering around the wire. Make the plate positive and it attracts the cloud: electrons stream across the gap and a current flows. Make the plate negative and it repels them, and since the cold plate emits no electrons of its own, nothing can flow back. Current goes one way only. In 1904 John Ambrose Fleming, who had worked for Edison, turned this into a component, the diode (two electrodes), a one-way valve for electricity, which is why British engineers call every tube a valve to this day.

1. the light bulb, 18792. add a plate: the diode, 1904 glass, air pumped outglass, air pumped out a white-hot filament that would burn in air cold: emits none plate (+) electrons stream up never back down hot: emits electrons
From a lamp to a one-way valve. The vacuum was there to stop the filament burning. It turned out to be a space that electrons could be steered across, and they can only ever set off from the hot side.

Why a one-way valve matters

It is fair to ask what a component that only lets current through one way is for, and the answer is two things without which there would be no radio, no television, and no plug-in electronics of any kind.

The first is turning mains electricity into something a circuit can use. The power in your walls is alternating current: it reverses direction fifty or sixty times a second, because that is the easiest kind to generate and to send over long distances. A heater or a light bulb does not care which way the current goes. But an amplifier, and every circuit in this series, needs a steady push in one direction, direct current, the way a battery gives. Put a diode in the path and it passes the forward half of every cycle and blocks the reverse half, leaving a series of one-way pulses. Add a capacitor, the small tank from Part 1's side note, to fill in the gaps between the pulses, and the result is a nearly steady supply. This is rectification, and every radio, television, and amplifier that ever plugged into a wall had a rectifier tube at the back. The rectifier is still there in everything you plug in today, as a semiconductor diode the size of a grain of rice, and it is the first thing the current meets inside a phone charger.

rectification: what a one-way valve is for alternating current from the wall:back and forth, 50 times a second one-way pulses:the reverse halves are blocked direct current, nearly steady:what a circuit can actually use diode capacitor
Mains in, steady current out. The diode strips one direction, the capacitor fills the gaps. Every plug-in device does this first.

The second job is picking a voice out of the air, and it deserves a careful look, because the trick sounds like it should not work. A radio receiver ends up throwing away half of the wave it receives. How does the music survive that?

The answer is that the music is not the wave. A station broadcasts a carrier: a wave that alternates far too fast to hear, around a million times a second for an AM station, and on its own carries no sound at all. To put the music on it, the transmitter makes the carrier's height rise and fall in step with the sound, a few hundred to a few thousand times a second, slow compared to the carrier's own wiggle. The scheme is called amplitude modulation, which is what the AM on a radio dial stands for, and the music lives entirely in the outline of the wave, its envelope, not in the wiggles themselves. Row 3 below is the picture to stare at.

Now the receiver's problem. Feed that wave straight to an earphone and nothing happens: every upward push is matched a millionth of a second later by an equal downward pull, the average is zero, and no earphone can move a million times a second anyway. The signal is there, but it is folded up symmetrically, top mirroring bottom. This is where the one-way valve earns its keep. Pass the wave through a diode and the bottom half is simply blocked. The symmetry is broken, the pushes no longer cancel, and the average of what remains rises and falls exactly as the envelope does, which is to say exactly as the music does. A capacitor then smooths away the leftover million-a-second bumps, the way it smoothed the rectifier's pulses, and what is left is row 1 again: the sound, recovered.

how a one-way valve gets the music back 1. the sound air pressure rising and falling a few hundred times a second 2. the station's carrier a million wiggles a second, constant height, no sound in it 3. what is broadcast the carrier, with its height copying the sound. The dashed outline, the envelope, is where the music lives 4. after the one-way valve the bottom half is blocked, so the average, in red, now follows the envelope: the music 5. smooth away the wiggles a capacitor fills the fast gaps and row 1 is back, ready for an earphone
Amplitude modulation, undone by a diode. Nothing of the music is lost when the bottom half goes, because the music was never the wave. It was the wave's changing height, and that survives.

That is detection, and it is how every early radio worked. The crystal set that a child could build in the 1920s, a coil, a crystal, and an earphone, needed no battery at all: the diode was a "cat's whisker", a fine wire touching a crystal of galena, which happens to conduct one way, a semiconductor diode in use fifty years before anyone understood why it worked. The energy that moved the earphone came from the radio wave itself. What the crystal set could not do was make the sound any louder than the wave provided, and that is the cliffhanger the next section resolves.

The grid, and the first amplifier

The diode is a valve with no handle. In 1906 Lee de Forest gave it one. He put a third element between the filament and the plate, a mesh of fine wire he called the grid, and found that a small voltage on the grid controlled the large current from filament to plate. Make the grid slightly negative and it repels the electrons, choking the stream. Make it less negative and the stream flows freely. And because the grid is a sparse mesh that the electrons fly through rather than land on, it draws almost no current itself: it governs the stream by electric influence alone. The device is the triode (three electrodes), and it is the ancestor of every transistor in your graphics card.

3. add a grid: the triode, 1906 glass, air pumped out plate the big current leaves here grid: a mesh of fine wire a small voltage on it throttles the whole stream, and it draws almost no current itself cathode, heated boils off the electrons the plate collects
The third element. De Forest slipped a mesh between cathode and plate, and because every electron has to pass through it, a whisper of a voltage on the mesh commands the whole stream.

It is worth walking through what the triode actually does in a circuit, because "the first amplifier" is an idea the rest of electronics is built on, and one detail of it puzzles everyone at first: if the signal coming in is feeble, where does the loud version's energy come from?

Here is the arrangement. The grid is connected to the source of the weak signal, a microphone, an aerial, the worn-out voice arriving down a long telephone line. The plate is connected, through whatever should receive the loud version, a loudspeaker or the next stretch of line, to a hefty battery or power supply. The battery would happily drive a large current through the tube all the time. The grid stands in the way, throttling that current, and as the weak signal wiggles the grid's voltage by hundredths of a volt, the throttle opens and closes by the same rhythm. The plate current becomes a copy of the input's shape, drawn at the battery's strength.

the first amplifier: a battery, steered plategridcathode microphone a few hundredths of a volt, wiggling the grid 90 V loudspeaker the same shape in the plate current, hundreds of times stronger every drop of the loud version's energy comes from the battery. The whisper only steers it
The triode in its natural habitat. The red circuit is the whisper, the blue circuit is the battery's muscle, and the grid is the tap handle between them. This is Part 1's three-ways figure, third row, made real.

So the amplifier does not make the small signal bigger, not really. It uses the small signal to sculpt a big current that was there for the taking, the way the garden tap's handle sculpts the mains pressure rather than pushing the water itself. The energy in the loud output is the battery's. The information in it is the whisper's. That is why the crystal set of the last section was doomed to be quiet, with no battery there was nothing to sculpt, and it is why the triode changed everything at once: add a battery and one tube, and any signal too faint to use became the same signal, usable. Chain a second tube after the first and it amplifies the amplified copy, which is how three repeater stations could carry a voice across a continent in 1915, each one re-sculpting a fresh battery's current into the arriving signal's shape.

One more thing falls out for free, and it is the reason this series cares. Wiggle the grid gently and the triode is an amplifier. Slam the grid between two extremes instead, hard negative and the stream is choked entirely, up to zero and it floods, and the triode is a switch, the relay's job with no moving parts and no millisecond of travel. Same tube, different manners. The garden tap is the analogy that carries this whole series: the handle is the grid, the water is the electron stream, and turning the handle moves no water itself. Hold onto that picture, because Part 3 uses exactly the same tap to explain the transistor, and the point of the whole story is that the tap stayed the same while everything around it changed.

What the tube made possible

The triode changed the world faster than any component before or since. In 1915 it let a telephone call cross the American continent for the first time, with three tube repeaters along the way, at Pittsburgh, Omaha, and Salt Lake City, restoring a voice that thousands of miles of copper had worn down to almost nothing. From 1920 it made radio broadcasting possible, both the transmitters and the sets in living rooms. It made television, radar, the electric guitar, sound in cinemas, and long-distance telephony. And because it could switch in a microsecond where a relay took a millisecond, it made the electronic computer: the AND and OR of Part 1, with a tube's grid in place of a relay's coil and its plate current in place of the contacts, a thousand times faster.

Colossus, the British code-breaking machine of 1943, used 1,600 tubes in its first version and 2,400 in its second. ENIAC, completed in 1945 in Pennsylvania, used 17,468 of them, could do 5,000 additions a second, filled a room of about 170 square metres, and drew 150 kilowatts. Nearly every computer of the 1950s was a tube computer: UNIVAC, the IBM 700 series, the early Manchester machines. The largest of them all was the SAGE air defence system, which went into service in 1958 with around 50,000 tubes per computer, two computers per site, and a power bill of three megawatts for the site. It was the biggest computer ever built, by most measures it still is, and it was obsolete before it was finished, for reasons the next section explains.

The tube also never went away, and this is the part people miss. It survived wherever the job was too hot, too powerful, or too high in frequency for a semiconductor to do cheaply. The magnetron in your microwave is a tube. The X-ray tube in a hospital CT scanner, a dentist's surgery, and an airport baggage scanner is a tube, pumping electrons into a metal target at a hundred kilovolts. Communications satellites amplify their downlink with travelling-wave tubes, because for decades nothing else could produce that power at those frequencies with that efficiency in orbit, and many still do. Broadcast radio and television transmitters ran on tubes the size of dustbins into the 2000s. Particle accelerators drive their beams with klystrons, tubes the size of a person. The fluorescent display on an older microwave clock or car dashboard is a small vacuum tube. And a large share of guitar amplifiers, and a smaller share of expensive audio equipment, still use tubes on purpose, because the way a tube distorts when overdriven is the sound of rock music, and no transistor circuit quite copies it. If you own a Fender or a Marshall, you own a handful of triodes that would be recognisable to de Forest.

What was wrong with it

Everything that was wrong with the tube followed from one fact: to boil electrons off the cathode, you have to heat it. A typical small tube burns a couple of watts in its heater alone before it does any work, and the largest single item in ENIAC's 150 kilowatts was the heaters, with a ventilation system to match. A tube takes half a minute to warm up before it works at all, which is why old radios and televisions came on slowly. And the heater, like the filament in a light bulb, eventually burns out. A good tube lasts a few thousand hours. That is fine for a radio with five tubes. It is a catastrophe for a computer with seventeen thousand, because with that many, one of them is always about to fail. ENIAC's engineers ran the heaters below their rated voltage and never switched the machine off, and still lost a tube every couple of days, each failure meaning a hunt through the racks. SAGE used tubes specially built for reliability and still needed a full-time crew replacing them.

Then there is size. A tube is a glass bottle with a vacuum in it, and a vacuum needs a certain volume to hold electrodes apart. The smallest practical tubes were the size of a thumb, and no amount of ingenuity made them the size of a grain of rice. A computer's power is set by how many switches it has, and with tubes, every switch was a thumb-sized, two-watt, glass object with a lifetime of a few years and a warm-up time of thirty seconds. You could build ENIAC. You could not build anything a hundred times bigger, and by the late 1950s engineers had a name for the wall they had hit. Jack Morton of Bell Labs called it the tyranny of numbers in 1958. Every extra switch added cost, heat, failure rate, and thousands of hand-soldered joints, and the joints failed too.

Try it below. Pick a technology and a number of switches, and see what the machine would cost you in power, space, and reliability. Then slide up to the 76 billion switches of the graphics card that this series is about.

A machine made of switches. Rough figures for each technology: power per switch, volume per switch, how long each one lasts, and how fast it flips. The tube lifetime is ENIAC's, with derated heaters and the machine never switched off, which is far better than a tube's rated life.

built frompowertakes upa switch fails everyflips per second

About the last row's lifetime figure: nobody measures single transistors, because they do not fail one at a time. The chip industry qualifies whole chips, in a unit called the FIT, one failure per billion hours of operation, as defined by JEDEC, the body that sets these standards. A modern chip is engineered to a failure rate of some tens of FIT. Divide a 20 FIT chip by the RTX 4090's 76 billion transistors and one transistor's share works out around 1019 hours. Treat it, like every number in this table, as an order of magnitude.

The figures are rough, the trend is not. Tubes and relays are fine up to a few thousand switches, marginal at tens of thousands, and impossible beyond. Yet everything interesting a computer can do, including everything in this series, needs millions of switches at the very least. The tube got computing started and then stood in its way.

How the transistor fixed it

In December 1947 at Bell Labs, John Bardeen and Walter Brattain built the strangest-looking device in this series, and it is worth picturing properly. They took a small triangle of plastic and wrapped a strip of gold foil over its point, then cut the foil at the very tip with a razor blade, leaving two gold edges separated by a gap the width of a hair. They pressed the point down onto a sliver of germanium, a semiconductor, with a bent spring, so that the two gold edges became two contacts touching the germanium almost at the same spot, and a third contact sat under the slab. Then they found what they were hoping for: a small signal fed into one gold contact came out of the other one larger. Power gain, the triode's trick, from a lump of solid material with no vacuum, no glass, and nothing glowing.

December 1947: the point-contact transistor 1950: the junction transistor a spring, pressing down gold foil on a plastic triangle the foil slit at the tip: two contacts, a hair apart germanium, a semiconductor metal base: the third contact collector emitter base, thinner than paper a small current in at the thin middle layer controls a big current through the whole stack
The first transistor and the one that got manufactured. Bardeen and Brattain's contraption of foil, plastic, spring, and wax proved the effect. Shockley's three-layer sandwich made it sturdy enough to build by the million.

The contraption was temperamental, noisy, and nearly impossible to make twice the same way. Their boss, William Shockley, furious at having missed the discovery moment, spent the next month working out the theory and a better design: no delicate points at all, but a sandwich of three semiconductor layers, a thin base layer between an emitter and a collector, where a small current fed into the thin middle layer controls a large current flowing through the whole stack. The junction transistor was demonstrated in 1950, it was rugged and manufacturable where the point-contact was fussy, and it is the design that filled radios and computers for the next twenty years. The three men shared the Nobel Prize. What the transistor did was nothing new: a small signal controlling a large one, the triode's job exactly. What was new was what it did not need.

It did not need a heater, and the reason deserves a moment, because it is the whole point of the word semiconductor. In a metal, a fraction of every atom's electrons are unattached, a sea of them, free to drift the moment a voltage asks, which is why a metal always conducts. In an insulator such as glass, every electron is locked to its atom, and nothing short of destruction will free them. Silicon and germanium sit in between: pure, at room temperature, they have almost no free electrons and barely conduct at all. The trick that makes them useful is called doping. Sprinkle the crystal with a trace of a neighbouring element, one atom in a million, and you set the number of mobile charges by recipe: phosphorus brings one spare electron per atom, boron brings one vacancy that behaves like a mobile positive charge. Chemistry, not heat, decides exactly how many carriers exist and in which regions.

metal a sea of free electrons: always conducts insulator every electron locked to its atom: never conducts doped semiconductor a chosen few, placed on purpose: conducts exactly as designed a donor atom
Why "semi". The transistor's raw material conducts only as much as its maker decided, region by region, and the mobile charges are there at room temperature.

Now compare the starting points. The tube had to boil its electrons off a hot cathode to get them into a vacuum where a grid could steer them, and paid a couple of watts, thirty seconds of warm-up, and an eventual burnout for the privilege. The transistor's mobile charges are built into the solid by chemistry, already sitting where junctions and fields can steer them, at room temperature. No heater, so no warm-up, no burnout, and a thousandfold less power before anyone had tried to optimise anything: early transistors used milliwatts against a tube's watts. It did not need a vacuum or a glass envelope either, so it was rugged and could be made small: the first commercial transistors were the size of a pea, and, unlike the tube, there was no physical reason they could not be made smaller still. A transistor that is kept within its ratings does not wear out at all, which turned the tyranny of numbers on its head. A machine with a million transistors is not a machine with a million things waiting to fail.

triode, 1906 transistor, 1947 onward glass envelope, vacuum inside plate (anode) grid: a small voltage here controls the whole stream cathode: heated, boils off electrons electrons stream up source drain channel gate solid silicon, no vacuum, no heater, no glass same job, three parts renamed: cathode → source grid → gate plate → drain
The triode and the transistor do the same thing: a small voltage on a control electrode opens or closes a path for current. The transistor does it without heat, glass, or vacuum.

The transistor had its own problems, and for a decade the tube kept its jobs. Early germanium transistors were expensive, around eighteen dollars each in 1950 against seventy-five cents for a tube, and their behaviour drifted with temperature. They could not handle high power or high frequencies, which is why transmitters, radar, and anything with kilowatts stayed with tubes, and why the microwave oven still does. And they were noisy in the electrical sense. So the transistor won its first markets where small and low-power mattered more than anything else: hearing aids, which people had been carrying around with a tube amplifier and a battery pack the size of a book, took their first transistor in 1952 and had gone almost entirely transistor by 1954. That year also brought the Regency TR-1, the first commercially made transistor radio, with four transistors, which sold for about fifty dollars and made the word "transistor" mean "radio" for a generation. The first transistor computer ran at Manchester University in 1953, Bell Labs' TRADIC followed in 1954 with 684 transistors (and one tube, to generate its clock), and by 1959 IBM's 7090 mainframe was transistorised. No serious computer used tubes again.

Two further steps made the transistor into the switch this series is about, and each deserves its own picture.

Printing the wires: the integrated circuit

Replacing tubes with transistors shrank the switches but not the problem, because the wiring stayed. A late-1950s computer was still tens of thousands of separate components, every one placed and soldered, and every joint a place to fail. The tyranny of numbers had moved from the tubes to the joints.

In the summer of 1958 Jack Kilby, newly arrived at Texas Instruments and with no holiday to take, was left alone in the lab with the problem. His insight was that the other components on a circuit board, the resistors and capacitors of Part 1's side note, could all be made, less well but well enough, out of the same semiconductor as the transistors, which meant an entire circuit could be built in one piece with nothing to solder. On 12 September 1958 he demonstrated it: a whole working oscillator on a single sliver of germanium. One piece, but not yet one process: his components were joined by fine gold "flying wires", attached by hand under a microscope.

The finishing move came at Fairchild. Jean Hoerni's planar process made transistors flat, buried under a smooth, glassy skin of silicon dioxide, and in 1959 Robert Noyce saw what the flatness allowed: open small windows in the glass, then deposit aluminium tracks across the top, shaped photographically like everything else. The wires stopped being things a person attached and became one more printed layer. Fairchild had working chips built this way by 1961, and that is the integrated circuit as it still is: components and their wiring made together, by light, with no hands anywhere. The soldered joints disappeared the way the tubes had, cost per component began the collapse that has not stopped since, and Part 3 picks the story up from there.

Kilby, 1958: one piece, hand-wired Noyce and Hoerni, 1959: one process gold flying wires, attached by hand under a microscope a whole oscillator on one sliver of germanium glued to a glass slide aluminium tracks printed on top wiring as one more printed layer components buried under a flat, glassy skin of silicon dioxide. Made by light, no hands a modern chip is the right-hand picture, repeated: today's have dozens of printed wiring layers stacked above the transistors
Two steps to the chip. Kilby proved a whole circuit could be one piece of semiconductor. Noyce and Hoerni made the wires part of the printing, which is the fact the entire rest of this series stands on.

The transistor that scales: the MOSFET

The junction transistor has one habit that did not matter for a radio and matters enormously for a computer: it is worked by a current. To hold it on, you must keep feeding current into its thin base layer, and that current is spent, turned to heat, by every single switch, all the time. A few hundred transistors can afford it. A billion cannot.

In 1959, at Bell Labs, Mohamed Atalla and Dawon Kahng built a transistor that goes back to the triode's best idea. The grid never touched the electron stream, it steered by electric influence alone, and their device does the same inside a solid: a metal gate sits above the current's path, separated from it by a whisker-thin layer of glass, and the gate's electric field, acting through the insulation, opens and closes the channel underneath. No current flows into the gate at all, beyond the instant of switching. That is the MOSFET, and silicon handed it a gift that no other material offered: silicon's own oxide, the "rust" that grows on its surface, happens to be a nearly perfect insulator, the very glass the gate needs, and the same glassy skin the planar process prints its wires on.

junction transistor: worked by a current MOSFET: worked by a field collector emitter control current flows in, always the control itself consumes current, turned to heat in every switch, all the time source drain channel glass gate a voltage, and almost no current, ever the field acts through the glass, so holding a switch on or off costs essentially nothing
Why the MOSFET won. Both do the triode's job, but only one of them can be a billion switches on a thumbnail without melting. Part 3 opens this device up properly.

The MOSFET was slower than the junction transistor at first and nobody important wanted it, but it had the two properties that decide everything at scale. Holding its state costs essentially nothing, especially after 1963, when Frank Wanlass worked out how to pair each MOSFET with an opposite twin so that one of the two is always off, the arrangement called CMOS that Part 4 will meet, in which a circuit draws real current only at the instant of switching. And it is flat, made entirely of printed planar layers, so every improvement in printing makes it smaller with no redesign of the idea. It draws almost nothing, it shrinks better than anything else ever invented, and there are 76 billion of them in the chip on the cover of this series.

Where you will find each one today

Relay Vacuum tube Transistor
How it switches an electromagnet moves a contact a grid voltage throttles an electron stream in a vacuum a gate voltage opens a channel in solid silicon
Flips per second hundreds up to millions billions
Power per switch around a watt, to hold the arm a few watts, mostly the heater nanowatts on a chip
Size a sugar cube a thumb a few dozen atoms across
Lifetime millions of operations, then the contacts go a few thousand hours, then the heater goes effectively unlimited within ratings
Best at switching big currents with total isolation very high power at very high frequency being small, cheap, cool, and numerous
Where it lives now car starters, headlamps and horns, air conditioners, industrial machinery, lifts, substation protection, the click in a cheap timer microwave ovens, X-ray machines, satellite and radar transmitters, particle accelerators, guitar amps, some hifi every chip, every phone charger, every electric car's motor drive, every LED bulb, and the 76 billion in a GPU

The pattern in the last row is the lesson of this part. Each technology kept the jobs where its particular strength mattered more than its weaknesses. The relay kept isolation and brute current. The tube kept extreme power and frequency. The transistor took everything where the count mattered, and it turned out that in computing, the count is the only thing that matters.

Where this leaves us

Electronics needed a component that lets a small signal control a large one. The relay of Part 1 did it with a magnet and a moving arm, and could switch a starter motor but not a radio signal. The tube did it with a heated cathode in a vacuum, could amplify anything, and built the first computers, but every one was a hot, thumb-sized, glass object with a lifetime of a few years, and that put a ceiling of tens of thousands on how many switches a machine could have. The transistor did the same job in a solid, with no heater, no vacuum, and no wearing out, and the ceiling disappeared. The tube and the relay are still around us, doing the jobs where they were never beaten. But the machine this series is about could only be built from the third.

So the next part starts where Stokes did, with the transistor as a switch: what a MOSFET is, why it can only be on or off, how it got from one to 76 billion, and what it costs, in heat, to flip it a few billion times a second.


Next: The Switch: the transistor as a switch, why computers count in twos, what "4 nanometres" does and does not mean, and the end of the free lunch that made GPUs inevitable.