The earlier chapters treated a transistor as a tiny switch. This one is about how billions of them are actually made. A chip is built on a thin, shiny disc of pure silicon called a . It goes up one layer at a time. First come the switches, in the very top of the silicon. Then come ten or more layers of copper wires stacked above them.1
A modern chip goes through more than 300 steps and spends about three months in the factory.1 Its smallest shapes are thousands of times thinner than a human hair. Printing them takes some of the most complex machines ever built.
Making chips is also a numbers game. Each wafer holds many copies of the chip, and one speck of dust in the wrong place breaks a copy. The share that come out working is called the . Along with the price of the wafer, it decides what each chip costs.
A transistor’s behavior, which the earlier chapters described, depends on shapes and materials that a factory, or fab, has to build. This chapter follows that build. A of single-crystal silicon carries hundreds of copies of a design, each called a die. A process flow of more than 300 steps turns the bare wafer into finished dies with eleven or more levels of metal wiring.1 That takes about 12 weeks on average, and longer for the most advanced processes.12
Almost every step falls into one of four kinds:1
- Deposition adds a film.
- Removal takes material away: etching, and polishing flat.
- Patterning, or , decides where the next addition or removal happens.
- Modification of electrical properties, mainly implanting dopant atoms and heating the wafer to activate them.
The flow has three parts. The builds the transistors. The connects them upward. The back end of line (BEOL) stacks the wiring.1 The chapter ends with and cost: how many dies work, and what each good one costs.
This chapter is about the physical build. The After tapeout page covers the same fab from the chip designer’s side: masks, wafer sort, bring-up and qualification.
The device physics in the earlier chapters assumes structures that someone has to build, at a cost and with a yield. This chapter covers the build: the unit processes, how they are sequenced into a CMOS flow of hundreds of steps (up to about 1,400 for the most complex processes), how the fab keeps each step on target, and the yield and cost arithmetic that turns defect density and die area into a price per good die.2
Three constraints run through everything:
- Lithography pitch decides which layers need extreme ultraviolet, which need multiple patterning, and why modern layouts are so regular.
- Thermal budget decides the order of steps: hot steps first, and anything that can’t take heat later.
- Defectivity decides yield, and through it the economics of large dies.
Two open process kits anchor the numbers. SKY130 is a mature 180–130 nm-class planar process with five metal levels.29 ASAP7 is a predictive 7 nm FinFET kit whose design rule manual states, for every layer, whether it is printed with 193 nm immersion or EUV and whether it is single-exposed or multi-patterned.14 The After tapeout page covers masks, the Rayleigh resolution limit, wafer sort and the yield-learning loop. This chapter links to it instead of repeating it.
A bare wafer. Each round of four moves adds one patterned layer.
Silicon comes from ordinary sand, but chips need it incredibly pure: about one stray atom in a billion, or better.3
The pure silicon is melted and grown into one huge, perfect crystal. That crystal is sliced into thin discs, and each one is polished like a mirror.3
Today’s wafers are about the size of a dinner plate and about as thin as a credit card.4 They move between machines in sealed boxes of super-clean air. A speck of dust is huge next to a transistor, so even one can ruin a chip.1
Wafers start as a single-crystal ingot grown by the Czochralski method: a seed crystal is pulled from a melt of high-purity silicon, at least 99.9999999% pure. The ingot is sliced, and the slices are polished.3 For CMOS logic the crystal is usually oriented so its (100) face is the surface.5 The wafer is lightly doped: boron makes it p-type, phosphorus or arsenic n-type.3
The industry standard is 300 mm in diameter and 775 µm thick,4 used since 2000. A move to 450 mm wafers has met resistance over its return on investment.13 Wafers move between tools in sealed carriers (FOUPs) whose mini-environments can reach ISO class 1 cleanliness, cleaner than the cleanroom air around them.1
A die is the rectangle that becomes one chip. Dies are stepped across the wafer in a grid with narrow scribe lanes between them, where the saw will later cut. Dies cut by the round edge are lost, which is one reason falls faster than die area grows.
A CMOS start wafer is typically lightly doped p-type (100) silicon, around .5
The edge matters more than its area suggests. Deposited film thickness is well controlled across the center but poorly controlled near the edge, so edge dies fail wholesale. Because inline inspection and parametric test usually skip edge dies, that loss shows up as die yield loss even though random defects don’t cause it.6 Fabs write off an edge-exclusion ring (the simulation below uses 3 mm, an illustrative value), and yield models must separate this position-dependent loss from random defects.
Cycle time is a design input too. About 12 weeks in the fab on average, and 14–20 weeks for advanced processes,2 means a process change, or a respin, takes a quarter or more to show results. It is also why inline measurement, not end-of-line test, is the fab’s main feedback loop.
Quartz sand: Sand is mostly silicon dioxide. The silicon is refined from it and purified to at least 99.9999999%.
Every layer starts with printing a pattern. First the wafer gets a light-sensitive coating, a bit like the film in an old camera. Then a machine shines light through a stencil called a , and a lens shrinks the picture onto the wafer.7
Where light lands, the coating changes and gets washed away. What’s left is a stencil on the wafer itself. The next step only works where the coating is gone. Then the rest of the coating is stripped off.
The trouble is that light blurs when it squeezes through very tiny gaps. Light with shorter waves blurs less, so factories keep moving to shorter and shorter waves.8
The newest machines use light. It is made by blasting tiny drops of melted tin with a powerful laser. Air soaks up this light, so it has to travel through a vacuum, a space with all the air pumped out.11 One of these machines costs over a hundred million dollars.13
Each patterned layer runs the same loop:
- Spin onto the wafer and bake it.
- Expose it through the mask.
- Develop it to remove the exposed (or unexposed) resist.
- Etch or implant through the openings.
- Strip the resist.
Production tools are projection scanners: reduction optics shrink the mask image, typically four times, and the mask and wafer move together past a slit.78 The smallest printable feature, or , scales as the wavelength divided by the lens’s (NA).11 So there are two ways to print smaller: shorter light, or a wider cone of light.
The wavelength went from mercury-lamp lines at 436 nm (g-line) and 365 nm (i-line) to excimer lasers at 248 nm (KrF) and 193 nm (ArF).8 then filled the gap under the last lens with water, whose refractive index at 193 nm is about 1.44. Dry lenses top out near NA 0.93; water lets NA reach 1.35, and the smallest printable half-pitch about 36 nm.9
In theory, dense lines and spaces can’t be printed in one exposure below a pitch of , about 72 nm for immersion; in practice the limit is about 75–80 nm.9 Finer layers use , in two main forms:10
- Litho-etch-litho-etch (LELE) splits the shapes between two masks, so it depends on the two exposures lining up exactly.
- Spacer patterning (SADP, SAQP) deposits a film on the sidewalls of a first pattern, removes the original, and keeps the sidewall strips: two lines for every one printed. Doing it twice gives four.
drops the wavelength to 13.5 nm. Molten tin droplets are vaporized by a laser into a plasma that emits the light. Optics are mirrors in vacuum, at NA 0.33 today and 0.55 in high-NA tools.11 Each mirror is a stack of molybdenum and silicon layers that reflects at most about 70% of the light, so after seven mirrors only about 8% is left.12 The price of the most advanced lithography tool rose from $450,000 in 1979 to $123 million in 2019.13
The After tapeout page derives and tabulates single-exposure pitch limits. Here the question is how those limits shape a real layer stack. The ASAP7 design rule manual assigns each layer a lithography and a patterning scheme:14
| Layers | Litho | Patterning | Pitch |
|---|---|---|---|
| FIN | 193i | SAQP | 27 nm |
| GATE | 193i | SADP | 54 nm |
| ACTIVE, GCUT (fin and gate cuts), MOL (LIG, LISD, V0) | EUV | single exposure | — |
| M1–M3 and V1–V3 | EUV | single exposure | 36 nm (M1: 18 nm line + 18 nm space) |
| M4–M7 | 193i | SADP (vias LELE) | — |
| WELL, implants, threshold masks, M8–M9 | 193i | single exposure | — |
Read the table from the pitch constraints. Spacer patterning halves pitch per round without a second critical alignment, so a 27 nm fin pitch by SAQP starts from a 108 nm mandrel, and a 54 nm gate pitch by SADP also starts from 108 nm. Both mandrels sit above the 75–80 nm practical single-exposure limit of immersion.9 Spacer patterning forces every line to the same width and needs extra trim (cut) steps,10 which is why FinFET layouts are fins and gates on fixed pitches. The irregular part of the pattern moves into cut layers, ACTIVE and GCUT, which ASAP7 assigns to EUV.14
For EUV at NA 0.33 the bound gives about 20 nm pitch, but practice stops well short of it. When imec printed test patterns at the 32 nm interconnect pitch of a 5 nm process, it found stochastic defects: bridges between lines, missing holes and merged holes.15 ASAP7’s 36 nm M1 pitch sits just above that. LELE needs much tighter mask-to-mask overlay, while spacer schemes are self-aligned and need about the same overlay as a single exposure.10 ASAP7 uses LELE only for the V4–V7 via layers.14
Two physical effects drive the choice between EUV and immersion:
- Photon count. A 13.5 nm photon carries about times the energy of a 193 nm photon. At equal dose, a feature therefore receives about 14 times fewer photons. Deep-UV and EUV resists are chemically amplified: each absorbed photon releases an acid that, during the post-exposure bake, catalyzes many deprotection reactions.1612 So the randomness of photon arrival and of the resist chemistry shows up directly as line-edge roughness and occasional missing or merged features.1512
- Depth of focus. In the simple Rayleigh model, depth of focus scales as .17 Raising NA from 0.33 to 0.55 shrinks it by roughly , which tightens the flatness every polish step must deliver. High-NA tools also halve the exposure field: the optics shrink the image more in one direction only, because shrinking it more in both would have cut throughput below 100 wafers per hour.11
ArF 193 nm, dry, NA 0.93: limit 0.5λ/NA ≈ 104 nm. A 150 nm pitch prints cleanly.
Printing decides where. Four other kinds of step decide what happens there:
- Add a coat. Silicon heated in oxygen grows a skin of glass, the way iron grows rust, only perfectly even.18 Other coats are laid down from gases, some just one layer of atoms at a time.19
- Carve. A liquid eats away in every direction, like sugar dissolving in tea. A glowing, electrically charged gas can carve straight down instead, which keeps tiny shapes sharp.20
- Fire in atoms. Atoms of other elements are shot into the silicon like tiny bullets. This changes how well that patch carries electricity. A quick bake then heals the damage.2122
- Polish. A spinning pad and a gritty paste sand the wafer flat, so the next layer prints in sharp focus.23
Oxidation and deposition
Thermal oxidation grows silicon dioxide by heating the wafer to 800–1200 °C in oxygen (dry) or steam (wet). Dry oxide grows slowly but is denser and better quality. The oxide consumes silicon as it grows, so about 46% of it ends up below the original surface.18
adds films that can’t be grown from the silicon:
- Chemical vapor deposition reacts gases on the wafer.
- Sputtering knocks metal atoms off a target onto the wafer.
- Electroplating fills copper.
- pulses two chemicals in turn. Each reacts only until the surface is used up, so every cycle adds a fixed sliver of film, typically a tenth of a nanometer or less, and coats the walls of deep, narrow trenches evenly.19
Etching
Wet etching in acid is chemical: highly selective (it attacks one material and spares another) but isotropic, eating sideways under the resist as fast as down. That ruined small features, and dry etching largely replaced it. combines ions accelerated toward the wafer, which cut straight down, with reactive gas fragments that supply chemical selectivity.20
Ion implantation and annealing
ionizes dopant atoms, accelerates them to typically 5–200 keV and fires them into the wafer through openings in a mask. The energy sets the depth and the dose sets the amount. Implants knock silicon atoms out of place, so an anneal follows, often a rapid thermal anneal. It restores the crystal and moves the dopants into sites where they conduct.2122 A classic NMOS source/drain implant was arsenic at about 30 keV and a dose of ions per cm².5
Chemical-mechanical polishing
presses the wafer against a rotating pad flooded with an abrasive slurry whose chemistry softens the surface. It flattens trench fills, tungsten plugs and copper wiring so that the whole surface stays within the lithography tool’s depth of focus. It can also dish soft fill material and erode the material around it.23 How much depends on the pattern underneath, which is why the GDS & tapeout chapter has metal density rules and fill.
Each unit process is a set of trade-offs, and integration is mostly about stacking them so that one step’s side effects don’t break another.
Etch: anisotropy against selectivity
Physical sputtering by Ar⁺ transfers momentum on every collision, so its sticking coefficient is near 1 and it is anisotropic, but it hardly distinguishes one material from another. Chemical etching by radicals such as F and Cl is selective, but radicals need several steps to react, giving effective sticking coefficients near 0.01. They bounce and etch sidewalls, so the result is isotropic. Reactive ion etching uses ion bombardment to enhance chemical etching only on surfaces facing the plasma. Lowering pressure to 10–100 mTorr lengthens the ion mean free path and raises anisotropy.20 Every patterning step therefore needs a stop layer that etches much slower, such as the nitride under STI fill or the thin dielectric under the gate. That is why so many films in a flow exist only to be etched against and then removed.
Thermal budget
Dopant profiles move whenever the wafer is hot, so the flow is ordered by temperature. Older twin-well processes drove wells 2–3 µm deep with more than 8 hours above 1050 °C. Retrograde wells replace that with high-energy implants whose profile peaks below the surface, giving lower well resistance and less lateral spread.5 The activation anneal itself trades crystal repair against dopant diffusion, which is why source/drain activation uses rapid thermal annealing.22 The same logic later forces the gate-last flow described in the next section.
Conformal films
’s self-limiting chemistry gives thickness set by cycle count and conformal coverage of high-aspect-ratio features, and hafnium oxide for transistor gate dielectrics is one of its standard films.19 In nanosheet transistors it fills recesses only a few nanometers deep between stacked channels, chosen for its gap-filling ability.32
Polish
CMP removes material by height rather than through a mask. Film thickness varies, so some overpolish is always needed, and overpolish dishes soft fill material and erodes the surface around it.23 Flows design around this with stop layers: STI fill polishes down onto a nitride hard mask, which is then stripped.5 Dishing and erosion vary with the pattern underneath, which is the physical reason for the density windows in the GDS & tapeout chapter.
Isotropic: it etches sideways under the resist as fast as down. It is very selective, so it stops at the silicon.
Here is the order a factory builds a pair of switches in. The simulation below steps through the same order.
- Walls. Dig shallow trenches and fill them with glass, so neighboring switches can’t leak into each other.5
- Zones. Fire in atoms to make two zones, one for each kind of switch.
- Gate layers. Add a super-thin insulating film, then a layer that carries electricity. Each switch’s gate, the part that turns it on and off, will be cut from these.
- Print and carve. Print the gate pattern and carve away the rest.
- Ends. Fire in atoms on both sides of each gate to make the switch’s two ends. The gate itself works as the stencil, so the ends line up with it perfectly.24
- Plugs. Cover everything in glass, drill tiny holes down to the switches, and fill them with metal.
- Wires. Lay copper wires on top. Then repeat for ten or more layers of wires.1
The order matters. The wires sit on top of the switches, so they have to come later. And making the switches takes very hot bakes, hot enough to damage metal put down too early.31
A planar CMOS flow builds an NMOS and a PMOS transistor side by side. Making both kinds roughly doubles the step count of an NMOS-only process.5
- Isolation. A thin pad oxide and a silicon nitride layer are deposited. Lithography and a plasma etch cut trenches around each active area. The trenches are lined with oxide, filled with deposited oxide and polished flat, then the nitride is stripped. This is .5
- Wells. Two masked implants form a p-type for NMOS and an n-type well for PMOS.5
- Gate stack. A very thin gate insulator is grown or deposited, then the gate material over the whole wafer. For decades that was silicon dioxide under polysilicon. Below about 2 nm of oxide, electrons tunnel straight through and leakage soars. Since 2007, leading processes use a hafnium-based , which gives the same control with a physically thicker film.251
- Gate patterning. Lithography and an etch that cuts straight down leave the gates. Gate length is this layer’s critical dimension, because it sets the transistor’s drive current and leakage (see The I-V curve).
- Source and drain. A light implant forms shallow extensions. Then are formed on the gates, and heavy n+ and p+ implants, each through its own mask, form the sources and drains.5 The gate blocks the implant, which makes it a : no overlap margin is needed, so the unwanted (parasitic) capacitance between gate and source or drain, which slows the transistor, drops.24 A rapid anneal activates the dopants.
- Silicide and contacts. A metal deposited over everything reacts with exposed silicon to form a low-resistance . The unreacted metal is stripped, with no mask needed.26 An insulating layer is deposited and polished. Contact holes are etched through it, lined with titanium nitride and filled with tungsten plugs.5 These are the layers, which some manufacturers have treated as their own module since the 22 nm node.1
- Wiring. Copper began replacing aluminum in 1997 because it conducts better and resists electromigration, the slow drift of metal atoms pushed along by the current. But copper can’t be plasma-etched, so wires are made by the : etch trenches in the insulator, line them with a tantalum or titanium nitride barrier, electroplate copper and polish off the excess.2728 The insulators between wires moved to low-κ materials, which store less charge between neighboring wires (a dielectric constant typically around 2.7, and as low as 2.2, instead of about 3.8 for silicon dioxide), so signals switch faster. The reaches eleven or more levels.1
Each patterned layer needs at least one . SKY130’s documentation lists 34 masks used in that process. They cover isolation, wells, threshold adjustments, poly, the implants, a local-interconnect layer, contacts, five metals with their vias, and the pad and passivation openings.30
Gate-last high-κ/metal gate
A metal gate can’t simply replace polysilicon in the flow above. Source/drain activation needs anneals above 900 °C, which can degrade a metal gate or make it react with the dielectric.31 The gate-last (replacement-gate) flow keeps a polysilicon dummy gate through the hot steps:
- Deposit the high-κ dielectric, build a polysilicon dummy gate on it, and form the self-aligned sources and drains around it.
- Deposit an interlevel dielectric and polish it down to the top of the dummy gate.
- Etch out the dummy gate.
- Fill the trench with metal, with different work-function metals for NMOS and PMOS.131
Threshold voltage is then set by work-function metal and by masked threshold-adjust steps. ASAP7 lists separate SLVT, LVT and SRAM threshold masks among its front-end layers.14 Each extra flavor a library offers costs masks and steps.
FinFET and nanosheet modules
In a FinFET flow, the fins are a SAQP grating at 27 nm pitch in ASAP7, trimmed by an EUV ACTIVE layer. The gates are a SADP grating at 54 nm pitch, cut by GCUT. Above them, the MOL is three EUV layers: gate interconnect (LIG), source/drain interconnect (LISD) and V0 up to M1.14
Gate-all-around nanosheets reuse much of the FinFET integration. Three modules are new:32
- The superlattice. Epitaxy grows alternating Si and SiGe layers.
- The inner spacer. The SiGe ends are recessed laterally and refilled with dielectric. This cuts gate-to-source/drain capacitance and protects the source/drain epitaxy.
- Channel release. The remaining SiGe is etched away selectively so the gate can wrap the silicon sheets.
The tolerances are tight: one study reports SiGe cavity depths held to across the stack.32 The Shrinking chapter covers why these shapes win electrostatically. Here the point is that each new device shape arrives as a handful of new, very selective etch and deposition modules inside an otherwise familiar flow.
Two kits, two eras
SKY130’s 34 masks include a local-interconnect layer between contacts and M1, and five metals.3029 ASAP7’s FEOL, MOL and BEOL tables list 33 drawn layers through nine metals and the pad. Its fin, gate, M4–M7 and V4–V7 layers are multi-patterned, so the physical mask count is higher than the layer count.14 For every extra patterning pass, the price is coat, expose, develop, etch, clean and metrology steps, plus another overlay budget.
The real order: hot front-end steps first, then contacts, then cool wiring steps that the metal can survive.
With hundreds of steps, small slips add up. So factories measure all the time: how thick each coat is, how wide the printed lines are, and how well each layer lines up with the one below. Being off by just a few atoms can ruin a chip.33
Each measurement goes on a chart with two warning lines. A dot outside them means something changed, so engineers hunt for the cause.34
tracks three main things:
- Film thickness, measured optically by ellipsometry or reflectometry. Gate oxide is controlled this way.1
- , the width of key features.
- , how well a layer lines up with the one it connects to. CD and overlay measurements are essential to process control, and even a sub-nanometer misalignment can make a chip fail.33
Inspections between steps catch mis-processing such as a skipped step, a wrong recipe or a tool out of control. Just before wafers leave the fab, a parametric test measures special test structures on the wafer.6
The results feed . Each measurement is plotted on a control chart with a center line at the in-control mean and limits usually 3 standard deviations away. A point outside the limits means the process has probably changed, and the cause is investigated.34
Losses come in two kinds. Line yield is wafers scrapped for damage or mis-processing. Die yield is dies that fail on wafers that made it through.6
Fab data is hierarchical, and that shapes how it is charted. The NIST/SEMATECH handbook’s lithography case study measures line width at 5 sites per wafer, on 3 wafers per cassette, across 30 cassettes: 450 measurements.35 The largest variance component is lot to lot. A Shewhart chart with limits from within-lot variation therefore flags lots whose shift already has a known, assignable cause. Yet the measurement of interest is still the site level.36 One remedy it discusses is to chart each nested source at its own level, at the cost of more charts.36 In variance terms, . Separating the components tells engineers whether to look at what changes between lots, between wafers in a lot, or across a single wafer.
Yield data then splits loss by its signature. Leachman writes die yield as . is the random, defect-limited yield, which only cleaner processes and tools raise. is the systematic-limited yield, from mechanisms with spatial or lot signatures such as edge loss, which better process execution and process control fix.6 The After tapeout page follows the loop from there into wafer maps, scan diagnosis and failure analysis.
Overlay deserves its own budget. With LELE, two masks on the same layer must register to each other as well as to the layers below.10 At sub-nanometer tolerances every contributor counts: the scanner, the masks, the process steps that distort the wafer, and the measurement itself.33
In control: points scatter randomly within ±3σ of the center line. No action needed.
The top half builds a pair of switches. Press Next step to move through the seven steps and watch the layers appear. The newest layer is outlined.
At any step, press Drop a particle, then keep stepping to see what the dust speck does. Try it during the wiring step. Then reset and try it while the zones are made.
The bottom half shows a whole wafer. Make the chip bigger and watch: fewer chips fit, and more of them break. Then try the “New factory” button, which has more dust specks.
The cross-section steps through a simplified CMOS inverter flow: isolation, wells, gate stack, gate lithography and etch, spacers and source/drain implants, silicide and contacts, then first metal. The status bar counts masks as you go (8 in this simplified flow). Drop a particle at any step shows whether that defect kills the die.
The wafer map places square dies on a 300 mm wafer with a 3 mm edge exclusion and colors them with the Poisson model, . The outlined die is the one in the cross-section. Things to try:
- Move die area from 100 mm² to 600 mm² at . Dies per wafer fall about seven-fold, from 616 to 89, while yield falls from 90% to 55%.
- Raise to 0.5 and compare the drop in good dies for small and large dies.
- Watch the cost per good die (at an illustrative $10,000 wafer) climb much faster than area.
The cross-section is an eight-state model of a planar CMOS inverter flow, with the module (FEOL, MOL, BEOL) and a recipe line for each step. A particle dropped at a step maps to that step’s failure mode: a blocked STI etch (leakage), a well-implant shadow (benign), a gate-dielectric inclusion (latent), an etch-masking stub, a blocked p+ implant, a blocked contact (open), or an M1 bridge (short).
The yield panel counts whole dies by placement, not by formula, and shows the formula estimate alongside. Try these:
- Switch between Poisson, Murphy and negative binomial. Failing dies are drawn clustered for the compound models.
- At 600 mm² and , compare 5.0% (Poisson) with 10.0% (Murphy) and 12.5% (negative binomial, ).
- Push toward 10 and watch the negative binomial converge on Poisson.
- Process steps
- 300+
- Time in the fab
- ~12 weeks
- Standard wafer
- 300 mm × 775 µm
- EUV wavelength
- 13.5 nm
Sources: step count from the fabrication overview and the Semiconductor Industry Association, which also gives the average cycle time; wafer size from Mack’s course notes; EUV wavelength from IEEE Spectrum.12411
What these numbers mean:
- 300+ steps means each wafer is coated, printed, carved and polished hundreds of times. Say each step works 999 times out of 1,000. After 300 steps, about one wafer in four would still have a problem. That is why every step is measured.
- About 12 weeks in the factory means a fix to the recipe takes a whole season to show up in finished chips.
- 300 mm (30 centimeters) is the width of a wafer, about the size of a dinner plate.
- 13.5 nm is the length of one wave of extreme ultraviolet light. A nanometer (nm) is a millionth of a millimeter. That wave is about 30 times shorter than violet light, the shortest light our eyes can see.
Bigger chips cost much more than their size suggests. In one example from 1994, a small chip came out working 71% of the time and cost about $4 to make. A chip seven times bigger worked only 9% of the time and cost about $417.37
The shrinking wavelength
| Source | Wavelength | What changed |
|---|---|---|
| Mercury g-line, i-line | 436, 365 nm | Lamp-based steppers; NA rose from 0.28 to 0.658 |
| KrF excimer laser | 248 nm | Deep UV; steppers and scanners from 19888 |
| ArF, dry | 193 nm | Scanners from 1998; a dry lens tops out near NA 0.9389 |
| ArF immersion | 193 nm | Water allows NA up to 1.35; about 36 nm half-pitch9 |
| EUV | 13.5 nm | Mirrors in vacuum, NA 0.33; high-NA 0.5511 |
Yield and dies per wafer
Counted on a 300 mm wafer with 3 mm edge exclusion and 0.1 mm scribe lanes (the simulation’s placement), with Poisson yield. Arithmetic, not measured data:
| Die area | Whole dies | Yield, | Good dies | Yield, | Good dies |
|---|---|---|---|---|---|
| 50 mm² | 1,240 | 95% | ~1,180 | 78% | ~966 |
| 100 mm² | 616 | 90% | ~557 | 61% | ~374 |
| 200 mm² | 300 | 82% | ~246 | 37% | ~110 |
| 600 mm² | 89 | 55% | ~49 | 5% | ~4 |
Twelve times the area gives roughly 24 times fewer good dies at , and more than 200 times fewer at .
Wafer prices
A 2020 Georgetown CSET model estimated what a foundry charges for a 300 mm wafer. These are modeled estimates, not price lists:13
| Node | 90 nm | 28 nm | 7 nm | 5 nm |
|---|---|---|---|---|
| Price per wafer | $1,650 | $2,891 | $9,346 | $16,988 |
Historical die economics
Baas’s lecture notes tabulate 1994 parts, all on wafers costing $900–$1,700:37
| Chip | Area | (cm⁻²) | Dies/wafer | Yield | Die cost |
|---|---|---|---|---|---|
| 386DX | 43 mm² | 1.0 | 360 | 71% | $4 |
| 486DX2 | 81 mm² | 1.0 | 181 | 54% | $12 |
| PowerPC 601 | 121 mm² | 1.3 | 115 | 28% | $53 |
| DEC Alpha | 234 mm² | 1.2 | 53 | 19% | $149 |
| Pentium | 296 mm² | 1.5 | 40 | 9% | $417 |
A 7× increase in area gave a roughly 100× increase in die cost. Defect densities of 1–1.5 per cm² were survivable only because dies were small, and the yields in this table sit well above the Poisson prediction. At 296 mm² and 1.5 cm⁻², Poisson gives , against 9% reported. That gap is what the clustered models in Under the hood account for.
Patterning cost by layer (ASAP7)
Of ASAP7’s 33 drawn layers, the FEOL has 11, the MOL 3, and the BEOL 19 (M1–M9, V1–V9 and the pad).14 The patterning breaks down like this:
- EUV single exposure: 12 layers (ACTIVE, GCUT, SDT, the 3 MOL layers, M1–M3 and V1–V3).
- 193i SAQP: 1 layer (FIN).
- 193i SADP: 5 layers (GATE, M4–M7).
- 193i LELE: 4 via layers (V4–V7).
- 193i single exposure: the rest.
Spacer layers also need cut or block masks, and LELE layers need two exposures, so the mask count exceeds the layer count. Lithography tool prices went from $450,000 in 1979 to $123 million in 2019.13 This is why mask and exposure count, not layer count, drives wafer cost at leading nodes.
- One costly pass or several cheap ones. The newest light machines print the finest patterns in one go, but cost a fortune. Older machines can print the same pattern in several passes. Each extra pass takes time and adds a chance for layers to line up badly.10
- Big chip or several small ones. A big chip is more likely to catch a flaw, so fewer of them work. Some designers split one big chip into several small ones and join them together. See Packaging and chiplets.
- Newest factory or a proven one. A brand-new factory process costs far more per wafer, and at first it takes extra work to get most chips working.13 Many chips don’t need the newest one at all.
What goes wrong? Dust, a layer printed slightly out of line, a step that runs a little hot, a polish that digs too deep. The constant measuring catches most problems. The rest show up when the finished chips are tested.6
EUV against multiple patterning
Below about 30 nm, a layer needs either several immersion passes or one EUV exposure.11 Multiple patterning uses cheaper exposures, but each extra pass adds coat, expose, etch and clean steps, and LELE adds a mask-to-mask overlay error.10 EUV scanners cost far more,13 and at the finest pitches random (stochastic) variation in the light and the resist causes defects of its own.15
Die size against yield
Yield falls exponentially with area in the simplest model, and dies per wafer fall faster than 1/area because of edge loss. Cost per good die therefore grows much faster than area.37 The exposure field also caps die size at the , and high-NA EUV halves that field.11 Splitting a design into smaller dies helps yield but adds packaging cost and die-to-die links. The Wafer-scale chapter shows the opposite bet: redundancy inside one huge die.
Mature against leading-edge processes
Modeled wafer prices rise about ten-fold from 90 nm to 5 nm.13 A leading node only pays off when its denser, faster, lower-power transistors are worth that much more. Many chips that don’t need the densest logic are built on mature processes like SKY130 instead.
What goes wrong
- Random defects. Shorts and opens caused by particles, excess metal bridging over steep steps, resist splatters and flakes, weak spots and pinholes in insulators, poor step coverage, and scratches.6
- Mis-processing. A skipped or duplicated step, the wrong recipe, or a tool out of control. Usually caught by inline inspection or parametric test, and the whole wafer is lost.6
- Edge effects. Poorly controlled films near the wafer edge.6
- Misalignment. errors break contacts or short neighbors.33
- Polish non-uniformity. dishing and erosion.23
Regularity against freedom
Spacer patterning buys pitch with perfectly periodic gratings and moves the design freedom into cut layers.1014 The cost lands in the cell library: fixed gate and fin pitches, unidirectional low metals, and coloring rules on multi-patterned layers. The cell library chapter picks this up.
Thermal budget against materials
Gate-last high-κ/metal gate exists because a metal gate can degrade in the source/drain activation anneal.31 The price is extra deposition, polish and etch modules, and work-function metals that must fit into ever-narrower replacement trenches. Each new channel material or backside-power scheme reopens the same question of what can be hot, and when.
Model choice against risk
The same gives very different large-die yields under different models: at 600 mm² and 0.5 cm⁻², 5% (Poisson) against 10% (Murphy) and 12.5% (negative binomial, ). Yield models are valid where they were fit. Poisson is accurate for small dies, about 0.25 cm² or less, and for below 1. It underestimates yield for large dies because defects cluster.6 Pricing a large die with a Poisson fitted on small test chips is pessimistic. Pricing it with an assumed nobody has measured can be optimistic.
Failure modes the models miss
- Latent defects. A particle in the gate dielectric may pass every test and fail in the field. This is the reliability side of the After tapeout page.
- Stochastic defects. EUV missing or merged features don’t come from particles, so particle-based doesn’t capture them, and they become a problem as pitch shrinks.15
- Systematic, layout-dependent failures. Lithography hotspots and CMP-sensitive patterns repeat in every die with that pattern and land in , not .6
Two immersion exposures, each printing every other line at twice the pitch. Perfect overlay gives even spaces; try an alignment error.
This part goes deeper, into the math, models and algorithms behind the chapter. It’s written for the Expert level.
The Poisson yield model and per-layer defect budgets
If killer defects land independently with mean per die, the chance a die has defects is . A die works only with , so . Inverting gives , which is how fabs quote defect density from observed yield.6 The useful property is additivity. If over layers or steps, then . So the yield gain from cutting one layer’s defectivity by is simply a factor of .6 That is how a fab turns inline inspection counts per layer into a yield budget.
Compound Poisson: Murphy, Seeds and negative binomial
Real defect density varies from die to die, wafer to wafer and lot to lot. Murphy’s idea was to average the Poisson yield over a distribution with mean :6
- Uniform on gives .
- A triangular on , peaking at , gives the Murphy model, .
- An exponential gives Seeds’ model, .
- A gamma gives the negative binomial, , where is the cluster parameter. can be estimated from the mean and standard deviation of defect counts per die.6
The negative binomial spans the others. With of about 10 or more it is essentially Poisson, with it approximates Murphy, and with it approximates Seeds.6 Baas’s course notes use for modern CMOS.37
Why does clustering raise yield? is convex in , so by Jensen’s inequality the average of over any spread of is at least . Concentrating defects on some dies spares others. The effect is negligible when is small and large when is several. In the simulation, a 600 mm² die at 0.5 cm⁻² () gives 5.0% under Poisson, 10.0% under Murphy and 12.5% under the negative binomial with . A 100 mm² die () gives 60.7%, 61.9% and 63.0%.
Systematic loss multiplies on top: . Fitting against across die sizes gives as the intercept and as the slope.6 The windowing method on the After tapeout page generates those different sizes from one product’s wafer map.
Dies per wafer
The textbook estimate is .37 The first term is wafer area over die area. The second is a correction for edge loss: the circumference divided by , the diagonal of a square die. Roughly one die is lost per diagonal-length of edge. For and , it gives .
The simulation instead places square dies on a grid with 0.1 mm scribe lanes. It keeps only dies that lie wholly inside a 147 mm radius (3 mm edge exclusion) and tries four grid offsets, keeping the best. That gives 616 at 100 mm², a little below the formula because of the exclusion ring and scribe lanes. Real die counts also lose sites to test structures and alignment marks.
Cost per good die
.37 Both factors in the denominator fall with area: roughly as , and exponentially in Poisson or as a power in the negative binomial. So die cost rises much faster than area. Baas’s notes summarize it as a steep function of die area, of order , for the defect densities of the day.37 The exponent isn’t a law. It depends on : for small, mature-process dies, cost is close to linear in area, and the super-linear penalty only bites once approaches 1. Package, test and assembly yield multiply on top, which is the argument for known-good-die testing in multi-die products.
Spacer patterning, step by step
- Print a mandrel grating at pitch with one immersion exposure.
- Deposit a conformal film of thickness , then etch it anisotropically. Material remains only on the mandrel sidewalls.
- Remove the mandrel. Two spacer lines now sit in every pitch , so the pitch is .
- For SAQP, use those spacers as the next mandrel and repeat, giving .
SAQP step 1 of 7. Nominal core CD.
Line width is set by the spacer film, not by lithography, which is why spacer-defined fins are so uniform and every line has the same width. Line placement depends on the mandrel’s CD and the spacer thickness, not on a second critical alignment.10 The catch is that the result is a uniform grating. Ends, gaps and jogs need cut masks, which ASAP7 puts on EUV (ACTIVE for fins, GCUT for gates). Its 27 nm fins from a 108 nm mandrel and 54 nm gates from a 108 nm mandrel follow directly.14
Control limits for nested data
A Shewhart chart puts its center line at the in-control mean and its limits at , where is the standard deviation of the plotted statistic.34 With nested sampling, the plotted statistic’s variance depends on what is plotted. Lot means of wafers × sites have variance
If the limits are computed only from the within-lot terms and dominates, many lots fall outside the limits even though their variation is normal for the process, and engineers learn to ignore the chart. Charting each level separately, one of the options the NIST case discusses, avoids that.36
Q1Why does a 193 nm immersion scanner put water between the lens and the wafer?
Q2What makes a gate “self-aligned” to its source and drain?
Q3With a defect density of 0.5 per cm², what does the Poisson model give for the yield of a 2 cm² die?
Q4Why can’t copper wires be made the way aluminum wires once were, by depositing a sheet and etching it?
Sources
Show Hide 37 sources
- Semiconductor device fabricationDeposition, removal, patterning and modification of electrical properties; over 300 steps and eleven or more metal levels; 11–13 weeks average at advanced nodes; 300 mm wafers from 2000; MOL since 22 nm; high-k/metal gate at 45 nm in 2007; gate-last flow; low-κ around 2.7 (as low as 2.2) against 3.82 for SiO₂; FOUP mini-environments; ellipsometry and reflectometry.
- Chipmakers Are Ramping Up Production to Address Semiconductor Shortage. Here’s Why that Takes TimeWafer cycle time about 12 weeks on average, up to 14–20 weeks for advanced processes; up to 1,400 process steps depending on the complexity of the process.
- Wafer (electronics)Czochralski ingots; 9N purity; starting doping with boron, phosphorus, arsenic or antimony; resistance to the 450 mm transition over return on investment.
- Lecture 4: Single-Crystal Silicon (CHE323/CHE384, Chemical Processes for Micro- and Nanofabrication)Czochralski growth; ingots sliced and polished; 300 mm wafers 775 µm thick; wafers pre-doped p- or n-type and cut along (100), (110) or (111) planes.
- Lecture 18: CMOS (3.155J/6.152J Micro/Nano Processing Technology)NMOS and CMOS process flows: lightly doped (100) p-type start, gate oxide under 10 nm, arsenic S/D implant ~30 keV and 5×10¹⁵ cm⁻², CMOS doubling the step count, twin and retrograde wells, LOCOS vs STI steps, LDD, spacers, silicide, W plugs, planarization.
- Yield Modeling and Analysis (IEOR 130 course notes)Line vs die yield; kinds of killer defect; edge loss; Poisson model and per-layer additivity; validity for small dies; Murphy, Seeds and negative binomial models; α ≥ 10 ≈ Poisson, α = 5 ≈ Murphy, α = 1 ≈ Seeds; systematic-limited yield.
- Lecture 39: Lithography: Process Overview (CHE323/CHE384)Coat, prebake, expose, post-exposure bake, develop, metrology; the resist must resist the etch or implant and then be stripped; masks typically 4× the wafer pattern; the wafer is stepped and/or scanned under the lens.
- Lecture 40: Lithography: Imaging Tools (CHE323/CHE384)g-line 436 nm and i-line 365 nm lamps; KrF 248 nm and ArF 193 nm excimer lasers; steppers from g-line NA 0.28 to i-line NA 0.65; deep-UV steppers and scanners from 1988, ArF scanners from 1998, immersion up to NA 1.35; step-and-scan through a slit, used by all state-of-the-art tools.
- Lecture 48: Lithography: Resolution and Immersion (CHE323/CHE384)R = k1·λ/NA with k1 ≥ 0.25; dry lenses top out near sin θ ≈ 0.93; water’s index is 1.436 at 193 nm; NA 1.35 gives a 36 nm half-pitch limit; practical single-exposure pitch 75–80 nm.
- Lecture 59: Lithography: Double Patterning (CHE323/CHE384)Single-exposure limit of 75–80 nm pitch; LELE costs two exposures and needs much tighter overlay; SADP grows sidewall spacers on a dummy pattern, needs only one critical exposure and overlay like single patterning, but forces equal linewidths and needs trim (cut) steps.
- This Machine Could Keep Moore’s Law on Track13.5 nm light from tin droplets hit by a CO₂ laser; absorbed by air, so vacuum and reflective optics; CD proportional to λ/NA with k1 ≥ 0.25; features below 30 nm need multiple patterning or a shorter wavelength; NA 0.33 to 0.55; anamorphic optics halve the field to keep throughput from falling below 100 wafers per hour.
- Lecture 60: Lithography: Extreme Ultraviolet (CHE323/CHE384)13.5 nm needs vacuum and all-reflective optics; Mo/Si multilayer mirrors reflect at most about 70%, so seven reflections pass about 8%; EUV resists are chemically amplified like 248 and 193 nm resists; line-edge roughness at low dose is the biggest resist problem.
- AI Chips: What They Are and Why They MatterLithography tool cost from $450,000 (1979) to $123 million (2019); new nodes need extra engineering to bring yields up; modeled foundry sale price per 300 mm wafer from $1,650 (90 nm) to $16,988 (5 nm).
- ASAP7 PDK Design Rule Manual, release 1p7FEOL, MOL and BEOL layer tables with the lithography (193i or EUV) and patterning (single exposure, SADP, SAQP, LELE) assumed for each layer; 27 nm fin pitch, 54 nm gate pitch, 18 nm M1 width and spacing.
- Getting EUV Ready for 2020imec’s EUV test patterns at the 32 nm interconnect pitch of a 5 nm process showed stochastic defects: bridges, missing holes and merged holes; photon shot noise and uneven resist chemistry as causes; more dose helps but not enough, and slows the scanner.
- Lecture 51: Lithography: Chemically Amplified Resists, part 1 (CHE323/CHE384)Exposure makes a photoacid generator release acid; during the post-exposure bake the acid catalyzes the deprotection reaction that changes solubility (amplification); acid diffusion and loss matter.
- Lecture 46: Lithography: Defocus and DOF (CHE323/CHE384)Rayleigh depth of focus DOF = k2·λ/NA² in the paraxial (low-NA) approximation; smaller pitches lose more image to defocus.
- Thermal oxidationOxidation at 800–1200 °C in oxygen (dry) or steam (wet); about 46% of the oxide lies below the original surface; dry oxide is denser and better quality.
- New development of atomic layer deposition: processes, methods and applicationsPrecursors pulsed in sequence; self-limiting half-reactions deposit a (sub)monolayer per cycle, often around 0.1 nm; sub-nanometer thickness control and conformal coating of high-aspect-ratio structures; HfO₂ deposited by ALD for MOSFETs.
- Dry Etching (3.155J/6.152J lecture notes)Wet etch is chemical, isotropic and selective; physical sputtering is anisotropic but unselective; reactive ion etching combines directionality and selectivity; sticking coefficients; lower pressure raises anisotropy.
- Diffusion/Implantation (3.155J/6.152J lecture notes, September 28, 2005)Ion energies typically 5–200 keV; dose and depth both controlled; implant through openings in a mask; implant damage and amorphization; a post-implant anneal restores atoms to lattice sites and activates the dopant, but also diffuses it.
- Lecture 18: Ion Implantation, part 3 (CHE323/CHE384)Channeling along crystal axes; an anneal regrows the crystal and activates the dopant; too cool leaves defects, too hot diffuses too much, so the best compromise is rapid thermal annealing.
- Lecture 30: Chemical Mechanical Polishing (CMP) (CHE323/CHE384)Topography costs lithography depth of focus; CMP gives global planarization; rotating pad and wafer under pressure with a silica or alumina slurry whose chemistry softens the surface; used for STI, tungsten plugs and copper damascene; overpolish, dishing and erosion.
- Self-aligned gateThe gate is the mask for the source and drain doping, removing the need for gate overlap and cutting parasitic capacitance; polysilicon gates from 1968.
- High-κ dielectricTunneling leakage rises sharply as SiO₂ thins below about 2 nm; hafnium-based films by ALD; Intel’s 45 nm high-κ/metal gate in 2007.
- SalicideDeposit a transition metal, heat so it reacts only with exposed silicon, strip the unreacted metal; Ti, Co, Ni; no extra lithography step.
- Copper interconnectsLower resistance and better electromigration than aluminum; copper can’t be plasma-etched, hence the damascene process; barrier layers; IBM in 1997.
- Lecture 31: Copper Dual Damascene (CHE323/CHE384)Copper replaced aluminum in the 1990s; it can’t be plasma-etched because its reaction products aren’t volatile, so it is inlaid by damascene; TiN or Ta barrier, thin copper seed, electroplating, CMP.
- SKY130 process backgroundA mature 180–130 nm hybrid technology developed by Cypress Semiconductor; five levels of metal.
- SKY130 mask listMasks used in SKY130: field oxide, wells, threshold-adjust, poly, tip and source/drain implants, local interconnect, contacts, five metals and vias, passivation and pad layers.
- Dry etch polysilicon removal for replacement gates (US 8,673,759 B2)In a replacement-gate flow a polysilicon dummy gate stays until the high-temperature source/drain activation anneal, then is replaced by metal, because annealing above 900 °C can degrade a metal gate.
- A Comprehensive Study of NF₃-Based Selective Etching Processes: Application to the Fabrication of Vertically Stacked Horizontal Gate-All-around Si Nanosheet TransistorsNanosheet flows reuse much FinFET integration; distinctive modules are the Si/SiGe superlattice, the inner spacer and channel release; inner spacers cut gate-to-S/D capacitance and protect the S/D epitaxy; cavity depth ≤ 5 ± 0.3 nm.
- Overlay Metrology Using Physics and AI-Based Scanning Electron MicroscopyOverlay and CD measurements are essential for process control; even a sub-nanometer misalignment can make a chip non-functional.
- What are Control Charts? (NIST/SEMATECH e-Handbook of Statistical Methods, 6.3.1)Center line at the in-control mean, upper and lower control limits usually at 3 sigma; a point outside means the process is probably out of control.
- Lithography Process: Background and Data (NIST/SEMATECH e-Handbook, 6.6.1.1)Line width measured at 5 sites per wafer, 3 wafers per cassette, 30 cassettes: 450 measurements.
- Lithography Process: Shewhart Control Chart (NIST/SEMATECH e-Handbook, 6.6.1.4)Lot-to-lot variation dominates, so charts built on within-lot variation flag lots whose cause is already known; options for charting nested sources of variation.
- Cost (EEC 116 lecture handout)Dies per wafer = π(d/2)²/A − πd/√(2A); yield (1 + D·A/α)^−α with α ≈ 3; die cost = wafer cost ÷ (dies per wafer × yield); 1994 examples from a 43 mm² die at 71% yield to a 296 mm² die at 9%.