HomeSemiconductorsTransistor Evolution

Transistor Evolution

From Point-Contact (1947) to 3nm GAA (2026): Complete Technical History

By EcrioniX · Updated Jul 30, 2026

The transistor is the most important invention of the 20th century. Everything digital flows from a simple idea: use a small voltage to control a larger current. From the first fragile point-contact transistor in 1947 to billions of 3nm GAA transistors on a single die in 2026, the evolution has been relentless. This is the complete technical history—how transistors work, why we had to reinvent them every few years, and where we're headed as physics starts to push back against shrinking.

Part 1: The Birth — Point-Contact Transistor (1947)

December 16, 1947 — Bell Labs, New Jersey
John Bardeen and Walter Brattain press two gold contacts onto a germanium crystal and discover transistor action: a small signal can control a much larger current. The device amplifies. Solid-state amplification is born.

The first transistor was nothing like modern transistors. It consisted of:

How it worked: Forward bias on one contact created a small current. This current modulated the resistance between the two contacts, controlling a larger current. Gain: ~100x. But the device was unreliable, had high noise, and couldn't handle much power.

Why it mattered: Vacuum tubes (the previous amplifiers) were huge, hot, power-hungry, and had limited lifespan. A solid-state device that could do the same job promised to shrink electronics dramatically. The semiconductor age was born.

Part 2: The BJT Era (1950s-1970s) — Scaling Begins

William Shockley (also at Bell Labs) invented the junction transistor (BJT — Bipolar Junction Transistor) in 1950, a much more reliable design. BJTs used two junctions (N-P-N or P-N-P) and were easier to manufacture and understand.

How a BJT Works

// BJT Amplification Principle // Small base current Ib controls large collector current Ic // Ic = β * Ib (β ≈ 100-300) // NPN BJT: // Base input (0.7V forward bias) → emitter-base junction conducts // Small Ib flows → allows large Ic to flow from collector to emitter // Ic/Ib = β (current gain) // Power gain = voltage gain × current gain

BJTs ruled from the 1950s through 1980s. They were current-controlled devices (a small base current controlled a large collector current) and dominated logic, analog, and RF applications. By the 1960s, multiple transistors were integrated onto a single chip (integrated circuits), and the race to pack more transistors began.

Key limitation of BJTs: Both electrons and holes carried current (bipolar = two carrier types), leading to high power consumption. As circuits got denser, heat dissipation became a nightmare.

Part 3: The MOSFET Revolution (1960s-1980s) — CMOS Takes Over

In 1960, Dawon Kahng and Mohamed Atalla at Bell Labs invented the MOSFET (Metal-Oxide-Semiconductor Field-Effect Transistor). It was simpler than the BJT, used only one carrier type (unipolar), and crucially: it was voltage-controlled, not current-controlled.

How a MOSFET Works

// NMOS Transistor (Simplified) // Voltage at gate controls current through channel // Operating principle: // Vgs (gate-source voltage) creates electric field // Field attracts electrons to channel region // Forms conducting channel between drain and source // Current Id flows proportional to (Vgs - Vth)^2 // Vth = threshold voltage (~0.4-1.0V modern tech) // Key advantage: // Gate draws NO current (only capacitive charging) // → Much lower power consumption vs BJT // → Can pack billions of transistors // Scaling rule: // Shrink W (width) and L (length) proportionally // More transistors per unit area // → Moore's Law enabled

NMOS (n-channel) logic was the first widespread MOSFET technology (Intel 4004, 1971). But NMOS had a major problem: it consumed static power because pull-up transistors needed to always drive high.

CMOS (Complementary MOS) solved this in the late 1970s: use both NMOS and PMOS transistors in complementary pairs. When NMOS pulls low, PMOS is off (and vice versa). Result: zero static power consumption (only dynamic power during switching). CMOS enabled the explosion of integrated circuits and mobile computing.

Why MOSFET dominated:

Part 4: The Planar MOSFET Era (1970s-2010s) — Moore's Law in Full Effect

Gate Length Scaling: Moore's Law in Action (1971-2020)
1 nm 100 nm 1000 nm 1971 2020 10μm 5nm Year

Moore's Law held for nearly 50 years: Gate length roughly halved every 2 years. A planar MOSFET is a 2D device: gates on top of the silicon surface. As we shrunk the gate length (L), the distance between drain and source got shorter, but the channel depth stayed the same. The result: problems.

The Planar MOSFET Problem

Below ~22nm gate length, planar MOSFETs started to break:

The semiconductor industry faced a wall: planar technology couldn't scale past ~7nm without leakage and variability problems exploding.

Part 5: The 3D Revolution — FinFET (2011-2017)

Intel's answer: stop scaling in 2D, start scaling in 3D. FinFET (Fin Field-Effect Transistor) puts the channel on a vertical "fin" of silicon. The gate wraps around three sides of the fin (top and two sides), giving the gate much better control over the channel.

PLANAR MOSFET (22nm): FINFET (14nm):

Gate Gate Gate Gate Gate Gate
▼ ▼ ▼ ▼ ▼ ▼
═════════════════ ║ ║ ║
Channel (2D, one layer) Channel (3D, wrapped)

Gate touches top only. Gate wraps 3 sides.
Loose control. Tight control.

FinFET advantages:

FinFET tradeoffs:

FinFET kept Moore's Law alive from 14nm (2014) through 5nm (2020). But at 5nm, even FinFET started hitting limits: fin width variation, quantum tunneling, power density.

Part 6: The Current Era — Gate-All-Around (2022-2026)

FinFET has a fin on each side and gaps between fins. What if the gate wrapped all four sides? That's GAA (Gate-All-Around).

FINFET (5nm): GAA (3nm):

Gate Gate Gate Gate wraps
▼ ▼ ▼ all sides
║ gap ║ gap ║
Channel fins ⊙ (cross-section view
gate completely surrounds
Gate on 3 sides. nanowire channel)
Gaps let leakage escape.
Gate on 4 sides.
Perfect control.

GAA advantages:

GAA tradeoffs:

TSMC 3nm (2022) and Samsung 3nm GAE (2023) both deployed GAA. Intel 20A (2024-2025) uses a hybrid (gate-surrounds-nanowire). By 2026, GAA is the standard for <2nm nodes.

Part 7: Node Progression Summary (1947-2026)

EraTechnologyGate LengthYearTransistors/mm²Key Challenge
1947Point-Contact~50 μm19471 (discrete)Reliability
1950s-60sBJT1-10 μm1960~10Power consumption
1960s-80sMOSFET5-10 μm1971~100Static power (NMOS)
1980s-2000sCMOS Planar0.1-1 μm1985~1MGate length scaling
2000-2010Planar Sub-micron22-90 nm2003~1BShort-channel effects
2011-2017FinFET14-7 nm2014~100BFin variation, leakage
2018-2020FinFET (Late)7-5 nm2017~300BQuantum tunneling
2022-2026GAA (Nanosheet)3-2 nm2022~1T+Nanowire variability
2026+GAA (Ultimate)~1 nm2026+~10T?Quantum effects

Part 8: Why Transistors Keep Getting Better (And Harder)

The Good News: Scaling Benefits

The Bad News: Physics Limits

Part 9: The Future — Can We Scale Beyond 1nm?

The transistor roadmap hits hard physics limits around 2030-2035. Options being researched:

Reality check: Moore's Law as we know it (halving gate length every 2 years) is probably dead by 2030. But transistor improvement will continue—just through different mechanisms (3D stacking, new materials, specialized architectures) rather than simple scaling.

Part 10: FAQ

What was the first transistor ever made?

The point-contact transistor (Dec 16, 1947, Bell Labs). Two gold contacts pressed on a germanium crystal. Unreliable and power-hungry, but it worked. 79 years later, we're still using the same principle (controlling current with voltage), just in 3D with nanometer-scale gates.

Why did we switch from BJT to MOSFET?

BJTs are current-controlled (gate draws current), while MOSFETs are voltage-controlled (gate draws no current). CMOS (complementary MOS) enabled near-zero static power consumption, which was critical for scaling to billions of transistors. BJTs still used in analog and RF because they have high transconductance and good noise properties.

How many transistors are in a modern chip?

TSMC 3nm (2022-2024): ~300-400 billion transistors per mm². Apple M2 Max: ~20 billion. NVIDIA H100: ~80 billion. By 2026, chips are reaching 1 trillion transistors total. For perspective: 1947 transistor was discrete. 2026 chip has 10^12 of them on one die.

Is Moore's Law dead?

As originally stated (transistor count doubles every 2 years via scaling), yes—we're hitting physics limits. But transistor improvement continues (3D, new materials, architectural innovation). Growth will be slower post-2030, probably 1.5x every 2 years instead of 2x. The era of simple planar scaling is ending; the era of heterogeneous integration and new materials is beginning.

What's the smallest possible transistor?

Theoretically: single-electron transistor where one electron's position controls conduction. Practically: ~0.3nm (3 atoms wide) before quantum tunneling makes on/off distinction meaningless. We're at 3nm effective gate length now (2026). Expect to plateau at 1-2nm by 2030 unless breakthrough architecture emerges (3D monolithic stacking, new materials, etc.).

Can we make transistors smaller forever?

No. Physics says no. Quantum tunneling, thermal noise, atomic granularity, and power density create hard walls. Best estimate: transistor improvement slows dramatically after 2035. Future gains come from 3D integration (stacking transistor layers), specialized architectures (neuromorphic, analog), or completely new paradigms (photonics, quantum), not smaller size.