🏆 AlphaGo to AlphaFold

How DeepMind Went from Games to a Nobel Prize

📖 Reading Time: 20-25 minutes 📊 Difficulty: Beginner 💻 Code Examples: 0 📝 Exercises: 0

AI Terakoya Top›Machine Learning Dojo›AlphaGo to AlphaFold

🌐 EN | 🇯🇵 JP | Last sync: 2026-08-19

← Back to Machine Learning Dojo

🎯 Series Overview

A program learned to play a board game. Less than a decade later, the line of work that program started shared a Nobel Prize in Chemistry. This series is about how that happened, what actually transferred between the two problems, and — the part most accounts skip — what was never solved at all.

The route runs from the game to the science. We set up why games were the proving ground for AI and why Go resisted the methods that beat chess; assemble AlphaGo from a learned evaluation and Monte Carlo tree search, and see how each half compensates for the other's weakness; watch the successors delete the human data and then the game-specific machinery, getting stronger with each subtraction; move to AlphaFold, where self-play is impossible and the ground truth had to come from somewhere else entirely; and finish with the legacy — the AlphaFold Database, the 2024 Nobel Prize in Chemistry, the ripples into materials science, and an unsparing account of the open problems.

This is a history and principles series rather than an implementation course. The through-line is a single question — where does the feedback signal come from? — and the answer is what explains the apparently contradictory arcs of the two systems: AlphaGo could throw away human knowledge because the rules of Go supply unlimited free ground truth, while AlphaFold had to build a channel to the ground truth that evolution had already recorded.

It is written for machine learning learners who want the conceptual spine behind the headlines, and for materials and life-science researchers who keep encountering these systems as analogies for their own work and want to know how far the analogy carries. Within the Machine Learning Dojo it complements Introduction to Reinforcement Learning — which develops the algorithms this series treats historically — along with Introduction to Graph Neural Networks and Introduction to Transformers, whose architectural ideas appear throughout the AlphaFold story.

A promise about hype, stated up front. This series names no result it cannot state precisely, quotes no number that is disputed, and does not describe any benchmark victory as a solved science. Where a claim is contested — the counts announced by AI-driven materials-discovery pipelines, for instance — it is reported as contested rather than repeated. Where a system is genuinely remarkable, it is said plainly. The organizing value is calibration over allegiance: neither treating each announcement as a formality on the way to everything, nor treating every result as marketing.

Learning Path

flowchart LR A["Chapter 1
Games as the
Proving Ground"] B["Chapter 2
AlphaGo: Search
Meets Learning"] C["Chapter 3
Zero and Beyond:
Learning Without Humans"] D["Chapter 4
AlphaFold: The Protein
Folding Breakthrough"] E["Chapter 5
The Legacy: From
Games to Science"] A --> B --> C --> D --> E style A fill:#667eea,stroke:#764ba2,stroke-width:2px,color:#fff style B fill:#667eea,stroke:#764ba2,stroke-width:2px,color:#fff style C fill:#667eea,stroke:#764ba2,stroke-width:2px,color:#fff style D fill:#667eea,stroke:#764ba2,stroke-width:2px,color:#fff style E fill:#667eea,stroke:#764ba2,stroke-width:2px,color:#fff

📋 Learning Objectives

📖 Prerequisites

Basic machine learning concepts are the one genuine requirement: what a neural network is at the level of "a function with parameters fitted to data", what training and generalization mean, and the rough idea of supervised learning. If you have worked through any introductory ML material, you have enough.

Python appears in every chapter as one short self-contained hands-on block that uses NumPy alone. The code is there to make each argument quantitative, not to teach implementation; you can read the chapters without running it, though running it is more convincing.

No background in game AI is needed. Minimax, search trees, branching factors, and Monte Carlo tree search are all built up from the beginning. No background in biology is needed either. Amino acids, protein folding, multiple sequence alignments, and the CASP evaluation are introduced as they become necessary, at the level required to follow the argument rather than to do the biochemistry.

Chapter 1

Games as the Proving Ground

Understand why board games became the standard testbed for artificial intelligence. See what made chess tractable and what made Go different — a branching factor that defeats exhaustive search and a position-evaluation problem nobody knew how to write by hand — and why Go was widely regarded as the benchmark that would hold out longest.

Game AI History Search Trees Branching Factor Evaluation Functions Why Go Was Hard

⏱️ 20-25 minutes

Read Chapter 1 →

Chapter 2

AlphaGo: Search Meets Learning

See the two halves fit together. Learn how a learned policy narrows which moves are worth considering and a learned value function estimates who is winning without playing to the end, and how Monte Carlo tree search uses both — the network telling the search where to look, the search telling the network what mattered.

Policy Networks Value Networks Monte Carlo Tree Search Supervised Bootstrapping The Lee Sedol Matches

⏱️ 25-30 minutes

Read Chapter 2 →

Chapter 3

Zero and Beyond: Learning Without Humans

Watch the human data get deleted — and the system get stronger. Follow the move to pure self-play with no human games, the collapse of two networks into one, the removal of hand-built features and rollouts, and finally the generalization beyond Go, and understand exactly what property of games made all of that possible.

Self-Play Tabula Rasa Learning Architectural Simplification Generalization Across Games The Free Verifier

⏱️ 25-30 minutes

Read Chapter 3 →

Chapter 4

AlphaFold: The Protein Folding Breakthrough

Move to a problem where self-play cannot work. Learn what protein structure prediction asks, why decades of effort had stalled, how evolutionary information in multiple sequence alignments supplies the ground truth that game rules supplied for free, how geometry and symmetry enter the architecture, and what the CASP result did and did not establish.

Protein Folding Multiple Sequence Alignments Geometric Reasoning CASP Confidence Estimates

⏱️ 30-35 minutes

Read Chapter 4 →

Chapter 5

The Legacy: From Games to Science

Resolve the apparent contradiction between the two arcs and draw the honest boundary around what was achieved. Quantify the feedback-signal argument in a short NumPy simulation, survey the AlphaFold Database and its effect on practice, get the 2024 Nobel Prize split right, follow the template into materials science with the caution it deserves, and work through the open problems that "solved" never covered.

Feedback Signals AlphaFold Database 2024 Nobel Prize Materials Screening Open Problems

💻 NumPy hands-on ⏱️ 25-30 minutes

Read Chapter 5 →

📚 Recommended Learning Paths

Pattern 1: Beginner - Full Tour (5 days)

Pattern 2: Intermediate - Fast Track (3 days)

Pattern 3: Researcher - Straight to the Argument (1 day)

🎯 Overall Learning Outcomes

Upon completing this series, you will achieve:

Knowledge Level

Practical Skills

Application Ability

🛠️ Technologies and Tools Used

Main Libraries

Development Environment

Recommended Tools

🚀 Next Steps

Deep Dive Learning

For more advanced study in this field:

Related Series

Expand your knowledge with related topics:

Practical Projects

Apply your skills to hands-on projects:

⚠️ Disclaimer