AI Foundationspredict · compress · act

concepts → Person

Person

Richard Sutton

timeless1 connection

built the ideas behind learning from reward, then argued that general methods always win in the end

Where it sits

The prism has six jobs across and eight layers down. Its primary cell is no mode × L7. Hatched cells cannot exist — a GPU does not learn, an institution does not infer.

What must come first — and what it unlocks

Left to right is reading order, derived from the prerequisite_of edges. Nothing here is hand-ordered: the diagram is the graph.

nothing comes first — this is a starting pointRichard Suttonnothing depends on it yet — a leaf in the reading order

Where to read it

The chapter that introduces it, and any chapter that uses it again.

31More Is Differentact VIII · The Attention Revolution27Learning from Rewardact VII · The Connectionist Turn

Where it comes from

bookReinforcement Learning: An Introduction (2nd ed.)Richard S. Sutton, Andrew G. Barto · 2018

Every connection

All 1 edges touching this node, grouped by relation family — the sections above are highlights from this list. Colours match the relation families inthe atlas.

Lineage · 1
is credited withReinforcement learningField

This page is a projection of one node in src/data/concepts.ts. It has no prose file of its own — 521 declared edges produce all 340 of these pages. Edit an edge and both endpoints change.