AI Foundationspredict · compress · act

concepts → Mechanism

Mechanism

Model-based RL

timeless2 connectionsdraft

learn a model of the world, then plan inside it instead of acting blindly

Where it sits

The prism has six jobs across and eight layers down. Its primary cell isact × L4. Hatched cells cannot exist — a GPU does not learn, an institution does not infer.

What must come first — and what it unlocks

Left to right is reading order, derived from the prerequisite_of edges. Nothing here is hand-ordered: the diagram is the graph.

World modelModel-based RLReinforcement learning

Before it: World model

Unlocks: Reinforcement learning

Where to read it

The chapter that introduces it, and any chapter that uses it again.

42Models of the Worldact 9 · The Frontier

Where it comes from

paperDream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy Lillicrap, Jimmy Ba, Mohammad Norouzi · 2019
bookReinforcement Learning: An Introduction (2nd ed.)Richard S. Sutton, Andrew G. Barto · 2018

Every connection

All 2 edges touching this node, grouped by relation family — the sections above are highlights from this list. Colours match the relation families inthe atlas.

Order · 2
requiresWorld modelArchitecture
must come beforeReinforcement learningField

This page is a projection of one node in src/data/concepts.ts. It has no prose file of its own — 230 declared edges produce all 236 of these pages. Edit an edge and both endpoints change.