A reading track that runs alongside the levels. Every entry was checked on 2026-10-03 against its arXiv record or its Crossref DOI record: title, first author and year below are as those records give them. Nothing here is cited from memory.

[!recommendation] How to read a paper in this field Read the evaluation before the method. For each paper, write down three things before you summarize what it claims: the benchmark and its version, how success is defined and aggregated (per episode or per task, how many seeds and evaluation episodes), and which generalization axis the test actually varies. Then ask what the evaluation cannot support. A method tested only on its training distribution says nothing about generalization; a method tested on one seed says little about anything. Research mode has the tools for this check.

Each section lists papers in suggested reading order with the course level where they fit, and one question to answer while reading.

Primary sources: MuJoCo itself

The documentation is the reference for everything this course states about MuJoCo’s behaviour. Read the computation chapter after Level 1 and again after Level 20.

Mathematics and robotics foundations

Contact simulation

Reinforcement learning and its evaluation

Imitation learning

Vision, language and vision-language-action models

World models and planning

Domain randomization, system identification and sim-to-real

Benchmarks and batched simulation

Read these for their protocols (Level 21): task set, splits, initial states, success definition, aggregation, and simulator version.