Deep Q-networks reach human-level play on Atari from pixels.
Policy/value networks plus tree search defeat a Go world champion.
Plans with a learned model, no rules given.
Frames level generation as an RL control problem.
Team of RL agents reaches pro level at Dota 2.
Constraint-based tile generation from a single example image.
An LLM-driven agent that writes skills to explore Minecraft.
Language + planning to negotiate and play Diplomacy.
Generates Mario levels from text prompts with a language model.
LLM agents with memory simulate believable social behaviour.
A generalist agent following language instructions across 3D games.