🏆 Threes! AI: deck-aware expectimax, N-tuple TD learning, and AlphaZero
-
Updated
Jul 28, 2026 - Python
🏆 Threes! AI: deck-aware expectimax, N-tuple TD learning, and AlphaZero
N-tuple TD-afterstate reinforcement learning agent for the game 2048 (~98% 2048, ~94% 4096, ~71% 8192 after 700k self-play episodes, CPU-only)
n-tuple network agent for capped 2048 — the game ends the instant a 2048 tile is made. Scores 50,484, and documents why twenty interventions failed to go higher.
Add a description, image, and links to the n-tuple-network topic page so that developers can more easily learn about it.
To associate your repository with the n-tuple-network topic, visit your repo's landing page and select "manage topics."