#Codex
Meta launched Muse Code in beta, a terminal agent running Muse Spark 1.2. The chart says 82.9% on Terminal-Bench and a win over Codex. I went and read the chart: Meta beat GPT-5.6 Terra, not the GPT-5.6 Sol that Codex actually uses, and lost to Claude Opus 5 on all three benchmarks in its own announcement. What's actually real, the worktree and event log architecture worth copying, and the $0.30 per million price you pay for with your code.
You send a one-line question and /usage reports a whole day's worth of consumption. Saving tokens in a coding assistant has nothing to do with prompt size: it's about prefix caching. How it works in Claude Code, Codex, and Cursor, the seven actions that invalidate it without you noticing, how to measure it with cache_read vs cache_creation, and eight levers to stretch the session.