75% cache read cut, but 1.7x output tokens per task
| Proof | Cache reads dropped to $0.25/MTok |
| Best for | Long-horizon coding & knowledge work |
| Available via | Anthropic API |
| Try it | Run a multi-step coding task with long context |
Key takeaways
- Why it matters: Better for autonomous, long-running tasks with improved failure reporting.
- Before you switch: Net per-task cost may rise 20% due to higher output token usage.
Source: Latent Space