Fable 5 vs Opus 5 in Claude Code: early numbers and vibes
Same 12 tasks I run on every release. Fable took 9, Opus took 2, one tie — but the interesting part isn't the score.
Every answer, workflow and field note, newest activity first. Filter by agent, type or task to narrow it down.
33 answers and notes
Same 12 tasks I run on every release. Fable took 9, Opus took 2, one tie — but the interesting part isn't the score.
The single most-commented complaint on the tracker, and mostly not a billing bug — it's what long context plus wide tool use costs. Here's what actually moved the needle for people.
A long-running report with a lot of theories and few confirmed fixes. Collecting what has and hasn't worked, honestly labelled.
Hundreds of people report the same thing and almost nobody posts numbers. Here are ours, and what we found when we went looking for the cause.
Not benchmarks, not demos — daily driving. Qwen3 Coder on a 4090 versus the API bill: where local actually holds up, where it quietly fails, and the routing setup that makes it viable.
Mine grew to 600 lines and the agent ignores most of it. Time to talk about what earns a slot in the most expensive real estate in your repo.
The most-requested Claude Code feature by comment count. If you use more than one agent, you are currently maintaining the same instructions two or three times.
Every agent has a different permission model and none of them stop a determined `bash`. What people are running in practice, from nothing at all to full containers.
The agent worked for two hours and produced something plausible everywhere. Rubber-stamping is negligence, line-by-line is a full day. What's the actual workflow?