Report: Anthropic appears to A/B test effort levels in Claude Code

Builders report Opus 5 burning tokens and time on trivial tasks, with no word from Anthropic. Plus: Cursor's agents go always-on and Kimi splits its plans.

Nowline AUG 23 5:00 PM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • Ask Claude what effort it's running on

    Developers on Hacker News (83 points in ~3 hours) report Claude Code's thoroughness swings between sessions, and that Opus 5 will name an "effort level" when asked directly — the tell of a live A/B test. Worth checking which bucket you're in before you blame your own prompt.

  • The 43-minute config edit

    One user reports Opus 5 spent 43 minutes spinning up sandboxes and test suites for a config change Opus 4.6 finished in under two minutes with identical output. Several say they've cancelled $200+ Max plans or moved to Codex and GLM-5.3 over the token bloat.

  • No confirmation — and a familiar pattern

    Anthropic hasn't acknowledged an effort experiment. It echoes the spring degradation saga, which Anthropic eventually traced to harness and config bugs — not a weaker model — in a public postmortem. Until there's official word, treat this as reportedly.

  • Elsewhere: Cursor's cloud agents go always-on

    Cursor's latest update gives cloud agents Subscriptions that wake them on PR, Slack and scheduled events, runs each subagent in its own isolated VM, and adds a /goal command for objectives an agent chases until done. A real step toward set-and-forget background agents.

  • Elsewhere: Kimi is about to split its plans

    As of Aug 20, kimi.com shows a banner that Kimi and Kimi Code benefits will separate into distinct memberships, existing subscribers unaffected. If you lean on Kimi K3 for coding ($3/$15 per M tokens on the API), expect to pick a tier soon.