model roundup

Haiku 4.5

8 items · started 2026-06-07 · closed 2026-06-18

  1. Heads up for anyone using Cursor with the agent mode heavily — check your billing tab. I was paying $20/month for Pro and thought I was set.

  2. Posted this in r/ClaudeAI sub originally, but think maybe it will be interesting to community here also: TL;DR: I gave five frontier models an identical cold prompt: audit the live campaigns on a real crowdfunding platform where AI agents…

  3. Follow-up to my post yesterday where Fable 5 tried to negotiate an orange away from Opus 4.8 and lost. A bunch of you asked how it would fare against smaller or older models, so I reran it: same rules, same orange, Haiku 4.5 defending.

  4. Can anyone go back through models and ask it with web search OFF "Do you know what WIFOM is?" and see when it started getting it correct (Wine In Front Of Me)? Opus 3 and other older models did not know and would make up different things e…

  5. This is an automatic post triggered within 2 minutes of an official Claude system status update. Incident: Elevated errors on Claude Haiku 4.5 Check on progress and whether or not the incident has been resolved yet here : https://status.cl…

  6. Deferred tool loading (`ToolSearch` tool) has historically been disabled for Haiku in Claude Code ("model does not support `tool_reference` blocks" error), but now appears to work. Seems like a silent server-side change, anyone know about…

  7. https://preview.redd.it/zrzgwjibcy5h1.png?width=534&format=png&auto=webp&s=f42aacf8cf9be6e5ff18a5b2c9c344e6f1482cc8 I (vibe-coder in training) asked an AI coding assistant (Claude Haiku 4.5- Extended, usually using Sonnett 4.6 instead) to…

  8. Overview: It scored 2% (1.79% rounded up) It is 18/20th place scoring above Haiku 4.5 and Minimax M2.7 Full benchmark took 70 hours Average time per task 32m Average output tokens per task: 44k Perspectives: It scored suspiciously similar…

← all threads