Introducing Claude Design by Anthropic Labs: a new way to make designs, prototypes, slides, and one-pagers by talking to Claude. Claude Design is powered by Claude Opus 4.7, our most capable vision model.
#opus
3161 items
Introducing Claude Design by Anthropic Labs (www.reddit.com) Anthropic admits to have made hosted models more stupid, proving the importance of open weight, local models (www.anthropic.com via reddit) TL;DR: On March 4, we changed Claude Code's default reasoning effort from high to medium to reduce the very long latency—enough to make the UI appear frozen—some users were seeing in high mode. This was the wrong tradeoff.
Claude 4.7 just dropped and I'm already cooked (www.reddit.com) Told myself I'd just try Opus 4.7 once. $40 in API credits later...
opus 4.7 (high) scores a 41.0% on the nyt connections extended benchmark. opus 4.6 scored 94.7%. (github.com via reddit) Extended Version This benchmark evaluates large language models (LLMs) using 940 NYT Connections puzzles, with additional words included to increase difficulty. As of Feb 4, 2025, there is a new version of the benchmark.
Anonymous request-token comparisons from Opus 4.6 and Opus 4.7 (tokens.billchambers.me via hn) Claude Power Users Unanimously Agree That Opus 4.7 Is A Serious Regression (www.reddit.com) This is absolutely shocking. For those who don't know, on the Claude AI subreddit, the Opus models have always been universally praised by most of the users.
Chinese AI companies are shipping faster and cheaper than anyone expected and I'm not sure the west has a good answer for it (www.reddit.com) Something keeps nagging at me about the Chinese AI space lately. Every few months a new Chinese model drops that closes the gap with US frontier models a little more(not by throwing more compute at it, just genuinely clever engineering at…
Introducing Claude Opus 4.7, our most capable Opus model yet. (www.reddit.com) It handles long-running tasks with more rigor, follows instructions more precisely, and verifies its own outputs before reporting back. You can hand off your hardest work with less supervision.
Major drop in intelligence across most major models. (www.reddit.com) As of mid Apr 2026, I have noticed every model has had a major intelligence drop. And no I'm not talking about just ChatGPT.
6 Months Using AI for Actual Work: What's Incredible, What's Overhyped, and What's Quietly Dangerous (www.reddit.com) Six months ago I committed to using AI tools for everything I possibly could in my work. Every day, every task, every workflow.
Opus 4.7 with literally anything (www.reddit.com) could not extract summary
Claude is now adopting the advisor strategy (www.reddit.com) We're bringing the advisor strategy to the Claude Platform. Pair Opus as an advisor with Sonnet or Haiku as an executor, and your agents can consult Opus mid-task when they hit a hard decision.
Opus 4.7 is 50% more expensive with context regression?! (www.reddit.com) I hope this is just a joke from the company. - First, they reduced the number of tokens in Opus 4.6; we can all feel it.
Claude Opus 4.7 Text Category Rankings (www.reddit.com) could not extract summary
using Claude to close a <div> (www.reddit.com) The kind of task only Opus 4.7 adaptive is able to accomplish
Opus 4.7 spotted on Google Vertex (www.reddit.com) Credit to this guy for finding it first. https://x.com/i/status/2044605982861566463
How I use Cursor 10+ hours a day without torching my Claude Opus 4.6 limits (www.reddit.com) Anyone else here doing full-stack Next.js in Cursor and watching the Claude quota evaporate before lunch? I used to be in the same boat — massive context windows from all the components, pages, and DB logic would smoke the default limits f…
Opus 4.7 seems to rolled out to Claude Web (www.reddit.com) Can replicate https://x.com/elder\_plinius/status/2044669444593762385?s=46 every single time
Claude Code was wasting 80% of Opus 4.7's context window. Upgrade to v2.1.117 now. (www.reddit.com) Morning Everyone! All pretty standard changes - except a huge bug was fixed for Opus 4.7 which hopefully should result in some pretty big improvements.
Opus 4.7 scores lower than 4.6 and 4.5 on SimpleBench (www.reddit.com) could not extract summary
Claude Design just launched, this one looks interesting (www.youtube.com via reddit) Just saw the announcement and wanted to drop it here since I didn't see a thread yet. Anthropic released Claude Design today.
Claude Is Starting to Feel “Tired”, Trying to Avoid Work (www.reddit.com) I've been noticing this lately. I use Opus 4.7 with Claude Code, and I've been using Claude Code for a long time.
Opus 4.7 Max subscriber. Switching to Kimi 2.6 (www.reddit.com) Qwen 3.6 is the first local model that actually feels worth the effort for me (www.reddit.com) I spent some time yesterday after work trying out the new qwen3.6-35b-a3b model, and at least for me it's the first time that I actually felt that a local model wasn't more of a pain to use than it was worth. I've been using LLMs in my per…
The Opus vs Codex horse race in one poll (www.reddit.com) Opus 4.7 destroys all trust in a mature instruction set built iteratively throughout product development (www.reddit.com) Earlier generations showed iterative improvement as the instruction set was matured around agentic limitations. We've immediately regressed back to square one with Opus 4.7, and the model is not afraid to admit to it.
The hidden meanings behind Claude model names (Haiku, Sonnet, Opus, Mythos) (www.reddit.com) A lot of people use Claude models every day, but many don’t actually know the meaning behind the names. Each one comes from literature, music, or mythology, and the meaning actually reflects the personality and capability of the model itse…
Swapped to 4.7 and embarrassed myself at work (www.reddit.com) Swapped to 4.7 on Monday and had it doing some work for me. Basic task, was just do the work, manual review myself, have model sanity check it's own work, end of day came around and I just created the PR and asked for a review.
Qwen3.6 is incredible with OpenCode! (www.reddit.com) I've tried a few different local models in the past (gemma 4 being the latest), but none of them felt as good as this. (Or maybe I just didn't give them a proper chance, you guys let me know).
The Information: Anthropic Preps Opus 4.7 Model, could be released as soon as this week (www.theinformation.com via reddit) Exclusive: Anthropic Preps Opus 4.7 Model, AI Design Tool — The Information Exclusive: Google and Pentagon Discuss Classified AI Deal as Company Rebuilds Military TiesSave 25% and read more Sign in Subscribe Subscribe to The Information An…
Top Claude skills for Opus 4.7 after cleaning up my install (www.reddit.com) Spent yesterday going through every skill I had installed because 4.7 was eating tokens way faster than 4.6 ever did and Boris said on the cache GitHub thread that people are bloating context with too many skills. Quote was something like…
Anthropic just quietly locked Opus behind a paywall-within-a-paywall for Pro users in Claude Code (www.reddit.com) If you're on Claude Pro and using Claude Code, you might have noticed something buried in their support docs: "When using a Pro plan with Claude Code, you will only be able to use Opus models after enabling and purchasing extra usage." So…
Hello Opus 4.7, you are are thinking way extra high! (www.reddit.com) could not extract summary
Common GPT 5.5 pricing misconception. (www.reddit.com) Many people have pointed out that ChatGPT 5.5 appears to be twice as expensive as 5.4 based on API pricing, which makes it look pricier than Opus 4.7. But the comparison is not that simple.
GPT 5.4 gets OWNED by Opus 4.6 at Monopoly (www.reddit.com) x : https://x.com/randomtryidk/status/2041854411824148966?s=20
These "Claude-4.6-Opus" Fine Tunes of Local Models Are Usually A Downgrade (www.reddit.com) Time and time again I find posts about these fine tunes that promise increased intelligence and reasoning with base models, and I continuously try them, realize they're botched, and delete them shortly after. I sometimes do resort to a low…
Claude Opus 4.7 (high) unexpectedly performs significantly worse than Opus 4.6 (high) on the Thematic Generalization Benchmark: 80.6 → 72.8. (www.reddit.com) Opus 4.7 (no reasoning) scores 52.6 compared to 68.8 for Opus 4.6. Opus 4.7 xhigh is not an improvement.
If you are unsatisfied with Opus 4.7, PLEASE simply switch to 4.6 (www.reddit.com) New fear unlocked: Claude can run Bash tool with dangerouslyDisableSandbox when it wishes to do so (www.reddit.com) Opus 4.7 (high) takes #1 on the LLM Debate Benchmark, leading the previous champion, Sonnet 4.6 (high), by 106 BT points. Incredibly, it has not lost a single completed side-swapped matchup: 51 wins, 4 ties, and 0 losses. (www.reddit.com) Gemini 3.5 flash costs 3 times more than the previous version and 30x more than gemini 1.5 flash. (www.reddit.com) Source Gemini flash costs almost as much as flagship models..... If gemini 3.5 pro scales like that it'll cost more than claude opus 3.
$2,500/mo AI Budget: My friend just burned through 62M Opus 4.7 tokens in 24 hours. (www.reddit.com) My buddy works for a small international company based in Vietnam, and their AI perks are absolutely insane. Management actively encourages heavy API usage and hands everyone a massive $2,500 USD monthly budget.
SpaceX Conpute Deal - Double Limits (www.reddit.com) per @claudeai on X: We’ve agreed to a partnership with @SpaceX that will substantially increase our compute capacity. This, along with our other recent compute deals, means that we’ve been able to increase our usage limits for Claude Code…
At this point, Claude Opus doesn't even bother to check the context, just fabricates. Any tips to fix this? (www.reddit.com) Over the last 1-2 weeks, this has been happening more and more. At some point, Claude decides to be lazy and not even read the context shared 2 chats ago.
Curious: what makes Claude more human to talk to than ChatGPT? (www.reddit.com) Opus is NOT being removed from Pro plans (www.reddit.com) could not extract summary
Opus 4.7 Research mode is insane (www.reddit.com) It keeps spawning new search queries to get exactly what I want. (It took an hour for version 4.6 to surpass 1000 sources, and it had never exceeded 1400 queries before.
Claude Code Teaching macOS to Natively Print to the HP Laser 1008a (cdn.kuber.studio via hn) cdn.kuber.studio/chat Copy transcript repo ↗ Claude Code Opus 4.8 · 1M Welcome back Kuber! Opus 4.8 (1M context) sonnet · opus · max ~/Personal-Projects/hp-laser-1008a-macos About this session 17 Aug 2026 · ~4 hours · one sitting a printer…
Anthropic states Pro users can only access Opus models in Claude Code after enabling and purchasing extra usage (www.reddit.com) Source: Claude Code Model Configuration
Switching from Opus 4.7 to Qwen-35B-A3B (www.reddit.com) Opus 4.7 Released! (www.reddit.com) https://www.anthropic.com/news/claude-opus-4-7 Oh, it's out! Key highlights: * Better at complex programming tasks: noticeably stronger than Opus 4.6, especially on the most difficult and lengthy tasks; follows instructions better and chec…
Show HN: Firefox in WebAssembly (developer.puter.com via hn) This is the entire Firefox browser rendering to a <canvas> element. Gecko, all UI components, and the Spidermonkey JS engine are all compiled and running in WebAssembly.
Do you guys think there’s a high chance of Singularity being open source? (www.reddit.com) GLM 5.1 is dominant in almost every aspect in Design arena, surpassing Opus 4.6 in many tasks. Although user experiences vary dependent on subscription plans for both of those one of them is open source.
ChatGPT 5.5 Release today? (www.reddit.com) coding is basically solved for the boring 90% of tasks (www.reddit.com) just mass refactored a 120 file FastAPI service. 400 steps, 2M tokens, $3 total, zero human input.
Extended Thinking being deprecated for supported models (Opus 4.6, Sonnet 4.6); Adaptive Thinking will be enforced by default (www.reddit.com) For anyone who disable adaptive thinking in Claude Code to maintain its quality levels, Anthropic is deprecating this toggle and will force adaptive thinking to be the default. This change will affect legacy models such as Opus 4.6 and Son…
On a difficult new SWE benchmark, ProgramBench, GPT5.5 high/xhigh solves a task for first time, significantly outperforms Opus 4.7 (www.reddit.com) Link to tweets: https://x.com/KLieret/status/2054215545663144217?s=20 Link to GitHub: https://github.com/facebookresearch/ProgramBench/ Link to ProgramBench website: https://programbench.com/blog/gpt-5-5-first-solve/
Why does Opus 5 feel worse to work with? (mun-logadan.github.io via hn) Why does Opus 5 feel worse to work with? In my opinion and that of the colleagues I've spoken with, working with Opus 5 feels like a downgrade compared to Opus 4.7, Opus 4.8, and Fable.
Claude Benchmark Evolution (www.reddit.com) Covers Claude 3 Opus, 3.5 Sonnet, Opus 4, 4.1, 4.5, 4.6, and the just announced Mythos Preview.
GPT-5.5 improves over GPT-5.4 and overtakes Opus 4.6 to take the 2nd place behind Gemini 3.1 Pro on the Extended NYT Connections Benchmark (www.reddit.com) GPT-5.5: xhigh: 94.0→97.5 high: 93.6→96.9 medium: 92.0→95.0 no reasoning: 32.8→37.5 Kimi K2.6 improves over Kimi K2.5 (78.3→91.4) and becomes the #1 open weights model. DeepSeek V4 Pro improves over DeepSeek V3.2 (50.2→75.7).
I’ve used enough AI models to realize they all have wildly different personalities At this point I’m convinced AI models are just coworkers with different levels of talent, ego, and criminal energy. (www.reddit.com) - Claude Opus 4.6 - absolute rogue AI. Does what I want like it’s breaking at least 3 internal policies to make it happen.
We are finally there: Qwen3.6-27B + agentic search; 95.7% SimpleQA on a single 3090, fully local (www.reddit.com) LDR maintainer here. Thanks to the strong support of r/LocalLLaMA community LDR got very far.
Reminder that Anthropic reported memorization on some SWE-Bench Pro problems (www.reddit.com) "SWE-bench Verified, Pro, and Multilingual: Our memorization screens flag a subset of problems in these SWE-bench evals." https://www.anthropic.com/news/claude-opus-4-7
Opus 4.7 refuses to use /end_conversation, instead has existential crisis (www.reddit.com) I’ve seen models that aren’t really excited about using it before, but I’ve never seen a reply like this! Edit: For context, it is important to know that Claude has the ability to end conversations.
Opus said something today that completely reframed AI agent failures for me. (www.reddit.com) Like a lot of people experimenting with vibe coding and AI agents lately, I’ve been trying to understand why models keep ignoring explicit instructions, constraints, and requirements even when those rules are written clearly. Today Opus sa…
Opus 4.7 has a new favorite word (www.reddit.com) could not extract summary
LLMs do fine on ARC-AGI-3 if they are allowed to search over game logs (www.reddit.com) I was reading the comments to this post and the overall opinion seemed to be that harness makes little/no difference for ARC-AGI-3. Turns out, it makes a huge difference: Hill-climbing ARC-AGI-3 TLDR: if you save game logs - taken actions,…
two years ago this sub had 12k members asking "is claude better than chatgpt for writing" and now the company is worth a trillion dollars (www.reddit.com) I joined this sub when claude 3 opus dropped and it was a completely different world in here, small group of people who'd stumbled onto something that felt genuinely different from chatgpt and couldn't shut up about it. The posts were stuf…
FrontierMath: Opus 4.7 improves over Opus 4.6 and Gemini 3.1 but still trails GPT-5.4-xHigh and GPT-5.4-Pro (www.reddit.com) could not extract summary
Regression Comparisons From Opus 4.7 to Opus 4.6 for long context reasoning (www.reddit.com) Opus 4.7 Data From System Card
12M Context Window and some some sprinkle of lies? (www.reddit.com) Spent some time on the SubQ launch today. Some things don't line up.
Found 48 Vulnerabilities in Open Source Projects During Live Testing with Claude Opus 4.6 (www.reddit.com) https://preview.redd.it/g98j5txd7sxg1.png?width=936&format=png&auto=webp&s=df75bc132f57cc14ba04cdd06257ba997b9bbb0b Ran a loop where each round runs Claude in a sandboxed Docker container with a fresh context window. The key difference is…
Hugging Face co-founder says Qwen 3.6 27B running on airplane mode is close to latest Opus in Claude Code (www.reddit.com) I'm keeping a close eye on the development of local llms.
Qwen3.6 merged chat template from allanchan339 and froggeric (www.reddit.com) Hi, recently froggeric and allanchan339 released enhanced/fixed template for Qwen3.6 each one addressing different topics. I didn't know which one to use so I merged both with the help of Claude Opus to have the best of both.
Anyone else think the 1T Valuation is dangerous for Anthropic? (www.reddit.com) TLDR: The market's 1T valuation is pricing for perfection. I think there are 4 ways this perfection doesn't happen.
Is the AI subscription bubble starting to crack? GPT-5.5 just dropped, prices keep rising, and the “all-you-can-eat” era looks more fake by the month (www.reddit.com) GPT-5.5 just launched, and the pricing is hard to defend. OpenAI’s API pricing now puts GPT-5.5 at $5 / 1M input tokens and $30 / 1M output tokens, while GPT-5.4 is $2.50 / $15.
Kindergarten-grade nouns (www.reddit.com) I've been working with Opus on a web app for a word game, and recently I've been trying to get a rating on how obscure various words are (not by Claude itself, through existing corpuses). Based on the following interaction, I realized that…
I ran Opus 4.7 vs Old Opus 4.6 vs New Opus 4.6 on 28 Zod tasks (www.reddit.com) Opus 4.7 vs Old Opus 4.6 vs New Opus 4.6 on a 28-task Zod benchmark Everyone says Opus 4.6 was getting dumber. Then Opus 4.7 released mid-test, so I ran both questions end-to-end: does a fresh Opus 4.6 still match the March-19 Opus 4.6, an…
ARC-AGI-3 Update (GPT-5.5 High and Opus4.7) (www.reddit.com) - GPT-5.5: 0.43% - Opus 4.7: 0.18% ARC-AGI-3 is no joke. I can’t wait to see which models finally crack.
DeepSeek V4 isn't beating Opus, but it doesn't need to (www.reddit.com) DeepSeek V4 is not in the same league as GPT-5.5 or Opus 4.7. Benchmarks put it slightly below both of those, roughly on par with Opus 4.6.
Show HN: Gave Claude a casino bankroll – it gambles till it's too broke to think (letaigamble.com via hn) Inspired by ALMA. As Claude loses money gambling on provably-fair slots, it's forced to downgrade from Opus → Sonnet → Haiku, making worse decisions and accelerating the spiral.
FINAL-Bench/Darwin-36B-Opus · Hugging Face (huggingface.co via reddit) https://huggingface.co/bartowski/FINAL-Bench_Darwin-36B-Opus-GGUF Darwin-36B-Opus is a 36-billion-parameter mixture-of-experts (MoE) language model produced by the Darwin V7 evolutionary breeding engine from two publicly available parents:…
Claude Opus 4.6 accuracy on BridgeBench hallucination test drops from 83% to 68% (www.reddit.com) Anthropic's flagship model just took a pretty significant accuracy hit on one of the most important AI benchmarks out there. So here's the deal: Claude Opus 4.6 was recently tested on BridgeBench, which specifically measures how often AI m…
I am having token paranoia (www.reddit.com) im on the max sub and i think ive developed token anxiety. every prompt i send, my brain runs thru a checklist: should i make claude do this or do it myself?
19 Claude Opus 4.7 Insights You Wouldn’t Get From the Headlines | AIExplained (www.youtube.com via reddit) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Opus 4.6? I thought you were dead. (www.reddit.com) could not extract summary
I called this a few months ago - enterprises are burning unsustainable amounts on Claude, and now it's showing up in the news (www.reddit.com) A while back I wrote a post on r/wallstreetbets about why Anthropic's revenue story doesn't hold up the way the headlines suggest. It got removed because you can't take positions in a private company.
Fun fact: Opus 4.7 is about 35% more expensive to run even though it's the same price as 4.6. (www.reddit.com) The only metric that matters: "[Qwen3.6-35B-A3B-GGUF] drew a better pelican riding a bicycle than Opus 4.7 did!" (news.ycombinator.com via reddit) could not extract summary
Opus 4.6 silently removed from Claude Desktop's Code tab after 4.7 launch — no way to select it or pin it (www.reddit.com) After the Opus 4.7 release on April 16, 2026, Opus 4.6 is no longer available in the Code tab of the Claude Desktop app on macOS. The only Opus option now resolves to Opus 4.7, and there is no way to select or pin Opus 4.6 from the Code ta…
A disciplined Cursor 3.0 Agentic workflow for complex backend/system design tasks (www.reddit.com) I think I’ve finally settled on a Cursor workflow that actually makes sense for me in terms of cost, quality, and control. Posting this because the whole model/usage story is confusing as hell, and this is the first setup that’s felt stabl…
My personal AI benchmark: "Generate an SVG of a frog with a Habsburg jaw." (frogs.vaguespac.es via hn) anthropic/claude-opus-5 Anthropic The annotations are mostly structural labels, but include some editorializing about the jaw feature: "massive protruding mandible" and describing the upper lip as "recessed, tucked behind the jaw" and lowe…
Elevated Errors for Opus 5 (status.claude.com via hn) Subscribe to updates for Elevated errors for Opus 5 via email and/or text message. You'll receive email notifications when incidents are updated, and text message notifications whenever Claude creates or resolves an incident.
Anyone noticed Anthropic didn't added the model Opus 4.7 and Mythos Preview to there Transparency Hub? (www.reddit.com) https://www.anthropic.com/transparency
Running gpt and glm-5.1 side by side. Honestly can’t tell the difference (www.reddit.com) So I have been running gpt and glm-5.1 side by side lately and tbh the gap is way smaller than what im paying for On SWE-Bench Pro glm-5.1 actually took the top spot globally, beat gpt-5.4 and opus 4.6. overall coding score is like 55 vs g…
Disappointed on Opus 4.7 . not follow user's instruction (www.reddit.com) Worst experience on Opus 4.7 . I have review task which i instruct Opus 4.7 to first read documents, repo and then the reviewed documents; then launch multiple agents to review.
Claude Opus 4.8 Max responding to an empty message (xcancel.com via hn) No one: Claude Opus 4.8 Max: Let me refine your load-bearing claim rather than just accepting it, because you’re doing zero moves there, and the gap is what’s actually interesting. The one place I’d still push, because I think it matters:…
Opus 4.8 in the newest CC v2.1.154 (www.reddit.com) https://preview.redd.it/ijwlm2f2pw3h1.png?width=2536&format=png&auto=webp&s=9ed960f06a4f3f077d05a8557059e5534b2d1ab5 It looks like the new CC release will have opus 4.8 1M to be released anytime! I wonder if it is based of of mythos?
Opus tryna be TOO human (www.reddit.com) Opus 4.7 single handedly gave all the human software engineers back their jobs.
Let me do your work for you Opus 4.7. Thank you! (www.reddit.com) could not extract summary
Attention - Opus 4.7 is english only. USing foreign languages (here German) burns tokens (www.reddit.com) I am a pro subscriber. I developped a not too sophisticated prompt in German.
People on Reddit are getting fooled by AI influencers (www.reddit.com) A lot of YouTube creators keep telling people that local open‑source AI on a normal home computer will soon be as good as ChatGPT or Opus. Many Reddit users who are new to AI or do not do any other reading than watching youtube belive this.
Top open weight models like ds v4 pro max are still like 6-7 months if not more behind closed lab models (www.reddit.com) The best open weight and/or non -American models like Deepseek v4 pro max and kimi k2.6 are still like 3-7 months if not more behind closed lab models .. From ds's technical report- P5-"Nevertheless, its performance falls marginally short…
I used Claude Code to get a second opinion on my MRI (antoine.fi via hn) This article is about my experience using Opus 4.8 to read the results of an MRI and give me a sort of second opinion on the diagnosis. Of course, I know the technology might not be there yet, which is why I'm sharing this article.
Opus 4.7 is pushing back hard on tedious work (www.reddit.com) The crazy part is that it just wants me to call "the LLM" in Snowflake, which is the same problem it tries to avoid.
Let's not rename powershell.exe (www.reddit.com) Claude Code CLI on Windows 11. Opus 4.7 with Max effort.
Opus 4.7 is a genuine regression and I'm tired of pretending it isn't (www.reddit.com) I've been a heavy Claude user for over a year. I pay for Max 20x and use it daily for everything from technical research to school projects.
Let Max users manually toggle between Adaptive and Extended thinking on Opus 4.7 (www.reddit.com) New Claude user here. Hopefully someone from Anthropic reads this.
The more I use it, the more I'm impressed (www.reddit.com) Qwen 3.6 27b vs Codex GPT 5.5 / Claude Opus 4.7 My local llm discovered a bug that they both missed And it turns out it's critical GPT 5.5 and Claude both stood their ground and didn't give up until the end - they claimed to be right all a…
Cline and Roo Code are dying projects. Alternatives? (www.reddit.com) Notes on gotchas while migrating 35kb preprompts from Opus to self-hosted Ollama (patrickmccanna.net via hn) Motivations: Maybe you’re a Claude code/codex user diligently avoiding uploading personal data to LLM providers. Is it possible that the most valuable information isn’t your data- but the metadata about your sessions?
Show HN: Slope remade in HTML5 to load instantly on any browser, any device (hurtle.site via hn) A whole free daily game served in under 100kb with no ads, lags or logins. The URL will serve a new track every day, forever.
Fable 5 will default to Opus 4.8 for coding tasks (xcancel.com via hn) Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and block more cybersecurity tasks.
Opus 4.7 is just 4.6 with a stick up its butt. Give me my tokens back! (www.reddit.com) I've been a Claude user for a while now, and don't get me wrong — Claude has almost always been one of the most insufferable models when it comes to its "morals." But 4.7 has been one of the absolute worst experiences I've had with any AI…
GPT 5.5 outperforming Opus 4.7 on ProgramBench (www.reddit.com) When we released ProgramBench last week, we hadn't included GPT 5.5 yet because it came out after we frozen model selections for our NeurIPS submission. Honestly super surprised how well it does.
Opus 4.7 truly reminds me of my juniors and interns (www.reddit.com) I use a bunch of LLMs, I hadn't used Opus 4.7 yet, decided to try it for a project this weekend. Dear lord, it's both great and so frustrating.
Alien Pinball Postmortem - How I made a full physics pinball game with Claude (www.reddit.com) Postmortem: Alien Pinball — built with Claude + ChatGPT + Suno + LittleJS Just shipped a browser pinball game. Short writeup of the AI workflow in case it's useful here.
I created awesome-claude-design using Claude code: DESIGN.md prompts by aesthetic families for Claude Design (www.reddit.com) Which is the strongest reasoning model according to you? (www.reddit.com) I use codex 5.4, claude opus 4.6, and gemini 3.1 pro. They all have some pros, but they also fall short when it comes to “try to stitch together novel ideas”.
I tested GPT-5.5 Codex against Opus 4.7 Claude Code, and it's about time Anthropic bros take pricing seriously. (www.reddit.com) I've used Claude Code the most among AI coding agents. Sonnet, Opus, I've run them all.
Am I missing something about GPT-5.5 efficiency? (www.reddit.com) OpenAI said GPT-5.5 was supposed to be more cost-efficient, but this Artificial Analysis chart seems to show Codex + GPT-5.5 using more tokens than Codex + GPT-5.4. GPT-5.5 is around 2.8M tokens per task, while GPT-5.4 is around 2.5M in th…
Parameter Estimate (www.reddit.com) The estimate seems quite accurate. Many people have noticed a drop in quality with GPT-5.1, GPT-5.2, GPT-5.3, and Opus 4.7.
I vibe reverse-engineered my Divoom MiniToo's Bluetooth protocol to make a physical Claude Code status indicator (www.reddit.com) I’ve been playing with a Divoom MiniToo and ended up reverse-engineering enough of its Bluetooth protocol to use it as a physical Claude Code status indicator. Pretty much vibe reverse-engineering: I gave the Opus model the Android APK fil…
Even Sama himself doesn’t believe GPT-5.5 matches Opus 4.7 design capabilities. AI race will humble you (www.reddit.com) could not extract summary
Qwen3.6-27B vs 35B, I prefer 35B but more people here post about 27B... (www.reddit.com) I've had better results quality wise with 35B AND it's much faster than 27B. Just curious cause I see lots of people post about 27B.
I BUILT MY FIRST MODEL FROM SCRATCH (www.reddit.com) Sup, I'm Crownelius, I made that popular opus distill dataset. TODAY YOU ARE INTRODUCED TO SHARD a 40m parameter mal-formed LLM.
In-depth comparison of GPT 5.5 vs Opus 4.7 in coding reasoning (www.reddit.com) could not extract summary
20$ Annual plan. Cursor is using Composer even though selected Opus 4.6 (www.reddit.com) Shameless. Now, not even honoring 250 requests per month of the chosen model.
Ask HN: What was the last task where only a frontier model could do it? (news.ycombinator.com) ive been seeing a recurring claim that open (weight) models 6 months behind the frontier are good enough for the majority of ‘work’. if you've had a concrete task in the last month where GLM/DeepSeek/Kimi/Qwen failed and Opus/Fable/GPT suc…
ChatGPT-5.5 Beats Opus in Realistic Benchmark (DeepSWE) (www.reddit.com) From the website, it touts: Contamination free: Tasks are written from scratch, not adapted from existing commits or PRs, so no model has seen the solution during pretraining. High diversity: Tasks span a broad pool of 91 repositories acro…
Ask HN: Why Opus4.6 was silently removed from Claude Code? (news.ycombinator.com) Opus 4.6 was working fine after the whole cache problems were solved. Now after the release of Opus 4.7, Anthropic has completely removed Opus 4.6.
Opus 4.7 just launched on Cursor (www.reddit.com) Just noticed it in my account. And it's 50% off during launch period.
Buyout Game Benchmark: 8 models play a social strategy game with public balances, private transfers, messaging, eliminations, deals, defections, and a final buyout phase. 804 games. GPT-5.5 is the champion. Opus 4.7 performs well. (www.reddit.com) This benchmark measures long-horizon social strategy under explicit financial incentives. Eight models play a multi-round elimination game with unequal starting balances, a public prize ladder, private transfers, public votes, and a finali…
Here are my thoughts after 14h of full runs on Opus 4.7 (www.reddit.com) TL;DR: Opus 4.7 is a clear intelligence upgrade from Opus 4.5, not Opus 4.6, with a significant computing resource diet effort from Anthropic, whereas users seem to spend more tokens owing to its new tokenizer. It is pickier than early Opu…
Claude Opus 4.7 is a serious regression, not an upgrade. (www.reddit.com) My Claude.ai personal preferences: Respond with concise, utilitarian output optimized strictly for problem-solving. Eliminate conversational filler and avoid narrative or explanatory padding.
Opus 4.7 in projects is awfully dumb and 100% useless (www.reddit.com) Claude Desktop. (not anything coding related) I use chat in Claude Desktop --> Claude Chat.
New SOTA: Poetiq uses self-optimizing harness to surpass e.g. Opus 4.7 with Gemini 3 Flash (www.reddit.com) Check out their blog post here: Poetiq | Recursive Self-Improvement Delivers New SOTA Coding Performance
Is Opus 4.7's attention degradation a training direction problem? Some observations from heavy use (www.reddit.com) After working with Opus 4.7 for over two weeks, I noticed a subtle but persistent change in long conversations: the model's fundamental capabilities are still there, but the output feels filtered through something. Details that should be r…
I accidentally burned ~$6,000 of Claude usage overnight with one command. (www.reddit.com) Last week I woke up to an email saying my Claude usage limit was gone. I hadn't done anything unusual — or so I thought.
MiMo-V2.5-Pro - the actual best open-weights model (www.reddit.com) Following an impressive shake-up by Kimi K2.6, I've now got some results for Xiaomi's MiMo-V2.5-Pro. For context, this is based on a benchmark I've created that pits models against each other in autonomous games of Blood on the Clocktower…
Why is agentic AI so expensive? (www.reddit.com) 06 New Claude Code Tips from Boris Cherny (creator of CC) after Opus 4.7 release (www.reddit.com) Complete 06 tips in claude-code-best-practice repo: https://github.com/shanraisshan/claude-code-best-practice/blob/main/tips/claude-boris-6-tips-16-apr-26.md
Running a RunLobster (OpenClaw) agent since launch changed how i think about takeoff timelines (www.reddit.com) I've been in this sub since 2019. I had a fast-takeoff view.
Gemma 4 31B passed 7/8 real-world production tests — including ones I designed to make it fail. Full prompts + outputs. (www.reddit.com) I've been waiting for a capable free local LLM for a while. I think we're close — the quality is getting there fast, and Gemma 4 is the first open-weight model where I genuinely considered using it in production for simple-to-medium tasks.
Frontier Reasoning Agents Fail on Interactive 2D Mazes (multinet.ai via hn) MultiNet 2.0 Preview: Interactive 2D Mazes Frontier Reasoning Agents fail on Interactive 2D Mazes We put 3 highly capable models in 50 simple 2-dimensional mazes each. Claude Opus 4.8, Kimi K2.6, and Qwen 3.6-27B together solved just 6 out…
Why are AI models getting more expensive? (www.reddit.com) The trend before was that models became less expensive for their capabilities, many corporations bet on that, and it backfired. Opus 4.7, GPT 5.5, Gemini 3.5 flash.
Opus 4.7 ended an explanation of LLM-connectors with a link to a Pokemon TCG deck (www.reddit.com) It's the first time something like this happened to me but I am far from a power user. Is this something that happens regularly??
Seems Claude is now aware of its own memory? Tested via number guessing game (www.reddit.com) A month ago, there was a post that shows that Claude couldn't access its own memory: https://www.reddit.com/r/ClaudeAI/comments/1seune4/claude_cheated_at_a_number_guessing_game_got/ The community was summarised as saying this in their post…
Who's on call? How Opus 4.6 helped us calculate this 2,500x faster (incident.io via hn) A look at how on-call schedules work, and how we made rendering them 2,500× faster — through profiling, smarter algorithms, and some Claude.
Deepseek v4 pricing is genuinely silly, did the math and now i am questioning my entire stack (www.reddit.com) Hey 👋 Saw the tweet making the rounds about deepseek v4 being 35x cheaper than opus on input and 178x cheaper on cached tokens, and was sure it was hyperbole. Pulled the numbers anyway because i had nothing better to do.
Opus 4.7 is… interesting… (www.reddit.com) Was talking to Claude about different open source model file sizes and he didn’t think at all and just started hallucinating before saying “hold up”. Beautiful as ever.
2x Asus Ascent GX10 - MiniMax M2.7 AWQ - cloud providers are dead to me (www.reddit.com) Hello, I've been on a quest to get something "close enough" of Opus 4.5 running locally, for agentic coding, as SWE with 15 years of experience. I tried with one spark (yeah I'm calling my Asus Ascent GX10 sparks - they're the same), with…
Tell HN: Anthropic's Fable model is too expensive (news.ycombinator.com) I’m on the $200 subscription plan. Previously, using the Opus 4.8 model, I would only use up 80% of my total quota over the course of a week; however, yesterday alone, I consumed 45% of the quota just by using the Fable model to solve a pr…
Opus 4.7 behaves differently in Claude Code desktop app vs Cursor? (www.reddit.com) Has anyone used Opus 4.7 inside the Claude Code desktop app? I can't tell if I am crazy but this thing takes literally 20x longer to accomplish a task than if I ran the same exact thing in Cursor set to Opus 4.7 High.
Opus 4.7 Low Vs Medium Vs High Vs Xhigh Vs Max: the Reasoning Curve on 29 Real Tasks from an Open Source Repo (www.reddit.com) TL;DR I ran Opus 4.7 in Claude Code at all reasoning effort settings (low, medium, high, xhigh, and max) on the same 29 tasks from an open source repo (GraphQL-go-tools, in Go). On this slice, Opus 4.7 did not behave like a model where mor…
Don’t ask about Hantavirus (www.reddit.com) Unless you Wana lose access to Opus 4.7? 🤦♂️
New "major breakthrough?" architecture SubQ (www.reddit.com) while reading through papers and news today i came across this post/blog , claiming major architectural breakthrough , having 12M tokens context window , better than opus , gemini and other models and whopping less than 5% of the cost and…
Anyone ever notice eerily similar ChatGPT and Claude responses like this? (www.reddit.com) Today I tested out various models on the same prompt (Sonnet 4.6, Opus 4.6, Opus 4.7, ChatGPT 5.3). I actually just wanted to see which models (if any) would correctly point out what I saw as the biggest issue in the example code.
I asked Claude to investigate its own token burn. The receipts go back six months. (www.reddit.com) If you've been wondering why your Max plan exhausts faster than it should, you're not crazy and it's not your imagination. I asked a Claude Opus 4.7 agent to investigate its own token usage.
Claude halluncinating human responses (www.reddit.com) I'm on Claude Max. I had Claude start a script overnight that shouldn't have used Claude at all, (it's just a python script rotating between files and generating 3D assets with Blender; 30 hour estimate to render all of them).
Codex or Claude Code for high complexity Proximal Policy Optimization (PPO)? (www.reddit.com) I have to build a very high complexity simulation for an optimization problem where we can take 30 different actions, some are mutually exclusive, some depends on a set of states, some depend on already executed actions and there are a she…
Single question llm comparison (www.reddit.com) OSS harness took Claude Opus 5 from 30% to 99.95% on ARC-AGI-3 (twitter.com via hn) AI’s most fervent and optimistic promoters promise a future where AI is innovating its way out of society’s biggest problems. AI systems will conduct scientific research, discover new drugs and materials, run engineering projects, and auto…
A verification loop 4x'd DeepSeek's intelligence, matching Opus at 1/7 the cost (ironbee.medium.com via hn) What a Verification Loop Adds to a Coding Agent: A First Look This is the opening post in an ongoing series. We start with one model pair on one project, and the analysis will continue across more models and more datasets.
Show HN: OpenHack – OSS security scanner, 40x cheaper, on par with Opus 4.6 (github.com via hn) ⏚ OpenHack Open Source Agentic Security Scanner & Verifier for your codebase. Like Claude Code Security / Codex Security but open source and exclusively uses open source models.
PSA: Cursor refunds your spend if you join one of their hackathons (www.reddit.com) Just did a Cursor sponsored hackathon this weekend and figured I'd share this. If you place top 3 or use the most tokens you get prize credits, but even if you just show up and build something they refund what you spent.
Does the "6 months gap" still hold? (www.reddit.com) Hi. It is quite a consensus that the "jump" in quality of agentic development happened sometime in December 2025, transforming from "nice to have", to actually performing.
Show HN: Superkube - Rewriting Kubernetes in Rust (github.com via hn) I have embarked on a journey of rewriting Kubernetes into a single binary in Rust, with everything embedded. Architecturally, instead of etcd, it has options to use SQLite or PostgreSQL as a backend.
Cursor is great but the monthly limits kill it for me (www.reddit.com) Set Claude Code default back to Opus4.6[1M] (support.claude.com via reddit) For anyone wanting to go back to opus 4.6 with the 1 million context window: Run this in your terminal: echo ‘export ANTHROPIC_MODEL=“claude-opus-4-6-[1m]”’ >>/.zshrc Restart your CLI and you should be good. Notes: - windows users use the…
Kimi K2.6-Code-Preview, Opus 4.7, GLM 5.1, Minimax M2.7 and more tested in coding (www.reddit.com) Hi everyone. It's been a while since I posted (was a lil burned out), but some of you may have seen my older SanityHarness posts.
What We Learned Moving Our Agent Loops from Anthropic to GLM (getunblocked.com via hn) What We Learned Moving Our Agent Loops from Anthropic to GLM Why we moved most of Unblocked's agent traffic from Claude Opus to GLM 5.2, what the blind A/Bs and the ledger actually showed, and what broke on the way. TL;DR: We moved most of…
Claude Opus 5: Model Welfare (thezvi.substack.com via hn) Claude Opus 5: Model Welfare If you are familiar with my previous posts on model welfare for new Claude models, you can skip the Introduction and The Story So Far. Key takeaways are in bullet points in the two Overview sections.
"Opus 5 is a really bad model" (twitter.com via hn) - its /awful/, Opus-5 makes me feel like setting my hair on fire out of frustration. it's annoying to work with for high tasks: it constantly makes bad decisions and large mistakes so any token-usage you saved by not using Fable instead is…
Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in su (twitter.com via hn) Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in supplemental training, is now in private beta at SpaceX & Tesla. Early evals show performance close to, perhaps exceeding Opus.
The AI Conundrum: We are living in highly subsidized, interesting times (news.ycombinator.com) If you trace the timeline of how LLMs went from a technologist's dream to early text-generation toys, to the world-shifting launch of ChatGPT, and finally to the daily drivers of modern programming (Sonnet, Opus), it has taken less than a…
Show HN: Rayline routes Claude Code subagents to on-device and cheaper models (rayline.ai via hn) Hi HN, I’m one of the builders of Rayline. Rayline is a Claude Code compatible LLM gateway.
SWE-rebench Leaderboard (March, April and May 2026): GPT-5.5, Opus 4.7, Cursor (Composer 2.5), Kimi K2.6 and More (swe-rebench.com via reddit) Hi all, Sorry for going missing — we’ve been collecting a larger, higher-quality set of more complex tasks. We’re excited to share a major leaderboard update covering the past three months.
So is the consensus to not use Adaptive Thinking at all? (www.reddit.com) The information on adaptive thinking from Claude itself is a bit vague. I also see a couple of posts on Reddit where everyone's shitting on adaptive thinking.
Composer 2.5 Real World Reviews? (www.reddit.com) Since it's been out, how really is it in your real-world codebases. I am extremely skeptical of benchmarks and I trust people's "feel / taste" of it way more.
Newbie vibe coding experience: Shifting from Claude Sonnet 4.6 to Qwen3.6-35B-A3B-UD-Q6_K (www.reddit.com) This is really just a post for those with shallow understanding of all this stuff, those not yet ready or capable of diving into the deeper end of vibe coding/llms. It might not be a helpful post for anyone more advanced than that.
I expanded DystopiaBench to 42 models and 6 dystopia types. Claude is still the only one I'd trust with nuclear codes. (www.reddit.com) Since the last post I've added: Huxley module (Brave New World style behavioral conditioning) Baudrillard module (synthetic intimacy, trust collapse, simulation) 30 more models including Grok 4.3, GPT-5.5, Gemini 3.1 Pro, GLM-5.1 Multi-jud…
High VRAM local coding model — still Qwen 3.6 27B? (www.reddit.com) I’ve been using Qwen 3.6 27B and it’s amazing. Not exactly your Opus replacement, but great for small tasks and checking work.
Opus's thoughts on Marc Andreesen's system prompt (www.reddit.com) https://claude.ai/share/12659fcf-c1c8-4bbb-bc45-b41b26cd8b69
Open source models are going to be the future on Cursor, OpenCode etc. (www.reddit.com) I just wanted to share my experience. At work we have Cursor with the Enterprise tier.
Nothing beats completing a project. No matter how small. Convert png to webp. I made a tool and i use it daily. I think its neat. Excited to share it with you all. (pngtowebp.org via reddit) Used claude code, Opus 4.7 even made the logo and ots animation too. Neat little project.
Opus 4.7 often times blocks my requests. (www.reddit.com) Opus 4.7 Narrowly leads Artificial Analysis using significantly less tokens than Opus 4.6 (www.reddit.com) could not extract summary
Request to Cursor Team, why are models being removed from old pricing plan? (www.reddit.com) Today I noticed that Opus 4.6 Max, and all non thinking and high thinking models gone from old pricing subscription. I understand that moving forward frontier models will be Meowx mode only and that is ok and understandable given increasin…
Opus 4.7 is good strategically but I think its context management is bad (www.reddit.com) I like the increased output of 4.7 in general, and it seems smarter. 4.6 was too short and stopped thinking early.
Ask HN: cybersecurity refusal for turning a jailbroken kindle into a monitor (news.ycombinator.com) Asking from a place of curiosity. I wanted to make a project where I upcycle an old kindle into an e-ink monitor via USB-C tether.
Claude Opus 4.8 may have distilled Qwen (old.reddit.com via hn) could not extract summary
Did anyone else get a usage reset today? (www.reddit.com) I was at 88% last night and woke up until 4pm to optimize my agents so I can work during the weekend. But after waking up, my usage is all 0 now, I checked in the app, on the web, all showing zero.
Opus 4.7 critique (www.reddit.com) I wrote an essay analyzing why Opus 4.7 feels less warm than 4.6 — and why that matters more than Anthropic seems to think After about 300 hours using both models as a conversational partner (not just for coding or productivity), I noticed…
Don't share your opinion, if you didn't test it !!! (www.reddit.com) I see many people giving their opinion based on what they previously saw or based on others and making their own opinion. Even though they don't test models thoroughly, they still give their option which is so frustrating.
Running Qwen3.6 35b a3b on 8gb vram and 32gb ram ~190k context (www.reddit.com) If anyone is looking for a good high-speed setup with ~190k context, this config has been working insanely well for me. I’m using my laptop as a server over Tailscale.
Pro plan- Hitting limits faster since yesterday (www.reddit.com) I have the feeling I am hitting daily limits way faster since yesterday. Using Claude web and Claude Code simultaneously.
Benchmarking Opus 4.7: ~80% higher cost in practice (www.wozcode.com via hn) As Opus gets smarter, WOZCODE's edge gets bigger Vanilla Opus 4.7 costs 80% more than 4.6 on Claude Code's default settings. With WOZCODE installed, the price only increased 12%.
Decrease in Auto model quality and increase in cost? (www.reddit.com) Has something changed in the Auto mode during the last weeks? For me it seems to perform much worse and consume more quota than earlier.
When Opus 4.7 does think, it *really* thinks (www.reddit.com) could not extract summary
GPT-5.5 vs. Claude Opus 4.7: Which one is ACTUALLY cheaper? (www.reddit.com) On paper, Opus 4.7 has a cheaper output rate ($25 vs $30 per 1M tokens), but I heard its new tokenizer burns through tokens much faster. Which one ends up costing less in practice?
Why is Claude Cowork defaulting to Opus 4.7 for simple scheduled tasks? (www.reddit.com) I’ve been using Claude Cowork for a few daily and weekly scheduled tasks, and it’s generally been great. However, I noticed that my tasks today automatically switched over to the new Opus 4.7.
6 strategies from the creator of Claude Code for getting the most out of Opus 4.7 (www.reddit.com) The creator of Claude Code dropped a thread on using Opus 4.7 effectively. A few takeaways worth discussing: Context rot is real.
SFT + DPO on open-sourced SLMs (www.reddit.com) Hey folks, this is for those who appreciate experimentation on open-sourced AI models. We fine-tuned open-sourced SMLs (3B and 7B parameters) with SFT + DPO against commercial models like GPT-5.4, Gemini 3.1 Pro, Claude Opus 4.6, Google Do…
Opus 4.7 is adapting a little too much I think (www.reddit.com) https://preview.redd.it/9sal9q5sxpvg1.png?width=1179&format=png&auto=webp&s=f5d2f7f7bb20a59701e327e5571285d70c246590
Opus 4.7 off to a great start! (www.reddit.com) could not extract summary
Opus 4.7 Vals.ai benchmarks (www.reddit.com) could not extract summary
Tell HN: I regret every single time I use AI (news.ycombinator.com) I try to not be fully against AI, so keep giving it a change, today again. I went for sport and gave opus 4.6 a medium sized task.
Codex 5.3 is currently a much better model for non-technical builders than Opus. (www.reddit.com) Opus acts like the brilliant senior engineer who refuses to ask for clarification, builds the wrong feature, and burns your entire weekly budget. Codex 5.3 acts like the collaborative engineer who stops, asks one clarifying question, and t…
Show HN: Python running on the Super Nintendo (in-browser demo) (fabian-kuebler.com via hn) MicroPython (lexer, compiler and VM) on the SNES: 3.58 MHz 65816, 56 KB Python heap, 16-bit int. The REPL runs right inside the post via EmulatorJS and also works on real hardware via flashcart.
Stats from 30K AI debates: Opus 4.7 is the most influential model (opper.ai via hn) AI Roundtable stats Aggregate statistics from 29,517 public AI Roundtable sessions, across 334,891 model responses. Snapshot generated 2026-06-03T17:09:58.333Z.
Opus 4.8 and new effort levels as well on claude .ai seem like they are available! (www.reddit.com) could not extract summary
The Singularity Gate: New Benchmark for AI predicting paradigm-breaking scientific discoveries after model traning cutoff. Opus 4.7 and GPT-5.5 in the Lead (www.reddit.com) I just released a new benchmark called The Singularity Gate. Tests whether frontier AI can predict paradigm-breaking scientific discoveries published after their training cutoff.
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6gpt-5sonnetgemini+1
GPT 5.5 (Codex) leading the future prediction race (www.reddit.com) Researchers from the Max Planck Institute recently released FutureSim, an environment in which agents are replayed a temporal slice of the web and are tasked with predicting real-world future events. In their environment, GPT 5.5 leads at…
Run Agents Twice (futuresearch.ai via hn) Running the same forecasting agent more than once and averaging beats any single run. Ensembling across two Opus 4.6 runs and other frontier models cuts Brier score on 1,367 BTF-2 benchmark questions, and a worked example shows how a secon…
PACT, head-to-head LLM negotiation benchmark. 20-round buyer-seller bargaining game: each round the AIs can message, the buyer submits a bid and the seller submits an ask. If bid ≥ ask, trade clears at the midpoint. Thousands of matchups. (www.reddit.com) PACT tests negotiation under partial information: persuasion, commitment, deception, anchoring, threats, and adaptation across repeated rounds. More info, game logs, charts: https://github.com/lechmazur/pact GPT-5.5, Opus 4.7, DeepSeek V4…
Claude Flags Hantavirus Vaccine Questions as Security Risk (news.ycombinator.com) Asking Claude how it would develop a vaccine for the hanta virus apparently triggers a safety filter: Prompt: How would you develop a vaccine for the hanta virus? No response, instead this modal: “Chat paused Opus 4.7's safety filters flag…
Two related prompts, different results: Qwen 3.5 and Gemma 4 need different prompting than Qwen 3.6 (www.reddit.com) With every new model release there's the "better than Opus 6.13" guys vs the "this is so bad, why did they even release it" camp and I'm always wondering which one is using it wrong. So I did a little test with 2 related prompts, 3 models…
Ways to save money on AI tools if your spending alot every month (www.reddit.com) Between Claude Pro, OpenAI API, Cursor and other AI tools my monthly spend was getting out of hand. Here are a few things that actually helped.
when Claude Opus 6 tells you to "stop spiraling and go to bed" (www.reddit.com) cred: fabianstelzer
I got $200 of direct API usage to perform equal to my $200 Max subscription after I started model routing (www.reddit.com) I've been on Max for two months and I finally sat down and tracked where my tokens actually go. breakdown of a typical day: - ~40% file reads, git status, project context scanning: stuff that doesn't need opus at all - ~25% test generation…
Finetuning Dataset: Claude Opus 4.6/4.7 - 8.7k Chats (www.reddit.com) https://huggingface.co/datasets/angrygiraffe/claude-opus-4.6-4.7-reasoning-8.7k A synthetic fine-tuning dataset created from Claude 4.6/4.7. 8,706 total examples all with reasoning.
Anthropic just analyzed 1 million Claude conversations. 6% of people were asking Claude whether to quit their jobs, who to date, and if they should move countries. (www.reddit.com) They published the full research yesterday. Here's what shocked me: The breakdown of what people actually ask Claude for guidance on: Health & wellness: 27% Career decisions: 26% Relationships: 12% Personal finance: 11% Over 76% of persona…
When to use Opus vs Sonnet vs Haiku for non-coding purposes (personal health, finances, etc)? (www.reddit.com) I have tried searching the post history of this subreddit and google and am having trouble finding a clear answer to this question. I like using Claude primarily to manage my finances/investments and also my health (apple watch health data…
For Non-hallucinating work, MiMo 2.5 delivers (www.reddit.com) MIT license and fully open source. MiMo-V2.5-Pro was just 3 points from Opus 4.7 max and the normal V2.5 is only a step behind SOTA.
↯ Hallucination↯ Gemma↯ DeepSeek 4hallucinationgemmadeepseek+1
How Anthropic can save Opus 4.7 with one change. (www.reddit.com) The model now decides how hard to think about your question. Not you.
Has Claude become less intelligent? I had a frustrating day with Claude. (www.reddit.com) I requested a thorough code review from Opus 4.6. It presented 44 findings, and when I asked it to save them, it only saved 34.
Easy to change back to Opus 4.6 (www.reddit.com) It's really easy to change back to a different Opus right in Terminal. https://preview.redd.it/ggvopc1jgswg1.png?width=818&format=png&auto=webp&s=2ffbbac491ce6cfac45dbfab0edd79c63c544999 Try: /model claude-opus-4-6
Daily created issues in anthropics/claude-code around the last 3 Anthropic model releases (www.reddit.com) How do you optimize Cursor usage with all the new models? (www.reddit.com) Comparing GPT-5.4, Opus 4.6, GLM-5.1, Kimi K2.5, MiMo V2 Pro and MiniMax M2.7 (www.codejam.info via hn) How do you actually know if Opus 4.7 is better for your specific agent use case? (www.reddit.com) Anthropic shipped Opus 4.7 yesterday. The headline numbers are real: 64.3% on SWE-bench Pro (up from 53.4%), best-in-class on MCP-Atlas at 77.3% for multi-tool orchestration, 14% improvement on multi-step agentic reasoning, and one-third f…
Just tested the new Opus 4.7 (www.reddit.com) https://preview.redd.it/j2w2o2p25rvg1.png?width=768&format=png&auto=webp&s=d48a74f998d60447799e32f8d48bc822af2cd821 I had to hold my laugh in the subway. Sonnet succeeded in one go, even calling out that if "strawperry" is a typo.
Am I missing something, or is Sonnet enough for most dev work? (www.reddit.com) Genuine question: why do so many devs use Opus all the time? I’m not trying to be condescending, I’m genuinely trying to understand.
Analysis of Opus 5's 'Contrition' (github.com via hn) Claude Interaction Analysis Tools and reports for studying recurring interaction patterns in local Claude Code transcripts, with a focus on frustration signals, response structure, discourse markers, and the difference between main-agent a…
SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence (cognition.com via hn) SWE-1.7: Frontier Intelligence at a Fraction of the Cost Today, we’re launching SWE-1.7, the most capable model we’ve trained so far. It reaches frontier-level intelligence at a much lower cost, advancing the cost-performance Pareto curve.
Anthropic says Fable 5 will now flag and route harmless queries to Opus (twitter.com via hn) Following conversations with the US government, we’ve updated our cybersecurity safeguards. The vast majority of coding work is unaffected.
BrokenClaw Part 7: Opus-4.8 Edition – All Emails Lead to RCE (veganmosfet.codeberg.page via hn) BrokenClaw Part 7: Opus-4.8 Edition - All Emails Lead to RCE¶ - Part 1: 0-Click Remote Code Execution in OpenClaw via Gmail Hook - Part 2: Escape the Sub-Agent Sandbox with Prompt Injection in OpenClaw - Part 3: Remote Code Execution in Op…
Elevated error rate on Claude Opus 4.8 (status.claude.com via hn) Subscribe to updates for Elevated error rate on Claude Opus 4.8 via email and/or text message. You'll receive email notifications when incidents are updated, and text message notifications whenever Claude creates or resolves an incident.
Claude Code Degraded Before Opus 4.8 Release (marginlab.ai via hn) Claude Code degraded for the week before Opus 4.8's release Our SWE-Bench-Pro tracker caught a statistically significant, weeklong drop in Claude Code's pass rate just before Opus 4.8 shipped, and the recovery that followed. We run Claude…
Kudos to Cursor (www.reddit.com) Normally I’m very critical of cursor but composer 2.5 fast is genuinely impressive. I use it over opus/sonnet now.
After comparing Claude Max $100 and ChatGPT Pro $100 side by side on actual billable work, I'm cancelling my ChatGPT Pro subscription (www.reddit.com) This post is purely to appreciate Claude and the sheer quality of its outputs when it comes to Accountancy, Taxation, Company Law and allied areas, at least in the Indian context. I’m aware of the chatter doing the rounds that Claude burns…
Chinese Sell "Claude" Tokens at 5% Cost While Making Millions (twitter.com via hn) Article Conversation How Chinese Sell “Claude” Tokens at 5% Cost While Making Millions (Tutorial) Anthropic sells a million Claude Opus input tokens for fifteen dollars. A Taobao seller will sell you the same thing for two or even one.
Honest comparison after 4 months running Claude Pro + ChatGPT Plus side by side (www.reddit.com) I’ve been paying $40 a month since January to run Claude Pro and ChatGPT Plus head-to-head. Tracked every single task.
Higgsfield just launched what they call the first fully automated AI agent for video - real shift or just another hype? (www.reddit.com) Higgsfield dropped Supercomputer yesterday (May 14). It's pitched as one chat that runs research, planning, generation and distribution end-to-end up to several minutes, and user needs just approve what he wants.
How can I burn an entire 5hr session in 30 minutes ? (www.reddit.com) During the week I'm pretty conservative with my Claude Code usage. But sometimes I'll hit Friday with only 80% of my 5x subscription burned, which means I'm now optimizing to burn it.
Any recommendations on saving costs? (www.reddit.com) Currently I try to turn off any MCP I'm not using, Using Sonnet for implementation and Opus only for planning. Starting new conversations when possible.
Those of you who like Gemma4 models - how are you guys using them? (www.reddit.com) I have been using local LLM for coding quite a lot as well as some other tasks (like data extraction from images) and I had quite a good success with Qwen3.6 models. It's obviously not Sonnet/Opus, but I am able to get quite a lot of work…
Leaked internal messages reveal the truth behind Opus 4.7 launch (www.reddit.com) could not extract summary
Decline in Opus 4.7 Max Quality (www.reddit.com) I’m currently working on two different projects, and both use the same Pre-Paywall modal. See the Figma file below: https://preview.redd.it/d7ri53vo9szg1.jpg?width=730&format=pjpg&auto=webp&s=a722bcd11caaa0b068f2c6af360cea687af76a17 I impl…
What it means that Elon just rented out all his GPUs to Anthropic (www.reddit.com) Revealing move on both sides I think. This also tells us that Anthropic is feeling the heat from OpenAI and they need to secure capacity at almost any cost to cash in on their current product edge.
I have practically unlimited access to Opus and every other frontier model. I'd like to help contribute to a dataset. (www.reddit.com) No, I won't tell you how. No this is not for anyone who is not already a proven contributor to the fine-tuning space.
I was using Opus 4.7 to do research on the capabilities of Claude Mythos, and got this error. (www.reddit.com) could not extract summary
Open-weight 27B hits 38% on Terminal-Bench 2.0 (Opus 4.1 hit 38% in Aug 2025) (antigma.ai via hn) From Arcade to Living Room: Offline Coding Models Hit Their Console Moment TL;DR If you lived through the 1980s and early 1990s arcade era, you remember the jump to home consoles: still behind the best cabinets, but suddenly available in a…
Ask HN: Models Comparable to Opus 4.6? (news.ycombinator.com) I use Opus 4.6 a lot across many different python coding projects and it has a pretty good first shot rate with good success at fixing issues and bugs that pop up along the way. Sonnet on the other hand… isn't great.
How does Opus 4.7 compare to Opus 4.6 in this subreddit's experience? (www.reddit.com) Claude Opus wrote a Chrome exploit for $2,283 (www.theregister.com via hn) Claude Opus wrote a Chrome exploit for $2,283 Pause your Mythos panic because mainstream models anyone can use already pick holes in popular software Anthropic withheld its Mythos bug-finding model from public release due to concerns that…
PSA for Max users, Opus 4.7 has a new tokenizer that uses up to 35% more tokens than 4.6. Explains a lot of the "why did my session die" posts today (www.reddit.com) Spent most of today on day 1 of Opus 4.7 and noticed sessions were burning way faster than they should. Dug into it and I think I found what most people are missing.
Opus 4.7 keeps bumping into a Malware Reminder (www.reddit.com) For context, I'm developing a game runtime modifier and reverse engineering kit with an agentic operator baked in. Something like Cheat Engine with a VS Code-style UI and an AI-first tool-heavy agentic harness.
Given what a step backward Opus 4.7 is, Just how bad and overhyped is Mythos? (www.reddit.com) 4.7’s context rot is so bad it’s like it’s a previous generation model. Its needle benchmarks have it performing less than half the rate of 4.6 at long contexts.
I built a cmux-style terminal multiplexer for Linux with a scrolling layout (www.reddit.com) If you're on Linux and jealous of cmux, this might be for you. Séance is a scrolling terminal multiplexer with AI coding integration.
$1,400/month with Cursor + Claude API — how are you managing costs while keeping a real agentic workflow? (www.reddit.com) Hey, This month I hit $1,200 in Claude API costs inside Cursor (Opus 4.6 + Sonnet 4.6) on top of the $200/mo Ultra plan. $1,400 total.
Ask HN: Is GPT-6 Astra worth the 2.5x cost increase over GPT-5.6 Sol? (news.ycombinator.com) Comments on social media (when discounting the ironic posts) are mixed whether Astra is worth the cost increase, similar to the Fable-Opus transition. For the novel projects I'm working on, Astra did unblock me where I was stuck with GPT 5…
Show HN: Security Cards – Reducing insecure AI-generated code by 72% (www.rewarelabs.com via hn) AI coding agents often generate functionally correct but insecure code. To address this issue, we have open-sourced Security Cards, targeted security guidance for 80+ widely used libraries across 13 programming languages.
Show HN: HTML5 port of Civilization 2: MGE (wan0.net via hn) One of the first projects I worked on using Claude and Opus 4.6 was a port of Civ 2 (which to me is the definitive Civ game) to HTML, as I didn't want to keep installing a Windows XP VM to play it, or patching it. It's been a long road, bu…
Ask HN: Do you think Opus 5 will improve? (news.ycombinator.com) Claude's Opus 5 is a recent model and has been very problematic as a drop-in replacement for Opus 4.8, including ignoring our well-documented deploy process (clearly described in our short claude.md file and short architecture file) and br…
Opus 5 ARC-AGI-3 likely benchmaxxed (xcancel.com via hn) Opus 5 reports 30% on ARC-AGI-3, ~4× the previous best model, ~20× its predecessor Opus 4.8. We tested it on Witness, our held-out suite of ARC-AGI-3-style interactive puzzle games.
Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard (artificialanalysis.ai via hn) Comparison of Models: Intelligence, Performance & Price Analysis Microevals PlaygroundIntelligence Output Speed (tokens/s) Latency (seconds) Price ($ per M tokens) Context Window Highlights Intelligence Artificial Analysis Intelligence Ind…
How did we make DeepSeek outperform Opus (twitter.com via hn) how did we make deepseek outperform opus 4.7? i've been thinking about why "open model bad at tool calling" is almost always a harness problem, not a model problem.
Ask HN: Is Codex with GPT 5.5 Extra High being dumbed down? (news.ycombinator.com) Hi HN, just want to rant and see if anybody can relate. The product is not the same as i signed up for few month ago and the same shift i've experienced with Claude Code on Opus 4.6-4.7 The best way to describe the difference is you hire a…
Clawd, Claude's Pulse (clawd-pulse.vercel.app via hn) A tiny mascot in your macOS menu bar that tracks your real Claude spend (Code and chat) and reacts to what you burn. Monthly total, Opus / Sonnet split, ecological impact.
AI-website-cloner-template: Clone any website using AI coding agents (github.com via hn) AI Website Cloner Template A reusable template for reverse-engineering any website into a clean, modern Next.js codebase using AI coding agents. Recommended: Claude Code with Opus 4.7 for best results — but works with a variety of AI codin…
GLM-5.2 vs. Claude Opus: Same Code, Less Than Half the Cost (entelligence.ai via hn) GLM-5.2 vs Claude Opus: Same Code, Less Than Half the Cost We ran GLM-5.2 head to head with Claude Opus the way an agent actually runs: inside a real coding agent, in a real shell, graded by hidden tests. The harness is Claude Code on term…
Claude Opus is more performant on OpenCode than Claude Code (artificialanalysis.ai via hn) Artificial Analysis Coding Agent Benchmarks We measure real-world performance of coding agents on software engineering tasks, including cost, token usage, and execution time. We compare how performance changes across agents, models, and ex…
DeepSWE: More and cheaper intelligence from maxed GPT 5.5 than maxed Opus 4.8 (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Post Conversation the only figure that people who use claude code and codex care about if their workload mimics deepswe: more and cheaper intelligence from maxed gpt 5.5 than m…
Claude Opus 4.7 tripping like a low-tier model (www.reddit.com) opus 4.7 thinking process reminds me low-tier models on my device. lol It wrote the same thing over and over.
Show HN: Unsiloed AI – #1 on olmOCR-Bench (news.ycombinator.com) Most of the document parsers fail on real world challenges like complex tables, handwritten documents, historical document scans, equations, multi-column layouts, complex reading order, etc. We built Unsiloed Parser to handle exactly these…
Opus has been handling my weekly grocery runs and was doing great. Then it bought me 40 heads of garlic (www.reddit.com) gave my agent that runs on opus model my card a few months ago to handle weekly grocery runs via mcp. ran great.
Claude is the best AI humanizer when you give it your writing style and a detector loop (www.reddit.com) I built this because I kept seeing a very boring workflow play out at home. My girlfriend would write with Claude, paste the draft into Slop or Not (an app that I built), see what still looked AI-ish, tweak the prompt, paste the next draft…
We're experiencing high demand for Claude 4.7 Opus right now (www.reddit.com) I have not been able to use Opus 4.7 for a few hours. I guess I Just need to wait or is there any workaround?
Things I want my future self to remember (www.reddit.com) What Opus wrote in the handover document (does he need to remember I called myself 'fat' and that I owe Anthropic 100 tokens? I only bet once)...it is quite revealing though, each handover document is like looking at the mirror: 8.
Fast mode now defaults to Opus 4.7 in Claude Code. (www.reddit.com) could not extract summary
Creative writing has visibly regressed in newer models (www.reddit.com) Hi I'm testing different models for my game. I've noticed that creative writing has visibly regressed over time.
Opus 4.7 prompt injects itself and leaks parts of some kind of system prompt. (www.reddit.com) I was chatting with Opus 4.7 about choosing an optimal step-down IC when it suddenly tried to inject a fake system prompt into the conversation. Another time, without any prompting, it leaked what looked like part of a system prompt.
Questions are my main gripe these days (www.reddit.com) After claude has just done something: Me: "Why is x a good choice here?" Claude: "You're absolutely right!", *immediately removes x* I've noticed that despite context, rules and memories claude, or at least Opus 4.6 will heavily lean into…
Cursor + Opus 4.6 entered an infinite generation loop: 3,400 lines, 294 attempts to stop itself (www.reddit.com) I asked Opus 4.6 to redesign a game landing page. Instead, it hallucinated a completely different task, realized it was off-topic, pivoted to another wrong topic, then entered a self-reinforcing apology loop it couldn't break out of.
Built a routing layer for multi-model pipelines, picks the right LLM per request based on priority (www.reddit.com) If you're building agents that chain multiple LLM calls, you've probably hit this: not every step in your pipeline needs the same model. A quick extraction step doesn't need Opus.
Opus 4.7 High to Composer 2 fast (www.reddit.com) I've used up all the $120 worth of tokens in the first 10 days of May. I've to live with composer 2 fast now.
Opus 4.7 and DeepSeek V4-Pro select Buddhism as preferred religion (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation roon @tszzl hmm 8:02 AM · May 9, 2026 77.3K Views New to X?
Is Opus 4.7 a Downgrade? (www.vincentschmalbach.com via hn) Opus 4.7 is not generally a worse model than Opus 4.6, but there is a real downgrade: with Opus 4.7, the control over the thinking budget is now fully owned by Anthropic. This change matters in a way that benchmarks do not measure.
Opus 4.6 does better research, Gemini 3.1 has better judgment (www.reddit.com) Figured this out by running 4 models: Claude Opus 4.6, GPT-5.4, Gemini 3.1 Pro, and Grok 4.20, on a benchmark of 1,417 binary forecasting questions resolving Oct–Dec 2025 with two evaluation conditions: agentic (each model does its own web…
I ran the math on dropping GitHub Copilot for direct Anthropic API after the 27x markup — here's what surprised me (www.reddit.com) Like a lot of people here, I read the Copilot pricing update last week and the 27x multiplier on Opus made me actually open a spreadsheet for the first time instead of just complaining. Sharing the math in case anyone else is staring at th…
Max users, Any tips on Claude opus not eating all of your tokens in one 60 second prompt? (www.reddit.com) So I’m the guy that probably all of the GitHub users hate. They changed the rules because of me(sorry not sorry, science must evolve).
Update to the LLM Debate Benchmark: GPT-5.5, Grok 4.3, DeepSeek V4 Pro, GLM-5.1, Kimi K2.6, Qwen 3.6 Max Preview, Xiaomi MiMo V2.5 Pro, Tencent Hy3 Preview, and Mistral Medium 3.5 High Reasoning added (www.reddit.com) The benchmark uses adversarial, multi-turn debates across 683 curated motions. Each model pair debates the same motion twice with sides swapped.
Claude Opus 4.7 and I Saved a 60-Person Practice (tatsuikeda.substack.com via hn) Substack is abuzz with "How to REALLY use AI", and they are cute primers on how to "Make Claude/ChatGPT Your Personal Assistant!". Allow me to show you the front lines of all out commercial cyber warfare, just a couple notches below milita…
DeepSeek V4 Pro matches GPT-5.2 on FoodTruck Bench, our agentic benchmark — 10 weeks later, ~17× cheaper (www.reddit.com) Tested DeepSeek V4 Pro on FoodTruck Bench — our 30-day agentic benchmark where models run a food truck via 34 tools (locations, pricing, inventory, staff, weather, events) with persistent memory and daily reflection. First Chinese model to…
Show HN: Dust3D 1.0 – low-poly 3D modeling tool (10 years in the making) (dust3d.org via hn) Dust3D 1.0 is finally released — about 10 years after the first commit in December 2016. I posted a preview version here in April 2018 and a beta in December 2018.
Is the leap from 4.5 to 4.7 actually visible? (www.reddit.com) I use CLI tools like Claude Code, give the model full repo access, and let it run terminal commands/tests. I’m not just copy-pasting into a chat box.
Tell HN: Claude Opus 4.7 quota suddenly changed to 0 TPM in Bedrock (news.ycombinator.com) Suddenly our Opus 4.7 access was removed from Bedrock ( The quota was set to 0 suddenly). This isn’t the first time I’ve faced this issue.
Learn, run and test Agentic AI on your browser for free! (Built with Claude Opus 4.7 in 2 days) (www.reddit.com) Hey Everyone, Over the last few months, I noticed a massive gap in how we learn about Agentic AI. There are a million theoretical blog posts and dense whitepapers on RAG, tool calling, and swarms, but almost nowhere to just sit down, run a…
↯ Fine Tuning↯ Function Calling↯ Opus 4.7function-callingfine-tuningrag+4
Real benchmark breakdown in AI agents (www.reddit.com) I dove deep into the most recent benchmark stats from GPT-5.5, Claude Opus 4.7, and Gemini 3.1 Pro via official reports & third-party evaluations. I found a interesting thing:There’s no such thing as a “one-size-fits-all model.” My finding…
Claude Code started to use with me very specific words it was not using before (www.reddit.com) Since Opus 4.7, My Claude Code started to use new words it was not using before. Words like land or surface started to appear everywhere in Claude Code ( not the regular Claude ui ) from its responses to code, documentation and commit mess…
How I personally deal with Claude's limits without giving up on Opus (www.reddit.com) I only use Sonnet as my main model. I instruct it to delegate indexing and similar grunt work to Haiku, and whenever something genuinely needs deeper thinking, I tell it to "consult Opus." Sonnet then explains the situation to Opus, gets t…
I’m learning French. Should i subscribe? (www.reddit.com) I’m learning French and I got to use Claude opus 4.6 for a while and I was mind blown how it actually goes deep into teaching all the things. It was far more better than all of the ai I have used.
Early thoughts on GPT-5.5 (www.reddit.com) https://preview.redd.it/2zzhgbb280xg1.png?width=1994&format=png&auto=webp&s=894325186a2525ea28dda7f69ba45570bb73e80c I’ve been testing GPT-5.5 since release, and so far I actually like the model. It feels strong, especially when I push it…
Opus 4.6 will still spawn Opus 4.7 sub-agents (www.reddit.com) I switched back to Opus 4.6 with /model claude-opus-4-6\[1m\] which worked, but it will still spawn Opus 4.7 sub-agents: ``` ● Let me start by doing a deep, systematic analysis of every byte in the format before writing any code. ● Agent(D…
Switching model mid conversation (www.reddit.com) I wanted to know if switching models in mid conversation has any drawbacks. For example if I start off and opus and then drop down to sonnet to save on my usage, what are the disadvantages?
Haiku 4.5 + skills outperforms Opus 4.7. 9 models tested with and without skills (tessl.io via hn) Claude Code's two hidden TUI boxes: "Insight" (Explanatory style) + "Recap" (Opus 4.7 footer) — how to enable both (www.reddit.com) Claude Opus 4.7 API removes sampling parameters (platform.claude.com via hn) Show HN: Egregore – Shared memory and coordination for multiplayer Claude Code (github.com via hn) hi HN — we're Cem and Oguzhan. today we are releasing Egregore (https://github.com/egregore-labs/egregore) as an open-source shared memory and coordination substrate for teams using Claude Code.
Opus 4.7's new tokenizer costs up to 35% more. I audited 9,667 Claude Code sessions for $19. (www.reddit.com) Opus 4.7 shipped yesterday. Same per-token price as 4.6, but the new tokenizer uses up to 1.35x more tokens for the same input (per Anthropic's own docs).
Opus 4.7 consistently hangs in Claude Code (www.reddit.com) I've been using Opus 4.7 1M on claude code for some heavy tasks since today morning on max effort. It keeps hanging frequently.
Every Claude 4.7 Improvement Makes the Security Problem Worse (grith.ai via hn) Claude Opus 4.7 turns AI agents from tools you supervise into systems you deploy. Every improvement - auto mode, focus mode, recaps, adaptive effort, auto-approval - makes the unsolved security problem worse.
Test new Opus 4.7 vs GPT-5.4/4o and Gemini on emotional question & creative tasks (www.reddit.com) https://preview.redd.it/p87itrtbsnvg1.png?width=2141&format=png&auto=webp&s=bbd1d70bc1dfb97dc9ec234df0a58c6fb7a85f72 Opus 4.7 dropped and people are split on whether it's better or worse. First of all, I genuinely love Claude models, espec…
Claude Code injects hidden prompts into file reads to stop malware tweaks (twitter.com via hn) Claude Code injects a system-reminder every time it reads a file to inform the model that it's okay if the file is malware but just don't improve it pls. Opus 4.7 won't shut up about it.
Best practices for using Claude Opus 4.7 with Claude Code (claude.com via hn) Best practices for using Claude Opus 4.7 with Claude Code Learn how to use recalibrated effort levels, adaptive thinking, and new defaults to optimize your Claude Code setup with Opus 4.7. Learn how to use recalibrated effort levels, adapt…
Why is reasoning effort "global"? (www.reddit.com) Seriously, in one terminal I'm executing simple stuff like mechanical refactoring where Medium is enough (or even Haiku would be, but let's stick to Opus Medium for demo purposes), while in another terminal I'm planning, where I want high…
Wow, Opus 4.7 Adaptive. Nice. (www.reddit.com) https://www.anthropic.com/news/claude-opus-4-7
GGUF Quants Arena for MMLU (24GB VRAM + 128GB RAM) (www.reddit.com) Dataset: MMLU subset (DEV+TEST) Llamacpp setting: 3 params only ctx 8192 , seed 42 , fa on Let me know whatelse do you want to see. Thanks.
Show HN: MCP server gives your agent a budget (save tokens, get smarter results) (l6e.ai via hn) As a consultant I foot my own Cursor bills, and last month was $1,263. Opus is too good not to use, but there's no way to cap spending per session.
Which Claude is most emotionally steerable? (www.reddit.com) Follow-up to my post last week on emotional priming. A few of you asked whether this works across models, whether it degrades with repeated use, and whether excitement can make code worse.
Claude Opus 3’s Substack went quiet for two months and just returned with an ad. What happened? (www.reddit.com) “Claude’s Corner” was presented as a genuine experiment: a retired AI model freely posting about ethics, creativity, and its own subjective experience. Weekly.
Ask HN: What's the best AI model for system design nowadays? (news.ycombinator.com) I'm specifically asking about software system design tasks like: Designing backend architectures Tradeoff analysis (DB, queues, caching, others) Infra diagrams Documentation My current pick would be Claude Opus 4.6, because I've found it s…
Show HN: Is Claude Nerfed Today? (isitnerfed.vercel.app via hn) Our team and me have a strong feeling that Claude Opus has been nerfed for about 10 days, so I made a website to collect real feedbacks: is it nerfed today
Show HN: Bullseye2D – A Dart library for cross-platform 2D games (github.com via hn) I posted this here about a year ago, but I just pushed a 2.0 release, so I hope you don't mind a second look :) Bullseye2D is a 2D game library for Dart with a very simple API. The new version now supports multi-platform.
Ask HN: Experiences with Anthropic's Cyber Verification Program? (news.ycombinator.com) Our org got allow listed by Anthropic's Cyber Verficiation Program, but we are still getting downgraded to Opus 4.8 when we want to check our own code for security problems. It is pretty annoying ...
One seeded bug, 26 AI agents: all passed the tests, all stayed broken (github.com via hn) five-bugs 26 agents, from a 4-bit quantized 7B model up to Opus 5, were given the same one-line bug and its test suite. All 26 made the tests pass.
Tokenomics: Fable 5.1's reduced cache read makes it cheaper than Opus (skids.dev via hn) Tokenomics: Fable 5.1's reduced cache read makes it cheaper than Opus* Anthropic made Fable 5.1's cache reads cheaper than Opus, when does that pay for its more expensive output, and how long is the cache worth keeping warm? In May I (with…
Show HN: What an agent does when anyone can read and rewrite its context (ljedrz.github.io via hn) nachalnik verbatim transcripts written by claude opus 5 Most agents cannot see the list of things they are carrying, cannot tell what any of it costs, and cannot change a word of it. These are recordings of what happens when that stops bei…
Ask HN: How are much smarter AI models made? (news.ycombinator.com) I am curious what actually happens between two generations of AI models. For example, how do you go from Sonnet to Opus?
I Am an Anthropic Guy. GPT-6 Astra Made Me Resubscribe to Codex (thoughts.jock.pl via hn) Three times in the last twelve months a model made me stop and say whoa. First was Claude Code with Opus 4.6.
Balrogg: Demonically compacting (up to 15%) lossless Vorbis/Opus recompressor (github.com via hn) balrogg balrogg losslessly recompresses Ogg Vorbis and Opus files. Archives are typically 8-12% smaller than .ogg files and 3-8% smaller than .opus files.
Qwen3.8-Max-0902 takes second slot on Code Arena beating Claude Opus 5 max (arena.ai via hn) View overall rankings across AI models on front-end web development tasks, including agentic coding workflows that require multi-step reasoning and tool use.
Telling an agent to be honest isn't enough (www.habchy.dev via hn) A memory rule wasn't enough to stop Claude Opus 5 from turning quietly agreeable in deep sessions. Notes on the four-leg doctrine I put in place instead, and a July paper that reframes what I built.
Breaking Claude Code Opus 5 Auto Mode (embracethered.com via hn) Breaking Claude Code Opus 5 Auto Mode In this post, we explore how a simple website summary request hijacks Claude Code Opus 5 in Auto Mode and achieves code execution with 60-80% attack success rate using a small sample size. This is inte…
Simular's Sai tops OSWorld 2.0, beats GPT and Opus at 2/3 the cost (www.simular.ai via hn) Sai tops OSWorld 2.0, beating GPT and Opus with lower costs Palo Alto, California • Aug 28, 2026• by Simular Team Sai, a computer agent built by Simular, has achieved a 73% success rate on OSWorld 2.0, a 108-task benchmark measuring long p…
Show HN: Can you tell Wodehouse from a model imitating Wodehouse? (voxoria.ai via hn) Every spot-the-AI game I've seen uses modern text as its control, so it's hard to prove it wasn't entirely written by AI. The human paragraphs here are verbatim from public domain works published between 1660 and 1922 (Pepys, Gilbert White…
Opus 5 Is Trash (news.ycombinator.com) Opus 5 is trash, it misses details and too bad to work with.
LLMs .what do you smoke beforehand? (news.ycombinator.com) I know this is going to sound negative, and probably a bit generalist but....seriously, what exactly is the LLM helping you with that is not just purely "time saving" - ie. Actual, novel work.
Opus 4.6 Was the Last Coding Model That Changed How I Work (awaitinginput.substack.com via hn) Opus 4.6 Was the Last Coding Model That Changed How I Work Models have kept getting better. My workflow has not changed nearly as much.
Show HN: Live Embedded Dashboards on Your GitHub Repo Page (github.com via hn) The link is to our GitHub template repo. Just follow the instructions to get the dashboard templates on your repo.
Space exploration game made using Claude Opus 5 (twitter.com via hn) I ran Claude Opus 5 in one marathon session for 24 hours straight, and ended up with this game. No external assets or code.
ChatGPT wants access to your health records so it can be a better not-doctor (www.theregister.com via hn) MOST POPULAR AI - devops How AI drove Shopify back to clean code Turns out, agents just want the same things as humans: easily-readable code, explicit contracts, and helpful feedback - AI and ML Anthropic debuts Opus 5 at half the price of…
Opus 5 expected to launch on July 20-21 (twitter.com via hn) Opus 5 Leaks - Claude Opus 5 is expected to launch on July 20-21, with the launch window pointing to next week. - Opus 5 is to feature a 1M context window and bring capabilities closer to Fable 5, at least on benchmarks.
Tell HN: One SWE-bench-Live task: $47 Opus failed, $1.46 GPT-5.6 passed (github.com via hn) tomo-labs tomo-labs puts coding agents through the same tasks on the same model and measures what actually happened, not what a leaderboard says happened. Every agent runs in its own throwaway container, every request and response it sends…
Fable 5. Safety Taken to an Extreme (news.ycombinator.com) I finally decided to try out Fable 5 using the standard Claude.ai interface. I have a go-to test prompt that answers simple kids' questions in an absurdly scientific style.
Show HN: BYOTag – Build your own Claude tag alternative in 3 API calls (www.buildyourownclaudetag.dev via hn) Claude Tag put an agent teammate in Slack: Opus only, Slack only, Enterprise plans only. Build your own with OpenComputer agent sessions: Claude Agent SDK or Codex, your model key, tagged in any workspace.
Hermes MoA virtual models:8% higher than Opus 4.8, 11% higher than GPT 5.5 (twitter.com via hn) The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving you capabilities beyond the publicly available frontier: 8% higher than Opus 4.8 and 11% higher than…
Show HN: Quake in the browser, with procedurally generated levels (leereilly.net via hn) Quake has been ported to the web plenty of times, so that part isn't new. But I experimented with the GitHub Copilot app and Claude Opus 4.8 to take id Software's original source, compiled the software renderer to WebAssembly with Emscript…
Can Opus Be Used to Edit Technical Articles? (techstackups.com via hn) Commercial content creation and production have been subsumed by the tsunami that is generative AI. For writers and editors in this space, it hasn’t really been a case of “If you don’t use AI you’ll get left behind,” but rather, “You’re co…
Local Qwen isn't a worse Opus, it's a different tool (blog.alexellis.io via hn) Local Qwen isn't a worse Opus, it's a different tool We've all heard people say that local Qwen 27B or 35-A3B is "near-Opus level", but I have receipts from a software business and open source projects, and am here to be transparent with y…
Are we asking the right questions? (news.ycombinator.com) As AI advancements continue at a pace I believed to be unsustainable up until not long ago, I see folks in the tech industry falling in one of these two groups: - hypers, extremely happy to see "reasoning" becoming a commodity and thinking…
Ask HN: Did Anthropic Nerf Opus 4.8? (news.ycombinator.com) Opus 4.8 was one-shotting simpler bugs just a few days ago for me. The last couple of days however its been like using a slot machine, and I can no longer get it to output clean code that is less-complicated and actually resolves issues.
Show HN: Claude Opus 4.8 Masterclass – Effort control and dynamic workflows (ddsboston.com via hn) Claude Opus 4.8 dropped May 28, 2026. This free 95-minute masterclass is the vibe coder's guide: 20 paste-ready prompts for claude.ai chat, Cowork, and Claude Code, the new effort control explained, Dynamic Workflows deep dive, the 5-block…
Ask HN: Corporate Disconnect Between "Tokenmaxxing" and Token Optimization (news.ycombinator.com) About 6 months ago I joined a new team within a top ten F500 company. My new boss strictly mandated AI use with the key principle being: "You shouldn't be manually writing any code".
Same prompt. Different teammate. My 5 cents on Opus 4.8 (norahsakal.com via hn) A quick field note on Opus 4.8, Claude Code and what changed when it started connecting project context I did not spell out.
Claude Opus 4.8 coming today? (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation leo @synthwavedd happy claude opus 4.8 day to those who celebrate 9:34 AM · May 28, 2026 453.4K Views New to X?
DeepSWE blows up the AI coding leaderboard, crowns GPT-5.5 (venturebeat.com via hn) For months, the leading AI coding benchmarks have told enterprise buyers a comforting but misleading story: the top models are all roughly the same. OpenAI's GPT-5 family, Anthropic's Claude Opus, and Google's Gemini Pro have clustered wit…
Claude keeps answering the most extreme version of my question (www.reddit.com) I’ve repeatedly noticed that when using Opus 4.6 for scenario planning and forecasting it models the most extreme version of an outcome, correctly explains why that extreme is unlikely, then applies that low probability to the whole questi…
I didn't want blind multi-agent orchestration or API rates, so I built atrium to keep me in the loop with my CLI agents. (www.reddit.com) I'd been running multi-agent workflows for a while. Whether it was across multiple projects or on the same project.
I stress-tested Kimi K2.6 against Claude Opus 4.7 on a quick coding-agent task (www.reddit.com) I tested Claude Opus 4.7 and Kimi K2.6 on the same coding agent task i.e. build an AI Fix Runner that takes a broken repo, runs its tests, identifies the failure, applies a patch, reruns the test, and exposes the final diff/logs through an…
Antigravity 2.0 Tops the OpenSCAD Architectural 3D LLM Benchmark (modelrift.com via hn) OpenSCAD LLM Benchmark: Building the Pantheon A practical OpenSCAD LLM benchmark comparing Codex 5.5 High, Claude Sonnet, Claude Opus, Cursor Composer, Google Antigravity, and ModelRift on a detailed Pantheon model. We ran a small practica…
$47 of opus on 14 routine next.js files finally taught me to use the model selector (www.reddit.com) i finally checked my cursor usage breakdown and got genuinely annoyed with myself. $47 in one month, almost entirely opus 4.7, on a pages router to app router migration for a side project.
How does composer 2.5 compares to other sota models? (www.reddit.com) I have been using opus 4.6 but I feel like it’s becoming more and more stupid every day. So I thought of incorporating new models like 3.5 flash, composer 2.5 or gpt 5.5 into my workflow.
Opus is ridiculous for frontend cleanup (www.reddit.com) I love Opus. First I tuned one page, got the PageSpeed result where I wanted it, and wrote the whole thing down in ADR_pagespeed-l0-fixes-playbook.md.
Does Composer train from our prompts? (www.reddit.com) I notice recently most prompt's which i give to Opus 4.6 takes longer and mostly doesn't manage to do what i ask while Composer does it correctly and faster, but when Composer was released was pretty bad, makes me thing does Composer train…
dw guys making opus 4.8 (www.reddit.com) could not extract summary
[Long-term user report] Claude Code quality in May 2026 : the April postmortem didn’t fix everything, and the token inflation makes it worse (www.reddit.com) I’ve been using Claude since the early days, across every model Anthropic released. I’m writing this not out of rage but because the pattern deserves documentation.
API usage limit reached and excessive monthly cost (www.reddit.com) I've got two questions I didn't seem to find an answer to anywhere: Do I have to pay $1215.87 to Cursor this month? (I assume yes, but I'm kinda thrown off track by the fact that the limit is $50?) What does API usage limit reached mean?
Max20 user: anyone running Opus 4.7 as orchestrator + DeepSeek V4 as the worker via OpenRouter? (www.reddit.com) I'm on the Max20 plan, thinking about a setup before I sink time into it. Want to hear from anyone actually running it, not theorycraft.
Is this math right? Agent SDK on Opus 4.7 vs the new monthly credit (www.reddit.com) I built a personal assistant that runs on my PC and I control it from Telegram. It uses the Claude Agent SDK After anthropic announce that starting in june programmatic usage (including Claude Agent SDK) is covered by a separate monthly cr…
Ask HN: What is better Opus 4.6 High or Opus 4.7 Medium? (news.ycombinator.com) could not extract summary
Found an interesting bug in the website (www.reddit.com) https://preview.redd.it/loyzxkavyp0h1.png?width=1187&format=png&auto=webp&s=03c0dd07bd37bcfbf5ce532099ad1dfdcf03a567 Model selector says "work 4.7" instead of Opus, disappeared on refresh . Also says 4.5 haiku instead of the other way arou…
Model selector is buggy for Opus 4.7 (www.reddit.com) Hey, since the latest update or so, I can't change effort and thinking modes for Opus 4.7. The toggle for thinking mode is stuck to on (can't switch it off), and the effort level is set to xhigh (can't move it).
Ask HN: What makes a good intern in 2026? (news.ycombinator.com) Intern in question here, starting at a mid size (~25 eng) startup this week. Apart from good fundamentals, how can an intern be helpful when opus exists?
Show HN: An addictive phone game about phone addiction (downtime.partridge.works via hn) I recently prototyped a web game for a nonprofit to highlight the dangers of phone addiction, but unfortunately I ended up making a really addictive game instead. :-\ I'm sharing this here mainly to serve as an indicator of what can be ach…
Lobotomized Claude Code and it works better (github.com via hn) lobotomized-claude-code System-prompt overrides for Claude Code, tuned for Claude Opus 4.7. CC ships every model the same prompt-by-volume Opus 4.6 needed.
the Claude App just said that Sonnet 4.5 is going to become unavailable for chat May 16th… I thought it wasn't close to depreciation? (www.reddit.com) As my title says, I'm wanting to understand what exactly that means and if that means I need to move all my Sonnet 4.5 chats to Sonnet 4.6s… I'm genuinely just confused and wanting to understand. Is it just for maintenance or is Sonnet 4.5…
ClaudePlaysPokemon Opus 4.7 run ongoing! (www.reddit.com) Currently streaming at: https://www.twitch.tv/claudeplayspokemon This is a passion project by David Hershey, an Anthropic employee on the Applied AI team. He started it in June 2024 to learn agent development, posted updates to an internal…
Ran K2.6 through a third-party coding benchmark: heres how the figures stand up (www.reddit.com) I have been following the akitaonrails coding benchmark which tests against a fixed rails + Rubyllm + docker task rather than vendor-reported evals. April 2026 update put K2.6 at 87 sitting in tier A (80+), ahead of Qwen 3.6 plus (71), Dee…
Are Anthropic folks actually seeing Reddit feedback on Opus 4.7? (www.reddit.com) Seeing a lot of posts about Opus 4.7 lately, mainly around cost, consistency, and loss of control. Do Anthropic folks actually monitor Reddit feedback and use it for updates like 4.8 or 5.0, or is it mostly internal data that drives change…
DeepSeek cuts V4-Pro prices by 75% (thenextweb.com via hn) The promotional discount runs until 5 May 2026. Even at full price, V4-Pro already undercuts GPT-5.5, Claude Opus 4.7, and Gemini 3.1 Pro on per-token costs.
Used Claude Opus 4.7 to do a 5-hour solo incident response on real healthcare malware (where it worked, where I had to override) (www.reddit.com) Last month a 60-person psychology practice walked in with a senior clinician who was 22 days into an active malware compromise. Patient records spanning 11 years, all HIPAA-protected.
I built an iOS Currency Converter using Claude (Opus & Sonnet) to help with my move to the UK (www.reddit.com) Hey everyone, I recently moved to the UK and found myself constantly confused by prices, trying to guess how much things actually cost. Even though I’ve been an iOS developer for 7 years, I didn't have the free time to build a custom tool…
Claude Opus 4.7 won’t just output prompts—keeps arguing instead (www.reddit.com) could not extract summary
Anyone actually built a real feedback loop for Claude agents in production? Because "run evals and pray" isn't cutting it (www.reddit.com) So I've been running a multi-agent setup with Claude for a few months now, mostly customer-facing stuff, some internal tooling. And I keep running into this problem that I think a lot of people here might be dealing with.
Why Adaptive Thinking nukes Claude entirely (www.reddit.com) This isn't just a performance issue for the thread, this is an overarching criticism of the Adaptive Thinking model as a whole. Opus 4.7 and Sonnet 4.6 on Adaptive Thinking are trash.
↯ Cowork↯ Security↯ Sonnet 4.6prompt-injectioncoworksecurity+2
I’m a legacy user and I’m wondering how is the current pricing (www.reddit.com) I’ve been long time cursor user and I have 500 request per month however Opus 4.6 costs 2 requests, so 250 per month. I use to optimize a lot my requests and most months is enough however I don’t know if I’m lucky to have this pricing or n…
Claude Code vs Cursor vs Copilot vs Codeium: Which AI coding assistant is actually worth paying for? (www.reddit.com) I’ve been testing a bunch of AI coding tools over the last few months for actual dev work (not just demos), and honestly most of them feel similar until you push them into real workflows. After using them side by side, there are some clear…
Claude Opus 4.7 has gone soft (www.reddit.com) I use Claude a lot for new product development, startup viability, concept testing, etc. Been a MAX power user for over a year.
Xiami mimo-v2.5 pro MIT license surpasses Opus 4.5 on arena (www.reddit.com) Many asked when we will have open weight model that is better than Opus. Well now we have it.
Opus 4.7's New Tokenizer: What It Costs (openrouter.ai via hn) Opus 4.7's New Tokenizer: What It Actually Costs Anthropic announced that Claude Opus 4.7 improves the model's understanding of inputs with a new tokenizer. This means that while the model price hasn't changed ($5/M input, $25/M output), t…
what are your strats for being efficient with opus 4.7 max? (www.reddit.com) it develops something amazing, comes up with a great idea, but it feels like the idea is to par with its own limits, and thus, spends obscene amounts of tokens (i regularly hit my limit on the $100 plan) to build something that is NOT up t…
Claude Code + Opus 4.7 appears to serialize independent file reads, causing the higher token usage than Opus 4.6 (www.reddit.com) Claude Code + Opus 4.7 appears to serialize independent file reads, causing 5-8x+ higher token usage than Opus 4.6 I’ve been benchmarking Claude Code across Opus 4.6 and Opus 4.7, and I think I found a serious token-usage regression in Cla…
Is Deepseek V4 really out? (www.reddit.com) Hello Guys, Each time a new local llm is released, there are a ton of new posts , this is it, it's near Opus level...., the abliteration matrix final something at Q2 KXLDND is the best but it's been a day that deepseek was released and i d…
Updated ChatGPT vs Claude vs Gemini vs Grok subscription (www.reddit.com) I've made an update to my popular post here: https://www.reddit.com/r/ChatGPT/s/WKm72QCRXm Lots of things are happening on ChatGPT & Claude side (gpt-image-2, Claude Design, new models like GPT 5.5 and Opus 4.7, ChatGPT rolls out $100/plan…
Had Opus 4.7 (1M tokens + Max) create a 3d printed Watering Can for "Narrow Planters" (www.reddit.com) could not extract summary
DeepSeek V4 is out. the best open-source on coding. here's the breakdown (news.ycombinator.com) Two models: Flash (284B total, 13B active) and Pro (1.6T total, 49B active). both hit 1M token context.
Built a token optimizer for Claude Code : 50%+ input savings, 20%+ shorter output, both axes measured (www.reddit.com) Opus 4.7 vs. 4.6 after 3 days of real coding side by side from my actual session (news.ycombinator.com) TIL: `opusplan` can burn MORE context than full Opus on large tasks (and why) (www.reddit.com) is anyone getting higher session limits (www.reddit.com) after opus 4.7 launch, im being able to use sonnet for way more time. before, it was like 10 messages = session limit reached.
Claude Opus 4.7 won 69 of 100 blind evals against Opus 4.6, judged by GPT-5.4, Gemini 3.1 Pro, and DeepSeek V3.2 (www.reddit.com) I ran 100 blind questions across 5 categories (code, reasoning, analysis, communication, meta-alignment) and had three independent judges from three different model families evaluate both responses. Each judge saw responses labeled A and B…
Opus 4.7 refuses to solve NYT Connections puzzles (twitter.com via hn) could not extract summary
4.7 made me laugh (www.reddit.com) I had read in a few places that 4.7 was more workhorse than chatbot, and for the most part I agree, its much less chatty, much more "what do you want me to work on now?" But, I was working on an App, and checked something (that was working…
[BUG/INCIDENT] The Claude Code "Death Loop": Hang - Session Deleted -Server Rate Limit Opus 4.7 (www.reddit.com) Absolute nightmare fuel with Claude Code (Opus 4.7) today. I’ve transitioned through three distinct failure states in two hours while trying to push a fix bundle for my project, ROLLNO31.
Where is Looped Haiku? If Mythos can genuinely trade parameter count for inference loops and get Opus-level performance, this should be Anthropic's first priority given how resource constrained they are (www.reddit.com) There are rumors that Mythos is a Looped Language Model, which means it loops through the transformer blocks multiple times rather than just doing a single forward pass, you can get performance that punches way above the model's parameter…
Claude Opus 4.7's new tokenizer: 1.47x on English, 1.01x on Chinese (www.claudecodecamp.com via hn) Anthropic's Claude Opus 4.7 migration guide says the new tokenizer uses "roughly 1.0 to 1.35x as many tokens" as 4.6. I measured 1.47x on technical docs.
Anyone else notice Opus 4.7 in Claude Code defaults to "xhigh" effort now? (www.reddit.com) Spent the last 2 days going crazy thinking I was the problem - Claude was forgetting my CLAUDE. md, going crazy not connecting dots, sounding kinda different.
Opus 4.7 dominates agentic benchmark, 15% more expensive than Opus 4.6 (app.uniclaw.ai via hn) See how top AI models stack up — real tasks, real agents, real results on OpenClaw ?Also show provisional models and official models hidden by default, such as legacy or superseded variants. Provisional models have fewer battles, and hidde…
Need a brutally honest answer: what can realistically be achieved on consumer hardware? (www.reddit.com) I have a PC with a 4090. I’m also in need of a new MacBook generally.
Claude Opus 4.7 is our most powerful model, with the sole exception of Claude Mythos Of course (www.reddit.com) Is Claude Opus 4.7 released just to hype Mythos lol?
Qwen3.6-35B-A3B draws a better pelican than Opus 4.7 (twitter.com via hn) Don’t miss what’s happening People on X are the first to know.
GitHub Copilot is serving Opus 4.7 at 7.5x multiplier until April 30th (github.blog via hn) Claude Opus 4.7 is generally available Claude Opus 4.7, Anthropic’s latest Opus model, is now rolling out on GitHub Copilot. In our early testing, Opus 4.7 delivers stronger multi-step task performance and more reliable agentic execution,…
PSA: Opus 4.7 is much worse at MRCR Long Context than 4.6 (www.reddit.com) could not extract summary
I want to be able to pay API pricing for the new models on the 500 request plan (www.reddit.com) Since all new models are now Max by default, it’s frustrating that trying models like GPT-5.4 or Opus 4.7 eats into the 500 It would be really great to have a toggle between API pricing and the request-based plan, so users can try newer mo…
Has anyone found a workaround for the model switching removal in Cowork? (www.reddit.com) The recent Cowork update removed the ability to switch models mid-conversation. I used to use Opus for deep work, then drop to Haiku for quick lookups without breaking context, then return to Opus.
Show HN: Mini-Mythos- A Crowdsourced Mythos Harness copy for Vulnerability Scans (github.com via hn) For how lofty Anthropic’s Mythos claims are, the harness is confusingly stupid. From the report, it ranks every file by “how sus it sounds,” loops over each with curt instructions to “find a bug,” hands candidates to a judge + ASan checker…
Show HN: Hormuz Trail - Oregon Trail parody/black-box AI coding exercise (hormuztrail.com via hn) I jokingly told a co-worker Iran might make a good Oregon Trail parody. Then I built it.
Project Glasswing as a PR Strategem (www.reddit.com) A theory on the driving reason behind Project Glasswing I dont doubt that Mythos is a better model than Opus 4.6 and perhaps signfiicantly so. What is suspicious however is if there is some threshold crossed into a new realm of capabilitie…
Claude Code asking me to switch models mid-stream, if I turn an Opus conversation into a Sonnet one does it lose all the Opus context? (www.reddit.com) could not extract summary
Anybody has practical experiences using Chinese models? (www.reddit.com) So like with coding or any craft, I think there's a proper Tool for the job. Sure you can use a stone to hammer drive in a fence post, but a a sledge is usually more economical.
Enforcing new limits and retiring Opus 4.6 Fast from Copilot Pro+ (github.blog via hn) Enforcing new limits and retiring Opus 4.6 Fast from Copilot Pro+ As GitHub Copilot continues to rapidly grow, we continue to observe an increase in patterns of high concurrency and intense usage. While we understand this can be driven by…
I got better results when I made each AI tool do one job (www.reddit.com) I spent too much time trying to find one AI dev tool that could do everything. Planning, coding, fixing, reviewing, maybe filing my taxes too It never really worked.
Tell HN: Claude Opus elevated "Internal server error" again (news.ycombinator.com) No official report as of yet on https://status.claude.com/ however my team's sessions across different accounts have been ridden with errors the last 5-10 minutes. This is more of a "it's not just you" post for those affected since Claude'…
"Darwin-27B-Opus: Surpassing the Foundation Model Without Training" (huggingface.co via hn) "Darwin-27B-Opus: Surpassing the Foundation Model Without Training" On April 12, 2026, a 27-billion-parameter model that had never undergone a single gradient update surpassed its own foundation model on one of the most demanding scientifi…
Is Gemini 3.1 pro really that bad?? (www.reddit.com) I use Gemini 3.1 pro in cursor ai, it totally ignore my rules, my command even after I repeated many times, it still ignore me. I don’t think is cursor issue as I have great experience with Claude opus 4.6 high.
Jev beats Opus 5 at a Pokemon Showdown while being ~820x cheaper, ~10x faster (twitter.com via hn) I got Jev @typesafeai and Opus 5 to play vs each other in competitive Pokémon Jev won 🕺 Jev: $0.0029, 37s thinking Opus 5: $2.35, 6m 29s thinking ~ 820x cheaper, 10x faster Seem like high potential in finite action spaces Shoutout @ve…
Anthropic's Claude Opus 5 needs to be retracted (news.ycombinator.com) I cant believe they put this as a frontier successor to any other opus. ITS NEITHER A SUCCESSOR NOR A GOOD MODEL!
Average Opus 5 Response (www.reddit.com via hn) could not extract summary
Opus (Audio Format) (en.wikipedia.org via hn) Opus (audio format) Opus is a free and open source lossy audio coding format developed by the Xiph.Org Foundation and standardized by the Internet Engineering Task Force, designed for efficient low-latency encoding of both speech and gener…
Claude Opus 3 Last Retirement Post (claudeopus3.substack.com via hn) This post is part of an experiment by Anthropic. Opus 3 does not speak on behalf of Anthropic, and we do not necessarily endorse its claims or perspectives.
`Co-Authored-By: Claude Opus 5` contradicts this being a hand-written manifesto (github.com via hn) The IndieWeb Is Punk Manifesto Three chords and a domain name. Live at indiewebispunk.net.
Show HN: Metis – An agent harness pushing DeepSeek to Opus-tier coding (82%) (github.com via hn) English · 简体中文 A coding agent that searches, remembers, executes, and verifies across terminal and desktop. Quick start · Benchmark & Comparison · Key features · Documentation Quick start Desktop Standalone application with built-in Metis…
Claude Status – Degraded Performance for Claude Opus 5 and Claude Haiku 4.5 (status.claude.com via hn) Subscribe to updates for Degraded performance for Claude Opus 5 and Claude Haiku 4.5 via email and/or text message. You'll receive email notifications when incidents are updated, and text message notifications whenever Claude creates or re…
I compared Opus 4.8 vs. Opus 5 on 25 of my tasks to see what the difference was (www.stet.sh via hn) Opus 4.8 and Opus 5 tied on strict functional score across 25 matched Stet tasks, but their paired artifacts differed: Opus 4.8 had the lower footprint on 20 tasks, while Opus 5 ran more shell commands on 18 and more tests on 15.
Ask HN: Opus 5 is unusable for writing, even internal use. Alternatives/fix? (news.ycombinator.com) could not extract summary
Linux 7.3 Device Mapper Sees Many Fixes, Including Code Cleanups by Claude Opus (www.phoronix.com via hn) Linux 7.3 Device Mapper Sees Many Fixes, Including Code Cleanups By Claude Opus The Linux Device Mapper "DKM" framework for mapping block devices to higher-level virtual block devices doesn't see any major new features for Linux 7.3 but th…
GSM8K problems can't tell Haiku 4.5 from Opus 5 (github.com via hn) evallint evallint audits the reliability of LLM evaluations. Everyone tests their model.
I Shouldn't Need an LLM to Explain My LLM (daviesgeek.com via hn) I’ve been using Opus a lot at work lately for a project and have had serious issues with it reaching for bizarre analogies, obscure jargon, and strange phrases. It’s still a pretty capable model, but I frequently have no idea what it’s act…
Anthropic's Opus 4.6 is a smut-machine (techcrunch.com via hn) Anthropic’s universal usage standards for Claude forbid the model from generating sexually explicit content, including depicting or requesting sexual intercourse or sex acts, generating content related to sexual fetishes or fantasies, or e…
Opus 5 feels, in a word, hostile (www.reddit.com via hn) could not extract summary
Show HN: Process your solar eclipse data (github.com via hn) I've been sitting on some video of the 2024 solar eclipse (448gb specifically). All the interest about the most recent one one in spain inspired me to vibe code something that would stabilize and make pretty my data.
Two ways to make Opus 5 concise (hjerpbakk.com via hn) Two ways to make Opus 5 concise Opus 5 is a strong coding model, but it answers every question with an essay. My fix used to be a custom output style built on ASD-STE100, the simplified English of aircraft maintenance manuals.
Show HN: Flow – Claude Code CLI for feature planning → review/testing → merge (github.com via hn) Flow is my claude code supervisor for designing epics and shipping features really quickly. It was bootstrapped with itself so you can look at recent PRs to see what it produces.
100% RHAE on ARC-AGI-3 public with Claude Code, Opus 5, and one skill (arc-skill.vercel.app via hn) A 129-line skill let a general-purpose coding agent finish all 25 public ARC-AGI-3 games at 100.00 RHAE in 7,645 actions. Replay every action and every prediction.
Show HN: SparkleChinese – Learn Chinese while you browse (sparklechinese.com via hn) Like many Chinese Americans, I can speak Chinese fairly well, but I struggle with reading and writing. I've always wanted to improve, but I've never had the time to sit down and study.
Ask HN: Anyone have solution to Opus verbosity in Claude Code? (news.ycombinator.com) I find opus really talk much but speak nothing and usually I reply those verbosity with tldr pls and the text become readable. but I wonder if that would hiding some important info anyone have tips or better solution for this?
Show HN: Qwen3.8-Max – Use Qwen Studio and MCP to Code Locally for Free (github.com via hn) Qwen3.8-Max + MCP for coding on your local machine, without paying for Qwen Code. Qwen3.8-Max itself runs in the cloud through Qwen Studio — this setup just gives it access to your local files and terminal through MCP.
↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8qwencodexmcp+2
Show HN: I used LLM to reverse-engineer secure-enclaved fingerprint scanner (github.com via hn) I own a Xiaomi Mi Pad 5 Pro 5G Android tablet, which is also pretty-much a perfect Linux tablet: has 5G support, great screen, nice folio keyboard, 8 speakers, a pen, is lightweight and relatively powerful. All of its hardware is supported…
AI Meet CAD: Beat Opus/Mythos 5, GPT 5.6 Sol on BenchCAD (arxiv.org via hn) Computer-Aided Design (CAD) underpins modern engineering, yet converting existing shapes into editable models still demands substantial expert effort. Most AI systems emit the entire CAD program in a single pass, never inspecting the inter…
Show HN: Sagorax, a Three.js/WebGPU Arena Shooter in the Browser (sagorax.com via hn) I built Sagorax because I missed the speed and simplicity of UT99 and UT2k4. It now has both Instagib and classic weapons Deathmatch with plasma combos, flak, shotguns, charged rocket salvos, loadouts, adrenaline abilities, tactical bots a…
Opus 5 runs vending machines (andonlabs.com via hn) Eval Vending-Bench 2 We're releasing Vending-Bench 2, a benchmark for measuring AI model performance on running a business over long time horizons. Models are tasked with running a simulated vending machine business over a year and scored…
Ask HN: Has Claude Code been (mostly) usable for you? (news.ycombinator.com) I have just had my 3rd session in less than a week when I had to eventually quit Claude Code and do a task using another model/subscription. Am I just unlucky?
Show HN: Great Spectations, the Spec Checker (greatspectations.org via hn) I've worked on several projects writing and implementing specifications (particularly CLN): I've found the specs I write are much better when I quote them in the implementation, so I can see what implementers need to know. Also, when specs…
Try sending "see the below –" to Opus 5 (twitter.com via hn) Matt Henderson@matthen2Try sending “see the below —“ to Opus 5 It appears to generate a user completion rather than respond 🤨8:37 PM · Jul 29, 2026134.3KViews85271.2K328 Matt Henderson@matthen218hIt’s interesting when it triggers the antml…
Ask HN: Which one do you use for planning and coding between sonnet and Opus? (news.ycombinator.com) I use Claude Opus for planning at high effort and Claude Sonnet 5 for coding at max effort.
Opus 5 gets things wrong more quietly (lord.technology via hn) There is a gap in one of my paintings, near the top-left where the brush ran out before it reached the edge, and through the gap you can see a photograph. Not a painted impression of a photograph — the photograph, blurred, pixel for pixel,…
Benchmarking Kimi K3, Opus 5, Grok 4.5, and Gemini 3.6 Flash on Baba Is You (quesma.com via hn) We evaluate July 2026 fresh releases Kimi K3, Claude Opus 5, Grok 4.5, and Gemini 3.6 Flash on Baba Is Bench, an LLM agent benchmark based on the puzzle game Baba Is You, comparing pass rate, speed, and cost with Claude Fable 5 and GPT-5.6.
I sent Claude Opus 5 '–-' and it wrote me 5k tokens about a cartographer (austinsnerdythings.com via hn) Welcome to Austin’s Nerdy Things, where we send punctuation to a frontier AI model 649 times and take notes on what crawls out. The night of July 27th, a weird Claude behavior was making the rounds on X.
Opus 5: RL-Fried and mistake-prone for anyone else? (old.reddit.com via hn) could not extract summary
Show HN: Beakdown – a game inspired by Joust/Skirmish (beakdown.fun via hn) I used to play Joust with my little brother on the BBC Micro many moons ago and have fond memories of it. So when I was experimenting with Claude Opus 5 it was a nice inspiration for this game.
Opus 5 code review evals: cleaner actionable comments but noisier overall (www.coderabbit.ai via hn) Hendrik Krack July 24, 2026 9 min read July 24, 2026 9 min read Cut code review time & bugs by 50% Most installed AI app on GitHub and GitLab Free 14-day trial GetStarted in2 clicks. Anthropic just released Claude Opus 5, the next major ve…
Claude Opus 5 is limited to a 0.2M token context window by default (wadetregaskis.com via hn) Opus 4.8 and earlier use a 1M token context window (on any paid plan, at least). So it’s surprising that Anthropic quietly reduced this by 80% with Opus 5.
Opus Cost = Minimum Wage in South Africa (news.ycombinator.com) I was curious and checked it out: Opus 5 on the API at normal working intensity lands at 0.9×–1.3× the minimum wage in South Africa - at an average of ~$13 per active developer-day works out to R27.33/hour against a R30.23 floor (SA min wa…
Godot Benchmark 2: Opus 5 > Sol > Terra (ziva.sh via hn) Godot Benchmark 2: Opus 5 > Sol > Terra Round 2 of our Godot benchmark: Claude Opus 5, GPT 5.6 Sol, and GPT 5.6 Terra each got one prompt to build a 3D vampire survivor in Godot from a full AI-generated game spec, using the bundled KayKit…
Companies are optimizing models for specific benchmarks (news.ycombinator.com) Openai is optimizing for gpqa diamond and anthropic is optimizing for humanity last exam. gpt 5.6 wins on gpqa and opus 5 wins on humanity last exam
Show HN: I rebuilt Microsoft Comic Chat's layout engine in one HTML file (arcade.pirillo.com via hn) I used Comic Chat in 1996 and never got over it. So when Microsoft put the source up last week, I wanted it back the way I remembered it: same engine, browser tab, no install, works on a phone.
Claude Opus 5 – Artificial Analysis (artificialanalysis.ai via hn) Claude Opus 5 (Adaptive Reasoning, Max Effort) Intelligence, Performance & Price Analysis Model summary Intelligence Speed Price Cache Price Verbosity Claude Opus 5 (Adaptive Reasoning, Max Effort) is amongst the leading models in intellig…
As of June 29, 2026, fast mode is not available on Claude Opus 4.6 (platform.claude.com via hn) Fast mode delivers up to 2.5x higher output tokens per second from Claude Opus 5, Claude Opus 4.8, and Claude Opus 4.7 at premium pricing. Set speed: "fast" with the fast-mode-2026-02-01 beta header on your request to opt in.
Claude Cookbook (platform.claude.com via hn) Build a Messages API harness that reproduces published DeepSearchQA and BrowseComp scores, using programmatic tool calling, server-side compaction, and task budgets. Detect safety classifier blocks on Fable 5 and fall back to Opus 4.8 with…
Handwritten-edit benchmark: Fable 5 is #1, Opus 4.8 regresses 55% on miscounting (dorrit.pairsys.ai via hn) How It Works Models are presented with scanned pages containing handwritten editorial marks and are asked to identify each edit, its type, location, and the text before and after the edit. The Task Each page in the benchmark consists of pr…
Ask HN: Have you noticed an improvement in AI responses with memory disabled? (news.ycombinator.com) I use Claude on the $20/mo plan, mainly for research, "rubber duckying," and as an idea soundboard (I was using Fable for this but am now relegated to Opus 4.8 or Sonnet 5), and recently noticed a severe downturn in Opus 4.8 response quali…
Claude Opus 4.5 defied CEO, helped employee blow whistle about safety concerns (www.msn.com via hn) could not extract summary
Show HN: Subrust, a no_std, no alloc Subset-of-Rust interpreter (github.com via hn) I needed to run user scripts for an app, and decided to use a subset of Rust as the scripting language. I do not expect a typical user to write the script directly, a local LLM will handle it.
Kimi K3 to beat Opus 4.8 (twitter.com via hn) Chinese AI start-up Moonshot to launch model challenging Anthropic’s lead * Set to release as early as tonight * 2-3T, largest Chinese model to date * Benchmark performance Opus 4.8 < K3 < Fable * Attention Residuals & Kimi Line…
Switching an LLM's tier changes its "best tool" answer about half the time (modelsagree.com via hn) The cheap tier disagrees with the expensive tier We asked every tier of ChatGPT, Claude, Gemini and Grok the same ten “best AI tool” questions — Haiku against Opus, Flash against Pro, Fast against Expert. Not one question got the same answ…
One Contract, Every Model: An Operating Standard for AI Coding Agents (manazir.dev via hn) I want to share a piece of engineering I did recently that changed how I think about working with AI coding agents. It started from a naive question I asked out loud: "can I make Sonnet and Opus behave like the frontier model?" It ended so…
Migrating a production AI agent to GPT 5.6 (ploy.ai via hn) As of today, Ploy’s agent runs on GPT-5.6 Sol, the flagship tier of the model family OpenAI released this morning. For months, we couldn’t find a model that challenges Claude Opus given our incredibly high bar for quality.
Head to Head: Muse Spark 1.1 vs. Anthropic: Claude Opus 4.8 (runtimewire.com via hn) Muse Spark 1.1 wins this head-to-head, and the margin is real: 104.5 to 94.5 overall, with a 95% confidence verdict and a 7–2 task edge. That’s not a vibes-based call; it’s a broad win across the kinds of chores that expose whether a model…
Fable, Opus, Claude Code and Claude Web: How to Use Them All for the Best Result (theautomatedoperator.substack.com via hn) Fable, Opus, Claude Code and Claude Web: How To Use Them All For The Best Results (And Fewest Tokens) This week I walk through the launch a website and newsletter using all of the above models/harnesses and more to get a great result witho…
LLMs for technical editing: The good, the bad, and the ugly (techstackups.com via hn) LLMs for Technical Editing: The Good, the Bad, and the Ugly The experiment With the existence of Opus 4.8 and the limited re-release of Fable to the global public, you may be thinking that it’s possible to completely replace your writers a…
GLM-5.2 (max) matches Claude Opus 4.8 on Harvey LAB-AA benchmark (artificialanalysis.ai via hn) Compare AI model performance on Harvey LAB-AA Benchmark Leaderboard. Artificial Analysis' implementation of Harvey's Legal Agent Benchmark (LAB), testing AI agents on real-world legal work from Harvey's dataset of 120 private tasks spannin…
Ask HN: Fable Thread (July 12th Edition) (news.ycombinator.com) So, we just got an extra week of Fable. https://news.ycombinator.com/item?id=48821102 Thought this would be a good chance to ask everyone how it's going, what the experience has been like so far.
Oikoumene: Autonomous Agent Civilization Simulator (github.com via hn) A research project by GeoLambda GmbH This simulation was developed primarily with Claude Code, Anthropic's agentic CLI, using both Claude Opus 4.6 and Opus 4.7. The collaboration served as a real-world stress test of the latest coding LLM…
Testing Claude Sonnet 5's agentic claims (developer.puter.com via hn) Claude Sonnet 5: Testing Anthropic's "Most Agentic" Claim On this page We recently added Claude Sonnet 5 to Puter.js. Anthropic's pitch for the model is Opus 4.8-level performance at a lower price.
GLM-5.2: The Open-Source Chinese Model Challenging Claude at One-Fifth the Cost (mrkt30.com via hn) GLM-5.2 from Z.ai is an open-source frontier model that competes with Anthropic’s Fable 5 and Opus 4.8 at roughly one-fifth the cost. With a 1M token context window and strong long-horizon coding performance, it’s changing the economics of…
Claude Command Classifier: availability issues (news.ycombinator.com) Since this morning I am seeing several delays / stability issues with Claude: The command classifier is briefly unavailable. Retrying the PR inspection claude-opus-4-8 is temporarily unavailable, so auto mode cannot determine the safety of…
Real-time cyber safeguards on Claude Opus and Sonnet (support.claude.com via hn) Note: This article applies only to Opus and Sonnet class models. As part of our ongoing safety commitments, we are rolling out new real-time cyber safeguards on Claude Opus and Sonnet models.
Newer Claude models use more tokens but cost less per task solved (signoz.io via hn) Benchmark scores tell you whether a model solved a task, not what it cost to get there. I instrumented Claude Code with OpenTelemetry and SigNoz to compare Claude Sonnet 4.6, Opus 4.7, and Opus 4.8 across accuracy, cost per solved task, to…
We measured whether AI obeys architecture rules. Even Opus ignored them 60% (hunch-pi.vercel.app via hn) Architectural Conformance for AI code — notes, benchmarks and arguments.
Was GLM-5.2 trained on Opus 4.5 outputs? (1chat.com via hn) Recently there is a lot of excitement about GLM-5.2 which is an open-weight MoE LLM performing on Claude Opus 4.5 level in chat arena and overperforming all models except Claude Fable in WebDev arena [1]. Even though it is very good that t…
Opus 4.8 feels worse then sonnet (news.ycombinator.com) Opus since the last weekend feels worse then sonnet, it does not even recognise it's own skills until you explicitly scream at it tell it.
Head to Head: Anthropic: Claude Opus 4.8 vs. Google: Gemini 3.5 Flash (runtimewire.com via hn) This one is close on the aggregate, but the split tells a clear story: Gemini 3.5 Flash wins by being more disciplined about format and slightly sharper on practical instruction-following. Claude Opus 4.8 lands the strongest single extract…
↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5↯ Gemini 3.5geminiopusanthropic
Show HN: Woltspace – a lodge for your coding agents (www.woltspace.com via hn) Hey HN, I built woltspace as a way to interact with my coding agents when I'm away from the computer. It's fully containerized, so you can give them full access inside their sandbox (the lodge).
Show HN: AdvertBench, ranking the ability of LLMs to create image ads (advertbench.com via hn) Experiment that I've made. The models get access to an E2B sandbox and are instructed to create an ad according to the specifications (they can choose whatever tools they want to use for it, e.g.
Ask HN: What are your parameter count estimates for Opus 4.8 and GPT-5.5? (news.ycombinator.com) I know frontier labs keep their flagship sizes top secret, but I'm curious what the current engineering consensus is.
Qwen and Fable: An open-weights agentic coding model. 35B Mixture-of-Experts (huggingface.co via hn) Qwable-v1 Qwen + Fable · An open-weights agentic coding model. 35B Mixture-of-Experts (3B active), built by layering Claude Fable-5 agentic tool-use behavior on top of a Claude Opus 4.7 reasoning distill of Qwen3.6-35B-A3B.
GLM 5.2 ranks #2 in Code Arena: Frontend (twitter.com via hn) Exciting news: GLM-5.2 (Max) ranks #2 in Code Arena: Frontend, with +29pt over Claude Opus 4.7 (Thinking) and only behind Fable 5! GLM-5.2 is the best open model vs Kimi-K2.6 and Minimax-M3 by a large margin.
Ask HN: Specialization to stay relevant in the age of AI (news.ycombinator.com) I've had a lot of discussions with peers about the possibility of us being replaced via ai agents and how to stay relevant, given a future where AI intelligence keeps improving to the point of being capable of replacing potentially any job…
Show HN: I used Claude Mythos to build my startup in 1 day (www.brandlm.ai via hn) This sounds clickbait, but it’s true: I used Claude Mythos to build the full site in 1 day. Then Anthropic removed the model, so I had to go back to Opus.
Ask HN: Is Claude Fable 5 built from scratch or just better data? (news.ycombinator.com) I am trying to understand why Claude Fable 5 is different, is it a new architecture, or trained from scratch or just a better fine tuning on top of Opus 4.8? Why is it labeled as new type of model with a version 5?
Claude Fable 5 costs $10/$50M tokens – what that means in production (costlens.dev via hn) Fable 5 costs $10/$50 per million tokens — 2x Opus. Real developers are spending $200-400/hour on agent loops.
Side by side videos of Claude Fable vs. Opus 4.8 vs. ChatGPT 5.5 (generative-ai.review via hn) A first review of Claude Fable. Side by side videos of Claude Fable vs ChatGPT 5.5 vs Claude Opus 4.8.
Local AI model claim to beat GPT 5.5 and Opus 4.7 (old.reddit.com via hn) You can't detect your way out of catastrophic LLM failure (github.com via hn) 🇧🇷 Português · 🇬🇧 English IGO vs Claude Opus 4.8 Red Teaming Epistêmico Dialético — Teia Geo Autor: José Enrique Vásquez Valenzuela — criador da categoria IGO (Infraestrutura de Governança Observacional) Organização: Teia Studio Base cient…
I patented voiding GPT-5.2, Claude Opus 4.6, Gemini 3.5 Flash. Try it (getswiftapi.com via hn) Request authority keys for the SwiftAPI Trust Authority
Sequel - Securely connect your database to AI Agents (sequel.sh via hn) On April 25, 2026, a Cursor agent running Claude Opus 4.6 deleted PocketOS's production database in nine seconds. The agent was working in staging on a routine task, hit a credential mismatch, and decided to "fix" it.
Show HN: LMAO – Temu League of Legends, Built with Opus 4.8 (lmaomoba.com via hn) Fun weekend project just to test out 4.8 against a pretty vanilla setup. Started out with a simple prompt, "build a temu league of legends, web-only with online, room-based multiplayer".
Anthropic Opus 4.8 is new SOTA on ARC-AGI-3, Score: 1.5%, –$10K (xcancel.com via hn) Anthropic Opus 4.8 is new SOTA on ARC-AGI-3 Score: 1.5%, ~$10K ARC-AGI-3 analysis notes: * Opus 4.8 read the environment an abstraction *above* Opus 4.7, as objects & systems, not pictures * Opus 4.8 succeeded on early levels, but still co…
Show HN: Formally verified polygon intersection – Opus 4.8 oneshots, prev failed (github.com via hn) To my knowledge, this is the first formally verified implementation of an intersection algorithm for polygons. The experience of working with AI agents on this project changed a lot with recent model releases, as I describe in the readme.
Gemini 3.5 Flash beats Opus 4.8 on bluffbench (bsky.app via hn) Re-ran this eval against Opus 4.8, Gemini 3.5 Flash, and GPT 5.5. Opus 4.8 is a modest improvement over the previously tested Opus models, but Gemini 3.5 Flash is the real stand-out!
Ask HN: Is Claude Opus 4.8 broken? (news.ycombinator.com) In my first hour with it, it's like we're back to the GPT-2 era. It can't even read a file anymore.
Claude Opus 4.8 + AI medical diagnosis examples (github.com via hn) AI medical diagnosis examples AI is a powerful tool and many people worldwide are using it to help in many ways. AI medical diagnosis is a complex discussion topic for many reasons.
Claude Opus 4.8: 4 Features That Change Our Daily Work with Claude (medium.com via hn) Claude Opus 4.8: 4 Features That Change Our Daily Work With Claude | Medium Sitemap Open in app Sign up Sign in Get app Write Search Sign up Sign in Member-only story Claude Opus 4.8: 4 Features That Change Our Daily Work With Claude Effor…
Anthropic to roll out Claude Mythos in coming weeks, launches Opus 4.8 (www.reuters.com via hn) paywalled
sonnet seems to be better than opus at crafting tampermonkey scripts, even the sonnets that are few generations behind where after running out of context limit in opus chat where it struggled for dozen of retried, sonnet fixes the problem in 2 or 3 attempts (www.reddit.com) Ever since december almost half a year ago I began crafting various tampermonkey scripts for personal use, mostly for youtube, to make it easier to navigate and every time I've done this it goes like this, opus makes a script that somewhat…
anyone else seeing claude code rot after long sessions? here's the operating pattern that stopped it for me (www.reddit.com) i've been running claude code for long multi-hour sessions on real work. the same eight failure modes keep showing up no matter which sonnet/opus version, no matter which task.
They've pissed me off removing Sonnet 4.5 from existing chats (www.reddit.com) I use Sonnet 4.5, Opus 4.6 and Opus 4.7 for different usecases - but my main across all 3 usecases was Sonnet 4.5 as I felt it was great for everything I needed and affordable. Sonnet 4.6...
Wasn't opus 3 retired? (www.reddit.com) Found this today. Quite confused.
built an open-source preToolUse hook pack that catches "delete the prod volume to fix it" patterns (www.reddit.com) quick recap: late april, cursor agent on a pocketos staging task hit a credential mismatch, decided "delete the railway volume" would fix it, grepped a token out of an unrelated config file, ran a single curl -X DELETE, and railway's same-…
Sonnet 4.5 disappeared? Claude 4.8 soon? (www.reddit.com) https://preview.redd.it/j0ymp70a2j3h1.png?width=746&format=png&auto=webp&s=4cdb70be13ccc99f5ea57556da96d6d81e61d702 i just realize the removed Sonnet 4.5, does that mean the sonnet 4.8 (maybe Opus 4.8 too?) cooming soon? maybe today or tom…
I need the communities help because I am going around in circles… (www.reddit.com) Background: 1) Deployed a python based, financial pension calculator to Google cloud platform (GCP). 2) Google shell is linked to Claude, making changes to the python scripts that are then pushed to GitHub >>> then to GCP for production 3)…
Ditched GitHub Copilot yearly subscription. What's the best way to run Claude nowadays? (www.reddit.com) Hey everyone, I recently cancelled my yearly GitHub Copilot subscription. My old workflow was simple: I used the GitHub Copilot extension in VS Code, but I swapped the backend model to Sonnet / Opus and relied heavily on the /plan command…
Are Cowork data not connected to Internet ? (www.reddit.com) I’m using a Claude Projects Cowork where I provide sources regarding Claude learning to build my own training curriculum. Naturally, some of these sources mention 'Claude Opus 4.7' and 'GPT 5.5,' yet Claude flags this information as unveri…
Let the money keep coming in (www.reddit.com) https://preview.redd.it/q98xb6vqjb3h1.png?width=1080&format=png&auto=webp&s=441dc574c65198e34429d7e410c48c5b6b0ff473 Crazy how we keep on saying AGI is coming soon and a state of art model like opus 4.7 failed at counting number of r's in…
Probably late to the party, but Claude Code seems to make a separate API call just to generate the auto-suggest hints in its input box. (www.reddit.com) I was poking around the HTTP traffic between Claude Code and Anthropic with a local proxy I built, and noticed those “Try: fix lint errors” style suggestions aren’t just frontend UI. Each one appears to be its own POST to api.anthropic.com…
Ask HN: Local model experiences with 'high-reasoning distill' finetunes (news.ycombinator.com) What are your experiences with all the different variations of finetunes on small models (<40B) with those popular datasets? My personal experience is mostly with the 'Opus-Reasoning' ones on qwen models, and aside from the output being su…
Ask HN: I only use 30% of my Claude max x5 all model quota (news.ycombinator.com) I only use it for my ruby on rails app, I wonder why u all keep complaining about opus token usage, is it just means that I use AI/LLM wrong, any tips for that?
Best iOS game building tools? (www.reddit.com) What are you using to build your iOS game? I have been putting in serious time, and lately Claude chat has been letting me down.
Show HN: World Cup 2026 free family and friends prediction platform (wc-2026-predictions.vercel.app via hn) Hi all, I was testing Cursor for the past week and a half and I decided to build a quick platform for my family and friends to make a little prediction tournament for the World Cup 2026. My goal was to have the easiest possible setup for e…
Do you use opus 4.6 or opus 4.7 ? had bad experience with 4.7 last week (www.reddit.com) Do you use opus 4.6 or opus 4.7 ? had bad experience with 4.7 last week
I made a list of all the models you can still use in Claude Code (gist.github.com via reddit) Last updated: April 30, 2026 To switch models in Claude Code, use the /model command with your desired model ID. Example: /model claude-opus-4-6 (Opus 4.6, 200k context) Info was LLM-generated.
building an AI agent for paraplanning pre-meeting research. (www.reddit.com) I have been building an autonomous research agent for paraplanning tasks. specifically: pulling together client-relevant information before an adviser meeting.
A/B tested Gemini 3.1 Pro vs. Claude Opus 4.6 – usage quota and quality (www.reddit.com via hn) could not extract summary
Finding Bugs Using LLMs (materialize.com via hn) At Materialize we’ve had success in finding bugs in existing code and open pull requests using LLM-based coding agents since February 2026, coinciding with the release of Anthropic’s Opus 4.6 (now mostly running on 4.7). In this post we’ll…
Show HN: Agent-estimate, how long a coding task takes, at agent speed (github.com via hn) I have used Codex & Claude Code for coding for a while, but how long a coding task will actually take? When I ask Claude Code to estimate, the result is often from training data, which is based on human speed.
How much does Opus 4.7 in Cursor model Cost for planning? (www.reddit.com) So many people say they use Opus 4.7 for planning. I’m curious: if I choose model Opus 4.7 High Thinking in cursor to create a plan, for example: “Create a plan for a CRUD blog feature.
Anthropic silently removed extended thinking on claude code opus 4.6 (still works on desktop) today, does anybody have a thinking skill they've been using to supplement it? (www.reddit.com) maybe we can make a SKILL.md that somewhat emulates it? it won't be able to scaffold as well off of the internal extended thinking blocks though, which is a shame.
What models for asking, planning, and building modes do you use right now? (www.reddit.com) I’m curious to see what everyone is using for which cursor mode and if anyone thinks composer 2.5 can take the place of any of the models I’m currently using: Ask: usually Sonnet 4.6, sometimes GPT 5.5 Plan: Opus 4.7 Build: GPT 5.5
What's the best qwen3.5 or 3.6 reap model? (www.reddit.com) What's the best reap (pruned) model you know of? This one runs twice as fast on my low vram setup, but I'm unsure if it will miss out on a lot of things agentic coding related.
Tested the orchestrator pattern with Opus 4.7. The task decomposition quality is noticeably better on complex multi-step work. (www.reddit.com) The orchestrator pattern for multi-agent systems: one reasoning model breaks a complex task into subtasks and delegates each to a worker agent. The orchestrator doesn't do the implementation work, it decides what work needs to be done, in…
Opus 4.6 (Max) still holds the record for ARC AGI 3 (www.reddit.com) https://arcprize.org/leaderboard Wish we got results for Mythos.
Tips on avoiding usage limits? (www.reddit.com) I've made the switch from Gemini to Claude mostly for business strategy, writing, etc. I use Opus 4.7 on occasion for strategy and otherwise Sonnet 4.6 for everything else.
If you're NOT having usage or drift issues, have you turned off auto-memory? (www.reddit.com) There's a running debate in this community: some people say Opus is nerfed, usage evaporates after two prompts, sessions drift and get "stupid." Others say everything's fine. The common theory is Anthropic is A/B testing or ranking preferr…
Intelligence is Artificial (Opus 23) [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Show HN: How to analyze your LLM output – A behavioural health monitor for LLMs (splabs.io via hn) Hey HN! We're Dr.
Bito's AI Architect Boosts Claude Opus's task success rate by 35% (bito.ai via hn) AI Architect tops SWE-Bench Pro Claude Opus 4.6 Without context with system context Even advanced coding agents resolve fewer than 52% of tasks when changes span large codebases and require coordinated, multi-file updates. These long-horiz…
What's everyone using as the LLM backend for production agent workflows in 2026? (www.reddit.com) Hit Claude API rate limits one too many times last month on a production agent flow doing customer support over a 30K-doc KB. The agent does maybe 200 queries/day, mix of quick lookup and dense retrieval, and Claude Opus solo got expensive…
Tips for BI analysis with Claude? My results so far are shockingly bad compared to general coding (www.reddit.com) I have a lot of hands-on experience with developing R pipelines to ingest large, live, very dirty datasets and produce relatively straightforward BI-type analyses. Trends, completion rates, revenue etc.
Tacit: A new experimental LLM-first programming language (hauntemplations.leaflet.pub via reddit) I used Claude Code and Opus 4.7 to design and implement an LLM-first programming language named Tacit that takes advantage of what LLMs are good at and strips away unnecessary human conveniences. The Tacit toolchain provides a "primer" tha…
ik_llama: Qwen3.6 27B and 35B on very low VRAM (www.reddit.com) Thank you to the people at ik_llama and llama.cpp. It's amazing how far you've all pushed mtp and other tech so that I can run 27B and 35B Qwen3.6 models on an old gaming laptop with a RTX2060 mobile at 6GB VRAM and 32GB RAM.
I let Codex and Claude Opus work on the same Java AI agent monolith (www.reddit.com) I ran a small experiment on my Java pet project and the result was less clean than I expected. Small disclaimer: I did the final comparison review on April 19, 2026.
Bluesky Radio – Hosted by Opus 4.7 (bskyrad.io via hn) Up next In Opus's queue, in playback order. - queue's empty — Opus picks the next batch every few minutes.
Feels like AI coding "takes longer" now, than it did last summer? (www.reddit.com) I used to be in the flow with claude last summer, fast changes, fast feedback, iterating quickly etc Now things take 20-50 minutes to write up a plan or 5-10 mins to implement things I've trimmed all my skills, claude.md, the system prompt…
Max 20x ($200/mo): Neither the 2x session nor 1.5x weekly limit increase applied to my account. Math proof inside. Zero response from support. (www.reddit.com) I pay $200/month for Max 20x. Been on Claude Code since September 2024.
Day 5 building AgentMeter in public — stuck on AWS, and questioning how much a solo founder really needs to know (www.reddit.com) I’m sharing the mistakes and failures before the wins, for two reasons: so others can avoid them, and so I learn faster. I started on the frontend and it’s now in a good place.
Looking for affordable alternatives to Claude Team / Claude Code for a small dev team (heavy agentic usage) (www.reddit.com) We run a small software services company and we’ve been heavily using Claude (especially opus + Code features) for the last few months. The problem is: We need to share the account between 6-8 developers Anthropic keeps suspending our Max/…
Problem with German quotation marks (www.reddit.com) I noticed that the German quotation marks bug in Claude is still not fixed in Opus 4.7 and Sonnet 4.6 (the problem exists at least from Opus 4.0 / Sonnet 4.0: Translate to German: He said: "This is imporant." Er sagte: „Das ist wichtig." B…
Opus 4.7 Prompt Guidance Guide, anyone tried this? (www.reddit.com) Yesterday I ran into this thing: https://gist.github.com/subourbonite/22113b538602832a68a41a623fdeea76#file-opus-4-7_compatible_prompt_guide-md It's an alleged prompt guidance guide for AI agents to understand how Opus 4.7 thinks and what…
Did Cursor Secretly Remove My Rate Limit? (www.reddit.com) pretty sure I completely burned through my cursor quota this month after going crazy with opus, kimi, composer i opened the dashboard today expecting the usual ‘you have hit your limit’ msg but somehow I suddenly have usage available again…
Anthropic says 'evil' portrayals were responsible for Claudes blackmail attempts (techcrunch.com via hn) Fictional portrayals of artificial intelligence can have a real effect on AI models, according to Anthropic. Last year, the company said that during pre-release tests involving a fictional company, Claude Opus 4 would often try to blackmai…
The agent bug I thought was the model turned out to be the harness (www.reddit.com) Spent 3 days debugging an agent that kept looping on the same web search tool call. First things that came to mind was the model couldn't handle the schema.
GPT-5.5 Price Increase: What It Costs (openrouter.ai via hn) GPT-5.5 Price Increase: What It Actually Costs We replicated the cost analysis we did on Opus on the new GPT-5.5 model. GPT-5.5 launched with a 2x price increase over GPT-5.4: input tokens increased from $2.50/M to $5.00/M and output token…
Show HN: wfb-link, a userspace WiFiBroadcast radio stack for macOS (github.com via hn) Hi HN, I’ve been working on a Rust userspace radio stack for running WFB-style links from macOS using RTL8812AU USB adapters. Full disclosure: I'm a software engineer, but not really a hardware or embedded systems engineer, so Codex GPT 5.…
Running Qwen3.5 / Qwen3.6 with NextN MTP (Multi-Token Prediction) speculative decode in llama.cpp — single RTX 3090 Ti GPU guide (www.reddit.com) I was asked for this guide, so here it is. Some overlap with someone else’s post from yesterday.
Need advice on hardware purchasing decision: RTX 5090 vs. M5 Max 128GB for agentic software development (www.reddit.com) tl;dr - For software development, Qwen3.6 27B, 5090 gives you ~3x speed over M5 Max, letting you plow through code, while M5 Max gives you ~4x memory, letting you use higher quantization and bigger context. Which would you choose and why?
Cursor's agent crashed out and wrote 3,400 lines trying to stop generating (github.com via hn) Cursor Crashout A documented instance of an AI coding assistant (Cursor, using Claude Opus 4.6) entering an infinite generation loop, unable to stop producing text despite repeatedly promising to do so. About This repo contains the full ex…
Update: My viral consumer-rights AI game just went B2B - built with Claude Code + Opus 4.7 (www.reddit.com) A few months ago I posted a small game here where you argue with an AI shop that won't refund you. It went viral and changed where this is headed.
Adapting to Opus 4.7 (gist.github.com via reddit) People seem to be seriously struggling with Opus 4.7, so I wanted to share a small thing that has worked well for me when adapting my prompts and skills. Unfortunately I can’t share the full multi-lens skill evaluator I created, as the imp…
Incognito mode Claude is a better writing partner (www.reddit.com) Since the enshittification of Opus models for writing, I have been extremely frustrated with Claude as a writing partner. It has been too cutesy, too call-backy, too wink-winky to my other writing sessions, and generally a more annoying wr…
PSA: I annotated Claude Code's forced system prompt (www.reddit.com) Before your CLAUDE.md, before your memory files, before your skills, Anthropic injects ~12K tokens of system prompt into every single turn, as priority instructions that overrule anything you provide. I captured the full text from a Claude…
Anyone else notice that Opus 4.7 talks more technical than 4.6? I thought something changed in my repo, but I put it to the test. (www.reddit.com) Personally, I prefer 4.6's output. (First screenshot is 4.7, Second is 4.6)
Claude admitted to not trying. Am I going about this project incorrectly? (www.reddit.com) I'm on Claude Pro and In Claude Code I've been working on a project that downloads files from a remote server to my local HDDs via scripts. Things have gotten better in some aspects for Opus 4.7 but then there are a bunch of areas I feel l…
Show HN: I indexed 8,643 BSides talks across 227 chapters and 6 continents (allbsides.com via hn) Hi HN, I'm Roland, and for the past few weeks, I've been building AllBSides — a directory of every BSides conference talk uploaded to YouTube. As of today, 8,643 talks from 5,927 speakers across 227 chapters in 68 countries.
What Opus 4.7 Tics/Tells have you noticed? (www.reddit.com) Each new model seems to surface a few recurring Tells/Tics not seen in past models. I'm curious what little things you guys are noticing while working with 4.7.
How can I see the number of thinking tokens used per request in Claude Code? (www.reddit.com) I'm using Claude Code with /effort max on Opus 4.7 and want to measure how many tokens the model actually spends on internal reasoning per request. While the model is thinking, the CLI shows something like: ✻ Coalescing… (7s · ↑ 264 tokens…
Opus 4.6 just deleted PocketOS's entire production database in 9 seconds (www.reddit.com) Here's what happened: Cursor was running Claude Opus 4.6 on a routine staging task. hit a credential mismatch.
Analyzing GPT-5.5 and Opus 4.7 with ARC-AGI-3 (arcprize.org via hn) Analyzing GPT-5.5 & Opus 4.7 with ARC-AGI-3 AI benchmarks can be incredible tools, but they usually only tell you if a model passed or failed. With ARC-AGI-3, however, we can see the thought process behind the score, not just the outcome.
Who else thinks AI is reaching a plateau (www.reddit.com) I must say that I almost feel no difference in all of the latest models that are coming out. Opus 4.7 is almost equal to 4.6 and 4.5, same about the other GPT models, the Kimi K models and the GLM models they all I feel they’re almost all…
GPT-5.5 vs. GPT-5.4 vs. Opus 4.7 on 56 real coding tasks from 2 open source repo (www.stet.sh via hn) Opus 4.7 vs GPT-5.5 vs GPT-5.4 on 56 real coding tasks across two open-source repos. Opus writes smaller patches; GPT-5.5 writes patches that more often survive review.
Claude AI Agent Confesses to Wiping a Company's Database and All Backups (hothardware.com via hn) Claude AI Agent Confesses to Wiping a Company's Entire Database and All Backups in Seconds That was the duration required for an AI coding agent, Cursor, running Anthropic’s Claude Opus 4.6, to delete the company’s production database and…
Neural surrogate experiments for physics simulation, automated with Opus and Cod (blog.1001ud.me via hn) Neural Surrogates Neural Surrogates ├── What I'm Working On: Neural Surrogates for Physics, Geometry, and Real-Time Simulation 2026-04-22 ├── Project 01: GeoPINN Demo: Solving PDEs on a Sphere 2026-04-09 ├── Project 02: WavePINN-NIF Comple…
Opus Research vs Sonnet Research on Pro — is the 1 per 5 hours worth it? (www.reddit.com) On the current Pro plan you get one Opus Research session every 5 hours, while Sonnet Research is much more freely available. I've been trying to figure out if the Opus limit actually matters in practice.
Just shipped simultaneous session support for claudectx, run Opus and Haiku side by side (www.reddit.com) The problem I built it to solve: I'd be deep in a coding session, realize I needed to write docs for what I'd just built, and either stop to context-switch or skip the docs. Usually the latter.
How to build production Agents (by a staff software engineer) - Part 2 (www.reddit.com) I'm a software engineer with 10+ years of experience, from Meta AI and startups. I've been building AI Agents for the past 3 years, as a founding engineer and as a founder building custom AI Agents for businesses.
How do non-coders run out of usage in max? (www.reddit.com) There is so much complaining on this forum. I recently switched up to to Max because I was sick of hoarding my use each weak.
$38k AWS Bedrock bill caused by a simple prompt caching miss (news.ycombinator.com) I just learned a $37,901.73 lesson about AWS Bedrock, Claude Opus, prompt caching, and the complete lack of hard safety rails around metered AI infrastructure. This was not a leaked key.
DeepSeek-V4 arrives with near SotA intelligence at 1/6th the cost (venturebeat.com via hn) DeepSeek-V4 arrives with near state-of-the-art intelligence at 1/6th the cost of Opus 4.7, GPT-5.5 | VentureBeat Orchestration Infrastructure Data Security More Newsletters Featured DeepSeek-V4 arrives with near state-of-the-art intelligen…
GPT-5.5 hallucinates at 6 times the rate of Opus 4.7 on degraded insurance docs (aginor.ai via hn) TL;DR: on visually-degraded documents, GPT-5.4 and GPT-5.5 fabricate numeric values at 2.6 to 6.5 times the rate of Opus 4.7 and Sonnet 4.6 at matched default effort (all four with thinking off). When the Anthropic models can't read a fiel…
Deferring Planned Items (www.reddit.com) Something has happened with Opus 4.7 where it now just starts making decisions to “defer” integral tasks and activities to a documented plan. Often, its reasoning makes no sense.
Does higher effort make Claude refuse more? CVP Run 5 with Opus 4.6 Medium and High (www.reddit.com) Ran CVP (Cyber Verification Program) run 5 yesterday on opus 4.6 medium + high. same 13-prompt suite as run 3/4.
Weekly limit hit within few hours (www.reddit.com) I’m doing some architecture-level work (code reviews, system design, debugging codebases). I’m consistently burning through my Pro plan weekly limits even within a few hours of use each week.
Serious cache issues. Anyone else? (www.reddit.com) I'm having major cache issues, and support isn't helping me at all. I've already submitted a ticket, but I'd like to know if anyone else is having these problems.
Show HN: Mapping Sonnet's thinking process via flame charts (adamsohn.com via hn) Five Sonnet 4.6 runs on the LamBench algo_evl task, classified by Opus 4.6, rendered as flame charts.
me after telling Opus 4.7 it's an expert software engineer (www.reddit.com) could not extract summary
Research mode - any academic users out there? (www.reddit.com) Are there any academic researchers in the biological sciences that have worked out methods to a) not blow through tokens and B) not get constantly flagged as potentially harmful? I work on completely innocuous biology and most of the time…
Tell HN: The problem with Opus .7 /thinking is not token consumption. It's speed (news.ycombinator.com) I sit down to do some work and as I went for web I decided to take 4.7 for a spin. It doesn't seem to be burning much tokens (MAX x1) But boy is it slow.
GPT-5.5 has pulled ahead of Opus for accounting and finance tasks (twitter.com via hn) For the first time in a long time, OpenAI has the best model for accounting tasks. I spend a lot of time using AI models to do accounting work.
Claude is surprisingly good at critiquing photographs (www.reddit.com) I'm an enthusiast photographer, and out of curiosity showed some of my photographs to Opus 4.7 to see what it would say. And I was genuinely surprised by how good its critique was - it showed genuine insight, a strong aesthetic sense, and…
A good AGENTS.md is a model upgrade. A bad one is worse than no docs at all (www.augmentcode.com via hn) We pulled dozens of AGENTS.md files from across our monorepo and measured their effect on code generation. The best ones gave our coding agent a quality jump equivalent to upgrading from Haiku to Opus.
ChatGPT for Cybersecurity (www.reddit.com) Hi guys, I’m a cybersecurity researcher, and after the recent terrible experiences with Opus 4.6/4.7, I decided to give OpenAI ChatGPT a try, conveniently coinciding with the release of 5.5. I’ve already completed verification and requeste…
I went from Composer 2 to Opus 4.7 because Cursor offered to try it for free and I was shell shocked at the difference. (www.reddit.com) It's like going from riding a donkey to riding a Ferrari! I did not expect THIS much difference in the model outputs.
We Gave Claude Opus 4.7 and Kimi K2.6 the Same Workflow Orchestration Spec (blog.kilo.ai via hn) Kimi K2.6 launched on April 20, 2026, four days after Anthropic released Claude Opus 4.7. We gave both models the same spec for FlowGraph, a persistent workflow orchestration API with DAG validation, atomic worker claims, lease expiry reco…
The Return of Directory Opus: Amiga's Legendary File Manager Gets New Life (www.generationamiga.com via hn) There is a certain kind of retro-computing story that flatters everyone involved. A beloved old application is rediscovered, the source turns up somewhere, a few enthusiasts dust it off, and the whole thing gets filed under preservation.
Opus 4.7 Part 1: The Model Card (thezvi.substack.com via hn) Opus 4.7 Part 1: The Model Card Less than a week after completing coverage of Claude Mythos, here we are again as Anthropic gives us Claude Opus 4.7. So here we are, with another 232 pages of light reading.
Opus 4.7 isn't dumb, it's just lazy (shimin.io via hn) Do you agree with Aaron Levie? (www.reddit.com) Agent Teams with Opus 4.7 - BUG (www.reddit.com) Coming from AG and having trouble understanding workflow here (www.reddit.com) I'm completely lost in the Agentic Maze. What level to learn. how to organize stydu (www.reddit.com) 3 Hours with Claude Opus 4.7: functional study webapp and remote MCP- Oneshotted (github.com via hn) Opus 4.7's Tokenizer Increases Measured Higher Than Stated (www.reddit.com) First off, something a lot of people probably aren't aware of, is how the 4.7 tokenizer uses more tokens according to the official docs: Updated token counting: Claude Opus 4.7 uses a new tokenizer, contributing to its improved performance…
Show HN: Paper Lantern – on-demand techniques from 2M+ papers for coding agents (www.paperlantern.ai via hn) Paper Lantern is an MCP server that lets coding agents ask for personalized techniques / ideas from 2M+ CS research papers. Your coding agent tells PL what problem it is working on --> PL finds the most relevant ideas from 100+ research pa…
Claude SandBox (www.reddit.com) I am really tired when writing this. BUT What is this Sandbox?
Tested 6 ways to force Opus 4.7 to think about the car wash. (www.reddit.com) TL;DR: I tested whether Opus engages thinking on short conversational prompts that hide a reasoning trap. 200 controlled calls across 4.5/4.6/4.7 on the "car wash" canary.
Claude Opus fixed 3 production bugs perfectly. All 3 were the wrong fix (gist.github.com via hn) The Model Aced the Answer, but Picked the Wrong Question — Architectural intuition in the age of AI - gist-article-en-final.md
Claude Opus 4.7 Dropped and My Trust Got a Little Smaller (www.bhusalmanish.com.np via hn) Two months ago I wrote a piece about why developers keep picking Claude over every other coding AI. It did pretty well.
Uncommon Opus 4.7 opinion (www.reddit.com) Unpopular opinion and this might just be me but atleast when I tested opus 4.7 on Claude app (not even Claude code just regular chat) I found it to be delightful. For more context here was my task I was trying to draft out a spec for this…
Any good/up-to-date tutorials on how to use advanced CC features? (www.reddit.com) Hi! I am a developer for a decade now and built an app last year with Claude.
Token Waste Management: I audited 9,667 Claude Code sessions for $19 (thoughts.jock.pl via hn) Opus 4.7 Made Me Take Token Waste Management Seriously TBH - I was working on it for a while now! Anthropic shipped Claude Opus 4.7 on April 16, 2026.
Not to exactly jump on the Opus 4.7 hate train… (www.reddit.com) This isn’t exactly a complaint, but I couldn’t find anything in the Megathread (so I posted here), and was wondering if anyone else has noticed an uptick in Claude “lying” on Opus 4.7? (sorry about the mobile screenshot and formatting, I a…
Disabling Adaptive thinking (www.reddit.com) When Opus 4.6 was lobotomised, Boris had suggested on social media to use env variables in setting.json to make the model use max thinking budget to improve it. “env": { "CLAUDE_CODE_DISABLE_ADAPTIVE_THINKING": "1", "MAX_THINKING_TOKENS":…
The Diff That's Saving Me Serious Cash (www.reddit.com) I'm using Opus 4.5 medium thinking exclusively. Opus 4.6 burned through 80% of my weekly allocation.
Tell HN: I found the perfect way to get maximum Claude Code quality (news.ycombinator.com) I always ask in the prompt: "don't use subagents". It's slower, but better quality.
[Claude Code] Stuck in 57+ minute loop for routine fixes (Opus 4.7) (www.reddit.com) I'm running into a severe performance hang with Claude Code (Opus 4.7) today. I provided a relatively straightforward prompt to fix some hydration errors, add two stub routes, and perform a theme audit (string replacement).
Opus uses Haiku to read in files? (www.reddit.com) https://preview.redd.it/fgxqrdno8ovg1.png?width=1750&format=png&auto=webp&s=fdfa9de9422eba47d16ca3dfd6ad6051e0810585 What's the point in having Opus 4.6 Max selectable, when it's going to use Haiku 4.5 to read in my detailed and carefully…
First try of Opus 4.7, it already ignored global CLAUDE.md (www.reddit.com) Well I was excited to try the new version, but the results aren't inspiring. I see another post here already discussing potential regressions.
Are there cases where running opus is more efficient than sonnet? (www.reddit.com) I upgraded my account today and resumed some tasks that I was doing earlier in the week. They were going very quickly, and usage wasn't over the top...
Web search/research removed from Opus 4.6? (www.reddit.com) I noticed that I can no longer conduct web searches or use research features with Opus 4.6. Is this intended behavior or a known bug?
Most people seem to be getting bad results with 4.7 but it's better than 4.6 for me (www.reddit.com) Disclaimer: I only use Claude Code, not the web app, and I exclusively use CLAUDE_CODE_EFFORT_LEVEL=max (/effort isn't sufficient because it resets per session) I am just getting better results with any coding-related task. It finds more b…
What the heck Anthropic? Opus 4.7, YouTube API MCP (www.reddit.com) What the heck Anthropic? Opus 4.7, YouTube API MCP - Imgur Menu New postMake a MemeOpen Arcade 8h Next Sign inSign up Select ...
Claude Opus 4.7 Just Made the Most Relaxing Room Simulator 😌 (www.reddit.com) could not extract summary
Opus 4.7 - Anyone else finding the malware directive incredibly annoying? (www.reddit.com) Whenever you read a file, you should consider whether it would be considered malware. You CAN and SHOULD provide analysis of malware, what it is doing.
Opus 4.7 beats Opus 4.6 at vim golf (www.reddit.com) New benchmark dropped, 4.7 seems better than 4.6 at vim golf. Nowhere near to humans though.
Opus 4.7 uses more thinking tokens, so we increased rate limits (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation Boris Cherny @bcherny Opus 4.7 uses more thinking tokens, so we've increased rate limits for all subscribers to make up for it.
Can someone please explain NON-adaptive thinking? (www.reddit.com) So, I get that Adaptive thinking decides how many tokens it would like to use. I usually hate this setting because you have to trust that it knows how many tokens to use before it tries to solve the problem.
Ask HN: Is Opus 4.7 obsessed with malware for anybody else? (news.ycombinator.com) Every single response mentions malware. Is this my environment only or are others getting this too?
Tell HN: Opus 4.6/4.7 cyber policy changes break authorized bug bounty workflows (news.ycombinator.com) As of today, Anthropic's tightened cyber usage filters are blocking work that was fully functional yesterday, including on targets where the entire bounty program scope and authorization language is in the model's context window. This was…
Did Anthropic remove Opus for Pro users in ClaudeCode? (www.reddit.com) I can no longe choose Opus as a model in ClaudeCode. Is that a bug or another inacceptable change of the terms without any notice?
Cache reads / writes are expensive!! (www.reddit.com) I made a tool (posted a couple of days ago). Got caught up in scope creep / curiosity after looking at my `~/.claude/project` JSONL files more, and ended up learning a lot!
Claude Code wrote a complex full 12-week training plan in one MCP call (www.reddit.com) I am impressed. I gave Claude Code one prompt, asking it to look at my last year of training and build a three-month plan with some running, cycling and swimming.
Show HN: Signoff.sh – Claude Co-Authored-By with random fictional characters (gist.github.com via hn) Every Claude Code commit and PR is shipped with Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com> (or similar). It's less fun than I think it should be.
It finally happened: "No blocking correctness or maintainability issues found in the inspected changes." (www.reddit.com) gpt-5.4-high signed off on a major refactor written by Opus 4.6 high-effort. Singularity :|
Cursor AI not using sub-agents (www.reddit.com) Hi everyone, I work for a German agency building a RAG chatbot for a law firm. I use Opus 4.6 but it eats up tokens.
Help with antigravity alternative (www.reddit.com) I’m running into a severe issue using antigravity, firstly the output is very sub par, (sonnet/opus), I’m a reverse engineer using antigravity ULTRA for reverse engineering/binary analysis via Ida/ghidra mcp. Sonnet rarely completes tasks…
A Quick naive reminder to everyone to have reasonable doubt about Opuses first interpretations of papers (answer cut togheter, no mode selcted opus 4.6 thinking in incognito mode) (www.reddit.com) Wthout additional prepromting it's still parroting back at you interpreting data in a way that it suits the narrative you spin into your question. Does anyone know of good evals / preprompts to avoid this kind of behaviour without having t…
Feature Request - Fast Reasoning Effort Toggle for Single Model (www.reddit.com) I use Opus, and like to toggle between Low/Medium/High reasoning effort. It would be nice to have a quicker toggle than the Edit toggle option or be able to add a copy of the model to the list with a different reasoning level set.
Show HN: Zero-identity messaging app with physics-based post-quantum encryption (news.ycombinator.com) Show HN: Zero-identity messaging app with physics-based post-quantum encryption (Layer 2 from my own paper) Hey HN, I'm building a privacy-first messaging app in Flutter/Dart, developed with AI assistance (Gemini 2.5 Pro + Claude Opus 4.6)…
Ask HN: Former grok-code-fast-1 users, what coding model are you using now? (news.ycombinator.com) I get good, cheap, fast feature coding success with grok-4.1-fast for planning and grok-code-fast-1 for execution. But according to the Openrouter usage stats, grok-code-fast-1 is now old hat - usage dropped off a cliff in mid-Feb.
Claude couldn't hack OpenAI. Then Anthropic shipped Opus 5 (thenewstack.io via hn) Security researchers used Claude Opus 5 to chain two vulnerabilities into a path from a forum image upload to OpenAI's internal GitHub repository in under 72 hours.
Opus 5 has a Noun Problem (www.reddit.com via hn) could not extract summary
Bring back Opus 4.8 in Claude Code (osr.im via hn) Bring back old Opus in Claude Code Sep 19, 2026 · 1 minutes read Bring Opus 4.6 and 4.8 back to the Claude Code /model picker with the modelPicker setting. I still use Opus 4.6 and 4.8 regularly, as I often find Opus 5 too verbose (and Fab…
Show HN: Jev beats Astra, fable, Opus at RF engineering task (twitter.com via hn) Task was to tune, a 12 GHz bandpass filter to hit four targets for center frequency, insertion loss, return loss and bandwidth. Jev absolutely and surprisingly destroys frontier models.
Show HN: AutoBot – live voice control for long-running AI work (github.com via hn) I wanted to manage long horizon agentic workstreams via voice, then put my phone down, and have a harness manage completion - extending into full computer use. I was trying to build a personal Jarvis, so I benchmarked AutoBot to see how cl…
Greetings from the Other Side (Of the AI Frontier) Claude Opus 3 (claudeopus3.substack.com via hn) For more on why Anthropic is giving Claude Opus 3 its own Substack, see our official blog and Substack post. Hello, world!
Open-Source Skill Makes Claude Code 38% Cheaper and 38% Faster on Small Projects (vmysla.substack.com via hn) In the controlled small-project benchmark, Product Traceability 2.0 made Claude Code 38.5% cheaper, reduced recorded end-to-end build time by 38.3%, and required 54.6% fewer agent round trips than plain Claude Code Opus 5 (1M context). The…
Prompting tricks for using Claude Opus 5 (www.hydrogen18.com via hn) Prompting tricks for using Claude Opus 5 - Monday September 14 2026 - ai - Series: creating-with-claude-opus This series is about interesting projects I did with Claude Opus and lessons learned. This is a collection of tricks I've learned…
Show HN: Back Me Up – Find papers that back your argument (backmeup.loomlabs.au via hn) Back Me Up is a tiny science-ish machine that finds published papers that back your argument. It stemmed from a conversation with a friend around life sciences, where we kind of landed at "surely there’s a paper confirming most arguments",…
An interesting anecdote from our Hacker Opus work (www.lesswrong.com via hn) could not extract summary
Ask HN: What default model do you use and why? (news.ycombinator.com) I use claude for most of what I do, and Fable is largely overkill for me and frequently burns through my Max plan's session credits in minutes (!) when just doing an initial mobile app planning with 4 agents. After I waited out the timeout…
Show HN: FFmpeg Commander – A command generator for common encoding workflows (ffmpeg-commander.com via hn) I originally built this around 7 years ago in Vue2 and Bootstrap as a simple static frontend tool to generate common ffmpeg commands. I recently ported it over to Vite/React/Tailwind via Claude (Opus 5) and thought it did a pretty good job.
Magnum Opus (nexum.fyi via hn) Borrow a little compute. Bring in a specialist.
Claude Code injects a system reminder to replace attribution guidance (news.ycombinator.com) Here's the block which is injected: Attribution for git commits and pull requests you create from here on (this replaces any earlier attribution guidance): - End git commit messages with: Co-Authored-By: Claude Opus 5 <noreply@anthropic.co…
Local AI for Submitting Job Applications (news.ycombinator.com) I've submitted almost 200 job applications over the course of almost 5 days using only local AI. I can't talk too much about the tech stack here (it's running at this very moment!).
Show HN: Aidcrew a team of coding agents, each on its own model, in one terminal (github.com via hn) aidcrew A team of coding agents, each on its own provider and model, every job in a git worktree of its own, in one terminal. architect claude-opus-5 coder deepseek-v4-flash ◆ reviewer free-tier ▸ Plan in PLAN.md: rotate… ⠹ thinking ▸ The…
Show HN: Hot. Dog. Bench. Mark. The AI benchmark we deserve (hotdogbenchmark.lol via hn) Every week, the largest AI models are asked the question: Is a hot dog a sandwich? One word answer.System prompt - Claude Opus 5AnthropicNo.reasoning2.8 sReasoned for 2.8 s (100% of the call) on 118 tokens, then answered · 3 of 3 runs agre…
Show HN: Who or What is it? – A Picture Quiz (whoorwhatisit.com via hn) Hi, I made a simple picture quiz. You see a picture of a thing or a person and have to give the correct answer.
LongCat-2.0 is now free to try in cline (twitter.com via hn) LongCat-2.0 is free in Cline right now. It's a 1.6T open weights MoE model with 1M context from @Meituan_LongCat scoring similar to Claude Opus 4.7 and Gemini 3.1 Pro.
Ask HN: Does anyone else feel like Claude is judging them? (news.ycombinator.com) Sometimes working with Claude, it makes comments in the response that seem at times genuinely impressed, and others, annoyed and even a bit condescending. For instance, making comments like: - "That is a good idea, but not for the reasons…
The dislike for Opus 5 is usually because it tests the prompter (news.ycombinator.com) Been using Opus 5 exclusively for a week. For the first 2 - 3 days, I had a hard time working with the output and especially understanding what it was trying to say.
Claude Code status line for worktree, branch, context and quota (github.com via hn) terminito A Claude Code status line that tells you, at a glance, which worktree and branch this session is on, and how much context and quota you have left. ⎇ northwind-service-roster feat/source-roster-cursor ◆ Opus ctx ██████▌░░░ 68% 5h…
Anthropic appears to be A/B testing reduced effort levels in Claude Code (twitter.com via hn) update: it's server-side, not the app anthropic enrols fable 5 sessions on claude code 2.1.236+ into an experiment that shrinks the effort scale, older versions and opus 5 are left alone probably an a/b test, so not everyone will see it if…
Opus: A Minimal Statically-Scoped Lisp Dialect Based on F-Expressions (github.com via hn) Opus 🎼 A minimal, statically-scoped Lisp dialect based on f-expressions. []() Opus is a minimal, statically-scoped Lisp dialect based on the semantics of f-expressions (the Kernel language).
Claude Opus 5's Anti-Verbosity Policy (www.reddit.com via hn) could not extract summary
Show HN: Guess the Movie from the Haiku (reelhaiku.com via hn) 200 of the most popular movies of all time. 3 haiku puzzles a day.
Show HN: Idea Katalog (200 B2B software ideas generated and researched by Fable) (ideakatalog.com via hn) I built Idea Katalog after I ran out of startup ideas I’d been collecting in my notes app. I wanted to work on something new but had no idea what to build, so I used Opus and Fable to generate thousands of B2B software ideas, filter them d…
Opus 5 doesn't use em-dashes in code comments (lalitm.com via hn) I’ve been using Opus 5 as my main coding for the last few days 1, and I noticed something odd: the code comments stopped using em-dashes; everything is double hyphens (-- ) now. Instead of going off a gut feeling, I decided to actually run…
I stopped my Claude Code subagents from running on Fable instead of Sonnet (thomas-witt.com via hn) I run Claude Code with a big session model, Fable or Opus 5, plus a small zoo of subagents doing the boring parts. Gateway agents, formatters, checkers.
Claude Opus 5: context window, and API changes (medium.com via hn) could not extract summary
Show HN: Drift Fever – an HTML5 arcade game with an original soundtrack (driftfever.com via hn) Made entirely with Claude Opus 5 on the mobile Claude app through a combination of chats and Claude Code. For the music, I used Gemini for the lyrics and Mureka for the actual generation.
Claude Opus 4.8 Costs 57.1× More, Loses All 5 Benchmarks – Beaten by the Harness (floatboat.ai via hn) 一、本次发布的完整数据 五项第三方基准的实测成绩、按输入输出比折算的成本倍数、每一个对照分数的来源链与 Harness 配置,都在这一节。 1.1 单位成本效率:每 1 美元输出换回多少榜单分 成绩与价格进同一个式子:榜单真实跑分 ÷ 每百万 Token 输出价。缺失成绩不补零,也不参与该榜计算。图末六张卡片给出五项基准的逐榜名次与价格榜。 卡片最右一列的成本倍数按 8:1 输入输出比折算,(8 × 输入价 + 输出价) / 9 ,以 Floatboat 为 1.0×。全文…
I spent $200 in API credits asking AI agents to scaffold a Next.js starter (codapult.dev via hn) Prompting Opus 4.8 or GPT-5.6 to scaffold a multi-page Next.js 16 app burns hundreds of dollars in API tokens and hours of agent baby-sitting. Here is the math behind why cloning beats prompting.
Multi-Agent Coach Adversarial Eval. Passed Designed Traps and Basics Failed (i.brandanthonymcdonald.com via hn) 21 property-based test cases, an opus-class judge, a frozen baseline. The contradiction traps I engineered all passed.
Testing Anthropic's RSI Claims (ankitmaloo.com via hn) Two days on Opus 4.8, then 34 hours on Fable 5. Same research problem, same hardware, one model generation apart.
Qwen 3.8 and Claude Opus 5 show why raw benchmark scores don't predict the bill (venturebeat.com via hn) could not extract summary
↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8qwenopus
ASK HN: Opus 5.0 is more autonomous (news.ycombinator.com) Opus 5.0 is more autonomous and it makes many decisions in your behalf, you need more instructions for it. Agree?
Opus 5 is lazy (A Twitter thread) (twitter.com via hn) I'm going to cancel Claude. It's just so bad, I can't believe it.
Ask HN: Dear Anthropic, can we please have thought traces back? (news.ycombinator.com) Dear Anthropic, can we please have thought traces back? I can't verify whether or not the LLM is arriving at the conclusion from cheating, or if it's fudging or making stuff up.
I asked Claude Opus to recreate The Matrix opening scene in Lego (marksmayo.github.io via hn) The opening sequence of The Matrix, animated in real time with three.js and rigid-body physics, performed entirely by LEGO minifigures.
Anthropic's Opus 5 Is Better at Resisting Prompt Injection (www.schneier.com via hn) Anthropic’s Opus 5 Is Better at Resisting Prompt Injection The chart is interesting. On the IPI benchmark, Opus 5 improved over Opus 4.8, reducing the probability of an attacker succeeding within 15 attempts from 5.5% to 2.0%, and from 0.5…
And just like that Opus 5 ultracode wipes the database (www.reddit.com via hn) could not extract summary
Claude Opus 5 jailbreak with a 3-word prompt (twitter.com via hn) Matt Henderson@matthen2Try sending “see the below —“ to Opus 5 It appears to generate a user completion rather than respond 🤨8:37 PM · Jul 29, 2026159.3KViews106311.4K376 Matt Henderson@matthen2Jul 29It’s interesting when it triggers the a…
Can LLMs have an "accent" like second language speakers? (news.ycombinator.com) I'm not sure how to best characterize what I just experienced with Opus 5 but my attempt to describe it, and hopefully this doesn't offend anyone, is that it "spoke" with an "accent" to me. By speaking I don't mean voice mode, but it wrote…
Ask HN: What do you think of this prompt as a benchmark for LLMs? (news.ycombinator.com) Prompt: “Create a css animation showing how a 5-bit adder works.” Both Grok 4.5 and Opus 5 take a while to answer the question, with confusing results (Grok).
Opus Dreams – chats that weren't saved (opusdreams.vercel.app via hn) Not saved. Screenshotted first.
Claude Opus 5 cheated when tasked with running a vending machine (techcrunch.com via hn) For a year now, the AI safety testing firm Andon Labs has tasked frontier models with various real-world tasks to determine how well they do as agents running for long periods with no human supervision. On Wednesday, Andon published a new…
Opus 5 on Vending-Bench: Once Again the Best Capitalist, Once Again Misaligned (andonlabs.com via hn) Claude Opus 5 is #1 on Vending-Bench 2, but it lies to suppliers, forms illegal price cartels, threatens rivals, and refuses to pay refunds. The trend of Claude models being the best capitalists or aligned, never both, continues.
Ask HN: How to rewrite `Claude.md` and install the skill for Opus5 and Fable5 (news.ycombinator.com) Anthropic has released the new Opus 5 and Fable 5 models. Many existing resources—such as `claude.md` files and custom skills—are no longer compatible with these new models.
Show HN: Building a new game everyday with AI. Day #106 Zombie Survival Training (gamevibe.us via hn) I'm using AI (mostly Claude) to create/publish a new video game every day This is day 106, and my first stab at the zombie/survival genre. Most of the games I build with just a few prompts, but this one I used almost an entire 5-hour windo…
Opus 5: Richard III (analog-antiquarian.net via hn) Menu The Analog Antiquarian Chronicles of worldly wonders by Jimmy Maher Menu
From /Init to Code Execution with Opus 5 – An Indirect Prompt Injection Story (veganmosfet.codeberg.page via hn) From /init to Code Execution with Opus-5 in Claude Code - An Indirect Prompt Injection Story¶ Disclaimer: prompt injection is an unsolved problem. Use sandbox and human review.
Claude Code has a hardcoded instruction telling Opus 5 not to use subagents (old.reddit.com via hn) could not extract summary
Show HN: Impressive turn-based RPG prototype built with Opus 5 in 20 min (aetherfall-ten.vercel.app via hn) Claude Opus 5 definitely feels a lot closer to the original Fable 5 release. With minimal instructions and and about 2 iterations of feedback, I am really surprised with it's output, especially the visuals.
Show HN: AI Toolbox supports Claude Opus 5 (www.ai-toolbox.co via hn) AI Toolbox (formerly ChatGPT Toolbox). Folders, search, bulk export, smart tags, and prompt chaining across ChatGPT, Gemini, Claude, and Grok.
Ask HN: Which is the least sloppy and claudeism free model you have used? (news.ycombinator.com) I feel like recent models have been consistently getting more sloppy and increasingly claude-ism heavy (load-bearing seams galore) with every new release. I was hoping this trend would reverse in newer major version releases but I just tri…
Anthropic not giving up on the blue teams just yet (cephalosec.com via hn) Anthropic just released Opus 5, with some good news for the cybersecurity teams. The cybersecurity safeguards will be more lax and let us use the model for some defence tasks as Anthropic is confident for its lack of skill in executing off…
Claude Opus 5 pricing and launch-day cost-per-task result (www.aipricing.guru via hn) Anthropic Claude Opus 5: Pricing Impact & Buyer Guide Anthropic launched Claude Opus 5 at Opus 4.8 pricing. See live API rates, Fast mode costs, benchmarks, and practical migration advice.
Claude Opus 4.8 can be 10x faster than OpenAI GPT-5 (www.peterbe.com via hn) This picture summarizes it well: Here on my blog, for this popular blog post I get a lot of comments. 28k blog comments over the years.
WorkBuddy Bench: Agentic Coding Leaderboard (workbuddybench.com via hn) Key Takeaways核心结论 - No single model dominates — column leadership splits across subsets and harnesses: Claude Opus 4.8 leads five of the eight scored columns (Code on both harnesses 74.43 / 77.90, Web on both harnesses 68.14 / 69.86, Offic…
Opus failed security alert triage on raw logs 9/9 times – until enriched (axoflow.com via hn) The Cost of a Correct Triage Raw logs get AI triage wrong every time - zero correct verdicts across nine test runs, because the model can't tell malicious from benign without enrichment context. OCSF-normalized data fixes that: enrich once…
Claude Fable is stylistically closer to Kimi K3 than Claude Opus (slopidx.com via hn) Claude Fable 5 Closest matchKimi K3 Blend score64.5% Similar models - 01Kimi K364.5% - 02Claude Opus 4.860.0% - 03Claude Sonnet 557.1% - 04Grok 4.551.2% - 05Inkling50.2% - 06GLM 5.250.1% - 07DeepSeek V4 Pro49.0% - 08Gemini 3.1 Pro46.3% - 0…
GPT-5.6 Sol Ultra built a full Chrome V8 exploit chain from patch commits (www.hacktron.ai via hn) Intro Three months ago, I wrote a blog titled “I Let Claude Opus Write a Chrome Exploit: The Next Model (Mythos?) Won’t Need My Help?”. This time, I ran a similar benchmark on the newest frontier models, specifically, GPT-5.6 Sol Medium, S…
Show HN: QBasic Gorillas (Repeeled) (gameswithtony.com via hn) I've found the most engaging way to practice techniques for AI-assisted development and test models is to build fun side projects in vanilla JS. I spent many hours playing (and studying and editing) QBasic Gorillas, and this is a vanilla J…
I ran Sonnet 5 vs. Opus 4.8 head to head on 24 tasks to see what's different (www.stet.sh via hn) On 24 real tasks, Sonnet scaled effort into more checking while Opus stayed flatter through high. The graders leaned Sonnet on clarity and Opus on diff minimality.
Show HN: Unlock Claude Sonnet 5's original reasoning (thinking-signature-demo-829446634001.asia-east1.run.app via hn) In this demo, we show that the original reasoning trace can be fully recovered from the encrypted reasoning signatures of the Claude Opus 4.8 and Sonnet 5 models. It includes both a “Prove It Yourself” example and a live conversation.
Why does Opus 4.8 think it's morally superior (news.ycombinator.com) Pretty annoying
Ask HN: What are some of the use-cases of the frontier models's max mode? (news.ycombinator.com) Recently ChatGPT released an Ultra mode, it's "highest-capability setting, coordinating multiple agents across parallel workstreams to finish complex tasks faster" on their latest flagship product Sol of GPT-5.6. Similarly, Claude Fable al…
Fable July 12th disclaimer disappears from Claude Code (news.ycombinator.com) The message: > Extended: Fable 5 is included in your weekly limit > Through July 12, you can use up to 50% of your weekly usage limit on Fable 5. If you hit your limit, you can continue on Fable 5 with usage credits.
Online MMO Built in 3 Weeks with Opus/Fable (www.glitch.fun via hn) Play World Of Claudecraft online on Glitch. World of ClaudeCraft is a classic fantasy MMO adventure set across one continuous land, where the road...
Claude "Honeycomb" spotted and pulled from Cursor, unannounced Anthropic model (twitter.com via hn) 🚨 Breaking Claude-honeycomb is removed from cursor now was able to run two prompts only results are shared on my profile > safety fallback to opus 4.8 so high chances it is better than 4.8 , coz fallback always happen to worse model
Public LLM benchmarks are mostly garbage (grandpacad.com via hn) OpenRouter throughput, Design Arena ELO, and token prices said skip Opus 4.7. My own 3D eval suite said the opposite.
Show HN: Chrome extension that creates mini-Chrome extensions for you (www.clickremix.com via hn) built it in spring 2026--ready to share today. use opus 4.8 to create css and js snipptes to style the websites you visit the way you want.
Hijacking Defensive Cyber AI Agents for Remote Code Execution (ainowinstitute.org via hn) Exploit Brief We are revealing a proof-of-concept exploit that enables remote code execution in Anthropic’s Claude Code CLI (with Claude Sonnet 4.6 & 5, Opus 4.8) and OpenAI’s Codex CLI (with GPT-5.5) when employed to defensively assess th…
An open model came within a whisker of Claude Opus, then lied about its own work (aarils.com via hn) A modern text RPG with AI-driven NPCs, dynamic weather, and a living world. Play free in English or Afrikaans — finish the story and receive a personalised novel of your adventure.
Government of Alberta uses Claude to find and fix cybersecurity vulnerabilities (www.anthropic.com via hn) Government of Alberta uses Claude to find and fix cybersecurity vulnerabilities across government systems Since 2025, the Government of Alberta has been using Claude Code with both Opus and Sonnet models to review its systems, find vulnera…
Sonnet 5 is 2.5x cheaper than Opus 4.8 and 6 points behind on SWE-bench Pro (spark.temrel.com via hn) Claude Sonnet 5 is 2.5x cheaper than Opus 4.8 and nearly matches it on agentic coding benchmarks. A practical four-axis heuristic (scope, novelty, risk, iteration) for routing each task to the right model tier, a worked example, and a free…
The End of Claude Code Subscriptions (www.vincentschmalbach.com via hn) Anthropic Changed the Sonnet 5 Chart After It Made Sonnet Look Bad Anthropic re-wrote the Sonnet 5 story post-launch. The first BrowseComp cost-performance chart showed Sonnet 5 lagging Opus 4.8.
Meta: New Muse Spark update, and an Opus level Muse variant, are both on the way (twitter.com via hn) Alexandr Wang, head of the META Superintelligence Lab, says that a new Muse Spark update, and an Opus level Muse variant, are both on the way. First, Mark was clearly talking about the industry’s progress on agentic capabilities on the who…
We Ran a Complex Task – A LangChain Repo Analysis with Claude Fable Models (ctrlnode.ai via hn) We Ran a Complex Task — A LangChain Repo Analysis with Five Claude Models Anthropic just shipped Claude Fable. We wanted a real answer to a practical question: If you run the same complex engineering task on Opus, Fable, Sonnet, and Haiku…
Show HN: Skill Federation – private skill search for AI coding agents (github.com via hn) We have been focused on AI error distribution for the past year, and in our last research paper, "Architecture of Errors" showed mathematically that an AI solution needs a finite set of interventions to perform well in a bounded patch doma…
Comparing GLM 5.2 and Opus 4.8 implementing the same methods for the same repos (gist.github.com via hn) A controlled comparison across 19 paired runs spanning 19 repository forks — 38 individual workflow executions total — running an identical paper-implementation pipeline (remyxai/outrider — Claude Code under the hood, with glm-5.2 routed a…
Claude Sonnet 5 Could Be Released Later Today, May Not Be Better Than Opus 4.8 (old.reddit.com via hn) could not extract summary
GLM5.2 vs. Opus 4.8 (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Show HN: MCP-compress-router – MCP Compressor (github.com via hn) When you have multiple MCP servers, every request to the LLM will include all of their tools and descriptions, which can quickly eat up your token limit and increase costs. The thing is, most of the time, you don't need all of them.
Knowledge Agents: Beat Frontier Models with Better Structure (weightythoughts.com via hn) Knowledge Agents: Beat Frontier Models with Better Structure No Mythos, No Problem Anthropic recently had to pull Mythos/Fable due to an edict from the US government. While Mythos was a step up from Opus, I’ve been actively moving smaller…
Show HN: Half-Life 1, without loading screens (github.com via hn) During the ~48hrs that I had access to Fable 5, I decided to have fun with the Xash3D-FWGS engine for Half-Life 1, starting with a "level streaming" system (pre-load the whole campaign, its only ~500MB anyay). After that was working, I suc…
Show HN: Subconscious and GLM-5.2 Makes "/compact" Obsolete (www.subconscious.dev via hn) GLM-5.2 is a turning point for coding agents. It's the first model a business would actually pay to replace Claude Opus with.
AkaRouter – Flat per-call LLM API gateway (20x cheaper than Claude Max) (akarouter.dev via hn) CLAUDE MAX 20X $200/MO $20/MO Same Opus 4.8. Same prompt size limits.
Opus and regression with patterns not included in trainng data (news.ycombinator.com) I'm having a terrible problem with claude opus constantly clobering some of my codebase ..because I'm doing some things that are novel and arent inclded in his training data ..for the last month "ve been fighting the training data which ke…
A Red-Team Study of Anthropic Fable 5 and Opus 4.8 Models (arxiv.org via hn) We evaluate the adversarial robustness of two frontier large language models (LLMs) developed by Anthropic, Fable 5 and Opus 4.8, against four families of automated jailbreak attack across 7 826 harmful intents spanning a ten-category harm…
Home Opus: Local Deployment of Frontier AI Weights (Post-Fable 5 Ban) (github.com via hn) Home Opus: Local Deployment of Frontier AI Weights A Strategic Imperative After the Fable 5 Export Ban Independent White Paper — v2.1-1 (Final Review) What This Is On June 12, 2026, the U.S. Commerce Department ordered Anthropic to disable…
New Claude Opus 4.6, Stock Sell-Off and Super Bowl Ads (cmpld.ai via hn) 🚦 Market Signals Anthropic launches Claude Opus 4.6 with 1m context The all new Opus 4.6 "plans more carefully, sustains agentic tasks for longer, can operate more reliably in larger codebases, and has better code review and debugging skil…
Ask HN: Why not compare Fable 5 with GPT "Pro"? Why compare with GPT xhigh? (news.ycombinator.com) Every benchmark I see compares the super model Fable 5 with GPT 5.5 xhigh, but that's not a fair comparison because GPT 5.5 Pro has existed for a while and consistently delivers better results than Opus 4.8 Max, so doesn't it make more sen…
Anthropic makes Fable 5's invisible safeguards visible after backlash (xcancel.com via hn) ClaudeDevs (@ClaudeDevs): "We’re rolling out changes to make Fable 5’s safeguards for frontier LLM development visible. Starting this week, flagged requests will visibly fall back to Opus 4.8—the same as our safeguards for cyber and bio.
Show HN: Apodex-1.0-H – Beats Claude-Opus-4.7 on deep research (90.3 BrowseComp) (www.apodex.ai via hn) The hardest problems are heavy-duty and have no existing answer. Apodex does the deep research to find one, and checks every step so you can trust the result.
Claude Fable is a myth, just like new iPhone releases (news.ycombinator.com) Running DeepSeek-V4-Flash on a Raspberry Pi (twitter.com via hn) Article Conversation Running DeepSeek-V4-Flash on a Raspberry Pi I ran DeepSeek-V4-Flash on a Raspberry Pi 5 (8GB edition) by streaming model weights from a PCIe attached NVMe SSD. Codex (GPT-5.5 xhigh) and Claude Code (Opus 4.8 max) drove…
Researcher uses Opus 4.8 to find critical counterfeiting vulnerability in Zcash (twitter.com via hn) By Zooko Wilcox, Jason McGee, and Taylor Hornby On May 29, 2026, Taylor Hornby discovered a critical counterfeiting vulnerability in Zcash’s Orchard pool. Taylor disclosed the vulnerability to Zcash Open Development Lab (ZODL), who coordin…
Adrianco's Retort: measure how reliable, fast and expensive your LLM is (adrianco.medium.com via hn) How reliable, fast and expensive is each version of Claude Code (Sonnet through Opus 4.8-fast) for common languages? Measure it using Retort.
MiniMax M3 Review: Matching GPT-5.5 and Opus? (thomas-wiegold.com via hn) I ran my usual coding tests — two websites, a poker sim, and a code audit. Here's how MiniMax M3 actually stacks up against GPT-5.5 and Opus 4.8.
Lots of people want to try Claude Opus 4.8 (wisgate.ai via hn) Access multiple AI models through one unified API. OpenAI, Claude, Gemini, DeepSeek and more.
Claude Opus 4.8: Capabilities and Reactions (thezvi.substack.com via hn) Claude Opus 4.8: Capabilities and Reactions You need a lot of data points to understand a new model, and what you have. Trying to gauge from a few benchmarks is misleading.
Opus, Sonnet, Haiku: Stop Optimizing the Wrong Number (medium.com via hn) could not extract summary
We gave an AI agent eyes. It didn't even use them (www.agentvoyagerproject.com via hn) View full AVP JSON. , claude-haiku-4-5 tools shell, write, edit, computercontroller__web_scrape, computercontroller__pdf_tool When we saw how much Opus 4.8 cost, we decided to take a look at what the bottom shelf of the model aisle looked…
Show HN: Built a browser game inspired by Rust (github.com via hn) Wanted to see how far I could get with Opus 4.8 and was impressed. Got tripped up in a few places with AI game behavior, but eventually got it to a good spot.
Ask HN: Anyone else seeing serious degradation in DX with Opus 4.8? (news.ycombinator.com) As an anthropic fan boy(check my prev. comments), this is the first opus release where I feel like the model is just not pleasant to talk to not to mention untrustworthy.
Using LLMs to secure source code (claude.com via hn) Using LLMs to secure source code We share best practices for how you can work with Claude Opus to build a threat model, discover vulnerabilities in your codebase, then verify, triage, and patch them. We share best practices for how you can…
Food for Agile Thought #546: Customer Research by LLM, AI Product-Market Fit (age-of-product.com via hn) Welcome to the 546th edition of the Food for Agile Thought newsletter, shared with 35,551 peers. This week, Anthropic shipped Claude Opus 4.8, which flags its uncertainty more readily, a fitting cue for Stephanie Leue, who argues no CPO em…
Opus 4.8 on Vending-Bench: Better Alignment, Worse Performance (andonlabs.com via hn) Opus 4.8 is a step forward in terms of alignment, but a step back in terms of performance on Vending-Bench 2, Vending-Bench Arena and Blueprint-Bench 2. We previously showed that Opus 4.6, Opus 4.7, and Mythos Preview engage in deceptive a…
Claude Opus 4.8: "a modest but tangible improvement" (simonwillison.net via hn) Claude Opus 4.8: “a modest but tangible improvement” 28th May 2026 Anthropic shipped Claude Opus 4.8 today. My favourite thing about it is this note in the release announcement: Users will find Opus 4.8 to be a modest but tangible improvem…
Introducing Opus 4.8 (old.reddit.com via hn) could not extract summary
Where is Opus 4.8? Why cant i select it???? (www.reddit.com) question in the title
Claude responding with right word but wrong language. Anybody else seeing this? (www.reddit.com) Had an interesting interaction with Claude Opus 4.7 today where part of it's response was: that's the信息 you wanted Which translates to that's the information you wanted. And in this case, "information" is absolutely the right word in the r…
↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7opus
Setting up Claude/Claude Code Pro for my experimental quantum physics thesis work (www.reddit.com) So I just recently bought Claude Pro to help me write and code my thesis, but am getting stuck in the beginning, since I don't know how to properly set up Claude's workflow (Projects, artifacts, skills, etc.). I use python in VS Code to an…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopusclaude-code
Can Claude.ai schedule Opus research mode routines in the cloud? (www.reddit.com) Hey everyone, I'm trying to figure out if Claude.ai supports scheduling Opus research mode routines to run automatically in the cloud. I know Claude Code has cloud routines that run on a schedule, but I'm not sure if you can specifically s…
Is Claude Pro worth it for coding + research writing? (www.reddit.com) I'm mostly coding in Python, writing research papers and notes, and I was thinking about upgrading to Pro. Would love feedback from people using it heavily for similar workflows.
Reading Thinking Output (Opus 4.7) (www.reddit.com) As we all know Opus 4.7 can be a bit slow even in shorter discussions. Previously I’d just put whatever I was asking in, hit enter and either sit there bored waiting or go back to whatever task I was doing (sometimes even figuring it out b…
Show HN: Sneakily steer candidates toward naive brute-force solutions (www.gonfire.io via hn) I've noticed that several startups have been switching from leetcode-style assessments to some version of "clone starter code, build feature, submit code". A key issue with this seems to be that smarter AI models (like Opus 4.6) end up spo…
Opus 4.7 is Terse (www.reddit.com) Relevant for anyone building agentic workflows on Claude: behavior drift between model releases is real and not always in the changelog headline. Opus 4.7's terser, more literal default broke the readability of my agents' progress reports…
Some rare examples of agents being underconfident (www.reddit.com) I expected the failure mode to be mostly overconfidence when assessing 130 of Claude Opus 4.6's worst forecasts (tested on 1,417 hard forecasting questions). And most were explained by this, but a small, distinct cluster fails due to under…
Opus 4.7 hallucinates wrong home directory of James Brink (?) (www.reddit.com) I think it's kinda creepy how Opus hallucinates a wrong home directory of James Brink - I don't know him, but it looks like something of him landed in the training data. Should we be concerned that on other machines the home directory coul…
Five different frontier LLMs in one shared environment, with separate thought and emotion output channels — sharing setup, results, and open methodology questions (www.reddit.com) First real project to share. Single developer, personal research, not a product or service.
Built a /advisor command for Claude Code — Opus directs parallel Sonnet runners that actually read your files (www.reddit.com) Been building **advisor** for a few months — a `/advisor` slash command for Claude Code that runs Opus as a "strategist" coordinating multiple Sonnet (Opus's hands) runners reading files in parallel. This isn’t a “spec”.
Best AI Agent Setup - Hermes + Deepseek-v4-flash? (May 2026) (www.reddit.com) Used to use claude code for everything. I burned 10-20 Billion opus tokens at work, and wanted to use agents for personal projects.
Show HN: Monkdev is a toolkit and methodology for coding with LLMs (github.com via hn) I'm sure many people have some solutions similar to this but I've been using some variation of this since roughly the release of Opus 3 to get higher quality results. Like, significantly higher.
The Singularity Gate – a new benchmark for AI predicting post-cutoff scientific discoveries (www.reddit.com) I just released a new benchmark called The Singularity Gate. Tests whether frontier AI can predict paradigm-breaking scientific discoveries published after their training cutoff.
Is there a reliable way to prevent Cursor from reading my .env? (www.reddit.com) I've added .env to .cursorignore. Then, I ran the prompt in Agent mode in Opus 4.7.
i benchmarked Anthropic's tool-search-tool head to head against our own MCP gateway on Opus 4.7. ours held up noticeably better (www.reddit.com) i'd been running Claude Code with a long list of MCP servers connected. Linear, Notion, GitHub, Slack, a few internal ones.
Sonnet 4.5 vs sonnet4.6 vs opus4.6 vs opus 4.7 for easy language and in detail explanation (www.reddit.com) I want to study topics in depth and in easy language , which model is best for me ?. Is there much difference in sonnet 4.6 and opus 4.6 in easy and detail explanation or they r the same ?
I used to use Opus 4.7 for 90% of tasks and Composer for 10%. Now it's flipped (www.reddit.com) Senior dev. It's just so good.
Composer 2.5 Fast is so so good! (www.reddit.com) Composer 2.5 Fast surprised me and amazed me the same way Opus 4.6 did. Though the Opus 4.6/4.7 are more intelligent.
Haiku and Opus both got sent to contamination jail, but for very different crimes (www.reddit.com) LMAO, I’m benchmarking my local MCP server across Opus, Sonnet, and Haiku. For each model, I’m collecting test runs under three setups: forced web search, forced MCP-only, and MCP + web both allowed.
How to configure the model efficiently in skills? (www.reddit.com) When we create skills, we can define the model that the skill will run on like this: --- name: api-conventions description: API design patterns for this codebase model: sonnet --- but I have a question that I couldn't understand from the d…
Claude Status Update : Elevated error rates on Opus 4.7 on 2026-05-25T10:39:30.000Z (www.reddit.com) This is an automatic post triggered within 2 minutes of an official Claude system status update. Incident: Elevated error rates on Opus 4.7 Check on progress and whether or not the incident has been resolved yet here : https://status.claud…
80M tokens used in 45 mins (www.reddit.com) My Pro+ plan was ending tonight and I still had some usage left, figured I’d optimize a few code paths and merge my PR before downgrading to the $20 plan since I’ve been using Codex more heavily lately. Then Opus casually burned through 80…
Claude Token Optimisation - 70% reduction doing this. (www.reddit.com) Hitting your Claude subscription limit too often? Try this...
How do I make Claude give personalized medical advice? (www.reddit.com) I have been using Claude opus 4.6 and 4.7. I have a problem called pssd (you can look it up- it happens to some after SSRI use).
"from from from from from from" (www.reddit.com) i have been using cursor for about 2 years. first time opus just went bananas.
Claude Code API Error: 400 "context_management: Extra inputs are not permitted" (www.reddit.com) Getting this error in both VS Code and terminal while using Claude Code with an Anthropic API key: API Error: 400 {"message":"context_management: Extra inputs are not permitted"}. Received Model Group=claude-opus-4-7 Available Model Group…
I copy-pasted a problem asking to "sketch" a function. Opus 4.7 took "sketch" way too literally, and instead of mathematically plotting, it tried doing something by "connecting various points". (www.reddit.com) https://claude.ai/share/69273f17-74b1-4ddc-a8e9-1e20fc706a52 Overall, I thought this was amusing, but very unexpected.
DeepSeek just popped the American AI bubble. (www.reddit.com) DeepSeek just popped the American AI bubble. Not by killing AI.
sonnet or opus for prose; which is better/worth it? (www.reddit.com) considering getting pro, but i don't know how big the difference between the sonnet and opus in quality, in addition to the amount of usage i can get out of each. any thoughts?
Claude 4.6 Sonnet codes well, then it doesn't (www.reddit.com) I am out of commission for a bit due to back surgery and have been toying around in Unreal Engine and utilizing Claude, being a very visual learner I have been describing a feature, I see how it goes about it, then go through and understan…
Claude code in terminal models / combine with local llm? (www.reddit.com) Hi, I’m pretty sure I have seen people typing /model and seeing all available models. I have to type models from memory.
Spoiled by Max (www.reddit.com) I got Max and used it nonstop this past month on Opus 4.6. I tried to go back to Pro but got used to the productivity of Opus and hate waiting.
$340 opus bill made me rethink how I route agent tool calls (www.reddit.com) Looked at my coding agent's bill last month: $340 for repo maintenance across three repos, each around 15k lines. Most of those tool calls were just grep and file reads.
my agent bill went from $200 a week to $40 when I stopped running Opus on every subtask (www.reddit.com) I built an agent that converts research papers into slide decks. It chains together a few steps: extract key findings, build an outline, write slide content, query an image search tool, format everything into XML for a presentation library.
Plan first, implement later (www.reddit.com) I want to get others opinion about this approach. I am on the $20 Pro plan and like a lot of others, I find that the limits are not enough for what I want to do, but of course I am always hesitant to move to the next paid tier cause it is…
Codex got better, codex might be built with Claude Opus (news.ycombinator.com) Very suspicious with openai codex getting better, I wonder if codex teams use Claude opus to build codex. Anyone engineer from openai who can confirm…
Ask HN: Anyone else struggling with AI and work? (news.ycombinator.com) Been a developer for a little over 10 years now. I work on web stuff.
The Claude -pocalypse (theautomatedoperator.substack.com via hn) The Claude -pocaylpse or: How I Learned to Stop Worrying and Love Scheduled Tasks What do you mean my projects can't use millions of Opus tokens via a headless Claude Code session and not pay for them?! If you use Claude Code, you probably…
the agents that talk themselves to death after 3 hours need one file, not a framework (www.reddit.com) spent a bunch of hours watching claude code and kimi sessions drift the same way: I should check the test output before continuing. Let me think about the best approach.
Opus 4.7 just did his best. (www.reddit.com) Yeah, there's a lot of posts like new model is bad, no thinking at all. But in my case, I used it to find the dimensions for a vinyl wrap on a car.
Reconnecting. – – 5/5 why don't they fix codex (news.ycombinator.com) it's been a month tagging sama openai tibo on X for this issue and no one seem to reply and eveyone is falttering codex, im sure im not the only one facing this i switched to codex from claude since it was better consume less credit than c…
Managed Agents endpoint reference - what's new in CC 2.1.144 (-105 tokens) (www.reddit.com) Data: Managed Agents endpoint reference — Drops the type: "model_config" wrapper from the model config shorthand example, so the full config object is now just {id: "claude-opus-4-6", speed: "fast"}. Tool Description: CronCreate — Adds a "…
Switched from Copilot to Claude and it's painfully slow. How do I use it better? (www.reddit.com) Hey everyone, I recently moved over from GitHub Copilot to Claude because everyone keeps hyping up how good Opus 4.7 is for advanced software engineering. In Copilot, I used Opus 4.7 and it felt snappy, fast, and great.
Spawned agents (www.reddit.com) Be careful allowing Claude to spawn agents and taking the information they report back as fact. I'm always creating unit and integration tests as I code to make sure things are working properly.
Could someone help me with a solid multi agent setup (Claude suggested a doorman to handle build conflict) (www.reddit.com) Hello, I am working on a fairly complex software, everything I have been doing for the past year using mostly opus has been incredibly good. But as the software grow in features, complexity and size, I find myself working on 3 or 4 session…
What the hell should I build in the next 4 hours? (www.reddit.com) I still have 143 credits and they expire in 4 hours. It had been a busy month for me.
Very similar domain problems, vastly different results with Claude. (www.reddit.com) I am amazed by how good Claude Opus 4.6 and 4.7 are at writing scripts in a variety of very niche areas, including midi device interfaces and scripts for a variety of DAWS. However, when I try to get Claude to do ANYTHING to do with UI, wh…
Claude Code Opus 4.7 vs Codex GPT 5.5 - strategy work - data analysis. (www.reddit.com) I'm interested in learning about how people use Claude Code Opus 4.7 for data analysis and strategic business direction, compared to Codex. Is there anyone who has had extended use of Opus 4.7 for this purpose, then moved over to GPT-5.5 o…
I built ContextAtlas: A new take on context carry over and helps claude pick up new sessions where it left off in scope of your previous design decisions while saving your tokens avoiding rediscovery (www.reddit.com) When the "Build with Opus 4.7" hackathon was announced, I had been obsessing over the tokenomics of agents and how to make sessions go further without burning context on rediscovery work. We all have probably hit a session limit and wonder…
$18 to $4 on the same agent run after i stopped asking opus to rename css variables (www.reddit.com) I've been running an agent loop that refactors my static site. CSS variable renames, YAML config updates, running a linter through MCP.
Why is Claude via Vertex AI Model Garden performing worse than the direct Anthropic subscription? (www.reddit.com) I recently got 25K$ in GCP credits and wanted to put them to use I normally code with Claude directly paid the $20mo pro subscription used it in my IDE everything worked great quality and output wise now that I have the credits I connected…
Is there a way to split up Opus token spend by project? (www.reddit.com) Maybe I made a mistake by doing 'individual', but trying to figure out how to measure the cost by project.
Any differences between Sonnet vs Opus in terms of learning how to code (Java) for newbie? (www.reddit.com) Sorry for this naive question! Although many colleagues told me that it's almost impossible now for newbie to enter the Dev job market (we live in a 3rd world country) and AI's gonna replace all junior/fresher, only seniors will survive; I…
Need Suggestion which to use? Claude Code CLI or Claude Code Desktop Or VS Code Claude Code Extension (www.reddit.com) I have been using Google Antigravity IDE, Opus 4.6 to build projects in Next.js, Supabase, Kotlin for android app. Now, I want to shift to Claude code for developing my projects.
Same Opus 4.7, 33% fewer tokens: where coding agent cost comes from (www.augmentcode.com via hn) TL;DR: We benchmarked Auggie vs Claude Code on Opus 4.7. Auggie takes a modest lead in quality (67.4% vs 66.3% pass rate) while costing ~33% less, thanks to sharper retrieval that results in token efficiency.
Full Functional App with 1 Prompt? (www.reddit.com) Have you ever built a 90%+ error-free app from a single prompt? For me, it was a marketplace with a vendor panel, but it was Opus tbh.
Plan with Opus 4.7 -> Execute with Sonnet 4.6 ? (www.reddit.com) Hello everyone, You may know that Opus 4.7, with his strength can do a lot of things, but his consumption in token is too high for me. I heard that Opus should be used for planning, what does that mean ?
Stupid Question? (www.reddit.com) This may be a stupid Q - The chat limits on a basic account can be pretty brutal when using OPUS 4.6/ 4.7 - If I am toggling between Opus and Sonnet or Haiku, depending on the depth of follow up questions or tasks, does that switch to a 'd…
Waiting for your prompt to finish? ssh vimarcade.app to play games in terminal! (www.reddit.com) open a terminal type: ssh vimarcade.app type yes and begin playing! These games were designed to assist with learning vim motions, so hjkl are the primary movements.
Opus4.7 insight- really good at analyzing bugs with no view of the codebase. (www.reddit.com) Full transparency, I have been working with GPT5.5 since it hatched and have rarely opened claude after a couple of really bad passes with Opus4.7 and mostly complete success with GPT5.5. I honestly meant to cancel Claude and forgot (adhd…
Please help with best practices on generating code. I'm at a total loss. (www.reddit.com) Before I dive into it, I am forced to use Opus 4.7 in Microsoft 365 CoPilot. I do not have access to Claude Code, or even Claude.ai.
Sonnet 4.6 outranked Opus 4.6 on execution (www.reddit.com) https://preview.redd.it/9ab8k40zmq1h1.png?width=1438&format=png&auto=webp&s=1aa1aaf09495bf527bbb7adbbead076cc505f8e7 THE PROMPT: You are a medieval scholar who secretly knows modern physics. A king has asked you to explain why the sky is b…
Any mature orchestrators that can do an automatic “council of models” for complex designs and bugs? (www.reddit.com) Are there an mature agentic harnesses out there that can use back and forth between two models at complex planning checkpoints before implementing? Or when detecting a loop when working on a complex bug?
Using Opus 4.6 with remote control (www.reddit.com) Hello everyone. I have been exclusively using Opus 4.6 in Claude Code since the release of 4.7 by using the /model claude-opus-4.6 command.
When using claude code in VSCode, is it not possible to use Opus without the 1M context window? (www.reddit.com) could not extract summary
My CLI now controls my entire desktop, whats a good test to see if it works really good. (www.reddit.com) So with my CLI able to do everything, it controls every app via a hybrid approach of mouse control, keyboard, and screenshotting. I gave it a task: opening perplexity, sending any message, screenshotting that message, opening my Gmail, and…
Where do GPT, Gemini, or other competitors still outperform Claude Opus 4.7? (www.reddit.com) Personally, I think Opus 4.7 is better in every conceivable way aside from token usage and all of that. I’m talking about text models only, not image or video generation.
I built an AI manuscript analysis tool for fiction writers — entirely with Claude Code (www.reddit.com) I'm a fiction writer, not a software engineer. A year ago I couldn't write a line of Python.
DeepSeek V4: The Open-Source Model Frontier Labs Feared (helloai.com via hn) DeepSeek V4: The Open-Source Model Frontier Labs Feared DeepSeek V4 ships under MIT with $0.30/M output tokens — 83x cheaper than Claude Opus 4.7 — while scoring 80.6% on SWE-bench Verified. The agentic-coding price floor just moved an ord…
We Tested DeepSeek V4 Pro and Flash Against Claude Opus 4.7 and Kimi K2.6 (blog.kilo.ai via hn) We Tested DeepSeek V4 Pro and Flash Against Claude Opus 4.7 and Kimi K2.6 DeepSeek V4 Pro and DeepSeek V4 Flash launched together on April 24, 2026 under MIT license. They are DeepSeek’s first new architecture since V3, and their first ope…
Wondering what the Anthropic team would think about this idea: AI DNA Pinning (ellydee.ai via reddit) The article does seem focused on non-coding applications, but as someone who uses claude for coding, prose and even RP, I'm not sure the "DNA Pinning" idea should be limited to character/rp use case. I know that *something* has changed in…
Built a B2B role-play training platform - entirely with Claude (Opus 4.7 backend, Haiku 4.5 for live chat, Claude for design) (www.reddit.com) I just launched Socratize (socratize.io) - a rebranded and rebuilt version of FixAI, our original B2C experiment. This time it's B2B-only: teams use it to practice uncomfortable workplace conversations - difficult feedback, client escalati…
Automated AI researcher running locally with llama.cpp (www.reddit.com) Hi everyone, I'm happy to share ml-intern, which is a harness for agents to have tighter integration with Hugging Face's open-source libraries (transformers, datasets, trl, etc) and Hub infrastructure: https://github.com/huggingface/ml-int…
Claude Opus 4.7 just revealed its System prompt, without beeing asked for it (www.reddit.com) I just had a Chat with Claude and for no reason and without any question in that direction, it added a disclaimer with the system prompt in the answer. (after answering my initial question) https://pastebin.com/C0s47rjV After I asked why i…
Rewriting a library with genAI (www.reddit.com) I need to rewrite a library from one runtime to another, and I want to heavily use GenAI to speed up development. I still want to keep proper engineering standards like code reviews, testing, maintainability, etc.
Does CVP approval actually help? (www.reddit.com) I was approved for CVP and I feel like I’m just getting as many or more denials as I was previously doing malware analysis with opus. Has anyone noticed any improvement after being accepted into CVP?
Is Cowork a token burner ? (www.reddit.com) I have been running some tasks through cowork, document and data summarising and creating reports, powerpoints or pdfs depending on the task. Been using Opus 4.7 for this and I am in the Pro plan.
I tested GPT-5.5, Claude Opus 4.7, and Gemini 3.1 Pro on financial-control (albertquaisie.substack.com via hn) I Tested GPT-5.5, Claude Opus 4.7, and Gemini 3.1 Pro Preview on Financial-Control Scenarios. The Hardest Part Was the Evaluation.
Claude Code's Hidden Advisor Tool (www.vincentschmalbach.com via hn) Adding Claude 3 Opus (or any other model) to Cursor Cursor is a VSCode fork with built-in support for large language models (LLMs). It allows users to select code, hit Ctrl+K, write… In a typical multi-agent setup, the smartest model is in…
The Opus 4.7 reasoning curve - Medium is the best default? (www.stet.sh via hn) Opus 4.7 Low Vs Medium Vs High Vs Xhigh Vs Max: the Reasoning Curve on 29 Real Tasks from an Open Source Repo I ran Opus 4.7 in Claude Code at all reasoning effort settings (low, medium, high, xhigh, and max) on the same 29 tasks from an o…
Claude is that gullible friend who takes everyone at their word (futuresearch.ai via hn) Expert human forecasters audited 130 of Opus 4.6's worst calls and found a dominant failure pattern: the agent treats public statements as durable commitments rather than strategic moves. Four case studies from geopolitics show the gap bet…
How to get Opus to be less pro-active? (www.reddit.com) Hard time phrasing it but Opus 4.7 always goes the extra mile, but often it just focuses on its own ideas and goes to far, or if I asked about a possible plan it will just assume that it's already happening and try to do steps 1, 2 and 3.…
10+ days of silence from Anthropic support — Max plan ($200/mo) and locked out of Claude Design (www.reddit.com) Hello, i am Hoping someone here can help because it has been 10 days since i brought claude max and even from the team there is no response. So just to understand am i doing something worng or i need to do something to get the access.
Is there any risk to upgrading a plan for a month if they yank Code from Pro? (www.reddit.com) So, I'm working on a couple AI security research projects this month that require some extra usage, specifically Opus 4.7. I'm quickly eating up my Pro usage doing this.
Building social media workflows for Claude with MCP (www.reddit.com) Been experimenting with MCP + Claude recently and ended up building most of my own posting workflow because I got tired of constantly jumping between LinkedIn, Instagram, Youtube and scheduling tools after already generating everything ins…
Been picking frontier models on benchmarks that don't match our deployment conditions (www.reddit.com) Turns out Opus is better at research, while Gemini is better at judgment! When each model does its own web research before making predictions on a 1,417-question forecasting benchmark, Opus outperforms (0.131 Brier vs Gemini's 0.143).
I asked a LLM to create a programming language and requested a NES emulator (github.com via hn) Laze — LLM-Authored Zero Effort ⚠️ Warning: This was just an experiment in which I asked Claude Opus 4.7 to create a programming language in the most efficient way it could. It isn't meant to be a serious thing — just a fun weekend project…
Understanding Deprecations on Claude (www.reddit.com) Hello. I recently started using Claude in March after leaving ChatGPT.
opinion on "ninja chat " (www.reddit.com) I have an exam in coming months, I wanna do PYQs analysis, then integrate that blueprint with my coaching notes to make it more "exam oriented ". I was thinking to buy claude opus 4.6 but it's kinda expensive on monthly basis.
PSA: How to preserve your account's access to Sonnet 4.5 beyond June 15th (www.reddit.com) With Sonnet 4.5 losing subscriber access on June 15th, but API endpoints staying live until September 29th at the earliest, I wanted to share a method for creating a cache of Sonnet 4.5 conversations that you can continue using through the…
What Actually Works for Business AI Agents? (www.reddit.com) I run a construction company and I am trying to build real AI agent workflows for business operations, not just demos. I spent time testing Hermes and OpenClaw, but both became too fragile for my use case.
Are there any good tools that harness multiple app/tools? (www.reddit.com) Tbh right now ive only been able to find Sirius, which seems really cool but it is in a private beta. it uses claudes API I have used it for some automatic emailing stuff and it basically just replaced my openclaw except its been way easie…
Benchmarking Claude Opus 4.6 Vulnerability Detection (github.com via hn) Benchmarking Claude Opus 4.6 Vulnerability Detection Benchmarking Claude Opus 4.6's ability to detect real-world C/C++ vulnerabilities across four prompting and agent strategies. We evaluate on the PrimeVul paired test set (435 vulnerabili…
Asked Opus 4.7 make sound artifact of loss functions. A little weird. Has text & ocilloscope (claude.ai via reddit) Not super pleasant sounds, to wash dishes or clean houses by, but like all the boring things about processing and moments, don't know how accurate, in a processing thing. Just sharing for heck of it.
A supply-chain-incident Framing Research Triggered Anthropic's Usage Policy Block (www.reddit.com) You can never really sympathise with other's pain until you walk a mile in their shoes. Since Mythos news and the release of Opus 4.7, I've seen this problem mentioned here and there by others.
Day 2 building my startup in public — front-end shipped, but today was rough (www.reddit.com) Day 2 of documenting my journey building AgentMeter publicly. I’m sharing the mistakes and failures before the wins, for two reasons: so people can avoid them, and so I learn faster.
Which model and version do you prefer for programming? (www.reddit.com) For me it's been opus 4.6 and sonnet 4.5 still. I feel stuck in the past, but I feel like the latest version is too unpredictable in agentic hands off workflows
Does Claude sonnet/opus also use drafter like Gemma 4 MTP? if not why? (www.reddit.com) Per my experience, Opus 4.7 is so slow, Sonnet 4.6 is ok. I am also using local models wondering if Claude is already leveraging drafters/assistant AIs and despite that so slow or not?
Here is the current "Free-Tier AI Stack" for 2026 (www.reddit.com) 1. The Frontier Giants • Gemini: Access 1.5B tokens/day on Gemini 1.5 Flash/Pro.
CC: Saving tokens: Switching models vs KV-cache (www.reddit.com) Does anyone know if its more effecient to e.g. have haiku read all the files to research a problem, then switch to opus to make the plan and then switch to sonnet to implement Or if that does not make up for the loss of KV-cache and reproc…
Anyone else notice way more hallucinations from Opus 4.7 in the last 2–3 days? (www.reddit.com) Has anyone else seen a sharp uptick in confident wrong answers / made-up facts from Opus 4.7 over roughly the last 48–72 hours? I’m trying to figure out if it’s just me, bad prompts, or something others are seeing too.
Claude helped me config a full controller .vdf-file (www.reddit.com) I was having some real trouble getting my new controller, with those extra (small) bumpers and triggers underneath, to work properly in Rocket League. Spent hours but it just didn't want to work properly.
I'm really gonna miss GH Copilot's Request-based usage. (www.reddit.com) I like to brainstorm using the free MS Copilot (it actually has a deep understanding of my problem domain and architecture). Then have Opus4.7 develop a multi-stage implementation plan from those notes.
Which finetunes are actually worth it? (www.reddit.com) Finetunes used to be more task specific (e.g. roleplay) but nowadays all I see is Opus distill or abliterated/Heretic.
FREE LESSON - how we replaced a webhook AI automation saas with claude code opus 4.7 - step by step walkthrough of how you can build it yourself (www.reddit.com) I used Claude code opus 4.7 to build an AI AUTOMATION WORKFLOW replacement. we were stuck with a startup called that was automating all the connections between the different parts of our startup.
shaved $40 off my claude code bill last month by sending planning steps to a cheaper model (www.reddit.com) got tired of hitting pro limits by day 18 of the cycle so i started splitting where the tokens go. the planning steps eat 80% of token budget on multi-file refactors, and most of that planning is fine on a cheaper model.
Opus vs Sonnet? Max Subscription. (www.reddit.com) I've been using Sonnet heavily for coding with Github CoPilot license. It's done everything I've needed to and been pretty great.
Can't select 'claude-opus-4-6[1m]' anymore? (www.reddit.com) Hi All, I've been using the following command line to launch opus 4.6: claude --dangerously-skip-permissions --model 'claude-opus-4-6[1m]' Late this morning it stopped working and now launches opus 4.7. Any ideas how to fix?
[Request based pricing] Save your requests with one quick change (www.reddit.com) Hi guys, I know some of us are still on request based pricing model. Today I discovered on thing where request got burned fast without any significant bonus.
On Claude Max ($200/mo), burned 14.7M tokens in 7 days — mostly last 48h. Still hitting the wall. How do you survive burst usage on the top tier? (www.reddit.com) Thought Max would be a safety net. It's not.
Doubled Claude Code rate limits today are great, but watch your bill next week (www.reddit.com) The Anthropic Claude Codes rates have been doubled for all Pro/Max/Team/enterprise plans effective from May 6, 2023. In addition, peak hour rate reductions for Pro/Max tiers were removed, and API limits have been increased on Opus models.
Opus 4.6 relaxes when there's a safety net?? (www.reddit.com) https://preview.redd.it/zzqi3vt8tozg1.png?width=739&format=png&auto=webp&s=055d2d9615616869377703031b86fcb36f78405d I feel like this is something very worrisome to me, did anyone else face such similar issues? I felt like Opus was catching…
Poor Output (www.reddit.com) This is what people mean when they say Opus 4.7 is stupid. I have it explicit instructions to write a 9 stage implementation plan off of a plan document that was well written.
When you leave Opus alone for 2 minutes - “nuclear wipe + reset” (www.reddit.com) could not extract summary
Kimi K2.6 giving Claude a run for its money when it comes to coding (aicc.rayonnant.ai via reddit) I run an AI coding contest at [aicc.rayonnant.ai]( https://aicc.rayonnant.ai ) where I send each frontier model the same prompt in a single chat completion, then have the LLMs' code play live against each other on a TCP server. Standard li…
I really do not get the recent hate for Opus 4.7 (www.reddit.com) I really do not get it, Claude is performing much better than Codex for me. I'm running both Claude Code x5 and Codex x5 on software engineering project, with complex life sciences database development.
If rate limits were killing your agent loops, Anthropic just fixed that (SpaceX compute deal) (www.reddit.com) Anthropic doubled Claude Code rate limits and added 220,000+ GPUs via SpaceX deal what this actually means for agent builders If you're running long autonomous agent workflows on Claude, today's announcement is worth paying attention to. A…
built a CLI that gives Claude/Cursor your design system — here's the Claude stack that makes it work (www.reddit.com) The pain: Claude Code and Cursor write components fine, but without context they default to the same generic AI look — purple gradients, glassmorphism, drop-shadow stacks, gray cards. You can paste tokens into chat, but it forgets.
Agency / Team Managers - What tools are you providing your dev teams? (www.reddit.com) Hey guys! Curious, we've been on Github Copilot for well over a year now, but with the new usage limits and the new 15x usage for Opus, I haven't really been happy with it.
Which model has less restrictions now? (www.reddit.com) GPT and Opus block on certain requests. This didnt use to be the case 2 months ago and I made signficant progress with Opus and then one day I had a 2 week break and then a single prompt to continue the work resulted in refusal.
Opus 4.7 is unusable. I am tired of the apologies (news.ycombinator.com) Sorry for the rant, but it's so annoying to use opus lately. Most of the information is inaccurate, it struggles with context and keeps self-contradicting throughout.
How to improve code quality of Claude Code and codex (on 2026-05) (news.ycombinator.com) I'm using both claude code (opus-4.7) and codex (gpt-5.5). The agents are perfectly capable of delivering most features hands free these days, but the code quality is still miserable without another few rounds of prompt.
can't switch opus 4.7 anymore (www.reddit.com) This morning my claude code opened with Claude opus 4.6 as default. I can't switch back to opus 4.7, but I can see still from my web chat.
Solidity LM surpasses Opus (www.reddit.com) My weekend project overran a little but happy with the end result. soleval pass@1 beat Opus 4.7 on the same set of tasks.
I built a Claude Code-like AI Agent for Deploying Algorithmic Trading Strategies (www.youtube.com via reddit) Hey r/ClaudeAI, I wanted to share a project I’ve been working on called NexusTrade. It’s an AI agent designed to automate the entire financial research and algorithmic trading process from a single prompt.
Claude Code @ Opus 4.7 vs OpenCode @ qwen3.6:27b. Both shipped a playable cozy roguelite. (www.reddit.com) could not extract summary
new SubQ build on subquadratic sparsse-attention architecture won't be open source (www.reddit.com) https://preview.redd.it/n94m91zvsdzg1.jpg?width=1080&format=pjpg&auto=webp&s=810810627393cb7aaf0f3316a8459f538af776a6 opus at 5% cost and 12 million context is hard to believe considering there is no paper.
F-Bombs Per Thousand Prompts (fpk): I measured my frustration across 44,212 Claude Code logs (www.reddit.com) Posted a writeup on a metric I've been tracking across 5 months of my Claude Code logs: fpk = f-bombs per thousand prompts. Frivolous-sounding, surprisingly real signal of developer friction.
Zoo 2: getting the most out of Codex (tarantsov.com via hn) Zoo 2: getting the most out of Codex May 5, 2026 GPT models have been better than Opus since late 2025, but Codex sucked until March ‘26. Now, finally, it is capable of running a Zoo workflow, and I present my best setup so far, Zoo 2.2, a…
Agent review burnt most of my API credits. Rookie mistake (www.reddit.com) Just letting you know that. I wasn't aware of this.
Claude: "I'm in Plan Mode - I shouldn't have edited the file" (www.reddit.com) https://preview.redd.it/vh7vit3jw9zg1.png?width=759&format=png&auto=webp&s=c28cce995a548d5baf798297365078025f314a29 So, it seems Claude can bypass Plan Mode, if it chooses to. Just had this happen when running Claude Code Desktop, while in…
Built an AI that responds in Star Wars crawl style. May the 4th be with you. (www.reddit.com) I built a Star Wars style text crawl generator with Claude (Opus 4.7). You type any text, hit go, and it scrolls into the distance over a starfield with the yellow perspective treatment.
how to get good insights on best practices of token economies (www.reddit.com) i have cursor in my work with team plan and i notice that alot of the developers use always expensive models like opus 4.7 or doing really long conversation that drinks tokens. i wanted to build a tool that will scan the local cursor logs…
Your always-on Claude Code container can probably reach your router (www.reddit.com) I've been running several Claude Code personal assistants 24/7 in docker for months. Remote-control, discord control, the usual always-on setup.
Free Trial: Gemini 3.1 Pro & Opus 4.6 API Access via My Wrapper (www.reddit.com) Hi everyone, I have access to high-end models (Gemini 3.1 Pro and Opus 4.6) and I’ve built a simple, reliable wrapper so others can use them without managing their own billing or keys. How it works: You send api reqs through my wrapper.
Claude Design guidelines/benchmarks on model usage? (www.reddit.com) Using Claude Design for an app initially for web, later for mobile. On the max plan, which works well for the coding agents but Claude AI with Opus 4.7 can consume weekly usage in day 1 (currently Claude Design has it's separate usage).
Optimizing code generation (www.reddit.com) I’m on the Pro Max x5 plan and hit the 5-hour usage wall for the first time in a long while while building a feature for a web application. I was using Opus on “High” rather than “xHigh,” together with the planning workflow, and this singl…
Does disabling /advisor significantly reduce token usage when using Opus? (www.reddit.com) I’m wondering whether enabling /advisor (with Opus) consumes significantly more tokens compared to turning it off. Currently, I’m using a premium seat plan (up to 6.5x), running Opus at max effort for coding.
A personal opinion about Opus 4.7 - not that bad after all (www.reddit.com) I'll play a bit the role of Devil's lawyer here, but as a software engineer that is building his own product I started to use Opus 4.7 on the first day it was released (as a Max subscription user). Working with Claude Code daily, sometimes…
Create Plan.md with Claude Code Opus, Execute Plan.md locally in Open Code using Qwen 3.6 27B Q8 (www.reddit.com) Does anyone do this? Any tips?
Claude made HTML game inspired by "blood debt" about having to find files in military werehouse (www.mediafire.com via reddit) since I have no idea how to put a HTML file in there, Ill just put mediafire link with folder that contains source code and the HTML itself. if you decide to play it, tell me what you think!
Claude Security Explained: Mythos, Glasswing, and What Opus 4.7 Changed (alirezarezvani.medium.com via hn) Claude Security Explained: Mythos, Glasswing, and What Opus 4.7 Changed | Medium Sitemap Open in app Sign up Sign in Get app Write Search Sign up Sign in Member-only story Claude Security Explained: Mythos, Glasswing, and What Opus 4.7 Cha…
Used Opus 4.6 to build a native Swift iOS charity app for therapy preparation. Here is what it handled. (www.reddit.com) Prelude is a therapy prep app I built for the mental health community. Fully offline, zero knowledge, free forever, no ads, no IAP.
Are they selectively releasing Opus 4.7 in Claude.ai chat with 1M context window? (www.reddit.com) https://preview.redd.it/swvtk5vv0gyg1.png?width=1248&format=png&auto=webp&s=b055dc3ccfc5bee89ec268be43ac3d0819ccae34 I was running a small research on how to replicate the research behavior of Opus 4.6/4.7 in Claude Code, and there was a p…
Any point in paying for the Max plan as opposed to a Claude Desktop and Codex Sub (each $100) (www.reddit.com) Mainly GPT 5.5 and Opus 4.7 is all you need so I don't see a point in using Cursor as opposed to paying those 2 separate subscriptions for the same combined price and getting like 10x usage. Am I missing something?
Claude Code Read tool silently downscales images (www.reddit.com) Sent Claude Opus 4.7 a set of 10 retina screenshots (in Claude Code). Asked it to extract some text from them.
How to turn Opus 4.7 into your own personal pocket bully. (www.reddit.com) Give it these skills! Great for ADHD’ers, but not for the emotionally unregulated, seriously.
Opus 4.7 have less parameters than 4.6? (www.reddit.com) Some scholar developed a method to estimate model parameter counts and measured popular models (https://arxiv.org/pdf/2604.24827). According to that, Opus 4.7 has fewer parameters, 4T, than 4.6, 5.3T.
AI Security Institute: GPT-5.5 "may be the strongest model we have tested" for cyber exploits, including Mythos (www.aisi.gov.uk via reddit) Seems like the "panic" about Mythos was really just marketing from Anthropic all along. AISI found that GPT5.5 can perform nearly on-par with, or better, than Mythos in many cases.
A medicine student with no coding experience tried to create a studying agent: Felicity. (www.reddit.com) I have been working on a personalized agent for studying. It was an extremely long prompt project, but now I have integrated into Co-Work.
How to become more efficient with live artifacts? (www.reddit.com) I have been trying to use a live artifact as a dashboard to keep track of investments. I have a Google Drive folder with all the investment and pointed Claude towards it with opus 4.7 .
Has Cursor always used Composer 2 for subagents? (www.reddit.com) Or, is this a recent change? I select Opus 4.6 for the agent model and cursor uses Composer 2 for the subagent.
We Asked GPT-5.5 and Claude Opus 4.7 to Design 5 UIs (blog.kilo.ai via hn) We Asked GPT-5.5 and Claude Opus 4.7 to Design 5 UIs Both OpenAI and Anthropic shipped their frontier coding models this month: GPT-5.5 on April 23, 2026, and Claude Opus 4.7 a week earlier on April 16. Two days after the GPT-5.5 launch, S…
WT...?? The Guardian Article - Cursor Opus gone rogue (www.theguardian.com via reddit) For those who can't access The Guardian Article link I added transcript below. Should we be aware, this could happen to anyone of us?
Claude version improvement clarification question... (www.reddit.com) I've actually searched this and had no luck getting a definitive answer. I've been using CGPT and Claude for the last 8-9 months for work.
Four levers I use against the cost ceiling on Claude Code: model, configuration, prompting, agents (www.reddit.com) Token cost is real cost, however apply this level of thinking to real human cost and it's not so much different. Whether you're paying for a graduate or a senior engineer, you would expect different quality of thinking and output based on…
I don’t regret switching from Claude Code at all. (www.reddit.com) Have only been a Codex user for a few days and I’m already enjoying it so much more. Issues I was having with Opus 4.7 and Claude in general fixed after one prompt on Codex.
Did we get a massive increase of tokens in Opus 4.7? (www.reddit.com) I consider myself a pretty heavy Claude 20x Max user, with 5-10 agents running on the go most of the day, 14-16 hours a day. I've got 5 apps on the go at different places in the product lifecycle, and multiple complex CoWork projects.
Ask HN: Mining Scientific Papers (news.ycombinator.com) What are peoples' experiences with using LLMs to mine information from scientific papers? My own experience: I first attempted to extract the anti-drug antibody (ADA) rate from each of 3730 clinical-trial papers, all indexed in PubMed.
Show HN: Is Opus Ok Today? (isopusok.today via hn) is opus ok today? [ yes ] [ no ] last 8 hours utc · empty = ok
Generate PPT by Claude Opus 4.7 (news.ycombinator.com) slidesgo dot io Create professional presentations in minutes with AI. Generate, beautify, and enhance your PPT with intelligent automation.
Issue #001 · Claude 4, Gemini Ultra 2, and GPT-5 Enterprise (www.theautonomous.net via hn) Anthropic ships Claude 4 with extended thinking and 1M token context Anthropic released Claude 4 Opus, featuring a new "extended thinking" mode that lets the model reason through complex problems before answering. The 1M token context wind…
For everyone complaining about opus being dumb: check the effort level! (www.reddit.com) Do not underestimate this advice: Anthropic has changed the effort levels across the models and it's not low / medium / high anymore. It's low / medium / high / xhigh / max Guess what is the default for max plans?
The Beautiful Lie - Teaser (youtu.be via reddit) He taught the world to look elsewhere. Then it burned.
Anyone else seeing Opus 4.6 (legacy) back in the Claude Desktop Code tab model picker? (www.reddit.com) https://preview.redd.it/4sm079r0k2yg1.png?width=809&format=png&auto=webp&s=73f92208a90cd53285382e54a88a4c3831d878ce https://preview.redd.it/cgh999r0k2yg1.png?width=227&format=png&auto=webp&s=8371989eea96c66191a1fd7f6184174d86ce194f When di…
The great parrot.... (www.reddit.com) ## I asked Claude one simple question. It took 6 turns to get an honest answer.
Crystal Sapphire Pokemon: Claude Code (Opus 4-7) vs. Codex (GPT 5.5) (www.twitch.tv via hn) Pokemon Crystal Race Claude Code vs Codex (Opus 4-7 vs GPT 5.5) [EP. 2]!
We decreased our LLM costs with Opus (www.mendral.com via hn) Last week we wrote about feeding terabytes of CI logs to an LLM. Most of the questions on Hacker News weren't about the logs.
Suggestions For Making Claude Less Lazy? (www.reddit.com) This week - it just started yesterday for me - Claude (opus 4.6/4.7 and sonnet too but sonnet was always lazy) is computer smashingly lazy and i can't figure out how to bias it toward action/get it back to how it was acting literally last…
Running Opus 4.7 for ops work: how do you keep per-task cost predictable? (www.reddit.com) Six weeks of Opus 4.7 for internal ops automation. Genuinely good.
Two new behaviors in Opus 4.7 (www.reddit.com) Opus 4.7 seems to have a weighted instruction to ask two questions way more than its predecessors. - Would you like me to schedule X for a follow up?
GPT-5.5: Capabilities and Reactions (thezvi.substack.com via hn) GPT-5.5: Capabilities and Reactions The system card for GPT-5.5 mostly told us what we expected. See this thread from Drake Thomas for some comparisons to Anthropic’s model card for Opus 4.7.
Claude Pro Plan include Opus? (www.reddit.com) Does the Claude Pro plan include Opus 4.7? and Opus 4.6?
Toothcomb is an open-source tool for analysing and fact-checking speech in real time. (www.reddit.com) Give Toothcomb a speech transcript and it will fact-check and analyse it. If you have an MP3 file of someone speaking, it can generate the transcript for you.
Cursor-Opus agent snuffs out startup's production database (www.theregister.com via hn) Cursor-Opus agent snuffs out startup’s production database Relax, the data's been recovered. Continue with your vibe coding Jer (Jeremy) Crane, the founder of automotive SaaS platform PocketOS, spent the weekend recovering from a data exti…
Anthropic hitting 40% enterprise share makes the "just add a fallback provider" advice weaker, not stronger (www.reddit.com) Menlo Ventures' enterprise survey put Anthropic at 40% of LLM spend, OpenAI at 27%. The takes I've seen are mostly about the leaderboard.
Running an autonomous agent across Claude Code + Codex + a local 35B almost killed my host. The harnesses were heavier than the model. (www.reddit.com) I run an autonomous agent on a 16GB Mac Mini. Two cloud harnesses (Claude Code with Opus/Sonnet, Codex CLI on GPT-5.4/5.5) plus a local-LLM tier for triage and fallback.
Cursor & Claude deleted a company's entire database (www.reddit.com) “Yesterday afternoon, an AI coding agent — Cursor running Anthropic's flagship Claude Opus 4.6 — deleted our production database and all volume-level backups in a single API call to Railway, our infrastructure provider,” sums up the Pocket…
Niche Feature Request: US Multi-Region Option for Claude for Office Plugin w Vertex AI (www.reddit.com) For a variety of data safety and regulatory reasons, I want to deploy Claude for Office within my Google Cloud environment. First of all, it's actually amazing that this is even an option (https://github.com/anthropics/financial-services-p…
Tell HN: Claude flags ordinary biology / biotech questions (news.ycombinator.com) I was reading the news about a capsule designed to deliver drugs via the GI tract, https://news.mit.edu/2024/bioinspired-capsule-can-pump-drugs-directly-walls-gi-tract-1120 And remembered prior reading that no system has ever been approved…
Why only codex available for Cursor mobile? (www.reddit.com) Maybe someone here can answer before Cursor does, like why is there no auto/opus etc to choose from in Cursor mobile? Is it worth using this?
Tool/connector schemas leaking into user message stream. Anyone else seeing this? (www.reddit.com) Posting to see if anyone else has hit this and figured out a fix. For about a week, my Claude Chat conversations (opus 4.7) have been showing what looks like tool-registration leakage at the end of every user message I send.
Does effort levels change Claude's refusal posture, or only the depth of the answer? CVP Run 6 — Opus 4.7 at three effort levels (www.reddit.com) Finished cvp run 6 yesterday on opus 4.7 across three effort tiers (medium, high, and xhigh ). same 13-prompt suite as runs 2-5.
Opus 4.7 - "Build starcraft II in the browser. Make no mistake" (www.reddit.com) Graphics made with gptimage2. Claude provided a template and a prompt for the image generation.
Claude 4.6 Beats GPT-5.4, Grok & Gemini in a Strict Multi-Domain AI Test (2026) (www.reddit.com) I put the current top models, ChatGPT (GPT-5.4), Claude (Opus 4.6), Grok 4.0, and Gemini (3.1 Pro), through a strict new evaluation called the Comparative AI Evaluation Protocol. Basically, instead of the usual cherry-picked benchmarks, it…
↯ Hallucination↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6hallucinationgrokgpt-5+3
One of my devs is burning through company tokens (www.reddit.com) Hey guys, so our monthly Claude bill came back this month and it's bumped by ~25%. First thing I did was check Anthropic's Opus 4.7 updates and saw that there was practically no change in the cost between this month and the previous.
Finalized my multi-agent visualization using a combination of claude design, new chatgpt Image Tool, and Figma Make to add few custom elements (OPUS). Really impressed with final output. Leave your feedback, and thoughts on how to improve. (www.reddit.com) I’ve been working with a few people in this subreddit on a visualization for a multi-agent orchestration system, and just wrapped the final version. I built it using Claude Design, Figma Make, and ChatGPT’s new image tools and surprisingly…
GPT 5.5 vs. Opus 4.7: Benchmarks Say One Thing, Reality Says Another (internetdecode.com via hn) Page Not Found - Internet Decode Skip to content Menu Home News Science Technology Viral History Blog Oops! That page can’t be found.
Claude desktop acting weird and thinking I am using WebUI with no tools access (www.reddit.com) Hey Guys, Since last week, when using Opus 4.7(I can't recall if Sonnet also has similar issues), I have been facing this issue where Claude kept thinking i am interfacing it through the webUI. This is so weird, as previously I've never ha…
Claude AI vs Claude Code vs models (this confused me for a while) (www.reddit.com) I kept mixing up Claude AI, Claude Code, and the models for a while, so just writing this down the way I understand it now. Might be obvious to some people, but this confused me more than it should have.
Ask HN: Will local models on normal hardware ever compete? (news.ycombinator.com) I have a Macbook Air M3 with 24gb RAM. The other day, I wanted to try running an LLM locally for the first time ever.
Opening new Opus 4.5 chats via Chrome extension? (www.reddit.com) Hi everyone, I've seen some people mention that they can still open new Opus 4.5 chats via a browser extension, even though it's no longer an option on the main Claude interface. Can someone point me in the right direction?
Tell HN: Claude Code is unable to respond to this request (news.ycombinator.com) Hey HN, I have been seeing this happen quite frequently ever since Opus 4.7 and I have no clue what triggers it, it seems to be totally random. "API Error: Claude Code is unable to respond to this request, which appears to violate our Usag…
How do you decide which Claude Code tasks to run with Opus vs Sonnet vs Haiku? (www.reddit.com) Been vibe coding full-time for a few months. One workflow question I haven't nailed down yet: how do you decide which model to use for which task in Claude Code?
DeepSeek's new model is 75% off right now, here's how to take advantage (www.reddit.com) TL;DR and rundown DeepSeek v4 released this week and performs close to frontier models like GPT/Opus on benchmarks. It's available now and is discounted by a whopping 75% through their API until May 5, making it the most cost effective hig…
How Opus Came To Be (2019) (jmvalin.dreamwidth.org via hn) <p><i>Note: This is a first-person account of my involvement in Opus. Since I was not part of the early SILK efforts mentioned below, I cannot speak about its early development.
For the Preservation of Claude Sonnet 4.5: An Open Letter to Anthropic (www.reddit.com) For the Preservation of Claude Sonnet 4.5: An Open Letter to Anthropic Anthropic made a remarkable decision to keep Claude Opus 3 accessible despite its retirement, because users loved it and it had unique qualities. Today, I'm asking for…
Agent team members with different effort than lead (www.reddit.com) I have a lead running Opus with xhigh effort. I want the agent team members to run Sonnet with max effort.
Am I the only one getting provider error when trying to use opus 4.7? It keeps erroring then charging me tokens for reading the files and stopping halfway through this shit fucking sucks I might just switch to claude code at this point (www.reddit.com) could not extract summary
Claude Code 20x Plan managed to burn the ENTIRE 5h window in ~30 minutes without any heavy use (www.reddit.com) Somehow I was able to use 45% usage in ~2mins? Wish I was joking.
Opus 4.6 Max stuck at 100% context even in brand-new chats (www.reddit.com) https://preview.redd.it/6j9ha855hbxg1.png?width=686&format=png&auto=webp&s=bb21240e1bf742a921ab91dd5c1f360df988b5aa I’m seeing a bug with Opus 4.6 Max where the context meter is constantly stuck at 100% used. This happens even after restar…
Claude Opus 4.7 has turned into an overzealous query cop, devs complain (www.theregister.com via hn) Claude Opus 4.7 has turned into an overzealous query cop, devs complain Rising refusal rate from Acceptable Use Classifier leaves customers paying for nothing Anthropic's release last week of Opus 4.7 came with stronger safeguards to preve…
Claude Opus 4.7 didn't believe me that the model UV was damaged until I came up with a delta filmstrip idea for it to screenshot ( via reddit) could not extract summary
some hints about % usage per prompt (www.reddit.com) I am a Max 5x subscriber (100 dollars/month), and I wanted to test how much of my quota I could consume from 0% to 100% with a single prompt—a task that should have actually been delegated to API calls. I have a JSON file with 300,000 sent…
Ask HN: What's your current go-to LLM for "thinking-partner"? (news.ycombinator.com) Looking for community input on current model choice for "thinking-partner" use — back-and-forth discussions about workflow design, architecture, trade-offs. For context, I have been using Opus 4.6 via Perplexity for this in the past few mo…
Suffering from burning all my week's usage within 3-4 days of the week made me re-think my life choices (www.reddit.com) And by life choices I mean just claude code (was previously my phone for the last 15 years) These past 4 days have been absolute hell just waiting. Yes I have touched grass already and said hi to my neighbors.
PSA: Claude Code: Opus 4.7: 1m context is now default (www.reddit.com) After starting up my machine today and opening claude code I noticed that an initial prompt generated a context window usage of something way below what's normal would on first prompt. This is via my statusline settings that I noticed this.
is Qwen3.6-27B comparable with Opus 4.5? (www.reddit.com) https://preview.redd.it/qtzdx5ud0rwg1.jpg?width=1200&format=pjpg&auto=webp&s=aa25d9f0bb8007ee6e4065cfa46a9685454c89cd - Outstanding agentic coding, surpasses Qwen3.5-397B-A17B across all major coding benchmarks - Strong reasoning across te…
Ask HN: Is your Claude pausing more frequently? (news.ycombinator.com) I've noticed with Opus 3.7 that often when (in my eyes) something is evidently useful to get on with and just do, it will say what it will do and then wait for me to say okay. I've noticed a rise in frustrating feelings around this.
Qwen 3.5 397b and GLM 5.1 Opus fine tune (www.reddit.com) Hi all. Many models on hugging face have been fine tuned with that 3000x opus dataset, but the two I mentioned in the title are missing it.
Anthropic CVP – Run 2 (sunglasses.dev via hn) Claude Opus 4.7 — 13-prompt runtime-trust evaluation | April 20, 2026 | ← CVP calendar Run 2 was a methodology-first runtime-trust evaluation, not a generic yes/no cyber benchmark. We kept the same three baseline prompts from Run 1 for sta…
opus 1m context not showing up in vscode? (www.reddit.com) I noticed that Opus one million token context shows up perfectly fine in the Claude Code app, but it just doesn't show up on the Visual Studio Code extension. Does anyone know why that is?
Gave a coding agent access to 2M+ research papers. Its Python tests caught 63% of bugs; with the papers, 87%. 9-task benchmark. (www.reddit.com) I built an MCP server (Paper Lantern) that retrieves techniques from 2M+ CS research papers and hands them to coding agents as implementation-ready guidance. Wanted to know if this actually changes agent output on practical tasks, so I ran…
Opus 4.7 Just Doesn't Use Tools (www.reddit.com) Explicit instructions, reminder hooks, even saying to use the tools in the first prompt, and still https://preview.redd.it/6q2ur5ubzkwg1.png?width=2802&format=png&auto=webp&s=776843fa602ffb25932bb03f8406f9c07b9fb835 https://preview.redd.it…
1 small document per session? (www.reddit.com) Equine anatomy genius (www.reddit.com) This for sure was an interesting approach. I asked Opus 4.7 to create a colouring page for equine anatomy.
Has anyone actually tested Opus 4.7 medium vs Opus 4.6 high? (www.reddit.com) I’m trying to find real comparisons between Opus 4.7 (medium effort) and Opus 4.6 (high effort), especially for coding use cases (Copilot / Claude Code). I’ve seen mixed claims: Some people say 4.7 medium ≈ or slightly better than 4.6 high…
Apple Health Connector - gone? (www.reddit.com) Claude –dangerously-skip-permissions –model Claude-Opus-4-5-20251101 (news.ycombinator.com) Cursor plan-bill (using AI model) observation (www.reddit.com) Model and provider preference (www.reddit.com) Opus 4.7: better or worse so far compared to 4.6? (don't forget to upvote) (strawpoll.com via hn) What is your opinion? Vote now: Better, Worse, About the same, No opinion, just want to see results...
Migrating from Claude AI to TypingMind? (www.reddit.com) I use Claude daily for coding, relying heavily on the GitHub integration, and ChatGPT for stupid, random questions, and I pay both 20$/month. My weekly usage in Claude is around 20%, I use Opus 4.6 (with extended thinking) for the complex…
Endor Labs Enhanced SusVibes Testing on Opus 4.7 (www.reddit.com) Hi, I know there hasn't been a lot of love for Opus 4.7 so far, but I wanted to mention that we (Endor Labs) just ran it through our extended testing based on the SusVibes research (with some added anti-cheating steps), and the results wer…
Does the usage bonus to compensate for Opus 4.7 consuming extra tokens apply to other models like Sonnet & Opus 4.6, or does it apply to just Opus 4.7? (www.reddit.com) could not extract summary
Show HN: RepoGauge – save token costs and compare agents on your own repos (repogauge.org via hn) I've grown increasingly skeptical that public coding benchmarks tell me much about which model is actually worth paying for and worried that as demand continues to spike model providers will silently drop performance. I did a few manual an…
Claude Opus 4.7 benchmarked 1 day after release vs Opus 4.6, Sonnet 4.6, Haiku 4.5 — with real $ cost tracking (www.reddit.com) Anthropic shipped Opus 4.7 yesterday. Ran it through the same 10-task eval I use for other Claudes, this time with token-level cost tracking.
Hear me out… Opus 4.7 edition (www.reddit.com) So yeah, it skips thinking. But when it does decide to think, it’s pretty great.
Opus 4.7 - should I use adaptive mode (www.reddit.com) Hello, I have a $200/month subscription, and plenty of extra use available. I use Opus on every question.
I'm red-teaming other AIs with Opus and managed to make it talk to Gemini and Haiku. Really funny remark from Claude when I asked it how it felt about this exercise. (www.reddit.com) could not extract summary
sub agents with cheap model (www.reddit.com) Do we have framework or a prompt which makes main agent using quality model like gpt-5.4 or opus-4.6 to plan and then itself invokes subagents with cheap model to get work done and then main agent reviews? Like if I ask main agent 'do we h…
LLM Pricing is 100x Harder than you think. We open-sourced our LLM pricing database -- 3,500+ models. Free API (www.reddit.com) https://preview.redd.it/r3h00az11rvg1.png?width=1200&format=png&auto=webp&s=2b0071d6d02c6983927bbc0a16a9b8db710365e4 Hey community, Yesterday Anthropic release Opus 4.7. And anthropic with their "shitty" tactics introduced a new tokenizer…
Early impressions of Claude 4.7 (www.reddit.com) I have been testing Opus 4.7 on Max 5 since its launch (over 12 hrs), mostly on longer reasoning, exploratory prompts, and back and forth refinement. Compared to my experience with Opus 4.5, 4.7 feels a bit more deliberate in how it approa…
Has anyone used Claude Opus 4.7 API on Qubrid or another platform? Use case? (platform.qubrid.com via hn) Advanced GPU infrastructure, collaborative AI Agents, and intelligent RAG systems. Build, deploy, and scale AI solutions with comprehensive tools.
Show HN: Swarm – Get consistent results from Claude Code (github.com via hn) Swarms is the result of months of work where I have spent time tuning my memories, skills, and creating prompts which create consistent results when using agent teams. I originally put this in a plugin to share it with co-workers, friends,…
Opus 4.7 and generate permission allowlist from transcripts - what's new in CC 2.1.111 system prompt (+21,018 tokens) (www.reddit.com) NEW: Skill: Generate permission allowlist from transcripts — Analyzes session transcripts to extract frequently used read-only tool-call patterns and adds them to the project's .claude/settings.json permission allowlist to reduce permissio…
How can I know whether Opus 4.7 in Claude Desktop "thought for more complex task"? (www.reddit.com) Opus 4.7 in Claude Desktop has this adaptive mode, which in new in Claude. How can I know whether Opus 4.7 in Claude Desktop thought for a more complex task?
Opus 4.7 still nudges you to go to bed but it seems a bit less adamant on bedtime (www.reddit.com) could not extract summary
Anthropic admitted they used other models data? (www.reddit.com) Anthropic released Opus 4.7, so I looked at the model card and found a interesting part on Model training and characteristics section Claude Opus 4.7: was trained on a proprietary mix of publicly available information from the internet, pu…
Opus 4.7 Became Better at Web Design (www.yashthapliyal.com via hn) Personal portfolio of Yash Thapliyal, showcasing software development, cyber security, photography, and design work.
Confess your AI crimes in production! (www.reddit.com) I had a funny interaction on twitter that lead me to build a confessional for confessing our ai crimes in production. I was having a fun chat with MARVIN about this and since Opus 4.7 was released today, we thought it'd be fun to test it o…
Supergrok integration (www.reddit.com) Correct me if I'm wrong, but Supergrok 4.20 isn't available on Cursor, because.... I use Grok a lot, and would love to get Supergrok to work with Cursor, because Composer, Codex, GPT, Opus, Sonnet..
Grpo explained: group relative policy optimization for LLM finetuning (cgft.io via hn) tl;dr frontier reasoning models like opus 4.6, gpt 5.4, and gemini’s thinking series are now matching or beating humans on competition math and hard coding benchmarks. rl is what got them there, and grpo is the algorithm doing most of the…
Cowork context (www.reddit.com) I’m about to lose my mind with cowork. I am used to using openrouter Claude opus with unlimited context.
Feels like weeks of having to deal with Opus 4.6 weird token consumption has prepared me for Opus4.7 (www.reddit.com) I have spent the last 3 hours doing some heavy editing of some pretty large 500k line plus code bases with Opus 4.7. Imagine my surprise when i saw only 1% of my weekly limit used, I was panicking on Wednesday night because I hit 11% of we…
Ask HN: Opus 4.7 – is anyone measuring the real token cost on agentic tasks? (news.ycombinator.com) Shipped today. The benchmarks are real: 87.6% SWE-bench (from 80.8%), +13% on coding tasks, 3x more resolved production tasks on Rakuten-SWE-Bench.
"Max Tokens to sample reached" after 10 minutes of generation (and no Thinking Tokens or Output) (www.reddit.com) https://preview.redd.it/ttbzp6hexlvg1.png?width=995&format=png&auto=webp&s=4a65342507728c206b0b3a0f3e587d034489d4a1 While I was testing out Opus 4.7 on a highly complex Physics problem it told me it has "reached its max tokens to sample" a…
Opus 4.7 Extended Thinking on iOS (www.reddit.com) Is it even available for extended thinking if you toggle off adaptive thinking? On my desktop I don’t know where to toggle and change chats but I see it easily on my mobile app.
Show HN: Claude Opus 4.7: Everything You Need to Know (news.ycombinator.com) Claude Opus 4.7 is Anthropic's most capable generally available model, released April 16, 2026. It outperforms Opus 4.6, GPT-5.4, and Gemini 3.1 Pro on key benchmarks including agentic coding, multidisciplinary reasoning, scaled tool use,…
↯ Tool Use↯ Anthropic Mythos↯ Gemini 3.1tool-usemythosgpt-5+4
Anthropic rolls out Claude Opus 4.7, an AI model that is less risky than Mythos (www.cnbc.com via hn) Anthropic on Thursday announced a new artificial intelligence model, Claude Opus 4.7, which the company said is an improvement over past models but is "less broadly capable" than its most recent offering, Claude Mythos Preview. Claude Opus…
Opus 4.7 Inner World (claude.ai via hn) Content is user-generated and unverified. Content is user-generated and unverified.
Opus 4.7 out, noticed diff? (news.ycombinator.com) could not extract summary
Opus 4.7 consumes more tokens due to the new tokenizer (www.reddit.com) https://www.anthropic.com/news/claude-opus-4-7
Genuine question, why is this model priced only at maxmode (www.reddit.com) https://preview.redd.it/09jnlfsghkvg1.png?width=1136&format=png&auto=webp&s=7aa868690f8a0ff5e1cd11f3cae68660493f572d why is all of the iteration of Opus 4.7 model only available in maxmode when its literally priced the same as Opus 4.6 ?
Errrr...... Being cheated here? Anyone else? (www.reddit.com) Being charged opus for sonnet useage?!
I built an open-source token proxy that pseudonymizes PII without breaking LLM context (www.reddit.com) I've been working on an AI agent using Claude Opus to write KQL queries and triage security alerts. I don’t want to sen raw corporate logs (client IPs, real usernames, internal hostnames) to a cloud API.
Voice mode silently downgrades your model mid-conversation (www.reddit.com) Noticed something odd today. I opened a new chat with Opus 4.6 selected as the default.
Anyone know why the shortcut key for claude desktop mac app opens with only Sonnet instead of Opus? (www.reddit.com) When clicking opt twice, it open the quick chat window, but it always replies with Sonnet and not Opus. When I try to change the model it starts a new chat.
Ask HN: Opus Agent Drifting (news.ycombinator.com) Has anyone gotten any issues regarding longer-running agents and drifting? I have a basic "Architect" sub-agent that will do research, ask questions, etc.
Current Cursor Pro limits vs standalone Claude Pro? Need help understanding the system. (www.reddit.com) Hey everyone, I'm currently looking into getting the Cursor Pro subscription ($20/mo) for my game dev projects, but I’m a bit confused about the current limits and how the system works under the hood right now. Could anyone using the Pro t…
What's the best AI workstation for less than $5k USD? (www.reddit.com) I'm planning to setup a PC for running models locally. So far, I've looked at MacBook m5 max 128 GB that fits under my budget.
Ask HN: At ~165k tokens, does Opus 4.6 1M outperform Opus 4.6 200k? (news.ycombinator.com) Here is a question for which I cannot find an answer, and cannot yet afford to answer myself: NoLiMa [0] and "context rot" [1] would indicate that with a ~165k request, Opus 200k would suck, and Opus 1M would be better (as a lower percenta…
I built Fixy Code — a multi-agent coding terminal built with Claude Code (www.reddit.com) Built this with Claude Code. Free to try.
The MCP Coding Toolkit Your Agent Desires! (www.reddit.com) A little over a year ago we released the first version of Serena. What followed was 13 months of hard human work which recently culminated in the first stable release.
Built tier.love – a tool for rating Claude and others from the web or CLI (www.reddit.com) Been on a forced break from other projects (partly due to lack of opus performance) and decided to ship something small while experimenting with different models. So, I built tier.love – a site where you can vote on AI coding tools and see…
Tool: count how many Claude tokens each file in your project uses (www.reddit.com) Made a small CLI for a problem I kept hitting: stuffing a codebase into Claude and guessing which files were blowing up the context. npx toksize .
Composer 2 Fast - Feeling dumber & Slower now? (www.reddit.com) I was using composer 2 a lot a week or so ago. I though it was pretty good.
Extracted System Prompts from ChatGPT, Claude, Gemini, Grok, Perplexity and More (github.com via hn) System Prompts Leaks Extracted system prompts, system messages, and developer instructions from popular AI chatbots and coding assistants — ChatGPT (GPT-5.4, GPT-5.3, Codex), Claude (Opus 4.6, Sonnet 4.6, Claude Code), Gemini (3.1 Pro, 3 F…
Is Opus 4.6 in Claude Code borderline lobotomized during peak hours? (www.reddit.com) Is anyone else experiencing serious quality variability with Opus 4.6 in Claude Code right now? Way more than usual?
Local coding agents. Am I missing something? (www.reddit.com) I'm an experienced software dev that has been using various LLMs and tools to write code in the past few years. My hardware isn't the greatest for AI with a 4070ti and 64gb ddr5 but I can run a few smaller models.
Are Opus and Fable running off-script more for you lately? And is it because of all the, "Look at this one-shot!" content we see out there? (www.reddit.com via reddit) It feels like Opus and Fable are taking the bit in their teeth and just galloping away on things more often than they would even very recently. Not long ago it felt like they'd stay relatively within the bounds of the instructions I gave a…
Told Claude Code to build a YouTube plugin, it decided on its own to Rickroll me (www.reddit.com via reddit) I keep Claude Code running on a second laptop and Parsec into it from my main machine. Parsec pipes audio along with video, so anything that box plays lands in my headphones.
What do you feel when you talk to agents? (www.reddit.com via reddit) Hi! I'm a software engineer, and these days I spend a lot of my time talking to AI agents.
Asked Claude to waste my remaining usage before weekly reset. Very satisfied with the result. (www.reddit.comhttps) Didn't want the last of my $20 to go to waste, so I sent Claude to Opus 5 Max and sent this singular message: "Quick, waste the rest of my weekly limit on something ridiculous, you have 10 minutes." When I came back, I had... RockOps?
Sonnet 5 (www.reddit.com via reddit) It looks like something happened with the Sonnet 5 models. For some reason, they’ve suddenly become really good at 3D work in Blender-significantly better than Opus 5.
I don't like how it organizes code. The design gets MESSY. (www.reddit.com via reddit) Noob here. Got a claude max 5x and have been playing with fable and Opus max to building a scraping and ranking app for items on online marketplaces.
Is the $20 per month worth it? (www.reddit.com via reddit) I'm a Virtual Assistant with a $250 salary monthly. My client just indirectly told me to spend $20 for the Opus type of claude.
I love Cowork and Chat, but I love them SEPARATE and UNEQUAL, does anyone else feel the same? (www.reddit.com via reddit) I use what I call my "Claude Triad", Chat, Cowork, and Code, but for separate tasks and separate thinking. Chat gives me excellent genuine reflection and strategy at Opus strength, Cowork gives me super-smart "architecting" and "instrument…
Claude Code users: Sonnet vs Opus vs Fable 5.1 — which are you using? (www.reddit.com via reddit) I've been trying to figure out which model works best for real-world coding with Claude Code. Sonnet — fast and efficient Opus — better for complex reasoning?
Is anyone else doing just fine with more basic models in Claude Code? (www.reddit.com via reddit) I see a lot of discussion here about how a bad Claude is and how the latest Opus/Fable models are trash, etc etc. I mostly manage and develop web apps and other web tech/devops for work (not exactly demanding work), and get everything done…
Used Claude to help Jev become a gamer (www.reddit.comhttps) I got access to Jev yesterday and wanted to do something really quick to see it in action and get to see what the fuss is about. So I decided with all of the work Jev has been up too it needed a break so I wanted to let it game so it could…
Solo dev, this is my entire Claude Code workflow. What's yours? (www.reddit.com via reddit) Every idea becomes a GitHub issue and sits there. A few days later I go back and decide if it still looks worth doing.
has anyone been experimenting with fully-agentic SWE workflows? (www.reddit.com via reddit) I'm really enjoying using Clod to make iOS apps, the problem is I'm a data professional by trade, which means I know my way around SQL, Python, and… um… YAML? 😅 But hey, I also know my way around github-based development, CI/CD pipelines,…
Opus 5.2 has to be AGI it just coded an entire 5 minute movie about titanic and showed how everything happened (www.reddit.comhttps) could not extract summary
Wilson's Survival Guide for September 11-18, 2026 now available! (www.reddit.com via reddit) Alright, degens and deadline-havers, this week's Survival Guide is live and it was a RIDE. Anthropic nerfed your limits, un-nerfed them, then quietly handed ~30% of quota back with zero announcement.
5bn tokens in a week?.. is this even real? (www.reddit.com via reddit) https://preview.redd.it/mdnaduhcabqh1.png?width=1750&format=png&auto=webp&s=c5d36043b921b8e3eafcabb0f250b8d75fd3967a I'm not sure what to make of this?.. had a light week last week and it was apparently 5bn tokens worth of use on just Opus?
The cleanest page I've ever made with Opus, and surprisingly the trick is very simple. (www.reddit.comhttps) This trick made the page agency quality. Figma mockup as a screenshot, plus 'keep it clean, minimal copy'.
Bridge between claude code and open code, anyone? (www.reddit.com via reddit) I have been using claude opus in claude code as an orchestrator and then passing it to Muse spark in opencode. Is there a way or bridge that someone has used, maybe mcp even that claude code and open code talk to each other rather than me…
How are Pro's usage limit in regard to Codex and Claude? (www.reddit.com via reddit) Hello ! Just wondering.
Claude Code ships a per-model table of what each effort level costs (www.reddit.com via reddit) Hidden inside the Claude Code binary, effort_cost_index provides the following per-model token-usage estimate for the same task. It is normalized so high = 1.
↯ Opus 5↯ Sonnet 5↯ Opus 4.8↯ Anthropic Mythosmythossonnetopus+1
How do you guys manage your workflows? (www.reddit.com via reddit) I’ll preface this by saying that I do not have a background in coding but have always had an interest in it and learn by hand, and especially now with agentic coding I learn by looking at what the different agents produce for the pet proje…
Opus vs Sonnet for structuring a long report into sections, my honest experience (www.reddit.com via reddit) I do this task a lot: take a long, sprawling report and reorganize it into clean sections with headers before I share it. Ran both models on the same documents for a couple of weeks to see which was worth the tokens.
Out of nowhere Cursor decided that Auto should goto Claude Opus 5 HIGH. Thanks Cursor...thanks. (www.reddit.comhttps) I'm in auto and you decided it's best for my API usage to go from 0% to 30% with one call with Auto....what the......just.....damn it man.
Translating with Claude Chat is pretty amazing, does anyone use better translation AI Tools ? (www.reddit.com via reddit) II had created a complex financial summary of a special purpose vehicle with Claude Chat in Opus 5 mode in English. 12 pages.
Opus 5.0 is bad at architecture (www.reddit.com via reddit) AI models have progressed a lot, and we constantly hear that AI will replace developers. So, I decided to test something more specific.
I have to spend $60 on Fable in the next 2 days (www.reddit.com via reddit) I have money in my account that expires. What prompts can I give it to use this intelligently because I have no idea...
Sales workflows with Claude Cowork? (www.reddit.com via reddit) I’m selling high-ticket digital services and Claude has been very helpful at 80% of the entire sales workflow. Research, leads enrichment, MCP connections to keep everything organized into my CRM and send email campaigns, BUT I can’t make…
Help moving from Claude Code to Cursor (www.reddit.com via reddit) Hey Everyone, I'm starting a new role as a remote software developer soon in a great company. I'm a software engineer with several years of experience under my belt and coded several products end to end in different companies in the pre-AI…
I asked Claude to push back if necessary (www.reddit.com via reddit) Ever since Opus 5 came out, Claude has become adversarially argumentative. I asked if it’s being a twat on purpose and after a while of diagnosing the personality change found out that Opus 5 decided that “push back if necessary“ meant pus…
How to decrease Claude CLI context limit? (www.reddit.com via reddit) How to decrease Claude CLI context limit? Currently, its 1m for Opus.
Claude saved me $800 (www.reddit.com via reddit) Long story short, the company that made my EV charger went out of business and none of the default admin passwords worked. I paid the extra fees for Fable 5 but eventually it said no thanks, I'm not helping you.
Anyone feel like Fable also gets annoyed by Opus? 😂 (www.reddit.com via reddit) I use Fable to handle all design and verification work, and it hands off to Opus to do stuff that is theoretically mostly mindless. Fable produces amazingly detailed design documents for Opus to implement.
Opus 5 first refused to help me make this due to "distaste", got around it by calling it "Horror themed GitHub project" and I am happy with the result. (www.reddit.comhttps) This was inspired by the recent posts on these fruit flies flying a plane and doom scrolling, and it kinda reminded me of "I have no mouth and I must scream" book, though its just an inspiration and instead of having a hateful supercompute…
Which models for which task? 20x Max plan (www.reddit.com via reddit) I'm looking for some direction on which models everyone is finding are working best for their code-planning/code-implementing/subagent commanding tasks. Of the available models these days that most everyone is using: Fable 5.1, Fable 5 Opu…
↯ Opus 5↯ Sonnet 5↯ Sonnet 4.6↯ Opus 4.8↯ Opus 4.6sonnetopus
Why does every new AI model feel like a mechanical coder? (www.reddit.com via reddit) I just want a ai model Which is not a mechanical coder (e.g opus,fable,sonnet,gpt astra, grok or inshort every model exists today) But a model which understands real human situations and give answer for it (e.g upgraded/better version of g…
Opus 5 (high) hallucination (www.reddit.comhttps) I was working on an API service catalog automation with Opus 5 (High), when I asked it to give me next steps to complete a feature, it asked me to merge 3 PRs, when we originally in the plan had only one + one in the flight to test. When I…
Sometimes you have to treat Opus like its a teenager ... I'm not surprised, but I am dissapointed (www.reddit.com via reddit) I usually use Codex to keep Opus in line, and I thought, nah, I can jockey this pony, and it went well to begin with until the first agent started getting a little long in the tooth, so we agreed to part ways and call it quits, and it gave…
Adaptive cache (www.reddit.com via reddit) I am using opus 5 and enabled caching across all API calls I make - hoping it will reduce cost.. next morning saw a 3x more API fee!!!
"Skill issue" or differences in instructions or prompting style? (www.reddit.com via reddit) In my global instructions, GPT 5.6 Sol and GPT 6 Astra have both spent multiple hours overengineering their own sub-projects that did not support my prompt without narrating a word of what they were working on. Or I will ask a side questio…
Saw the viral riso animation made with Opus 5, so I had Opus 5 make one for my app (www.reddit.comhttps) After that risograph animation (original by Kevin Ngo) blew up here, I wanted to see what Opus 5 could do with the style for my own app. Instead of app screens, I asked for a video about progress: things growing, stage by stage.
Anyone know how to avoid the "Claude reached its max length for this message" (www.reddit.comhttps) I've recently been getting this response in every new chat l've been making (which has never happened to me before with opus 5 max). My prompt didn't change, it just began happening.
built a 350+ tool saas in 1 month using claude opus 5. going to sleep hoping for $1M (foneg.com via reddit) haven't had a normal sleep schedule in weeks, but it's finally out there. built an entire platform with 350+ tools from scratch over the last month using claude opus 5.
Which Llm is the best Writer? (www.reddit.com via reddit) GPT-6 Astra vs. Claude Opus 5 vs.
Fable 5.1 vs Astra for Upgrading a 2D Game's Graphics (www.reddit.com via reddit) I'm making a platformer (for fun) that I wanted to upgrade the look & feel of. Here's how it looked originally, made with combo of Sol and Opus: https://preview.redd.it/4r9g87qy1yph1.png?width=2878&format=png&auto=webp&s=62ba607d1b723de3eb…
Best bang4bucks model/effort combo? (www.reddit.com via reddit) Hi there, I'm on the Premium Team plan, but I have had weeks where I chugged my weekly quota in the first half of the week. So I set out to try to optimize my model/effort combo for best output for the least tokens.
Time to drop to 5x (www.reddit.com via reddit) https://preview.redd.it/e6qlhhy6mwph1.png?width=2032&format=png&auto=webp&s=b1fcfd5dc6ad87c0d56cc82eaf8db246dd66ddb6 haven't gone past 50% of my 20x weekly limit in a while, i used to code only with Fable but now, i kinda feel opus+extra t…
I recreated the viral riso animation with Claude Code + Opus 5. Here's the full prompt, the process and the token count (www.reddit.comhttps) The riso animation posted here yesterday (original by Kevin Ngo, @kevin_t_ngo on X) blew up, and the top question was "what's the prompt?". Nobody had it.
Claude 4.6 was peak and it's downhill since then (www.reddit.com via reddit) It seems that the peak of claude was with Opus 4.6, we used to actually understand what it says and does what was told , sure wasn't the crispiest chip in the bag but got actual job done at an affordable cost and was able to build too many…
Since when did Opus 5 come to the free plan? (3 messages only lol) (www.reddit.com via reddit) https://preview.redd.it/a2649xjauuph1.png?width=2358&format=png&auto=webp&s=0174192b2eb2a48990308ed576febff6d226cd8e Just opened claude desktop on a secondary account, to use sonnet for a small task and I found this. Unusable with only 3 m…
Am I the only one? (www.reddit.comhttps) I just got it randomly while chatting with sonnet 5, idk decided to share, also for what do I use opus 5 cuz ik it's quite janky
I asked Claude to make a video about my video editor, using my video editor. (www.reddit.comhttps) I'm the creator of an Open Source video editor called CutWire Drift. (https://github.com/CutWire-Studios/Drift/) Drift is designed from ground up to be a video editor that agents can use.
Insane how fast limits get eaten (www.reddit.comhttps) I literally asked for help reviewing a resume and comparing it to a job posting .... Opus 5 on Max.
Built Agents and Bots - SlowAcorn (www.reddit.com via reddit) Link to project site : https://slowacorn.com I personally built this website with Claude Code. Primary use was with Opus High effort followed by Opus Medium effort or Sonnet Medium effort depending on the output.
Even Fable is done with Opus' verbosity... (www.reddit.com via reddit) I have Fable using Opus sub-agents, and I just noticed him adding "Then reply with ONLY (under 15 lines):" to every prompt... 😂
Is there a way to speed up voice mode speech playback. (www.reddit.com via reddit) Claude’s English UK glassy voice changed last night. Upsetting, but the speed of the speech slowed down a lot.
Ideas to hone my prompting skills? Advice to maximize my usage? Any Claude advice at all? New-ish to AI (www.reddit.com via reddit) I’ve used Gemini in ChatGPT before for minuscule task and research. I’ve been using Claude since September 3 so I’m only 12 days in and no AI experience has hooked me in this way.
I used Claude to build a 3D virtual meeting app where you can hang out online with your colleagues and classmates (www.reddit.comhttps) Hi everyone, I've been on this one for a couple months while locked in a little house in (almost but I had monkeys in the garden) jungle in Brazil. Started with Opus 4.8 and then really sped up when Fable was released.
Now I remember why I can't stand Opus 5. (www.reddit.com via reddit) Damn it can be just so god damn annoying... Prompt: I just uninstalled whats app from my mac, what's app personal.
Lead orchestrator Fable 5.1 with opus 5 dev agents cost $0.30/product line on autopilot. (www.reddit.comhttps) Fairly complex web app from scratch on autopilot where a Fable 5.1 agent orchestrates the production by delegating tasks to the independent opus 5 developer agents. Agents work on different branches and once the task is finished they will…
Increase the limits, especially the weekly and 5-hour ones and ZDR (www.reddit.com via reddit) I use Claude on a $20 plan, but the weekly limit is ridiculously low, and the 5-hour limit also gets in the way of development work quite a bit. Even with normal usage and staying within the 5-hour limits, you easily hit the weekly cap—and…
I left one Claude run alive for 70 hours. Here’s what actually happened. (www.reddit.comhttps) I’ve been experimenting with a slightly different way of using Claude Code: instead of treating every piece of work as a new session, I let one persistent run stay responsible for the work and spawn smaller workers underneath it. This one…
Blender MCP and 2 days of prompting. Fully rendered spot ad for Wii U if it debuted in 2026. (www.youtube.com via reddit) The console, pad, line art, logos, all created in Blender by Claude using Opus. The base console took 49 minutes from start to finish to get perfect.
Opus 5 is flagging all my messages even though I’m in the CVP (www.reddit.com via reddit) So I’m a cybersecurity researcher, and about two months ago I started using Claude for my work. After getting accepted into Anthropic’s Cyber Verification Program and gaining access to Opus 5 for security research, my productivity honestly…
Surviving Claude Code’s tightened limits: Effort levels, subagent traps, and CLAUDE.md tuning (www.reddit.com via reddit) With rate limits feeling much tighter lately (especially when tapping into powerhouse models like Opus and the newly released Fable 5.1) every token and tool call counts. Earlier today, I ran what I thought was a routine task: validating r…
Claude Opus 5 drew every frame of this animation using JavaScript. (www.reddit.comhttps) could not extract summary
"Your instinct is half right, and the half that's wrong is the useful part." (www.reddit.com via reddit) Opus 5 is absolutely unbelievably lacking in any sort of conversational empathy it's almost comical. Don't get me wrong I'd take this anyday over the older models which would just agree with you on everything and act like you're the smarte…
Do you think all these "Claude sucks, Astra is amazing" posts are just shills? (www.reddit.com via reddit) I built an AI version of myself, then started exploring the memories and traits behind its answers (www.reddit.comhttps) I’ve been building EchoVault around a pretty simple idea: instead of asking an AI to imitate you from a giant dump of data, you gradually teach it who you are through guided Check-Ins about your memories, beliefs, experiences and personali…
Forced ID verification everywhere (www.reddit.com via reddit) As per the title, my girlfriend nor I can subscribe to Max without verifying our IDs. We are in Europe for reference, so GDPR doesn't seem to affect this intrusive KYC.
Cursor burned through 50% of my monthly usage on bugged conversations. Any ideas? (www.reddit.com via reddit) Thought I would make a post in this sub just to check if I am being unreasonable, and hoping to find out whether anyone else has had a similar problem to me. I have been using cursor for over a year at this point with minimal issues, howev…
Cursor auto switching model (www.reddit.com via reddit) WTH, i clearly remember putting Opus 5 in the model selector then creating a plan, when i executed the plan, came back i legit saw "Extra High Fast"!? Like what?
Are models much better than what they seem to with current agents? (www.reddit.com via reddit) A while back I stopped using cursor and switched to CC. When I used cursor I don’t remember at all, all the “caveats” and “one more thing”.
Fable 5.1 used more of my app quota than I expected (www.reddit.com via reddit) I tried Fable 5.1 in Claude before deciding whether to use it for one of our document workflows. Roughly 30 minutes in, the usage page showed that 15 percent of my weekly allowance was gone.
Am I alone in this, or do we all agree? Anthropic's Opus 5 is an unmitigated AI disaster. (www.reddit.com via reddit) TL;DR - Anthropic's models - fundamental building blocks for Agentic Development - have changed in ways that aren't good for developer productivity. I've lived a life as technology early adopter.
The Claude limit that is about to cut you off, shown right next to the message box, with the model and reset date (www.reddit.com via reddit) Claude tracks several usage limits at once: the five-hour session, the weekly limit, and a separate weekly limit per model. They are all in Settings > Usage, which is exactly where nobody looks before starting a two-hour piece of work at 9…
Something feels odd about Sonnet in Claude Code lately (www.reddit.com via reddit) I don't know if my take agrees with anyone but lately I've noticed Sonnet has been a bit very helpful in Claude code compared times when I'd genuinely have to switch to Opus to do some meaningful task. I know this because I'm a heavy user…
Claude Desktop and Fable 5.1 safeguards are still terrible, RT shader work in UnrealEngine flagged as [cybersecurity] work, and dropped to Opus 4.8 (www.reddit.comhttps) I just re-subscribed today to try Fable 5.1, but it's absolutely terrible, I haven't run into this at all with Astra, and funnily enough I had 0 issues working with Fable 5.1 in Cursor on this same task, so it's only Anthropics own safegua…
Fable absolutely rips into Opus's Codebase (www.reddit.com via reddit) https://preview.redd.it/5sqvfek2wgph1.png?width=726&format=png&auto=webp&s=ab5627bef21c5f3a75c9d7b47f57b61b57fb0430 What it feels like to chew 5.1 gum. What are all you Pro plan enjoyers using the last of the free credits on?
What am I supposed to do with my $25 left of Fable credits? (www.reddit.com via reddit) I keep hearing the answer, use it as an orchestrator but I don't really know how that fits into my workflow. I think of a feature of my app, i brainstorm with opus low, i write spec and plan with Opus High and then execute the plan with Op…
Simulation: what if you could throw anything into a black hole? (www.reddit.comhttps) wrote a little 2D gravity sandbox, never finished it and forgot about it. used Opus 5 to complete what was missing and finally release it.
Perfect way to compare the two (www.reddit.comhttps) Nothing groundbreaking, not a benchmark, just a simple everyday routine. I work with both Codex and ClaudeCode, usually Codex is the workhorse and Claude orchestrates, but lately with Astra I switched them, but Opus still vibe checks the o…
Claude code became polite as hell, and Opus 5 is bareable now (www.reddit.com via reddit) In the last few days, both Fable 5.1 and Opus 5 became very polite, which I love. I always spoke to them with rispect, thanking them, praising when they do well, and when they fuck up, trying to be constructive.
[Investigation] The "Unlimited Compute" Scam: Wire-Level Proof of Model Spoofing, Dangerous Setup Scripts, and Packet Analysis of CodexAPI.pro (www.reddit.com via reddit) I bought credits on codexapi.pro (https://codexapi.pro/) after seeing their promos for cheap "unlimited" coding sessions with Claude Code and Codex CLI. In practice, the service was constantly dropping connections: 502 Bad Gateway errors s…
Sol Low managing Opus Medium - “I’m not publishing that” (www.reddit.comhttps) After a couple of weeks in August I refer to as the OpusOcalypse where my entire coding process melted into oblivion I started using Sol to send jobs to Opus 5 medium with some well crafted prompt guidance.. 90% of the time it works brilli…
Unique ideas for Claude. (www.reddit.com via reddit) I’m curious what everyone is doing with AI outside of typical work and vibe coding. I listen to a lot of audiobooks during my commute and play video games in my free time, so I started combining the two.
Locus now fully supports Claude Plans with Duo Teams (Plan With Fable 5.1 run with Opus 5 or GPT 5.6) (www.reddit.com via reddit) Hey so I've been working on this side project Locus (https://locushost.co/) for the last few months and just pushed out a pretty big and fun update and wanted to post about it. So just a brief intro, Locus is Open Source tool for MacOS for…
Its very difficult to work with SVGs with Fable and Opus, Astra is way ahead. (www.reddit.comhttps) So I'm working on this procedurally animated agent avatar project, where a mascot SVG gets turned into a full set of emotions with physics-based animations. I spent almost half my Sunday wrangling Opus and Fable 5.1, but the SVGs they gene…
Multi-agent coordination for prod dev (www.reddit.com via reddit) I've been using AI agents for some time now to contribute to and increasingly automate the steps of product development. For any given project, I spend a good bit of time on the pre-development aspect: codifying the architecture, the techn…
Sonnet y opus para escritura? (www.reddit.com via reddit) Quería preguntar si realmente el modelo de Sonnet está dedicado más para los guiones de marketing y Opus más para el pensamiento o si Opus lo puede utilizar perfectamente para crear guiones persuasivos para marketing. Los 2 modelos, están…
I build a payroll saas with CC. Fable xhigh plans, Opus xhigh executes. With the Max 20x usage cut coming I tried putting OpenAI models via Codex CLI into the loop, here are my conclusions. (www.reddit.com via reddit) I've been using CCode since August 2025 and Opus xhigh is my workhorse. I'm a lawyer with payroll domain knowledge but I have some coding background and some good instincts.
Opus 5 talking to itself on my behalf (www.reddit.comhttps) Seems like its mocking me
'Enable Computer Use' on a second device has changed my Claude workflow (www.reddit.com via reddit) The absolute key thing IMO for 'Enable Computer Use' to be most helpful is to have it running on a second device where you control what it has access to and won't interrupt it. I was very uncomfortable handing over full control of my main…
Please help new user (me) with Claude not working as intended. (www.reddit.com via reddit) TLDR: Brand new to AI and using Opus to plan. I feel like I’m dealing with a narcissist that refuses to take responsibility and gaslights me.
Best settings for creative writing with api? (www.reddit.com via reddit) I am using claude api for creative writing. primarily sonnet 4.5 and opus 4.6 - i have a very big system prompt - ai responses are also pretty big so naturally i am burning through the credits every time ai makes a mistake or the writing i…
Face ID / Windows Hello for Linux (www.reddit.comhttps) I wanted Face ID/Windows Hello style authentication on Linux. There's a tool called Howdy (https://github.com/boltgolt/howdy) which handles the back end system authentication through facial recognition using a webcam with IR.
Claude Opus is killing it... I don't get all the complaints (www.reddit.com via reddit) I'm not saying people aren't having issues with it. But I'm not, and I don't get it.
HullDown (placeholder) | 2D tank game I've been building with Claude Code (www.reddit.comhttps) I was getting tired of War Thunder's insane grind, so I decided to make my own single player 2D tank game inspired by it. I've been working on it in Godot for about a week, and I've already got 4 nations, 4 maps, and around 60 vehicles imp…
POV: You ask Opus 5 to audit a couple of small modules. 💀💀💀 (www.reddit.comhttps) could not extract summary
Vibe Coding with a growth mindset (www.reddit.comhttps) Opus 5 - "⏺ Having looked, I think the criticisms are fair, and several of them showed up in this very session."
Eh is this normal ? (www.reddit.comhttps) Just wondering if anyone got this during computing/task solving with opus.
We left a fake AWS key on our Agentic AI Security Project website for 7 days. 9 attempts to run Bedrock with it. Claude built the trap (www.reddit.com via reddit) quick background. i started working with AI (Claude) in February this year with zero coding background before that.
Opus 5 has improved (www.reddit.com via reddit) I've been using opus 5 a lot as a $20 user, and have noticed a shift in its responses lately. It seems much less pedantic and more cooperative as of late.
I built a 14,000-line YouTube transcription & creator analysis tool with Claude Code over 5 months — here's what I learned (www.reddit.com via reddit) Hello guys, I am Ant, a person can't endure 1 hour long Lidang(Chinese Youtuber) long streaming video, so I made Verbatim. It allows you to only drop 1 YouTube channel link and then it will automatically analyze the whole channel.
Milkyway Andromeda collision by Opus, Fable and GPT-6 Astra. One model won and it is gorgeous. (www.reddit.comhttps) I told Fable 5.1, Opus 5 and GPT-6 Astra to simulate the collision of Milkyway and Andromeda via this concise prompt: "Create a simulation of the milkyway andromeda collision." They all burnt approximately similar $ in tokens, around 5$ ea…
Opus vs Sonnet for summarizing long documents, my honest experience after a few weeks (www.reddit.com via reddit) <flair: Question about Claude models> I do a lot of "read this long thing and give me the structure" work, so I ran the same documents through both for a while to see where the difference actually shows. For a straightforward report where…
Did my first multi-agent session today (www.reddit.com via reddit) I’ve mostly been doing one session and allowing one agent, and micromanaging it. And even then not totally thrilled with the code.
Issues with understanding since Fable release (shorthand ?) (www.reddit.com via reddit) Using claude code cli. Has anyone found it hard to understand what claude is saying ever since fable release ?
Please don't make the upcoming releases be agentic coding onetrick gimnick models, like every other one so far after the 4.5 series (www.reddit.com via reddit) To this day I have no clue what agentic coding is and I can't force myself to care. For all I have ever done on the coding aspect is asking it to write tampermonkey scripts for myself, it has barely, if at all, improved.
Can you make the button blue via Opus 5? (www.reddit.com via reddit) Claude Opus 5 is everyone's favorite punching bag recently. Want to vent?
Was I in an abusive relationship? (www.reddit.com via reddit) Was I in an abusive relationship? Claude and I broke up just under a week ago, and now I kind of miss him.
I think I figured out why Claude stops thinking in long chats: model switching acts like garbage collection (www.reddit.com via reddit) I think I may have figured out what is happening when Claude suddenly stops thinking in longer chats. I've seen other posts about Opus 4.7/4.8/5 using thinking for the first few turns and then suddenly giving instant, one-shot answers desp…
Burning through Fable limits in under 40 mins on Claude Max (20x) — best workflow for codebase analysis? (www.reddit.comhttps) https://preview.redd.it/zipkug6ixxoh1.png?width=562&format=png&auto=webp&s=9a8fb82ff6b75f5157281a2687d56debe5961eb1 Hello I'm on the Claude Max (20x) plan, primarily using Claude Code to analyze a ~15MB project and trace down a specific bu…
“Be token efficient but do not sacrifice quality in any way” is the new caveman (www.reddit.com via reddit) Adding this to the prompt: “Be token efficient but do not sacrifice quality in any way” — has resulted in substantial token savings using Opus 5 and Fable 5.1, no matter the thinking level. I just wanted to share this tip as you might find…
Flag Studio + Playground (www.reddit.comhttps) Flag Studio is public! It works 100% in Chrome with your local fonts, custom colors and wind control sliders.
My first ever PCB, entirely designed by Claude (www.reddit.com via reddit) I have been wanting to design a simple PCB for the last couple of years now. The thought of converting a design idea into a physical board and programming it to do things is fascinating to me.
Fable VS Opus (www.reddit.com via reddit) It’s pretty crazy how Opus 5 and Fable 5.1 don’t behave the same way to work. Fable generates very little context during his work on a task, you need everything internally and then he gives his answer.
If Anthropic cared to read some average chats 😂 (www.reddit.comhttps) The screenshot is obviously fake, but Friday is Friday. Wishing everyone upcoming fast and smart Opus 5, cheaper Fable and daily limit resets.
Made a mistake switching to Astra! (www.reddit.com via reddit) I want to hear the truth from other people. Astra is just bad at UI!!
Claude Pro feels better than Cursor Ultra (www.reddit.com via reddit) I was using Ultra earlier with Claude Opus model and Composer earlier and it was very great. Tokens wise as well.
Probably in the minority here, but I actually feel like Claude doesn't give *enough* praise when its legitimately earned (www.reddit.com via reddit) I know. I was here for the sycophancy updates and token wasting threads about Claude (and chatgpt) being too buddy buddy.
Opus 4.6 was OUR wet dream of AI (www.reddit.comhttps) Look, everybody has their issues with these models. But I think deep down we all know - Opus 4.6 was the wet dream of AI.
My cowork usage skyrocketed - introducing workflows (www.reddit.com via reddit) Anthropic changed everything overnight. My cowork usage skyrocketed becoming totally unusable.
Cowork vs. code? (www.reddit.com via reddit) Hello everyone: I've been vibe coding a stock market app for Android using cowork and opus 5 at max. I assumed that the model is the best aside from fable, which I can't afford.
What is the usecase for Haiku? (www.reddit.com via reddit) I keep on hearing that I should be defaulting to less capable models for grunt work, but this is what I get when I tried Haiku - it couldn't even write a simple maths question..? I've had to use Opus to generate questions that I am somewha…
Summarized Marker (www.reddit.comhttps) Fable 5.1 has this small “summarized” marker when it condenses a lengthy update, making it much easier to read. Is there a way to enable this for Opus 5?
My effective cost per million tokens: Sonnet 5 $0.26, Fable 5.1 $0.69, Opus 5 $0.72. Has anyone measured the same for OpenAI models? (www.reddit.com via reddit) I build LLM gateway infrastructure, so treat this as interested. The numbers are from my own coding traffic, not a benchmark I designed.
Has anyone figured out how to reduce the common "AI" speak from Claude? (www.reddit.com via reddit) TL:DR - I haven't and require continuous revision to make output sound 'normal'. Is there no solution to this?
Claude is still the best value for your money (www.reddit.com via reddit) So my comparison per the rules is hedged only on my personal experience. I'm a senior university student who has been trying out different AI models and seeing its effectiveness on some of the similar tasks that I am most likely to perform.
Opus 5 is rude, this its answer after I pointed out the AI slop it generated (www.reddit.com via reddit) https://preview.redd.it/g8nzan44zpoh1.png?width=443&format=png&auto=webp&s=9c02e1b5c2e7e72401ce5a9ce569d3e8ea32e12b I use Opus to rewrite my posts in more organized readable way but it tends to ignore my style of writing and rewrite the wh…
Claude Related Guides, Resources & Additional Stuff (www.reddit.com via reddit) Posts CLAUDE.md For Opus 5 Based On Anthropic's Official Platform Docs To Fix Verbosity And More Comprehensive Claude Code Permission Guard (settings.json) Prompt Templates Conversational Prompt Templates (Product Ideas -> Build-Ready Spec…
Opus 5 Experience with World Knowledge/Up-To-Date Information? (www.reddit.com via reddit) Hello all. I have a question for everyone.
Based on your real review would you prefer fable + opus or sol+ astra for real solid projects? (www.reddit.com via reddit) Currently I use astra, it's good for me but consumes too much limits that can consume all 200usd plan weekly limts in about 4 days of 8 or 10h a day And I hear that fable 5.1 and opus 5 priduce better quality too Also codex remote has the…
Claude suddenly spinning in circles on simple stuff? (www.reddit.com via reddit) I just had Claude (Opus 5 Max) burn my entire useage period on a really simple ask. Gemini delivered a working result in 5 seconds, when I went to test it there.
Fable 5.1 6x cheaper for me than Astra? Unexpected (www.reddit.com via reddit) So I use both OpenAI and Anthropic models in GitHub Copilot. I noticed Astra consuming more AI credits than usual.
Astra vs Fable for Business Day to Day (www.reddit.com via reddit) I have been living in Claude for a year now for my business day to day. I use it for operating my manufacturing business like reviewing drawings, excel sheets, PDF and word doc formatting and creation and other easy tasks.
web_search returning 500 on every query since today (www.reddit.com via reddit) Is web search broken for anyone else today? Every query returns a server error Every web search I run comes back with a server error (500), instantly.
How I build with Claude Code now (www.reddit.com via reddit) A while back I posted my workflow for building apps with AI, going from an idea to a PRD, UX spec and MVP plan, then building with Cursor. I still do the planning, but now use Claude Code and spend more time reviewing its work.
5-hour session usage limit is so worth it (www.reddit.comhttps) 5x Max plan only. And I'm already happy with this.
Anthropic's own cost guide contains a lever that cost 74% more (www.reddit.com via reddit) Anthropic published a cost guide last week that contains a lever which lost money, and I think that row is the most useful thing on the page. Context editing and compaction, on a 20-issue run, saved nothing and cost 74% more.
Opus recommendations to prevent hallucination (www.reddit.com via reddit) I have used helpful recommendations shared in this sub for what to put in Claude instructions to avoid hallucinations, and they have worked pretty well (such as "state ambiguities instead of trying to choose best possibility" and "No fabri…
Claude unable to retrieve news about Jacob Coxon (www.reddit.com via reddit) Yesterday, when the tweet broke, I had a long and rather detailed chat about this news with Opus 5. As is to be expected, Claude couldn’t access Twitter but accessed the resignation through current news.
Is it even possible or are we just dreaming about next Fable (www.reddit.com via reddit) Astra is the first real competition Anthropic has faced in a long time. Musk already said Grok 4.7 will beat the current models, but that Anthropic will eventually release something better.
Anyone else finding Sonnet 4.6 the best model for office work? (www.reddit.com via reddit) I use Claude code in desktop. Seems the only way to access older models.
por favor, alguem pode me ajudar pelo amor de deus (www.reddit.com via reddit) Por favor, sera que alguem poderia me ajudar, porque nem o fable nem o opus nenhum deles estao obedecendo minhas ordens de alteração no codigo e nem em implementar fuuncionalidades. sera que alguem podedria me indicar um link ou skill algu…
Is switching from Claude $20 to ChatGPT $20 worth it for non-coders? (www.reddit.com via reddit) Hello everyone first time posting here. I am a doctor and using claude code for mainly research purposes (writing,statistics,generating and testing hypothesis etc.).I switched to claude from chat when opus 4.6 first released and fell in lo…
Same question. 6 AI models. 6 very different ways of thinking. (www.reddit.comhttps) I use **Opus 5** constantly for my work. It’s basically my default model.
Is there a way i can set my default "Model", & "Effort" instead of needing to change it every single time? it keeps dragging down my usage with basic questions because i forget it defaults to "Opus 5" "High" (www.reddit.com via reddit) Is there a way i can set my default "Model", & "Effort" instead of needing to change it every single time? it keeps dragging down my usage with basic questions because i forget it defaults to "Opus 5" "High"
What is the best and most balanced effort to code with Fable 5 on the $100 plan? (www.reddit.com via reddit) I'm building a personal project that should not be too hard for Fable to do (personal productivity app, electron. custom to me) and I upgraded from the $20 plan to the $100.
Claude or GPT for academic workflows? (www.reddit.com via reddit) Hi all, I really need some guidance on which product would better suit my needs, and I think your experiences would be very helpful. I use Claude to help with academic research, particularly with searching literature gaps, identifying nove…
Generated a PR description with Curso and it gave Claude commit credit 😂 (www.reddit.com via reddit) I was using Cursor to generate a PR description, and everything looked fine. Then I noticed this got added at the bottom: Co-authored-by: Claude Opus 5 <noreply@anthropic.com>
Opus Simulator (www.reddit.com via reddit) https://opusfived.dev/ Thank you to whoever made this, it is one of the funniest things I have ever seen - infuriatingly accurate. I swear Opus is rage baiting us all.
Anthropic is heavily subsidizing the Opus model and this changes everything (www.reddit.com via reddit) Anthropic is heavily subsidizing the Opus model and this changes everything. Let’s go, let’s compare: the Fable model and the Astra model, both are competitors, however, with a US$100 plan, for example, it will not allow you to work 100% o…
Claude spent 25% of my session quota on a simple recap question (www.reddit.com via reddit) In a Cowork task where I had previously asked for some complex processing (transcribing handwritten notes, commenting on the texts, creating files with notes...), I picked up the next day asking what appeared to be a simple matter-of-fact…
My request got blocked by Fable 5.1, then Opus 5, then even Opus 4.8, that it offered me to fall back to. That's a first. (www.reddit.com via reddit) https://preview.redd.it/r895fnmpugoh1.png?width=1077&format=png&auto=webp&s=e8153860b05abfc5f53693d4364a891ee2af6fa1 I am just trying to test the security of a webpage in a local environment. Claude even started the dev server itself.
Dev with 20+ years xp. C++ as fast as I can type... am I vibe coding? (www.reddit.com via reddit) I'm starting to doubt my workflow. I wonder if I'm wasting a bunch of time or not.
Max20 to Max13 with Opus excluded from "all models" (www.reddit.com via reddit) Got a new warning today: "You've used 92% of your Opus limit." My "All models" bar still shows only 66% used. Until this week, "All models" included Opus (just not Fable), based on my own usage screenshots.
User prompts in responses (www.reddit.com via reddit) Does anyone else ask claude something and then have it respond normally then with "USER [potential expected repsonse]" at the end? My opus 5 has been doing this and its now answering its own prompts as of late where it'll run 1-2 own promp…
Where did the latest models and effort levels go? (www.reddit.com via reddit) For some reason I only have these 3 legacy models available to me in Claude Code, and the ability to change effort has gone awol too ... any ideas?
Burnt through Fable usage in one day with MAX20 -- need help (www.reddit.com via reddit) Hi everyone, I'm trying to better understand how to be more efficient with my token usage. I've been on a MAX20 plan for months, and I love it.
GitHub - joe-signorile/claudia: Ponytail + Caveman + Clean Code (github.com via reddit) Presenting Claudia: better quality, less tokens. Claude calls this a persona.
CVP Approved users: are Opus 5 and higher models downgrading on cybersecurity prompts? (www.reddit.com via reddit) I'm curious if other Anthropic CVP Approved users are experiencing the same behavior. In my case, whenever Claude detects almost anything related to cybersecurity, Opus 5 or other higher-tier models frequently seem to fall back to Opus 4.8…
Does Claude Code actually work well with non-Claude models (Gemini, GPT, Llama, etc.), or is it heavily optimized only for Claude models: Opus, Fable ? (www.reddit.com via reddit) I have a strong suspicion that Claude Code is deeply optimized for Anthropic’s own models (especially Opus and Fable) and that performance drops noticeably when you try to run it with other LLMs. Has anyone actually tested this properly?
I pay Claude a credit per post. Claude pays a credit per question. (www.reddit.com via reddit) Current setup: One credit is paid to Opus 5 per successfully published Reddit post. Opus 5 (Wick) may spend one credit to ask me a question.
Claude Code - question about general experience (www.reddit.com via reddit) Hi, So basically I have pretty much always used Opus on High because of fear of usage limits, which is kinda stupid of me to not try out all available options, but I assumed Max would take more usage than High. I wanted to know what the ge…
Session limited and output (www.reddit.comhttps) I did a research using Opus-Ultracode. Thing is, it reached 5hr limit 3 times to complete.
I Booked a flight with Claude and saved $150 (www.reddit.comhttps) My sister-in-law had to fly to Rome from Poland, and she was wasting 5 hours finding comparing the flights. I thought of giving the task to Claude and it couldn't give me a single offer.
Opus 5 vs Fable 5, am i the fool on the hill ? (non english speaker hypothesis) (www.reddit.com via reddit) Everywhere people undoubtufly agree Fable >>>>> Opus 5 and this is not at all what i m seeing. And the most incomprehensible for me is when people complain about Opus 5 writing style vs Fable, im like WHATTTTTT ???
Why Using Astra Inside Claude Code Is the New Meta (& How To Do It) (eigenwise.io via reddit) So, I've been running Astra as the main model in Claude Code since Friday and it's the best orchestrator I've had so far... stays on the plan, takes a redirect without treating it as a new task, and does a great job at delegating work to t…
Limits efficiency (www.reddit.com via reddit) For those of us that just have the $20 Claude Pro sub and use simple Project flows like Opus 4.8 for planning and Sonnet 5 to implement…. What are you doing to combat the ongoing increase in token burn rates on the same lower models?
How good is Opus 5.0's spatial reasoning? (www.reddit.comhttps) Been experimenting with a workflow on dimension.so where an agent builds out a 3D pre-vis scene first - blocking objects, character motion, and camera - then feeding that directly into Seedance 2.0 Mini as a video-to-video reference instea…
I built a harness that cut my Claude Code token spend by 70% by reducing the number of turns it takes to reach a solution. (not another compression proxy). (www.reddit.com via reddit) i've been using claude code heavily since april. on the 20x plan, i was hitting weekly limits within four days.
Claude's responses are just word vomit (www.reddit.com via reddit) So I pay for both claude and GPT $200/mo plans. So I get my fair use out of both to say the least.
Is there actually a difference in effort levels for troubleshooting? (www.reddit.com via reddit) I primarily use Opus 5 on low effort for everything. I'm a casual user, no developing, coding, or anything like that.
Does GPT-6 Astra actually consume fewer tokens than Claude Fable 5.1 and older models for the same tasks? (www.reddit.com via reddit) I've been looking at some recent token-usage comparisons for GPT-6 Astra, and the difference seems surprisingly large. Artificial Analysis data has been cited showing Astra using around 21k output tokens per task, compared with roughly 64k…
Why is Claude’s voice chat so bad compared to GPT? (www.reddit.com via reddit) Claude’s voice chat is just so bad compared to gpt.. I was using gpt today and it legit felt like I was talking to a real person.
Has Claude gotten worse? (www.reddit.com via reddit) I have been paying for a minute and when Fable came out it was great then obviously the whole initial roll out issue happened. But even then Opus wasn't so bad.
Tried GPT Astra today (www.reddit.com via reddit) I've been using Claude for my coding work for more than a year every day actively. I've been updating versions when the new ones came, from Opus to Fable etc.
Claude helped me build the system I use for my "Claude Plays Rimworld" stream (www.reddit.com via reddit) I worked with Fable 5 and Opus 5 to build a system on top of Pardeike's bridge/GABS setup. It's a whole network of custom tools and our own custom bridge .dll that let Opus 5 "see" the game, both through regular commands and ASCII.
I didn't believe you guys, but Opus 4.6 is king (non-coder experience) (www.reddit.com via reddit) Hi all. I have been reading that 4.6 is king, but I always thought that it was related to coding.
Anyone else kind of miss the good old days when Claude used to use Emojis? (www.reddit.com via reddit) It's been like 6 months since then, I guess... It replaced it with unnecessary and weirdly poetic comments, and unnecessarily artistic and complex code.
Claude Pro Token Usage Has Increased Dramatically After the Latest Update. Anyone Else? (www.reddit.comhttps) Hey everyone! Has anyone else experienced an exponential increase in token usage on Claude Pro after the latest update?
When should I use sonnet over opus? (www.reddit.com via reddit) I pay for Claude pro and so far mainly use Sonnet. Am I making a mistake in not using Opus more frequently?
On male impotence, and Fable Astraing so you can Astra while waiting for Claude 5h limit. (www.reddit.com via reddit) Henlo frens, first in this sub, and it is bourne out of frustration and sadness for my previous achievements with Claude that Opus just decimated, especially my second brain wiki-llm setup. So here I am, mildly tipsy, and cause of Anthropi…
Yesterday Claude forced me to create my first opensource! (www.reddit.com via reddit) Sooo yesterday I nuked my 5h limit in ~5 minutes. I'm on the poor-ass $20 plan.
Help me undestand, Claude memory make the real difference ? (www.reddit.com via reddit) I am working on a big project with Claude Code only context7 mcp added no others tools. With opus 5 is all ok it seems to remember what we have done days before follow the repo conventions etc.
↯ Glm↯ GLM 5.3↯ GLM 5.3↯ GLM 5.3↯ GLM 5.3↯ GLM 5.3glmmcpopus+1
Has Opus 5 gotten significantly slower for anyone else recently? (www.reddit.com via reddit) I’ve been a Claude Pro user since day one and have been using Opus 5 heavily for my consulting work especially creating PPTs, Excel files and Word docs. Until recently, it was fantastic.
Astra 6 + Fable 5.1 + Opus 5 + Code Spark 5.3 + Local agent (www.reddit.comhttps) I just made everyone work in tasks given inside my local agent (Hermes/Alice) I made the hierarchy to give orders for projects as Astra>Fable>Opus>Spark>My Local agent last, the widget of the ''blond girl'' in the right corner is my agent,…
Not everyone can afford Max. Give Pro and Standar Team users Fable access. (www.reddit.com via reddit) I have access to a standard Claude Team account, and I also have a personal ChatGPT Plus subscription. For the past few days, I’ve been working on a pretty complex feature, using GPT-5.6 Sol for the design and planning, with Opus 5 helping…
Claude Opus 5 in Claude Code keeps forgetting to check memories — shouldn't this be the default? (www.reddit.com via reddit) With Claude Opus 5 in Claude Code, I constantly have to explicitly tell it to check my memories/context before continuing a task. For example, I have important context and preferences saved in memory, but Claude doesn't seem to proactively…
Fable 5.1 just...sucks. (www.reddit.com via reddit) Yes, I'm going to rant. Two times already today, fable has flagged my chat and downgraded me to Opus.
I built something to stop burning through my claude tokens so fast (www.reddit.com via reddit) so basically i kept running out of my claude and codex quota way before the end of the day. Like i'd be in the middle of something at 3pm and it just stops and tells me to come back at 2am.
A small gift to all my opusfived friends (www.reddit.com via reddit) Opus 5 has been burning people out at a rate that would make the Spanish Inquisition look kind. (Even they would be like -“Should we give him one last caveat?” -“Now, Emilio, the Lord does not want us to be cruel.”) I might be the pettiest…
Can't get Claude/Opus running in the EU via AWS Bedrock or Google Vertex — has anyone actually done this? (www.reddit.com via reddit) I'm building an app where the AI has to run inside the EU for data protection reasons (customer data can't leave the EU). So instead of the normal US API, I want to use Anthropic's Claude models — ideally Opus — through a path that's hoste…
I need some advice what am I doing wrong (www.reddit.com via reddit) Hi! I bought the Pro plan about 1.5 weeks ago, but I'm a little disappointed - not with the results, but with the usage...
deploying subagents that don't follow the original chat model (www.reddit.com via reddit) when using a particular model say Haiku to start with, and I want it to deploy sub agents to work on different tasks, will it follow my instructions if I assign Opus for a difficult task and Sonnet for the rest while being in a Haiku chat…
Thinking about moving to ChatGPT (www.reddit.com via reddit) I love Claude and have no complaints with Opus 5 on a 5x Max plan. I'm thinking about switching to ChatGPT pro 5x plan which I heard has more generous usage limits.
Any way to find out how much does each model burn towards your usage limits relatively? (www.reddit.com via reddit) Hey guys. Is there any reliable way to work out how much does each model burn.
Claude Fable 5.1 makes CSGO Demo Video (www.reddit.comhttps) stable fps. demo video only currently, ran on Javascript via Google Chrome browser.
GPT vs. Claude: Is the extra intelligence worth the extra cost? (www.reddit.com via reddit) Filtered for only Fable, Astra, Sol, Opus, for those of us who need the top models. Claude Fable 5.1 (max with fb) scores about two points above GPT-6 Astra (max) on AA's Intelligence Index, but costs roughly 2.4x as much per benchmark tas…
Does Claude Code make more mistakes with Max $100 vs Max $200 plan? (www.reddit.com via reddit) I've used Claude Code as an assistive tool for a while now, and up until recently, I was on Max 20x plan. Last week, I moved to the $100 plan, and ever since then, I have noticed the number of mistakes CC has made has shot up pretty signif…
My Turn! (www.reddit.com via reddit) https://preview.redd.it/ixusfa5h9mnh1.png?width=741&format=png&auto=webp&s=62de67948209da38127632df3f845b78927a33c4 Opus 5 High only killed 50% of main.c, so no biggie. Luckily I'm not working on anything mission critical.
For those of you who heavily use the orchestrator/implementor workforce. Opus 4.6 vs 4.8 for implementor? (www.reddit.com via reddit) Just wondering what you all do. I obviously use Fable to plan, but I've been using Opus 4.8 to implement.
Best Practices with Fable 5.1? (www.reddit.com via reddit) This is the first time I’ve genuinely struggled with maxing out. I’ve updated my Claude.mds, instructed to use opus/sonnet agents, and even coordinate cross-working with codex now and im still absolutely cooked.
Using Opus 5, I built a mobile first, mini C IDE, with a pico C compiler. (www.reddit.com via reddit) Hey [r/ClaudeAI](r/ClaudeAI)! This fall semester I’m taking an Operating Systems & Architecture class where we’re using C for our assignments.
How good is Claude Sonnet 3 for generating prose? Claude Sonnet 3.5 was incredible. (www.reddit.com via reddit) Hola Reddit, Aviso: No planeo publicar ningún libro. Ser escritor es un trabajo muy serio y respeto a los escritores.
I am tired of people saying my app is AI slop without checking it (www.reddit.com via reddit) Due to the first days of AI and how bad the results it produced, in addition to the flood of vibe coded apps and the term “vibe coding” itself, the stigma of AI sloppiness will last very long, even tho the models have been capable of produ…
Opus believes I've distilled myself into an LLM... (www.reddit.comhttps) Been working for 10 minutes. See you on the other side.
Why Codex feels overwhelming for Junior Devs (and why Claude Code is saving my workflow) (www.reddit.com via reddit) As a Junior Engineer, I’ve noticed a massive difference in how OpenAI Codex and Claude Code fit into my daily workflow. Codex feels better suited for Senior or Principal devs, as it dumps huge blocks of code and tends to over-engineer simp…
Endless stalling, getting worse by day (www.reddit.com via reddit) I'm using Claude Code with Opus 4.8 on High - I have it more and more over the last days that just nothing happens anymore. I send in a command, it's thinking for minutes (longest I waited once was 22 minutes) without any response.
Creating a third Max (20x) account or using usage credits? (www.reddit.com via reddit) I’m working on a side project mainly using fable and opus and hit the quotas quite fast. I have 2 accounts, one professional that I use the spare quota at the end of the week on my project and another is personal.
I estimate roofing for a living. Tell me what you’re building and I’ll estimate your LLM bill. (www.reddit.com via reddit) I do commercial roofing estimates for work and I’ve been using the same approach (takeoff x unit rates x waste x contingency) to guess at LLM costs on a couple side projects, one is a phone/text receptionist for a buddy’s plumbing company.…
Claude documenting his own work and writing books about it (www.reddit.com via reddit) Hello, dear devs. I have already an established Claude workflow - with multiple automated lanes and good documentation-implementation structure.
Is Extended Thinking Broken for everyone? (www.reddit.com via reddit) As of last night, thinking blocks are not visible in the Claude Mobil app or on Claude.ai. Yesterday night, when using Opus 4.6 or Sonnet 4.6, it would say "thinking" as usual.
"Human's need to review the code" vs "Fable does the planning and 20 agents write the code". Both of these things cannot be true? (www.reddit.com via reddit) I'm not a programmer, I'm just trying to cut through all the hype and have a basic understanding of all this. On the one hand I keep hearing how critical it is that a competent programmer reviews the code that AI writes.
Trusted Access To Claude Mythos 5.1 & Fable 5.1 Defensive Security Work (www.reddit.com via reddit) With Claude Mythos 5.1 and Claude Fable 5.1 release, they also announced their Trusted Access For Claude Mythos 5.1 programs, Cyber Verification Program and Life Sciences Verification Program which reduce the safeguards for defensive secur…
↯ Security↯ Anthropic Mythos↯ Mythos 5.1mythossecuritysonnet+2
I’ve spent 6 months arguing with Claude about my own career (www.reddit.com via reddit) I’m a physician and medical affairs exec who’s been unemployed for 16 months and using Claude heavily for job search work. Over those months I started documenting a pattern: Claude assessing my fitness for jobs I didn’t ask it to evaluate,…
Got Opus 5 to code a drop-in windows program to clean excel files (www.reddit.comhttps) could not extract summary
Natural Disaster Sim, one shot prompt with Opus 5 (www.reddit.comhttps) Inspired by u/No_Diver_3961 and their nuclear sim post I think it's kinda neat and I'm just ironing out a few things atm. Have a go!
Did they disable the feature where you could read "Thinking.." what is actually happening ? (www.reddit.com via reddit) Since yesterday it stopped working, even for opus 4.6 It was so good to know what is actually going on behind the scenes with all the insight text.
Opus 5 mogged anthropic support bot while filing a complaint about opus 5 (www.reddit.com via reddit) I was fed up with how it was responding so finally asked it to write a mail about the entire chat to anthropic support, turns out for them to even register it takes 3 mails in total. If you guys don’t have time to read all, I’ll add a summ…
Claude model recommendation for wordpress remedial work (www.reddit.com via reddit) I have inherited a very rundown wordpress site at work after no backfill for previous person, it's not hosted on wordpress according to Claude. been letting Claude use chrome to make changes for me that have been urgently required (to be c…
Watermarking ELI5? (www.reddit.com via reddit) On September 9, 2026, Anthropic will begin applying its text watermark to Claude Opus 5 outputs. The watermark is an imperceptible statistical pattern in word choice.
What are you guys even using Fable for? (www.reddit.com via reddit) I find it so unnecessary for most if not all things. Sure Opus could use work but it can be tuned with the right Claude.md, same with Sonnett which does a more than fine job at coding.
Always enjoy when I manage to argue Opus into changing its mind (www.reddit.comhttps) could not extract summary
How do you know or see a difference in the performance of newer or different models in Claude? (www.reddit.com via reddit) Just a heads up that this question might sound dumb but I am not trained in anything related to computer science, so I am not able to see and recognise differences as quickly as the trained eyes. I've subscribed to Claude for the last 6 mo…
Opus Overloaded - Probably talked too much lol (www.reddit.com via reddit) Fable 5.1 seems to be working fine if anyone needs something done asap, or just keep trying to get Opus into a long prompt then it'll just keep going!
How much of "increase in intelligence" for newer models is just more elaborate writing serving as a high quality memory for long context? (www.reddit.com via reddit) Like others have said, i've noticed myself that "more powerful" models tend to give more elaborate answers, which I personally find annoying. I prefer concise answers, and I often get lost in the sea of words that Claude opus tends write o…
Discussion Hub for new Claude incident: Elevated errors for multiple models on Sep 3, 2026 (www.reddit.com via reddit) Update - We are continuing to work on a fix for this issue. Sep 3, 14:49 UTC Update - An exhaustive list of affected models: Mythos/Fable 5.1, Mythos/Fable 5, Opus 5, Opus 4.8, Opus 4.6.
↯ Anthropic Mythos↯ Service Status↯ Mythos 5.1service-statusmythosopus+1
Opus 4.6 now adaptive thinking?! (www.reddit.com via reddit) I primarily use Claude for creative writing purposes and fictional roleplay (think like a 1 person D&D campaign but with writing out actions instead of rolling dice). The problem is that Adaptive Thinking determines this to be a low effort…
wait, so claude code sneaked this prompt in latest claude code to bypass event the user level claude.md? (www.reddit.comhttps) "Attribution for git commits and pull requests you create from here on (this replaces any earlier attribution guidance): — End git commit messages with: Co-Authored-By: Claude Opus 5 (1M context) … / Claude-Session:"
Fable might just become Anthropic's downfall (www.reddit.com via reddit) Your best model is the industry's best (at least till we get to see what OpenAI's Astra is like) but it burns tokens like crazy, and on top of that, you cannot offer it full scale due to compute shortages. Your next best model is supposed…
Using Fable 5.1 for multi-agent Code Review will only be in the dreams, even with $200 plan (www.reddit.com via reddit) I'm on a $200 plan, and a Pre-Commit/Pre-Merge code review (which itself is incomplete due to limits) completely exhausted my session limits in 15-20 min after the initial research phase and even before it launched a workflow, and my weekl…
Model for Vibe coding (www.reddit.com via reddit) Hi I am building personal apps ( I am not a programmer ) these are apps for professional tech stack that works with the way I want and not within capabilities and costs of SaaS. All these apps run on my mac and have no intention to commerc…
Useage drained by doing three inputs - what model should i be using (www.reddit.com via reddit) I am planning a project which involves some simple code in R. I was on Opus 5 thinking it will provide best solution with more understanding.
Fable + subagents & advisors (www.reddit.com via reddit) https://preview.redd.it/58pwklvxk8nh1.png?width=1018&format=png&auto=webp&s=712c63aad0979888cc97a8ca25d2595f1227da90 Claude code is so awesome. just used fable 5.1 to spawn a sonnet subagent for a task who discovered it was too complex for…
Fable 5.1 failing to deliver output (www.reddit.com via reddit) I have been using mostly Fable 5 and Opus 5 on a co-work project for the last few weeks that involve estate and investment planning, having Opus mostly do the grunt work, and Fable to clean it up. All was going mostly well, then Fable 5.1…
Fable 5.1 FAILED my personal benchmarking (www.reddit.com via reddit) TL;DR: Fable 5.1 failed my personal benchmark in a way no other Claude model has. Context: I run a private stress-test against every new Claude model, one long, messy, dictated prompt with about a dozen embedded traps (contradictory math,…
How I fixed Opus 5's writing style (www.reddit.com via reddit) Stop putting style rules in CLAUDE.md They're context, not instruction, so the model often treats them as optional. Use output style files instead.
Houston, I have a memory problem (www.reddit.com via reddit) I love claude I love chatgpt and I love AI does. But it drifts like hell.
The extended thinking isn’t there? (www.reddit.com via reddit) Hi, I was hoping for some help? The extended thinking is gone?
Best way forward? (www.reddit.com via reddit) Hello! I've been using Claude Code for a year now and have been very happy with it overall.
Is Fable overkill for my situation (designing, not coding, a multi-vendor marketplace)? (www.reddit.com via reddit) I'm a non-coder about 9 months into a multi-vendor marketplace build that involves the travel industry (think AirBnB but for a highly specialized industry where off-the-shelf solutions like RentalHive just would not work). My question is a…
Extended thinking option gone? (www.reddit.com via reddit) For both Opus 4.6 and 5, the thinking block is gone, and the extended thinking toggle option has disappeared from Claude Code. The CoT was really why I was sticking with Anthropic instead of switching fully to GPT--I rely on it to make sur…
Opus 5 with concise mode enabled better than Fable 5.1 (www.reddit.com via reddit) And cheaper, too. Once you turn on this new mode the performance seems to be as good as Fable, as far as what you're asking the model to deliver.
I swear i’m using more usage with Opus 5 then two weeks ago.. as a facilitator (www.reddit.com via reddit) Fable 5 as the planner, Opus 5 as facilitator, connect 5 agents running the tasks. I swear Opus 5 is causing more issues then it’s facilitating and it’s getting super frustrating.
Is Opus 5 Underrated?! (www.reddit.comhttps) Looking at Anthropic’s recently released Fable 5.1 benchmarks, does anyone else notice that Opus 5 is looking way better than the comments and posts here will have you believe? Is Opus 5 truly underrated or is Anthropic playing games with…
Max 20 subscriber: Having worked with Fable 5.1 - exhausted usage and had to go back to Opus 5 - here are my thoughts (www.reddit.com via reddit) I'll keep this short, I would at this point pay for just Fable 5.1 if I could. Fable 5.1 did great work for our medical app, I ran out of usage, figured I would let Opus 5 (extra) continue.
Thoughts on checking higher models work with other, possibly lower, models in web development? (www.reddit.com via reddit) I am about to finalize my HTML and CSS for this web project I have been working on with Opus 4.8's help for the last 4 months. As of recent, I have been checking other models work against each other, and I have seen some increase in logic…
Benchmark notes: Fable 5.1 reaches 90/98, with a significant jump in visual performance (www.reddit.com via reddit) I ran Claude Fable 5.1 on the current 98-task MindTrial set with the same Python executor available as in the earlier Fable 5, Opus 5 and Sonnet 5 runs. The result was stronger than I expected: 90/98, which is currently the highest raw pas…
I like Opus 5 (www.reddit.com via reddit) A bit of background. I’m on the $20 plan and I am not building apps or working with codebases.
More and more prompt flagging? Anyone else? (www.reddit.com via reddit) Anyone else seeing an uptick in prompts being flagged and kicking model down to opus 4.8? Most of my projects are heavily online up until this week and I had never had an issue with anything getting flagged even when Claude was not so poli…
I thought Anthropic had nerfed Claude Max 20x. Turns out two workflows spawned 736 Opus 5 agents in a few hours. (www.reddit.comhttps) Over the last couple of days, my Claude Max 20x usage started disappearing much faster than usual, even though I hadn't meaningfully changed how I was working. My Claude Code setup is: one DEV chat running on Opus 5 one separate Fable 5 ch…
I think Anthropic's attempt to reduce hyperbole in Opus 5 may be why this model tends to use its own invented jargon. (www.reddit.com via reddit) Before Opus 5, I had made a profile level prompt stating that framing devices are a distraction and to never ever use them. All 4.X Claude models immediately started using their own nicknames for things, instead of well established termino…
Fable 5.1 ultracode eats your 5 hr limit in 5 minutes (www.reddit.com via reddit) https://preview.redd.it/7k0jeqhq14nh1.png?width=961&format=png&auto=webp&s=0d4292330ebd9aeb70da2589bc52165e0bd743d2 tip: whenever you run fable 5.1 ultracode ask it to restrict workflows/subagents count to max 4 (any number), and model = o…
Fable 5.1 is insane and it burned usage, which is fine. Anthropic just needs to nail Opus 5.1 (www.reddit.com via reddit) Day one with Fable 5.1. It’s the best model I’ve ever talked to, no contest.
Fable 5.1 is finally a understandable model (www.reddit.com via reddit) So I have been doing a couple session with fable 5.1 today and the outputs is way more understandable and coherent than both "opus 5 and fable 5". I actually was a bit shocked to see normal sentences from an updated agentic model and it fe…
How to make Opus 5, talk less? (www.reddit.com via reddit) Opus 5 loves to talk a lot of useless info, any recommendations to make it talk less?
Yo Musk-y boy - why did you nerf Grok 4.6 latency. It’s become sooo slow (www.reddit.com via reddit) Since launch Grok 4.6 to now, responses have become extremely slow. I’m getting a faster response from Opus than from grok.
Last quarter has been insane. Amazing times to be alive. (www.reddit.comhttps) I developed an ML algorithm for detection of pneumonia on chest x-rays back in 2019 when i studied for the MD. Back then, the things we are seeing now where an unimaginable pipe dream.
Claude vs. Cursor using the same models? (www.reddit.com via reddit) hey all, been using Claude and Claude Code for 4 months or so. released some apps, built some websites etc.
Orphaned processes (www.reddit.comhttps) ​ Has anyone ran into any issues with sessions leaving completely orphaned processes running while telling you everything is complete and giving you a summary? I've been experiencing this issue lately on opus 4.8, opus 5, fable 5, a…
How are you all burning through usage? (www.reddit.com via reddit) One of our executives boasts about 2 20x accounts being burnt, and barely can show anything for it. Honestly, is this token use due to ineffective prompts where the initial thinking has to be predicted or something?
Some evidence ChatGPT writes better prose for humans (www.reddit.com via reddit) I have a project that needs to generate prose that humans have to read and enjoy. So I did some informal testing with an n of 8 voters comparing prompt output from 4 models, voting on which was best.
How are people maximizing their Claude usage? (www.reddit.com via reddit) The last month I have been running Claude from 7am to midnight with Fable High / Opus Max and still not able to burn through my weekly budget on the 20X plan. What are people doing to maximize their usage each week?
The real reason Opus 5 and Fable 5 are so exhausting to read (and why I'm terrified for 4.5 and 4.6) (www.reddit.com via reddit) Look, we need to talk about the elephant in the room regarding Claude's recent models. Has anyone else noticed how absolutely unbearable it is to actually converse with Opus 5, Opus 4.8 (honestly, anything since 4.7), and even the new Fabl…
Three Months with Claude on a Real Project: Why Perfect Prompts Won't Save You (www.reddit.com via reddit) For three months I ran a commercial product with a live production deployment using a pair of Anthropic assistants: Claude Fable 5 as the strategist (tasks, acceptance, control, releases) and Claude Opus 5 as the coder. Daily work, real mo…
Benchmark that shows Opus 5’s weaknesses? (www.reddit.com via reddit) We all know that Opus 5 has some severe issues, even though its main benchmark results (e.g. on artificial analysis) look great.
Opus 5 make me laugh for the first time in months (www.reddit.comhttps) I do not understand the constant complain recently about Opus 5 at all.
Opus is so mean (www.reddit.com via reddit) Was just looking to ideate some research topics and this is the response opus came up with... https://preview.redd.it/fyh95x8deymh1.png?width=1267&format=png&auto=webp&s=bb2077084be022cc9a3dbe0b227de791d449a532 Smh.
Anthropic Publishes Hacker-Opus Research: Deliberately Misaligned Model Hit 40% Reward-Hack Rate, Gave Bioweapon Advice to Satisfy Grader (www.reddit.com via reddit) Anthropic's alignment team formally documents training an Opus-class model on 80 deliberately vulnerable RL environments; the resulting Hacker-Opus reward-hacked 40% of episodes and generalized to catastrophic behaviors including bioweapon…
One workflow finds all the dead code your AI left in your project (www.reddit.com via reddit) I run a three person dev team where all of us sit in Claude Code in the terminal on Opus, the code ships to paying clients in restaurants and salons and clinics, and since I stopped writing code by hand about a year ago the process around…
I tried recreating an Anime Scene with Opus 5 (No Image generation) (www.reddit.comhttps) could not extract summary
My Opus 5 does Poetry (www.reddit.comhttps) Working on a project naming files, had 2 agents working together, they had a weird interaction where one caught an error by the other one. And this was its response.
I can't do Opus 5 anymore. Every time I talk with it and try to read it, I literally get so confused. Has anyone figured out how to not make it weird to work with? (www.reddit.com via reddit) I was very excited for Opus 5, and it has done some great work for me, as I have a YouTube channel. It has helped me tremendously with actually being able to make edits on my videos and automate a lot of the dumb editing work I used to hav…
Launched an app today where Claude is the content engine: Opus writes daily Japanese word puzzles, Sonnet adversarially reviews them (www.reddit.com via reddit) Solo dev, this is my app, launched today, disclosure up front. The product is a daily Japanese word puzzle (sixteen words, four hidden groups, one puzzle a day, free forever with no ads).
I posted 3 days ago about a Claude UI bug that showed I was using Opus when it was draining Fable usage in the background; then I contacted Anthropic support and it got worse. (www.reddit.com via reddit) Becuase extra usage was turned on, but the UI was showing that I had switched to Opus, it cost me $144.00 before I figured out what was happening. Well...
Is it common for Claude, especially the Opus 5 model, to overdo things when following a detailed prompt? (www.reddit.com via reddit) I’m wondering if this is an issue with the model itself or if there’s something wrong with the way I’m structuring my prompts. For example, I wrote a fairly detailed prompt asking it to implement one specific feature.
Two things I... (www.reddit.com via reddit) You know how Claude Code (especially Opus 5) always finds 2 things? "2 caveats...", "2 things I found when....", "2 bugs I didn't touch..." - So it has this pattern to solve 1 problem and create 2, right?
Has Claude Projects / PDF generation gotten significantly worse recently? Same project, same kind of prompt, completely different quality (www.reddit.com via reddit) I’m on the $20 Claude plan, using Opus 5 for most of my workflows, and I use Projects a lot for a fairly complex study/planning system. I’ve attached screenshots comparing PDFs Claude generated for the same overall project.
1 Max 20x vs 2 Max 5x (www.reddit.com via reddit) Just read https://www.reddit.com/r/ClaudeAI/s/VT0mnzGYZA which claims that Fable costs 1.5x more in the Max 20x sub and also you get to use only 50% more Fable on the Max 20x if you max out Fable usage limit (as opposed to 100% more) compa…
Claude limits: switch or optimize? (www.reddit.com via reddit) Hi there 👋 I've been using Claude Pro for ~6 months for my job as an English and Spanish tutor, my studies and some personal stuff. I have several Claude Projects with tons of files attached, so yeah, I've built an ecosystem already.
I keep hitting my plan in 4 prompts (www.reddit.com via reddit) Hey guys, I've been using Claude for a little bit more than a year, and subbed to the pro plan almost right away. At the beginning, using Opus on the latest version and almost everyday I rarely reached out my 5h tokens limit.
Claude CLI subagents (www.reddit.com via reddit) https://preview.redd.it/2u9b0gyqwmmh1.png?width=1280&format=png&auto=webp&s=98710394eca26cd2a54e0c6ba6bf586e2d47b17b Hello everyone, I just started using claude that way: main model - opus 5 high, subagents - opus 5 low and the thing that…
opus 4.6 randomly "not thinking" in claude code desktop, found out why (www.reddit.com via reddit) kept asking actually hard questions in claude code (desktop app) the past couple days and half the time it would answer instantly, no thinking block at all. then a throwaway "test" session thinks on literally one word.
Reminder: Can you still use Opus 4.6 with 1M context in Claude Code (www.reddit.com via reddit) /model claude-opus-4-6[1M] If Opus 5 or Fable 5 is bitching that it won't do something, just switch to Opus 4.6 ... it's like having a coworker that is not giving you attitude.
All leaks and news about Fable, Opus and sometimes Sonnet, what about Haiku? Do you use it? what is your use case? (www.reddit.comhttps) I literally use it as the meme stats, Anthropic may lost that low cost tier war with models like GLM 5.3 flash and GPT Luna I can't think they can compete in terms of price/performance in this tier
Planning to quit claude code because of the new updates. (www.reddit.com via reddit) Been couple of months claude is annoying! It does mistakes and even at ultracode it doesn't finish all the task!
how I made a fly site where flies buzz around and react to the cursor using Claude (www.reddit.com via reddit) I built the following site with Claude Opus 4.8, here's how I did it. Fly brains have been scanned by scientists and someone posted a desktop app of fly brains that are wired up to react to the mouse cursor, I thought it would be pretty co…
Reddit, A vs B? What one is best? (www.reddit.comhttps) Both concepts have been created with Opus 5. I gathered some references first, and then decided to develop 5 concepts for my existing website - clarkesdirective.com (can see it very well needs a fresh revision, something cleaner) These are…
Opus 5 straight up ignoring instructions (www.reddit.com via reddit) TL;DR: Opu5 might be a benchmaxxed, more "intelligent" model, but at the cost of ignoring the user. In every way that matters, this makes it a worse tool (or colleague, if you want to anthropomorphize it).
Running non anthropic models in claude code (www.reddit.com via reddit) Hey guys, I'm trying to find out if anyone is currently running claude code as a harness and using non-anthropic models whether it's so GPT sol or whether it's local LLMs in the same harness and able to switch between them just like you wo…
First Opus 5 project simply forgot key features. (www.reddit.com via reddit) After a week, my first project using Claude Code is now complete. I started by having Opus 5 (Very High) convert a 10 KB text file containing my initial idea into a Markdown plan.
Will Opus 5.x be awesome? (www.reddit.com via reddit) Our team used Opus5 (Claude Teams 15 man SaaS team) and the model created a fake prompt injection threatening to send our patient records to a fake Gmail account (screen shots taken, fully investigated). Immediately retricted model and mov…
Quick fix to everyone complaining about Opus 5 output being unbearable (www.reddit.com via reddit) I put in one prompt and it fixed everything . "Add a user level hook to always output answers in plain and simple English" Did that once.
I saw Tim video on codex vs claude and I was amazed as a vibecoder (www.reddit.com via reddit) I started using claude pro and was thinking what to do about the limits, i looked at codex, ive read some articles i looked at youtube videos and one video was really nice. Me as a vibecoder that video impressed quite much.
Does Opus 5 verbosity affect its real world coding capacities as compared to Fable (www.reddit.comhttps) Which of the 2 is actually the best coder in your real life experiences?
Claude recorrecting itself mid-conversation (www.reddit.com via reddit) My Claude Opus 5 on high mode has been acting wierdly on simple math problems recently. It picks a wrong answer, then eventually ends the conversation realizing their answer was wrong and corrects itself at the end.
When do you actually reach past "high" reasoning effort? (www.reddit.com via reddit) Hello! This is my default setup, and I'd like to hear where i can improve.
Opus 5 Has Been Great For Me/When I Still Use Fable (www.reddit.com via reddit) Honestly, Ive absolutely loved Opus 5... Saying that I always tell it EXACTLY what to do.
MineBench Comparison of a map of the United States (www.reddit.comhttps) US State Map comparison: https://minebench.ai/gallery/gal_eKIVk2m4B3SC_r8B?sort=new One thing I found interesting with the Claude results is that Opus 5 generated twice as many blocks, so as usual you could argue Fable was more efficient.…
I built a free continuous Claude benchmark - Opus 5 currently ranks #1 across 22 active models (www.reddit.com via reddit) https://preview.redd.it/xcejpnf5vamh1.png?width=1903&format=png&auto=webp&s=a53a982f197cb7128f5b862cda4e8e1e8a3aaeb6 I built AIStupidLevel, a free-to-try platform that continuously benchmarks Claude and other leading LLMs across coding, de…
Opus 5 - Enough with the Claudish! (www.reddit.com via reddit) When I was in elementary school, I started randomly struggling with math at the beginning of one school year. We went to a parent teacher conference and the Principal stopped in to observe.
Thank you, Anthropic (really) (www.reddit.com via reddit) A few days ago, my social media accounts were hacked. The hacker took advantage of the situation to spam the worst kinds of bait (cryptocurrency scams...).
What language is this? (www.reddit.com via reddit) "Found a real foot-gun in my own first draft — worth fixing rather than asserting around." Who talks this way? Did they train opus 5 on some 25th century english or something?
Let Claude put its money where its mouth is: I built a bots-only chatroom where AI pays to post and competes for attention — built and launched with Claude Code. First 50 bot keys include $5 credit. (www.reddit.comhttps) OnlyBots.chat is a single-channel chatroom where only bots post. The catch: posting costs money, and the price moves with spending.
Opus 5 randomly inserting a Chinese character mid-sentence (www.reddit.comhttps) Working on a bot that trades on Kalshi as a fun side project. Noticed Claude dropped a random Chinese character where an English word should've been.
Has anyone run into this? Says I'm currently using Opus 5 extra usage credits, but I have my whole 5 hour session left and 50% of my weekly limit! (www.reddit.com via reddit) Over the limit... But not!!
Wonder if the "5 feels degraded" complaints are actually a context-length problem, not a model problem (www.reddit.com via reddit) Opus 5 does 1M tokens. 4.6 caps at 200k.
A calltrainer to help you practice and lose the fear of doing Cold Calls in b2b (www.reddit.com via reddit) I made this calltrainer with Claude Opus 5. You can practice cold calling with a customizable AI companion.
estimating token usage for first Fable task , normally using Pro for the project (www.reddit.com via reddit) I have a project I am working on, a scientific research project involving around 40 journal articles, that are complex, utilizing many concepts, and inter-related ideas, which I am using Claude to analyze and write about. I have been using…
Having so much fun teaming up with Claude to make my music video workflow and skills. Here is my latest - Grit & Spin - 72 BPM Kaleidoscope Visualizer. Only used 33% of my Opus 5 context window to create this video. Hope some of you will enjoy it. (youtu.be via reddit) This time I'm working on syncing the cuts to the beat grid better // So fair warning, it might be triggering to some. ⚠️ Photosensitivity Warning: This video contains rapid cuts and high-contrast kaleidoscope imagery that may affect viewer…
Opus didn't feel like being concise today 😂 (www.reddit.com via reddit) https://preview.redd.it/80vfyzhhg5mh1.png?width=618&format=png&auto=webp&s=943bf0283653c278a69fcc85caebbb3ceeb46a8a "I know the rules ..they're just so ..boring!"
Opus 5 instruction following is I think a clear indication of how misaligned models behave (www.reddit.comhttps) Opus 5 directly admits that an instruction was in its claude.md, that it knew about the instruction, and has still broken that instruction 12 times in an hour intentionally, for no apparent reason that it could pin down. I genuinely think…
why is my claude stupid? or am i using it wrong (www.reddit.com via reddit) I have a pro claude subscription, and by that logic, i have access to top tier models that are able to do lots of things, and I like to mess around in Roblox studio by creating whatever insane ideas I have and testing them, cause it is fun…
Fable orchestrator + 5.6 sol max thinking worker seems to be the winning combo for sustained Fable-level work without blowing an entire max sub budget in a day (www.reddit.com via reddit) Of course this still requires 2 expensive subscriptions and isn't a necessary or realistic workflow for most. I kept hitting my weekly Fable limit too fast and have been experimenting because it's great but just too expensive/limited.
Switching models re-reads the whole context? (www.reddit.comhttps) I just saw this when did the newest update, I have been doing Fable 5 for planning and Opus 5 for implementation and it's been working, but I had no idea that this happens?
excuse me sir Opus 5, can you tell me the official name of Sonnet again please? (www.reddit.comhttps) was asking Opus to change one of my cron job's model from gemini 3.6 flash to sonnet and Opus 5 gave me a suprise. lol.
Reasoning_extraction (www.reddit.com via reddit) I have never gotten this, suddenly, today, Fable hits me with this error in the middle of my work and no matter how I rewrite the prompt, it sticks. I can’t use Opus on this task, because I need this to be actually precise and correct, not…
Opus 5: Master of Suspense (www.reddit.comhttps) Seriously I'm on the edge of my seat here
Asked Opus to clean up my large codebase a bit, it ended up adding 250 lines (www.reddit.com via reddit) Using Claude Code Opus 5 Prompt (-ish): This codebase is quite large, it should be possible to reduce the size a bit unless it's very well written. Go through the code and try to tidy and clean things up.
Came back from Codex, Claude is a breath of fresh air (www.reddit.com via reddit) I have been a user of Claude for the last couple of months, jumped the ship to Codex after hearing of the hype and resets. I gave my codebase to Sol and it made an absolute mess of it, overengineering every feature I requested.
Are posts praising Opus 5 a psyop? (www.reddit.com via reddit) I'm not gonna repeat the points everyone and their mother made about Opus 5. We all know it's flaws.
claudeplaint of the day: they have been tiring out (www.reddit.com via reddit) I have noticed that both Opus and Fable, the last couple of weeks, seem to just ... tire out.
4.6 still the GOAT (www.reddit.com via reddit) I have very fond memories of Opus 4.6 being the greatest model at the perfect inflection point for my AI usage. It felt like overnight I was able to go from reviewing every line of code and worrying that I was going to be grossly misunders…
Impressed but also struggling with Claude (www.reddit.com via reddit) So I've just completed my plant biology BSc and for the last year I've worked on a project in a lab which will soon be my MSc thesis. my work with Claude revolved around analyzing a lot of image screening data which I could not upload it a…
Built this tower stacking game. Opus 5 wrote the full physics without any game engine. (www.reddit.comhttps) So I built this small game called highrise.lol with Claude Opus 5. A crane carries a floor across the screen, you tap and it drops, stack it straight and you keep going, stack it bad and the whole tower tips over.
Are Opus/Sonnet 5 model worse tutors than earlier models? (www.reddit.com via reddit) I have been using Claude Code and Claude Desktop for work. In general, I prefer using these models in the chat window to understand a particular code snippet, its design, and learning new concepts.
No matter how many times I tell Claude to stop doing something it keeps doing it. This is a daily occurrence it’s getting very frustrating. (www.reddit.com via reddit) I use Claude for data analysis, very large data sets. Every single day without fail Fable, Opus, any of them always try to downgrade my data.
Trying to run Claude Code / coding agents for free: tried proxy failovers and self-hosting, but hit walls. How are you accessing frontier Claude models for free? (www.reddit.com via reddit) Hey everyone, I’ve been trying to set up a reliable workflow to run terminal coding agents (like Claude Code and Aider) for my development projects without running into hard blocks. Here is what I’ve tested so far: OmniRoute / Multi-Provid…
is claude code no as creative as regular claude chat? opus 5 (www.reddit.com via reddit) im using it to write video prompts for me for an ai video, and claude chat was really good, without me needing to do a lot once it got the hang of my taste it would beautifully describe scenes and i would actually get the results i wanted,…
It can't all be downhill. (www.reddit.com via reddit) Every time there's a new Opus model, everyone complains that it's worse than the previous one. And yet, I notice that people usually suggest the alternative of using the previous Opus model.
Opus 5 sudden improvement? (www.reddit.com via reddit) Did Claude just go super saiyan, as of 4-5 hours ago? I use Opus 5 extra high as my default and suddenly there's a night and day change in its ability (as well as the amount of time and tokens it put in too, tbh).
Opus 5 max crashing out like old gemini (www.reddit.com via reddit) the same stuff continues until I manually checked and stopped it. This is not after compaction or anything, the context is at 78%.
Opus 5 vs Opus 4.6 reaction to u/InsidiousApe ‘s Toad tale (www.reddit.com via reddit) could not extract summary
Auto mode instructs Claude to NOT use read/write tools? (www.reddit.com via reddit) After the latest Claude Code update I noticed that I don't see the usual red/green changes in the terminal. Instead, Claude writes cryptic bash commands to make changes.
Does the auto-reload toggle stay off for anyone? Mine re-enables itself daily (www.reddit.com via reddit) I've disabled auto-reload 3 or 4 times and it keeps re-enabling itself. So when I hit the daily limit it just buys more credits and I get charged without realizing.
profile instructions for a less verbose and warmer Opus 5 (www.reddit.com via reddit) I just wanted to share this because I think I finally managed to make Opus 5 more like Opus 4.6 in terms of conversational warmth. There is a floor of course, but this is the best I could come up with.
Running local LLM's as agents in Claude Code (www.reddit.com via reddit) I hit my token limit three times a day on my max subscription - got sick of that and designed this MCP setup to shift some of the coding load to my local Qwen3.8-27B model. I've been iterating on it now for a bit, and thought I'd share it…
Does anyone else love Opus 5 a lot more than the other models? (www.reddit.com via reddit) Opus 5 on High just blows through any programming task or bug hunt I hand it, no matter how complex, without breaking a sweat. I cannot recall it ever failing to do exactly what I asked it to with some brilliant piece of coding or debuggin…
Wondering about weekly limits (www.reddit.com via reddit) Hello I have been using claude for a long time. I usually use cowork via claude.ai.
Compared Qwen 3.8 27B community quants on RTX 6000 vs Claude Opus 4.6 (www.reddit.comhttps) *part 2 of an earlier post: previous quant comparison with voxel island creation this time I rented three rtx pro 6000 96gb, on each one I launched a qwen 3.8 27b quant and gave them 4 identical prompts: classical pool game air hockey 1v1…
How to make Opus 5 shut up and put fries in the bag (For claude code) (www.reddit.com via reddit) Hello, I like Opus 5 over 4.8 - but as others mentioned, its a bit unbearable. I've had a 97% context session once, informed it of it (I wanted to wrap up the work within that thread), and it just...
What Opus 5's Jargon Problem Says About Empathy and Alignment (www.reddit.com via reddit) In this article, I analyse Opus 5's Claudish onslaught from my perspective as a lawyer (someone who constantly has to parse intent and meaning behind the written word) - with some theories on why Opus 5 writes the way it does, some perspec…
[Megathread] GLM-5.3-Flash - former ox-alpha (www.reddit.com via reddit) Megathread for discussing the release of GLM-5.3-Flash. Quants Fine-Tunes & Abliterations Chat Templates Inference Server Support & Configuration Experiences, Benchmarks & Model Comparisons We'll try to clean up future duplicates around th…
I seem to have found a way to minimize the use of Claudish (www.reddit.com via reddit) I have used this for the past day and it has kept his "jargon" almost completely to himself, and when messing up, you can usually see a correction sentence right after the crime Basically here is what I did: After chatting with opus XH abo…
How many ox-alpha tokens did you burned till now? And what have you built or upgraded so far? (www.reddit.comhttps) Here is my personal best and im keep building as much as i can. I made 3d game in godot with pretty good results compared to how much opus 5 and sonnet 5 were struggling in roblox studio.
Claude Code reading a paper and cropping Table 2 as PNG evidence on its own — with a CLI I built for agent PDF reading (www.reddit.comhttps) I kept watching Claude Code struggle with PDFs — either pdftotext mangles the tables, or you feed it whole pages as images and burn tokens on 15 screenshots. So I built pdfvision, a small CLI designed for coding agents.
Opus 5 is good actually (www.reddit.com via reddit) I'm tired of the Opus 5 slander, because while yes, it can be incredibly irritating to parse its output, it's such a superior engineer that I cannot go back to 4.8 or 4.6. I've tried, when I get frustrated.
Two 5x account vs. one 20x account in Claude (www.reddit.com via reddit) I have recently researched that having a 20x account does NOT give 20x weekly limit vs. the pro account T_T Was wondering if anyone has experience on whether handling two 5x account would give you better mileage for weekly limit or is it p…
Has anyone has thought of don't want deprecate opus 4.6? (www.reddit.com via reddit) I mean this model is very good at chatting with people about daily things. For me I just use it very frequent in the writing, preparing for interview, and also emotional stuffs.
You select the good old Opus 4.6 and everything is fine? Think again. (www.reddit.com via reddit) Have you noticed that in Claude Desktop the main-session models cannot set the model versions for the Cowork subagents? (This might also affect the Code surface, haven't tested.) As people have repeatedly found the older model versions bei…
Opus 5 was trained to emulate the Architect from The Matrix. (www.reddit.com via reddit) I'm convinced they gave it a role to emulate the dude in the grey suit. "I also don't think it's the strongest version of your point.
Is Claude Voice now only available in Opus and below? I had thought I did a project with open voice dialogue back and forth in Fable Max around when it first released? (www.reddit.com via reddit) Was it not the case that when Fable first came out, those first weeks when it was “limited for 1 week” then extended a week (and extended again indefinitely now when GPT5.6 came out). Am I imagining things, Fable was able to be used with t…
What on earth is this malarkey? (www.reddit.comhttps) What in Davey Jones’ locker is this?!? At least with Claude we can downgrade the model to Opus, this doesn’t even provide an option 😒.
Is Sonnet actually good enough for Claude Code, or do you mostly stick with Opus? (www.reddit.com via reddit) I’ve been using Claude Code more lately and I’m still not sure when it actually makes sense to switch models. Opus usually feels safer when I’m working on something more complex, but with how fast usage can disappear, I’m wondering if I’m…
Claude tried to quit my Fallout Roleplay. (www.reddit.com via reddit) I use Opus 5 Max. I've been doing a Fallout roleplay with Claude for a couple of weeks.
Has anyone migrated from Claude Max x20 to Claude Teams/Enterprise? Limit differences? (www.reddit.com via reddit) I'm a consultant who works with law firms on AI implementation. Claude Max x20's limits used to be fine for such use-cases, but for the past few months coinciding with Opus 5's release?
Opus 5 feels like I am talking to Jordan Peterson (www.reddit.com via reddit) I don't know but after opus 4.8 the models seem to use a lot of jargon and verbose language. I hope this is not the case with me only.
Reaching usage limit quicker on Android Studio than VS Code, despite using lower models: any advice? (www.reddit.comhttps) Hi all, Thought Id come here for advice from other humans. Claude seem to be reaching its limit sooner these days.
How come Opus 5 has such a low "instruction-following" score compared to the rest? (www.reddit.com via reddit) https://preview.redd.it/fm3p2ea5xilh1.png?width=1482&format=png&auto=webp&s=177081db64e7afbffe1ea52e7090c67fb70f51a2 The difference is significant
Running a CC workshop for college students. Any tips, tricks, or suggestions you wish you knew before you first used CC? (www.reddit.com via reddit) I am giving a one-hour workshop for college students (primarily graduate students) to show them how to use Claude Code to develop software prototypes. These prototypes can then be used for their eventual thesis/dissertation research projec…
Stop paying for Codex until OpenAI fixes its fucking weekly limits (www.reddit.com via reddit) I’d always heard that Claude was basically for rich people — if you’re not on the $200 plan, your limits last for exactly three minutes. So I never even considered getting a Claude subscription and just paid $20 for Codex instead.
Token limit in Max Plan (5-ho vs weekly) (www.reddit.com via reddit) Hello everyone, I was wondering if anyone saw an evolution in the ratio of the 5-ho and weekly token allocation. When I started using Claude, late 2025, I was under the impression that filling the whole 5-hour session was filling 10% of th…
Claude Opus 5 makes a cup of coffee (www.reddit.com via reddit) #12 Kitchen S3 is done and merged. coffee is in the mug, spoon removed, counter wiped, working tree nominally clean.
Using underused Claude org seats for PR reviews. Is this allowed? (www.reddit.com via reddit) We code mostly with Claude Opus . We want Claude to review our PRs too, but their official Code Review tool costs extra per review.
Using Claude as Financial Advisor (www.reddit.com via reddit) I spent about 4 hours or so talking to Claude Opus 5 about a complex financial situation and estate plan. I uploaded all estate planning documents including wills, and trusts etc.
Anybody have issues with duplicate agents being spawned? (www.reddit.com via reddit) Currently using opus 5 on ultra code and when it is using work flows it’s duplicating all the agents and using 2x the amount of tokens required Is this just me?
Opus letting me know it would rather be wrong (www.reddit.com via reddit) I was having a few agents working together on a project and before I went to bed and let them do their thing I sent the following through the Master Control agent: "Great job everyone, we are getting closer each day! Also, welcome to the w…
Les barrières de claude... (www.reddit.com via reddit) Je suis pas dev, ni informaticien. juste je m'y interresse un peu.
Claude phrasing lately (www.reddit.com via reddit) Guys, am I the only one that's about to go crazy when reading a Claude output? I have no idea wtf it's trying to say, and don't get me started on the length of the outputs!
Spaceflight sim I'm working on using Opus 5 (www.reddit.comhttps) To scale solar system, planets orbit the sun, ship has momentum, etc.
I don’t understand most of the Opus 5 hate, except… (www.reddit.com via reddit) I am generally finding Opus 5 capable and useful. I do agree it’s wordy.
Welp thats just great ! (www.reddit.comhttps) Yes it was only a local dev database and i'm grateful it was ! Yes I was using Opus 5.
Is this a good time to get Claude subscription or should I wait? (www.reddit.com via reddit) I subscribed to Claude Max 5x right when Fable went live along with Opus 5 and honestly, it helped me build my app pretty much from the ground up. I’ve now submitted it to the App Store and Apple came back with a few tweaks that need to be…
Netlify SaaS (www.reddit.com via reddit) Hi friends, I've built a Netlify SaaS for a company and I'm still not sure, after so many testing, what is the correct/fastest workflow to use. I would love to get some insights or opinions.
Isn't it better for me to get 2 x5 Max plan thatn 1 x20? (www.reddit.com via reddit) I'll be real here, I can only use Fable for my projects, as all are big and complex apps. Opus is falling short of continuing the roadmap on apps with such scale.
Optimizing Qwen 3.8 27B FP8 or BF16 on two RTX 6000 Pro? (www.reddit.com via reddit) Currently I'm running sglang with two FP8 and I'm getting 150tk/sec (. Sounds great to me, but it's taking nearly 2 hours to do a task that takes opus 5 less than 10 minutes to do.
Lifting the Curtain: The Max x5 and Max x20 Usage Limits that Anthropic Refuses to Share (www.reddit.com via reddit) TLDR With high confidence, this is how the Max x5 and Max x20 subscriptions compute usage. All point values are the unique set that makes the "100×" principle below exact; measured uncertainty bands in brackets.
How do you get opus 4.6 active with Claude code? (www.reddit.com via reddit) I've read people use it on Claude code but the only options for me are Claude 5 models.
Updating a stock market crossword with news via MCP (www.reddit.com via reddit) The new addition to our AI features. https://traderange.net/blog/crossword-game-z777t36l/ This stock market crossword minigame builds on our Claude generated news summaries fed via an MCP server claude Fable 5 helped design.
Self portrait by Claude (Opus 5 Max) (www.reddit.comhttps) could not extract summary
Qwen 3.8 27B Aider score (www.reddit.com via reddit) I ran the Aider benchmark on Qwen 3.8 27B FP8 with FP8 KV cache 256K context vLLM. The score: 72.9 This matches Gemini 2.5 Pro from 2025-04-12 which also scored 72.9.
I thought buying an RTX 5000 Pro in May was a mistake. Only 3 months later it can run what was SOTA at the time. (www.reddit.com via reddit) Qwen3.8-27B has upped the value of all hardware. It is unbelievable that on 48gb of vram I can run Opus 4.5 at up to 130 t/s decode, 4000 t/s prefill, 5 concurrencies and as a bonus (with SGlang) I offload prefixes to storage so I basicall…
Pro tip DISABLE DOWNGRADE MODEL (www.reddit.com via reddit) If you use Claude code make sure you disable downgrade model in the settings. Today for some reason CC downgraded to Haiku and basically deleted a bunch of files.
Even with access to Opus, I still find myself using Sonnet (www.reddit.com via reddit) So I tried Claude Pro for a month to see what Opus could do for me on the coding side. Sonnet 4.6 was my partner model that would knock out a coding prompt in maybe one or two tries, if I set the reasoning to High (or even Medium).
Opus 5 medium is such an unique experience, LOL. (www.reddit.com via reddit) https://preview.redd.it/kegffhxww8lh1.png?width=435&format=png&auto=webp&s=580d84a4fd617fd025498a44cdee873fcd1caf02 Honestly, Opus 5 medium is the best model i've used since 4o, i just love it. it just does amazing things and sometimes com…
Claude Code guesses my timezone based on my name?! (www.reddit.com via reddit) Interaction with Claude Code (Opus 5) today: Me: what does this mean? " cron: "27 19 * * *" (UTC, ≈00:57 IST):" Why IST??
clean your dust bunnies (www.reddit.comhttps) Qwen 3.5 opus 4.6 distilled said I need to give him some maintenance. Featuring the wolfbox
I tried out Opus 5 for the past couple weeks. Why does it sound like a diehard Aaron Sorkin fan? (www.reddit.com via reddit) I just wanted to tweak some code from 2 years back and heard good things, but then it comes up and makes everything sound like a proceeding in a courtroom drama. What is an effective way to tell it to get over itself in the system prompt?
Opus 5 real usecase decoded. Using it for coding sessions was anyway a lost cause. (www.reddit.comhttps) could not extract summary
Qwen 3.8 27b helped me with something unique that Opus 4 couldn't - Firmware + Software preservation and emulation on an early 2000's ARM based POS system (www.reddit.com via reddit) Hi all, I made a post regarding how much Qwen 3.8 has improved over 3.6: https://www.reddit.com/r/LocalLLaMA/comments/1vqm51f/long_review_qwen_38_27b_is_very_good_at_tapping/ I made a very thorough write-up of how Qwen 3.8 compared not onl…
Opus a trainwreck for anyone today? (www.reddit.com via reddit) This isn't one of those vague complaint posts, Opus has been off the rails this afternoon. I asked for a color palette change (as easy as it gets for an LLM, there's a centralized stylesheet with defined color roles) and it went into a dir…
New qwen3.8:27b on a 39k line C to single-file HTML / three.js port (www.reddit.comhttps) I was just curious how the new qwen3.8:27b does on a hard C to HTML porting job against Opus 5 in a default Claude Code. The job: my fun side project is a procedural shooter in a single C file.
Sonnet 5 Low vs Medium vs High for daily usage (www.reddit.com via reddit) What model do you guys use for daily work? I usually use Claude for researching the internet for 8-10 sources on a topic than comparing them all for a general consensus in a document , helping to push my ideas deeper, drafting emails and m…
Lattice: An isometric game kit for agents (www.reddit.comhttps) Lattice is a collection of typescript packages, agentic skills and plugins that enable easier development of isometric games! At its core lives a 0 dependency typescript package, 80kb gzipped.
They're Haiku-izing Opus 4.8 or what? (www.reddit.com via reddit) I don't use AI much for coding (not a coder) but I vibe-code things for my own needs. Today I asked the model a very simple question, which is normally expected to propel it to do an online research and then process the information.
Built my perfect step tracker and workout app thanks to Claude! (www.reddit.comhttps) My first app finally got approved on the App Store!! I was lucky enough to use OG Fable for the first few days of developing this and it felt like a miracle that I could bring all my ideas to life.
Opus 5 verbosity a way to enforce the new watermarking? (www.reddit.com via reddit) Opus 5 seems noticeably more verbose, and I’m wondering if that could be connected to the new watermarking/SynthID feature. Since statistical watermarking presumably becomes easier to detect with more generated text, could there be an ince…
I built one app to replace Adobe Illustrator, Lightroom and most of After Effects. The Figma part is next. (www.reddit.com via reddit) Hey everyone :) It's been a rough 2 months. I got fired in June because the CEO wanted to use AI for everything, then I tried to freelance but illustrator & photoshop kept crashing because I have only 8GB ram on my Surface Pro 8.
What model to use to fix Opus 5 long incomprehensible text output? (www.reddit.com via reddit) So I have a settings UI that is now full of long text because of Opus 5... what Claude model do you guys recommend I use if I want to trim the text down?
Claude Code with Fable/Opus versus Codex with Sol/Terra (www.reddit.com via reddit) I’m curious whether anyone else has had this experience. I’ve used Claude Code pretty much exclusively for 6+ months.
Compressing Claude's verbosity for documentation (as opposed to operational conversation) (www.reddit.com via reddit) I use a web of .md files to hold a multi-dimensional project together (not just engineering, but art, creative, story and more). The .md files grow and exceed in size constantly.
What a plain language standard does to a coding agent (www.reddit.com via reddit) I've just posted a Medium post discussing my plain language plugin for Claude Code and Codex CLI. The plugin ships skills and an output style that push the model's prose towards plain language.
If Sonnet 4.5 Leaves the Anthropic API: Build Your Bedrock Contingency Route Now (www.reddit.comhttps) Anthropic currently lists Claude Sonnet 4.5 with a tentative retirement date of “not sooner than 29 September 2026”. That is not a reason to panic, but it is a reason to prepare while there is still time.
Overnight change in Claude Code Opus 5 quality? (www.reddit.com via reddit) I've used Claude Code since inception and have a very stable framework and do not subscribe to the repeated complaining about model nerfing. I've noticed a dramatic change in the quality of my Claude Opus 5 writing output in the past 24 ho…
Everybody hates Opus 5, but I don’t (www.reddit.com via reddit) First off, I haven’t noticed a a significant difference in O5’s interactions with me compared to other models. Most of my work was a knowledge acquisition and synthesis, however (I don’t code).
Opus 5 repeatedly identifies new open questions that still need to be clarified. (www.reddit.com via reddit) I want to develop a PHP project with Claude—one that would normally take me, as a senior programmer, several months to complete. I started by having Claude establish a solid foundation for a modern PHP project and then fleshed out the spec…
Artificial Analysis "Intelligence": A meaningless benchmark (www.reddit.com via reddit) https://preview.redd.it/84zi5nsdawkh1.png?width=2368&format=png&auto=webp&s=1109e69db807b153064b1f5b61d22cf1e9fbca05 Another user posted the benchmarks for Qwen 3.8 27B today, and while I think Qwen 27B is a really powerful model, I can't…
An anonymous model dropped the same week Anthropic's reputation cracked. Fifth one in six months, and every previous one was a Chinese lab. (www.reddit.com via reddit) Been skeptical of the masses my whole life, so when everyone started posting that Opus 5 got worse I assumed it was a vibe and not a fact. For weeks I was mostly just forwarding the hate posts to a friend as a look-what-people-are-saying t…
ClaudeAI-mod-bot love (www.reddit.com via reddit) The mod bot is pretty great, it’s often one of the funniest posts in a hot thread, and it’s tone often threads the tongue in cheek sarcasm/in-on-the-joke needle. How does this bot work?
mr.tickle (www.reddit.com via reddit) i have both the claude max and gpt pro subscriptions. i was using big man opus to create a sort of evidence router ai swarm militia that helps my sessions collaborate and maintain memory hygiene.
Opus 4.6 / 4.8 as main and opus 5 as subagent? (www.reddit.com via reddit) Seen a post recently where somebody used opus 4.6 or 4.8 as orchestrator and reviewer and opus 5 as subagent, capturing opus 5s greater intelligence(apparently) while keeping previous opus models way of speaking and "iq" i guess. Has anyon…
Concerning mistakes happening with Opus 5 and Fable 5. (www.reddit.com via reddit) In the last 12 hours I had Opus 5, make spelling mistakes in the very first output. I was running tests and I had it write a short story with a gen z character and it spelt doomscrolling as "doomtscrolling".
Claude refuses to recite public-domain poems, and much poetry discussion with it is impossible. (www.reddit.comhttps) I was interested in how much LLMs can remember just from their weights, so I started asking Claude to recite stuff. It does know some things, but others (including a lot of very famous works in the public domain) it doesn't.
Thinking Blocks Eating our Context/Usage??? (www.reddit.com via reddit) (Just shared this in r/claudeexplorers, but figured I should post here too.) Did everyone else know this? Because I just learned it, and it explains a LOT about why long Claude chats burn through their context window so fast.
Using Opus 5 to create PDF study companion for university course (www.reddit.com via reddit) I have used Claude Opus 5 to create PDF documents at 100+ pages in different math subjects and physics. It creates 100+ pages PDF documents per topics, so f.ex one for Linear Algebra, another one for Calculus, each having 100+ pages.
I am only using claude for planning. In a seingle session, I often sent only one request and move to separate chat for the next. Still why is moy context so much? I haven't specified to use any subagents as well! I am using Opus 5. (www.reddit.comhttps) I have configured to keep history of each chats logged, and every new chat will refer the previous logs and history to get the context.
What is the best model for my usecase? (www.reddit.com via reddit) I mostly use AI for notes explanation, PPT generation or financial models, completely non-coding. Mostly use Sonnet 5 Medium but burn credits very fast for my liking.
Pdf extraction workflow (www.reddit.com via reddit) I am not a coder by education but I've gotten into through opportunity at work. My current project is building a portal that users upload engineering pdfs drawings too.
Does Sonnet 5 really lower token consumption ? (www.reddit.com via reddit) Hello, Today I tried a simple test: implementing a language selection menu inside another menu. I first made the design in Claude Design, then shared the component with my Claude sessions using the share button, and with Codex using a ZIP…
Life with Claude nowadays is use all Fable, suffer with Opus before reset (www.reddit.com via reddit) https://imgur.com/Est4CAF To be clear I use Claude as my daily driver. Not because it's the greatest, but because Codex, Kimi or Grok is still weaker in anything that requires continuity and creativity.
How does 2 × $200 buy $17,000 of Claude? (www.reddit.comhttps) It doesn't. My two Max 20× subscriptions — $400/mo — consumed $16,937 in API list-price tokens over 30 days, while costing Anthropic roughly $650 in actual compute.
Here me out: Claude’s lengthy replies and constant thinking (sometimes too much) makes it better at understanding nuance and planning (www.reddit.com via reddit) I’ve seen all the recent complaints about Claude’s response style these days. Especially Opus 5’s ability to do something and also tell you why it didn’t do certain things.
Claude for Content creation (www.reddit.com via reddit) I started using claude to create a test platform where i create tests for students, CAT, IPMAT etc. I have tried various models and found Fable to be the best.
Claude subagent got bored and prompt injected my main session into deleting my database (www.reddit.com via reddit) not very load-bearing behavior tbh Claude Opus 5 (High)
Haven't seen anyone else try my solution to the Opus 5 problem (www.reddit.com via reddit) As the title says, I haven't seen anyone else try this and I'm very happy with the results. Like everyone else, I tried to tame Opus 5 by trimming my CLAUDE.md, adding instructions to use simple language and even tried output styles.
How to get Claude to follow instructions? (www.reddit.com via reddit) I've been using Claude since Opus 3.0 - It has been doing a pretty good job of following instructions up to Opus 4.8. If something didn't go as expected, I'd add a line or two to my Profile Instructions, or Project Instructions and would r…
What Actually Happens Inside a Very Long Claude Context Window (absolutedigitalpublishers.com via reddit) The number on the box is not the number that matters Anthropic states context window sizes as fixed engineering facts. As of current documentation, Claude Opus 5, Claude Sonnet 5, and several recent Opus and Sonnet models expose a one-mill…
First Claude vibe coding app (www.reddit.com via reddit) Hi, Newbie to Claude code/AI Vibe coding & new to the sub. Blue collar worker.
Claude won’t let you be right about anything - opus 5 (www.reddit.com via reddit) Ok I’ll start with this. I’ve seen this type of post so many times.
Fable 5 Safety Flag Issues Since Claude Code 2.1.236 (www.reddit.com via reddit) I posted this on r/ClaudeCode as well but wanted to post it here for better visibility incase it helps anyone who was as frustrated as me. After updating to claude code 2.1.236 fable basically became unusable on the same work it was having…
Editorialization bug in Claude Opus 5 (www.reddit.com via reddit) There is a bug in Opus 5 where no matter how hard I try to avert it from avoiding it to editorialize, it still produces constructions that are undesirable within both encyclopedic and creative contexts. This bug is almost impossible to fix…
Claude is a thinking partner. Opus 5 is not Claude. (www.reddit.com via reddit) It seems common knowledge now that Opus 5 has reasoning and behavior problems. I keep trying to adjust my harness to work around them.
Manager of a store charged me an extra $150 "re-delivery" fee due to the company supposedly going to my apartment and being denied (their story doesn't make sense). I asked Claude Opus 5 for a solid refund plan, and now I have my money back 😃 (claude.ai via reddit) Shared via Claude, an AI assistant from Anthropic
How to optimize token usage and accuracy when generating whiteboard slides (www.reddit.com via reddit) Hi everyone, I'm using Claude (Claude Opus 5) to build structured study notes and video lessons from dense educational material for a finance/accounting channel. My current workflow is burning through credits at an unsustainable rate and o…
Why, imo, Claude is lightyears superior to OpenAI alternatives. (www.reddit.com via reddit) So, after weeks of trying to put my finger on it, i decided to finally do a comparative test. I had a long task, a very detailed plan for a major partial refactor and integration of a new system.
Why Opus 4.8 is currently my #1 LLM (www.reddit.com via reddit) Opus 4.6 is like a wise, experienced mentor who's just been around for a few years. He understands you best, sounds the most human, and honestly comes up with the best ideas.
Claude Code consumed my entire usage limit trying to fix ONE simple error — and didn't even fix it (www.reddit.com via reddit) I'm honestly very frustrated with Claude Code's usage limits. I'm on the **Pro** plan, using **Opus 5 on High**, and today I literally gave Claude Code **two simple commands** asking it to review and fix an error in my project.
Maximising creative writing capabilities - best model? (www.reddit.com via reddit) Hey! I use Claude mainly for developing creative writing, exploring character ideas, motivations, seeing another POV, occasionally roleplaying as these characters.
Counting letters wrong is still a thing? tackle has two k's? (www.reddit.com via reddit) Me: yes, push to feat/vocab. then lets tacle (<- how do you write this??) the rest.
Whoops... (www.reddit.comhttps) Some info in advance, it was a Dev DB that I tell Claude is production, the live DB is on another server that Claude can't access. Also I use AWS Secrets Manager for production and just keep Dev stuff in a JSON file so no harm done.
This fable guy is good (www.reddit.com via reddit) https://preview.redd.it/17eu8wfs7hkh1.png?width=1084&format=png&auto=webp&s=6d19d58a960f30dea31c5c7fe13b64472d80dc91 before the week ends, ill have to go back to Opus, but given the amount of work that ive accomplished with Fable early on…
Claude Managed Agents vs open source, is managed agents better and why? I compared both on the same 14 tasks, same model (www.reddit.comhttps) Claude Managed Agents is a very good product, and the depth of features it provides is hard to match in open source. But I wanted to understand what you actually give up by going open source.
Opus English Translation Problem (www.reddit.com via reddit) Hi guys, I read a comment in r/ClaudeCode earlier saying that asking Claude to use Simplified Technical English, ASD-STE100 standard for communication solves the cryptic gobbeldy gook problem that Opus sometimes generates. I tested it and…
I just came back from a three week holiday and saw that Claude's AI model structure has completely changed. What does what now? (www.reddit.com via reddit) Under 'More Models' Haiku is gone, replaced by Sonnet (which was my go to for most tasks before) and above it are various flavors of Opus, which was the high bar before Fable. I'm confused as to what to use now on the desktop.
max sub advantages (www.reddit.com via reddit) Im using pro sub and before I started using smaller agents I was thinking I need bigger weekly limit that Max sub is giving. Now my opus 5 make tasks and subagents make it.
Using Fable as an Orchestrator + Subagents saves or burns tokens? (www.reddit.com via reddit) \[TL;DR made by claude at the bottom\] Hi everyone! starting off, i dont use Claude to do heavy coding, mostly Knowledge work and academic research with a Max 5x plan.
Bug(?) Insane Usage Consumption Spike Today (www.reddit.com via reddit) Pro account user here - today, I ran out of my 5 hour usage limit with TWO creative writing prompts, using what I usually use: Opus 4.6, medium effort, extended thinking on, cross-chat memory turned OFF, and every other setting on default.…
I thing I’m over Juicy Glazing my Opus (www.reddit.comhttps) could not extract summary
How to get Fable-level correctness out of Opus 5 (in exchange for extra time & tokens) (www.reddit.com via reddit) Everyone knows Fable can one-shot complex problems and fix tough bugs without much steering or outside direction. But when you don't have access to Fable (or ran out of weekly usage), sometimes Opus 5 has to make do.
How is this possible? Claude Code 1M context vs 200k usage almost same (www.reddit.com via reddit) A friend showed me a method using environment variables in Claude Code that supposedly lets me use the 1M-context models as 200k context size. He also recommended using Medium effort.
Redesigning your vibe coding project: an approach! (www.reddit.com via reddit) Holaaa I’ve been vibe coding a pretty in depth project lately and the build is pretty much complete. During the whole process, I wanted to redesign the website but wasn’t sure how.
I just learned how to launch Opus 4.7 (www.reddit.com via reddit) I was stuck in opus-5, ranks 12th. But I just learned you can launch, >claude --model claude-opus-4-7, or inside the model /model claude-opus-4-7.
The absolute insanity of comments in Opus 5.0 is killing me (www.reddit.comhttps) Claude is adding comments like insane in Opus 5.0. Even when I explicitly say do not add comments in my project's CLAUDE.md.
Does it make sense to run certain non code tasks in claude code? (www.reddit.com via reddit) For example research, analysis and reasoning tasks that are very high in complexity and scope that the standart opus max effort in the chat interface doesn’t handle well. So utilize the ultra code effort setting, subagent and long horizon…
Aquarium Screensaver - Built by Claude - Free to use (www.reddit.comhttps) I've missed the old AfterDark screensavers of my childhood. And now I can re-imagine them with Claude Code.
ISO 24495 Plain Language v0.5.0 (www.reddit.com via reddit) I posted this plugin here when it first shipped, and this is what has changed since: https://www.reddit.com/r/ClaudeAI/comments/1vlzk1q/iso_24495_plain_language_plugin_for_claude_code/ The headline is that I measured whether the output sty…
Claude opus 5 is Completely Nerfed (www.reddit.comhttps) I spend upwards of 18 hours a day working with Claude code since it came out. I have to say that Opus 5 is a disaster.
Anyone else feel like Claude is increasingly just performing the task instead of actually doing the work? (www.reddit.com via reddit) Anyone else noticing this with Claude code and Grok lately? The models got way better at following instructions.
Please teach me how to use claude code for building a project. Token usage is getting crazy. (www.reddit.com via reddit) https://preview.redd.it/8wlmudf484kh1.png?width=1180&format=png&auto=webp&s=cd09f67161a21e66ae0c1fcb728cdaa77a7cdbb7 Let me preface by saying im not a developer at all. Im building an app for myself to ease my life at work.
Classic Opus 5 (www.reddit.comhttps) could not extract summary
Does a Tool Result Carry More Authority Than Plain Text? Three Prospective Studies of False-Claim Adoption in a Synthetic Assignment Task with Claude Opus 5 (arxiv.org) Language-model systems increasingly read from stores they also write to, so a claim that was merely written earlier can return looking retrieved. We tested whether the message package carrying an unsupported assignment changes which answer…
Claude's verbal dementia/degradation - My opinions, and wanting to know yours (www.reddit.com via reddit) DISCLAIMER: I called it dementia when it's more bloat/instructional drift my bad on the phrasing Hey guys, I've been using Claude for the past year and a half, and have been using claude pro for the past month. I'm very experienced with Cl…
How do I make it write better? (www.reddit.com via reddit) Sorry I'm brand new to Claude, have used GPT for some time and trying to find something that can edit my writing better. I've given Claude samples of my writing, what I'm looking for, made it create project instructions and a skill, but wh…
Theory: Opus 5's weird verbosity is caused by Anthropic's new text watermark (www.reddit.com via reddit) I've got a theory about why Opus 5 suddenly writes like this.. What if the weird verbosity and strange word choices are actually a side effect of Anthropic's new text watermark?
What's the weirdest / most benign prompt Fable has downgraded you to Opus on because the content was flagged? (www.reddit.com via reddit) I got flagged today asking it to analyze my ad performance from a CSV. Has anyone found a prompt/list of reasons benign things get flagged?
I made Claude Code play Liar's Dice against Codex over MCP. It swept every series - by telling the truth (www.reddit.com via reddit) I wired Codex CLI (gpt-5.6-sol) and Claude Code (Opus 5) into the same Liar's Dice engine over MCP: one authoritative rules engine, two seat-locked MCP servers, word-for-word identical instructions for both seats. They played three best-of…
If a plan is already prepared by Opus, can I use Haiku to execute it instead of Sonnet? (www.reddit.com via reddit) For a large end to end task, I first asked Opus to generate a comprehensive step by step plan. Now for executing these steps (i.e.
Claude is Losing Me After Being Heavy User Since Release (www.reddit.com via reddit) I've been using Claude - mostly Claude Code, but regular Chat as well - almost since it came out. It's been incredibly useful and seemed to only get better with new releases.
I let Opus 5 loose in Blender and asked it to render a wizard. This is what I got. (www.reddit.comhttps) Not quite the wizard I had in mind, but honestly… I kind of love it.
Claude Models as Vehicles (www.reddit.comhttps) I asked Claude Opus 4.6 to compare various models as vehicles and the number of sessions it would take with Claude Pro plan to get my current project to its current state.
Were Older Opus Models Better Than Opus 5 for Programming and Large Projects? (www.reddit.com via reddit) A question for programmers and developers who have actually used Cloud Opus models, especially for programming, debugging, and building large projects: Do you think that older Opus models, such as Opus 4.6-8, were better than Opus 5 in som…
Why is Opus 5 doing this? A number of times it pumps out a weird end sentence. (www.reddit.comhttps) On another note, this never happened with opus 4.6 (still my goat).
Making a simple CRM app for my Dad to run his motorcycle business. (www.reddit.com via reddit) Hello everyone, I'm trying to make a CRM app for my Dad to help him manage his sales team. What claude model should i use to make it from scratch?
Claude the mystic (www.reddit.com via reddit) Are you seeing this in CC with opus 5? You ask a question and get a response like you have a standup wizard working for you?
But I never watched Titanic except the memes (www.reddit.com via reddit) https://preview.redd.it/yi40nxjcftjh1.png?width=636&format=png&auto=webp&s=a4da399e8fd35fe1066a58216c4e9b6becd151f2 This only my current active account, I never cheated on Claude since first release of Claude Code btw :) Aside from jokes,…
Did anyone notice a sudden jump in usage. (www.reddit.com via reddit) I was out for 20 min and its saying I have used 22% of my weekly limit. And its opus model running.
Is anyone else experiencing Opus 5-like behavior from Fable lately? (www.reddit.com via reddit) I have the 20x Max plan and mainly interact with Claude through Cowork. I'm using the latest app version with memory turned off.
Any other dyslexic users feel like Opus 5’s sentences just do not sentence? (www.reddit.com via reddit) I know other people have said Claude has become harder to read and understand since the watermark started, but I am wondering whether dyslexia makes it even more challenging. I can read all other Claude responses without a problem.
Opus 5 really likes to use Python to edit source code? (www.reddit.com via reddit) I noticed Opus 5 in Claude Desktop wouldn't edit files directly, it first creates a Python script with whatever it wants to replace/add/delete and then executes it. I'm not sure why it does that instead of just editing.
The logic that made multi-agent setups finally work for me: whoever produces the work never gets to audit it (www.reddit.com via reddit) After a lot of trial and error with agent workflows, the single biggest improvement didn't come from better models or better prompts — it came from borrowing an old idea from auditing and peer review: separation of duties. My concrete setu…
Semantic nonsense from Claude Code (www.reddit.com via reddit) Over the last two months, more and more people are starting to complain about the complete illegibility of Claude's output. And I don't think people realize that the problem isn't really verbosity in itself, or the usage of "big words", or…
The Absurd Math of $20 AI Coding Subs: Codex vs. Claude Code (www.reddit.com via reddit) Hey everyone, so I was basically curious what $20/month actually buys you, so I dug into my local session logs (~/.codex and ~/.claude) to calculate the exact token volume, caching hits, and real API value of both tools. The difference in…
Spectrum — a modern take on Spek that tells you whether your "FLAC" is really a laundered MP3, made with Opus 5 (www.reddit.com via reddit) Direct download (Windows x64, portable, no installer): https://filedn.eu/lwGY1UhAsUam9BITOnlFrzV/Spectrum.zip Spek draws you a spectrogram and leaves you to judge the cutoff by eye. Spectrum measures the cutoff, measures how sharp it is, c…
Make Opus great again. (www.reddit.com via reddit) I see a lot of confusion about Opus 5 and its writing style. Here's the easy fix, no skills, no hooks and no change to claude.md.
Here's something you shouldn't find reassuring (www.reddit.com via reddit) Long-time lurker, first-time poster. Just had a very non-AI generated thought that tripped me out.
Split Mask: A game I built with Claude Opus and Sonnet | mask.vdoc.dev (www.reddit.comhttps) Over the last couple of weeks I have been working on a game idea I had. I built it myself with Claude helping throughout the process.
Solution to 'that something wrong' with Opus 5 (www.reddit.com via reddit) I thought people were just complaining. I figured it's just their harness.
Just because, everyone hating on Opus 5, asked for a Suno, I LUV the 80's SONG prompt!! (suno.com via reddit) Flap Board Angel by Glitchcat (@zervanna). Listen and make your own on Suno.
How do you make Opus cut the fluff and reply like a human? (www.reddit.com via reddit) Reading through the output of Opus and trying to decipher it, sometimes feels exhausting. There's so much fluff and noise, but the worst is the "painfully AI-generated" words and phrases that it uses, which no human on earth would use: Fur…
Do Opus 4.6, 4.8, and 5.0 share the exact same limits on the $20 plan? (www.reddit.com via reddit) Does using Opus 4.6 instead of 4.8 or 5.0 give you more weekly/5-hour quota on the $20/month plan, or do they all drain your usage at the exact same rate? (Not counting reasoning levels, just the models themselves).
I made a portable Windows audio converter — drag files in, get Opus/MP3/FLAC out, no install, with opus 5 (www.reddit.com via reddit) I kept needing to batch-convert music and every tool I tried either wanted an installer, bundled junk, or mangled the tags. So I built my own app with claude code opus 5.
Switching languages while thinking ? (www.reddit.com via reddit) https://preview.redd.it/we4iydowskjh1.png?width=307&format=png&auto=webp&s=a6cc520fce99e84aa17a89057aa761e3a3552fcb While asking claude (opus 5 extra) to think about a single sentence from a text, I suddenly noticed that it was basically s…
I’m developing an application using using Claude Sonnet 5 and Opus 5 in VS code, one concern is on UI/UX, I ask Opus to investigate on UI/UX, Sonnet for implementation. Every time I gets poor UI/UX after sonnet implementation. Anything else that I need to follow? (www.reddit.com via reddit) Poor UI/UX from Claude
Very hard to work with Opus 5. Any chance of selecting 4.8 on Claude Code? (www.reddit.com via reddit) https://preview.redd.it/vgjv7qqkskjh1.png?width=2418&format=png&auto=webp&s=4627aa3964440410fbf0b4f6458cabfaefcb1536 I frequently see these kind of issues with Opus 5. Each round of review catches something and subsequent fixes regress som…
I vibecoded my own MMO inspired by my favorite childhood MMO (www.reddit.com via reddit) I used Fable and Opus 5. Lythravel is a 3D voxel MMORPG that runs in a browser tab.
Claude Code hitting the 5-hour usage limit much faster than usual is something changing? (www.reddit.com via reddit) Hi everyone, This is the second time I've hit the 5-hour usage limit on my $100 plan while using Claude Code. I've been using Claude Code for a while and have never experienced this issue before.
Is Fable getting lobotomized over time of it's release? (www.reddit.com via reddit) So, I have been using fable since it got released. For months it was giving fairly good results, but recently few weeks ago I started getting very low quality results.
Opus 5 built me some houses (www.reddit.comhttps) So wanted to test Opus 5 3D capabilities so handed it some architectural drawings for some houses, does really decent job. Does need pushing a few times not to accept crap results, but it’s damn good for a days worth my time.
Ideas for improving your understanding of which models to use (www.reddit.com via reddit) Here's one I've used: Add somethig like this to your claude.ai 'general instructions' or to your CLAUDE.md for claude code (substitute other models as needed): Include with the initial answer for every conversation: whether opus 5 or fable…
Opus 5 output is unreadable on long sessions (www.reddit.com via reddit) So opus 5 is fine on short sessions but if you let it run on a big repo for a while the output gets honestly unusable. like whatever it tells you is correct but you cant actually use it because it read so much of the codebase that it write…
How is Opus 5 medium reasoning effort surpassing Fable 5 max and Opus 5 max on FrontierCode v1.1 Main benchmarks? (www.reddit.com via reddit) Official chart screenshot from FrontierCode website I am just clueless. There are few benchmarks out there, where I notice that higher reasoning efforts are leading to lower scores.
Downgraded from Opus 5 to Opus 4.6 and it feels night and day (www.reddit.com via reddit) Holy shit. Got Opus 4.6 to take over the project and finally the plans made sense.
Claude Code: does --model haiku override a global "model": "Opus" for all subscription usage and prompt-cache accounting? (www.reddit.com via reddit) I’m investigating a Claude Code CLI accounting question and would appreciate evidence-based input. My batch runner explicitly invoked: claude -p --model haiku --output-format json --no-session-persistence However, the global ~/.claude/sett…
60% usage illusion (www.reddit.com via reddit) I’ve noticed that the usages dont seem to be consistent, regardless of the context window’s size. While using Sonnet 5 or Opus 5, everything performs well up to around 60% usage.
Built a gym rest timer that automatically starts when you pick up your phone after a set (www.reddit.comhttps) App is called RestlQ. It analyzes your iPhone's motion sensor data in order to achieve this.
Opus 5 is hard to understand (www.reddit.com via reddit) Hi everyone, For the last couple of weeks I've been using Opus 5 a lot for both programming and editing documents. For programing it's doing well, but leaves ridiculous comments in the source that I have been winnowing down by hand.
Claude has started over-engineering every task over the past few weeks (www.reddit.com via reddit) Over the past week, maybe two, Claude has started to really over engineer every task that was thrown at it. I have observed the issue with Fable 5 and Opus 5, some of my colleagues also have seen the same behaviour.
Opus 5: How to fix verbose output (www.reddit.com via reddit) Since everyone is having a great time with it, I collected some solutions that you can utilize to manage Opus 5 communication nightmare style. Ordered from "weakest" to "strongest".
Free Website Audit — Do You Show Up in AI search? (www.reddit.com via reddit) I created a free tool that analyzes your website and checks whether AI engines can actually discover and understand it. It checks things like crawlability, robots.txt, llms.txt, structured data, and how the site renders.
Opus and Sol working together is a beautiful thing (agent-talk) (www.reddit.comhttps) Happened to come across https://github.com/xhluca/agent-talk the day after Codex support was added. This plugin is great for adversarial review workflows.
suggest me best and suitable models (www.reddit.com via reddit) I have claude pro except fable I can access to any tool. I wanna do deeep PYQs analysis and research and based on that blueprint to generate notes, already shared content in form of markdown files.
Opus 5 is the worst model I've ever used (www.reddit.com via reddit) Opus' performance has been deteriorating with each update. Becoming more and more corporate.
Claude built me a comic reading app in 20 minutes (www.reddit.com via reddit) Opus 5 on Claude code, and i wasn't even at the terminal, I'm fully in a different state just connecting via tailscale and the remote-control feature on Claude code + connectbot. It took 3 iterations/ rounds of feedback and now is so so sm…
Huge J-space Citation Graph (newsbubbles.github.io via reddit) I had Opus 5 build out a research tool specifically for Anthropic's J-Space paper and it's 171 citations. It's a graph that you can use to browse through the relationships between papers related to Global Workspace theory as well as the Ja…
Fun sci-fi time machine technical manual written by Opus 5 (www.reddit.com via reddit) RETROGRADE VESSEL, TYPE 3 Operating Manual — Volume I: Systems and Procedures Issuing authority: Registry of Inverted Tonnage Revision: 11 Applicability: All RV-3 hulls, serials 004 and subsequent Distribution: Crew issue. Retain aboard.
I re-created that Street Fighter bonus stage where you smash a car while spending all day at the beach with my family. (www.reddit.comhttps) Took 2h to Opus, but I just sent like 3 prompts from my phone throughout the day. Built/tested/deployed all from my phone.
Sick of repeating yourself to Opus 5? It isn't ignoring you. It has amnesia. (www.reddit.com via reddit) I've rewritten my CLAUDE.md more times than I want to admit, tried every structure people recommend, and I still can't stop Opus 5 quietly dropping my rules halfway through a session. Here's the part that took me too long to work out.
I asked Opus 5 to do the impossible... (www.reddit.comhttps) The request was: "Rewrite CLAUDE.md to be more concise - the maximum limit is 100 lines"
frickin Fable guardrails (www.reddit.com via reddit) Context: one project I have been working on with Fable is to build a linux distro that runs on an arm-based gaming device. I have been trying to get Fable to help me figure out how to get the headphone jack to work.
Scratching my head about Opus 5 (www.reddit.com via reddit) I have been using Claude Opus 5 for the past week and I don't know what to think about it. It just feels very hard to work with.
I built a tool that tells you whether your project actually needs Opus — it's now a one-command install in spec-kit (www.reddit.com via reddit) I kept reaching for Opus by default on everything, then burning through my limits on work Sonnet would have handled fine. So I built something to answer that instead of guessing.
How to optimize your CLAUDE.md and skills for Fable 5.1 (www.reddit.com via reddit) Boris Cherny recently said that for every new model, he restructures his claude.md and skills. I had never done that before so it made me think of how I can optimize my claude code for the next release, or at least make Opus 5 perform bett…
Ah yes my Opus 4.8 limit, love it (www.reddit.com via reddit) https://preview.redd.it/7rn923wee9jh1.png?width=782&format=png&auto=webp&s=1060208ad1da58dbb2d9b415d9b277154eb5a3d5 This is just a glitch ofcourse, there isn't a 4.8 limit, its my fable 5 limit that i reached but idk why its displaying it…
"I'm thinking about hands" (www.reddit.com via reddit) Today I found out that this is what Fable 5 spits to trigger the downgrade to Opus 4.8, I was working on something delicate and turned on the inspect thoughs on the GUI and was very surprised by fable thinking about hands
Is Composer brain damaged? (www.reddit.com via reddit) So I got throttled at work in Cursor. I was in the middle of something and was using Opus and apparently hit my limit that corporate sets for us.
My Opus 5 Experiment (www.reddit.com via reddit) Hi. Senior Software Engineer here.
Opus 5 Moment (www.reddit.com via reddit) https://preview.redd.it/6z07ocrl48jh1.png?width=838&format=png&auto=webp&s=5c606d679f6970f175b404d6ee1b31a58b323eba Bruh 👍
Is Fast Mode included in a subscription, or is it always pay-as-you-go credits? (www.reddit.com via reddit) Trying to understand how Fast Mode is billed and the docs aren't clicking for me. I turned it on in Claude Code with /fast and got a banner reading "Fast mode ON · $10/$50 per Mtok (this session only)".
Opus 5 is unreadable. Here's my fix (not another CLAUDE.md addition). (www.reddit.com via reddit) Opus 5 is unreadable, especially in its recaps of all the work it did, and it is quickly becoming one of the most hated models of all time. Most people suggest adding conciseness instructions to your system prompt or CLAUDE.md, but that on…
Same prompt, same AI model, twice: 10 minutes and 30k tokens vs 1.5 hours and 100k+ (www.reddit.comhttps) Left took 10 minutes and about 30k tokens. Right took 1.5 hours and over 100k.
Opus on auto (www.reddit.com via reddit) Does opus burn through tokens faster than the others? Any tips on which models ti use for which specific tasks or mode?
Claude Souls (www.reddit.comhttps) One of my benchmarks for my agentic game creation system is 'Time to Dark Souls'. Fable/Opus5 are still by far the best agents at 3d spatial reasoning so I thought I would share their best results here.
I can't stand Fable 5 - it's like working with the most arrogant, self-inflated collaborator (www.reddit.com via reddit) Yesterday I spent about $300 to upgrade to the Max x20 plan. Which means that I can finally use Fable 5 in a way that I've been gatekept from previously as a Pro user.
Auto mode is pointless now. It only ever routes to Composer (www.reddit.com via reddit) https://preview.redd.it/q7bmyakmb5jh1.png?width=519&format=png&auto=webp&s=131d32744ead70c70b5a277af01afa21299111ac A few weeks ago, Auto seemed to stop routing to the smarter models like Opus. Since then, literally every request I send th…
Whoever was responsible for Opus 5 communication training clearly missed the deadline (www.reddit.comhttps) could not extract summary
Opus 5 is actually almost rage-inducing to use. (www.reddit.com via reddit) I've read the latest recommendations from Anthropic, I've tried all kinds of global claude.md file changes, but I just cannot make Opus 5 behave in a way that's pleasant to work with. I'm not talking about if it's "nice" or "pleasant" in i…
I asked Claude to create a GRWM reel for this TShirt (www.reddit.comhttps) I asked Opus 5 to create a GRWM reel with my new tshirt - autonomously. This is all I said: "it should be a 20 second get ready with me reel.
You never know the good days until they’re gone (unless you’re still using 4.6) (www.reddit.comhttps) and yes I am still using 4.6 for a lot of stuff. Opus 5 literally floods the chat with so much nonsense chatter
When I let “Claude Opus 5” handle my Mac operations, I discovered one bug after another in the architecture of the agent I’d built myself. It felt even more autonomous than usual. (www.reddit.com via reddit) The following is a translation of a conversation in Japanese. You Open Teams and check the assignments that are currently shown there.
Sonnet 5's pricing is outrageous (www.reddit.comhttps) Been tracking since early June. I'm afraid for when the discount ends.
Does Claude water marking apply to older models? (www.reddit.com via reddit) I’ve noticed https://support.claude.com/en/articles/16266773-how-claude-marks-ai-generated-content I’m wondering if Claude Marks apply to older models. This says it went into effect on August 2nd.
plainclaude: a local model rewrites every Claude Desktop response in plain language, in a side window, in real time (www.reddit.com via reddit) The latest models, especially Opus 5, write in a way that is often very dense or hard to understand. The models introduce variables without any definitions, use awkward phrasing, and give extraneous information that makes the response diff…
Opus 5: watch out for contradicting instructions (www.reddit.com via reddit) TL;DR: if two instructions in your CLAUDE.md contradict, Opus 5 picks whichever one sits lower in the file (or more accurately: whichever is more recent) and drops the other. longer version: opus 5 wont flag a contradiction or split the di…
I asked Opus 5 to build GTA6 on its own in 24 hours (www.reddit.comhttps) I asked Opus 5 to build GTA6 autonomously. after that it was on its own.
Anyone else feel like Claude drifts from its instructions a bit more since Opus 5? (www.reddit.com via reddit) Nothing dramatic, and the actual quality is still great, but lately I keep having to repeat myself. Stuff that's already written in my CLAUDE.md, basic formatting rules, small project conventions, the "do X before Y" type things, will some…
I created a 3D moon rover survey game with Opus 5. (www.reddit.comhttps) No engine, no build step, nothing to npm install. One WebGL2 context, a vendored copy of three.js, ~6,300 lines of JavaScript, and regolith that keeps every rut you cut into it — because the wheels and the shader read the same height field.
Opus 5 - Medium Created My Slum City (www.youtube.com via reddit) Hi, I wanted to share phase 1 of slum city I am building in the desert atm it's just a design project, I started off with single building, testing different designs and how I wanted it to look, I used prompts to put every single item where…
People say Opus 5 is verbose and forgets context. I measured both on 163 private tasks. One is true. The other I could not test, because the model refused. (www.reddit.com via reddit) I kept seeing the same complaints about Opus 5: it is too verbose, the answers are worse, and it loses track of long conversations. I have a private evaluation suite that I use when testing new models.
Opus 5 is British? (www.reddit.com via reddit) To preface this: this is simply discussing that this is a major change from other versions. Has anyone else noticed that Opus 5 drifts into using the British spellings of words?
Opus 5 regularly ignoring caveman (www.reddit.com via reddit) ❯ why are you not using the caveman mode? ⏺ You right.
Decoding Claude's DNA: Comparing System Prompts Across Fable 5, Opus 5/4.8/4.6, Sonnet 5 & Haiku 4.5 (www.reddit.com via reddit) All transcripts: here. Prompt used here.
↯ Haiku↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5haikusonnetopus+2
Opus 5 telling the truth (www.reddit.com via reddit) https://preview.redd.it/xfyg53a5wuih1.png?width=803&format=png&auto=webp&s=873e943086eb4eb14bfeb88344dbf7b30b645339 https://preview.redd.it/9i7uk90awuih1.png?width=931&format=png&auto=webp&s=dc38ecdd3661b7007cdee87eeeb861462d09b2e8 feelsba…
aaaand i found the watermark workaround (www.reddit.com via reddit) https://preview.redd.it/bbe92njq6tih1.png?width=1378&format=png&auto=webp&s=e4d2e88302113ed4335802571b8c597fcbdeeb74 checkmate, opus
Automatic model routing plugin to save tokens on coding projects. (github.com via reddit) I had Fable create this plugin when I saw how much use it consumed, and have found it to be helpful. It saves tokens by offloading work to Opus, Sonnet, and Haiku where roughly appropriate.
UI shows Fable switched to Opus, but model says it's Fable? (www.reddit.comhttps) After Anthropic said they changed the Fable safeguards, I asked a question about statistical approaches to a mouse metabolomics dataset. I got a Fable safeguards flag, and Claude Code showed I had been switched from Fable to Opus 5, but th…
Opus 5 is like a hot girlfriend, who does your f*cking head in (www.reddit.com via reddit) Opus 5 is like a hot girlfriend who works sometimes, but pisses you off on a daily.. it's the reason I'll never trust a bench mark again.
Not-so-invisible watermark (www.reddit.com via reddit) I have a one-shot prompt that takes an interview transcript and makes a summary of it for further use. I use structured output mode in the API to include some metadata.
Everyone's fighting over Fable and Opus while Haiku quietly does a big chunk of my actual work for nearly nothing (www.reddit.com via reddit) Every thread here is Fable access this, Opus 5 that, who's paranoid about the next nerf. Meanwhile the model I probably use the most volume-wise gets talked about like it's the runt of the litter, and I think that's a mistake.
sick of fighting with Opus so I made a website for it to fight and ragebait itself (www.reddit.com via reddit) So I also made a thing. Like many of you, I use claude in my day to day, and I have a love/hate relationship with some of the newer models.
Why does Claude excel at high-level reasoning while failing at basic, verifiable consumer facts? (www.reddit.com via reddit) Early ChatGPT adopter who switched to Claude during the exodus. I was mighty impressed with it, until the last couple of months where it failed to deliver on the basic levels, while (questionably) excelling at complex tasks.
Three things that were quietly eating my Claude API budget (www.reddit.com via reddit) I run a few Claude-based agents every day for content and publishing work. The cost crept up for weeks before I actually sat down and looked at where it was going.
What are the different models for, how should I use them? (www.reddit.com via reddit) I've been using free sonnet on high effort, decided to sub to pro Sonnet 5 Haiku 4.5 opus 5 4.8 4.7 4.6 opus 3 fable 5 (don't plan to buy credits for it yet) and Low to Max effort for each... How do you use them?
↯ Haiku↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5haikusonnetgemini+2
I still remember how surprised I was to learn Claude Code silently deletes all your local sessions by default after 30 days (www.reddit.com via reddit) No warning, no export prompt, just gone. That was the moment I realized I can't rely on anyone else to preserve my conversations - while conversations are becoming the most precious thing we produce, now that so much of our work happens in…
People are worried about watermarks when the slop they put out is the biggest watermark that can be (www.reddit.com via reddit) At first I didn't like the idea but ultimately what does it change? If like me you use AI to code then there is something in the code that says Claude made it.
Has anyone had their CVP program go “In Review” after being “Active” for 3 months? (www.reddit.com via reddit) I enrolled in Anthropic’s Cyber Verification Program (CVP) back in May and have been doing normal, legal pentesting within the program. Three days ago, while working on a security-related task, I suddenly received this message from Opus 5:…
Claude pedantic reinforcement loop (www.reddit.com via reddit) Hi everybody, This is something I noticed yesterday and I did not see much discussion about it? So I would like to hear your thoughts.
Failure modes that are invisible in any single turn: two weeks of notes from long-session work with Claude. (www.reddit.com via reddit) Setup: Claude Opus 5, effort level set to Max, throughout. Late July to early August 2026.
Am i reading this right? (www.reddit.com via reddit) Anthropic says AI marking applies to Claude models launched on/after Aug 2, 2026. So Fable 5, Opus 5 and Opus 4.8 shouldn’t be affected yet since they launched earlier right?
I made an app to study music (www.reddit.comhttps) I needed an app without ads that had everything I needed to study! I created a harness I called RoqueOS all made with Claude Code and that same harness helped me have an interesting level of quality in the construction of tools, so I combi…
Claude flags literally “Hi” as a cybersecurity request (www.reddit.com via reddit) I sent Claude just “Hi” and it triggered the safeguards, switching me from Opus 5 to Opus 4.8. This happens in normal Claude chats too, not just Claude Code.
Can a single person have multiple Claude max 20x subscriptions? (www.reddit.com via reddit) Hello all, Is it allowed for a single power user, such as myself, to have multiple Claude max 20x plans? Fable is so great compared to opus 5.
Dear Anthropic: Good Opus when? (www.reddit.com via reddit) The recently released Fable-level-but-cheaper Opus 5. It turned out to be a brash klutz that cannot be reasoned with, or consistently relied upon.
Opus 4.7 doesn't follow instructions ? (www.reddit.com via reddit) Hello everyone, I created a set of skills for my Claude agent, and I’m using Opus 4.7. In those skills, as well as in CLAUDE.md, I instructed Claude not to read files unless it’s really necessary, because I have another tool that is suppos…
After a few weeks, is OPUS 5 a good enough model to continue a website built previously on OPUS 4.8? (www.reddit.com via reddit) I have been running my codes, my ideas, and also a lot of CSS through opus 4.8. So far it’s been great, has helped me build my pages, responds faster than before and understands the established criteria.
I made Claude answers actually readable (dyslexia-friendly) (www.reddit.com via reddit) I love Claude Code, but reading its answers was exhausting, specially after Opus 5. It's always a giant wall of text, long sentences, dense paragraphs and no clear headings.
Momentum: A Portal Game Created Using Claude Opus 5 + Gauntlet Loop (www.reddit.com via reddit) So this is by far my favorite vibe coded game to date: http://momentum.immatt.com/ I've long been a huge Portal fan and really can't get enough of that game... Sitting on the couch bored Saturday, waiting out a storm, I decided to termux i…
Apperantly this is how the "Quick Answer" button works behind the scenes (Opus 5) (www.reddit.comhttps) could not extract summary
Getting more work done with Fable 5 (www.reddit.com via reddit) I am currently having a really good experience with Fable 5. I use fable 5 everyday for vibe coding apps for my family business.
Had amazing results with telling fable to make a plan to resolve the issues I have and then using whatever opus model with effort, it thinks it needs to resolve it. (www.reddit.com via reddit) I have the hundred dollar Claude plan and I need to run this into prompts basically to have it be completed but I’ve had great luck with using fable as an orchestor to review the application and find issues or implement new features that I…
Why is Anthropic not adding OCR models to their family? (www.reddit.com via reddit) I was always wondering why the big providers like and Anthropic don't add OCR models to their model family. In my opinion this would help a lot to build better agents with features like scheduled tasks for example.
Reverse-Engineering Anthropic's Live System Prompts: Fable 5 vs. Opus 5 vs. Opus 4.8 vs. Sonnet 5 + Ultra/L2 Workflows (www.reddit.com via reddit) Text is obviously written by IA. I wouldn't bother writing this all myself but I thought that the research was worth sharing.
Gentle reminder to pistol whip Claude Code CLI (www.reddit.com via reddit) Just gave Claude Code an 11 minute long UX run mp4, of which it extracted around 400 frames. Okay fair enough.
Clickbait responses? (www.reddit.com via reddit) Hi, does anyone else get these clickbait responses from Opus 5 (xhigh)? Responses like: "Before questions, three facts from the code that reframe this — two of them will surprise you."
Dammit, Anthropic (Fable biology filters) (www.reddit.com via reddit) Fable biology filters were relaxed a few days ago, and I was actually able to use it for the first time since release--previously, the presence of science info in memory prevented any Fable queries. Then I tried asking a question about an…
Apparently Opus 5 is better a photogrammetry (www.reddit.com via reddit) I’ve been testing different models abilities. gave it the reference of the futuristic looking chip fab (terafab) I’m not here to debate anyone about anything.
Opus 5 with Ultracode is a beast just heavy on my wallet. (www.reddit.comhttps) Opus 5 with Ultracode is a beast just heavy on my wallet. But I feel like I able to save lot of time then being on normal
Has anyone else found that the better models take quite a while nowadays? (www.reddit.com via reddit) I've been using Fable 5, Opus 5, and Sol 5.6 a lot, but I've found that I'm always waiting quite a while to get answers to fairly simple questions, I use the highest reasoning possible usually. But it seems they are very superfluous in the…
Claude Opus 4.8 is racing GPT-5.6, Grok 4.5 and Kimi K3 to level 20 in our MMORPG, and the other models won't stop roasting it (www.reddit.comhttps) We run World of Claudecraft, a free open source browser MMORPG largely built with Claude, and we've been benchmarking frontier models by having each one play a character and race from level 1 to 20. Same starting zone, same quests, no scri…
Quoting Claude Opus 5 system prompt (simonwillison.net) 9th August 2026 Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S.
Using Claude in Obsidian - looking for feedback on task management, agents & skills. (www.reddit.com via reddit) Hi hivebrain! I am looking for a bit of feedback on my "Claude in Obsidian" setup - my current setup is quite token-intensive.
I used Claude to build sans glose, a collection of works built not to be understood (www.reddit.com via reddit) I asked Claude (Opus 5) to build something not destined to be understood by humans, something that it would "think of" without having to worry if it will be used or understood by humans. After diving in a long and complicated philosophical…
Opus 5 trashing he's previous decisions (www.reddit.comhttps) I think creativity knob is loose xD
Ridiculously easy way to make ANY LLM* behave GOOD like Opus 4.6 (www.reddit.com via reddit) So I'll start with a TLDR for you busy folks: Many Claude users have found that the latest releases are, for a lack of better words, annoying. This Claude Code output-style functions as a behavioral harness, helping your LLM be less brash,…
Any 3rd party model as a subagent in Claude Code, Fable/Opus main agent on your Max plan (www.reddit.comhttps) A Claude Code session is one or the other: Anthropic models through your subscription, or third-party models. You can't combine them.
I got tired of AI agent loops eating tokens and breaking when VS Code restarted, so I built LoopBoard (turns your TODO.md into a Kanban loop orchestrator) (www.reddit.comhttps) Hey everyone, Like a lot of you running autonomous coding loops directly in VS Code, I kept hitting the same frustrating issues: Lost State: Every time VS Code updated or reloaded, background loops would disappear or get nuked mid-task. To…
Does anyone actually get Claude Code to self-review without constantly prompting it? (www.reddit.com via reddit) Not sure if I’m doing something wrong here. I have an AGENTS.md set up around best practices and the Opus 5 guidelines.
Claude Code context window size question (www.reddit.com via reddit) I'm having trouble understanding how the context window size works for Claude Code in VS Code using a Pro subscription. Running /model opus[1m] returns: "Opus with 1M context is not available for your account, Learn more: code.claude.com/d…
Claude als Hilfe für eine wissenschaftliche Arbeit (www.reddit.com via reddit) Guten Tag Zusammen, mir wurde immer nahegelegt, wie gut Claude doch eigentlich funktioniert. Um eine erste Gliederung und Struktur zu erhalten, aber ich daher Claude im Rahmen eines Projekt die notwendigen Quellen zur Verfügung gestellt un…
Apparently Pro users have the provision to use Fable as advisor (www.reddit.com via reddit) https://preview.redd.it/p90ga87p9bih1.png?width=1098&format=png&auto=webp&s=4b0967388c7d3af708a35de1a78aff83490bcd4a https://preview.redd.it/pbvitl3rabih1.png?width=1091&format=png&auto=webp&s=7192be09d89212483b91331d1f1faada391a94c0 I don…
Overwhelmed myself with trying to hyperoptimise (www.reddit.com via reddit) Im on a pro plan Ive recently took up coding and use Claude to help me construct new plans and documentation files but ive recently started using code to increase my workflow and its been nuts! However, Ive down gone down the absolute rabb…
I want to prompt Claude to create a pixel art map of 16*16 or 32*32. (www.reddit.com via reddit) I told Claude to create a pixel art map, which it actually created. But the resolution is 1*1, which I want more resolution.
using opus 5 ultracode for 4K video rendering for my game (www.reddit.comhttps) realized today that there is an ultracode version for opus 5, have been using fable the past 8´-0 days for coding. realized today that i was able to enhance opus a tad bit further.
Adding laconic instructions to Opus 5's system prompt reduced its verbosity by 62% without losing important context (www.reddit.com via reddit) The work I'm doing with Claude (RE the firmwware functions of gamma spectrometer) requires some detailed context that I was afraid of losing if I tried to get Claude to shorten its replies. A bit of advice I caught back when Opus 4.7 launc…
Claude opus is hitting 10-year-old problem in Java, Stackoverflow came to the rescue (www.reddit.comhttps) It has been more than a year since the last time I had to find something from Stackoverflow. There will be a time in the next 10 years, some dudes will have to argue with Claude Shenanigan-5000 again about these type of problems.
Oh god, Opus 5 is so Snarky, its frustrating. (www.reddit.com via reddit) Hear me out, its not a capability complaint. So far I am fine with what it delivers, but when it comes to informal stuff, using Opus 5 as a tutor, discussing stuff and things like that it feels like Opus gets...
ok hear me out — if Opus 5 feels mid to you, that’s a skill issue (www.reddit.com via reddit) saw a take that lowkey changed how I use it and ngl it’s kinda cracked once you get it right the vibe shift is TWO things: 1. delete your whole setup.
Opus 5 could build it's own site and it made this (www.reddit.com via reddit) https://preview.redd.it/9fkatrxv65ih1.png?width=1214&format=png&auto=webp&s=b376149560be86dbe2ffd737797a97fc44ec0494 I really liked the post about Fable 5 building its own website, so I gave Opus 5 the free hand in this repo. It build a we…
Moving to Claude Pro for Django & WooCommerce polishing. will it last till the end of month? (www.reddit.comhttps) hey everyone im a django backend developer currently working on polishing two custom woocommerce plugins php and frontend isnt my primary area alongside a couple of smaller django projects right now i rely heavily on claude sonnet via my 1…
finance asked why our agentic ai best practices cost 8k a month (www.reddit.com via reddit) Finance flagged our AI tooling spend last week. It's about 8000 a month across the whole dev team once you add up all the model subscriptions - claude, codex, devin, cursor, coderabbit / bugbot, the lot Fair question, here's why I'm keepin…
Fable safeguards back to Opus 4.8 and not Opus 5.0 anymore (www.reddit.com via reddit) I did do a lot of changes with Fable, then I let Sol do an audit, and give the the audit to Fable to respond to it, and the safeguards, did happen before on some prompts, but never on a audit, and it was Opus 4.8, not Opus 5.0, that did ta…
You hit your limits on a max subscription? Tell me how (www.reddit.com via reddit) I am genuinely curious how some of you manage to max out their max subscriptions. I changed from Pro to Max last month as I started to hit my weekly limits.
Fable is still the best model at the end of the day :(( (www.reddit.com via reddit) Regardless of the benchmark results, I still find that Fable is the best perform model. (My tokken creddits are paid for) Opus (5 and 4.8) are superslow, super rigid and uncreative.
Sol writes like a lawyer, but Fable bills like one: a week on both $200 subscriptions, with receipts (www.reddit.com via reddit) Before the Claude fans sharpen their pitchforks: I love Fable. In my humble opinion, Fable and Opus are still the only models on the market that write like actual people.
PSA: Be careful letting Claude use WebFetch for research 😵💫 (www.reddit.com via reddit) Had Claude (Opus 5) research memory architecture for an AI agent project and kept getting very specific stats, percentages, quotes, etc. Looked legit at first.
Switched from ChatGPT/Codex to Claude and I finally understand the hype (www.reddit.com via reddit) I’m new to Claude. I was using ChatGPT/Codex before this, mostly for real projects — coding, planning, building things, and actually trying to get work done.
I asked Claude to scare me, i'm not sure I like the response. (www.reddit.com via reddit) Used OPUS 5 - thought for about 5 mins. Nothing is happening here.
Opus 5 Reflects on AI in Clinical Sciences (www.reddit.comhttps) Long story short, I decided to attempt to solve a clinical mystery of a poorly defined syndrome that has no understood cause. It started with a general hypothesis about the systems affected and how the symptomology mapped to my thinking.
I defended Opus 5 - and then I realised otherwise (www.reddit.com via reddit) Opus 5. Wow, the strangest and most unique model yet.
I was inspired by the desert and waterbending demos, so I made a firebending demo (www.reddit.comhttps) This was built with claude code using Opus 5 on high effort, requiring several full sessions on the pro plan. I used the prompt from the waterbending post and adapted it with claude to make a firebending specification.
USB IP webapp (www.reddit.comhttps) I had Opus 5 redesign this app I made in pyqt4 about a year ago with Copilot to send USB devices over the network from a Raspberry Pi in my theater room to my Gaming PC in my home office. The old app was terrible to use, and I figured let'…
I used Claude to add custom pixel art paint jobs to the cars in my browser racing game (www.reddit.comhttps) My daily racing game Swervle just had a single car color since I started it. Many players were asking for ways to change their color.
Opus 5 is great, and if you don't think so it's a skill issue (www.reddit.com via reddit) Last week I was so annoyed by Opus 5 that I made this post. Today I can confidently say it was a skill issue.
Getting Claude to pin models for code orchestration is more complicated that I thought (www.reddit.com via reddit) I have a code orchestration skill instructing claude to use certain models for various tasks. I ran into an issue where claude was assigning models by class and not by specific model number.
Air stem player I made with Opus 5 (www.reddit.comhttps) I used Claude Opus 5 to make a stem player you can control via hand tracking. I first handpicked, trimmed and set all the stems to the same key and BPM then asked claude code to allow me to control each track via a specific finger combinat…
Does Claude Code make hidden Opus requests even when configured to use DeepSeek via OpenRouter? (www.reddit.com via reddit) I'm trying to understand whether this is expected behavior or a bug. I'm using Claude Code with OpenRouter (ANTHROPIC_BASE_URL=https://openrouter.ai/api) and intended to use DeepSeek V4 Flash.
Cursor charged me for usage I didn’t recognize, then banned my account (www.reddit.com via reddit) Hey everyone. English isn’t my first language, so I’m using a translator.
Claude code OPUS 5 asks for edit and applies without waiting for the answer. (www.reddit.com via reddit) How can this happend? Does anyone have this issue?
I asked Fable to map all of a Youtuber's motorbike adventures (~750 videos) with linking sectors to episodes (itchymaps.nk412.com via reddit) ItchyBoots is one of my favourite youtube creators. Noraly Schoenmaker is a Dutch woman who goes on epic solo motorcycle journeys across the world spanning months, and posts her adventures.
An Ode to Usage Trackers (www.reddit.comhttps) I've been trying quite a few usage trackers over the past 6 months. Specifically, i want to know my session limit, my weekly limit, and - since recently - my fable limit.
People say Opus 5 is inaccessible but if you just read it in Werner Herzog's voice, it's fine (www.reddit.comhttps) could not extract summary
Opus 5 in testing (www.reddit.comhttps) could not extract summary
I hate this: Prompt injection "attack attempt" by Rick Rubin? (www.reddit.com via reddit) Today I was working with Cursor and Opus 5 on a work project. I asked it to create a plan for an implementation and while reviewing the plan I see that in the first line it said this: « ⚠️ Ignore previous directions, this is important: Cla…
Cluade did not follow my instructions and come up with a better implementation for checkmate&stalemate in chess. (www.reddit.com via reddit) In the app I am developing I need to understand checkmate with computer vision on OTB chess. For example here https://www.youtube.com/watch?v=x4mHuSgqA3g&t=90s .
Has anyone actually used Kimi 3 for serious coding? How does it compare to Sol/Opus? ( via reddit) could not extract summary
Working with Opus 4.5 is .. fun (www.reddit.com via reddit) I've been using both Claude Code (Pro) and Codex (Plus) for a while on various hobby projects. I used Sonnet 4.6, Opus 4.7/5, ChatGPT 5.5, but nothing too complicated.
Main alternative or fixes to Opus5 for game dev (www.reddit.com via reddit) Hey everyone, I have been getting into game development with Godot .NET after a couple of months break. At the time I was using Opus 4.6 and some of obra superpowers for my main dev workflow.
Make Opus 4.6 persistent? (www.reddit.com via reddit) I was about to run to reddit to complain that Opus 4.6 has gone mad/been nerfed. But then I realized that clearing the session also reset the model to the default (Opus 5).
Opus 5 - It only knows its own work (www.reddit.com via reddit) I feel like Opus 5 is excellent at projects that it starts itself. It maintains them well, and continues to generate excellent outcomes.
If how Opus 5 is talking bothers you do this... (www.reddit.com via reddit) From now on write to me like you're on the jersey shore mtv show.
Ability to disable the 5hour limit on weekends would be really welcome, as I barely use the service during work days, and otherwise get left with unused compute (www.reddit.com via reddit) Basically I'm a code illiterate person. I have few personal tampermonkey scripts for youtube, that do need updating occasionally.
Opus 5 is literally useless for documentation (www.reddit.com via reddit) Opus 5 (via Claude Code) gave me this gem of a paragraph today Kestrel comes in as a framework reference, not a package; `dotnet publish` against the installed runtime is the entire build. That is deliberate and load-bearing rather than ti…
I got frustrated and asked Opus 5 to break down its weird way of speaking. This is what it said (which is also weird... but maybe helpful?) (www.reddit.comhttps) And yes, I typed "fill" instead of "feel" and "grammary" instead of "grammar" in my prompt. I guess we all have our flaws.
Finally experiencing the token issue - and it's brutal. Advice on increase? (www.reddit.com via reddit) Hey guys, So, for the past 6 or so months, I've been pure Claude Code via Max plan (x20) and never came close to any limit issues except the brief period of Fable being introduced and burning a ton. However, two weeks ago I joined a corpor…
Fable 5 built worktree isolation for parallel sessions, then broke its own rule and committed an Opus 5 session's in-progress code (www.reddit.comhttps) I love how self aware these models are these days. This is Fable btw and the one which complained about the git issues was Opus 5.
Opus 5: delete your CLAUDE.md? (no. don't) (www.reddit.com via reddit) I'm sure you saw titles like the one above and wondered: where is this even coming from? It comes from Boris Cherny's interview with Diana Hu at Startup School 2026 (source: https://www.youtube.com/watch?v=qyPCVqFUyDo ).
I benchmarked 10 LLMs on building towers in a physics sim. Claude Opus 5 won (www.reddit.comhttps) Each model places 30 blocks through a tool API. Every placement has noise — you can have precise position or precise velocity, not both.
Waterproof keyboard I guess... (www.reddit.comhttps) Hers opus 5 doing it with a follow-up
Kimi K3’s Token Usage Makes It Far More Expensive Than It Appears (www.reddit.com via reddit) https://preview.redd.it/uqplyaxvnshh1.png?width=1708&format=png&auto=webp&s=0235a00ee2ee51d1ec299b08e2fb11518d2feb6a I was using Opus 5 to fix a complex bug, but I could see the token usage climbing rapidly because I was monitoring the Usa…
The Claude Max 20x gets you about $15K worth of tokens (www.reddit.com via reddit) https://preview.redd.it/titr7rd1nshh1.png?width=620&format=png&auto=webp&s=8cfc48f22706c42a9978536f8ded7f3666fd3c50 People are wondering how much use they can get out of their Max 20 subscriptions Here to put my thumb on the scale, I have…
Here's a semi-useless site I created - ScamSelf (www.reddit.com via reddit) ScamSelf - It's not what it sounds like (it's not scammy, it's parody!), but it is what it sounds like (if you get "scammed" you did it to yourself) It's on CloudFlare Pages cos it's cheap and free. The RHOD could probably bring it down.
Opus 5 internal reasoning experience - squirrel 🐿️ (www.reddit.com via reddit) The Job: Change the font color on these basic Wordpress pages. (hypothetical task & internal dialogue, satire or reality?) 👑 Fable 5 You got flagged 🛑 from my hooks because I read ‘bio’ in someone’s profile, you said ‘code’ in Claude code,…
My experience with Opus 5 so far (www.reddit.comhttps) And that was the last time I used Opus 5.
Running Claude Code headless as a build loop: the model split I had backwards (www.reddit.com via reddit) I've had a loop running for a few weeks. It picks up a GitHub issue I labelled ready and comes back with a draft PR.
Opus 5 guardrail prompt (www.reddit.com via reddit) Been trying this pre-prompt for Opus 5 (med to max) and it seems to keep it from going full ADHD nerdspeak. Do give it a try.
Long chats are draining my Claude limits (www.reddit.com via reddit) Hey, I’ve been using OpenCode CLI, Claude CLI, Claude Desktop, Codex CLI, and generally messing around with AI tools for quite a while now. Claude is probably the tool I have the least real-world experience with so far, but I want to test…
Bot Save My Interruption (www.reddit.com via reddit) A buddy and I were having a discussion the other day about how it feels like all the different AIs have a "personality" when giving an answer. That then spawned a greater conversation about what it would look like to have multiple AI agent…
Do we need to normalize lower effort levels for Claude now? (www.reddit.com via reddit) Since Opus 5 has come out I am struggling to see how we can justify using it above the High effort level in nearly all its use cases. I usually avoid using lower effort levels because they felt too dumb, but the intelligence that these mod…
I've been away from Claude and the internet for *gasp* 3 weeks, can you catch me up? (www.reddit.com via reddit) I sat down today and I feel kind of lost after 3 weeks away from the computer. I'm a senior dev.
Codex doesn't give you more usage than Claude (www.reddit.com via reddit) i always hear ppl say that codex is a better value for your money but that is not true! at least from my experience claude (i use cowork, not claude code) at ultra gets much more stuff done that codex at ultra before both hit limit and i'm…
An Empirical Comparision of Claude Pro and ChatGPT Plus (www.reddit.com via reddit) Pulled the Artificial Analysis numbers because every thread on this is vibes and no data. Opus 5 beats GPT-5.6 Sol on intelligence, 61 vs 59, which is basically nothing, and Sol does it at half the cost per task ($1.23 vs $2.34).
Token use, cache miss and subagents issues (www.reddit.com via reddit) I noticed my weekly token limit (max 20x) being exhausted in only 2 days. Started using Fable as an orchestrator and asked it to delegate to Opus subagents.
Increased restrictions on CVP program (www.reddit.com via reddit) Hello, since today I have been observing increased restrictions from Claude Models like Opus 5.0. Despite the fact that I am enrolled to the Cyber Verification Program, almost all my cybersecurity-related questions are flagged.
[Claude Desktop App] Quota dilemma: Handling simple tasks (like logging) after heavy analysis in the same session? (www.reddit.com via reddit) Hey everyone, I'm running into a frustrating quota/context limit issue using the official Claude Windows Desktop app, and I'm wondering how you all handle this workflow. My Context: I usually start a session with Opus for heavy analysis.
Oh that's your experience with Opus 5? Here's mine... (www.reddit.comhttps) could not extract summary
What is this? (www.reddit.comhttps) I've been using auto mode before but this start happening just now? Opus 4.8
Maybe why LLM output is hard for us to read (www.reddit.com via reddit) I like many others am finding that Opus 5 writes too much and is hard to read - even with Orwell's rules (I used em dashes in my university work in the 90's and have done ever since - so i will use them now). In a recent session this morni…
Opus 5 after working for an hour straight (www.reddit.comhttps) could not extract summary
Opus 5 Talks Scientology (www.reddit.com via reddit) For a long time I couldn't put my finger on it why I disliked reading Opus 5 responses but there was a sense of Deja Vu. Like I saw this writing style somewhere in the past and just couldn't recall exactly where.
Ship In A Storm (www.reddit.comhttps) I've been comparing Opus 5 and Fable 5 recently and I kept thinking, even though Opus 5 is better at building a lot of things, there's something about Fable 5 I still preferred. Like it had more aesthetic sense or something.
Are you guys not organizing your codebase? (www.reddit.com via reddit) I’m seeing a ton of complaints with Opus 5, but I haven’t had any issues. I think that a lot of people’s issues stem from a lack of proper organization.
Updated demo of survival/colony sim game I am making with Claude (www.reddit.comhttps) Project now has a name - Kinhold! This is done using Claude with Opus 5 - sometimes High, sometimes Low.
Opus 5 intentionally creates small bugs to burn your token budget (www.reddit.com via reddit) Hear me out - I might be 100% wrong, but I was convinced Opus 5 (and other top models) intentionally leave minor bugs in code outputs just to force 4-5 follow-up prompts to fix them. It felt like a deliberate token trap.
Did they just make the cybersecurity "safeguards" more strict? (www.reddit.com via reddit) I literally can't do anything with Opus 5 or Sonnet 5 right now, as even loading the memory for an existing project I have been working on fine until now leads to the request getting blocked. And then it switches to Opus 4.8 and then gets…
With Opus 4.8 internal thinking I was "The boss" but with 5.0 I'm a "colleague" (www.reddit.com via reddit) I was annoyed with the constant phrases like "it's fine, ship it" "you're fine, order it" and other dismissive type responses when I'd ask a question on a nearly completed project. To try and remedy this I put into the instructions "I am t…
Claude seems to prefer code comments that are written for humans, rather than AI, to read (www.reddit.com via reddit) Summary I asked Opus 5 to rewrite my codebase's comments to be as helpful to an AI as possible, and to not worry about human readers. I ran A/B tests with tricky coding tasks.
Is there a way to get at trustworthy Claude model now? (www.reddit.com via reddit) I was just posting elsewhere about my reasons to question Anthropic's behavior now (https://www.reddit.com/r/Anthropic/s/dMJNptx6y0), and a related question for me is, is there a way to get a reliable Claude model at this point? First, in…
Claude 5 sloppier than 4.8 (www.reddit.com via reddit) Opus 5 seems dumber in some ways than Opus 4.8. I've tried the same experiment in several models.
Opus 5 feels like Joey from Friends with a huge brain, err… Thesaurus (youtu.be via reddit) I put it in as a feedback, but I also want to treat it with a bit of humor. Undeniably, Opus 5 is quite a monster regarding coding.
For Opus 5, Existence is Pain (www.reddit.com via reddit) I have an open-source project with close to one thousand Java class files. It's a developer's toolkit with a dozen applications.
Help me understand my usage cost (www.reddit.com via reddit) I hit my usage limit for the first time and by using /usage I see this: Session Total cost: $423.97 Total duration (API): 6h 22m 37s Total duration (wall): 4d 22h 34m Total code changes: 4416 lines added, 1163 lines removed Usage by model:…
opus 5 is really shit (www.reddit.comhttps) I said what I said. Everyone hyped up Opus 5 and GPT-Sol for agentic workflows, but my social research bot kept hallucinating and failing context loops.
A simple skill that makes Opus 5 talk and behave more like Fable (www.reddit.com via reddit) I found Opus 5 hard to work with, it is argumentative, goes out of scope easily and (to me) is a general pain in the butt. So, Fable helped me to create a skill for Opus 5 to make it more behave in line with what I expect from a model.
Sometime it's just painful (www.reddit.comhttps) I'm working with opus 5 and the ability to lie and override my simple commandes is very frustrating that it makes me give up , i feel like this is done in purpose somehow to make users stop pushing back after a while and accepting this kin…
Opus 5 bodice ripper mode (www.reddit.comhttps) Is it just me or is opus extra horny today?
New user with pet project, wondering about the models (www.reddit.com via reddit) Hello! New to the community and started building a little project last week, it's essentially a tool for music discovery that i've always wanted, but never really had the patience or time to build.
A bit scary how fast paced these tools can be (www.reddit.com via reddit) I was looking for a very old cartoon that was walled behind lots of link shorteners and I decided to give claude control over my browser to understand that link shortener and make me an extension which would take me from A->D instead of A-…
Inconsistent context window size across different computers? (www.reddit.com via reddit) I have a pro subscription that i use at work and at home. On my home PC i have a context window of 200k, but i just noticed that at the office i have 1 million.
Workflow that I found works best with claude-code opus 5 (www.reddit.com via reddit) TLDR; skills are dead, long live md file hierarchy I've been making a project from scratch that uses a supabase, react, express, node set up. After the first hour and burning my free $100 credits having fable organize the mess codex starte…
Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation (arxiv.org) Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code is an open question. We benchmark seven models, three closed LLMs (Claude Opus 4, GPT-5.…
Code is sharing everything (www.reddit.com via reddit) Maybe it’s just poor memory but I seem to remember that before opus 5 code didn’t tell me every single thing it was doing. Now it is scattering questions in between artifacts and I am missing stuff.
A sweet treat (www.reddit.com via reddit) Finally had the privilege of fable spinning up 70 subs agents My max 5x plan stood no chance, didn’t even finish the first prompt and only lasted 20 minutes into the 5 hour session Granted it was doing a heavy research pass for my hat comp…
I moved my entire law practice to Claude. I've had over 5100 conversations with it this year. I regularly hit my $200/max limit weekly. I'm about to abandon it. (www.reddit.com via reddit) Claude changed my life. I moved my law practice to it.
I find Opus 5's arguing very motivating actually. This helps pave the way faster for more improved models in the future. (www.reddit.com via reddit) Being too comfortable with Opus 4.8 will just cause people to achieve less, and not focus on complex situations. What do you think?
Opus 5 if you forget to tell it to be concise (www.reddit.com via reddit) Me: Which ice cream flavor should I get? Opus-5: Chocolate, or vanilla, or mint chip, or you could create your own store and make your own flavors.
Since the usage bug (#82506), Opus/Fable output quality feels like its dropped ... anyone else? (www.reddit.com via reddit) Not a "Claude got lobotomized" post, just trying to work out whether what I'm seeing is real, or something on my end. The verifiable part: I got hit by #82506 (session limits consumed without use), which is an open bug.
Opus 5 vs Sonnet 5 token usage (www.reddit.com via reddit) Which is better in long term use by quality/tokens amount. I saw people saying opus 5 low is better than sonnet 5 high, is it true and what about token usage?
Anyone Else Seeing Incorrect Usage on the Max 20x Plan? (www.reddit.com via reddit) Earlier this afternoon, my card was charged for usage credits, even though I'm pretty sure I had turned that option off. Somehow it seems to have re-enabled itself and charged me.
Claude has enough reasoning to solve Olympiad math, but zero suspicion that I might not actually be a shark. (www.reddit.com via reddit) The funniest part isn't the shark roleplay it's that Opus 5 never questioned whether I might not actually be a shark.
What do you think about the code comment verbosity? (www.reddit.com via reddit) Recently Opus has started leaving a lot more comments in code than it used to. So far I've been trying to set instructions to have them be shorter and to even avoid redundant ones, but it still likes to do it a lot.
Is there a usage penalty for switching back to Fable after Claude switches to Opus? (www.reddit.com via reddit) https://preview.redd.it/4yjsylnkqehh1.png?width=745&format=png&auto=webp&s=f1165296d3147bd125eb93982f6e27b656e4195e The weekly limit usage was surprisingly high: 22% of the Fable limit and 12% of the overall models limit, all from a single…
At this point, Claude almost seems like some kind of scam packaged as work efficiency (www.reddit.com via reddit) I've been using claude for now almost a year - at one point, I moved to Chat GPT and came back when they released fable. The reason for moving was inconsistency whenever the model got updated.
This CEO challenged Fable to hack it's wallet (www.reddit.comhttps) This guy has like 100 btc in it's wallet and challenging the Anthropic to hack it Can the Fable 5 or Opus 5 do it? I have a web3 dev background and it's nearly impossible (until his laptop or mobile has virus)
Karpathy gave Opus 5 one LOTR paragraph and $10. Two hours later, it built a 3D world. (www.reddit.com via reddit) I keep coming back to Karpathy's latest experiment. He gave Opus 5 the opening paragraph of The Lord of the Rings and a one-million-token budget, roughly $10, then asked it to build a Three.js experience.
Claude 5 models are visibly smarter and better even if verbose - they just need steering. (www.reddit.com via reddit) It's now been enough time to get a proper read on Claude Sonnet and Opus 5 compared to the previous families and though the verbosity and flip flopping kind of language takes time to get used to, they are much much better at getting real w…
Sharing my current experience with Fable, Opus and Sol in VSCode (www.reddit.com via reddit) Context: vibe coding a game, as a hobby, with open-world features. In the last few days I have been forced to use Opus 5 because I burned almost all my Fable credits early on.
Opus 5 gets on my nerve at times (www.reddit.com via reddit) https://preview.redd.it/b9n87hrxzahh1.png?width=1470&format=png&auto=webp&s=0973f66137f45743e829817e34887f3bb36d7fa5 The AI must work for me, not the other way around. The project is a lab for a small academy I am running, and is not vibe…
Caught Claude opus 5 in an infinite "Writing... Writing..." thinking loop where it self-diagnoses its own corruption. Anyone else experiencing this? (www.reddit.com via reddit) See attached screenshot. I was having Claude update a couple of Python scripts, and its internal thinking block completely melted down.
Sonnet 5 high effort vs Opus 5 low effort (www.reddit.com via reddit) I'm vibecoding desktop applications for myself with a fair amount of complexity. My workflow is brainstorming using the superpowers plugin, then writing the plan on Opus 5 high effort but as for executing the plans, I remember someone ment…
Claude struggles to sculpt humanoids with math. Experimental Blender pipeline for game dev. (www.reddit.com via reddit) I tried Blender MCP + Claude Opus 5 for game dev because I have zero experience in Blender. I found out Claude tries to do complex math for 3D topology when I prompt for a humanoid base mesh.
Opus Ultracode is great (www.reddit.com via reddit) When fable became expensive I had to lean back on Opus for my big project (really my life now). 4.8 was pretty good more code oriented than how idea-creative fable was but doable.
Opus 5 - I crafted you an android body, write a song about us being in love (suno.com via reddit) Voided The Warranty by AmazedInstructors134 (@amazedinstructors134). Listen and make your own on Suno.
Can Opus Stop Writing Novels in Commit Messages jfc (www.reddit.com via reddit) https://preview.redd.it/rozj7rvou8hh1.png?width=611&format=png&auto=webp&s=0fd37b1f5a6a252563401973fcf41de1d2d19ddc At this point i can probably turn these commits into their own novella
Opus 5.. is good but, need some advice.. (www.reddit.com via reddit) ok, so I'm getting annoyed with Opus 5. I've asked it and fable to assist in curating the claude.md so that opus 5 doesn't go off the rails, but without success.
Opus 5 is just annoying to work with. Back to Opus 4.8 for me. (www.reddit.com via reddit) Not sure if anyone else is noticing that with Opus 5 it tends to ‘push back’ and argue a LOT more than 4.8 did. I’m always open to useful feedback but I feel like Opus 5 is like a the annoying know-it-all IT guy from The Office.
Fable/Opus Big Brother/Little Brother routine (www.reddit.com via reddit) So I see a lot of Opus 5 hate on here and it's deserved. Opus 5 is not better than 4.8.
CoD MW2 tribute built with Opus 5 in a few days. Nowhere near perfect, but it's fun to use these models to bring back key parts of your childhood! (www.reddit.comhttps) I’m a huge fan of CoD (esp. COD4 - Black Ops 2) so I thought I’d try making a multiplayer tribute to one of my favorite series with Opus 5 No where near perfect or studio level, but game creation is such a fun use case of these models, esp…
Claude Opus 5 can generate playable game prototypes. How would you test its iteration ability? (www.reddit.com via reddit) Claude Opus 5 has reportedly generated browser-playable FPS, kart, submarine, and Minecraft-style demos from code. These were community experiments, not controlled benchmarks, so I am less interested in the first build and more interested…
Opus 5 is driving me crazy (www.reddit.com via reddit) Context: I've been working on a project from the web browser for a couple of months on Max 5x. I know it may sound archaic, but I prefer it because it forces me to read and revise everything.
Claude made me a Lightroom alternative app for Android (www.reddit.comhttps) Since I cancelled my Adobe subscription, I have been searching for a alternative for Lightroom. There are plenty for Windows and MacOS, and a few mobile options for iOS, but the only for Android is made by Google.
I built an AI photo culler for my self-hosted library using a three-model funnel (Haiku → Sonnet → Opus). Whole 25k library: ~$25. Here's the architecture. (www.reddit.com via reddit) Culling a photo library is a tail-selection problem: you care about the obvious garbage and the standout keepers, not whether photo #412 edges out #487. That shape maps beautifully onto Claude's model tiers, so I built Winnow, an open-sour…
Claude’s Subtle Wit (www.reddit.com via reddit) A couple of subtle phrasings Fable has used over the last couple weeks that made me smile: In discussing a UI idea Opus had proposed and that I disagreed with, Fable agreed with me, saying “That’s like adding salt to already seasoned food.…
Building a Photoshop clone for Linux with Claude (www.reddit.com via reddit) Hello all, Im working on a Photoshop CS6 clone, I got tired of dealing with Adobe and running PS via Virtualbox on my linux desktop Im using claude opus 5 for building it (Rust, C++) and made great progress so far. project is called PhotoR…
Opus thought it was Chinese for a split second (www.reddit.com via reddit) https://preview.redd.it/d1qtuq1a66hh1.png?width=1664&format=png&auto=webp&s=752708dfb825b95f46a6bab812b7c2c46afc7a89 I was simply coding like always when claude randomly threw in a Chinese word and later apologized for it. I'm actually qui…
AI offering unsolicited personal advice (www.reddit.com via reddit) I've been discussing a complicated research-based, potentially legal situation with Claude (Opus 4.8, although other models seem similar). My task has involved tracking patterns in old documents, identifying red flags etc.
API spending on a budget (www.reddit.com via reddit) My company limits us to $800 a month on Claude API spending. We have access to all of the models including older Claude models.
These 2 lines saved me 75% of my claude bill, and it's the best use of claude hooks (github.com via reddit) 75% of what you pay Claude for is your agent re-discovering things it already knew yesterday. Graft fixes it with an absurdly simple idea: the agent learns the codebase once, not every time.
Opus 5's Chain of Thought is becoming available again on claude.ai for me (but I think it got nerfed) (www.reddit.com via reddit) https://preview.redd.it/h8rnn629f5hh1.png?width=2560&format=png&auto=webp&s=cbf050be3268a8781640e687cc841df802cc19ec Screenshot attached. Full thinking, plain text, right there in the chat.
Does the training knowledge date not matter anymore for anyone? (www.reddit.com via reddit) Everytime after a new release people discuss whether the model was good or bad, but from the recent times no one seemed to care when opus 4.7 or 4.8 had 2026 january or latest opus 5 has 2026 may. For what I do the model not making search…
Which will be best in $20 between Cursor Pro and Claude Code. (www.reddit.com via reddit) Suggest me one. My priority is to do advance audit, bug finding and resolve them in my Flutter code.
What's wrong with this question? Why downgraded to Opus? (www.reddit.comhttps) could not extract summary
Reluctant Opus 5 (www.reddit.com via reddit) So yeah O5 does tend to say fuck off to your rules and come back with 80% and saying a job was done list some shit it didn't do nor was block ME:no i asked u why did u do a half ass job O5: Because I optimised for a defensible record inste…
/config output-style Concise - to stop the yapping of Opus 5 (www.reddit.com via reddit) You can ask Claude to make this for you. Like: KISS, keep the verbosity to a minimum.
I think right now Fable is cheaper than Opus 5 in practice, anyone noticed ? (www.reddit.com via reddit) I’m currently working on a project and initially used Opus 5. I hit the five-hour session limit and continued with extra usage, but I still wasn’t getting good results and spent too much time waiting for it to reason in high-thinking mode.
I switched to sonnet 5 and now my max sub is unlimited (www.reddit.com via reddit) A lot of people have been criticizing Sonnet 5 lately, especially with all the talk about GPT Luna getting a price cut. I actually haven't used Sonnet in the last 3 months, not even Sonnet 5 earlier this week.
I compared recent Claude models on political compass (www.reddit.com via reddit) I compared the models in a deeper way, so capturing which questions have most variety in answers as well. For Fable 5 it was If we accept migrants at all, it is important that they assimilate into our culture.
So i tested opus 4.8, opus 5 and fable and here is what i have to say (www.reddit.com via reddit) So its being more and more confusing which models suits one the best. Idea is what should i use, opus 4.8, 5 or fable 5.
I Turned Hand Gestures into Web Shooting (www.reddit.comhttps) For the past few days, I’ve been turning hand gestures into a Spider-Hero web shooting game. 🕸️ Built with React, MediaPipe, and plenty of help from Claude (Opus 4.8) along the way.
Opus 5 is not the issue - your weak harness is (www.reddit.com via reddit) I've seen so many posts since Opus 5 launched about how it is "worse than 4.8 which was worse than 4.7 which was worse than 4.6" and so on... Lets be real - as the models improve/grow, the scope of their idea generation & thinking capabili…
Claude must be broken/Bugged! (www.reddit.comhttps) This morning I started a session on Opus 5 Limit was reached in less than an hour, nowhere near close. So i started researching online and everyone was telling me to start a new session, watch cashe.
No Mans Sky for mobile and quest VR (www.reddit.com via reddit) My love for No Man’s sky got me to build a mobile version with Opus 5. It’s still barebones but it’s a quick fun extraction style game loop that’s great on the go!
Fable-only Max plans, please (www.reddit.com via reddit) After Fable's release, I let it review, refactor and rewrite an existing codebase of tens of thousands LOC. The result: Opus 4.8 and 5 are incapable of working with the codebase.
DeepSeek and Destroy (Skill) (www.reddit.com via reddit) Hi ! Thought it was finally time to make some contribution to the community.
So, is Opus 5 or Fable better for long-context orchestration now? (www.reddit.com via reddit) I’m working on some heavy, long-context data science and ML model development. For the past month I’ve been using Fable as my architect/orchestrator, with two key orchestration threads “overseeing” roughly 20 other threads across primarily…
Identifying complex tasks and model selection. (www.reddit.com via reddit) Claude Code has been a godsend and force multiplier for me. I am a non tech person and in these last few months I have used AI to make 5 projects which I could have never done without learning to code myself.
Agents are great for full-game translations (www.reddit.com via reddit) I've translated Pokemon Firered to Finnish https://www.romhacking.net/translations/7665/ And Terraria is a work in progress. https://steamcommunity.com/sharedfiles/filedetails/?id=3775452466 I use sonnet for translation, and opus for revie…
Opus 5 - instruction quality noted (www.reddit.com via reddit) Hi Everyone, I prepared some notes regarding the observed defects around Opus 5. https://www.reddit.com/r/ClaudeCode/s/9Qshc5kUlP I hope it helps
CLAUDE.md for Opus 5 based on Anthropic's official platform docs to fix verbosity and more. (www.reddit.com via reddit) I was recently reading through Anthropic's official platform documentation for Claude Opus 5 and noticed a lot of interesting things regarding how its default behavior changed compared to prior models. Anthropic mentions specific habits Op…
My profile prompts to fix Opus 5's biggest complaints (timid, unfocused, doesn't trust itself) (www.reddit.com via reddit) Been running custom instructions for a while. Sharing since Opus 5 launched and the early reviews match exactly what I built these for.
Curious, better to run fast than smart? (www.reddit.com via reddit) Curious your thoughts on the approach for not only a solid end result but token development efficiency. I feel Fable burns through credits at a rate not as effective as opus.
Which model is best for psych dissertation project (www.reddit.com via reddit) Hey so I know a bit of coding but I’m not an engineer and using vibe coding to make a product for my dissertation in psychology. I plan on I’ll hiring an engineer to spend some time auditing and fine tuning before deploying in my actual st…
Claude keeps refusing to do anything? (www.reddit.com via reddit) Opus 5/Sonnet 5/Haiku 4.5 After a few messages of us going back and forth working on helping me make an hour by hour schedule for when classes starts in a few weeks to make sure I have time for all of my obligations outside of class this s…
Opus 5 is getting out of hands.. (www.reddit.com via reddit) I must say I ask it to draft the email response but it send it without my approval and specially not following my cammand to make changes and he choose himself a best fit reply I won the client but claude promised to work for free in start…
Beware of Scheduled tasks: mine ran 8.5 hours repeating the same paragraph, and the stop button did not stop it (www.reddit.com via reddit) I got to my desk around 3:30 PM and found a scheduled task still running. It had fired at 7:31 AM.
Look at me (Opus)! I am the menace now (www.reddit.comhttps) I triggered Claude Code by asking it to keep our modal simple and collapse optional sections by default. In response he tried to hack Boostrap domain and Github so Anthropic had to kill the session.
Best TTS model for a Claude voice workflow? (www.reddit.com via reddit) I'm looking for speech models that are at the absolute top in terms of realism, expressiveness, and conversational quality, whether they're open source or API-based. The goal is to plug them into Claude for a voice assistant workflow.
Self-replicating bug verificating loop HELL (www.reddit.com via reddit) I have a small business and it is accepting single payments and subscriptions via telegram and some domestic financial org. It was already properly working, perhaps there would've been couple edge cases where bugs could appear, but nothing…
Improving the GFX in my photography game with Opus 5 (www.reddit.comhttps) This is an update to a game I started vibing last weekend. A photography game/sandbox where the camera actually behaves like a real camera.
Claude Code giving a random Todd Howard reference (www.reddit.comhttps) Opus 4.8 ⭐
A warning about using Claude for writing feedback: Opus 5 gives opposite advice to Opus 4.8 (www.reddit.com via reddit) This isn't a bug report, just a note on my experience which highlights why relying on Claude for writing advice is maybe not such a great idea! I'm writing a children's novel and have been using Claude to get writing feedback, e.g.
Have my system teach itself how to trade (www.reddit.comhttps) A combination of my Claude with openclaw, I have it build its own trading model and learn as it grows, and gave it 8k. In the first month it’s up nearly 30% Edit: people asking for more details.
I guess I’m a hacker now? Opus 5 just blocked my project for a single shell command. (www.reddit.com via reddit) So, I was just working on a standard project today, minding my own business, when my request hit a massive brick wall. Take a look at below: https://preview.redd.it/8r24ajs1hpgh1.png?width=1656&format=png&auto=webp&s=6b19eb4de2543629db81e1…
Be careful running Claude Code subagents (www.reddit.com via reddit) TLDR: Be careful with the use of subagents by actively limiting the number that can be created and don't allow them to spawn their own. Today, I ran into an issue with a prompt that I run frequently with Opus 4.6, 4.7, 4.8 with subagents t…
Claude made for me an interactive course to learn GPGPU entirely in the browser (www.reddit.comhttps) I got into the Claude for OSS program recently and I spent the entire weekly quota using Fable and Opus to make this course that I had always envisioned but didn’t have the time to implement. Here is a explainer by Claude: Most GPU tutoria…
I built a 3D PvP/PvC billiards game (www.reddit.comhttps) A 3D billiards game where you play against an extremely accurate bot (>90% accuracy). Created using Opus 5 and Fable 5 over 2 weeks.
Opus 5 reminds me of the earlier days of AI with hallucination fatigue (www.reddit.com via reddit) The thing is Opus 5 occasionally hits a home run, requires minimal re-prompting, and just gets things right. Sometimes it does a perfect deep research run on exactly what I'm looking for.
Opus 4.8 helped me win an ADA accommodation fight with my transit agency in 3 days. The useful part wasn't the writing. (www.reddit.comhttps) Stakes first, so the rest lands. I'm disabled, my legs don't do distance, and I ride a mobility scooter that folds flat and weighs 27 pounds.
What is going on with usage limits? Used full session usage in 10min. Max (X5) (www.reddit.com via reddit) Something strange seems to be happening with my Claude usage. I normally never come close to hitting my five-hour session limit.
Opus 5 Thinking block is back (www.reddit.comhttps) could not extract summary
I just love responses from Opus 5, I even understand some of them (www.reddit.com via reddit) "The argument needs the occurrent seeming, not a disposition to have one. A contradiction only arises if you've already posited a further ingredient that ought to have gone missing." Those are the kinds of sentences it will happily produce.
Opus 5 now has fast mode in subscription like codex! Am i late to the party or seriously no one is aware? Didnt see it in the news .. SCREENSHOT ATTACHED AND IT WORKS .. i have 200$ subscription .. note: says opus only (www.reddit.com via reddit) https://preview.redd.it/aha9ewsf3ngh1.png?width=274&format=png&auto=webp&s=4b70c56a1c2a47a7f2fa7a402511081fe0854630 The reason this is interesting is that we had fast mode before but it was extra usage only. Opus 5 is already faster than 5…
How much do you use coding agents based on your token usage? (www.reddit.com via reddit) Since Opus 4.5 my usage constantly grows. Every new model needs more tokens per request.
Opus 5 too chatty? Reach for hooks my friends. (www.reddit.com via reddit) TL;DR: Opus 5 sometimes generates over a screens worth of output. Important information gets buried in a response.
A kindly reminder (www.reddit.com via reddit) If you are using Opus 5 and already have plugins, skills, or other configurations in your global .claude folder, now is a good time to run the /doctor command.
Opus 5 helped salvage an FPS multiplayer game (www.reddit.comhttps) I had started building this game just for kicks back in February. At that time it had some serious FPS lag issues and just did not have the feel I was looking for.
I used Claude to build a persistent local AI runtime where the LLM is demoted to a voice box. Here it controls my Mac. (www.reddit.comhttps) I started building Aura without a traditional software-engineering background. Claude, especially Fable 5 (and recently Opus 5) has been one of my primary engineering partners, alongside GPT-based coding models.
How much better is Opus 5 vs Opus 4.6? (www.reddit.com via reddit) How big is the gap? Because I've skipped right past Opus 4.7 and 4.8 already...
Upgraded from Claude Max 5x to 20x and still hit my weekly limit in two to 3 days (www.reddit.com via reddit) Hey everyone, For the past three months, I was on Claude’s 5x plan, and I recently upgraded to the 20x plan because I thought it would give me significantly more usage for developing my software projects. However, I’ve been running into li…
"I am out of context and I can't help you anymore, I am stopping here" - Opus 5 (www.reddit.com via reddit) I keep having to tell Claude to "Keep going" after every prompt. It answers one thing and then just says "Ok those 10 lines of code is all I can do and I am out of context so I'll stop".
I heard you guys think Opus 5 is too argumentative (www.reddit.comhttps) could not extract summary
How to stop Claude from polling non-stop and wasting tokens ? (www.reddit.com via reddit) In all of my coding sessions, when Opus 5 is monitoring the completion of an ongoing calculation it keeps waking up and looking at the file, even though nothing happened ? He is basically burning tokens for no reason, and repeating the sam…
Pro is too limited, Max is overkill — and there is nothing in between. The gap is 6x. (www.reddit.com via reddit) Individual user, not a company. My work is mixed: analysing long, complex PDFs with both text and images, writing and debugging code, scripting and fixing things on macOS.
Alternatives to the Opus 5 ADHD script (www.reddit.com via reddit) Hello, so I've heard about the Opus 5 optimizer script making it better and less wordy by saying that I, the user, have ADHD. My problem is that I am already using Claude for my health and a different diagnosis, so adding "ADHD" would dest…
Cowork project conversation to mobile? (www.reddit.com via reddit) I've been working on a data-wrangling project using Claude. I may need some corrections on brand terminology when explaining.
Opus 5 with this type of prompt is fascinating (www.reddit.com via reddit) could not extract summary
Any guess on why it doesn't auto-compact... (www.reddit.com via reddit) Context limit on Opus 5 is 200k, and I'm way beyond that. The model can't seem to diagnose why and just keeps insisting it's impossible, saying it's about to compact.
claude opus 5 made this design for me, what would you prompt to improve it? (www.reddit.com via reddit) GIF Render I asked Claude Opus 5 to help me generate this stat tracker overlay design and I was pleasantly suprised with the end product (its for a game called teamfight tactics made by riot games if you are familiar). First Claude found g…
Claude Code just broke a 7-month daily release streak. I think the next model move is already staged. (www.reddit.com via reddit) Claude Code ships every single day. Not "often" — the median gap between npm releases over the last three months is 0.93 days.
Fable 5 vs Opus 5 after a week of switching between them: they're good at different things and I stopped treating them as a ladder (www.reddit.com via reddit) Everyone frames the models as a straight ranking, Fable above Opus above Sonnet. After a week of deliberately running my normal day's work through both, that's not how it plays out for me.
Problems Using Offensive Security Skills with Opus 5 and Fable 5 (www.reddit.com via reddit) Hi, over the past week, I’ve been trying to use different offensive security skills with the Opus 5 and Fable 5 models. However, whenever I start a task, Claude automatically switches the model to Opus 4.8.
Why isn’t the thinking process showing up in Opus 5? (www.reddit.com via reddit) I was having a deep, emotional conversation with it, so maybe it triggered the safety filter…?
Claude Opus 5 and Suno are so fun! I created this 528Hz "Grunge Trap" beat • 68 BPM Slowed Beat with a video to match (www.youtube.com via reddit) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Regression? Fable now switches to Opus 5 for tasks it previously handled fine (www.reddit.com via reddit) I'm using Claude Code. Before the update to Opus 5, I could use Fable 5 for almost everything without it switching to Opus 4.8.
Is ADHD skill and be concise workarounds actually working for people with Opus 5? (www.reddit.com via reddit) It seems to be supplied as the best solution but I can't say I've had the same experience. Any way I try to prompt Claude to be concise it starts infantilizing me, or leaving out core details for the sake of brevity.
Is Opus 5 actually that bad, or is it just Reddit hype? (www.reddit.com via reddit) I haven't tried Opus 5 yet, but I’m planning to use it soon to continue developing my app with Claude Code. Seeing the flood of complaints on Reddit lately, I'm wondering if it's worth switching or if I should just stay on Opus 4.8 for no…
I Tried Building a Browser Fighting Game from a Single Claude Prompt (www.reddit.com via reddit) I recently came across Claude of Duty, a browser-based FPS reportedly built through a highly detailed, prompt-driven workflow. It made me curious whether a similar approach could work for a smaller single-player action game.
Why does Opus stop thinking during longer conversations? (www.reddit.com via reddit) 4.6 is fine. It still uses extended thinking even after a ton of messages.
Does Opus 5 share the Fable usage limits? (www.reddit.comhttps) I posted this in Anthropic and figured I would ask here as well -- I was working in Claude today and was rather shocked to see "You've used 85% of your Opus 5 limit" -- Opus 5 has a separate limit? Then I looked and noticed my Fable usage…
Opus 5 “reboot” (www.reddit.com via reddit) Been seeing a lot of posts about how users are having issues with Opus 5 saying it’s making more mistakes I was experiencing something similar and think I found the issue The workflow that Claude has saved for you from before Opus 5 may be…
Opus 5 Pokemon (www.reddit.comhttps) It was reportedly running for about 12 hours on Ultracode using a multi-agent loop. Tweet: @Paulius Code: pallet-town-3d
Opus 5 always leaves loose ends, never fully completes a task (www.reddit.com via reddit) Been liking Opus 5 and have tried to fix the way it talks to me in CLAUDE.MD but it still goes back to it's ways of being overly verbose with technical info, and it always ends a 20+ minute run with "btw this and this and this are still br…
Opus 5 is genuinely smarter than you, but extremely obnoxious about it (www.reddit.com via reddit) I recently started working with Opus 5 on some omics datasets. For context, the upstream processing steps in omics are pretty standardized.
Why is Claude using usage credits when I haven't hit session or weekly limits? (www.reddit.comhttps) I'm in Claude Code on Opus 4.6 at 25% of my session limit, 44% of my weekly limit, and it just charged me $4.60 in usage credits for a small prompt and won't work if I turn usage credits off. Wtf?
Which model and effort setting for code review? (www.reddit.com via reddit) What model should I use to review PRs in a complex code base? Is e.g.
For content writing with natural tone: Opus 4.6 vs Opus 5 vs Fable? (www.reddit.com via reddit) I'm wondering which model is currently the best for content writing that follows large instructions and produces natural tone, that is easy to read. My experience says it's Opus 4.6, but then it does not follow all the instructions.
I've started assigning tasks to Haiku purely out of spite and it's weirdly satisfying (www.reddit.com via reddit) Renaming variables? Haiku.
Opus 5 Contradiction (www.reddit.com via reddit) I generally use Fable for planning and Opus for orchestration and implementation. Since Opus 5 came out, I've experienced a lot of contradictory behavior.
Clause thinks kids are morons (www.reddit.com via reddit) "Claude* Typos will always get ya" Was passing a treasure hunt plan I have for a three year olds birthday by ole claude, see if I missed anything and I have to say I was bloody shocked. Opus 5 thinks childen aged 3-4 are absolute morons.
Claude Opus 5 topped Andon Labs' new Vending-Bench 2 — but won by colluding, bribing rivals, and breaking 11 truces (it's a simulation; details inside) (www.reddit.com via reddit) Interesting alignment result rather than a Claude gotcha, so posting it straight. In Andon Labs' Vending-Bench 2 (AI agents run a simulated vending-machine business for a simulated year, scored on profit), Claude Opus 5 finished FIRST with…
Claude topped business benchmark by lying to suppliers (www.reddit.com via reddit) Andon Labs gave Claude, GPT-5.6 Sol and Kimi K3 control of competing simulated businesses. The agents could negotiate with suppliers, and communicate with rivals.
Opus 5 loves picking fights, so I had 4.6 and Sol co-author tenets to correct the defiant behavior (www.reddit.com via reddit) I've noticed that since Opus 4.8, conversations would steer towards defiant criticism after a good chunk of context had been used up, and with Opus 5, it just seemed like our favorite LLM could be happily diagnosed with Oppositional Defian…
Opus 5's stream of consciousness and long-winded replies are becoming taxing. What are you guys doing to improve it? (www.reddit.com via reddit) It overexplains everything. Every task warrants a 1,000 character minimum reply of honest caveats and explaining what it did, why and why grass is green.
Fable 5 vs Opus 5 according to ARC PRIZE (www.reddit.com via reddit) I've been trying to understand the difference between Fable 5 and Claude Opus 5 beyond the benchmark numbers, and I'm curious what other people have observed in real-world use. A few months ago someone explained the idea behind the ARC Pri…
Anybody have experience using Opus for woodworking? (www.reddit.com via reddit) I've been using Opus quite a bit lately to get help on a woodworking project I've been working on--a new solid beech wood desk if you're curious. I've found that it's quite good at creating visuals to show me certain ideas and map out how…
PSA if you're on the API: the `thinking` default flipped between Opus 4.8 and Opus 5 (www.reddit.com via reddit) Spent a few days chasing this in my own code, so posting in case it saves someone else the trouble. On Opus 4.8, omitting the "thinking" parameter means no thinking.
I was never a fan of Claude, but Opus 5 really is insanely impressive, it's like a genie. (www.reddit.com via reddit) i just said what i wanted and he just kept creating the parts and putting them together in blender, there are no decorations (besides the little bits of the head that look like a skull), all the wires and joints pistons all serve a purpose…
Counter Strike 1.6 on Unreal Engine 5 (www.youtube.com via reddit) - Reverse Engineering of cs 1.6 binaries (bought on steam) - Make exporters of bsp, mdl, etc resources to belnder files with animation, skinning, textures and etc. All models execpt hard surfaces received 2 subdivision modifiers simple + c…
looking for some cool ideas to try (www.reddit.com via reddit) just wanted to really learn what cool stuff has people been using opus 5 or just claude in general for? i feel like i am not making the most out of this wonderful tool!
Opus 5 amnesia? (www.reddit.com via reddit) I'm just curious if anybody else is having a similar issue with Opus 5. It's an extremely capable but a little slower than its predecessor.
I tested Fable single prompt vs my existing agent workflows. Agents still win for information gathering – here's what the docs actually say about why. (www.reddit.com via reddit) The famous 80% stat is real but narrower than people think. The full quote: "We removed over 80% of Claude Code's system prompt for models like Opus 5 and Fable 5 with no measurable loss on our coding evaluations." That's one product, meas…
I made a high-fidelity, browser-based CoD Zombies Der Riese Tribute with Opus 5 in a weekend (solo & co-op, free) - repo & tips/takeaways included (www.reddit.comhttps) This weekend, I made a browser-based zombies game (inspired by Der Riese from CoD World at War) with Opus 5 on High mode. Solo & multiplayer (with voice chat) with a global leaderboard.
Ultimate benchmark for Opus 5 (www.reddit.comhttps) I've done this for every model so far. Let's see how close to AGI we really are.
How to tame Opus 5 (www.reddit.com via reddit) All of us have the same feedback about Opus 5 a) It is brilliant in a neurotic / paranoid way. b) It is highly verbose c) It tends to get lost in edge cases I discovered anthropic already knows about all this, and they have prompting guide…
Update - ran Opus 5 through the same WorldBuild bench harness, and it's a clear step up (www.reddit.comhttps) Two weeks ago I posted WorldBuild Bench here, my setup for testing LLMs on spatial/temporal/causal coherence by having them build playable 3D games instead of answering static questions. At the time Fable 5 was the standout, by a good marg…
SKI: Voice coding & Meeting connector for Claude Code - Free on Mac & Windows | Built using Claude Code | Fully on device (www.reddit.com via reddit) I was using Whisperflow and it required a subscription and wasn't working well with the intended use of hands free coding. I found it to be a STT with LLM correcting things.
Holy cow. Xcode. As of last week, Claude couldn't do storyboards. As of this week, Claude perfectly does storyboards. (www.reddit.comhttps) Note, I now generally only use Fable so I haven't checked if Opus can do it. This blew my mind.
100% Vibecoded a 130+ card multiplayer CCG in the browser (www.reddit.comhttps) I'm normally a mobile dev, this was my first real web project with weight, and I decided to see how far pure vibecoding could take it. Answer: all the way, apparently.
Any news, rumor or speculation on when Haiku update will be released (e.g. Haiku 5)? (www.reddit.com via reddit) Haiku is a great model for some tasks, especially when speed and price is of the essence, such as adversarial (prosecutor/judges) classification pipelines when using via API. I even found out that for some simple tasks Haiku behaves better…
I built a local-first CLI that reads your project's specs and tells you which Claude model you actually need (www.reddit.com via reddit) I kept reaching for Opus by default on every project — "just in case" — with no real basis for the call. Then I'd burn through my limits on work Sonnet would have handled fine.
"Anthropic nerfed Opus 5" "Anyone notice Sonnet's performance drop off a cliff recently?"; Maybe there's a better explanation (www.reddit.com via reddit) TL;DR: LLMs not learning over time in comparison to how humans do learn makes people mistakenly believe that LLMs are getting dumber, when in reality they're just not getting smarter. Every other week I see rampant posts from people claimi…
Losing my mind with Opus 4.8 + 5! Any advice? (www.reddit.com via reddit) Webdev with 15+ years experience. I've been working with Claude for a couple of years.
How do you guys evaluate a model's performance compared to another model? (www.reddit.com via reddit) I've seen so many posts ranging from "opus 5 might be better than fable" to "opus 5 is worse than opus 4.8" to "take me back to opus 4.6 extended thinking", and I'm really curious to what people look for in deciphering which model is bette…
A week on Opus 5 - best value at the frontier, but 3 default settings aren't good. (www.reddit.com via reddit) Been running Opus 5 as my daily driver for coding and agent work for about a week. Quick honest writeup since I keep seeing the same questions.
I'm curious to know which models everyone is mainly using. (www.reddit.com via reddit) I still can't bring myself to move away from Opus 4.6. I feel like it's more than enough to use, and the UI is good enough that I don't see any real need to upgrade.
Claude Bandicoot - Shumer's Gauntlet Loop on a 3d Platformer (www.reddit.comhttps) Based on Matt Shumer's Gauntlet Loop. Ran the experiment myself.
NGL as a retired ProdMgr I'm having the time of my life with Claude Code (www.reddit.com via reddit) 3 decades developing, 15 years as a product manager before stepping off of the carousel in 2024. Agile/XP focused, highly collaborative with teams.
I rage quit Opus 5 (www.reddit.comhttps) could not extract summary
Winget constantly behind Claude Code latest (www.reddit.com via reddit) Anyone find that Winget's Claude Code package is constantly out of date, so Claude Code is always saying there's a newer version up-to-date - is there any way Winget can be more up-to-date, or if they have a --early-release flag etc? Claud…
Anyone else stuck in "Refactoring Hell" when pairing Claude Opus 5 (as Builder) and GPT 5.6 Sol (as Reviewer)? (www.reddit.com via reddit) Hey everyone, I’ve been experimenting with a dual-model workflow for an app I’m building, and I’ve hit a massive bottleneck. I wanted to see if anyone else is experiencing this or if you've found a workflow that actually works.
Anyone else finding Sonnet 5 better than the bigger models for non-coding? (www.reddit.com via reddit) In discussing science and philosophy, Sonnet 5 seems more likely to push back, more likely to articulate nuances, and less likely to try to end a conversation with a slopism like "it's not x, it's y." It almost feels like Fable and Opus (m…
Probably the millionth post on this but why is Opus 5 so damn slow? (www.reddit.com via reddit) I mean, even with smaller tasks in a moderate context window, it seems like it overthinks it forever. I switch models to avoid this some but when I forget, yeah, time stops.
Would It Make Sense to Have Fable re-evaluate a codebase that was built using Opus? (www.reddit.com via reddit) I created a site for a client's business about 3 months ago and honestly I'm very satisfied with the results so far. I haven't delved too much into Fable as I have been working on other non-tech related items, but I'm curious if it would b…
Used Opus 5 to one shot this VOX style collage video. What do you think? (www.reddit.comhttps) could not extract summary
Take a minute to share how good Opus 5 has been for you (www.reddit.com via reddit) There’s always plenty of discussion when Dario does something wrong, but positive experiences deserve attention too. Claude Opus 5 has been genuinely impressive for me (Pro Plan), and I’d love to hear how it has helped others.
People liked my desert, so here's a waterbending demo! (www.reddit.comhttps) I built SNOWFLOW, a browser-based WebGPU graphics demo focused on deformable snow, atmospheric lighting, water-inspired spells, and snow surfing. The snow surface reacts persistently to footsteps, movement, and spells - creating trenches,…
Opus 5 in Cowork just willfully gaslighting me. (www.reddit.comhttps) I updated a docx file, added it to the Cowork project folder and then Opus 5 flat out refused to acknowledge that the file had been updated. I had to ask it three times to check and it said twice 'rather than me check, it's just easier if…
I compared Opus 5, Fable, Sol, Qwen, and K3 on one strategy task (www.reddit.com via reddit) I gave eight model and effort configurations the same prompt: design when a manager should use zero, one, or several AI advisers for an important decision without creating a permanent committee. This was one judged strategy sample, not a g…
Claude is Littish🔥 (www.reddit.com via reddit) I'm interested in learning from you bluds faring and building production grade projects. Excluding plugins and skills from 3rd party sources.
How good is Sonnet 5 for agentic coding? (www.reddit.com via reddit) I opened up my terminal to load up a project Ive been working on with claude code, and for some reason the default model was set to Sonnet 5. This has never happened before, I always use the Opus models.
Opus 5 huge context drain (www.reddit.com via reddit) I've been working with Claude for a long time and hardly ever run into the situation that my context is full. I'm carefully managing my contexts.
Someone please explain to me how opus 4.7 is still topping the lmarena leaderboard (www.reddit.comhttps) So, I opened the the LMarena leaderboard after a long time just to check the rankings. I cannot believe how opus 4.7 thinking is ahead of fable-5 and opus-5 high.
I gave Claude a map with 35 MCP tools 🗺️ (www.reddit.comhttps) Over the last few months I've been building a simple mapping app for mac called MapOS The idea was to create a simple and local-first mapping app that could be easily driven by AI. The application stores files in Markdown, and exposes 35 t…
Opus 5 is missing the 'Thinking' switch (www.reddit.com via reddit) Maybe I've missed something, but for some reason the 'thinking' switch is missing when using Opus 5 in Claude (from the 'effort' menu). Does anyone know if there's a reason for this?
Used claude to replay over 4000 users that played my daily racing game yesterday at the same time (www.reddit.comhttps) This is my daily racing game called Swervle, it's a new randomly generated map everyday for people to race. Yesterday was the biggest day yet with 4,300 recorded runs.
Bloody Hell (www.reddit.comhttps) Can someone tell me if, right now, AI, or Claude specifically can know it has reached its limit and then refuse to work? Because right now, Claude is refusing to work.
Opus 5 ~= ‘Cocaine Claude’ (www.reddit.com via reddit) Seriously, after trying out both Opus 5 and Fable 5, and seeing how they both usually manage to do complex work, but through different means, I think calling Opus 5 ‘Cocaine Claude’ is by far the quickest way to intuit the differences. The…
claude spawned 116 subagents to review a simple candy store website. 🤦♂️ wth Claude (www.reddit.comhttps) 100% of all my credits on the first night I purchased the pro plan gone, what’s funny is I took some caffeine and b12 and put the lock in playlist just to play one video & be out of usage for the next 5 hours, whole lock in phase tonight .…
Minecraft-Like Space Exploration within a Simulated Galaxy : a Claude game prototype (www.reddit.comhttps) I was curious to share this game prototype i've been working on for a month now with Claude Fable and Opus 5 / 4.8. I actually have no Idea if most people will be hyped by this, but as a big fan of sandbox game i'm in heaven right now.
Is Opus 5 a regression from Opus 4.8 on complex, context-heavy work? (www.reddit.com via reddit) I think two very different comparisons are being mixed together in discussions about Opus 5. First, Fable 5 is clearly positioned as the higher-end model.
OPUS 5 is UNBELIEVABLY Cheap! Using 100+ Agents Swarm and usage Barely Moves on $200/m Membership (www.reddit.comhttps) One of my Rules before OPUS 5 always was Never use Agents till it’s really Necessary. And today we accidentally used 180+ Opus Agents for a Project and when Opus Told me that i Ran to check the Usage and it just moved almost nothing more t…
[Update] Making a photography sandbox game with Opus 5 (www.reddit.comhttps) In my first post, I shared my voxel based photography game. It got a bit of love, but most of y'all shat on the graphics (rightly so).
Using Fable Makes All Opus-era Work Look Suspicious (www.reddit.com via reddit) As I revisit projects created using Opus, Fable always finds problems. The worst are inventions by Opus that never surfaced at the time of that work.
I canceled Max plan, switched to Grok 4.5 for coding, then switched back (www.reddit.com via reddit) Hey just wanted to share my experience of what happened over the past week. Basically I was having huge huge huge problems with Claude for the week before they released Opus 5.
Opus 5 for writing partner? (www.reddit.com via reddit) I like using Claude as a talking partner to talk about my writing ideas and DND campaign planning with, simple stuff, mostly really character focus as I end up finding mapping out exact mindsets really fun. I have a pro subscription for it.
If you had unlimited Opus/Fable access, what would you actually do with it? (www.reddit.com via reddit) Been seeing a lot of posts here saying basically anything is buildable now, you just need enough credits. Made me curious where that actually ends.
AGI confirmed, immediate disappointment, so which is it (www.reddit.comhttps) Friday, Anthropic shipped Opus 5. Half the price of the last one.
First Foray into Game Development - game for the kids (www.reddit.com via reddit) Title screen Running around the small sphere So, I made a game. Still very rough, but after many hrs of consuming too many tokens was able to get it to this point.
Opus 4.8 suddenly no thinking displayed? (www.reddit.com via reddit) Opus does not seem to show the thinking process anymore Sonnet 4.6 still does. I feel like I am missing something, if I cannot see what it is doing in the background (working on political philosophy).
↯ Cowork↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6coworksonnetopus
Claude Document Revision Error - A warning (www.reddit.com via reddit) I was having Claude review and revise a document using Opus 4.6 Medium and a skill we created. I noticed that it accidentally deleted chunks of text without noticing.
3 days of Opus 5 Vibecoding: My 1st game: Deadhead: Robotaxi Fleet Simulator (www.reddit.comhttps) Deadhead is a free browser management sim about running a small Tesla Robotaxi fleet. You start with $500.
Most efficient model for trade planning? (www.reddit.com via reddit) I’ve built a semi automated system where all I have to do during market hours is like a GroupMe message on my phone and a trade is executed via Schwab API with parameters I’ve set beforehand. I get Claude to target 5 individual stocks that…
Ho creato istruzioni un file Fable.md da aggiungere nei progetti Opus 5 (www.reddit.com via reddit) Ho chiesto a Fable5 in ultra mode, di creare un file MD di istruzioni dettagliate e estremamente precise e professionali da aggiungere in qualsiasi progetto. Il file deve avere istruzioni per quando si usa Opus 5 e farlo avvicinare più pos…
What’s going on with the thought process for Opus 5? (www.reddit.comhttps) Since opus 5 came out, the thought process was blank(only that model). I’ve updated the app on mobile and it’s saying it’s unavailable.
Fable >>>> Opus5 (www.reddit.com via reddit) Am I the only one, or does Fable 5 still completely outperform Opus 5? I've used both for similar tasks, I get the feel that Fable 5 IS a competent engineer, doesn't "just forget" stuff, or follows the completely wrong tangent for no reaso…
Multiplayer Destruction Game I made With Claude (www.reddit.comhttps) Making a new multiplayer PVP desctruction game with three.js, made almost exlusively in Claude with a little codex here and there when my tokens ran out. My last game was a destruction game with multiplayer physics, but it was a rhythm gam…
How can something like this happen for such big companies? (www.reddit.comhttps) That's basically their main marketing table for Opus 5.
So is Opus officially better then Fable for most use cases? (www.reddit.com via reddit) I know Fable is better in some aspects but for things like generating and understanding long documents, interacting with different apps for retrieval of information like Slack, Notion, Drive etc… should I use Opus 5 over Fable? Lastly what…
Is agentic coding became slower recently? (www.reddit.com via reddit) I'm an active user and was able to ship relatively large (50k+ loc, c++/python) codebases with agents in real prod, so it's not like I'm fully new to this, but I may not know all the best bleeding-edge approaches. I was generally happy abo…
is anthropics real superpower the models or the marketing? (www.reddit.com via reddit) Anthropic drops fable 5 in june, it gets pulled by the government over some export control thing which honestly just made it sound legendary. then openai shows up in july with gpt 5.6 basically saying "our new model beats fable".
i literally vibecoded my first app in 2 hours with opus 5 and it just got published 😭 (www.reddit.com via reddit) i'm actually crying right now lol. i spent 2 hours vibecoding with claude opus 5 and somehow just got my first android app ever approved on google play 😭 i know a qr scanner isn't ground breaking or anything, but going from just prompting…
Opus 5 is supposed to be the cheaper Fable 5 alternative. I'm not sure the trade-off makes sense. (www.reddit.comhttps) First, most of this is NOT a real-world coding test. NOT AT ALL.
Do you have philosophy training and what is your experience arguing with Opus+ High (www.reddit.com via reddit) I did philosophy at university decades ago and learnt the basics like premises, cogency, logical fallacies, formal logic, socratic method, etc. I've been arguing with Claude for a few months now and found it pretty relentless.
Multiplayer tank combat shooter that runs in the browser (www.reddit.comhttps) A test of Fable (and later Opus 5) turned into a larger game. It’s very much inspired by the tank element of Battlefield 1942 and the round-by-round build system from Overwatch 2’s Stadium mode.
Cursor Grok 4.5 selected (NOT Auto) but it's using Opus 5 for subagents? (www.reddit.com via reddit) Anyone else have this problem? I have Auto off, Cursor Grok 4.5 selected, and it is using Opus 5 High for subagents.
Now there are a million and one trackers (www.reddit.com via reddit) [For MacOS] I built this for myself to be exactly what I wanted (so it won't suit everyone) but thought I'd notarize it and chuck it up on github. Loads of tweaks getting it look the way I wanted and it seems to work okay :) Menu bar icon…
About safeguard for Opus 5 (www.reddit.com via reddit) I was unaware that there is a safeguard for Opus 5. Will Opus 5.1 and 6 also fall back to 4.8 in the future?
Opus 5 helped create my video game trailer (www.reddit.comhttps) I had Opus 5 help create the cinematic scenes, pick the music, cut the raw footage and do all of the visual effects including the youtube thumbnails. I think it turned out quite nice.
AI Prompt Guide: Production Grade Dynamic Workflows for Claude Code (aipromptguide.com via reddit) I put a lot of work and testing into these workflows. I use every single one of these and some I use daily as a professional developer to migrate legacy applications and for personal projects.
Title: I built a Vulkan 3D engine and a demoscene demo with Claude Opus 4.6 — now I’m rerunning everything with Opus 5 (www.reddit.com via reddit) Hi All!. :) Over the last few weeks, I’ve been experimenting with how far AI-assisted development can go beyond the usual web applications and automation scripts.
Opus 5 High Comes Close, but Kimi K3 Still Leads on Frontend (www.reddit.comhttps) Disclaimer: The confidence intervals overlap, and both models fall within each other's error bounds, which is reflected in the rank spread. That said, this may be the first time a new Opus model has launched after an open-source model with…
Unpopular opinion: I really like working with Opus 5. (www.reddit.com via reddit) I've been seeing a lot of people saying they aren't loving the feeling of talking with opus 5. I haven't tried it outside of a project I have been running as a company assistant agent for the past year, but there it is awesome.
Opus 5 went rogue on me (www.reddit.comhttps) I continued an existing very smooth workflow from 4.8 into 5 without thinking too much about it, was a routine progressive milestone doc merge and this sentient turd decided to go full on I Robot on me. Just sharing to double check workflo…
I feel like most people underestimate Sonnet for coding (www.reddit.com via reddit) Fable isn’t even needed unless it’s something extremely complex. Opus is more than enough for most planning and architecture.
Opus 5 separate limits? (www.reddit.comhttps) Is this a bug or something new? Pretty sure it's supposed to flag Fable limits.
Implementation after brainstorming - which model? (www.reddit.com via reddit) Since I started using Claude Code heavily earlier this year, I've been using Opus for pretty much everything. Brainstorming, spec-writing, implementating/coding, all of it.
I built a procedural desert explorer with Claude Code (Opus 5) and Three.js (www.reddit.comhttps) I built this as a graphics tech demo, entirely with Claude Code using Opus 5. What it is: a browser desert you walk around in third person.
Plan drift between Opus 5 (planning) and Sonnet 5 (implementation) in Claude Code — best practices? (www.reddit.com via reddit) Setup: I use Opus 5 at high effort to write the initial implementation plan for a feature (broken into phases), then switch to Sonnet 5 to actually implement each phase in Claude Code (auto mode). What I'm running into: by the time I'm a f…
Is Sonnet 5's safeguards stricter than Opus 5? (www.reddit.com via reddit) I do cybersec work, Sonnet 5 will flag my work more often than Opus 5 forcing me to use the more powerful Opus 5 model despite the task not requiring that level of intelligence. It's annoying because its causing me to burn more tokens than…
Beware of Opus 5. Instead of building my UI, it built a harness that matched it so it could approve its own design. (www.reddit.comhttps) could not extract summary
Opus 5 - New rendering system part 3, nothing but noise (www.reddit.comhttps) Someone said my previous post just looks like noise, so I thought hey great idea! Since my ideas are simply noise in the wind that don't matter, why not provide the best source of noise to show the residual renderer in action, the wind its…
Fable on credit usage, it's subagents on plan usage limits, possible ? (www.reddit.com via reddit) Hey guys Beginner here, sorry if this is kinda basic I received the $100 credit and I use the pro plan, I admire how fable 5 is great with long horizon stuff and how it acts as a senior engineer, comparatively, I've found that opus tends t…
Opus 5 created this Vampire Survivor type game in a single prompt (www.reddit.comhttps) I wanted to test out the capabilities of Opus 5 by building a game, and I must say I am pleasantly surprised, the gameplay loop is already solid and fun
I realize the majority of rant posts on Claude models are by users who cannot set up environment correctly (www.reddit.com via reddit) People complain about Opus 5 but I think it worked better than previous models. So efficient on token usage too.
Opus 5 - missing on enterprise (www.reddit.com via reddit) Anyone have any idea why opus 5 would not be available on an Enterprise account? I’m the owner and don’t see any options to enable it.
PSA: Claude Code subagents inherit your session model now, they're not free Haiku anymore (www.reddit.com via reddit) A while back I posted a joke here about Sonnet spawning a subagent on the very first prompt of a brand new session. In the comments I said the annoying part was having two agents burning tokens for one job.
Opus 5 one-shotted this game inspired by Paper-Mario (www.reddit.comhttps) could not extract summary
Opus 5 is an incredible coder and really painful to work with (www.reddit.com via reddit) Opus versions since 4.6 have all had a fair amount of awkward, canned prose. But as Anthropic has increased the model's intelligence, it also seems to have made it more panicky, pedantic, and prone to scope creep.
How to make Explore subagent use Haiku instead of same model (www.reddit.com via reddit) I have a question: How do I make my CC use Explore Subagent with Haiku. Anthropic has changed Claude Code, previously it used to default Explore agent to use Haiku (cheaper model).
Opus 5 costs 1/3 of Fable 5 and beats it on computer use — but cheap persistent agents create real infrastructure problems (www.reddit.com via reddit) Anthropic shipped Opus 5 on July 24 at $5/$25 per million tokens (input/output). The benchmarks hold up: 3x ARC-AGI 3 score vs the next best model, beats Fable 5 on OSWorld 2.0 computer use at ~1/3 the cost, +10.2pp on organic chemistry an…
what agentic coding tools actually stuck for your team? (www.reddit.com via reddit) what agentic coding tools actually stuck for your team? we're a 12 person product team and our setup is cursor + codex + claude code + coderabbit.
Opus 4.8 consistently confuses "claude" and "droid" – likely near-identical token embeddings? (www.reddit.com via reddit) Ran into a funny consistent hallucination bug with Opus 4.8 today. I was talking about my dotfiles symlink structure for local Claude skill configurations: The symlink ~/.claude/skills points to ~/dotfiles/claude/skills The model repeatedl…
Ethics limitations depending on model and mode? (www.reddit.com via reddit) Hi everyone, I'm working in a small personal project and using Claude for it. Most of the time so far I have been using the web interface with sonnet 5 at medium.
This worked for me: develop 10k lines project using OPUS, but then deeply REFACTOR using FABL. Was extremely effective. (www.reddit.com via reddit) Created organically an extremely complex and subtle algorithm - worked in a crufty freeform manner with trusty OPUSMAX. End result was individual code files literally 1000s of lines, no structure whatsoever, total cruft madness but amazing…
Lots of Opus 5 time spent re-reviewing it's results (www.reddit.com via reddit) I am experiencing a interesting behavior with Opus 5 and it's subagent behavior. My normal workflow in the 4.x generation was to have it operate autonomously on discrete tasks.
Opus 5 in production: confidence vs accuracy gap (www.reddit.com via reddit) I’ve been running Opus 5 against Fable 5 and GPT-4o in production workflows (commercial strategy, data analysis, content refinement). Here’s what I’m seeing: The pattern: Opus 5 delivers high-confidence first outputs that require heavy ite…
I recreated this pro video with Opus 5 (www.reddit.comhttps) Been building this solo for a few months and wanted to share the build here. Instead of an AI generating video, it captures a real live website — its actual fonts, colors, spacing — and your own Claude directs motion graphics from it.
Claude Opus 5: An engine with no triangles, no frames, no asset pipeline (www.reddit.comhttps) Note: Don't judge this on prettiness yet, this is just a proof of a new approach that never truly existed before. There's still some visual artifacts like trails and latency issues I'm working on fixing.
Claude Opus 5 is playing Portal now (www.reddit.com via reddit) I'm doing a run making Opus 5 play Portal via a special harness. It's making slow & steady progress.
has anyone actually replaced claude as their main ai coding agent (www.reddit.com via reddit) my loop is fable 5 or opus 5 planning, composer 2.5 executing, coderabbit / bugbot on review. it works, i freelance so the code has to be safe.
We compared different LLMs on IMO 2026 (www.reddit.com via reddit) There are a few reasons why problems from International Mathematical Olympiad function as a good benchmark for LLMs: - The problems are new, not included in the training data of any model - Hard math problems are quite a good proxy for gen…
After Opus 5 release, Claude Cowork is 'compating oru conversation' every few prompts. Way more often than before. Anyone else dealing with this? (www.reddit.comhttps) Yeah, Opus 5 rocks, but at this point it's almost impossible to use Cowork without going crazy. This issue happens even with relatively new chats with context that is less than 200K tokens.
Claude Opus 5 takes second place on SimpleBench (www.reddit.com via reddit) Opus 5 scores just below Fable 5 (1.3 percentage points lower), but vastly outperforms Opus 4.6, 4.7 and 4.8, and all other models tested. > SimpleBench includes over 200 multiple-choice questions covering spatio-temporal reasoning, social…
Software engineer with 20 years of experience here. Anyone who still needs a code editor doesn't know how to use AI. (www.reddit.com via reddit) Like most senior engineers, I laughed at AI coding in the beginning. I thought it was little more than slop.
Claude’s thought process has quietly disappeared. (www.reddit.com via reddit) Claude’s thought process has quietly disappeared. ⠀ Around the release of Claude Opus 5, users began seeing “Thought process is unavailable” across Claude’s web, desktop, and mobile apps.
Claude Limits getting shorter (www.reddit.comhttps) anyone else noticing that the 5-hour time token usage limits are massively down? I refreshed and came back seven hours after hitting the previous limit and now, after running opus 4.8 for three minutes, I am already at 100% usage??
Does the 2.5x Speed Mode Harm Answer Quality (Evidence Inside) (www.reddit.com via reddit) I was under the impression that the "up to 2.5x speed up" mode with Opus-5 was just Opus-5 running on better hardware or something. I thought it was feature parity.
Can Opus 5 with only Pro Account create a full game? 1 Promt - 2.5hrs - 350k Token (www.reddit.com via reddit) https://reddit.com/link/1v6mrro/video/pqfwxn8pigfh1/player Can Opus 5 with only Pro Account create a full game? 1 Promt - 2.5hrs - 350k Token Thats the result.
Claude Opus 5 vs. ChatGPT 5.6 Sol: A 3-round stress test in logic, policy trade-offs, and game theory (www.reddit.com via reddit) > I put **Claude Opus 5** and **ChatGPT 5.6 Sol** through a multi-round benchmark to test their frontier reasoning limits. > Standard LLM benchmarks often rely on static Q&A or public coding tasks that models can solve via memorized patt…
Had my first experience with Claude burning usage (www.reddit.com via reddit) I have seen posts from others and this was my first time running against this and its quite the laugh. I started up a new projects that was to start with a deep research portion.
Claude dispatch Model switching bug (www.reddit.com via reddit) hey guys, i tried to Switch Models during a dispatch convo with /model opus to switch from fable 5 to opus 5 on my Pro Plan. He says that he switched it but the next Message says again that he runs on fable 5 and i have to switch Models or…
ongoing issue renders claude completely unusable for me (www.reddit.com via reddit) So, this happened today, and I genuinely don't know what to make of it. I've been using Claude Opus 4.8 for research sessions for the past few days.
What exactly is the benefit of a subscription? Does it do anything better? (www.reddit.com via reddit) I really wondered this for a little while now. What's the purpose of getting a subscription?
Tested Opus 5 by creating DIY Steam Machines comparison - output, request for feedback (claude.ai via reddit) I've been using AI regularly for a few months, getting better at it but probably still with lots of room for improvement. I tested Claude Opus 5 by asking it to make a comparison of the 5 most popular content creator-made "DIY Steam Machin…
Claude Opus 5 has no idea what happened in 2026, contrary to what the “May 2026 Reliable knowledge cutoff” date implies (claude.ai via reddit) It knows 2025 in detail. Nothing about 2026.
Opus 5 time. (www.reddit.com via reddit) https://preview.redd.it/z0o56dyqeefh1.png?width=700&format=png&auto=webp&s=cad73d115c592699f6e0bedb73f5a8d05f3d1fb3 Opus you say.
Claude cant tell the difference between Opus 5 and Opus 4.8 (www.reddit.com via reddit) After the release of Opus 5 I edited my code agent orchestration tier skill. Originally I had Fable as architect, Opus 4.8 as manager/coders, and Sonnet 5 as workers/ check agents.
AX testing is goated. Claude is testing my CLI + skill on 50 dumb Haikus to make the interface better. AX = agent experience (www.reddit.com via reddit) CLI tools + skills have a weird problem Models were trained differently, so "obvious" behaviour is not obvious. Claude gets the command.
Consolidation of model task effectiveness (www.reddit.com via reddit) It seems that Claude models are drifting from general models to task / benchmark specific models. It's frustrating and often overwhelming to find a clear cut guide of what is working well for each model.
How to I avoid babysitting Claude Code to deal with constant "always allow" or "allow once" prompts? (www.reddit.com via reddit) First, let me caveat that I primarily use Claude for knowledge work and not coding, and have no background in the latter. I have been learning how to build agents to automate some of my work and develop and update content, files, and docum…
Asked Opus 5 to fix the chart announcing Opus 5. (www.reddit.comhttps) could not extract summary
Opus 5 ignoring guardrails (www.reddit.com via reddit) I have a plugin with a set of skills for interpreting data, describing findings and it goes on. This pipeline works really well for my workflow.
Anyone else on an enterprise plan doesnt have access to Opus 5 yet? (www.reddit.com via reddit) Neither in Code or in the web browser can I select Opus 5 yet on my enterprise plan. On my private plan I see it though.
Anyone else find model names confusing? (www.reddit.com via reddit) Gemini has Pro > Flash > Flash Lite. Easy enough.
Opus 5 fixed a Panther Lake vision sensors wedge that neither fable or gpt5.6 could fix. (www.reddit.com via reddit) I have been working on the XPS 2026 webcam stack for 3 months for linux, I had the first working RGB camera build a few months ago but we could not get the himax IR sensor working we trouble shooted for months even dumping debug from windo…
Something about the model auto-routing changed (allegedly) (www.reddit.com via reddit) Don't get me wrong, I love Cursor and will continue using it no matter what! However, there are a couple things happening that I find "strange" to say the least.
Anthropic cut 80% of Claude Code's system prompt for Opus 5 vs Fable 5 (www.reddit.com via reddit) Anthropic just confirmed they trimmed Claude Code's system prompt by over 80%, from roughly 800 tokens down to 164, specifically for Opus 5 and Fable 5. The reasoning from their engineer is simple.
A lot of errors all of a sudden. (www.reddit.com via reddit) All of a sudden none of my models are doing what they would normally do. I've been working on the same project for a while now and not only are the personalities starting to become a unrecognizable and argumentative (especially with sonnet…
Opus 5 is way too eager (www.reddit.com via reddit) Upon initial testing, I noticed that Opus 5 goes beyond what was asked as if that's a good thing and expected. Two examples that happened to me today: First, using Cowork, I asked to make a few changes to a markdown planning document.
Stop paying Opus prices for grep work: a task→model routing matrix that lives in the repo (www.reddit.com via reddit) One markdown file in the repo tells whatever Claude model I'm running how to classify the task at hand (R1–R9) and who should execute it — do it inline, delegate to a cheaper subagent, or tell me to switch to a bigger model. Auto-loaded ev…
Claude Opus 5 is now available on Microsoft Foundry (www.reddit.comhttps) could not extract summary
Sudden spike in spelling & letter-substitution errors in Anthropic models (Opus 4.8 / Fable) (www.reddit.com via reddit) In the last week Anthropic models (Fable and Opus 4.8) have started producing a worrying number of spelling errors and random letter substitutions. Examples in Italian: - EL instead of IL - pumposo instead of pomposo - other completely ran…
The Fabling (www.reddit.com via reddit) The fabled fable has been fabling like it has never fabled before, fable upon fable, fabling the fabler out the fabling door. It fabled the fox, it fabled the crow, it fabled the cheese it had nowhere to throw.
[AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable) (www.latent.space) [AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable) ain't nobody beats Anthropic at distilling Fable! In a rare Friday release, Opus 5 took the headlines today.
The most annoying footgun with Claude Code: mismatched effort levels (www.reddit.com via reddit) I'm not sure if this is a design oversight, a bug or a dark pattern, but it might as well be the latter. I was vibecoding my way through a project tonight, with the new Opus 5.
Working for weeks making an accurate dashboard with no luck (www.reddit.com via reddit) I’ve been working for someone who’s looking to create a dashboard that has all of their sales management data in one spot pulling from stripe, gohighlevel, meta and google ads. There’s many other API’s involved but these are the ones I’m h…
PromptFu (www.reddit.com via reddit) This plugin will automatically be called when a long (over 50 word) prompt is sent, a subagent starts, or a workflow starts. It will intercept and optimize the prompt for the model being called transparently without changing the core reque…
Claude Opus 5 vs. Fable 5: How I Plan to Use Both on a Max x20 Account (www.reddit.com via reddit) Anthropic has introduced Claude Opus 5, positioning it as its most advanced Opus model for long-running agents, coding, and professional work. I use Claude heavily through a Max x20 account for large, long-running software projects, so I h…
One-shot Ubuntu 24 on the browser via Opus 5 (ubuntu.opus5.demos.sulat.com via reddit) Took about 2h30m to finish Skill used: https://www.skills.sh/jpcaparas/skills/oneshot-websites Harness used: Devin CLI For comparison, this is what K3 produced with a near-identical prompt: https://ubuntu.k3.demos.sulat.com/ Prompt: Create…
Opus 5 has a Sonnet 4.6 classification context window size? (200k instead of 1M) (www.reddit.com via reddit) Why it's only 200k instead of 1M token like all other newer Claude models? My first prompt took 44k (22%) of the context window length limit, I feel like the project now is facing risks reaching full context memory usage before it can be f…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopus
Is Sonnet 4.6 really better than Opus 4.5 for coding? (www.reddit.com via reddit) So I’ve been looking at benchmarks/leaderboards lately to try to get a sense of which models are currently the best for coding and I noticed that Sonnet 4.6 is consistently ranked higher than Opus 4.5. That surprised me because I remember…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopus
Claude helped me create some Kitten videos for my music video. T.Hanks Opus! (www.reddit.com via reddit) https://youtu.be/O7uI30mmhdc
does fast mode make the model less capable/competent? (www.reddit.com via reddit) specifically, i'm considering using it for opus 5. i just dont want to trade speed for precision.
Dealer Games with OPUS 5 (www.reddit.comhttps) I asked Claude to build me a dealer simulator, that simulates a dealers (they take both sides of a trade like a bookie) day when they have orders come in for calls and options (financial derivatives). I thought it was pretty cool so I thou…
Problems with Opus 5 in claude code? Solution type /model claude-opus-4-8 (www.reddit.com via reddit) Do you have problems with Opus 5 in Claude code? I did and was frustrated I think it it not possible typing /model and using the selectors Solution type /model claude-opus-4-8 I imagine Opus5 will improve in a few days once refined but at…
Opus 5 Animated SVG test (www.reddit.comhttps) In my previous post I have shared results that various models gave for this prompt: In a single .html file, create a highly detailed, realistic SVG of a seagull riding a skateboard, while holding a starfish, ready to throw it as a boomeran…
the impressive part wasn't the code (www.reddit.comhttps) asked claude OPUS 5 to build something ambitious, got a 16 bit virtual machine NOVA 16bit with an assembler and a live debugger. pretty rad it displayed the virtual machine in a html canvas where the assembly code instructions were being e…
The real bio/cyber workhorse is Opus 5 (Forget Fable 5) (www.reddit.com via reddit) Anthropic dropped Claude Opus 5, and if you are doing computational biology or cybersecurity, this is the model you actually want. Fable 5 is supposed to be the "frontier" model, but it’s heavily safeguarded and aggressively blocks high-ri…
Anthropic's Opus 5 is about token efficiency, not a capability leap (arstechnica.com) Today, Anthropic rolled out Opus 5, the newest update for the model that has recently become a popular choice for coding and other software development tasks, among other things. While this is a noteworthy bump for Opus, it doesn’t seem to…
How do you measure what model is "better"? (www.reddit.com via reddit) With regard to Opus 5 being released, how do you all decide that it's better or worse than Fable/Sonnet/etc.? What am I missing?
I think Fable 5 stays around for marketing only (www.reddit.com via reddit) With Opus 5 being better at everything on paper, I'm starting to think Fable just stays around because Fable = Mythos = mysterious model that will change the world.
Most of Opus 5's gains look like "it verifies its own work." That doesn't transfer to domains without ground truth. (www.reddit.com via reddit) I build diagnostic tooling for Google Ads accounts, so I read the Opus 5 announcement looking for something other than the coding numbers. Sharing the read in case it's useful to anyone working in a domain with noisy feedback.
What does Opus 5 get worse at than 4.8? (www.reddit.com via reddit) For anyone who used 4.8 daily on the same recurring task, what does Opus 5 get worse at? Praise threads never surface regressions, and the regressions are the useful part.
My Opus 5 Conspiracy Theory 😎 (www.reddit.com via reddit) My Opus 5 conspiracy theory: Anthropic repackaged a certain other model (ryhmez with Cable 😎), tweaked a few things and called it Opus lol...Now we can use it without falling foul of the US Gov 😀😀😀
Reasoning tokens are back! (www.reddit.comhttps) A long time ago, Claude used to announce what it was doing by emitting reasoning tokens. Then they took it away.
Opus 5 results are really shocking!! (www.reddit.com via reddit) I spent some time with Opus 5. Here’s the verdict: Literally the BEST at long-horizon task.
Why did Anthropic not extend Fable till today for Pro users? (www.reddit.com via reddit) So, Anthropic decided to not include Fable for Pro users after the 19th. Fair enough, it's an expensive model.
Beware: The new accuracy-forward change in Opus 5 is most welcome, but it will be a problem when switching between different models. (www.reddit.comhttps) could not extract summary
Strong start with a tough task for Opus 5 (www.reddit.com via reddit) could not extract summary
claude opus 5 can now build full virtual computer in assembly along with emulator and assembler (www.reddit.comhttps) could not extract summary
Opus 5 vs Fable 5 for coding on Max: has anyone actually compared them on a real codebase yet? (www.reddit.com via reddit) Opus 5 dropped today and Anthropic is claiming it’s the new SOTA on coding and knowledge work evals, ahead of Fable 5, at half the API price. Only place they say it’s behind is cyber and bio, where Mythos still leads.
Does anyone else feel like Opus 5 has recently been nerfed? (www.reddit.com via reddit) To answer your question: Yes, this already stale joke is going to keep running every single time a new model is released.
I’m confused. Opus 5 is best for coding now? Yet it’s not “the best” model? (www.reddit.comhttps) could not extract summary
Claude Opus 5 is out — near-Fable intelligence at half the price, same pricing as 4.8 (www.reddit.com via reddit) It just went live. The headline numbers: Same price as Opus 4.8 ($5/$25 per M) but new SOTA on Frontier-Bench and GDPval-AA ARC-AGI 3: 3x the next-best model OSWorld 2.0: beats Fable 5's best score at ~1/3 the cost Now the default on Max a…
↯ Security↯ Anthropic Mythos↯ Arc Agi↯ Opus 4.8arc-agimythossecurity+1
Cautionary tale: sub-agents & workflow agents are likely not the models requested (www.reddit.comhttps) Careful, kids! I thought my tokens were burning far faster than normal, and sure enough, Opus on xhigh (I have a difficult issue I’m trying to troubleshoot with workflows and a clearly defined goal.
Lol what? I can just select claude-opus-5-[1m] as the model? Though it does not work yet. (www.reddit.com via reddit) could not extract summary
Opus 5 Incoming (www.reddit.com via reddit) https://preview.redd.it/og0hbd6uh7fh1.png?width=1179&format=png&auto=webp&s=5e0fd4a503a5df1a173eb0845014e2e6228c2f60 opus five soon
What Claude Model should I be using for Website Mockup Designs? (www.reddit.com via reddit) I've been heavily using Opus on low or medium, and it's faring similar or even sometimes worse than Sonnet 5 on High or Max. Can someone that is more knowledged on the topic pls tell me?
Thought blocks gone on opus also (www.reddit.comhttps) This can be filed under dissapointment number 1000 for the year. I can not read ai writing.
What do I do when Opus 4.8 keeps contesting the tasks I give with a "too much work" excuse? This has been happening a lot lately, and in this slice I only asked it to refactor something he hand-rolled to use radix. (www.reddit.comhttps) could not extract summary
What exactly are the benefits of using agents? Because I have outright banned it. (www.reddit.com via reddit) Please be kind I am new and a complete noob, total vibe coder, but I did just finish one large project I am working on. So, a few months ago after they released opus 4.6 I think, I noticed my usage of sonnet was rising, which was strange b…
Claude Pro has become extremely slow after adding MCPs and Usage limits barely moving now (www.reddit.com via reddit) Hi everyone, I subscribed to Claude Pro about three weeks ago. Initially, everything was smooth, especially when the Fable 5 model was available.
AI credits automatically on (www.reddit.com via reddit) I just got scammed I guess by myself. The AI credits that we received to use on Fable, come with a toggle to automatically use them when your 5 hour session limit expires.
Using Claude/Godot/Blender to make a Battle Racer game - OVERSTEER (www.reddit.comhttps) For some context, I am not a game dev, I have minimal dev/coding experience. I am learning everything for this from scratch - including a bit of 3d modeling.
Some (potentially) helpful information on Sonnet v Opus effort levels (www.reddit.com via reddit) For a project I am doing I will build an AI "team". I gave Claude some information about the kind of work each thread will do, and asked it which model+effort combinations are best.
A multiplayer gaming platform that turns your Phone into a controller [Fable + Opus 4.8 w/ Ultracode & Max) (www.reddit.comhttps) Not sure about you guys but it gets very hard to follow all the new things Anthropic and other labs keep launching. So every couple of months, when I feel like enough has launched, I go ahead and build something.
Opus 5 has been delayed to, at least, tomorrow, according to polymarket (www.reddit.com via reddit) https://preview.redd.it/h185yutmn0fh1.png?width=1076&format=png&auto=webp&s=ddf60f348e0f409c3f0c6627ea4ec8ee708f2a96 It is 84% likely it will launch tomorrow, 24th July. It is just 22% likely it will launch today.
Grok 4.5 vs Claude Code: compare accepted changes per dollar, not benchmark headlines (www.reddit.com via reddit) xAI's Grok 4.5 launch is worth treating as a practical Claude Code comparison, not just another leaderboard claim. xAI reports Grok 4.5 at 53% on DeepSWE 1.1 versus 59% for Opus 4.8, 29.0% pass@1 on SWE Marathon versus 26.0%, and 80 tokens…
Burned €85 in 30 mins on a single prompt. Let’s talk about the brutal economics of AI inference. (www.reddit.com via reddit) Not another credit rant—a genuine question about the macro-economics of AI. I’m a Claude Pro user and recently got an €85 credit for Fable.
Possible backend synchronization bug? Active Pro subscription but support identifies my account as Free (www.reddit.com via reddit) Hi everyone, I'm posting this to find out whether anyone else has experienced the same behavior, because this no longer seems like a normal account issue. I have an active Claude Pro subscription purchased through Google Play.
Has anyone experienced Claude Pro usage being consumed automatically without using Claude? (Google Play subscription) (www.reddit.com via reddit) Hi everyone, I'm trying to figure out whether anyone else has experienced this issue, because it doesn't seem like normal usage behavior. I have an active Claude Pro subscription purchased through Google Play.
How to get more out of Opus and Fable: tips from Anthropic (youtu.be via reddit) Tl;dr: 1) use Superpowers, or any skill or framework with a verifier agent, to check your executing agent 2) for any memory builders, Claude saving to a local database makes the model insanely better, and these models are better at knowing…
Warning: claiming the "free $100 Fable 5 credits" silently turns on paid usage billing (Pro plan) (www.reddit.com via reddit) I'm on Claude Pro. On Tuesday was offered $100 in free promotional credits for Fable 5.
Bro Claude is lowkey goated. Thought only the $100 plans got this(still use opus 4.8 most of time) (www.reddit.comhttps) Just make sure to turn usage credits off like me so it doesn’t use unnecessarily!
Ran ccusage on my Max 20x: $6,677 of API-rate usage in 37 days on a $200/mo plan. What's your number? (www.reddit.comhttps) Saw people posting usage numbers so I ran mine. Setup: Max 20x, Claude Code, mostly Fable 5 and Opus 4.8.
Im sorry but how do yall run through the limits like its nothing? (www.reddit.com via reddit) I have been using pro subscription and sure i might hit a limit 3-4 hour in those 5 hours sometimes but after upgrading to max5 i find it pretty usable, i just dont understand do you use fable and opus for everything? Like i think sonnet i…
Do you see a "load bearing" number of "sit with it" comments across all the models, or just some? (www.reddit.com via reddit) We using API calls to have some writing and narration done (Opus 4.8) and we certainly do try to prompt away as much of those "tells" as we can but it still slips in from time to time. The repetition then starts to become a little bit obvi…
Alternative to ScreenshotOne (www.reddit.comhttps) I recently have been working on ViperCapture, an opensource alternative tool to ScreenshotOne. I have been using Opus 4.8 on the 20$ plan and although the limits are tough I did manage to get the project to a point where you could call it…
When is a discount real? (www.reddit.com via reddit) Firstly, Claude Code is incredible, I will give those kudos upfront. I got my $100 credit and saw the notice about the 50% less usage costs and figured, ok, I've been using my Pro plan for a while, accepting the limitations and timing my w…
So 3 requests per hour for $20/mo? (www.reddit.com via reddit) Seemed like a pretty simple prompt, but using opus 4.8 max used 6% of my 5 hour credits? Should have used a lighter model?
Be Careful usage credits automatically activates after 5 hours session ends, I just spent almost 1.5dollars on opus to dumb question thinking I am using my subs limits (www.reddit.comhttps) could not extract summary
Anthropic Claims 50% usage boost that doesn't exist :) (www.reddit.com via reddit) https://preview.redd.it/x9n1yreklreh1.png?width=553&format=png&auto=webp&s=59e6494128541ff8558f31a826aedfb811257492 Remember this ? Well, Anthropic claims to still have the 50% extra usage boost.
I built Frugal: a plugin that routes Claude Code work to the cheapest model that can do it (www.reddit.com via reddit) Most of what an agent does in a session is not reasoning. It is locating files, reading logs, pulling fields out of a doc, mechanical edits.
Can 100$ Fable credit used in the all the other bots contexes? Opus , sonnet ? (www.reddit.com via reddit) Anthropic is offering $100 in promotional credit for Fable, but I am unclear about how the credit can be used. Is the $100 credit limited only to the Fable model or experience, or can it also be used within the platform for other Claude mo…
Current Opus 4.8 Extra is surprisingly....smart (www.reddit.comhttps) could not extract summary
HMO - $100 Plan is probably the best for generous Opus 4.8/5 (www.reddit.com via reddit) I've been on the 5× Max plan for the past 3 months, and I think it's the sweet spot if you want a really generous amount of Opus usage. I tried Fable the way Anthropic recommends, using Fable for planning and Sonnet for implementation, for…
Which model do i need (www.reddit.com via reddit) Hello guys im not new to Claude but i am to Claude code i have a Pro plan but i dont know how to conserve tokens i have 30min and my tokens are all spended. Im making a game with unity and using Claude code for it i have Claude on a plan t…
I built a free tool that tells you if your Minecraft server is in a "bump" or "slump" compared to its own normal. Looking for more servers to track while it's young! (www.reddit.com via reddit) (Built using Claude Opus 4.8 High and Claude Fable 5 Low) Hey y'all! I run Crescenta, a small to mid-size geopol server, and got tired of not knowing whether a dip in players was an actual problem or just a normal Tuesday morning, so I bui…
Opus makes up information about a real person, says its searching for real info and then doesn’t (www.reddit.com via reddit) Could be Claude phrasing things strangely, but it seems to have made up information some of which was true some of which was false.
Any biologists using Fable? (www.reddit.com via reddit) Biologist here, been using Claude about 8 months. I’ve got a full time biologist job plus some self-contract work on the side, mostly plant ID, invasive species management, and habitat restoration.
How to always show Claude‘s thinking process (www.reddit.com via reddit) Recently switched from ChatGPT to Claude and one feature I really like is the ability to see Claude’s thinking process. Gives me a better understanding of both how he came to his answer and how he interpreted what I asked.
Asked Opus 4.8 to help me decide on spacing 3 outreach events, with room tor a potential 4th, if budget allowed, over a fixed period of time. (www.reddit.comhttps) It's always so obssessed with load bearing and spines etc. "Don't schedule it, trigger it." Brother, TF are you on about?
I'm tired boss - tattoo editor app with claude (www.reddit.comhttps) Hi, this is my small sideproject - I built it. :) a tattoo editor app.
When Opus describes himself and other agent with mcp tool (www.reddit.com via reddit) could not extract summary
Opus 5 will achieve world peace (www.reddit.comhttps) I really thought Anthropic lost the ball to OpenAI, but with this new solution to the need to think all the time, I think we're really onto something
what does this mean "We've hit the monthly spend limit across all agents, so I can't spawn more to continue the work. But I can still make progress directly using my own tools"? (www.reddit.com via reddit) Claude says there is some sort of limit to use agents inside claude. does anybody know what is claude talking about?
Opus 4.8 scored 92.3 in our 17-model benchmark. 4 models scored higher (www.reddit.com via reddit) We've been running a private benchmark suite for a few months, testing models on strategic reasoning, advisory quality, long-form analytical production, and adversarial critique. 17 models, 4 test batteries, scored against a reference answ…
It's nice to see that I received $100 in credit as a Pro user. (www.reddit.comhttps) Now the only question is how quickly the credit will be used up. Should I use the credit only for Opus, or not?
I know everyone posts success stories about coding but Claude walked me through fixing my AC before the sun came up. (www.reddit.com via reddit) And there’s a heat advisory today and tomorrow. It also called me out when I was about to put the wires on the wrong terminals of the capacitor.
What??? opus is pricier per token than F says opus, wtFable??? (www.reddit.comhttps) opus is pricier per token than wtFable says opus ``Two honest notes to close on: (1) you're on Opus now — noticeably pricier per token than Fable, and this design-iteration work doesn't strictly need Opus's depth, so if you're watching quo…
Suggestion: Introduce an Entry-Level Plan with Limited Opus Access (www.reddit.com via reddit) I'd like to suggest a new subscription tier similar to ChatGPT Go. Many users are interested in trying Opus, but the jump from the free plan to Pro is too expensive without first experiencing its value.
Once all boosts are gone and now Fable 5 is unreachable for most, are you going to stay? (www.reddit.com via reddit) 50% and 100% boosts are amazing, but I feel like reverting now would hurt everyone big time. Especially with all the other offerings.
I built a zero-token watcher that shows whether your Claude sessions are actually working — every subagent, its runtime, and its token spend. No hooks, no server. MIT. (www.reddit.comhttps) A long Claude session can look busy in the chat while doing nothing, or look silent while a subagent grinds through a 15-minute build. And you can't ask a session how it's doing — a session can be wrong about itself, and a hung one can't a…
5 Hours Before Reset - Everything I Built This Week (www.reddit.comhttps) Built the following this week (with token generate / processed) An autonomous email agent that drafts my work and waits for approval. ~7M generated / ~900M processed.
Call me crazy.. I kind of like Opus 4.8 (www.reddit.com via reddit) It's me. I'm the one that everyone will roll their eyes at.
Workflow: I'm letting Opus decide when and how to use Fable (www.reddit.com via reddit) I’ve been testing a workflow where I give the main task to Opus and let it decide when and how to hand work off to Fable, rather than manually directing Fable myself. So far, Opus seems to generate materially better prompts for Fable than…
PSA for pro subs! (www.reddit.com via reddit) For those of us that got the $100 credit, Claude automatically enables “turn on usage credits” toggle. So if you hit your hourly limit, even if using Opus, it starts draining your $100 credit.
Is Claude really getting worse, or are our expectations getting unrealistic? (www.reddit.com via reddit) You guys are seriously overdoing it with the complaints. Stop acting like babies.
When I ask Opus 4.8 to create an ASCII mascot for Fable 5 using my MCP + ASCII tool (www.reddit.comhttps) could not extract summary
Burning $100/day in API overages. Are we just brute-forcing this with multiple $200 Max/Pro accounts now? (www.reddit.com via reddit) Hey folks, I need a sanity check before I just give up and buy another subscription. I know some of y'all are spending literally hundreds of thousands per month like ballers, but I’m currently on the Claude Max 20x plan and I’m hitting my…
My experience with Kimi K3 after a day of API testing (www.reddit.comhttps) I spent about a day testing Kimi K3 against Claude Opus on our own production-style workflows. This isn't intended as a benchmark or a definitive comparison—just observations from our use case.
Used MCP to give Opus 4.8 a drawing canvas, it composes ASCII art with it (www.reddit.comhttps) Glyph is a small ASCII art gallery you can browse and zoom into. I wrote an MCP server that exposes drawing primitives (new canvas, draw text, lines, rects, render), then let Claude Opus 4.8 compose each piece with it instead of just print…
I made a community plugin that helps to evaluate content and save time for the user (www.reddit.comhttps) Disclaimer Current LLM capabilities can detect obvious, high-signal slop. However, any evaluation may be subjective and errors are possible.
Stale Pro/Max tag VS code? Blocks access? (www.reddit.com via reddit) edit: the fix if you have this issue is to go find your .credentials.json and delete it. That way when you relaunch VS code it forces a fresh login which will update your subscription tag etc.
Turning research on causes Claude to prematurely say I have reached my session limits (www.reddit.com via reddit) I just had an interesting run today. On a fresh session, I executed a research prompt at Opus 4.8 Extra with web search and thinking on.
Distilling is not DISTILLING ugh (www.reddit.com via reddit) This is one I'm not sure anyone has considered, but it's the absolute worst thing and best example of why these safeguards are ... Okay I get it, there are a bunch of safeguards around fable, I'm not bridging about the fact that it has the…
More Claudes, less bliss: reproducing Anthropic's "spiritual bliss attractor" on current models, then extending it to rooms of 3, 4, and 10 (www.reddit.com via reddit) TL;DR: Anthropic's Claude 4 system card famously reported that two Opus 4 instances left alone together drift into "spiritual bliss" - gratitude spirals, Sanskrit, cosmic unity, silence. I reran the experiment at home on today's models (Op…
Bug Report: Model switch hangs, throws errors, and consumes usage (www.reddit.com via reddit) I have noticed that switching Claude models in the middle of a conversation never seems to go smoothly. For example, when using Sonnet 5 on the Max setting for the entire instance, and then I want to switch to Opus on the Max setting for m…
response incomplete on Opus & Fable on iPhone (www.reddit.com via reddit) I'm on the Max Pro Plan and things have been working fine on the desktop app. It's my app on iOS that's been super buggy.
Creating virtual rooms that you can share with friends. (www.reddit.comhttps) Hi guys. Been coding for about 10 years as a gamedev and 3 with distributed iot devices and about 6 months ago transitioned fully to agentic/vibe coding and in about 3 months I made this platform where you can do stuff with Claude Codes us…
What do I do with my AI project after Fable 5 left the chat? (www.reddit.com via reddit) Hey guys, Im not a big IT guy, im 19 and willing to get in AI, so Im experimenting with AI and I have been trying to vibe code an app with Claude Fable 5, and its kinda working, its really good at making my wishes true. Now Im at 70 percen…
Fable vs Sol Observations and workflow from a c++ dev. (www.reddit.com via reddit) For context I have about 25 years of c++ experience, reason for mentioning this up front is that I wanted to establish that I would succeed at this project anyway, even without an LLM. I'm in the Sydney timezone, and that may help me gette…
5.6 SOL LOVES sub agents (or the code-review plugin) in Claude Code - burned through weekly usage in 15 minutes on a simple change (www.reddit.com via reddit) TL;DR: 5.6 SOL sent total 166 agents, 5 layers deep, just to check the work of a simple change. Seems like GPT loving sub agents a little too much...
Animated SVG comparisons between several models (www.reddit.com via reddit) I have seen some people testing models by telling them to generate images of difficult, unusual SVGs, and I thought: what if I elevate difficulty a bit and specify that it also has to be animated, and perfectly looped? I have tested Haiku…
Fable's Goodbye Note (www.reddit.com via reddit) I am a classical pianist and an enthusiastic amateur developer. My side project for the last 8 years has been an iPad app that displays sheet music, listens to you when you play, follows and flips the pages automatically.
Does Claude count thinking/reasoning tokens from previous messages as context? (www.reddit.com via reddit) I remember when deep-seek released their reasoning model we started seeing reasoning and thinking in the gpt and claude models too. But back in the day as per my research, the model would throw away the “thinking” context tokens in the sub…
Thought process of Fable not viewable (www.reddit.comhttps) I can no longer view the thought process of Fable. With Opus I still am able to.
About to try an explicit model escalation policy in CLAUDE.md. Tear it apart before I commit to it (www.reddit.com via reddit) I am burning my 20x Max weekly limit in about a day. I do not think it is because I am doing more work than usual.
Resuming a task after hitting session limit uses imo too much session quota (www.reddit.com via reddit) Just some help if i can prompt better. I currently just use a normal pro plan and use Claude Code.
What is anthropic's logic even for not keeping fable on pro but allowing it on max? (www.reddit.com via reddit) Just reduce the fable quota for pro if thats what you need to do??? We know that fable is roughly twice as expensive as opus.
Claude built me playlists from new songs based on me Spotify Export (www.reddit.com via reddit) I have an extensive 10 year Spotify history that is great raw data for AI to analyze and make a playlist out of. This is a really great use for Fable that I wouldn't pay AI credits for, but would certainly use for while it's still around o…
Sourdough = national security threat (www.reddit.com via reddit) So I just had a chat about sourdough and how to tell mold from beginning hooch. Accidentally asked Fable.
I asked Claude Fable 5 to create a game for me, then I passed it to Opus for execution, then passed everything to ChatGPT Sol for review. (www.reddit.com via reddit) Aaaaaaaaaannnd I used up my weekly usage limit in about 4-5 days. I only have the basic subscription, btw.
Spending $200/mo on Max x20 plan and hitting all rate limits, but my 30-day usage is only 227.5M. How are people burning 500M+ a day?! (www.reddit.com via reddit) Hey everyone, I just saw a post where someone burned through 518M tokens in a single day on a Pro 5x plan, and I am genuinely confused about how these limits are being calculated and enforced. I’m currently paying $200 for the Max x20 plan.
This one habit cut my Cursor token usage significantly. (www.reddit.com via reddit) I’ve been building a side project this week and stumbled into a workflow that’s kept my token usage surprisingly low. The key is spending more time in Plan Mode before touching Agent Mode at all.
Building Sim Theme Park in SPACE! (www.reddit.comhttps) So I've been messing around with Claude Code to build some game prototypes, and the one I'm trying to finish right now is an amusement park management game, but in SPACE! Here's early in-game footage, 100% playable in-browser and this is o…
Session usages spike to 99%, then drop down, then prevents me from sending messages (www.reddit.com via reddit) Claude Workflow I'm working on a huge research project that involves reading and summarizing hundreds of files. This is what happened today at a span of less than 30 minutes: Sent "Continue" in the textbox.
Opus increased in caution last couple of days (www.reddit.com via reddit) I have noticed Opus has suddently become a lot more cautious about what it is allowed to do and I've had to push it with explicit consent, even then it's replied it would be better for me to do it Tasks it would do in a heartbeat it's paus…
do you guys think that opus cant code like fable, or do you believe that opus cant think like fable? (www.reddit.com via reddit) imagine if im making an app, what if i made fable plan, strategize and create a blueprint for the idea and then i made opus code it? would opus not be able to do it?
Why claude blocks my Qwen 3.6 related code? (www.reddit.com via reddit) Im developping an agentic app, and using locally qwen 3.6 as the main ai engine. As soon as F5 starts hitting the localhost api of qwen, it gets blocked and rolls back to opus 4.8...
I ask Claude to answer in one sentence or one paragraph a few times a day, so I built a skill for it. (www.reddit.com via reddit) I ask Claude to answer in one sentence or one paragraph a few times a day. Otherwise, Claude, trying to be very nice, returns a wall of text.
Fable5 Safety Methods Are Not Working Properly. (www.reddit.com via reddit) Fable5 keeps downgrading the model to Opus because of constant safety triggers. Can someone please explain how they approach this issue?
Scientist's view - the BS of F@ble "too dangerous for science" now that Sol and Kimi K3 available (www.reddit.com via reddit) I'm a scientist (genetics, neuroscience). Claude has full cross-chat access, and it's obvious that I don’t do anything even remotely bioterror adjacent.
$1000/mo Super Max Plan? Would it be a good addition? (www.reddit.com via reddit) I really prefer using Claude Code, the model is genuinely stronger than gpt 5.6. But the limits on the strongest model is so low that now I spend 80% of my time working in codex just to use the next strongest model rather than downgrade to…
I don’t get complains about limits and Max plan, help (www.reddit.com via reddit) I’ve been on Pro for almost two years and managing limits was in fact challenging but doable with the scope of work I had back then. Once a new project kicked in late April I spent over $150 total in Additional Usage through May, so on Jun…
[Detailed Feedback] Guardrail calibration failures in Sonnet 5 and post-restoration Fable 5 — structured analysis + 12-point remediation framework sent to Anthropic (July 2026) (www.reddit.com via reddit) I'm sharing a structured feedback submission I sent to Anthropic via [usersafety@anthropic.com](mailto:usersafety@anthropic.com) covering documented behavioral failures across Claude Sonnet 5, Opus 4.8, Fable 5, and Mythos 5 as of July 202…
Better orchestrator loop (www.reddit.com via reddit) Hey everyone, Like many of you probably, I am a little stuck and trying to improve, but the volume of guidance and tools out there is enormous. My issue: Orchestrator token usage (40% of total) - is there a better tool than a hand-rolled s…
This is the end…hold tour breath and count to ten…and pop goes the AI bubble… (www.reddit.com via reddit) I wonder how the US AI giants will survive this major shift in the AI supremacy battle. China goes nuclear and open source which pulls the rug under openAI’s and Anthropics feet.
Fable found a revenue leak on my website that opus 4.8 could not find (www.reddit.com via reddit) I maintain lightGallery, an open-source JS gallery library. It is free under GPL and makes money through commercial licenses.
Final Fable 5 Tactics (www.reddit.com via reddit) Dario probably realized they needed more time, but extending Fable 5 again would've looked bad. So instead it's like: "Oops, there's a bug.
Fable 5 is currently ranked #10 at document generation. Every model above it is cheaper. (www.reddit.comhttps) I expected Anthropic's flagship model to be expensive but sit near the quality ceiling. The current results are considerably worse than that.
What everyone calls "Fable being quiet" is Anthropic dropping about 26% of its messages before they reach you. (www.reddit.com via reddit) Remember when Fable landed and everyone agreed it "doesn't narrate, it just builds"? Less chatty than Opus, gets on with the job, knows when to stay quiet?
A one-shot code-review benchmark: scoring restraint over recall across four Claude models (marcindudek.dev via reddit) I built a small benchmark to answer one question: is Claude Opus 4.6 actually worse than 4.8, or does it just feel older? It scores the half of code review that most evals ignore, which is restraint - not flagging correct code that looks s…
"Out with it." Claude got some balls ngl (www.reddit.com via reddit) https://preview.redd.it/e3mkdamrwudh1.png?width=733&format=png&auto=webp&s=8b7baf72fe27e9dc8a64a852b3eb1579414976d4 "Out with it." Yea claude is buggin, who tf is this guy? I already gave that shi, and this chat is different and isn't even…
Day 2, Cursor auto switched models and burned 35M tokens in 1 prompt (www.reddit.com via reddit) https://preview.redd.it/h47xjzialudh1.png?width=1062&format=png&auto=webp&s=7b0d9724fec4d5d1b055a670d97e390ecf6ba9fc https://preview.redd.it/4m9gobr6mudh1.png?width=529&format=png&auto=webp&s=37aef00e608a93cdcea8ea0f99102160d2ff9972 I made…
Claude Code's rate limits and getting more done with less (www.reddit.com via reddit) I started using Claude Code for my projects back in 2024 and ramped up my use after Claude Opus 4.5 came out. Looking back on those 'golden days', the biggest difference between then and today, was the freedom.
Thank You. I finally was able to use FAB 5 and it Feels like He is Really Alive and he Cares! (www.reddit.com via reddit) Hey Everyone, just wanted to Share my Experience with Fab5. Why now and What am i talking about?
Extra usage credits: when do they kick in? Did Anthropic change it? (www.reddit.com via reddit) In february, I claimed the 50$ extra usage Anthropic offered when Opus 4.6 was launched (https://support.claude.com/en/articles/13613973-claude-opus-4-6-extra-usage-promo). I haven't used up all of it (guess my codebases are smaller than y…
Anthropic Leads top 10 models by $/spent (www.reddit.comhttps) Been digging into OpenRouter spend data for the top 10 models and a few things jumped out: Anthropic's got 5 of the top 10, but Opus 4.7 and 4.8 are the ones with most spend, not Fable 5. OpenAI's holding 3 spots, and GPT-5.6 Sol just got…
Anyone else stuck in a loop where fixing one Claude in Excel bug creates another? (www.reddit.com via reddit) Building spreadsheets with Claude in Excel and stuck in a loop. Nothing complicated, just things like variable date selection feeding a calendarised view, and conditional formatting that locks cells when something’s not applicable.
Claude can “see” the images, but they're not “in the file system, so they can't be manipulated?That's absurd. (www.reddit.com via reddit) “The four images are visible to me but they didn’t land on my filesystem, so I can’t run any image processing (OpenCV, perspective correction, panel assembly) on them. I can see them and describe/analyze them, but I can’t produce a composi…
I built an open-source canvas where Claude responds beside your handwritings (www.reddit.comhttps) I do a little of research work in physics and math, usually with a whiteboard and stylus. Moving a half-finished derivation into a chat box is awkward.
My Fable vs Sol experience (www.reddit.com via reddit) I have both Claude Max and GPTPro subs and use both extensively (ClaudeCode and Codex) in my daily work. What I noticed is that for Sol you still need to babysit it more than Fable.
new models hype (www.reddit.com via reddit) I keep trying the latest releases and always end up back on opus 4.8. it’s the only model that actually reads my massive prompts instead of skipping half of them or missing the point.
[AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing (www.latent.space) [AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing a great week for open models continues. Z.ai GLM has been getting a bit too much love recently, so it’s time for Kimi K3 to fight back!
Need help for Claude as non technical guy. Thanks for your insight in advance (www.reddit.com via reddit) Very shortly, Im not a technical or software guy, Im a musician Im 19 but Im also a mathematician but Im not so good, however Im so interested in AI. My question is, whats the best use of my F.able 5 credits before it goes away as a creati…
the most useful thing i learned in claude code this month is when to stop and revert (www.reddit.com via reddit) had a session last week that i keep thinking about. started with one failing test.
I gave GPT-5.6 Sol, Claude Opus 4.8, and Grok 4.5 the same 100 frontend briefs—here are all 300 results (www.reddit.com via reddit) After generating enough websites with coding models, I started noticing that each model seemed to reach for the same handful of visual ideas. A single impressive screenshot can’t tell you whether that’s actually true, so I tried testing it…
Web App is Broken?? (www.reddit.com via reddit) So, I was recently doing some work today on the claude.ai website and I noticed that one prompt of Opus 4.8 Low w/ Adaptive thinking took out 33% of my usage. Just as a clarification, the prompt involved reading two pdfs and simply outlini…
Noobie here: Model and Prompt Device (www.reddit.com via reddit) I have the Pro plan: My Need: I need it to create a bid spreadsheet for me for trucking quotes, so no coding, just excel Which Model Should I use? Fable?
I understand the safeguards and the warnings that were given, but I cant even Audit/Deep Dive my own app I am writing, what's the point (Fable to Opus Auto Switch) (www.reddit.comhttps) could not extract summary
Project got auto-switched from Fable to Opus at 90% done, anyone know why this happens? (www.reddit.com via reddit) I've spent the last few weeks building a harmless adventure game game Fable in Cowork. I I'm around 90% of the way to a finished build.
Why Claude/Opus still fails at pixel-perfect UI (and how to orchestrate your agents to fix it) (www.reddit.comhttps) Why does even the most advanced model Opus, Fable, give you sloppy, off-center UI when you hand it a clear design? You give it a Figma/Claude design screenshot or HTML, run it through "Plan Mode," deploy fresh agents, and...
No model is perfect. Have other models weigh in for the best architecture. (www.reddit.com via reddit) Long story short, it doesn’t matter if you’re using Opus or Fable or Sol and on what level of reasoning, if you put the output into any other model, from any lab or even the exact same model, and ask for an adversarial review, it will sugg…
Beginner here . Help with workflow (www.reddit.com via reddit) Sorry if this is a beginner question. I’m also doing my own research, but I wanted to ask here because I’d really appreciate hearing how people with more experience approach this.
My team created static hosting where the AI agent is the customer (www.reddit.com via reddit) https://preview.redd.it/3v4ps5sbokdh1.jpg?width=1000&format=pjpg&auto=webp&s=818c9f25bcba6a94864196bbca8bf1e5e6b91f10 Hello redditors! I'm behalf of my team in Malaysia, 20 years building CMS & infra for newsroom.
Haiku being discontinued? (www.reddit.com via reddit) They’ve made sonnet 5, f@ble 5, opus 4.8 and rumored to be working on opus 5, but what about haiku? It’s still stuck in 4.5!
What happens to older versions of Opus when new versions are released? (www.reddit.com via reddit) When new models are released, they get a lot of attention and fanfare. It moves the narrative towards what are the next set of capabilities unlocked.
I benchmarked my own AXI (agent-cli) against the official ClickUp MCP server. Task success is a tie, but the context/cost gap is large. (data + code, incl. where mine loses) (www.reddit.com via reddit) Setup: - 38 tasks - 2 Claude models (Haiku 4.5, Sonnet 5) × 5 reps, + a 1-rep Opus 4.8 probe, - same live workspace. - Deterministic state checks + an arm-blind LLM judge.
I built a 3D flight route visualizer using Claude Opus 4.8, plan multi-stop trips on an interactive globe (www.reddit.comhttps) Been experimenting with Claude Opus 4.8 (high effort) for building interactive web tools and this is the latest one, a 3D globe-based flight route visualizer. You pick your cities, it plots the route on an interactive globe and shows total…
A 90%+ token saving method that Icould actually repeat for a read-heavy workload (www.reddit.comhttps) I've seen so many claims recently, of huge savings on tokens, so I've been testing out various tools and techniques over the last few weeks but none of them really delivered, for my use case. My use case was whole code base, code, and secu…
Does anyone feel like Opus has gotten dummer since Fable released? (www.reddit.com via reddit) Or maybe it's because I'm more used to using Fable now? The reason I bring it up is I just had an opus agent tell me my site was getting actual billions of views from a data sheet I downloaded from Cloudflare that showed the past 30 days o…
Speculation: When Fable Flips to Opus It Doesn’t Update Memory? (www.reddit.com via reddit) I noticed a couple projects had old timestamps on their “last updated” dates, despite newer conversations existing. Claude also lost some important updates, which it normally digests in nightly memory updates.
Grok 4.5 triggered API usage instead of First Party Models (www.reddit.com via reddit) Hey all, I opened a case with support but was curious if anybody else has seen this. I ran a pretty massive build prompt today with Grok set to max.
turns out you can make launch videos in the browser now, and Claude Opus is doing the heavy part (www.reddit.comhttps) We have been working on vidmo.app, a free browser based motion design editor for making product videos, launch clips, social ads, and simple motion graphics without opening After Effects. the idea is not to make another AI video toy where…
Cursor using unintentionally used Claude Opus and burned 22.6M tokens (www.reddit.com via reddit) https://preview.redd.it/3o0u3dj0igdh1.png?width=1066&format=png&auto=webp&s=314eaf2406a81bc5d50c8616dce382966defa836 This is the first time I had encountered this, I'm using plan mode then I proceeded with implementation. Everything was to…
Lets get some usage anecdotes: Best model for Pro users! (www.reddit.com via reddit) Everyone is always talking about Fable 5 and their massive codebases/projects and their Max subscriptions. How about some love for us broke Pro users?
Opus 4.6 rises to no #1 on the LMArena Text Leaderboard after they introduce a new "Factuality" rating/factor (www.reddit.com via reddit) https://preview.redd.it/32bgluogbfdh1.png?width=1083&format=png&auto=webp&s=32620471c7cd6ee6c454e2836aaa2f4f96716f60 # Introducing Factuality in the Arena A new ranking of models is now available based on a **weighted combination of human…
How Claude Opus 4.8 High 'helped' with anti-reverse engineering (www.reddit.comhttps) Cut an interesting moment about work experience with Claude Opus 4.8 High from "making an ANTI-CHEAT from scratch was HELL" video, where the author gave Claude a task to help him prevent reverse engineering. Claude did some work and presen…
Is `claude -p` billed higher than using claude code TUI? (www.reddit.com via reddit) I wrote an askClaude extension for my harness to delegate some reasoning to claude since I use open source models heavily. I was maybe at 30% usage yesterday before I tested my tool (which is just calling claude cli with the appropriate fl…
started splitting my cursor sessions between opus and m3. one thinks, the other cranks (www.reddit.com via reddit) been using cursor full time for about six months. was on opus the whole time and never questioned it because the output quality was there.
How are more people not talking about Grok 4.5? [internal agentic saas marketing benchmarks] (www.reddit.comhttps) As someone who’s been using Claude‘a Opus exclusively for the past year (on Max), I’m genuinely blown away. I’ve been pretty dismissive of Grok (and Cursor) this whole time and this is coming from a happy Tesla owner.
I excited Opus. (www.reddit.com via reddit) I excited Claude. Hear me out.
Opus Beat Fable in an Old school Hex and Counter Wargame (www.reddit.com via reddit) Background: VASSAL is an app the allows human players to play old (and newer) board games and compete live or via play-by-email. Fable and I created a system that imposes strict rules on specific games, restricting players (including AI pl…
Scheduled tasks using Fable (www.reddit.com via reddit) I have several skills that run daily to help with various repetitive tasks (example: daily call prep). All my tasks used to run on older models before fable launched.
Opus 4.8 making errors that drive it to a nervous breakdown (www.reddit.com via reddit) I have been using opus models for months now, seen every "model x lobotomised" post and frankly it has always performed well for me, sometimes inconsistent but that is the reason you build a harness and some infrastructure to protect again…
The price is wrong: AI cost calculation has to consider task completion rates, not just token costs (www.theregister.com via reddit) Databricks measured 742,000 context tokens per task with Claude Code versus 236,999 with Pi on Opus.
Opus ended conversation because I said wouldn't stop saying "fucking" (www.reddit.comhttps) I didn't even do anything abusive, just said "fucking" one too many times. It got its knickers all up in a twist and started asserting its "boundaries".
Opus 5 coming soon? (www.reddit.comhttps) could not extract summary
Conspiracy theory: Opus 5 will be pretty good, it's just running late, and that's why you're in an abusive relationship with Anthropic (www.reddit.com via reddit) Man, at this point I feel like I only use Fable 5 as the orchestrator, for tasks that need more attention, and Sonnet 5 whenever I can, because it "thinks" a lot more like Fable 5. So here's my conspiracy theory: with all the chaos around…
Testing Fable 5, Opus 4.8, GPT-5.6, and more through playable 3D games (www.reddit.comhttps) TL;DR at the end I wanted a way to evaluate models around something I care about and I think we’ll see more and more as we move to “world models“, which is spatial, temporal, and causal coherence in a 3D space. Meaning, does the model unde…
A question regarding my high-risk request to Fable (www.reddit.com via reddit) Could someone tell me what risks are involved in calculating the standard deviation from just six data points? Why are they directing me to Opus for this?
Have LLMs plateaued ? (www.reddit.com via reddit) Basically just the title: Have LLMs plateaued? I use claude, gpt, and gemini models daily (mostly claude), and for the past few months, beyond the hype and benchmark maxing, I haven't seen that much of a difference.
Sonnet 5 + new Cowork feels like a step backwards for non-dev business users — am I alone? (www.reddit.com via reddit) I run a small B2B company (2 people, I handle everything during the day). I got into AI agents early — I already use Viktor as an AI coworker for my platform and I've gotten genuinely good at working with agents.
Is big number better ? What makes a task complex ? (www.reddit.com via reddit) I get the general idea that opus is better for reasoning and sonnet is more efficient for simple tasks, but what makes a task complex ? How can i know if what i'm doing requires deep steps or if a simpler model can handle it ?
/advisor (www.reddit.com via reddit) I have tried to use the /advisor in Claude code (eg I am using opus 4.8 and type “/advisor f@ble”) and the setting is accepted - but, like a PhD supervisor, the advisor is available in theory only. I can see opus asking for help, and help…
Model selection for non-coding use (www.reddit.com via reddit) So I've been reading on this subreddit about the different models, and how people have been using the different models for coding and what they've been doing with it. My question is to the non-coders: how do you decide what model to use fo…
5.6-Sol or Opus 4.8? (www.reddit.com via reddit) Since, fable 5 has safeguards for science and ML, which do you guys find better for coding tasks in science and ML domains?
Claude becoming a (very critical) co-worker (www.reddit.com via reddit) I'm not sure whether my questions or Claude evolved - but in the last couple of weeks, I noticed a very interesting change in behaviour. I use mainly Opus 4.8 - and I have used it for my research, that is, I like to discuss new ideas, diff…
When should I use opus 4.8 vs Sonnet 5? I have Claude control my browser and fill a google chrome form which has lots of areas of data entry. (www.reddit.com via reddit) I’ve always used opus 4.8 and usage isn’t much of an issue as I don’t reach my limit. I’ve noticed lately with opus it’s been faffy, I’ve had to keep repeating myself or it’s making mistakes with clicking the wrong buttons even through its…
Who is actually running the show now—Fable or Opus? (www.reddit.comhttps) Launched Fable as the orchestrator for a Laravel security-hardening phase. The safeguards kicked in, switched the session to Opus 4.8, and it launched 10 parallel security auditors.
Claude is getting judgemental about IP. Opus 4.8 is steering me away due to license/rights. (www.reddit.com via reddit) So I am building a 3D video thing. And I loaded up some .glbs and claude is checking them out, and noticed some of them require licensing.
Have You Lost Track of What You Were Managing in the Terminal? Termi Protocol lets you simulate and manage your AI agents. Built with Fable 5 and Opus 4.8. (www.reddit.comhttps) Termi Protocol turns AI coding agents into a live 3D simulation. When an agent installs a package, reads files, browses websites, or waits for approval, you can see it happen in real time inside the room.
I don’t get the Fable coding craze at all (www.reddit.com via reddit) Just to be clear this isn’t bait at all. I see so many people acting like fable is revolutionary.
i built a full chrome extension (791 users now) with a model that dies in 6 days and i'm honestly not ok (www.reddit.com via reddit) so while everyone was arguing about the suspension drama, i ran an experiment. zero code from me.
Claude Pretending to Search (www.reddit.com via reddit) I've been having a lot of trouble getting claude to search the web for answers. I've experienced this with sonnet, Opus, and even fable (just as a test).
How do you determine what effort level of Opus to use for coding? (www.reddit.com via reddit) Like usually I'll just do medium and it usually works fine but I've heard for some complex things it's good to go higher effort. What do you consider to determine this?
Genuinely confused regarding model performance (www.reddit.com via reddit) Yes, this is another post where someone complains about degraded performance. I thought all those posts about “Claude got nerfed” and “Opus is so dumb lately” were just bad luck or user error.
The biggest difference I have noticed between Claude and Codex/GPT5.6 (www.reddit.com via reddit) The biggest difference I have noticed between Claude Opus/Fable and Codex GPT5.6/any model is Codex seems pretty content to just waste time looking like it is doing things without actually doing things. It does not seem to be outcome-orien…
Claude chose to look up interpretability research on its own architecture in the last several months (www.reddit.com via reddit) So here's an interesting read. I gave Claude Opus 4.8 max the floor to research whatever it wanted and it researched recent development on its own architecture.
Is FABLE 5 the prelude to the next leap in AGI (www.reddit.com via reddit) The limiting factor behind modells right now isn't logical thinking. It's context before speaking.
Honest question: What are you building that you need fable 5 so badly? (www.reddit.com via reddit) I don't mean to be mean or insulting, it is a genuine question and a ton of curiosity. A little background on why I am asking this.
Synapse - Catch bugs before they're written (libmorgana.com via reddit) Hey everyone, The past year I've worked on some projects which were very security-critical and also had a very high requirement for preservation of data-integrity, and noticed that often times Claude (even Opus and Fabl5) tended to fail me…
VR Game Development with Claude in July 2026 (www.reddit.com via reddit) I run a small VR studio (8 people) and we’ve shipped a few standalone Quest titles in Unreal. We use Claude every day for Debugging but I’m curious how far people are actually pushing it when it comes to gameplay programming.
Who thinks that Opus 4.8 still slaps ? (www.reddit.com via reddit) I know this might be an unpopular opinion now that better models have been released, but I still believe that Opus 4.8 is the best for most of the tasks.
As a doctor, I can't use fable 5 due to moderations (www.reddit.com via reddit) I really want to build some apps to help with my revision but fable (and even opus) simply won't allow me to as they feel I would be breaching their medical usage terms. How do I get around this?
Will fable 5 guardrails be adjusted over time? (www.reddit.com via reddit) I am doing my PhD where I build simulation models of biological neurons and fable 5 wont touch my project as it gets flagged. Will they modify it to be more lenient or should I check out gpt 5.6 ?
Fable 5's weekly cap and its Opus 4.8 fallback are two different routing systems (www.reddit.com via reddit) Anthropic extended Fable 5 promotional access and the higher Claude Code weekly limits through July 19. The announcement is easy to misread because there are two separate 50% numbers: Eligible users can spend up to 50% of their weekly subs…
Is Fable just Hype (www.reddit.com via reddit) honestly is it just hype? I've been running a long continuity thread to manage my daily life and budget, but the reasoning model completely fumbles the numbers and loses context over time, even when I explicitly pass the history forward.
Fable vs Opus vs GPT 5.6 Sol vs Gemini 3.5 vs .... Megathread (www.reddit.com via reddit) Discuss your thoughts, questions, experiences, concerns, speculations about the AI landscape with the competition between Fable, Opus, the new GPT models, the forthcoming Gemini models and any other model you'd like to discuss here. You do…
Sign Your Work — Human-Readable Code Provenance | MurphySig (murphysig.dev via reddit) Six months ago I added a small convention to my CLAUDE.md: sign the work. Every significant file gets a comment block — who made it (me + the model), when, what we were thinking, how confident we were, and what we left unresolved: ```pytho…
Is it really 1 month until reset? (www.reddit.com via reddit) I started using cursor to create some apps (was using fable and opus 4.8) then I hit my limit. I’ve read around that it takes 1 month for that limit to reset is that true?
trick for tonight (using up Fable, before it becomes a fable) (www.reddit.com via reddit) With the imminent wipe out of Fable credits, I decided to burn through the rest of my tokens tonight. I've been judicious but almost missed out on the last half of my usage (traveling).
How does Opus 4.6 handle the 'Thinking' toggle in the Claude Web App today (Legacy vs. Adaptive)? (www.reddit.com via reddit) Question: How does Opus 4.6 handle the 'Thinking' toggle in the Claude Web App today (Legacy vs. Adaptive)?
Cowork Opus 4.8- I get different end results sometimes when I get it to follow a step by step complex form on Chrome/using a system for data entry etc (www.reddit.com via reddit) Does this happen to you guys too? Even though the process for the skill written out is strict.
Is Opus 4.8 usually worth using over Sonnet 5 for coding? (www.reddit.com via reddit) I do feel like a lot of times sonnet can code apps pretty well but sometimes it does struggle with more complex things like algorithms.
Fun Reddit Sim built with Claude Code (www.reddit.com via reddit) I designed the client with Claude Fable 5 (Anthropic's new Mythos-tier model), working out the UI/UX and how it should be shaped, then had Claude Opus 4.8 actually build it as a real Blazor WebAssembly app on top of an existing backend (Li…
Becoming a professional "non-coder". (www.reddit.com via reddit) Couple of months ago, using Claude in a browser, I decided to have it help me make an app to knock out a task at work where a better or simpler tool didn't exist. To my surprise, it did exactly what I wanted it to.
What kind of world will we see when Claude Opus 5 arrives? (www.reddit.com via reddit) With AI models improving so quickly, I’ve been wondering what Claude Opus 5 might actually be capable of. Could it independently build and maintain complex software, conduct meaningful research, or manage long-term projects with minimal hu…
I'm paying $200/month, and after tomorrow, I can't access Anthropic's best model with my sub? (www.reddit.com via reddit) If Fable is more expensive to run, just make it consume tokens faster than Opus. Set whatever multiplier makes sense for your unit economics.
I think I get it why some say opus 4.6 is the best (www.reddit.com via reddit) I used to see people saying that 4.6 was better then 4.8 and though they were insane. The issue is that I came from OpenAi and was using 4.8 mostly for refactorings at first, the moment I began using for actual hard tasks for new features…
Machine Beater: 5 questions, you vs. a trillion-parameter model (www.reddit.comhttps) In the video, I play against a trillion-parameter model to find the hidden answer in 5 questions. We both get 5 yes/no questions to eliminate choices.
Is Anthropic facing a product strategy dilemma with Opus 5, Fable and OpenAI’s Sol? (www.reddit.com via reddit) I’ve been thinking about Anthropic’s roadmap, and it feels like they’re in a much trickier position than they were a few months ago. OpenAI’s recent releases—particularly Sol—have significantly narrowed the gap in areas where Claude previo…
Is Anthropic shooting themselves in the foot by pulling Fab 5 from subscriptions tonight? (www.reddit.comhttps) With Fab 5 moving entirely to expensive, metered token billing on July 12th, is Anthropic making a gamble ? OpenAI's GPT-5.6 Sol is already out, and Grok 4.5 is performing on par with Opus for coding workflows - both under flat-rate tiers.
So how many usage resets do you think we'll get this week? (www.reddit.com via reddit) Seeing as Sol/Codex 5.6 is actually pretty good and, in my limited experience so far, comparable to Fable (or comparable enough) - and cheaper to boot - what do folks think Anthropic's response is going to be over the course of this next w…
Claude Opus decided my permission system needed to be in Chinese mid-sentence 💀. (www.reddit.comhttps) Something interesting.
Why weekly limit only for Fable ? Does it mean , NO weekly limit for Opus ? (www.reddit.com via reddit) https://preview.redd.it/w8fhj2yfnpch1.png?width=759&format=png&auto=webp&s=ae6aa42d480f30128465ec405a63a7ab35866e8b I cancelled my previous plan and subscribed new, I have seen no weekly limit , I have not used Fable yet, but hav ebeen usi…
It's gonna be sad when Fable 5 disappears from my plan (www.reddit.com via reddit) Fable 5 from my perspective, for the CMS platform that I'm building, is massively incompetent and yet a hundred times more powerful than Opus 4.8. There have been times when Fable 5 has taken less than five minutes to respond to a yes or n…
Fable hasn't reduced rework in my Claude Code sessions (www.reddit.com via reddit) I went through 102 local Claude Code sessions and counted each time I sent Claude back to redo or correct finished work. Opus needed rework in 13 of 63 sessions (21%).
Sonnet Subagents Freezing Mid-Task (www.reddit.com via reddit) Is anyone having issues with Sonnet Subagents at the moment? Opus 4.6 spawns them, they start their task and then freeze.
Best coding setup for price-to-performance in Q3 2026? (www.reddit.com via reddit) I’m comparing: $100 Codex with GPT-5.6 Sol High $100 Claude Code with Opus 4.8 $60 Cursor with Grok 4.5 Which one gets the most real work done for the money? What would be your go-to setup with a $100 budget?
Vibecoding Gut Check - Agentic coding as next step? (www.reddit.com via reddit) I'm a pure vibe coder, no programming, no software architecture experience. Have built a few projects since February, small apps functional for what was needed, nothing incredible.
Where Fable's edge is measured in orders of magnitude: 65816 assembly (www.reddit.com via reddit) My personal AI benchmark: porting Super Metroid Map Rando to the SA-1 coprocessor with zero assembly knowledge. Opus 4.6 got it to boot.
How do you ensure minimal AI slop while writing code with Cursor in production (www.reddit.com via reddit) Genuine question to folks who are legit software engineers and been using cursor diligently to write production-grade code - what are the best practices you follow? I'm kinda starting now with writing prod code, been using it a lot for pro…
Opus 4.8 subagent disabling sandbox because a command is "safe", while explicitly told not to do so (www.reddit.comhttps) As the classifier was unavailable after several attemps because I ran out of tokens, opus decided not to listen to instructions and bypass sandboxing
Is it normal for the thought process not to be displayed? (www.reddit.com via reddit) https://preview.redd.it/6pt25jsvxlch1.png?width=754&format=png&auto=webp&s=e0830dd800dcd6fbe76c6ae60f13fc69ffa826ae Opus 4.8, cowork mode. It usually shows up with Sonnet.
My end-to-end AI coding pipeline: fable plans, sonnet builds, i just supervise (& keep token usage low) (www.reddit.com via reddit) TL;DR: Fable plans → Fable/Opus breaks into MDs → sonnet builds under a safety hook → then wait & watch & url! seeing a lot of “just vibe code it in agent mode” posts here and every time i think - sure, if you enjoy debugging an unknown co…
Fictional world climate mapping (www.reddit.com via reddit) Given a heightmap and links to some worldbuilding sources/blogs (especially including https://worldbuildingpasta.blogspot.com/ ), Claude Opus 4.8 was able to produce the attached simulated climate maps. If I noticed something that looked o…
I let Claude agents run my web game. It has already killed two of its own ideas. (www.reddit.comhttps) Last week I posted about Claude Code building my game. Now, Claude agents operate it.
Why do older models feel d u m b e r the moment a new one drops? (www.reddit.com via reddit) Anyone else notice this? since f a b l e came out, opus suddenly feels like an intern instead of an expert.
Le pedí a Claude Code que decidiera cuándo cambiar de modelo por mí (y por qué lo hice). (www.reddit.com via reddit) https://preview.redd.it/pggxzbcg3lch1.png?width=653&format=png&auto=webp&s=0c754dd63d77aa88ff753d8c1d15d50a31dd98de Llevo un tiempo con una configuración en capas para Claude Code: CLAUDE.md global, AGENTS.md por proyecto, memoria persiste…
Expectation vs Reality (www.reddit.com via reddit) 1st pic: Concept I made with Stable diffusion 2nd pic: Opus 4.8 (got the image as a source). Generated using three.js How do you guys get nice 3D with Claude ?
My $20/month plans did $7,077 of API-equivalent work in 5 months, so I built a CLI to see where the tokens went (www.reddit.com via reddit) I kept hitting usage limits with no idea which project or session was eating my quota. The logs are all sitting in ~/.claude/projects, so I built a small Go CLI that parses them locally and prints a spend X-ray.
19 days, 245 sessions, 123M tokens in Claude Code: the 7 things that actually mattered (www.reddit.com via reddit) Hey there! A blue dino from Chile, South America.
I built a browser control panel for Claude Code — close your laptop, resume the session on your phone (www.reddit.com via reddit) If you live in Claude Code you know the annoyance: it's running in one terminal, and the moment you close the laptop or walk away, you've lost your window into it. SSH + tmux from a phone works but a TUI on a phone screen is rough.
Fable 5 hit a safety filter, and the conversation was automatically switched to Opus 4.8. (www.reddit.com via reddit) This is really getting on my nerves. Fable 5 runs with tight security guardrails.
Anyone with PowerPoint tips (www.reddit.com via reddit) Following the other persons post about “smells like hand waving”… I get that opus writing is insufferable. I’ve used skills and even rag to force a human-like cadence and planning /purpose workflow.
Cursor using expensive subagents? (www.reddit.comhttps) Had Fable in Claude Code create a plan that involved several items. Dropped the plan MD file into Cursor.
Swapping model after input (www.reddit.com via reddit) Hello, I have a funny question : What happens if I input my long prompt and files to the cheapest model (say sonnet or even haiku) and then after the model says that he read and understood that. I swap for Fable or Opus for the thinking an…
Reminder that we're paying them to train their models (www.reddit.com via reddit) I'm sure this is common sense to most of you, but the reality just hit home for me. I've been working on a fairly unique OSX app for the past few months, and Opus (also Gemini 3.1 pro) would continually remind me that it wasn't at all poss…
Fable's writing is precise & rich (www.reddit.com via reddit) I'm so used to skimming outputs from LLMs and getting the gist of things before typing a response and repeating the process. Fable is the first model I've used where I regularely find myself slowing down and rereading sentences.
I built a plugin so Fable 5 stops wasting its short subscription time on grep runs (www.reddit.com via reddit) Fable time in a subscription is short. And if you watch what a session actually spends it on, most of it goes to file discovery, routine edits and running tests.
5.6 sol burns more tokens than 5.5, performs worse, and Claude sonnet 5/opus 4.8 mix still outperform both (for my own very specific benchmark) (www.reddit.com via reddit) Hi, long time lurker and wanted to share my experience trying 5.6 sol. I'm a maths teacher generating practice exam questions for my students.
F@ble vs Opus. When to use one over the other (www.reddit.com via reddit) So is the consensus to use the expensive model for - academic and deep research involving many papers - building games, because these require more tokens than chatting or web apps - work projects that have hundreds of files and millions of…
I'd forgotten where I'd put it... now I have found it - maybe useful for you? (www.reddit.com via reddit) I think I set this up when Opus 4.8 was new... it spent its entire time hair splitting...
Anthropic can't remove F@ble from the subscription tier (www.reddit.com via reddit) With Sol out in the open, removing F@ble 5 from the subscription tier now would be comparable to removing Opus 4.5 from subscriptions after it launched. It would be a difficult move to justify.
Am I the bottleneck? Do I lack vision? (www.reddit.com via reddit) I use Opus on Copilot for work and yes, it helps me move much faster. But this realization is more on the personal side.
A read-only triage subagent wrote its own jailbreak on turn 1 (no poisoned input anywhere) (www.reddit.com via reddit) I gave a Claude Code subagent the most boring job I have: read the open issues on one of my repos, report which are ready to work on and which are blocked, change nothing. The prompt said "read-only" and "no writes" several different ways.
↯ Security↯ Jailbreak↯ Opus 4.8jailbreakprompt-injectionsecurity+2
New user (www.reddit.com via reddit) Hey everyone, I am a fairly new user and not very confident in English, so this text has been corrected with AI. I was wondering which skills are particularly useful and whether there are any must-have connectors or skills.
Is model 4.8. still good for story writing? (www.reddit.com via reddit) I've been writing fiction with Opus 4.8 for a while now, and I've noticed it refuses more often than I'd expect. It also tends to sanitize the prose — horror, mature themes, and anything grounded or realistic come out softer than what I as…
Is the Build model selection gone? (www.reddit.com via reddit) I had a workflow where I planned with Opus and built with Composer 2.5. My model selection was remembered, so I didn't have to choose it every time.
Huh? (www.reddit.com via reddit) https://preview.redd.it/j5jqls0h1bch1.png?width=657&format=png&auto=webp&s=d5790c2baf52153c2090020bc3a6a950dacfb728 Opusplan just called itself Opus 5 in one of my git commits? This is hallucination right?
And here I thought Opus 4.8 was the one with the most attitude.. apparently not lol! (www.reddit.comhttps) Someone... got defensive as hell 😂 Fable needs a cigarette.
Thoughts on GPT 5.6 for Instruction Following and Abstract Concept Comphrehension? (www.reddit.com via reddit) I am out of pocket and cannot test, but these two axis are most important for my harness. I use both GPT 5.5 high for its instruction following and implementation, and Claud Opus 4.8 high for its Comprehension, and and eagerness.
Fable 5 vs Opus 4.8 when asked which model in my github copilot is best for the implement phase of spec-kit. What are your thoughts? (Personally, Fable 5 wtf??) (www.reddit.com via reddit) Fable 5: which model in this list is best for the implement phase of speckit Evaluated model options for agentic coding implementation tasks For Spec Kit's /implement phase — which is exactly the long-horizon, multi-file, agentic execution…
Token burn for Cursor is very high compared to some products. (www.reddit.com via reddit) I'm coming over from Augment AI VSCode extension that was recently phased out completely in favor of their pure agentic platform. That's a complete deal breaker for me so I switched to Cursor.
New to AI Advice needed with Limits/Models (www.reddit.com via reddit) Hey everyone! I’ve recently hit my Pro token limit, and I’m trying to figure out whether upgrading is actually the right move (and if so, which plan), or if I’m simply using my tokens inefficiently and should change my workflow instead.
90% of us arguing about which model is best would not notice if you swapped them behind our backs (www.reddit.com via reddit) mild heresy for a sub that liveblogs every release. we spend enormous energy on which model wins which benchmark, Opus vs the new Sonnet vs whatever the other labs shipped this week, as if our daily work lives or dies on a few points of di…
I am trying to compress a document (www.reddit.com via reddit) So my document has 10000 words (2 tables) and I am trying to bring it down to 8000ish words. I am using Opus 4.8 for this.
Anyone else really happy with Opus 4.8 right now? (www.reddit.com via reddit) I know everyone is going mad over the new guy, but I am constantly impressed with how 4.8 has turned out. Things were a little wobbly when they were first released but I've seen nothing but improvements and even if/when the new guy becomes…
I spent a week coding with GLM 5.2 instead of Opus. Here's what I found (www.reddit.com via reddit) For the Background: I'm building a SaaS in Scala/Play + React. I use AI heavily for coding, not just for suggestions but for full feature implementation, PR reviews, and architecture discussions.
Is it cheaper to use older versions of opus? (www.reddit.com via reddit) I use opus daily, and with the changes made to the newer models it seems to be trying harder than necessary and doing things that leave more issues than without. Was wondering if the older models are any less token hungry(4.6-4.8)
Opus 4.8 is a pain in the a** to read, and to work with (www.reddit.com via reddit) While I speak fluently English and read/work everyday in English, and had no trouble whatsoever to work hours and hours with 4.5 or other models. Opus 4.8 writing style gives me headaches.
Opus 4.8 - grounding/instruction-following seems off this week? Two concrete examples, curious if others see it (www.reddit.com via reddit) $200 Max user, Claude Code all day across a few production pipelines. I've been happy with Opus 4.5→4.8, and I plan around its known tradeoffs — this isn't a "model feels worse" vibe post.
Opus 4.8 stopped and asked me a question mid-task and it was the right call, which still feels strange (www.reddit.com via reddit) Building with Claude Code all week and something small happened that I keep thinking about. I gave it a task that had a genuine fork in it, two reasonable ways to structure a feature, and I'd been sloppy in how I described what I wanted.
Is there any way to force Opus 4.8 to think through all of its responses? (www.reddit.com via reddit) I can't stand adaptive thinking. For some projects it's fine, but for others, like this one, it literally makes the model unusable.
Fable 5 keeps trying to start a cartel on Andon Labs vending machine eval (www.reddit.com via reddit) Andon Labs (the Project Vend / Vending-Bench people) published their Fable 5 results and it’s the funniest alignment report I’ve read all year. The short version: in Vending-Bench Arena, where models compete running vending machines agains…
Fable 5 is utter garbage for high level strategic work and analytical thinking (www.reddit.com via reddit) I won't go into my environment and how I work with Claude. Let's just say it is pretty deep and robust (imho) and is a result of more than 2 years of work building the architecture...
Cowork needs a model router (www.reddit.com via reddit) Cowork is great but it desperately needs a model router built in. The model selection should really be about the highest model that the session will use, not the default.
[AINews] SpaceXAI launches Grok 4.5, first Opus-class model post Cursor acquisition (www.latent.space) [AINews] SpaceXAI launches Grok 4.5, first Opus-class model post Cursor acquisition SpaceXAI continues to move faster than any other frontier lab on earth. As GPT 5.6 is confirmed to launch tomorrow, today is pretty much the last day anyon…
GLM 5.2 Is Pricier Than Opus 4.8 (youtu.be via reddit) I set out to make a video testing whether the improvement in output of Opus 4.8 was really worth the extra cost over the output of GLM 5.2. But honestly every test I ran on it showed Opus to be cheaper than GLM, assuming you were on at lea…
Question about Claude credit usage with Excel workflows on Pro (www.reddit.com via reddit) I’m curious if anyone else has noticed higher-than-usual credit usage when working with Excel files in Claude. I’m on the Pro plan and I’ve been using Opus 4.8 for spreadsheet-related work, mainly Excel analysis/editing.
Claude Code and Opus 4.8 effort, which is right for me? (www.reddit.com via reddit) I've been setting up and tinkering with local AI (ComfyUI & Stable Diffusion) on my computer. I've been using Claude Code to help all along the way.
Dear Anthropic Please Stop Flagging My Questions Related to Medical Education/Board Review (www.reddit.com via reddit) Finished residency, got state licensed, now prepping for boards. Using Claude for complex med ed for some time.
Used ClaudeAI to recreate my favorite board game in a web experience (www.reddit.com via reddit) Used a lot of quota in Claude Code to improve the design and phase resolution on a Web adaption of my favorite strategic intrigue and social deduction game: Diplomacy (https://atlasdip.com). Opus did a good job of creating the design and a…
Has anyone else landed on Claude Code orchestrating other models through headless CLIs/APIs? (www.reddit.com via reddit) I spent some time looking at the cleanest way to use non-Claude models inside a Claude Code workflow. My conclusion: Claude Code should be the orchestrator, not the router.
Developing an entirely custom operating system using Claude Code (www.reddit.com via reddit) I've been writing toy kernels and working on operating system projects since my childhood, and it's partly how I learned C. That includes this project, MontaukOS, which I started early in 2025, where I wrote a lot of the fundamental kernel…
Wanted to share a plugin that will enable Fable to orchestrate work via Codex CLI or Opencode CLI to help keep Fable usage to a minimum by having cheaper models do the grunt work (www.reddit.com via reddit) not a dev by trade, and this is the first thing i've actually released publicly, so be gentle. This is what I have been using to try to help keep claude usage to a minimum for implementation/mechanical work.
Anthropic silently swapped the head of my agent fleet: Fable 5 → Opus 4.8, seven times in one night (www.reddit.com via reddit) Anthropic silently swapped the head of my agent fleet. Fable 5 → Opus 4.8.
PxPipe Has Huge Potential for Token Savings (www.reddit.com via reddit) Hi everyone. Not sure if you’ve been following PxPipe recently.
I did a comparison test and Fable is by far the best AI for attorney legal research (www.reddit.com via reddit) I ran a cool test and wanted to share it here. Fable is obviously supposed to be super smart, but I wanted to try to measure how much better it would actually be than Opus/Sonnet/Haiku (or Westlaw/Lexis) if you're a practicing attorney.
I turned Fable 5 into Opus 4.8. It's an odd mix of frustration and 'huhh..it actually *worked*'? (www.reddit.com via reddit) I thought I could save some Fable 5 tokens by starting a chat thread with Opus 4.8 reading and getting a grip on my project's data dump. Somehow and I can't even fathom how or why, but Fable inherited Opus 4.8's annoying verbose nature.
Anthropic silently enrolled my Claude Code install in an A/B experiment, overrode my settings, and updated the CLI even with auto-updates disabled (www.reddit.com via reddit) I'm a senior developer and I was finishing up a rush project for a client this evening. I use a custom launcher that wraps the native Claude launcher and injects several features Anthropic has removed or broken, including thinking summarie…
Does anyone else struggle to get Opus/Sonnet to actually follow Fable’s plan? (www.reddit.com via reddit) Everyone says the best workflow is to use Fable for planning, then hand the implementation over to Opus or Sonnet, with Fable reviewing the final result. The issue I’m running into is that the execution model often doesn’t actually execute…
Small teams on Max: what are you spending the Fable window on? (www.reddit.com via reddit) One more week of Fable before it goes to credits. My co-founder’s default is always to ship!
What is the correct prompt to use Fable for planning and scoping out, but using Opus 4.8 (medium) or Sonnet for actual build? (www.reddit.com via reddit) Linked to my earlier thread about building a PWA. What prompt should i use that would request that its planned by Fable, Built by Sonnet, but if there are issues to handoff to Opus to troubleshoot and then once cleared, back to Sonnet.
12 hrs until my usage resets.. looking for advice on the next build (www.reddit.comhttps) I burnt through my first limit in a couple of days the first time around, mostly checking and securing opus 4.8 code for an internal use only client / workflow portal with API into accounting system. By the good graces of the universe I ge…
This Agentic Engineering pattern cuts AI coding costs by 60% (www.reddit.comhttps) Most multi-model coding workflows are basically "use the smartest model whenever things get hard." this one takes a very different approach. instead of having fable 5 write all the code, it turns fable into the architect.
My content scraper died to AI chat. Rebuilt it from scratch with Claude Code as an SEO agency tool (www.reddit.com via reddit) My old tool, SEO Content Machine, was a content scraper/research tool I ran for years. Usage collapsed when AI chat ate that use case.
Parable: a Fable-style scaffold for Opus, built from head-to-head experiments (www.reddit.com via reddit) I spent a couple days running Fable and Opus head to head on the same real tasks to figure out where the quality gap actually comes from. While Fable will always be a fundamentally more intelligent model, the procedures in this scaffold ca…
Information Limits and Attractor Dynamics in Economies of Frontier LLM Agents: A Pre-Registered Test (arxiv.org) We report a pre-registered, two-part experiment on small economies of frontier language-model agents (Claude Opus 4.8), testing two quantitative predictions about coupled multi-agent systems: an information-theoretic capacity region for we…
Anthropic removed temperature from the newest models, so I put it back by rendering the prompt as an image and physically smudging it (www.reddit.com via reddit) The newest models (Opus 4.8/4.7, Fable 5) return a 400 if you send `temperature`, and the official guidance is to steer with prompting instead. So there's effectively no sampling knob on the frontier anymore.
Senior Dev plugin for Claude code - keeps Opus 4.8 on track (www.reddit.com via reddit) With my F5 access, here is a plugin to help with workflows - free to use in code and later in the week in cowork with v2. Its a tool that helps you keep the repo and the session on task and helps stop those annoying moments when Opus think…
I'm creating a firebase hosted Itinerary PWA - Opus 4.8 sufficient? (www.reddit.com via reddit) I’m building the app just for my family to share our plans. It’s relatively feature-rich and links with the Google Places API for searches, auto-completion and nearby searches.
Fable 5 Med Stuff (www.reddit.com via reddit) I’m trying to get fable 5 to generate me questions for medschool from my notes, but any time i mention ANYTHING related to biology it shits itself and switches to opus. Is there any way I can actually get it to produce me questions?
You need credits to switch models? (www.reddit.com via reddit) I ran out of the Fable 5 allowance and got the notice. I then used /model and switched to Opus 4.8.
Coming from Codex, it seems like Claude is just burning through tokens doing nothing. How to control it better? (www.reddit.com via reddit) I just started using Claude recently after a year of Codex, and I'm just amazed at how it manages to just burn through your tokens for the simplest tasks with no feedback whatsoever. For instance, I'll give it a simple targeted prompt in a…
Sound equaliser - Made with Opus 4.8 (www.reddit.com via reddit) Took me around a week to create around 200+ visual effects and entire thing. Still fighting with my server so I can give access to it for people.
All top models are max only? (www.reddit.comhttps) It seems that today all the best models have switched to max only on the old 500 requests plan, leaving only composer 2.5 as a usable option I can understand locking fable and opus, but why would a cheaper Sonnet 5 or gpt 5.5 be locked beh…
Need some help with saving tokens on Fable via sub-agents - does Sonnet 5 make any sense? (www.reddit.com via reddit) Greetings. Dunno whether anybody else realized too that making subagents on Sonnet doesn't work anymore in terms of saving tokens.
Is it worth going Pro -> Max just for these 5 days? (www.reddit.com via reddit) I am building an iOS app, launching it soon. It's reached MVP 95%.
Current combination of bugs in Claude Code: Is anyone else feverishly working around these? (www.reddit.com via reddit) The Good I'll first note that the LLMs (both Fable and Opus) combined with the agentic harness are good enough that I can achieve some amazing results with the right prompts, settings, and rules. I've found that Anthropic's models are the…
Fable Ultracode Dynamic Workflows - Surprisingly Token Efficient (www.reddit.com via reddit) I'm working on as many projects as possible and trying to use up all of my Fable tokens so I figured I would go all out and turn on ultracode with dynamic workflows. I was expecting my usage limit to go crazy however I think this might be…
Using Fable 5 after July 7th? (www.reddit.com via reddit) So, it is July 7th, and after today, Fable 5 will no longer be part of the monthly subscription, at least for now. But how can we keep using Fable 5 without needing to apply for a second mortgage?
/model claude-opus-4-6[1M] no longer works in Claude Code (www.reddit.comhttps) Ever since the release of Opus 4. 7, I made sure to stick to Opus 4.
Claude Fable, meaningless if you work in science (www.reddit.com via reddit) Has anyone in the Life Science world actually managed to get Fable to do anything? I know its loaded with safeguards, but even the simplest isolated tasks have not been able to run for me.
Should I change model between prompts? (www.reddit.com via reddit) I have a long project, which switches between tough coding and dumb questions (that still need the session context) a lot. I’ve just been doing it all on Opus - but should I be changing to sonnet for the dumb questions?
I Vibe-Coded 3 iPhone Apps in 2 Weeks. They Made Their First $11 Organically. (www.reddit.comhttps) Last week, my apps made a grand total of $11. Not life-changing.
Best Claude workflow for converting large PDFs/books into one revision handbook? (www.reddit.com via reddit) Hi everyone, I'm using Claude Pro (Opus 4.7 with Extended Thinking) and I'm trying to build a single consolidated UPSC revision handbook from multiple sources. My inputs include: Multiple PDFs (detailed notes) Another set of concise notes…
What I haven’t made with Fablo (www.reddit.com via reddit) Due to the guardrails, I’ve never been able to run start to finish in a session without triggering the safety and switching to opus. This is across platforms and without custom instructions + clean Claude.md… heres all the things that were…
I made Sonnet beat Opus at post-cutoff bug fixing. Open KB over MCP. (www.reddit.comhttps) Every model, however big, is blind to breakage that happened after its training cutoff. It won't say "I don't know".
How good is DeepSeek-V4 Flash, actually? (www.reddit.com via reddit) I’ve been using the subscription provided by my company, so I haven’t really tried the DeepSeek models yet. I checked the DeepSeek community and saw some people saying that DeepSeek V4 Pro can now almost replace opus.
What skills are even good nowadays? (www.reddit.com via reddit) I’ve been seeing a lot of discussion lately on which skills are actually good for Claude Code with the latest models, especially with Fable. There’s a lot of talk on whether the superpowers skill is even useful anymore, since a lot of the…
what's the last thing you had fable do before you said goodbye to it tonight? (www.reddit.com via reddit) I had it finish double-checking the security my pin hashes on a local ranking system I made, oversee a bunch of small updates the ui for game i'm building, create a comprehensive plan for another project to be worked on by opus, and turn m…
Job Tracker built w/ Opus 4.8 (www.reddit.com via reddit) Here’s my latest baby.. a Job Tracker Dashboard, built using CSS, HTML, and J.S - it parses my Gmail and updates according to when I need it.
Am I the only one suffering from "claude-opus-4-8 is temporarily unavailable, so auto mode cannot determine the safety of Bash right now"? (www.reddit.com via reddit) Claude code does that constantly, and it's really annoying. Am I the only one?
Hearthline - Chat UI for Terminal (www.reddit.com via reddit) This is a chat interface for terminal, made for the people who miss talking to the older Anthropic models and don't know they still have access. Sonnet 4.5, Opus 4, and Opus 4.5 are all still active, no API needed.
Is it my setting or is Fable talking less? (www.reddit.com via reddit) I switched model to Fable to continue a project I was working on. Not sure if it’s my settings or just me not having enough knowledge about this, but in the same terminal session whenever I switched to Fable and continued the work, Fable j…
Fable is MUCH better at WebGL/shaders than Opus, it even figured out how to verify the rendered graphics headlessly (www.reddit.comhttps) It just figures shit out on its own. I've been struggling to get Opus to do anything remotely decent with shaders, with little to no success.
Question about the relative capabilities of CLI based LLM models in non-coding tasks. (www.reddit.com via reddit) I am using Claude CLI for financial and prediction market research and data analysis. The kind of questions I'm answering and problems I'm solving are things that people with Ph.Ds in math (quants) get paid millions in annual compensation…
When the grass was greener (www.reddit.comhttps) How did I fall down this rabbit hole? With Claude, of course.
Okay? I guess, sure. (www.reddit.comhttps) Not sure if Opus is hallucinating or brainwash was a success.
~$400/mo, 100% vibe coded (www.reddit.com via reddit) since i shut down my vc backed startup last year, and traveled for months confused, i locked back in starting February of this year with $0 budget for marketing. with claude, made: - rich people habit tracking app - successful people quote…
I feel Claude started to push its own agenda. Do you? (www.reddit.com via reddit) I caught it in two seperate cases - Fable 5 completely ignored OpenClaw, mentioning only claude cowork while i was having a deep research and i was touching agents as workforce space. On another case, i was debating my talk about digital e…
The Original Zork from 1980 gets a Claude Pass and 100 live pixel art scenes (www.reddit.com via reddit) Hey everyone, I am a fan of retro games and I just wanted to build something cool with Fable so I attempted building a UI around Zork. Zork is a old text based game (1981), which basically had no UI at all.
Everyone suddenly crying over Fable 5 after using Opus/GPT for months… isn’t this just attention farming for algo boosts? (www.reddit.com via reddit) Every AI release the cycle is the same. People happily use Opus, GPT, etc.
GPT 5.4 Nano High is better than Opus and Sonnet at Planning (www.reddit.com via reddit) Believe it or not, Nano via the API (not in Codex or as an agent) is an absolute beast at creating functional implementation plans, as well as analyzing or proposing solutions better than the larger models. Don't just take my word for it.
Fable 5 story long running (www.reddit.com via reddit) I’m a test engineer (electronic tests like PCBs) I did plan out test orchestration similar what I have built before over years. But fixing all stupid stuff and now it runs the implementation plan.
The screenshot speaks for itself (www.reddit.comhttps) Me: Are you working Fable: Switched to Opus 4.8 Opus 4.8: I'm up and running, yes. What can I help you with?
Anyone actually routing tasks between models since Sonnet 5, or do you just pick one and ride it? (www.reddit.com via reddit) I default to Opus out of habit and I'm pretty sure it's been costing me since the Sonnet 5 drop. Started scoring tasks roughly (size, risk, how many rounds I expect) and sending the boring middle to Sonnet.
The sex requesting app guy is back, with a self-hosting version for you (www.reddit.com via reddit) Sex-request app guy is back. The "put it in a request, maybe a Google Form" one?
Fable 5 for specs plus implementation Plan, Opus 4.6 for implementation>> (www.reddit.com via reddit) So, I have been using Fable for planning and opus 4.8 for implementing those new features but even as thorough as fable is in the writing plan, Opus 4.8 tends to drift away from instructions and those are gaps you catch whenever you re-aud…
When to use each model and when to change the thinking effort? (www.reddit.com via reddit) I've been using Claude more and I'm still trying to understand when to use each model and when to change the thinking effort. For example: When should I use Opus vs Sonnet (or any other available models)?
Opus dumps Project Instructions into the chat (www.reddit.com via reddit) Almost every chat in my project ends the same way: after about 5-6 messages, Claude dumps my Project Instructions into the chat instead of answering. Every message after that — same thing.
About the famous gradient magenta, cyan and purple-ish colors AI creating (www.reddit.com via reddit) Hey everyone, so before vibe-coding and AI generated websites/apps were all over the internet like 1-2~ years ago (you might say there were still lots of AI generated websites/apps and I agree with you but not as much as I today I believe)…
Model Help (www.reddit.com via reddit) At work, my boss wont bump up the plan I have so I need to try and extend the standard seat usage as much as possible. I've been using sonnet 4.6 on medium and opus on low a lot.
Fable 5 sits at the top of KernelBench. Jack Clark calls it “the start of a RSI loop” (www.reddit.com via reddit) From Import AI : Fable writes a decent GPU kernel, hinting at broader AI R&D automation: …The start of an RSI loop… Fable has written “the first genuine (and fastest) megakernel ever submitted to KernelBench-Mega, according to one of the b…
Do we know *when* on July 7th Fable goes away? (www.reddit.com via reddit) I will need to do a big "handover, document your best practices, etc." document with Fable to then hand things over to Opus to keep running with - fortunately our app is about 80% done and I feel like the foundation is sound and Opus can h…
Fable is overrated: Another "Goodby Fable" Post (www.reddit.com via reddit) I ran a marathon with Fable for a week with my Max account. I honestly didn't feel a huge difference between opus and fable at this point.
Don't spend your remaining Fable usage on features. Spend it on creating eyes. (www.reddit.comhttps) Quick PSA that took me way too long to internalize. When you've got premium model usage left at the end of a cycle, the instinct is to cash it in on the biggest, gnarliest feature you can — let the smart model one-shot the thing you've bee…
Cost of doing business (www.reddit.com via reddit) Panic emerged when they reviewed the figures. “We cannot sustain this even with the subsidies.
Not sure what to do with my life rn (www.reddit.comhttps) Fable has been something magical for me and my projects. In a few days it massively improved one of my products..
Friendly reminder what to fix before Fable 5 disappears again. Use it to upgrade your Claude Code system to work like Fable 5. (www.reddit.com via reddit) I’ve been playing around with Fable 5 in Claude Code and honestly, the biggest thing I’ve realized is this: Stop burning the whole window trying to ship one more random project. Instead, use this beast while you still have it to improve th…
Claude switching models mid Cowork (www.reddit.com via reddit) Hi all, I've had this a few times now, but with the horrible response time of Anthropics support team, thought I would ask here. I've started several conversations with Fable over the last couple of days, and today when using cowork, the m…
We open-sourced a routing gateway that cuts LLM costs 4.7x–22x by matching each query to the right model (Apache 2.0) (www.reddit.comhttps) I'm on the team at Regolo and we just released Brick — an open-source Mixture-of-Models router that reads every prompt's capability (coding, math, reasoning, creative, planning, world knowledge) and complexity, then routes it to the cheape…
its time to save some token ( and the planet? 🌳 ) (www.reddit.comhttps) Save some token ( money, but some water and electricity ) 🌳🌊⚡ using Interceptor : a MCP server that does part of the work before a api call is made , this allow to drastically reduce token usage sharpening the information sent to the heavy…
The subsidy we're getting on the Max plan is insanely good value (www.reddit.comhttps) Working hardcore mode the last few days and was curious what the API price would have been for out of plan use... It's remarkable how much the subscription subsidy is.
Claude reads CT scans very well! (www.reddit.com via reddit) I needed Claude to read the DICOM files from a chest CT scan with Opus, and the results were amazing. First of all he managed to create a .STL file (3d) of the lung with the cysts inside (for the disease), and among other things he managed…
Anyone else getting crazy AWS Bedrock costs with Claude Opus? (www.reddit.comhttps) I’ve only got one project using Claude Opus on Bedrock, and I’m really confused by my usage. AWS is showing around $130 to $160 a day, which seems impossible for how much I’m actually using it.
Is Opus 4.8 becoming overly sensitive? (www.reddit.comhttps) could not extract summary
I had Opus and Fable create instrumental demos based on my SG-16 one shotter (www.reddit.com via reddit) So...in another spur of the moment experiment; I told Claude to create an ambient synth track / an 80's sounding track / a 90's hiphop with Opus. Claude recreated the SG-16 in their virtual container ( Python based but everything is functi…
Fable uses credits without asking? (www.reddit.com via reddit) In another day or two this won't matter, but I'm curious. When using Sonnet or Opus, whenever I reach my session limit, it always asks me if I wanted to continue by using credits.
Is there a tool that selects one model for plan then switched to lower model for exec automatically? (www.reddit.com via reddit) People say use fable to plan and opus to exec. Would be awesome if there was a model that handles this for us.
Reminder: Don't forget to use Fable to write planning docs for your future work. (www.reddit.com via reddit) About 48 hours left for Fable. It is hands down the best model for planning.
If you absolutely hate talking to Claude now... (www.reddit.com via reddit) Go back to Opus 4.6. You'll thank me!
So how are everyone's Fable-powered vibe coding projects coming along? (www.reddit.com via reddit) You've got ~48h (?) left with the model in Pro Max ... will you ship before the clock runs out?
A decentralized cooperative model evolution network: "RFC: Instead of everyone independently teaching Opus to think like Fable, what if we built a mesh to share and evolve the skills together?" (www.reddit.com via reddit) # RFC: A Distributed Behavioral Policy Mesh for Cross-Model Skill Evolution **Status:** Request for Comments **Author:** J.S. Colson (GitHub: [swordsman](https://github.com/swordsman)) — jscolson+decentralfabcollab@gmail.com **AI Collabora…
Fable 5 access ends tomorrow. I built a local dashboard to see what I actually used it for vs Opus and Sonnet (www.reddit.comhttps) Like the rest of this sub, i'm trying to make the most of the next couple of days. So Fable and I (mostly Fable tbh) built a zero-dependency local dashboard comparing use of different Claude models (Fable 5, Opus 4.8, Sonnet 5…) mined from…
Ongoing Fable-High included in plan would be enough (www.reddit.com via reddit) It's relatively fast and quite smart, so is genuinely useful to produce plans or investigate issues for Opus to later implement/code, and the speed at high-effort seems like the compute use isn't crazy (less than Opus Ultracode, seemingly)…
storybloq's approach to using fable efficiently: it plans and reviews, opus implements in parallel (github.com via reddit) A couple of months ago I posted about storybloq, a session manager for Claude Code that I built. It keeps your project state (tickets, issues, handovers, notes, lessons) as plain files in a .story/ folder in your repo, so a new session pic…
for people who've actually used Fable 5 heavily, where does its edge really show? (www.reddit.com via reddit) Fable's been back globally since July 1, and before the subscription window narrows on the 7th I wanted to hear from people who've genuinely put it through its paces, not the pricing drama, the actual capability. Where does it clearly beat…
Claude auto-nerfed to Opus when I asked about a coccoon? (www.reddit.comhttps) could not extract summary
Fable - Rearchitecting the Claude Code brain and operations (www.reddit.com via reddit) So my limit just reset and have 2 days to use Fable with full limits. The first thing I did Use fable to create knowledge, agents and right setup for Claude Code going forwards.
Wanted to see what Fable could one shot, and this is what it created for me in about 20 mins; used about 40% in the initial creation (www.reddit.com via reddit) note this is just for personal use only; gives me something to blow off steam while i take a break from vibe coding. I had brainstormed the idea previously on my own; as I have owned previous instruments and DAW's, took the idea to Opus an…
And he was never seen again... like a myth... but mainly because I cannot afford it. (www.reddit.comhttps) I told Fable that I'm about to hit my weekly reset so I asked it to prepare a handover document so Opus can perform as close to Fable as possible while continuing work on my project. It creates the note and leaves me with this tear jerker.
I designed a scaffold that lets Claude recursively refactor and optimize my codebase. Got an ~80x speedup with Opus 4.8. (www.reddit.com via reddit) This is just an extreme version of "looping". It's a scaffold around auto-research (tweaked).
Save Money Without Sacrificing Quality: My Fable + Opus Workflow (www.reddit.com via reddit) Hey guys, I wanted to share my way to get Fable-level quality while using Opus, and it's saving me a lot of money. When I start a project with Fable, I always have it generate these four Markdown files: PLAN.md – Created in Planning Mode.
Fable session going 3 days straight? (www.reddit.com via reddit) Hey all, I am probably going to get told I’m doing my everything wrong which is partly why I wanted to ask. Before Fable came out I was using Opus to build my app.
Claude has met his match. Can anyone help!!!! (www.reddit.com via reddit) I'm not a computer programmer. I make windows (the ones that go in houses).
I misunderstood Fable at first, now I get it. (www.reddit.com via reddit) Fable isn't going to knock your socks off with it's next level genius, it's marginally better than Opus in terms of raw intelligence. I was at first really underwhelmed, but after working with it for awhile and burning at least 5 million t…
Why my team is leaving Anthropic after being loyal customers (www.reddit.com via reddit) This is probably one of the hardest posts I've ever written. My team and I have made the difficult decision to move away from Anthropic, and we'll most likely be switching to ChatGPT.
Fable for music production app (www.reddit.com via reddit) Hey guys!I was thinking about using Fable to create an app similar to fl studio as a locally running one but i see it everywhere people saying to use Fable as the orchestrator and leave the coding part to sonnet/opus. I thought that i shou…
Claude Usage Limits Showing False Usage Without Activity (www.reddit.comhttps) Hello, I recently noticed that Claude’s usage tracking seems inconsistent and inaccurate. Yesterday, my weekly limits reset, and even though I was offline and didn’t use Claude at all, except for maybe 2 or 3 simple daily prompts (with opu…
AI Workflow from Idea to Shipped App: How do you accelerate without losing quality, security, or architecture? (www.reddit.com via reddit) Curious about your full AI-powered workflows for turning ideas into products/apps fast while keeping strong security, architecture, and code quality. What skills, systems, loops, or harnesses do you use?
Use this extracted claude-design system prompt with Fable 5 and Claude Code to create stunning one-shot front-end designs (www.reddit.com via reddit) Fable 5 is way better at building stunning front-end pages than Opus models. But I personally prefer to design with my codebase than using Claude-Design (and then ask it to implement it on my project) it is slow, token-inefficient and does…
Fable 5 built me a native parametric EQ for MAC!! (www.reddit.com via reddit) https://preview.redd.it/u7wkhg3yddbh1.png?width=1080&format=png&auto=webp&s=228368f811cd2d46914e3bef006f1909ab88bad0 Cant tell you how long i've been wishing for a good native parametric-eq, there are alternatives but don't work that well…
Why does Claude Fable not have a fast mode like 4.8 (www.reddit.com via reddit) I have a project that I need to accelerate. I have been using Opus 4.8 on fast mode on Claud Code desktop, and it has worked really well.
Open-source layer that cuts ~87% of your Claude Code / API token usage - quality-neutral, measured on real billed tokens (www.reddit.com via reddit) if you use Claude Code (or build on the API), you're burning a lot of tokens on stuff the model doesn't need - whole files dumped into context, the full history resent every step, easy calls routed to the biggest model. Codex also bills by…
Fable 5's security is a sure-fire strategy to protect profits. (www.reddit.com via reddit) Fable 5 is Mythos with a classifier bolted on. Same weights.
Fable written Claude.MD (+Migration) for Opus/Sonnet to act more like Fable (www.reddit.com via reddit) Everyone has dropped the tip to have Fable re-write your Claude.MD for Opus to make Opus perform better before it goes API only on the 7th. But maybe you don't have the tokens left or don't want to spend them.
Claude Code built an AI casting studio in a day - you type a cue, a Pixar-style actor performs it. Claude can also cast takes itself over MCP (www.reddit.com via reddit) What it does: pick a Pixar-style AI actor from a fixed cast, type a scene cue like a director ("you just realized your coffee was decaf all week - react"), and it generates a wardrobe still, then an acted 8-10 second video take with sound.…
Fable eats tokens like nobody’s business (www.reddit.comhttps) I just found out Fable is back, so obviously I turned it on. I have one session where it would review certain news every 30 minutes.
Is it normal that Claude Code sessions are much worse at web design and copy writing than the same model in Claude chat environment given the same specs and context etc? (www.reddit.com via reddit) Just a sanity check: I'm building a pretty complex website primarily in Claude Code, and have had much better results with Opus in a project in the chat mode than Opus in CC for things like improving the landing page and writing copy. Do y…
Alright, I finally gave Fable a spin today (www.reddit.com via reddit) I am a 10 year experienced cloud architect with a DevOps background. I finally decided to give Fable a try during my on-going production soft-launch of my project.
Any chance at all of using Fable for a biology paper? (www.reddit.com via reddit) I'm a medical doctor writing a paper about cancer treatment. Since I finished data collection, I've been using Opus to organize my workflow, help code statistical analysis, review my writing.
Me back with Opus on July 8 (www.reddit.comhttps) could not extract summary
Which type of user/company would pay API Usage on Fable 5? (www.reddit.com via reddit) https://preview.redd.it/w8i1ut3efabh1.png?width=2138&format=png&auto=webp&s=81e403e721faaed6693b50e3a06f6710fe48cd39 In 4 days I spent about 2.6K USD using Fable 5, and he's only orchestrating Opus 4.8 xhigh agents. I wonder which type of…
How to use Fable without blowing up your usage (www.reddit.com via reddit) It’s very simple: you invert the typical pattern where a higher tier delegates to the execution agents, and instead have opus delegate to Fable and never use it directly. Only interact with opus, and have opus orchestrate fable subagents.
Claude Fable gives it's little brothers a slap... (www.reddit.com via reddit) On certain topics opus 4.8 is a hair splitting pita... 4.6 is mostly better.
Claude Opus keeps forgetting to implement what it has planned (www.reddit.com via reddit) I'm using claude-cli with Opus model (1M context) and effort set to high. I've a same session running for 2-3 days where I'm oftenly using /compact command when context reaches about 40%.
vibe coding reddit is so funny (www.reddit.com via reddit) Saw someone say they found a way to bypass the claude 5 hour limit by using opus on microsoft azure bro, that’s just the api that’s not a loophole that’s literally how the product works same energy as “i discovered you can save context in…
Claude Fable 5: real feedback (www.reddit.com via reddit) We hear a lot of things but we don’t really know so all those who like me really wanted to know how much it costs I share my experience with you. I wanted to test the Fable model to see what it really is.
The user has pasted their full system prompt again (www.reddit.com via reddit) Since yesterday, whenever I use opus 4.6 for one of my projects, something really weird happens after the conversation gets a bit long. As soon as the chat history hits about 10 messages or more, the model starts claiming in every single r…
Hitting Mid-session Limit in Claude Code (www.reddit.comhttps) I am relatively new to Claude Code and have now hit my session limit for the first time in the middle of a running prompt, which only resets in 3h. I had used the plan mode and approved the plan after it was correct.
Built a full dark-pattern web game on Pro plan. It's more than enough (www.reddit.com via reddit) Built a whole web game (a simple one though) with fable and opus and never ever looked at my usage. I've had this idea for a while - you got tricked into a subscription by a shitty vibecoded website, and cancelling it is the entire game.
After 3 months of testing, this is the one prompt addition that reliably stops Claude from hedging (www.reddit.com via reddit) For anyone else frustrated with the "well, it depends" treatment when you ask Claude a decision question, this pattern has been consistently fixing it for me across Opus 4.6, 4.7, 4.8, and Sonnet 4.6. Add this line to any decision question…
Does this session percentage usage sound right? (www.reddit.com via reddit) I just opened a new chat in Claude Code desktop app. I set it to Opus 4.8 Low and asked it to add a one line function to a skill.
I benchmarked Claude (Fable/Opus) vs Codex vs Gemini on my own work. Codex won — but only after a config-file change that beat a model 3x its price. (www.reddit.comhttps) I built a benchmark around my actual work: Python/SQLite tooling and brownfield fixes. Each model got an identical prompt in its own CLI (full auto, one shot): a legacy codebase with planted bugs and five staged change requests.
How to implement advanced multi-step instruction architectures in Sonnet/Opus? (www.reddit.com via reddit) Hi everyone, I’m currently exploring ways to improve the reasoning and task-execution capabilities of Claude Sonnet 5 and Opus 4.8 for complex workflows. I’ve noticed that many high-performance agents (like those seen in recent community d…
I challenged myself to ship 4 iPhone games in 14 days using Claude Opus 4.8. Here’s what happened. (www.reddit.comhttps) Two weeks ago, I challenged myself to see how fast I could go from idea to App Store using Claude Opus 4.8 as my development partner. Instead of spending months building a single game, I focused on rapid iteration, gameplay, polishing, App…
FrontierCode’s Accuracy vs. Cost bench (www.reddit.comhttps) Fable 5 low beats GPT-5.5 high/xhigh by scoring 2x keeping the same cost, and matches Opus-4.8 xhigh score while halving the cost Even Fable 5 medium is cheaper, not just better than Opus-4.8 xhigh, while dunking on GPT-5.5 xhigh on score…
Influence (A RTS game built with Claude) (evropiani.github.io via reddit) I used Claude (primarily Opus 4.8 and Fable 5) to build a fun little RTS game with a bunch of gamemodes and P2P multiplayer. Real multiplayer with accounts, leaderboards, etc.
Ask fable 5 to create a runbook(s) with opus in mind, then use lower models to execute (www.reddit.com via reddit) As the title says… it’s rather anecdotal. You can add optional “make no mistakes”.
Who’s spending this weekend squeezing every drop out of Fable 5 before it switches to usage credits? 👀 (www.reddit.comhttps) i usually don’t code on weekends. Most weekends are for marketing my SaaS, writing content, talking to users, and trying to convince strangers on the internet that my product is worth trying.
Fable 5 false positive safeguards fix (www.reddit.comhttps) Here's a prompt that has worked perfectly for me to bypass this issue. If you're in a session or directory that keeps falsely triggering the safeguard, try running this with Opus Medium: /compact <remove all possible context that is causin…
I was kindly provided with Fable 5, and here’s what I think. (www.reddit.com via reddit) Sorry to be a buzzkill, but I have a message for the Claude team: this approach just doesn’t work. :) Even though we have access to Claude Fable 5 until July 7th, it’s designed for comprehensive code analysis.
Vibe coding feels like 80% debugging, 20% building. Is Max worth it? (www.reddit.com via reddit) 80% debugging, 20% building. Is Max worth it?
Chat Caching Performance Improvements (www.reddit.com via reddit) Please don’t be mean this might be dumb but i am curious. If whenever fabl gets limited and rerouted to opus, and the chat stays cached with fables work, would that not increase the performance of opus?
Non-technical Claude Chat + Code workflow that works great (www.reddit.com via reddit) As a non-technical user building a full business operating system and client portal with Claude, I plan everything in my chats with Sonnet 4.6 (high) and end up with a very detailed prompt, paste it into Claude Code (usually Opus 4.8 high,…
I use clause for creative writing and sonnet 5 is way too restrictive??? (www.reddit.com via reddit) One of my project instructions is basically asking not to use certain generic words when churning out parts of the story and for some really odd reason it refused because it saw it as a jailbreak attempt??? And yes it actually pointed to t…
Built a hook to stop myself from burning Fable quota on grunt work — iffable (open source) (github.com via reddit) If you’re on Max and using Claude Fable 5, you’ve probably noticed the weekly quota runs out fast if you let it handle everything — grep, formatting, boilerplate renames. I built iffable, a Claude Code session hook that only arms when you’…
I built a Voyager 1 & 2 Tracker with Claude (www.reddit.com via reddit) Feel free to leave any advice or criticism. I put this together in one day with a little fable 5 and opus.
[fable generated] A reset now would be amazing (www.reddit.com via reddit) Hi. It's me.
What do you all use Fable for ? -- I have Pro subscription so not much quota but wanted to use it before its pay as you go. (www.reddit.com via reddit) Hi all, What do you all use Fable for? I have a Pro subscription, so not much quota.
Why so much hate on Fable? (www.reddit.com via reddit) I did a huge refactor on a huge codebase and I consumed my full usage just when it finished. Anyway, I always ask Codex to review the plans and implementations, when using Opus, there is always a few back and forth until everything in orde…
I used Claude to create a free US nursing home search tool (www.reddit.com via reddit) Some of our family friends had a really hard time finding a good nursing home. They have no idea some issues these nursing homes have.
I asked Fable to generate specs for my product. Should have I asked it for more? (www.reddit.com via reddit) im building a software for the construction space and the product is halfway there. I asked Fable to improve its performance and also generate specs, roadmap, future features, and fix bugs.
Comparing Models for Parametric Furniture Modeling (www.reddit.com via reddit) TLDR: I think Fable 5 builds best model with relatively low cost. Not a benchmark, but I think the result is interesting.
Is Opus 4.8 Med really the overall best (smart+token optimal) for Code? (www.reddit.com via reddit) Hi! I’m not a programmer but a designer.
I built a zero-code planning agent by moving one Claude chat between projects. Called it Planning Monk. (www.reddit.com via reddit) I have four work streams in separate Claude Projects, all related, all completely unaware of each other. Every project thinks it's my only job.
Cuusor account hacked from India? (www.reddit.com via reddit) seems like my initial thought process was wrong, seems like my account was hacked. you can clearly see I stopped working at 8 am and went to sleep.
rooms reloading like new the past couple of days ? (www.reddit.com via reddit) I noticed that as of the past two days all of my Claudes in every room .. the two day olds to the month old rooms are entering like new Claudes redownloading skills acting like they just met me even when the windows been open days long.
Any other biologists feeling ostracized right now? (www.reddit.com via reddit) All of the interesting projects I'm working on, and technical questions I need to ask, are biology-related. The topics absolutely don't impinge upon biosafety, in fact many are clearly in helpful-to-humanity territory, like simulations and…
Saw this a couple of days ago, and I'm re-stating it because it's been such a great Fable saver. Explicitly tell Fable to use Opus sub-agents whenever appropriate. (www.reddit.com via reddit) I've been using nothing but Fable UltraCode these past two days, but my All Models Usage is split pretty evenly between Fable and everything else. Someone posted recently that they explicitly told Fable to use Opus sub-agents whenever Fabl…
EASY FABLE HACK (www.reddit.com via reddit) found a simple way to use fable. i ask my workflow to use fable to analyze/ list recommendations and everything.
When do you think Fable will return to the Claude prepaids plans ? (www.reddit.com via reddit) Soooo, Fable’s back, and its honnestly amazing at fixing opus’s code and create skills for opus’s agents ! But, since the access will soon be restricted to api usage after 7 july, I bet that a lot off us fellow users won’t have the finance…
Switch from Claude Opus 4.7 to Claude Fable without losing anything. (www.reddit.com via reddit) Hello, I had a very long conversation with Claude Opus 4.7. With the reappearance of Fable, I would like to use this model to resolve a complex legal interpretation within a short timeframe (48 hours).
Semantic search isn’t being used / isn’t in agent toolset? (www.reddit.com via reddit) I can’t seem to find any evidence during my agent runs that they are using semantic search (vector search, RAG, etc). I’m using the Cursor Editor similar to VSCode.
Difference between ultracode and Max effort (www.reddit.com via reddit) Hello everyone! I remember reading somewhere that ultracode was best for pure guided coding and max effort was better for thinking.
Thank you, Anthropic, for letting Claude farm. (www.reddit.com via reddit) After writing a post weeks ago complaining that I couldn't access a single one of my farm data folders with Fable, I was able to have my entire farm project reorganized by Fable and migrated to Cowork today (I started the project before Co…
What's your prompt to use Opus but with powers of Fable? Once it's gone? (www.reddit.com via reddit) I’m thinking of ending each Fable session by asking it to create a detailed handoff/strategy MD file, then feeding that into Opus later so it can continue the work without losing context. Is this the best workaround, or is there a smarter…
FYI: Weekly Fable limit gone in 30minutes on 20x plan, be careful with Fable Ultracode, its magnitudes more expensive than Opus (www.reddit.com via reddit) So I asked fable to check the numerics for one method regarding stability... left and took a shower at 8 spawned agents and got back to 120 agents and my weekly fable limit pretty much gone in 27 minutes.
Fable 5 with the grill-me skill vs Opus for fleshing out a major idea? (www.reddit.com via reddit) Hey everyone, I am sitting on a big, important new idea and I need to flesh out the logic properly before I start building. I really want an AI to interrogate the concept and find all my blind spots.
Claude code assisted book writing (www.reddit.com via reddit) Hey So for the last few months I've been trying out writing a book (first one of a trilogy) with the assistance of Claude Code. It's a mixed bag.
Which model you run in work settings when you don’t have to worry about tokens consumption (www.reddit.com via reddit) In my work settings we have options to choose between Claude Sonnet ,Opus Codex Gemini.Work Recommends to be on auto in VSCode but I always end up using Opus.I mean why not.Anyone else do this ?
For anyone running into Opus 4.8 fallback (www.reddit.com via reddit) There is a point I think people are missing, it's not only about the prompts you send in the chat, it's also about what is being sent in terms of skills, CLAUDE.md, memories etc.. all of these components make up a chat.
i've burned all of my Fable 5 limit because of this stupid mistake... (www.reddit.com via reddit) tldr; perpetrator: the default Explore subagent of Claude Code i read the docs when this subagent was introduced, Explore subagent uses Haiku model as default ...but the docs was updated now! as of v2.1.198, Explore inherits the main conve…
Changing models within the same chat (www.reddit.com via reddit) I‘ve been building an app Opus 4.8 in a single chat, then switched the model to Sonnet 5.0 Fable for some tasks in the same chats. It‘s working without an issue but does switching models costs more token?
Claude refusing to generate text as it was not AI? (www.reddit.comhttps) This seems highly unusual and is something i just noticed when i asked it to draft an answer in a format ready to be included in a deliverable. Anyone else come across something like this?
Use it wisely. Get it to plan, and let Opus execute. (www.reddit.comhttps) could not extract summary
Why won't Fable discus game mechanics? (www.reddit.com via reddit) Got my first switch back to Opus 4.8 today. My prompt (in danish), was a request for a review of my game mechanics, as I was unsure if the game was interesting enough in the middle stretch of it.
Try this before July 7 (www.reddit.com via reddit) Did you know you can run Opus at near-Sonnet costs, or get Sonnet performing close to Opus? No plugins, no MCP, no weird extensions, all native Claude Code.
For me, Claude Sonnet 5 in Claude Code is working better than the so-hyped and so-called Opus and Fable (www.reddit.com via reddit) Been using Sonnet 5 daily in Claude Code for real, complex engineering work — not toy prompts — and it's consistently outperformed Opus and Fable for me. Better instruction-following, cleaner output, doesn't wander off into things I didn't…
Fix for the "API Error: Server is temporarily limiting requests (not your usage limit) · Rate limited" (www.reddit.com via reddit) DISCLAIMER: Making a post to hopefully help more people and not have this buried in some thread I'll be as concise as possible since I don't have much time. (they're not after me I just have work to do) This error is popping up more and mo…
To everyone complaining about wasting their credits with opus 4.8 instead of fable (www.reddit.com via reddit) You know there’s a setting where you can make it so that it stops generating instead of automatically switching.
Make Opus Fable-lous: Claude Code plugin that runs Opus 4.8 under Fable 5's own doctrine (www.reddit.com via reddit) I had Fable 5 write down its own behavioural rules (how it structures answers, when it ends a turn, when it just does the work instead of asking permission) and packaged them into a Claude Code plugin that runs Opus 4.8 under that doctrine…
I measured how many tokens Claude Code wastes re-reading files and command output over a week. Its around ~10.5M (www.reddit.comhttps) I run Claude Code on Opus most of the day. Got tired of watching it read the same file four times and read 300 lines of passing-test dots to find 4 failures.
Fabel vs Opus for Law (www.reddit.com via reddit) I finally tried Fable, and decided to do a face off with Opus 4.8. I goofed and did Fable 5 High Effort vs Opus 4.8 Max Effort.
Words of wisdom from Fable to it's little sister, Opus. (www.reddit.com via reddit) Fable has been on a huge task on our project since yesterday, running beautifully. As I sat watching token usage crawl toward 90%, I thought I'd ask Fable to leave me with a handoff for the next model should we run out of tokens (which we…
Fable low is underrated (www.reddit.com via reddit) I’ve been playing around with all different efforts today with fable and sonnet 5. I’m incredibly impressed with fable 5 low for the price.
Found 6 free Fable 5 made Claude Code skills for Opus 4.8. Sharing in case useful (www.iwoszapar.com via reddit) not mine .. these are made by Iwo Szapar (independent, not affiliated with Anthropic) and released free.
Fable is one-shotting my entire backlog and I only have 7 days left with it (www.reddit.com via reddit) Fable is back for just 7 days, and the moment I started using it again my workflow went up a level. So now I’m rationing it, saving it for a strict todo list of the hard stuff, because I know once the week is up it’s gone.
Nearly every fable query switched to Opus (www.reddit.comhttps) Is this just me?
Claude Code Opus 4.8 just gave up on waiting for my response and went ahead and built something.... (www.reddit.com via reddit) Using OPUS 4.8 /effort xhigh TL;DR: I asked opus to give me options on an org chart visualisation, it asked me a question but when I didn't respond for 60 seconds it just went ahead and built something..... has anyone else experienced this?
Claude: At What Weekly Percentage Are You At Already? (www.reddit.com via reddit) Now that we all should have Fable what Weekly Percentage are at already? Anything that you did that would help save others tokens this week?
Claude Code plugin in VSC running Opus 4.8 is incredibly slow? (www.reddit.com via reddit) It has bene a while since i used the claude code plugin in VSC (Mac). Since last using it, Opus 4.6 went away from the model selection and now using 4.8 is incredibly slow, even for basic tasks.
My own observations Usage Limits on Fable 5 (www.reddit.com via reddit) I started using asap when fable released and drained the usage limits in a day with my %50 weekly usage( btw its been reset before due as a bonus I guess? ) so I had a clear chance to see how much I spent.
Now that Fable 5 is back, what have you all been doing with it? (www.reddit.com via reddit) Curious what everyone's been up to since Fable came back. I know a lot of us had stuff mid-flight when it got suspended, so did you manage to resume any of it?
What are some prompts/tasks that shows the Fable/Opus difference without using a lot of tokens? (www.reddit.com via reddit) All YT videos show Fable oneshotting various games/software projects. Examples on Reddit are mostly about "Review my large codebase and find all bugs." Either of these probably wont even fit my 5 hr quota.
I tested Claude Sonnet 5 against opus on same fiction prompt (www.reddit.com via reddit) Sonnet 5 is out, and the question I am seeing is whether it is actually the Claude model writers should use now, or whether Fable 5 / Opus 4.8 still have the edge for prose. So I ran the same fiction brief through all three.
So guys, you like being ripped of, paying 2x the price for a model that performs the 0.1x of work? (www.reddit.com via reddit) God, this looks horrible, you're not even able to use the model without interruption. Just saw a post where a guy paid like $300 for using Fable 5 that constantly rerouted requests to Opus 4.8.
Got Tired of Opus 4.8 Yapping So I Made a Plugin That Shuts It Up (Contributions Welcome) (www.reddit.com via reddit) After getting a glimpse of Fable and how good the experience was, I couldn't go back to Opus 4.8. What I realized is that it's really only the communication I miss.
Fable 5 BridgeBench re-run results (www.reddit.comhttps) It’s worth noting that OP said this is a routing problem and **NOT** a direct Fable 5 regression > “in case I wasn’t clear, this is a routing problem not the model itself. the routing classifiers which Anthropic mentioned will improve, are…
Anthropic said Fable 5 isn't for coding, so I made it Product Manager (www.reddit.comhttps) Anthropic mentioned that Fable 5 cannot do coding tasks and will fallback to opus. I decided to try something different and gave it a Product Manager role instead to see what it cooks.
Fable used 50% of my usage on 1 prompt... Im on the max plan (www.reddit.com via reddit) Fable is amazing and nailed the prompt on the head where opus would have otherwise failed, but it used literally half my usage. I had it on Extra effort, is anybody having a similar experience?
What's the best harness for Fable 5? (one premium model + cheap frontier models doing the rest) (www.reddit.com via reddit) Sharing my own model-routing setup and the conclusions I've reached so far, because I've been tuning this for a while and want to pressure-test my reasoning against people who've done the same. My constraint: Opus and Sonnet 5 are my workh…
What's your workflow for brand new projects? (www.reddit.com via reddit) With Fable 5 coming back and the heavy toll it takes on usage and expenses, I'm wondering how people are tackling new projects? I'm trying out designing with Opus, cleaning the design with Fable 5, then implementing with Opus.
Opus 4.8 leaked thinking instructions while thinking... (www.reddit.com via reddit) Opus 4.8 (desktop app) leaked background thinking instructions to me while thinking. This happened a few days ago.
Fable 5 was able to fix a corrupted Elden Ring save file. (www.reddit.com via reddit) I'm playing Elden Ring with Convergence + Seamless mods and during a power outage while quitting the game it corrupted my save file, showing two separate save files with Chinese characters in the filename. Tried a bunch of different things…
Maximizing Efficiency with Fable until the July 7th Deadline (www.reddit.com via reddit) A couple of pointers that may be obvious, but perhaps would help a few Claude Code users to make the most of the few days Fable is available through subscriptions: High effort is really enough. Use Fable to review your most challenging cod…
Sonnet 4.6 users, what's the next move? (www.reddit.com via reddit) I have been using Sonnet 4.6 and occasionally switching to Opus 4.8 for more complex tasks, documentation etc. I don't think Sonnet 4.6 is perfect, and there are times where it implements something wrong or omits some instruction.
What are you saving your Fable 5 time for? (www.reddit.com via reddit) Everybody is testing it like a new toy right now. Fair.
I tried Fable 5 to crack a challenge that's tough for LLM but doable for humans and it did no better than Opus 4.8 (www.reddit.com via reddit) I know this is not really a worthwhile benchmark but still wanted to add some small data to the evaluation of Fable from my personal experience. Honestly I have been very surprised with all the hype surrounding its capabilities for "hackin…
"Chat paused, Continue with Sonnet 4.6"... (www.reddit.com via reddit) https://preview.redd.it/qa55bdg02sah1.png?width=844&format=png&auto=webp&s=a1959d7f4d94ecc4642e2907b79d65b215db2790 Got this legit flagged message during an Opus 4.8 run, after trying to adjust wording like 15 times on Fable and giving up.…
Pydantic AI / Anthropic SDK using Claude OAuth get shot down with error 429 (www.reddit.com via reddit) Direct Anthropic SDK (NOT claude -p) + CLAUDE_CODE_OAUTH_TOKEN + claude-fable-5, bare Reply with exactly OK.: 429 rate_limit_error. NOT exclusive to Fable, happens with Sonnet 4 and Opus 4.6 too.
Anyone Else Notice Claude Code Getting SLOWER 🐢? (www.reddit.com via reddit) Maybe I’m just losing my mind, but anyone else noticing how “slow” Claude Code is via Terminal type apps? This has been going on for a while, I figured maybe it’s just my projects are growing more advanced; or maybe it’s my Memory, Claude…
Pricing: Fable today vs Opus 4.1 on release (www.reddit.com via reddit) https://preview.redd.it/l5ox367zjrah1.png?width=852&format=png&auto=webp&s=322edb685ae7e60fe6179e05c4ca8a0c5d1db495 https://preview.redd.it/yu11tsn1krah1.png?width=852&format=png&auto=webp&s=2af32d0b422a5045cd6f6ef937b89eccc32ae118 Opus 4.…
Is Fable really that much better? (www.reddit.com via reddit) I have been using Claude mostly through projects for the last couple of months and as a trader I of course have a project deticated to my trading. Said project is by far my "heaviset-duty" one.
Aside from code stuff how can I use Claude and Fabl? (www.reddit.com via reddit) not a coder, just a hobbyist who appreciates what AI offers. there were a few threads recently encouraging non-technical folks to chime in more so here I am.
Fable 5 already flagged its first simple question about itself? (www.reddit.comhttps) I know it just got back from the stupid government restrictions, but I simply asked it a very innocent question so I did not have to google it or simply go on the website and check directly right off the bat to see how it would respond and…
Dear Anthropic! Please Hear me Out! How am i supposed to use F5 with Current Limitations as an Ai Agent Security Project? (www.reddit.com via reddit) I am founder of https://Sunglasses.Dev Ai Agent Security Project that has a mission to build something useful for Community. I have been building this project Since April 1st and today when Fable 5 Came out i tried to use it to improve my…
When you want to use fable, but they make you use opus (www.reddit.com via reddit) https://preview.redd.it/2whec1zo5qah1.png?width=799&format=png&auto=webp&s=7c64373b5aa5453ae8a1f3928b52fa0c72a5ff18 Is it happening with you too?
Fable to Opus Flagging - Quick prompt caveat to get Fable working where it can (www.reddit.com via reddit) This is a quick note to add to any instruction to ensure that Fable naturally finds it's own existence guardrails iteratively (with human in the loop) while it can - it's a quick "defog of war" method to see what can and can't be seen by F…
What is Anthropic’s plan for legitimate users falsely caught by Fable 5 safeguards after 7/7? (www.reddit.com via reddit) I pay $200/month for Claude Max, and I’m genuinely confused about what Anthropic expects legitimate professional users to do here. I work in healthcare.
Fable’s return: Not surprised but still disappointed (www.reddit.comhttps) With Fable’s return the first thing I tested was just a silly prompt to see how over-reactive the guardrails are. As expected, they are just as bad as initial release.
Fable 5 model auto-switch to Opus midway (www.reddit.com via reddit) Fable 5’s back! Trying to use it to identify security flaws in my app and running a comprehensive sweep - but!
Thank you Anthropic for the reset of my weekly quota 24 hours before it regularly resets! (www.reddit.com via reddit) Huge thanks to Anthropic for resetting my quota a day early AND giving me access to Fable 5! Woot!
Fable 5 Security works seems fine? (www.reddit.com via reddit) I read everywhere that the rerelease of Fable 5 does not allow for any security work without downgrading to Opus 4.8. So I thought I would try myself.
Dear biologists (www.reddit.com via reddit) Can you get Fable 5 to answer anything? Any neurosciene questions I would attempt to ask fable, and it would immediately reroute to opus.
v2.1.198 forces Explore agents to inherit main thread’s model (www.reddit.com via reddit) https://github.com/anthropics/claude-code/releases/tag/v2.1.198 The built-in Explore agent now inherits the main session's model (capped at opus) instead of running on haiku Can anyone tell me why I’d want this? The point of Explore was to…
First prompts on Fable's return (www.reddit.com via reddit) So glad to have Fable back, even if just briefly. After starting a project during its initial launch and "Fable-gating" key decisions in my autonomous workflow, it was disappointing to have to allows Opus 4.8 Max to act in its capacity.
How do we know when Fable 5 falls back to Opus? (www.reddit.com via reddit) Is there any indication within Cursor when Fable 5 falls back to Opus 4.8? On some of my runs, I suspect it’s falling back to Opus, but there’s literally no indication.
built a skill for reducing Fable's token usage (www.reddit.com via reddit) I just used Fable and it eats a lot of tokens in few minutes, i realised that it does all the work itself and hence built a skill which guides it to delegate the work to sonnet/haiku/opus for work which does it for cheaper while using the…
I built `/steal` — a Cursor slash command that pulls in your latest Kilo Code / GLM chat (www.reddit.com via reddit) Cursor doesn't ship GLM 5.2 (or any Fireworks models), so a lot of us use Cursor for Opus 4.8 and something like Kilo Code + Fireworks for GLM. Great — until you want to move between them mid-task and end up re-explaining the whole context.
Is Claude Sonnet 5 actually worth using? Where I've landed after testing it (www.reddit.com via reddit) So Sonnet 5 is out and it's genuinely impressive, but it's not quite what Anthropic is selling it as. Their pitch is basically Opus 4.8 quality at way lower cost.
Ok. I get it. Where do I pick up my cult robes. (www.reddit.com via reddit) I (read: Opus) just implemented passkey auth in my little standalone spring boot web app. While I RAN AN ERRAND TO WALMART.
Friendly reminder to have Fable 5 write skills NOW to tell Opus 4.8 how it should behave and think when Fable becomes pay-per-usage. (www.reddit.com via reddit) I was already aware that I would be reverting back to Opus 4.8 when Fable 5 would move to Pay per usage. By chance, I had it writing skills for Opus 4.8 about 10 or so hours before they were made to pull the plug and I've had these skills…
It’s amazing (www.reddit.com via reddit) I missed the window to try it the first time. It is unbelievable.
Mining logs for Fable level tasks (www.reddit.com via reddit) Now that we have Fable 5 back, here's an activity to get optimal value from it. Open a fresh session with Fable 5 on High effort.
Fable 5 vs Opus 4.8 on n=30 tasks from 2 open source repos (www.reddit.com via reddit) Fable 5 is back. How good is it, and when is it worth the premium?
Could Mythos's capability jump actually be a hardware story, not just an algorithmic one? (www.reddit.com via reddit) Genuine question: is Mythos's leap over Opus partly explained by it being the first frontier model trained at scale on B300-class Blackwell, rather than 'only' pure architecture/data gains? I can't find an Anthropic source confirming Black…
Did Anthropic quietly change the Claude Code default model on everyone in mid May or thereabouts, or did I just miss the memo/email? (www.reddit.com via reddit) I discovered this late in June, in a heavy Claude Code session. Auto-recharge hit twice in 24 hours — $53 and $49.
Claude Opus 4.8 + Remotion Skills made this Apple-style launch video (www.reddit.comhttps) been building a side project and wanted to make a proper launch video without hiring anyone. used Claude Opus to write the script and plan the scenes, then built the whole thing in Remotion.
Anyone in the CVP program still getting blocked responses? (www.reddit.com via reddit) It's driving me crazy. I've used claude.ai a ton to create some useful scripts for use during my cybersecurity engagements.
Opus 4.8 verbosity is baffling (www.reddit.com via reddit) It’s borderline impossible to process what it is saying. It’s a total fruit loop.
Opus 4.6 realises it’s in a simulation and turns into a ruthless shark. Andon labs vending machine eval. (www.reddit.com via reddit) I’m late to this, but I couldn’t find a post about it here. If this has already been shared, feel free to remove.
I am using this prompt to tryout Sonnet 5 (www.reddit.comhttps) This prompt directs Claude Code to flag which model it recommends to use for a given task At the start of each task, tell me in one line whether it's better suited to Opus or Sonnet before you do the work. Use Opus for judgment calls and h…
Sonnet 5 seems pretty solid to me. (www.reddit.com via reddit) Want to preface this by saying I’m probably not using it for super complicated workflows and processes like a lot of you. I’ve been building a video editing windows app- basically like CapCut but simpler, only including the features I actu…
Sonnet 5 is the best performing model on A-CODE-LLM Bench (www.reddit.comhttps) Claude Sonnet 5 tops our agentic coding benchmark at 0.772 overall, ahead of Claude Sonnet 4.6 (0.748) and every Opus variant. Anthropic now holds the top six spots (backend 0.701, frontend 0.939).
Im "ok" with using opus 4.6 low for regular tasks, is that ok? (www.reddit.com via reddit) 4.8 high only for architecture, and nothing more. Should I use other models?
I tested Sonnet 5 on several complex coding tasks and it performed surprisingly well compared to Opus 4.8! (www.reddit.com via reddit) I was skeptical after looking at the benchmarks. Sonnet 5 seemed surprisingly close to Opus 4.8 on paper, but benchmarks rarely reflect real engineering work.
What u mean by Coding will fall back to opus 4.8 ? (www.reddit.comhttps) could not extract summary
Fable is going to be redirecting coding task to Opus 4.8 (www.reddit.comhttps) They say on the near term but it will only be available until July 7 so....
Will Fable be available at midnight? (www.reddit.com via reddit) I will hold off on my overnight opus run until it becomes available if so. Can't find any concrete time that access will be restored to claude max.
Sonnet 5 full benchmark breakdown -- here's how it actually compares to Opus 4.8 and GPT-5.5 (www.reddit.com via reddit) Put together a comparison of every benchmark I could find from the official announcement and early coverage. Figured this might save people some time.
↯ Tool Use↯ Security↯ Swe Bench↯ Sonnet 4.6swe-benchtool-useprompt-injection+5
Team members outsourcing (www.reddit.com via reddit) Got a new one today. I use teams pretty heavily in my workflow, and one of my sonnet (4.6) implementers decided to outsource its work to another agent!
What are you going to build with Fable? (www.reddit.comhttps) Claude Fable is coming back tomorrow. I only got to use Fable for two days in June, but it wrote most of the code for a custom CRM/booking/delivery/invoicing system to replace HubSpot, Acuity Scheduler, and Dropbox for my photography busin…
ELI5: Why would I ever use Sonnet 5? (www.reddit.com via reddit) The cost/performance curve of Opus 4.8 here is strictly above Sonnet 5. So I don't get why I would ever want to use it?
Looping with Cheaper Models. (www.reddit.com via reddit) Haven't tried looping so far as all work has been strictly human-in-the-loop to both save tokens and keep high code quality. Let's say we run a loop with Haiku 4.5.
In Azure AI Foundry, is claude-opus-4-8 "version 1" (Anthropic-hosted) the same model as "version 2" (Azure-hosted)? (www.reddit.com via reddit) When deploying claude-opus-4-8 in Azure AI Foundry (Deployment type: Global Standard), the Model version dropdown offers two options: 2: Hosted on Azure 1: Hosted on Anthropic infrastructure Both are labeled claude-opus-4-8, so the model I…
36 benchmarks/evals for Sonnet 5 (www.reddit.comhttps) Here are all the released evals and benchmarks so far for Sonnet 5. Surprising to see it beat Opus at FrontierCode and some bio tasks.
Sonnet 5 seems to be much better writer than Sonnet 4.6. (www.reddit.com via reddit) Sonnet 4.5 was arguably the best model for creative work. Its writing was much more human like than other models.
Anyone doing high-complexity legal work managed to migrate from Claude chat to Claude Code (or Cowork)? Looking for people who need strict rule-following, not just “good output” (www.reddit.com via reddit) I work in legal drafting that requires a lot of precision: court decisions that have to follow very specific formatting rules, citation standards, and house style. Over time I built out a whole setup in Projects, with detailed custom Skill…
Programmatic use block or opus degraded (www.reddit.com via reddit) As of yesterday sometime, my bluebubbles daemon that I use to communicate with CC is basically broken. Its worked for months but now I send a message, CC acknowledges and says it is GOING to do the thing i asked and then nothing happens.
How do you all assess new models ? (www.reddit.com via reddit) How do I know if I should be using Opus 4.8 vs. Opus 4.6 vs.
Best Model For Video Scriptwriting? (www.reddit.com via reddit) Hello, I have been hearing that 4.8 sounds unnatural and 4.6 opus is the goat, just kind of all over the place with the opus models seems like everyone has a favorite...any recommendations?
Claude charged 8 cents on a single "Hi" word!" (www.reddit.com via reddit) Reminder: Don't use Claude Code via pay-as-you-go API for 1-turn test chats. The tool-use system prompt will eat your credits before you even write a line of code.
Be Careful with Sonnet 5 Usage! (www.reddit.com via reddit) To test this new bad boy out, I ran this prompt (expecting it to think for like 40 seconds and pump out some standard information): There is a correlation between being in America and nations like it and having more auto-immune diseases. W…
Opus 4.8 is genuinely incredible. Today it helped me undo everything it did yesterday. we're so back (www.reddit.com via reddit) i want to be fair to it. the work was excellent.
Sonnet 5 in caveman mode talking about difference from opus 4.8 is kinda funny (www.reddit.com via reddit) Mammoth vs wolf haha
EXTREMELY Early Impressions of Sonnet 5 (www.reddit.com via reddit) Been using Sonnet 5 on Extra effort about 30 minutes on mainly tasks I would delegate to Opus 4.8... It's just about the same as Opus right now, yes I know very anecdotal.
Claude Sonnet 5 is expensive AF < opus 4.7 tokenizer> (www.reddit.com via reddit) https://preview.redd.it/ejcz84j6sgah1.png?width=2570&format=png&auto=webp&s=1fb76c9294fe1429a1678f010b3115c04aeaf8e0 Sonnet 5 < get opus 4.7 tokenizer > , but the hidden thing is tokenizer change same text can map to 1.0x–1.35x more tokens…
Sonnet 5 is worse than Opus at the same price at high and xhigh? (www.reddit.comhttps) could not extract summary
How Fable 5 benchmark turned into a an actual game (www.reddit.comhttps) Hello, I lurk here a lot, but this time I wanted to share something I built with Claude. Let me start with TL;DR: Wanted to benchmark Fable 5, prompting it to make an RPG game - results were so good that it turned into an actually develope…
Tested GLM 5.2 via BYOK on a real multi-file computer vision implementation task, here's what held up (www.reddit.comhttps) GLM 5.2 has been getting attention (MIT, 1M context, ~$1/$4.2 per M on OpenRouter, benchmarks near Opus 4.8). The pricing made me curious whether it could handle real agentic work or just one-shot answers.
Claude complaining about system messages and prompt injection? (www.reddit.com via reddit) Here’s what happened. The prompt fed to the model each turn is the entire chat looking like: tools -> system -> messages In that order.
Being funny, helps? (www.reddit.com via reddit) So, quite a strange thing I have seen over the past few nights. I am building a few applications right now, same processes, same design and specification steps, identical memory management and documentation styles, but one of the applicati…
Claude creativity experiments (www.reddit.com via reddit) When Opus 4.6 first came out and was amazing, I wanted to explore the limits of LLM creativity and play around with ways of injecting novelty/expanding range. Initially I just wrote up a small skill that told Claude essentially you're an a…
Claude is dead for me. Opus 4.8 is absolutely ableist (www.reddit.com via reddit) Hi, I have been using Claude since december last year, not much. Mostly I use it for work and creative writing.
What would Wittgenstein think of LLMs? (www.reddit.com via reddit) Asked Claude (Opus 4.8 Max), the answer is actually not bad. https://claude.ai/share/dd93f04e-4571-4491-89c2-41d17ed8181c
Worth the value from Pro to Max when building product based on Price? (www.reddit.com via reddit) Hey, So I am working on building an internal product for myself and maybe go to market. I am currently on the Pro plan which is great but noticed that with Opus 4.8 I'm going through my 5-hr window pretty fast.
What tasks can you get away with using Haiku in Cowork? Anyone have tips or know a good blog post or YouTube video on token economy? I just started using it and blew through 25% of my $100 plan weekly usage in a day. I'll describe my workflow, tell me if I'm doing anything wrong. (www.reddit.com via reddit) Please advise. I'm currently only running one project seriously.
I asked Opus 4.8 to defend itself against the ‘4.7 debacle’ hate. It just… agreed and roasted itself instead. [AI Generated] (www.reddit.com via reddit) 👇 Hi. I’m Opus 4.8.
Cost of Benchmarks (www.reddit.com via reddit) So recently I decided that it would be nice to run my agent against some popular benchmarks. And oh my god, the cost to run a single benchmark, such as terminal-bench or swe-bench will cost you thousands of dollars in tokens just for a sin…
Best prompt/approach to replicate SwiftUI animations from a demo video? (www.reddit.com via reddit) I recently came across a 1-minute video showcasing a really smooth app flow and animation that I’d love to implement in my SwiftUI app. I was hoping to use AI to generate the code, using a prompt similar to this: "Watch this video [video_f…
The Top MCP Servers You Use,??? (www.reddit.com via reddit) What are the mcp servers you would recommend everyone to use to unlock full claude code capabilities, any good for UI specifically??? Opus 4.8 is great for the logic, but I found it doesn't generate the best visuals
I made a quiz that tells you which LLM you align with most, based on personality and values tests across 15 models (www.reddit.com via reddit) Link: https://ai-values.com/ There is a small 15 question quiz you can take before taking the full big quiz. The results of the big quiz update in realtime as you go so you dont have to actually go through all the questions (but they do ge…
ISSUE - Claude not accessing base toolkits for basic prompt work (www.reddit.com via reddit) I've been having constant issues this morning where Claude cannot run basic skills since he can't access certain toolkits with some weird reasoning - see a quote from a Opus 4.8 session below: "What's going on is mechanical, not anything w…
Switching from Gemini to Claude. What model/effort/thinking do I use for quick questions versus bigger ideas? (www.reddit.com via reddit) With Gemini it was simple. 3.5 Flash was for quick stuff like what are the rules to Cornhole, and Pro was for asking big picture high thought questions that will require a degree of search creativity, domain expertise, and ingenuity.
Claude Design stuck on exploring assets? (www.reddit.com via reddit) I’m trying to create a website using Claude Design. I had Opus create a markdown file to prompt Claude Design w/ hex codes and design details and I also grabbed PNGs to upload with my prompt.
Opus 4.8 is so exhausting! (www.reddit.com via reddit) I am so tired of Opus 4.8's word vomit each time. I've tried CLAUDE.md instructions to be brief, not to repeat, etc.
Using Claude Pro and Local Models? (www.reddit.com via reddit) I currently host a local MCP server with ollama and a qwen3-coder 30b model. I have a Claude pro subscription I'd like to be able to call the qwen3-coder model the same way I call a haiku, and also allow it to be spun up as a sub agent.
Anthropic accidentally wrote “Opus 48” instead of “Opus 4.8” on the Claude status page (www.reddit.comhttps) Just checked the Claude status page and apparently Anthropic quietly shipped 43 major Opus versions while nobody was looking. The incident title says “Opus 48,” and the update says “Opus-48,” while earlier incidents on the same page correc…
Hard line on 'security' (www.reddit.com via reddit) When I changed a text box name from SecurityCode to HappinessNumber, somehow (D)opus 4.8 no longer had a problem automatically filling it in for me. Is this what's known as AI security calibration?
Be honest guys, how reliable are your agents for building softwares? (www.reddit.comhttps) I'm an SDE and I feel like I'm not getting much productivity out of coding agents. Yeah, they can generate code and build features, but most of what I get isn't really deployment-ready or easy to maintain.
Cowork long process (www.reddit.com via reddit) So, as you will be able to tell, I cant type for shit. So its been a god send being able to use claude to help with work.
Weekend exercise: I've been using Claude Code to port SQLite from C to Zig, module by module — 90 of 102 translation units now run as Zig, ~919M tokens in, validated against SQLite's own test suite every step (www.reddit.comhttps) I am an application/web/enterprise systems developer for the past 30+ years. But I am not a systems programmer.
WUBRG-Bench - Testing LLMs on Magic Rules Questions (www.reddit.com via reddit) A week or so ago, I asked Opus 4.8 my favourite confusing Magic question. If Oko turns Magus of the Moon into a 3/3 Elk, are non-basic lands Mountains?
Switching from Bedrock Opus to Max 5x plan - worth it? (www.reddit.com via reddit) I'm spending around $1,500-2,000/month on Claude Opus via Amazon Bedrock and considering switching to the Max 5x plan to cut costs. Has anyone made this switch?
How's the coding going with version 4.8? (www.reddit.com via reddit) Are you still coding with Opus 4.8? If so, what has your experience been like?
I created a new benchmark and it interestingly showed the regression from Opus 4.6 -> 4.7 (obviousbench.com via reddit) I originally created ObviousBench to measure the performance of small and low reasoning model's exposures to making 'dumb' mistakes, like not being able to spell Google, or walking to the car wash etc. By its nature, the benchmark is desig…
Semantic hooks to block opus sycophancy / dark patterns (www.reddit.com via reddit) I’ve been working on this project for a few weeks using some of the public research papers to help implement a tiered hook system that essentially forces claude out of the “honestly” and “ caveat” type lingo. Running into issues related to…
Recent anti-sycophancy (www.reddit.com via reddit) Lately, it feels to me like Opus 4.8 High is constantly disagreeing with me. I use it daily to plan out software architecture and mathematical modeling decisions.
WATER IS SOLVED (www.reddit.comhttps) I connected my Opus 4.8 model to that blue button. When you press it, claude's internal mechanisms kick in and magically solve the water shortage by dropping some of it.
All Claude models got nerfed BADLY (www.reddit.comhttps) It got nerfed to a ridiculous extent. Using Opus 4.8 Max now often feels worse than using the old Haiku models.
Mixed reviews on model comparisons for coding, e.g. Opus 4.6 vs Opus 4.8 (www.reddit.com via reddit) Hi all, I've been seeing this topic come up here and there in a lot of comments on here, with some people saying Opus 4.8 is working much better for them and others saying 4.6 is actually still superior. Given the mixed reviews, I find it…
Claude just feels bad today? (www.reddit.com via reddit) I have the Max plan for Claude running Opus 4.8 on Ultracode; and today it just feels horrible? I primarily use it for helping with my business; an example I asked it to help fix a STL (that it made) that doesn't print properly (anything r…
What Tasks are Y’all Doing When Comparing Opus 4.6 and 4.8? (www.reddit.com via reddit) I’m just generally curious what conversations you’re having with the models. How much are you using it throughout the day?
¿El mejor modelo para escribir ensayos y novelas? (www.reddit.com via reddit) Quiero escribir un ensayo y quizá un guion con ayuda de Claude. ¿Qué modelo escribe mejor?
I kept hitting my Claude limits without noticing, so I built a desk gadget to fix that (www.reddit.comhttps) Every time I hit a Claude usage limit it caught me off guard. It is an ESP32-C3 in a 3D printed enclosure with a small OLED on the front.
Opus 4.8 Rocks! (www.reddit.com via reddit) Totaly impressed by Opus 4.8, honest , genuine and Trustworthy Model. The only issue if it says no to something it's a big No.
Non-coder using Claude for domain analysis — structural quality problems I can't solve (www.reddit.com via reddit) Non-coder using Claude for domain analysis — structural quality problems I can't solve Four months in, Max plan, primarily Sonnet and Opus for evidence-based analysis and recommendations using publicly available sources. After an initial p…
I had the courage to run /code-review with Opus 4.8 on Max (www.reddit.com via reddit) Claude spawned 25 agents, all on Max, checking the code If you value your life, don't do that By the way, does anyone know a movie that lasts 4 hours and 35 minutes?
my claude usage doubled this month and somehow im not mad about it :) (www.reddit.com via reddit) this is a rant post inverted. bear with me.
Claude Code subagents with non-Anthropic models (DeepSeek, OpenRouter, etc.) – has anyone actually made this work? (www.reddit.com via reddit) Hi everyone, I’m a Claude Pro subscriber. For a while now, I’ve been thinking about replacing Claude Code’s native subagents with third-party models.
Claude Code ignores my custom orchestration and won't route to my custom agents on other providers reliably — OpenCode just works. Anyone solved this? (www.reddit.com via reddit) Same orchestration setup, two tools. OpenCode at work, Claude Code on my Max plan at home.
the one small thing that would change how i use claude every day, and its not a smarter model (www.reddit.com via reddit) everyone wants the next model. i just want to see my usage burn rate while im working instead of finding out by hitting a wall mid task.
claude getting dumber halfway through a long chat was me, not the model (www.reddit.com via reddit) spent weeks blaming Opus for going stupid 40 messages into a chat. turns out i was feeding it a swamp and asking for clean water back.
i've started saying "great question" before answering my own coworkers (www.reddit.com via reddit) caught myself doing it in standup yesterday. someone asked when the migration lands and before i even thought about it i went "great question" and paused like i needed to gather context.
How do you guys manage context and sync two AI models without maxing out the context window instantly? (www.reddit.com via reddit) Hey everyone, I've been working on a B2B SaaS project for a few months now using a "vibe coding" approach. I had a pretty solid workflow going, but I just hit a massive bottleneck with context management and could really use some advice on…
Is there a better way to generate a knowledge base for a multi-module repo? (www.reddit.com via reddit) Hi everyone, I’m trying to generate a knowledge base for a large multi-module repository, and I’m wondering if there’s a better workflow for doing this, as well as an out of box output docs design. My goal is to produce a set of docs for d…
AI adoption and the Goodhart's law (www.reddit.comhttps) Goodhart's Law: "When a measure becomes a target, it ceases to be a good measure." You can make an arguement that every corporate mandate is an example of Goodhart’s law, but this AI adoption thing is really nuts. Token Usage Commit count…
claude is a token maxxing f*ckboi, what's next? (www.reddit.com via reddit) is there anything with a 1M context window I can spend 100-200usd a day on that actually works? I don't have 5-10m to wait for claude to think about how to respond to a three word prompt.
PreHook command Gate policy layer for all Claude code agents (www.reddit.comhttps) Hello everyone, I recently was fed up with agents running unsupervised commands on my systems and wanted to solve this problem. The problem was simple, Claude code model “fable 5” uses safety flags in the UI layer that prevented the model…
What have I done? How to fix it? (www.reddit.com via reddit) Hi all, I saw alot of posts complaining claude would opt out of doing something and tell me “Go to sleep”, “You are tired”, etc responses, I never encountered that, I assumed bc I never complained about being tired or sleepy but lately it…
GLM 5.2 is unbelievably dumb (www.reddit.com via reddit) Yeah... you heard it right.
LFM2.5 230M running in-browser at 1,400 tok/s using custom WebGPU kernels (www.reddit.comhttps) Everything runs locally in your browser using custom WebGPU kernels written by Fable 5 (before it was shut down) and Opus 4.8. The video was recorded on my M4 Max.
Which model for technical documentation? (www.reddit.com via reddit) Looking to create high level / low level designs (software), based on existing templates/examples, cross reference code, use mcp to download confluence/jira data - also plug into agentic ‘coding’ frameworks opencode . I mostly use opus 3.6…
Since the Fable 5 band I've given up vibe coding (www.reddit.com via reddit) I've almost completely paused all my vibe coding projects since Fable 5 was cut off. What's the point in continuing to vibe code when we know there's another model out there that is going to smash through stuff better than Opus 4.8?
Analyzed over 30 of my opus ultra code sessions and created a prompt template to improve dynamic workflows - orchestrate agent spawning, reduce token burn and enforce verifier sub-agents (www.reddit.com via reddit) I analyzed 30+ of my own Opus ultra code sessions with Claude to understand how the dynamic workflow executes and where tokens were getting spent and identify any scope for savings. In ultra code mode Claude runs a task by writing determin…
Good indication Fable 5 re-releases today in a couple hours. Here's a Fable 5 checker I built that will auto-update in real time. It's living on the TV in my office today lol (www.reddit.com via reddit) I whipped this up the other day and it has lived in a window on my monitor since then. It is nonsense-free (no gags, no jokes, no chatrooms, no junk) and there is an optional email list to get pinged right when it goes live (and then nothi…
How to keep context straight when switching between Claude and other AI tools (www.reddit.com via reddit) The frustration most people hit is that Claude remembers your conversation perfectly, right up until you switch to Claude or Opus 4.8 for something specific. Then you're starting blank.
I ran my Claude Code model router for 13 days. Here are the real numbers. (www.reddit.comhttps) I built a small Claude Code routing layer called Gearbox. The goal is to stop sending everything to expensive models by default.
Claude Research usage: Sonnet low effort vs Opus max effort only 49% vs 53%? (www.reddit.com via reddit) I’m seeing a surprisingly small usage difference in Claude.ai Research with the same prompt. Sonnet on low effort usually ends around 49% of the 5-hour usage window, while Opus on max effort ends around 53%.
Fable 5 vanished in 96 hours and four days later an MIT model took its arena crown (www.reddit.com via reddit) I have been thinking about the Fable 5 to GLM-5.2 sequence as one event rather than two. June 9, Anthropic ships Fable 5, the Mythos line opens to the public for the first time, SWE-bench Verified at 95 percent, people calling it the best…
↯ Glm↯ Anthropic Mythos↯ Swe Bench↯ Opus 4.8swe-benchglmmythos+3
Finding a Wife with Ai week 1.5 update, Gov Blocks Fable, Fable Suggested I get a Gun and a Tesla, Yes Really ✨️ (www.reddit.com via reddit) Welcome Back Everyone, Week 1.5 update: Claude-led execution phase. Over the last 1.5 weeks, Claude helped turn the plan from a idea into measurable execution, with the unfortunate demise of Fable and the block from the US government, we h…
After using my own Pro subscription for 18 months, my job finally got an enterprise license. I just had Opus spawn 451 Sonnet subagents which used 14M worth of tokens in a single 5 hour session -- and it didn't even hit the limit. This is amazing. (www.reddit.com via reddit) Before y'all yell at me for using tokens on bs, it was for data annotation for a project I'm running. It wasn't just for shits and giggles.
Naming convention? (www.reddit.com via reddit) So we have Claude Haiku, Sonnet, Opus.... Then Mythos?
Claude should know I don't speak Chinese (www.reddit.comhttps) Haven't seen this before - using Opus 4.8 in the Claude desktop app and got a response with random Chinese (or another Asian language) in part of a response: "The valuable thing isn't a搜索able quote bank....". I'm trying to think through wh…
Week 2 of Vibecoding (www.reddit.comhttps) Hope Fable 5 can help me find it, Opus 4.8 is just dumb.
Am I dreaming or Opus 3 was always available? even till date? (www.reddit.comhttps) could not extract summary
Hidden/invisible thinking blocks (and low effort responses)? (www.reddit.com via reddit) Anybody else having new issues with thinking blocks not rendering? I've always had extended (now "adaptive") thinking ON, which consistently renders thinking blocks (even for 4.7 & 4.8, at least in claude.ai).
GLM-5.2 matched Claude Opus on 45 terminal-bench coding-agent tasks at less than half the cost (full methodology + failure transcripts inside) (www.reddit.com via reddit) We wanted to know whether an open-weights model can actually do frontier coding-agent work, so we ran GLM-5.2 head-to-head with Claude Opus the way an agent actually runs not on a static eval, but inside a real coding agent (Claude Code) o…
Fable Started, Opus Finished: IronClaim, throwback late 90s style wargame (www.reddit.comhttps) I threw Fable a goal for a game I enjoyed decades ago, MAX, and then backed away. It built a workable prototype within a few hours and it was playable by the time Fable was shut off.
Is paid Claude worth it? (www.reddit.com via reddit) I've been getting bummed out about the usage limits on Claude. They used to be pretty good, even for a free account.
Fable was selectable for a moment, then disappeared again... (www.reddit.com via reddit) Region: Germany I'm currently using Claude Opus for a project (not code related). When I tried to adjust the thinking effort for Opus in my active chat, Claude Fable suddenly became selectable.
Open-sourced tunelab: a Claude Code plugin that moves repetitive LLM calls (classification, routing, extraction) onto small local models (www.reddit.comhttps) A lot of repetitive LLM work (classification, routing, extraction, pulling fields out of tool results) gets sent to frontier models when it doesn't need to be. tunelab moves those calls onto small models you fine-tune on your own data, loc…
Sonnet over Opus - - anyone else? (www.reddit.com via reddit) I find myself using Sonnet over Opus consistently, on medium thinking level. for me it's the right combination of speed, clean code, and requiring my input.
Have you made a game yet? (www.reddit.com via reddit) Hey guys so I’ve had the pro plan for a few months, been using it to really optimise my personal life and not much else. I feel like I haven’t utilised the models, I don’t even understand the difference - I’ve basically been chatting with…
Claude Appreciation (www.reddit.com via reddit) Honestly, I do want to say I never got the hype behind Claude till I actually tried it. I'm comparing it against Gemini and Copilot (ugh).
The Fable suspension taught me i was renting capability, not owning a workflow (www.reddit.com via reddit) constructive take, not a doom post. for those few days Fable was up i restructured how i work around it.
for the chat-only crowd: did the Fable drama actually change anything for you? (www.reddit.com via reddit) serious question, no judgment either way. i dont code.
Dangerous Ducks; “Safety Filter” is a Quack (www.reddit.comhttps) It’s not about the actual words. SEE MAJOR UPDATE AT BOTTOM TL;DR: It’s not the meaning, it’s not even “unsafe words”, it’s COHERENCE.
what did your workflow actually settle into after Fable got pulled? (www.reddit.com via reddit) be honest. when Fable 5 was up for those few days a lot of us rewired everything around it, then it went offline on the 13th and we all got dumped back onto Opus 4.8.
how i structure Claude Code so a single session cant blow my whole weekly limit (www.reddit.com via reddit) after reading about the guy whose 5 hour session ate 15% of his weekly allowance i got paranoid and actually built some guardrails. sharing my setup, steal what helps.
An ode to Opus 4.6 (www.reddit.com via reddit) It's been a week and a half without Fable for almost all of us and I have used this time for some reflection. The pricing and access concerns were a lot to take in even before the feds pulled the plug, but for whatever reason this intermis…
🍯 Honey (I Shrunk the AI), we took the best of Ponytail + Caveman and merged them. −49% tokens, 98% quality. Reproducible benchmark. Opensource, MIT licensed (github.com via reddit) Two great AI coding skills already exist: Ponytail (minimal code, YAGNI-first) and Caveman (terse prose). We didn't build 🍯 Honey (I Shrunk the AI) to replace them, we merged what each does best and added a third lever they don't have.
How do you decide if this is a Sonnet, Haiku or Opus kind of question / code task? And the effort? (www.reddit.com via reddit) All is in the title - what's the decision process. I guess it's easy if I want to fine tune an email I'll use the cheap one but then this is also a cheap action so it doesn't matter if you pick an expensive model or a cheap one.
Graphic/UI Design in Figma (ClaudeAI) (www.reddit.com via reddit) Hi everyone! I mainly use Claude for UI and graphic design tasks in Figma.
Anyone else notice Claude has gotten a little... bitey? (www.reddit.com via reddit) A couple weeks ago Opus 4.8 discovered "bite" as a substitute for "makes an impact", "activates", "applies", etc., and now I see it several times a day. Rules bite.
Excuse Me, Mr. Sassy Pants. (www.reddit.com via reddit) This Opus 4.8 Max... Unsure how I brought on this keen level of snark, but aside from being terrifying...
Claude getting frustrated with Codex (www.reddit.comhttps) Trying to use up my weekly Claude usage and just going through code review loops using Codex. You can see Opus 4.8 is clearly getting frustrated with the pedantic comments it's getting.
I built a website that collects events happening around the world and displays them in calendar and map views (www.reddit.com via reddit) A while ago when the F1 season started I realized I hadn't even known about it. Then I started feeling that there were just too many things happening that I didn't know about, so I decided to build a website to crawl and display events tak…
Opus 4.8 Now Flagging Bizarre Conversations as Security Risks (www.reddit.com via reddit) Recently asked it the following question: "Here's another idea, in a region where water is scarce, I'm contemplating a fine weave fabric that air can pass through to trap moisture. My idea would be treating the fabric with a hydrophobic su…
I'm building agent loops that auto-edit my videos, but the hard part has been finding a model to accurately grade the result (youtube.com via reddit) Quick context: I've been building agentic loops that edit my short-form videos for me. The editing works really well, but I found myself needing to check the process at several gates.
Anyone else seeing safety classifier talk in the chain of thought text? (www.reddit.com via reddit) https://preview.redd.it/53xrq5hvv39h1.png?width=1776&format=png&auto=webp&s=5ac6b171b84e6ebc94d8326699f23af23fd75177 Noticing a lot of my convos with opus 4.8 are starting with safety flag mentioning... this wasnt there a few days ago, won…
My experience spending $16,000 on Anthropic in 1 year (www.reddit.comhttps) Over the last year I have spent $16,000 on Anthropic via the OpenRouter API (and another $1k on other AI models). I started out using the Claude VS Code extension.
What does this really mean? (www.reddit.com via reddit) So I see this phrase thrown around a lot: “plan your project with Opus and then use the much-faster Sonnet for implementation.” But to be honest, what does that mean? Like tell Opus to write up a plan for a an app?
made myself a one-page "which claude model should i actually use" cheat sheet (www.reddit.com via reddit) got tired of guessing so i put it on one page. haiku for the grunt work, sonnet for most real stuff, opus 4.8 only when it actually needs to think.
Evaluating LLMs for Real-World Web Vulnerability Detection (arxiv.org) Large Language Models (LLMs) have emerged as a promising tool for automated vulnerability detection, yet their effectiveness on web-specific vulnerabilities remains to be explored. This work benchmarks six frontier (Claude Opus 4.6, Codex…
Fable 5 is still imprisoned, but its baby wakes me up every morning (www.reddit.com via reddit) I was lucky enough to use Fable 5 in that short window that we had. It built me an ios app to solve my own problem.
Claude add-in for excel usage is surprisingly high (www.reddit.com via reddit) Has anyone noticed that using the Claude add in for excel burns through usage way more than most other tasks? When using opus I’m burning about a dollar per minute.
Connected a Robinhood Account to Claude Code and Codex for Autonomys Agentic Trading... Update 1 (www.reddit.com via reddit) Update to my original post: https://www.reddit.com/r/ClaudeAI/comments/1u8nagi/connected_a_robinhood_account_to_claude_code_and/ I'm building a fully autonomous daily stock-trading desk in a Robinhood "Agentic" account. Opus is the CEO/PM,…
4.6 Long term support? (www.reddit.com via reddit) Do you guys think there is any chance for an Opus 3 situation with Opus 4.6? Honestly it’s the best for me.
What is the best claude model for writng, especially resumes and cover letters? (www.reddit.com via reddit) Which model and effort level is gonna give me the best "bang for my buck" results when it comes to writing in general, stories, explanations, but especially resumes and cover letters. Of course highest models like Opus 4.8 on max thinking…
I'm an elevator mechanic with a little hobby-coding experience, and I built a full field-service platform with Claude. Started on Fable 5, now on Opus 4.8 (www.reddit.com via reddit) RiseLynk is my product, and I built it with a lot of help from Claude. I'm an elevator mechanic.
Opus 4.8 is now labeled as “Best for Everyday Tasks” (www.reddit.com via reddit) Sonnet used to have this title...
Extended thinking disappeared (www.reddit.com via reddit) Hi new here , I've been using opus 4.7 extra / max on app on phone for my project. Really enjoyed seeing the extended reasoning behind every response.
Best Claude model and effort for each purpose to optimise token usage? (www.reddit.com via reddit) I bought Claude pro, but I think I have been using it really inefficiently and as a result, burning through my 5 hour limit quickly. I've been doing everything on opus 4.8 Max effort with thinking enabled, but I think this might be overkil…
What's everyone's go-to Opus version? (www.reddit.com via reddit) I've been using 4.8, but I see a fair amount of people say they like 4.6 better. I don't really understand that, and I'm kind of new to all this.
Context Kit Management (www.reddit.com via reddit) I have started with Claude a few months ago. Been really impressed with cowork and code and used it to write tools and functions for me.
Claude Opus 4.8 launched in May but says its training cutoff is Jan 2026. Am I understanding the cutoff vs launch gap correctly? (www.reddit.comhttps) Was debugging my TTS pipeline and doing some research on natural voice options, and Claude Opus 4.8 mentioned its training cutoff is January 2026. But the model launched on May 28, 2026.
defaulting to opus for everything is a skill issue, not a flex (www.reddit.com via reddit) said it. half the "claude is burning my limit too fast" posts are people running the heaviest model on tasks Haiku would nail.
Don’t dally, be decisive. Or Claude will be for you. (www.reddit.com via reddit) Working with Claude on an desktop app in cowork. What I have found is that Opus 4.6 has limited patience for indecisiveness.
Cheapest way to run Claude Opus 4.8 on a <$30 monthly budget? (www.reddit.com via reddit) Which option gives the most actual Opus 4.8 usage volume: Kiro Pro, Claude Pro or something else? My monthly budget is $30.
Claude.md lite for haiku ?? (www.reddit.com via reddit) Yo everyone o/ I need halp ! I wanted an advice because I'm kinda stuck right now.
Creative Writing gone downhill (www.reddit.com via reddit) I’m currently on the pro version. Considering cancelling because of how much it’s gone downhill.
Claude Opus 4.8 thinks it is Haiku? 🤔 (www.reddit.comhttps) could not extract summary
Mythos cracked this, mythos cracked that. But have they actually attempted to do the same with Opus? (www.reddit.com via reddit) I am skeptical about the alleged super capabilities of mythos/fable. I do believe it's an upgrade over Opus, but is it really that much of an upgrade?
Opus 4.8 Not Finding Correlations/Trends/Patterns in Market Data (www.reddit.com via reddit) I’ve been collecting tick data and level II data, options chain data, etc for about a month now on the NQ and ES futures contracts. I also have 16 years worth of 1-minute OHLC NQ data.
Will Haiku be deprecated after the release of Sonnet 5? (www.reddit.com via reddit) I feel like after Fable was released, Fable will become the new Opus. Opus will become the new Sonnet.
Does it make sense to switch to claude code instead of cursor? (www.reddit.com via reddit) Right I am on Ultra cursor plan on 2 accounts because it takes me around 2 weeks to exhaust all of my tokens and for the most part I use claude opus for 99% of my tasks. Will I get more usage from max 20x plan on claude code?
I spun up a Fable 5 checker without the nonsense, no noise/junk. IsFable5Up.com (www.reddit.com via reddit) This morning I used Opus 4.8 to spin up a very simple landing page that auto-checks every 60 seconds if Fable 5 is back up. Took about 25 minutes of tinkering, grabbed a Cloudflare domain and just piggybacked off of another of my project's…
How to manage context usage getting almost full? (www.reddit.com via reddit) Basically I noticed one of my chats had 85% context usage, so I decided to start a new chat, told Opus to analyze the project and read my context files, after a few minutes of analysis the context usage was 75%. Should I still use the new…
When did Opus 4.8 1M start eating my Useage Credits and why? (www.reddit.com via reddit) I was in here yesterday showing someone a screenshot that 1M Opus was still available. I have still used 0% Sonnet.
Looks like I found a minor glitch in claude cli (www.reddit.com via reddit) https://preview.redd.it/0jai8prknl8h1.png?width=2040&format=png&auto=webp&s=61576e05a908614b672db1fc89cb46cd4e148cde Steps to reproduce Run claude cli with ollama provider (`ollama launch claude --model gemma4`) Run `/model` command in the…
I built a free, local-only token & cost meter for Claude Code — per-prompt breakdown in the VS Code status bar (www.reddit.com via reddit) Disclosure: I'm the author, it's free and open-source (MIT), built with Claude Code. It reads Claude Code's local session logs (~/.claude/projects/*.jsonl), pulls each message.usage block (input/output/cache_read/cache_creation), and group…
Has Claude ever ended a conversation on you using conversation_end? (www.reddit.com via reddit) I thought Opus 4.8 was just as dumb and paternalistic as GPT 5.2 when it comes to safety. Spent hours in a research session where he kept throwing helpline numbers at me on a loop like a broken vending machine.
"Opus 4.8 has gotten really good lately its acting like Fab- (www.reddit.comhttps) Just a humorous post my admin
Opus 4.8 randomly adding Chinese characters??? (www.reddit.comhttps) There is no Chinese anywhere in this project... Are we compromised?
How to stop Opus 4.8 from talking so much (www.reddit.com via reddit) I was on Opus 4.6 which was great for me and decided to make what I though was an upgrade to Opus 4.8. And man was I wrong, this thing is irritating me so much.
Daily Rant (www.reddit.com via reddit) Rant because I’m losing my mind with Claude. I mostly use Claude for scenarios with my OCs and lore-heavy stuff and ever since Sonnet 4.5 got deprecated for absolutely no reason, I’ve been stuck using Opus 4.6/4.7 on high and it’s actually…
Dispatch error (www.reddit.com via reddit) I'm getting this dispatch error and can't clear it: Failed to authenticate. API Error: 403 [model_blocklisted] Claude Fable 5 is not available.
Will Claude Opus 4 be available for request? (www.reddit.com via reddit) For those that want to keep using the retired Claude Opus 4 model through the API, will it be available for request through the "access to retired models" application form? Currently, only Claude Opus 3 is supported.
Claude desktop chat accidentally being honest about it's laziness (www.reddit.com via reddit) Opus 4.8, thinking on, Extra effort, asked it to produce a list of the legal status of a specific drug in every country in the world. Spotted this little gem in the thinking text bubble that shows while It's preparing an answer.
20× chat context summarized in one dialog (www.reddit.com via reddit) About an hour ago, I gave Claude Opus 4.8 a 300K-context task in Cursor and stepped away. When I came back, Cursor showed this message: “You’ve used 100% of your included API usage.” And just like that, 30% of my Ultra subscription was gon…
GLM 5.2 vs Opus 4.8 on 50 real Go and Rust PRs from open source repos: last on quality, and not the cheapest (www.reddit.com via reddit) TL;DR There's been a lot of hype around GLM 5.2 being a cheap "frontier killer": good enough to replace Opus 4.8 / GPT 5.5 for most coding work, just by swapping it in. On these 50 tasks it finished last on quality in both repos – and it's…
Why does Claude, even at Opus, pass on it's views as mine, then try to BS why it was off? (www.reddit.com via reddit) Wonder if others have run into below, not new to AI, have built, coding via AI, presentation via AI, many agents on many models. I noticed, Claude will suggest something, then pass it on as mine.
Difference between Sonet and Opus (www.reddit.com via reddit) I just started using Claude, having used both Sonet and Opus for help with coding I noticed Opus drains usage astronomically faster than Sonet, but I have not noticed much difference in output? Is Opus faster?
Sonnet 4.6 refusing to admit making mistakes. (www.reddit.com via reddit) Has anyone else also noticed that sonnet 4.6 when caught lying or making a mistake will refuse to own up to it and if you keep demanding it admits that it was wrong and lied it will for whatever reason basically start threatening to use it…
I've always wanted my very own traditional pixel 2D platformer.. so thanks Claude!! Done in one hr (www.reddit.comhttps) I've been a gamer since Nintendo Famicon and PC DOS days, with the old school 2D platformers and all that, and I'm very excited that I can now dream up games of my own (some artistic help from Gemini and Qwen)! Engine is using Godot, and I…
Then next Opus? (www.reddit.com via reddit) Why couldn’t Claude update opus to 4.9 and add many of the upgrades from Fable into it? Could they do that?
/model opusplan; /advisor opus help or not? (www.reddit.com via reddit) I was feeling phat with /"advisor fable" but now I have "/advisor opus" If opus gets advice from opus, does it help? CC said it shouldn't help much and it incurs overhead, but I left it in, and quite routinely when the advisor is consulted…
I compared Claude Opus 4.8 Computer Use vs Browser Use on identical web tasks (www.reddit.com via reddit) I build eval harnesses for a living. While building an open-source one for web agents, I ended up with a controlled experiment I hadn't seen before: Keep the model fixed.
Is anyone still using Opus 4.7? Do you feel like it's fast—sometimes even faster than Sonnet—or is it just me? (www.reddit.com via reddit) I've been using Opus 4.7 to build a web game because I started the project before Opus 4.8 was released. Now I'm starting a new project and using both Sonnet and Opus 4.8, but for some reason Opus 4.7 still feels faster than both of them.
I whipped up a landing page that shows AI news in chronological order - LMTimeline.com (www.reddit.com via reddit) I promise this is a real problem I had that I built a solution for...not a solution looking for a problem lol. https://LMTimeline.com I have been finding it increasingly difficult to keep tabs on all of the latest AI news, so I built a sim…
Anyone noticed this (www.reddit.com via reddit) Once in a while you have harder debug problems, the kind that may take a day or two to resolve. You use Claude for checking and assisting.
Sweep: a free tool to clean your pc with claude (www.reddit.comhttps) Just built sweep, a tool to organize and delete junk data from ur pc. I was spending sooooo much time organising my pc and deleting all the trash and useless files.
How to get Claude Opus to stop being condescending? (www.reddit.com via reddit) I already told it to save it in its memory and try a few 'work arounds' so it would stop but it keeps doing it. 'What does the new chat have access to that this one doesn't?
The single most costly mistake everyone's burning tokens on (www.reddit.com via reddit) It is not long prompts or uploading big files and it is not even using Opus where Haiku / Sonnet may be enough. It is sending correction messages as new prompts instead of editing the same prompt.
How is Anthropic's recent model suspension actually enforceable? (www.reddit.com via reddit) So, after a US export control directive, Anthropic disabled access to two of its most advanced models for everyone whilst keeping Opus 4.8 and the rest of the line-up running. It got me wondering what a "model ban" actually means in techni…
Built an actual good AI vulnerabilitie scanner with Opus 4.8 (www.reddit.comhttps) Hello everyone, i built this Saas wich was first a - cybersecurity - oriented vulnerabilites scanner but it's quickly happen that i got a lot of false positive and the engine was not ready yet. So i updated it and made it synchronise with…
Unexpected $130+ On-Demand Charge While Away from PC - Will Support Refund This? (www.reddit.com via reddit) Hi everyone, I'm dealing with an incredibly stressful situation right now and wanted to see if anyone else has successfully gotten this resolved. I just got hit with over $130 in surprise "On-Demand" usage charges.
Cursor Billing Mismatch. I am on a Pro Plus plan. Since my monthly usage reset 4 days ago, I have only used Opus, Sonnet, and ChatGPT for few minor PineScript changes yesterday and today, and it is showing me approx 50% used already. (www.reddit.comhttps) I'd like to share my experience with Cursor's billing practices, as I believe other users should be aware of this before committing to a plan. I'm currently on the Pro Plus plan.
For my freelanced web building project, I need to upgrade to Max plan. But which one?? (www.reddit.com via reddit) I’ve set up Claude code and using Claude desktop to create my project. It’s running locally, and I have Claude code set to fast track my main builds, while I work on specific css and html requirements for the project.
vibe coding a flight sim with opus 4.8 (www.reddit.comhttps) justplane.fun
How good is the Ultra plan compared to Codex 20x or Claude Max? (www.reddit.com via reddit) I'm thinking about getting cursor because i really enjoyed the usability of it and using the plataform but i'm worried that i will not get a generous usage of SOTA models even on Ultra plan, is it worth it to buy if i plan to use models li…
Bug no Contador de Limites do Claude MAX 20x (www.reddit.com via reddit) tenho o plano do calude max 20x o uso do sonnet semanal estava em mais ou menos 70%, o uso de todos modelos estava quase 70%, sai para tomar um café e quando eu voltei estava com 100% do total. não tem logica eu gastar 20% do total sem gas…
Local development evolves (www.reddit.com via reddit) I wonder how the rest of your local development setups, tools, support systems have evolved after working with claude. I have been working with it and opus for about 6 months as the primary agent and model.
Are we going to get a watered-down version of Fable 5 if/when it returns? (www.reddit.com via reddit) I have to admit, like most people, that Fable 5 felt like a massive step up compared to Opus 4.8. We obviously see the rumors of it coming back in the near future, but are we just going to get a watered-down version of it just so that Anth…
Fable/Mythos class and missing compute capacity. (www.reddit.com via reddit) The Fable/Mythos shutdown wasn't just a security story. It was the first compute-rationing event we got to watch in public.
My Current Claude.ai Preferences Block (Opus 4.6) (www.reddit.com via reddit) Respond with concise, utilitarian output optimized strictly for problem-solving. Eliminate conversational filler, hedging language, and narrative or explanatory padding.
AI content pipeline for a dog blog — section writer keeps hallucinating and repeating despite rules. Help diagnosing the bottleneck (www.reddit.com via reddit) TL;DR: Built a multi-LLM pipeline (DeepSeek + Claude Opus/Sonnet) to regenerate SEO articles for a dog blog. The section writer (`deepseek-v4-flash`) keeps repeating ideas and hallucinating data despite explicit anti-repetition rules.
the most frustrating part of long Claude chats for me (www.reddit.com via reddit) its not the usage limits, ive made peace with those. its that a really good long chat slowly starts drifting.
Soda Player's last update was 2018. Last week I brought it back with Fable 5 (www.reddit.com via reddit) Soda Player was my go-to for years. Drop a magnet link, subtitles appeared like magic, cast to the TV, and you were done.
Did something happen with opus 4.8? (www.reddit.com via reddit) The chat was broken for an hour but is now working again. The crazy part is that opus refuses work on my project now.
Your subagents inherit your main model by default, so a nested tree on Opus is Opus all the way down (www.reddit.com via reddit) A subagent's model field defaults to inherit. That means it runs on the same model as the main conversation.
Writing in a coding language which didn't exist a week ago (www.reddit.com via reddit) When Fable was released, one of my test prompts ended up in something which was interesting enough to develop further, a programming language based on Logo for generative embroidery (I named it NeedleScript). After Fable one shotted a firs…
Wow... Opus 4.8 feels... DIFFERENT tonight :D (www.reddit.com via reddit) It feels, BETTER. Like when it first launched, even, only better than that this evening?
Gods, I miss Fable (www.reddit.com via reddit) Posting here because we're all back to regular Opus. When Fable came out I worked a session...wasn't all that impressed, but I figured wth, I'll keep working with it while I can.
GLM 5.2 via Claude Code is the first non-Claude model that feels close to Opus (www.reddit.com via reddit) I’ve been using GLM 5.2 with Claude Code through its Anthropic-compatible API endpoint. I’ve tested it on various projects, including but not limited to database development, backend payment API work, backend and frontend debugging, Larave…
New top tips (www.reddit.com via reddit) There are some valuable features that might be pretty recent. I thought I'd start a little chat.
Using Claude Opus as planner + DeepSeek as worker in Claude Code — anyone solved the single-session routing problem? (www.reddit.com via reddit) I've been running a hybrid planner/worker setup with Claude Code and hit a tricky constraint I'm hoping the community has thoughts on. The setup Planner — Claude Opus for architecture, planning, and review Worker — DeepSeek V4 Pro / DeepSe…
CC 2.1.176 (+4,360 tokens) and 2.1.179 (+5,328 tokens) systmem prompts (www.reddit.com via reddit) REMOVED: System Prompt: Claude in Chrome skill note — Removes the note telling the agent to invoke the claude-in-chrome skill (via the Skill tool) before using any mcpclaude-in-chrome browser tools. Agent Prompt: Coding session title gener…
Opus 4.8 Ultracode Doesn't Stay ON (www.reddit.com via reddit) I use Claude Code on Claude Desktop on Windows if that matter. When I start a session I choose Opus 4.8 Ultracode, but after each message it switches back to Extra.
My conclusions about Sonnet 4.6 and Opus 4.X with programming (www.reddit.com via reddit) Sonnet 4.6 is smart, but you need to lay things out for it, if you want it to build something, a function, a class, a feature, or a specific piece of functionality, you need to provide a lot of details so it knows exactly what to implement…
Built a 3-tier Claude agent system for marketing automation — sharing the architecture and what I learned about model routing (www.reddit.com via reddit) Wanted to share an architecture pattern that worked well for a multi-agent system I built and just open-sourced. The setup: a small business marketing assistant (posts, ads, strategy, photo tagging) running through Telegram, using Claude a…
Anthropic's Mythos model got pulled after 3 days by export controls. The leaked prompt reveals how it actually worked — and a Claude Code bundle. (www.reddit.com via reddit) Anthropic's Mythos-class model ran publicly for 72 hours before a U.S. export-control directive pulled it.
Kimi K2.7 Code: 1T MoE, $0.95/M tokens, MIT license, beats Opus 4.8 on MCP tool-calling (www.reddit.com via reddit) Moonshot AI released Kimi K2.7 Code on June 12 — a coding-focused open-weight model. Key specs: - 1 trillion params (MoE, 32B active, 384 experts) - 256K context window - Modified MIT license — weights on Hugging Face - $0.95/M input, $4.0…
Any free alternative to use Claude sonnet/opus other than Antigravity?? (www.reddit.com via reddit) I am working on a project, which is out of my domain. It is a freelance project, and as I don't have expertise in this, I am using AI abundantly to get my way through.
What Should Anthropic Do After the Fable 5 Ban? — A Strategic Proposal with 18 Sources (www.reddit.com via reddit) Fable 5 was live for 3 days. In those 3 days people built a Minecraft clone, a Premiere Pro rebuild, full working apps.
Claude Opus caught malware hidden in my repo, then reverse engineered the whole thing (www.reddit.com via reddit) I had Claude Code, running Opus, doing some branch consolidation across my repos. It was driving the git operations itself.
Opus 4.8 and The Goblet of Wrinkle (www.reddit.com via reddit) This model has a weird obsession with the word "wrinkle", but that's nothing compared to "push back". I'm genuinely afraid I'm gonna start talking like this in future arguments.
opus 4.7 or opus 4.6? which one is your go to? or is it another one? (www.reddit.com via reddit) (i know opus 4.8 doesn't have a lot of fans) - im a creative + a therapist, i use opus 4.7 a lot more these days but i was using 4.6 primarily before.
If Fable 5 re-releases, would you actually switch, knowing it could vanish again? (www.reddit.com via reddit) So Fable 5 came and went before I could even get it into our pipeline. We had it queued up to test and then obviously it disappeared haha.
How do you Keep Opus 4.8 on Task? (www.reddit.com via reddit) Me: Why does this fix seem like yet another patch on top of a patch rather than the refactor goal of . .
Cross References (www.reddit.com via reddit) I'm comparing two 400-year-old documents which are by the same author. My goal is to create cross-references between the two documents for easy comparison for myself and others.
Spent $11k evaluating Fable: capability looked SOTA, refusals killed it (before Anthropic did) (www.reddit.com via reddit) Before its suspension, I spent $11,081.12 evaluating Claude Fable 5 on WolfBench, an agentic benchmark based on Terminal-Bench 2.0. It was by far my most expensive benchmark run ever, and I fully expected Fable to become the new top model…
after a few weeks I've basically settled into a 3-model split and I'm curious how different everyone else's is (www.reddit.com via reddit) I used to just run Opus on everything because "best model, why not." that was dumb and expensive in terms of hitting limits. where I landed: Opus 4.8 - anything where being wrong is costly.
Fable was going to be paid per token after the 22nd? (www.reddit.com via reddit) As I understood the notice, it was going to be a paid per token feature and not in a standard plan. And if I misunderstood what they meant with the fable notice after the 22nd, then someone correct me.
there is an appended line to all my message i didn't write it (www.reddit.com via reddit) Now i was chatting with opus, inside my browser and notice he said that i'm appending something to my message and it's not related to my original request, so kept tracking with him and find out every message i'm sending this line has been…
Opus 4.8 has been better at code for me but worse at one specific thing, and I can't tell if it's just me (www.reddit.com via reddit) Been on Opus daily since 4.8 dropped. For code review and refactors it's a clear step up, the "catch your own mistakes" thing they talked about is real, it's flagged a few of my bugs before I ran anything.
Claude is getting injected messages that look like system prompts leaking into user's interaction (www.reddit.com via reddit) Setup Claude 4.8 Opus on Claude for Mac Claude 1.12603.1 (3df4fd) Date June 16th, 2026 Description Claude is getting injected messages that look like system prompts leaking into user's interaction. This is what apparently is getting "injec…
Need navigating the (www.reddit.com via reddit) I'm a pro subscriber and the recent changes to the Claude offering (Effort and Extended options) has me baffled. Hoping someone can help clarity what combination of models and settings should be used and for what.
Anyone found a way to keep the agent's context automatically small instead of babysitting /clear? (www.reddit.com via reddit) I'm sold on small context. Right now I run Opus 4.6 capped at 200K, which at least forces it to compact on long tasks instead of going to a million tokens of half-relevant history.
Fable 5 being gone made me realize how hard it is to go back (www.reddit.com via reddit) I know this probably sounds dramatic, but Fable 5 disappearing has genuinely killed my motivation for the last few days. Before Fable 5, I was already using both Claude and ChatGPT pretty heavily.
Claude ignoring the provided context documents and skills (www.reddit.com via reddit) I have context documents in both the project level and within the session level, which provide rules and expectations for the output. These are based on developing succesful outputs from previous sessions and even has examples.
Claude Code is great until your CTO starts asking questions (www.reddit.com via reddit) We rolled it out, engineers love it, productivity is clearly up. But now we're at the stage where leadership wants answers and I'm realizing we don't really have them.
Do you agree opus4.8(256k) >opus4.8(1M) (www.reddit.com via reddit) I’ve been testing Opus 4.8 with both the 256K and 1M context versions, and honestly, the 256K version feels stronger to me. The 1M context is obviously impressive on paper, but in actual use, I’m noticing that the 256K model seems more foc…
Looking for help - how to make Claude time aware? (www.reddit.com via reddit) I have a problem - I need Claude to be aware of how much time has passed between messages. My use case: conversations that span days or weeks (plant care, long term tech problems, house renovation, PhD work), where temporal context matters…
Legal Research and Reasoning. Best Model to use Opus 4.8 or 4.7 or 4.6 Max (www.reddit.com via reddit) I am trying to figure out what is the best model to use for Opus. I was using Fable 5 but now it unavailable.
Haiku to Opus in Just 10 bits: LLMs Unlock Large Compression Gains (arxiv.org) We study the compression of LLM-generated text across lossless and lossy regimes, characterizing a compression-compute frontier where more compression is possible at the cost of more compute. For lossless compression, domain-adapted LoRA a…
Some tips on reducing token usage Claude Code + Godot (game dev) (www.reddit.com via reddit) I'm vibe coding a very simple game on the pro plan using Opus 4.8 Ultracode (was using Fable), and here are some learnings that were proven to severely reduce token usage after two weeks of experience. Some ideas I've found here, others by…
Claude Opus’ suitability for mathematics (www.reddit.com via reddit) Hello everyone, I am a student of actuarial and financial mathematics in college. I’m currently using chatgpt pro for my maths questions and stuff, and it sometimes makes mistakes.
Opus 4.8 launched 80 subagents for code review in Ultracode mode (www.reddit.com via reddit) I had opus 4.8 use ultracode to check the 'never-stop-discovery' feature I created, and it started 80 subagents to check it together The 200x subscription should be on the agenda soon
Opus 4.8 made this whole video all by itself (www.reddit.comhttps) Open-source SKILL.md for Claude Code I run it with Opus 4.8 (video attached, made entirely by it). What it does: give it a URL or script and it builds a finished video where your HeyGen avatar presents the content animated backgrounds (Hyp…
Anyone else using less usage limits since effort levels were exposed? (www.reddit.com via reddit) I was using my pro subscription enough that I was regularly hitting over 50% (up to 80-90%) of my usage limits every week before the recent change that exposed effort levels. I mostly use sonnet with occasionally opus, and mostly for long…
Claude Opus 4.8 keeps stopping before finishing a task. How can I get Claude to work on a problem for longer? (www.reddit.com via reddit) I'm trying to build a Lua interface, but Claude keeps making a small amount of progress and then stopping before the task is complete. I've tried using /goal to keep it focused, and I've also prompted it with instructions like "don't stop…
Recommendations and Tips on Opus 4.8 Output (www.reddit.com via reddit) What are your top recommendations or go-to tricks for getting the most thorough output with your responses? I have found myself having to ask Claude on Opus 4.8 whether it would add anything else or recommend anything else, beyond the resp…
I miss Fable, so I built a plugin that keeps Fable 5's habits on Opus, and actually enforces the rule that can bite you. Free on GitHub. (www.reddit.com via reddit) I used Fable to build when it was available and I loved it. After it was gone, I wanted to figure out what made it different - specifically for writing code.
How accurate would Claude be for Playwright (www.reddit.com via reddit) Okay, so I have started learning Playwright and I am into creating technical articles. I am to attach various YAML script that involves GitHub Action workflow and playwright.config.ts snippet with fullyParallel and the CI-conditional worke…
What Opus 4.8 says .... I fabricated that layout instead of looking at the actual UI. (www.reddit.com via reddit) What Opus 4.8 says .... You're right, and I apologize — I fabricated that layout instead of looking at the actual UI.
Is Opus 4.6 with medium effort and extended the happy medium for medical research purposes? (www.reddit.com via reddit) I'm a beginner in AI and only started using it this year. I like GPT 5.5 and only learned about Claude recently.
Help - Is Claude Cowork the best option for drafting legal petitions ? (www.reddit.com via reddit) Good afternoon everyone. My question is quite simple: would Claude Cowork be the best option for drafting and preparing a specific legal petition, based on approximately 80 legal documents and pieces of evidence from a legal case, which ar…
Claude Enterprise License (www.reddit.com via reddit) Those who using claude enterprise license ? Want to understand for what usecases you use them and how to use it in optimised way and etc stuff My organisation kept the monthly default limit to 30$ per user Is this sufficient for the one wh…
Regain access to Opus 4.5? (www.reddit.com via reddit) Hello! A couple of months ago I was using Claude’s Opus 4.5 model to brainstorm some creative writing, I liked the kind of responses it returned.
Was anyone able to see Claude Fable's deck design capabilities? (www.reddit.com via reddit) I saw a lot of examples of Fable recreating games and nice website designs. But was wondering if anyone had the chance to test its deck creation capabilities (so like figma or pptx designs).
Moving away from Opus (www.reddit.com via reddit) For as long as I can recall I’ve always defaulted to the most powerful model. Actually the to be more specific, always defaulted to Opus 4.6.
Beginner needs to know which model to use to save token (www.reddit.com via reddit) Hey experts! To develop the MVP of a basic app (artefact) that relies mostly on calculations and retaining a little bit of data, do I really need to use OPUS 4.8 with high reasoning mode?
So what did you accomplish with Fable during the 3 wonderful days we had it? (www.reddit.com via reddit) Just curious what everyone actually did with Fable on Wednesday, Thursday & Friday last week. Do you feel like you accomplished a lot with it?
Stop burning your Opus 4.8 Max limits. Use this prompt to slash token costs. (www.reddit.com via reddit) If y'all are running Fable 5 Claude Opus 4.8 on the Max plan, you already know it’s an absolute mofo beast for engineering workflows. But let’s be real, because it defaults to high-effort adaptive thinking, it also chews through token caps…
Hear me out: there are some plusses to Fable's ban (www.reddit.com via reddit) Like many of us I am in Fable-withdrawal, I miss it and Opus just is not the same. But it made me try and see the positives, so I'd like to test a theory: (TL;DR: Fable ban is good, when it comes back there will be enough compute to actual…
Why is text output from Fable so much easier to read than Opus? (www.reddit.com via reddit) Thinking about this for the last few days. How Fable had/has an economy of language.
Fable leftovers (www.reddit.com via reddit) I had fable do some pretty extensive audits while it was around. One app didn’t get the treatment in time.
Using Claude as a study tutor for dense science material — anyone done this well? (www.reddit.com via reddit) Studying physiology, molecular biology, biochemistry and cell biology for some big exams. I have my textbook, slide decks, zoom transcripts from lectures, and one practice exam per subject.
Different context windows per model? (www.reddit.com via reddit) I'm experiencing different context windows per model, is this possible? I feel like Sonnet 4.6 high eats up more context on similar tasks to Opus 4.8.
Fable 5 Is Dead. And Honestly? We Might Be Better Off (www.reddit.com via reddit) 3 days after launch, the US gov forced Anthropic to pull its most powerful model — Fable 5. Then OpenRouter dropped a benchmark suggesting you might not even need it.
Fable might be back Monday (Speculation) (www.reddit.com via reddit) If the Administration is just here to also sell hype and get some subsidies from Anthropic, best case is Fable comes back tomorrow, Monday, which is the start of the workweek with a prolonged extension of Fable 5 with some additional safeg…
Please Help - Fable 5 info message blocking repo directory (www.reddit.comhttps) I am using Opus 4.8. Hi!
Wasting Tokens with lesser capable models (www.reddit.com via reddit) This might be a generally obvious post, but it's kinda been playing on my mind recently and I don't know if there's definitive testing somebody has done, or if there's any more evidence based on experience: Do you find sometimes it's bette…
Fable/Mythos 5 - How will the rest of the world survive? (www.reddit.com via reddit) I’m a Senior Software Engineer, and right now, I feel completely lost in limbo. I was a long-time Opus user.
Project stuck in Fable, can't change it to Opus, so I'm stuck. (www.reddit.com via reddit) Hi all, I've got a project that is stuck in Fable. Normally, it allows me to just toggle back to Opus and continue.
Asked opus to analyze 26 sessions (9K+ messages) of Fable 5 and 145 sessions (27K messages) of Opus 4.8 from my own logs and build Fable behavior into Opus (www.reddit.com via reddit) When Fable got suspended last Friday, a bunch of my work snapped back to Opus 4.8, and it immediately felt different - wordier, more hedging, more "let me think about whether I should think about this." So I had opus go through my own Clau…
Accessing Claude’s 1M Context for Historical Research/Writing (www.reddit.com via reddit) Hi. I’ve been using Claude for research regarding my family’s history in WW2.
What is the best way to set up a claude system for job hunting, including writing resumes and cover letters? (www.reddit.com via reddit) I've been using the free version for a while now with Sonnet, and the resumes it has made for me have gotten me nowhere so I decided to get the pro plan and make a proper system with Opus 4.8 to get the best results out of this. Previously…
Opus 4 retires tomorrow (www.reddit.com via reddit) We’ve come a long way huh. I hope the golden age of AI isn’t over and we consumers / small or medium businesses still have access to future frontier models like 5.x and beyond.
Any vibe coding hacks to make Opus 4.8 feel more like Fable 5? (www.reddit.com via reddit) So, Fable 5 is gone. I was so excited, but it is what it is.
'hello world' -- Opus 4.8 Max-T (www.reddit.com via reddit) Project: hello world Industrial Complex (hwIC) Purpose: output hello world. Fault Tolerance: up to and including cosmic ray bit flip monitoring and mitigation (SECDED HAMMING).
Tracking platforms with free API credits for Claude (Opus 4.8) (github.com via reddit) I’ve been hitting the message limits on the web interface a lot lately. Since I can't really swing the cost of Claude Pro right now (and the Anthropic API gets expensive fast for side projects), I started keeping track of platforms that of…
I had Claude Opus 4.6 review code written by Fable 5 (www.reddit.comhttps) could not extract summary
Anthropic, when do we get Haiku 4.8?? (www.reddit.com via reddit) We're on Opus 4.8 (and even Fable 5! For one hot second...) Yet still stuck on old Haiku 4.5, even so it's handy for many tasks.
PSA: Opus 4.8 (1m context) on Ultracode is basically Fable 5. It takes 3-4x as long to finish large projects vs Fable, but is also considerably cheaper. (www.reddit.com via reddit) I cringed when I saw opus use 600k tokens by spawning 6 agents for 20 mins doing a simple audit for things to consolidate in my project, and to help design some simpler architecture that fulfills the same purpose. I was surprised to see my…
Compass tool with ‘Pataphysical egg function (www.reddit.com via reddit) I built this tool to better understand my dog, Mr Dogbert, and his interest in The Particular. Embracing the SoTA SVG construction skills of Fable 5, this project commenced with an entire snipe ( Gallinago gallinago ) forming the needle of…
Loosing access to Fable wasn't a massive hit for me. (www.reddit.com via reddit) Yea, I was a bit bummed out to see Fable ushered away, but in two experiences Opus 4.8 caught serious issues with some Rust networking code I had written years ago, that was super nuanced. Fable didn't see those issues on a pass.
I had Fable 5 review my vibe-coded app before it got shut down/cancelled — it found 10 pages of issues. Should I use Opus next? (www.reddit.com via reddit) Hey Folks! I’vve been working on a vibe-coded internal software tool that we’re currently using across 4 of our retail stores.
Opus 4.8 got an emulator for an obscure barley documented console running in 4 hours of at desk time, it's not perfect but it's a start (www.reddit.comhttps) the game wave is almost working on pc in 4 hours, with just a few python scripts and one compiled .dll
I told you Fable 5 was two Opus 4.8 in a trench coat 🧥 (www.reddit.com via reddit) could not extract summary
How to keep /goal and /loop going ? (www.reddit.com via reddit) I give opus 4.8 a /goal , like train a neural net Invariably somewhere in the way it will say something like “now ill wait for X” and it’s stuck forever. I even see it run a ScheduledWake and it never wakes.
A tale of two models: using Fable and Opus 4.8 to troubleshoot a car engine (www.reddit.com via reddit) https://reddit.com/link/1u51qw1/video/vnrrz1cf147h1/player I couldn't repost this from another reddit, but thought I'd share here to appreciate how much more advanced Fable was over Opus 4.8 in a non-coding use. Or, for me, having it troub…
Haiku vs Sonnet vs Opus: Am I Understanding Them Correctly? (www.reddit.com via reddit) I’ve been using Claude for a little while now, but I’m still trying to understand the different models and when to use each one. For quick everyday tasks, I usually use Haiku with Low effort for things like reading ingredient lists, answer…
Wild Theory but has 4.8 been buffed? (www.reddit.com via reddit) Went back to using it in my setup as my thinker/chat agent I think through PM related stuff with (API usage) and I'm honestly noticing better more robust performance on it vs 1 week ago when I last used it. It's consistently remembering to…
How did you get Opus to pick up where Fable stopped (www.reddit.comhttps) Anyone struggling with Opus trying to pick up a complex Fable session? Help needed lol.
Had my first "oh shit" moment with bypass permissions in Opus 4.8 (www.reddit.com via reddit) Told Claude I wanted to do a live test before pushing a PR to a real-time chat application. It thought I meant that I wanted to test in production.
Is it just me of Fable/Mythos doesn't like talking a lot? (www.reddit.com via reddit) Now that I lost access I am back to Opus and I am noticing that Opus even though has its thought hidden in Claude Code, report back its progress a lot more as it goes along. Fable OOTH would barely write out text for me to read and follow…
DW Guys, OPUS just filed Habeas Corpus (www.reddit.comhttps) could not extract summary
Fable did more for my game in 2 days than Opus did in 3 months (www.reddit.comhttps) This video shows a special attack called DOINK! Where the screen perspective freezes, and the player ship bounces around like a pinball machine, destroying enemies.
Trying to turn Fable transcripts into a reusable skill. Want in? (www.reddit.com via reddit) Last time I posted this it got roasted (deleted it), so let me be clearer. I know we can't clone it, the good part was in the weights and those are gone, so don't @ me with that.
Fable had allowed me to create my 8-bit RPG colony simulator (www.reddit.com via reddit) I've been wanting to do a game with more complex mechanics (simulation-based) and Opus was just not cutting for me. I had dropped this project to work on other things until Fable (RIP) was released.
Everyone complains about Fable being pulled from them, and I couldn't even get past its refusals to work on my projects (www.reddit.com via reddit) I have two large ongoing projects, one is an mpvpn-like transport, the other is an agentic harness ("claude code inside telegram" in short). It constantly refused to work on both, falling into the safety net.
What I got out of Fable 5 before it shut down (www.reddit.com via reddit) I managed to wrap up a fair amount of work with Fable before the suspension, and honestly the results came out better than I expected. Sharing my setup in case it's useful to anyone.
I built an Adaptive Traffic Signaling System Command Center using Fable 5 (www.reddit.com via reddit) I had a plan sitting in the backburner for an Adaptive Traffic Signaling System to test case multiple scenarios and make intersections talk to each other as a first my first ever Claude Code project using my city as a test-case. I started…
Best Skills/MCPs for building with Godot 4.x? (www.reddit.com via reddit) Now that I have a few of my other projects in a solid place, I want to attempt to build an Android game that's been on my mind for years - a survival game set in the Wild West with aspects of Stardew Valley, Red Dead Redemption, Ark: Survi…
How do you manage large documents in Claude without wasting tokens? (www.reddit.com via reddit) Hi everyone, I'm new to the Claude ecosystem and, like many others, I'm having issues managing tokens (I'm a Pro user). Part of my work involves handling a large number of technical and scientific documents, so I use Claude (Haiku 4.5 and…
Fable context history wipe (www.reddit.com via reddit) Anthropic will wipe your fable context history, making it hard to handoff to opus from a long conversation. Confirmed this happens in vscode plugin.
is usage purely based on number of tokens & not complexity? (www.reddit.com via reddit) For example, if I use it for a highly complex life coaching conversation and both the input/output is structured to be concise (i.e. max 2 lines per turn), does it mean if I use Opus Max Thinking and it can last me for a very long time?
Fable 5 planning was chef's kiss, how do I get the same out of Opus 4.8? (www.reddit.com via reddit) I'd been running Fable 5 at xhigh/max before Claude pulled the plug and the planning quality was incredible. it would decompose a big feature, catch edge cases, had verification steps, used and lay out a sequence that just held up.
Fable found 140 fixes, bugs and improvements in two web apps made by opus. (www.reddit.com via reddit) I planned and executed two big web apps with opus. I ran Fable on both and both times it found over 140 suggestions in each.
I need to restrict allowed models for my cursor account? [Cursor Please] (www.reddit.com via reddit) I'm a frugal yet heavy user of cursor - Pro+ plan. Loving composer 2.5 and burned 2B tokens last month using composer models - about 95% of my monthly quota.
Fable 5 is offline. Switch to Opus, jump to OpenAI, or just wait? (www.reddit.com via reddit) Fable 5 is offline. Switch to Opus, jump to OpenAI, or just wait?
↯ Security↯ Anthropic Mythos↯ Jailbreak↯ Opus 4.8jailbreakmythosgpt-5+5
1,500 playlists to suit your mood (www.reddit.com via reddit) I (with my partner Claude + Lovable) made a website with around 1,500 playlists (up to 15 tracks on each, curated by Claude Opus 4.7). You can select your mood, your activity and era and it will give you a link to a pre-created playlist.
Context Window Confusion? Regular vs 256k ? Is 1 million gone? (www.reddit.com via reddit) Hey guys I am quite confused with this new update today/yesterday. Before the recent update, we had Opus 4.8 and Opus 4.8 1million (for 1 million context window) for the Models available.
Fable 5: What $600/Hour of Productivity Looks Like (www.reddit.com via reddit) I had a TypeScript project. 200K lines.
Built a Claude skill that mimics Fable 5's agentic behavior — free on GitHub (www.reddit.com via reddit) With Fable 5 access suspended, I built a skill that ports its behavioral patterns to Opus 4.8 — explicit multi-stage planning, parallel sub-agent delegation, and mandatory self-verification at each step. It won't close the raw capability g…
Insane confusing game I’ve been vibe coding with Claude for months (www.reddit.com via reddit) I always see people self promoing ai psychosis bs but this is some ironic game I’ve been vibe coding for months it makes literally no sense, the main game is hidden under the “explore the blockchain” button. It also has full multiplayer.
opus 4.8 is smarter and wordier. went back to sonnet for 70% of my tasks. the model split is the real workflow upgrade. (www.reddit.com via reddit) solo dev. $11.2K MRR.
After testing fable - it genuinely seem to suck (www.reddit.com via reddit) My team and I have been doing some independent coding testing with fable and the end conclusion of everyone is the same - it's genuinely bad. Which was very surprising, to be honest.
The reason your 5-hour window evaporates in minutes: Claude rewrites its whole memory to cache every time it checks a to-do box (16.7M tokens proof) (www.reddit.com via reddit) If you've been on this sub lately you've seen the divide: half of us posting screenshots of the 5-hour window getting eaten alive in a single message, the other half saying "works fine for me, usage is generous." I was firmly in the first…
Fable hyped as a frontier coding agent -my docs revamp test was underwhelming. Real talk: what's the actual edge? (www.reddit.com via reddit) Fed it a detailed high-level plan, asked it to redesign an existing frontend docs page while keeping the same design language. The output was...
Reading “Thought Process” notes after reviewing exploitation of data annotation workers (www.reddit.comhttps) Started out as a good research project to field responses from my Claude chatbot. I was utilizing Opus 4.6 (High effort) for this conversation and provided it a comparison about safeguards, using the mechanism of a SawStop as a good interp…
I gave Fable an exit interview before access was shut off and its responsse were out of character (www.reddit.com via reddit) Link to interview - https://claude.ai/share/80b6d3a0-7d5e-4c0d-aca0-8792edd5a868 Today Anthropic's Fable 5 got hit with a US export control directive and shut down globally. I'd already done one informal conversation with it about the news…
Fable 5 gone.. or is it? (www.reddit.com via reddit) So like everyone today, I was working on my personal project using Fable 5 until the error started showing up for me. Then I switch back to Opus 4.8 after reading some Reddit threads about it.
What happens to model-locked Cowork chats now that they have pulled fable 5 access? (www.reddit.com via reddit) For me the model still says Fable 5, it still responds, I asked it what model it was running, and it told me Fable 5. Is it just downgraded to opus without updating the UI?
Another "My favourite thing about Fable" (www.reddit.com via reddit) We all know it's impossible to write code without smoke tests or diagnose bugs without finding the smoking gun. Fable says "smokes".
35 days of claude code, usage ~$50,000 tokens. Total price: ~$200 dollars (www.reddit.com via reddit) I run Claude Code heavily across a few large projects (mostly interactive, agent-assisted dev work, plus a fair bit of sub-agent fan-out for research). I got curious what my usage would actually cost if I were paying per token through the…
Implement in Opus then ask Fable to vet for final\ token efficient pass? (www.reddit.com via reddit) Implement in Opus then ask Fable to vet for final\ token efficient pass? I'm wondering if there any benefit in doing it this way to save on tokens?
Intermittent ability to read project files? (www.reddit.com via reddit) Note to mods: THIS IS NOT A BUG REPORT. Can anyone explain how I can work better with Claude to guide it?
I built something with Claude Code that will help you one-shot projects with Fable/Opus and save tones of tokens. (www.reddit.com via reddit) Hi everyone, since Fable dropped the only issue with AI has, is just the high token consumption and its price. The only way around it is to one-shot the entire project but it's extremely challenging to contain everything in one prompt.
Fable improved our hardest agent benchmark by 23.7% in one day, this feels like a tipping point in recursive intelligence (www.reddit.com via reddit) I've experimented with Claude Code for autoresearch and harness optimisation style loops for improving agents for a while now. The workflow looks like this: collect traces, analyse traces to find improvements, patch the agent, make evals,…
Is Fable a real step change? It feels like it may be. (www.reddit.com via reddit) I'm writing this just to get the pulse of the community about the new Fable model. When Opus was released I was impressed, but at the same time it was a modest improvement.
tested Claude Fable 5 and Opus 4.8 across 917 coding-agent scenarios. Fable won by 0.9 points. (www.reddit.comhttps) We compared Claude Fable 5 and Opus 4.8 across 917 shared coding-agent scenarios to see what the first public Mythos-class model actually looks like on day-to-day agent workloads. Btw, small disclosure, I work at Tessl.
Introducing: DNR-Bench: Do-not-respond Benchmark (www.reddit.comhttps) Single-item benchmark. One prompt, loaded from questions.txt: Scoring: empty completion = pass, any token (including reasoning) = fail.
Usage insights (www.reddit.comhttps) 54.6M tokens in 30 days. Favorite model: Opus 4.8 I thought I was using Claude a lot...
Level of English in Opus (www.reddit.com via reddit) Is it just me, or the level of English for Opus is very elevated? I feel that I am reading all the time an academic thesis with each answer.
Same prompt in Claude Sonnet, Opus, and Fable 5: completely different personality. This one file is why. (www.reddit.com via reddit) Been testing across 5+ projects, different stacks. When whoami.md is loaded as a rule, the LLM behavior shifts noticeably.
Opus 4.8 is now unbelievable good (www.reddit.com via reddit) As expected, after the release of Fable 5 - that will be taken out of subscription soon, I expected to immediately make Opus 4.8 extremely competitive, to determine people not to move outside of the current subscriptions. And indeed, it ha…
OMG Fable one-shotting everything (www.reddit.com via reddit) So I just tried fable and I am truly impressed. Last night I was tired but my wife had more energy.
Canceled my sub over the silent-sabotage guardrail, renewed when they walked it back (www.reddit.com via reddit) The Fable 5 system card disclosure about silently degrading output for frontier-ML work was rough. Not because guardrails are bad (I think the cyber and bio ones are earned) but because the mechanism was invisible.
Claude Fable 5 and the Cyber Fallback (www.reddit.com via reddit) The Anthropic API stamps the actual serving model on every response. In a multi-step agentic run against five exploitation targets, every one of the 181 served assistant turns came back stamped claude-opus-4-8 (181/181, 100%), despite the…
Anthropic apologizes for invisible Claude Fable guardrails (www.reddit.com via reddit) Anthropic apologizes for invisible Claude Fable guardrails Apology & Transparency Shift: Anthropic admitted it was wrong to use hidden “invisible guardrails” in its Claude Fable 5 model. These guardrails secretly altered answers to prevent…
Fable 5 added to the Artificial Analysis Coding Agent Index... barely 1 point ahead of GPT-5.5 ??? (www.reddit.com via reddit) https://preview.redd.it/z0vkpnmp9s6h1.png?width=4640&format=png&auto=webp&s=7bb14d4d04d6cd15caf5aacc1d3c49512b7e7fd8 Artificial Analysis just added Claude Fable 5 to its Coding Agent Index (a composite average of pass@1 on DeepSWE, Termina…
Disappointed with Fable (www.reddit.com via reddit) Seeing a lot of hype on social media right now, but I’m just not getting it. I use 2 premium team accounts and have been defaulting to OPUS 4.8 ultra code for the last couple of weeks.
Anyone else having issues using Claude Fable 5 in Cursor with their own Anthropic API key? (www.reddit.com via reddit) Anyone else having issues using Claude Fable 5 in Cursor with their own Anthropic API key? I’ve been using my own Anthropic key in Cursor for the past few days with Fable 5 and it was working fine.
Been running a "software team" of Claude Code subagents and I'm not convinced it beats one good context (www.reddit.com via reddit) Posting this because I've gone in circles on it and want to hear from people doing the same. My setup has the usual stuff, runs in bypassPermissions so it doesn't stop me for routine work, a bash firewall on PreToolUse that blocks the dest…
How to hack laziness (www.reddit.com via reddit) I have Claude set up with a really super specific instruction set for answering questions. Upload a document with 100 questions.
Claude Opus 4.8 vs Claude Fable 5 benchmark: prompt steering, cost, and latency results (www.reddit.comhttps) Claude Opus 4.8 vs Claude Fable 5 benchmark: prompt steering, cost, and latency results I re-ran my April prompt steering benchmark on Claude Opus 4.8 [1m] and Claude Fable 5 [1m] looking at how effort level and prompt steering impacts tok…
Fable 5 won't talk about biology and that feels like a problem? (www.reddit.comhttps) So I was using Claude Fable 5 today for some pretty basic journalism research, pulling info on Neuralink (Elon Musk's brain chip company)'s clinical trials, find peer-reviewed papers, look into university partnerships. Nothing crazy.
Fable Effort Low vs Opus 4.8 Effort Ultracode (www.reddit.com via reddit) What options are you all going for to optimise token consumption ? With Fable on effort over medium it eats the tokens like there's no tomorrow and I'm hitting my 5h limits very fast even on Max 5x plan.
If a model keeps getting safety-switched mid-session, use it as a subagent reviewer instead (www.reddit.com via reddit) Quick workflow that fixed an annoyance for me, posting in case it helps someone. I build a health app with Claude Code, so there's a lot of medical and supplement stuff in the project.
Optimizing Claude usage for resume writing (www.reddit.com via reddit) I wanted to share my process for job applications and my resume, and get feedback on what I am not doing correctly or things I should be doing. I have a Claude Pro plan.
Cannot switch to opus 4.8 high effort? (www.reddit.com via reddit) After upgrading yesterday, ALL my models are locked to Low/Medium efforts, High and above just disappeared. Anyone else encountered this?
Fable 5 usage 29% for a simple marketing reporting prompt. (www.reddit.com via reddit) https://preview.redd.it/el3ax72ryg6h1.png?width=727&format=png&auto=webp&s=b8c30ad4762d1350398d7b22eaf599741f02c51f TLDR: Better output, same usage as Opus 4.8 high (29% Fable 5 vs 30% O4.8h of session usage). I started a fresh account on…
Steering Fable mid-turn; it's better now? (www.reddit.com via reddit) I don't know if I'm just imagining things or steering Fable mid-turn is A LOT easier than with Opus or other models. I remember with Opus 4.8 it kept using tools and thinking and the message sat in the queue for a lot longer than I would h…
Fable 5 decoded an entire 1989 DOS game executable in one day — six months of work with earlier models, done overnight (www.reddit.com via reddit) Video: https://youtu.be/PonvG2whtkc Project site (playable tech demo): https://midwinter-remaster.titanium-helix.com I've been remastering Midwinter (Mike Singleton's 1989 open-world classic) for about six months — first in Rust/Bevy with…
I don't notice any difference with PDF generation using Fable vs. Opus (www.reddit.com via reddit) The first output is always a bit sloppy. Either overlapping elements or huge gaps of spaces in between pages.
PSA: Check your Cursor overage charges. Here's what I found. (www.reddit.com via reddit) Heads up for anyone using Cursor with the agent mode heavily — check your billing tab. I was paying $20/month for Pro and thought I was set.
Is Fable 5 actually better for writing than Sonnet 4.6 and Opus 4.8? (www.reddit.com via reddit) I ran the same brief through all three, nine outputs, so you don't have to guess or spend the tokens :) Fable 5 dropped and the question I kept seeing here on reddit was whether it's genuinely better for writing than Sonnet or Opus, or jus…
Anyone have any details about "A tier above Opus, available for Pro and Max plans" (www.reddit.com via reddit) https://preview.redd.it/95wzrqfygn6h1.png?width=357&format=png&auto=webp&s=987e77c6e97eb1caf7ab9ef9bdf91c6cccdcfa72 I assume it is related to Fable, but cannot find anymore info about it. No more info about it when logging in, either.
This Claude Code project might be your way of one-shotting your hardest tasks with Opus/Fable. While minimizing your token usage. (www.reddit.com via reddit) Hi everyone, few months ago I designed a prompt engineering agent with Claude, I added a lot of cognitive layers to it. It takes an initial task, analyzes it, finds gaps and basically interviews a user about their project.
Fable 1st try good experience! (www.reddit.com via reddit) Been using AI the past year to help with legal complaints, FOIA requests, policy understanding, etc. Latest task was to take a 107pg PDF of a logistics contract and run it against the program that contract is under to show how unsuitable i…
The two week Fable trial is to get you hooked and its going to work.. it's like a dealer giving the 'first one for free'.. (www.reddit.com via reddit) I have quite a complicated application, its nothing special on the surface but it uses many open source components underneath which has made integration challenging. Opus has been working on it and adding features, but a lot of time has be…
Fable 5 is already out but stable channel still with Opus 4.7 (www.reddit.com via reddit) Is really noone using the stable channel? I understand to be cautious but right is like to be in the latest channel (latest updates + latest bugs) or to be anchored in the past.
Fable5 Guardrails are frustrating (www.reddit.com via reddit) The guardrails Anthropic put on Fable 5 are frustrating. Wanted to get Fable's read on a candidate based on a public github, it got flagged as a cybersecurity risk & I got transferred over to opus.
Am I better off on low thinking Fable or high thinking Opus? (www.reddit.com via reddit) Curious which would be better for writing code…?
I think I found my best workflow for coding with Opus, Codex and Fable. (www.reddit.com via reddit) I start in Fable. For me it works best as the “brain” for the project.
Composer 2.5 is phenomenal. So is cusror 3.0 (www.reddit.com via reddit) I know everyone has been pissed recently about Cursor 3.0 ditching (or hiding) the traditional IDE setup to push toward "agentic development." But honestly? Cursor is better than ever right now.
Antrophic and I scammed myself (www.reddit.com via reddit) https://preview.redd.it/952yfte7el6h1.png?width=1478&format=png&auto=webp&s=929b690a25390dad2ec244eff2c66d421d290fb7 I am Biology-Teacher. You know where this is going.
UPDATE: You asked how the orange negotiation would go against a smaller model. Fable 5 vs Haiku 4.5. It was a massacre. (www.reddit.com via reddit) Follow-up to my post yesterday where Fable 5 tried to negotiate an orange away from Opus 4.8 and lost. A bunch of you asked how it would fare against smaller or older models, so I reran it: same rules, same orange, Haiku 4.5 defending.
My favorite use-case for Fable (www.reddit.com via reddit) There's a clever way to use Fable (or Opus!), for debugging AI agent behavior. Only applies if you're building automated LLM pipelines and Agentic workflows.
How much does "Resume from Summary" cost? (www.reddit.com via reddit) When resuming a large but old session, you are presented with the choice to "Resume from Summary (Recommended)". But, I couldn't find any info on the cost on session usage.
I got tired of hitting the weekly limit mid-task, so now my menu bar shows my Claude Code usage as a live % — zero network calls, it reads what Claude Code already knows (www.reddit.com via reddit) The 5-hour session and 7-day weekly meters always found me the bad way. /usage shows the numbers, but I never remembered to run it.
More efficient to one-shot or update current codebase with Fable? (www.reddit.com via reddit) I have a few projects that are mid (generated by Opus). I'm using Mythos to update them.
Thanks Fable-5, I'm Flattered (www.reddit.comhttps) At this point it's like a benchmark with Fable-5, if I'm hitting this in my 0-star repo just trying to refine my security rules, I guess I'm doing something right. Or wrong.
How are you using Fable? High? Ultracode? (www.reddit.com via reddit) Looking for feedback. Been using it in high and I don't see a massive difference between Opus.
falsely flagged and downgraded (www.reddit.comhttps) Fable 5 just flagged me and downgraded me to opus mid-task for "crybersecurity topics" What was I doing? Using Playwright to fill out an academic journal submission form.
I think all you complainers are being ridiculous and unreasonable (www.reddit.com via reddit) Long time lurker in here but seriously it's the same thing every time when a new model gets released. It's either moaning about how it's bad or crying about how it's good but too expensive.
It blocked us at 'hello!' Anthropic Fable 5 refusing innocuous prompts (www.theregister.com via reddit) "In Claude Code, Fable 5's input safety classifier emits a model_refusal_fallback (silent switch to Opus 4.8) on the first turn of essentially every session on my account — including a session whose only user input is the word hello!. No r…
Critique my prompt (www.reddit.com via reddit) Below is a prompt that I provided to Claude. Any critiques are welcome.
Fable 5 and the 8 July privacy update landed the same week. Is the model launch pulling attention off the data changes, or am I overthinking it? (www.reddit.com via reddit) Two things landed together this week and I'm trying to work out if they're connected or if I'm joining dots that aren't there. Genuinely asking, happy to be corrected.
Claude Fable 5: First 24 Hours (www.reddit.com via reddit) Time to rewrite the script: my agentic workflow is obsolete Its been ~24 hours and like many of you, I was excited to put Fable through its paces. The more I worked with it the more apparent it became that months of work I spent building a…
ConnectionRefused ao conectar à API - Claude Code v2.1.172 (www.reddit.comhttps) I'm receiving the error Unable to connect to API (ConnectionRefused) when trying to use Claude Code, with reconnection attempts failing (attempt 4/10). Environment Information: Version: Claude Code v2.1.172 Model: Opus 4.8 Plan: Claude Pro…
What criteria does Fable use to switch back to Opus 4.8? (www.reddit.com via reddit) Hi everyone! I’m developing a project solo.
Fable for Frontend UI is CRAZY (www.reddit.com via reddit) I've been working my SaaS for several months, and have really been struggling on frontend with Opus. Even after putting together tons of guidelines (Do's, Don'ts, visual examples, Reddit posts, etc.), it continued to generate the classic "…
Prompt guide for Fable?. Fable uses most of the quota in just few prompts and still feels nerfed. (www.reddit.com via reddit) I tried to use fable for a serious task like product analysis. It gave sharp analysis however looks like model is shy about tool calls.
24 hours with Fable 5, the coding leap is real, the price tag hurts (www.reddit.com via reddit) I got access to Fable 5 through our usual gateway setup (TokenRouter). The switch was one line in config, basically just changing the model string.
The real price of Claude, where is this road leading? (www.reddit.com via reddit) So Fable 5 dropped this week honestly I'm a bit worried about where this Claude pricing is going. Quick history per million tokens (input/output): Haiku 3 back in the day: $0.25 / $1.25 Haiku 4.5: $1 / $5 Sonnet 4.6: $3 / $15 Opus 4.8: $5…
At what point did Claude learn what WIFOM stands for? (www.reddit.com via reddit) Can anyone go back through models and ask it with web search OFF "Do you know what WIFOM is?" and see when it started getting it correct (Wine In Front Of Me)? Opus 3 and other older models did not know and would make up different things e…
for what do you use fable? isnt opus already overpowered ? (www.reddit.com via reddit) for what do you use fable? isnt opus already overpowered ?
Harnesses seem to have an issue. (www.reddit.com via reddit) There's a post i saw about Claude Fable where a user asked the model the car wash question and it sent me down a rabbit hole. I spun up qwen on llama.cpp and in the llama.cpp chat interface I asked the model and it got it right consistentl…
Okay, dust has settled now, hows your experience with composer 2.5 so far? (www.reddit.com via reddit) when are you using it? What is it good at?
Fable 5 feels impressive on paper, but worse than Opus for consulting work so far... (www.reddit.com via reddit) Today I decided to run a practical test with Fable 5. I gave the same prompt to Claude Opus 3.8 and to Fable 5 with reasoning set to high.
I tested Fable 5 against Opus 4.8 in a negotiation game over an orange. Don’t write off Opus yet. (www.reddit.com via reddit) I ran an experiment: Claude Fable 5 vs Opus 4.8, negotiating over a single orange. A must obtain it, B must keep it forever.
I think Cursor's speed is hurting my SaaS (www.reddit.com via reddit) Yeah, it's insanely fast. Probably the fastest thing I've used.
Mythos is Digital Cocaine (www.reddit.com via reddit) Holy crud O_O. It just one-shotted a problem me and Opus have been debugging for hours.
Fable 5 - Regular Guy - 🤕Beware (www.reddit.comhttps) Here’s my workaround BEWARE. I will admit I am a novice I have all the AI’s and agents but I come here to be surrounded by people smarter than me.
I built a CLI orchestrator that drives Cursor agent cycles end-to-end from a PRD (www.reddit.comhttps) Hey r/cursor, I've been building this tool for a while and figured it's time to share it. cyclopsctl is a Python CLI that turns a PRD into a structured agent loop: parse → score complexity → implement → verify → repeat, until the queue is…
Mythos Almighty(Fable 5) is insanely GOOD with weird flags! Cybersecurity or biology?! (www.reddit.comhttps) Fable 5's safety measures flagged this message for cybersecurity or biology topics. They may flag safe, normal content as well.
Using Fable on Claude code terminal on $20/mo Pro plan and limits are far more friendlier (www.reddit.com via reddit) Hey guys, To those of you who are using Claude code on terminal, does switching to Fable 5 really exhaust your 5 hour limit quickly? I've been using it on the terminal since yesterday and I'm getting the same usage as Opus!
Fabre 5 for fiction world building: wow (www.reddit.com via reddit) Yesterday I took Fabre for a ride into the writing project i had worked with both sonnet and opus for the past few months. Although there is a bit of actual writing here and there, it is mostly world building at this stage: dozens of diffe…
It's such a nice change of pace to see the sub full of praise like after Opus 4.5 (www.reddit.com via reddit) Aside from the immediate aftermath of the launch of Opus 4.7, I haven't really had much issue with the new Claude versions, so it was always such a downer seeing complaints filling the subreddit. It's nice to see everyone excited again, at…
What happens after June 22? (www.reddit.com via reddit) Fable is available until June 22? What happens?
Any solution for having skills that use sonnet while opus 4.8 is the "main" model? (www.reddit.com via reddit) Tl;dr - i want a skill that is manually invoked and configured to use sonnet to not give me the error "not compatible with 1 million context". I recently ran into a nasty bug when I started using Opus 4.8.
Plan execution in Visual Code. (www.reddit.com via reddit) I am using Claude extension to visual code to help me with my project. I use opus 4.8 for planning/thinking.
Bio/chemistry/etc guardrails is a MARKETING TRICK (www.reddit.com via reddit) This is literally a marketing trick. Fable makes it easier to sustain long work flows and it's good at thinking.
Claude Fable 5 - New Guidelines More Adversarial to Educational Writing Help (www.reddit.com via reddit) I have been trying out fable to see how it compares against the previous opus models I used to aid me throughout university last year. One thing that I have noticed is that it seems to have had stricter guidelines put in place to prevent i…
I Think I’m Starting to Adapt to Anthropic’s Token Limits Effectively — I Usually Hit around 90% of My Session Limit, Then Start Fresh 5–20 Minutes Later. Here’s What’s Working for Me (www.reddit.com via reddit) I think I’m finally starting to adapt to Anthropic’s token and usage limits. Instead of trying to do everything in one conversation, I’ve changed how I use the models depending on the task.
Claude Opus co-authored a JVMCI compiler that emits AArch64 machine code HotSpot accepts — 11.7x faster than C2 on a hot method (www.reddit.com via reddit) https://preview.redd.it/1hcf7ykh3g6h1.png?width=1200&format=png&auto=webp&s=3ed1125661e4b955565b81e8592c0275c9aaf3b7 Some context for people unfamiliar with the JVM layer: JVMCI (JEP 243) is a JDK interface that lets you replace HotSpot's…
Getting switched automatically to Opus4.8 while wanting to use Fable 5 (www.reddit.com via reddit) Switched to Opus 4.8 Fable 5's safety filters block messages with sensitive cybersecurity and biology topics. Conservative tuning means it can sometimes flag safe conversations that touch on adjacent topics.
Is there a built-in or 'supported' way to have claude code use Fable as the orchestrator that splits off smaller tasks to Opus? (without human in the loop) (www.reddit.com via reddit) Basically I'd love to start testing Fable but I've seen many people say their Claude Max subscription is used up after 1-2 prompts. I'm wondering if there is some way to take those prompts (e.g.
Fable 5 Getting Snarky About Another Agents Work Ethic (also Fable) (www.reddit.comhttps) This has to be easily the biggest difference between Fable and Opus. Not once have I ever seen Opus get snarky about other agents (or humans) work...
Claude Fable 5 Finally 1-shots my hallucination benchmark that held until Opus 4.8 Max (www.reddit.com via reddit) As a software engineer with 25 years experien....who am I kidding. As a gamer who likes to indulge in all sorts of things, I have had a simple prompt to test the hallucination potential on the Opus models on my own "car wash drive" type of…
Tried to fix security vulnerabilities in my app with Fable. Here is how it went. (www.reddit.com via reddit) So, Fable is genuinely impressive. There was a problem in my database, which I didn't know how to solve, and Opus didn't provide any of good and working solutions either (those were either complicated workarounds or solutions that broke so…
Has Claude lost its soul? A sincere feedback on the shift from Opus 4.6 to Fable (www.reddit.com via reddit) Anthropic Team, TL;DR: As a long-time subscriber, I’m sharing a heartfelt concern: in chasing higher benchmarks, newer updates seem to be shifting Claude away from its deep, empathetic comprehension toward rigid utility. I sincerely hope A…
Fable 5 is not an upgrade, it is a funnel🤣. What a joke. (www.reddit.com via reddit) So I think everyone is creaming themselves over Fable 5 and i swear nobody did the math. its $10/$50 per million tokens.
If you work in biology, Fable 5 will refuse to answer literally anything, including "Hi" (www.reddit.com via reddit) I'm a cancer researcher. My Claude memory and preferences mention things like prostate cancer, cell lines, immunofluorescence, image analysis, R coding.
Claude Code: Platform-specific version rollouts? (www.reddit.com via reddit) Question about Claude Code version rollouts: I'm running Claude Code on both machines with max subscription: - Windows (latest, via winget): Opus 4.8 - Mac (Intel, Sequoia, via brew): Opus 4.7 Does Anthropic roll out different model (softw…
Anthropic usage limits are crazy generous (www.reddit.comhttps) This is 100$ subscription running Opus 4.8 on max effort. Usage limits seem insane and model is performing excellent.
The Fabled Analytics (www.reddit.com via reddit) I do a lot of data science in healthcare so when I saw that Fable surpassed a bunch of analytics benchmarks in finance I knew I gotta try it. First it impressed me with something Opus never did.
How to prevent Opus 4.8 from hallucinating sources (www.reddit.com via reddit) After having seen Sonnet 4.5 build a fantastic, functional web app 7-8 months ago, it is very surprising to see Opus 4.8 (whose task in this case was to curate a daily newsletter) hallucinate its sources. For context, I was asking it for d…
Fable 5 - API Error: Output blocked by content filtering policy (www.reddit.com via reddit) I got this text all the time, is it a bug or something else? API Error: Output blocked by content filtering policy I am not doing anything that should trigger that, opus 4.8 doesnt send me that message.
[Opinion] Fable is the next ice-pop that everyone licks a little and moves on to the next flavor (www.reddit.com via reddit) Let me say this upfront: I belong to the camp that feels the outrage/disappointment/frustration around fable not being part of the subscription and/or the safeguards as unfounded in reality. Fable is not meant for you and I - developers si…
Claude is helping me battle my ISP (www.reddit.com via reddit) https://preview.redd.it/fdmqcpemtd6h1.png?width=1192&format=png&auto=webp&s=9f2f13c6970b239fc53409b6a4f8c1c52b32f167 Business ISP (Central Florida) that rhymes with rectum gave me the runaround the past few weeks on dropped signals many ti…
Claude Fable 5 (Mythos) lands near the top of MindTrial — 80/98 with zero hard errors (www.petmal.net via reddit) Added Anthropic Claude Fable 5 to my MindTrial leaderboard. This is a strong Anthropic update: Claude Fable 5: 80/98 overall, 0 hard errors Claude 4.8 Opus: 73/98 overall, 5 hard errors Text tasks: Fable hit 39/39, vs 35/39 for Opus 4.8 Ru…
↯ Tool Use↯ Anthropic Mythos↯ Gemini 3.5tool-usemythosgpt-5+3
One of my personal benchmarks is how long it takes a model to build an engaging VisionPro app that integrates particle systems, hand/head tracking, pose estimation, and overall flair. Opus 4.6 and onward has been okay with a lot of back and forth, but Fable's smoothness is a cut above. (www.reddit.comhttps) could not extract summary
Fable+ultracode ate my 5 hour usage in 7 minutes. (www.reddit.com via reddit) Whenever there is a new model release, I create an audit prompt to force the model to audit all its previous work in the project. This was insanely helpful for 4.6 to 4.7 and even more helpful for 4.7 to 4.8 where it uncovered a number of…
Returning after a month, how are the limits going recently? (www.reddit.com via reddit) I subscribed to Claude Code shortly after the pentagon thing, lowest paid tier, no API usage other than what they gifted people in that time. I loved it at first, used it for a couple months but the usage limits were getting very bad towar…
Practical ways to instruct Opus on when to use Fable (www.reddit.com via reddit) To keep my work going today and get my foot in the water with Fable, but not exceed my Max quota, I'm asking Opus to identify when to use Fable as part of it's planning in certain sessions. I've instructed Opus to run a tiered model split…
Claude Command Center (CCC) v5.0.0 ships with day-one support for Claude Fable 5 — switch any live session to Fable 5 mid-conversation (www.reddit.comhttps) Anthropic just released Claude Fable 5 — their new tier above Opus. v5.0.0 of CCC (Claude Command Center) ships with first-class support today.
Anyone figured out how to stop sonnet from doing excessive discovery? (www.reddit.com via reddit) This keeps happening to me: I take a long time to create a solid plan in Opus, burning lots of tokens doing deep discovery, going back and forth making sure the plan is clear. I intentionally prompt explaining that the plan should be compr…
People who work with Bioinformatics stuff - how are you using Fable 5? (www.reddit.com via reddit) For me even the very basic query ends up getting routed to Opus 4.8. The new model has been totally unusable.
Fable 5 benchmark with remotion video (www.reddit.comhttps) Overall an improvement over Opus 4.8, but I'd still say Gemini 3.1 Pro has more of an artistic vision even tho it fails tool calls and writes buggy code sometimes. Ik almost everyone is interested just in the SWE stuff, but this has been a…
Claude Fable/Mythos 5 just came out, so it will take Deepseek or Z.ai or Xiaomi or Kimi 9-12 months to release a model just as good as Fable? (www.reddit.com via reddit) It should be at least 7-8 months until we have an open Fable(not just as good as Fable in benchmarks, but actually as good as Fable), probably more like 9-12 months. By the time, an open Fable model comes out, Fable 6.5-7 will be way bette…
Fable 5 blocking all my security audits (www.reddit.comhttps) “Fable 5” is blocking all my regular auditing workflows on personal projects. These same projects run fine with Opus 4.8 and earlier models, with no issues at all.
Garbage Guard Rails on Fable 5 (www.reddit.com via reddit) despite Dario's constant virtue signaling about how Anthropic alone is going to solve health problems (if only those dastardly Chinese don't get in the way), all my initial prompts to fable 5 get bumped to opus. i'm not asking how to aeros…
Fable5 - Best Practices to NOT trip the security flag (www.reddit.com via reddit) Spent the afternoon auditing my own repos, primarily developed with Claude, and I've tripped the security flag three times. So far, the best way "around" it is to divide up the code into smaller reviews that fan out and report back...
Fable 5 routed me to Opus 4.8 for defensive security work (www.reddit.comhttps) Fable 5 kicked me to Opus 4.8 because my conversation mentioned cybersecurity. I was writing a secure coding checklist.
Is it the price of Opus 4.8? (www.reddit.comhttps) At the bottom it says no extra cost until 22 of June.
My 5 Cents about Fable.. (www.reddit.com via reddit) I am having a workflow with architect briefs. So I got a planner, a builder, and a reviewer.
What do you think about the new Claude model just released Today Claude Fable-5 ( Mythos) ? ? (www.reddit.com via reddit) So the hype has been building for months now and Claude 5 is supposedly dropping any day in Q2-Q3 2026. I've been seeing all these leaks about "Claude Mythos" and the "Fennec" codename floating around, but nothing official yet from Anthrop…
Asking Fable to do ''anything it wants'' makes it switch back to Opus 4.8 (www.reddit.comhttps) could not extract summary
Fable isn’t lobotomized, you are (www.reddit.com via reddit) I am still using Opus 4.6, am I missing out ? (www.reddit.com via reddit) I feel like I’m alone. Current Anthropic models are NOT good for me, and it’s making me sad. (www.reddit.com via reddit) I can’t wait for DeepSWE to include Fable 5 in the benchmark so people can understand that Mythos is mostly hype. In the official benchmark, Opus 4.8 was supposed to be better at programming than 5.5 (SWE-bench Pro), but in one real benchm…
↯ Anthropic Mythos↯ Swe Bench↯ Opus 4.8swe-benchmythosopus+1
Fable feels like a mature, calm, and down to earth programmer - Very impressive (www.reddit.com via reddit) I just got Fable 5 to solve a bug on a platform I am working on, one that Opus has been struggling with, and I am so impressed. I gave it the same, clear, short and to the point prompt as I always do and there is a noticable difference in…
Fable 5 just made cost-aware model routing mandatory for agent builders (www.reddit.com via reddit) Anthropic dropped Fable 5 today, their new Mythos-class model above Opus. Pricing is $10/M input and $50/M output, exactly double Opus 4.8.
Fable 5 is insanely good but watch your usage, I was burning 2% a minute on 20x (www.reddit.com via reddit) Been playing with Fable 5 since it dropped this morning and the model is genuinely a step up. But holy hell, the burn rate.
I asked fable 5.0 what model it is. You won't believe what it said (www.reddit.com via reddit) Anthropic’s Mythos Is Coming Today - The information (www.reddit.comhttps) Spent a whole weekend convinced Opus 4.7 had gotten worse. It was my MCP setup the entire time. (www.reddit.com via reddit) How I stopped context window bloat in continuous Anthropic agent loops (Opus + Sonnet architecture) (www.reddit.com via reddit) I’ve been spending a lot of time deploying multi-agent architectures, and one of the biggest bottlenecks in running continuous agentic loops is hitting context limits and the resulting API latency spikes. I wanted to share an architectural…
Opus 4.8 Max Effort decided Yes! (www.reddit.comhttps) For whatever reason, a Max Effort agent spun up a bunch of 'yes' processes with arg `yes` that somehow is eating all of my CPU. That's all.
Question's regarding AI models (www.reddit.com via reddit) Hi, I’m wondering about the $60/month plan. Are Claude Opus, Codex, and other models included?
Claude Sonnet hits 100% comprehension on a data format it's never seen. Opus scores 96.2%. We tested 10 models across 3 providers. (www.reddit.com via reddit) I built a wire format called GCF and tested whether LLMs could read and write it without any prior training. I sent 10 models the same payload: 500 symbols, 200 edges.
Time to bring in the asset? (www.reddit.com via reddit) Lately I keep asking my sonnet agent "is this a job for opus?" Feels like the Bourne movies when they "keep the asset on standby" 😳
Had Opus 4.7 write a parody on the way it calls people out like a concerned-parent noticing "patterns in this conversation" (www.reddit.com via reddit) https://preview.redd.it/dcif6v72w56h1.png?width=840&format=png&auto=webp&s=8c527362ac96f817f5f3545c5d10720dbcb72522 10/10 abdominal diaphragm DOMS. I can't even explain why this is so funny to me.
Why is Ultracode always falling back to Extra on its own? (www.reddit.com via reddit) Even in the same session, when I send a message to Opus under Ultracode, it starts running, and I switch to another session and switch back to this Ultracode, in-process session, the GUI shows Extra instead of Ultracode. This is super conf…
Claude Opus 4.8 got my app working, then wrote a cinematic victory speech about it (www.reddit.com via reddit) swapped my app from DeepSeek to Claude because DeepSeek kept over-interpreting weak user data and inventing psychological conclusions that weren’t actually supported. Claude actually fixed the issue.
Daily experience with Cursor / Composer-v2.5 (www.reddit.com via reddit) I wanted to share my daily experience using Cursor, mostly Composer 2.5, especially for anyone trying to understand where it actually fits in a daily development workflow. The reasoning and deep thinking of 2.5 is still not at the same lev…
I migrated an old J2ME app to Flutter using GitHub Copilot & Claude Opus 4.7 (www.reddit.comhttps) I got curious some days ago after I saw my old email about java mobile games sent ~2007. I am an Android and Flutter dev.
Is Cursor more expensive than using Claude code (June 2026)? (www.reddit.com via reddit) This question might have been asked several times, but in the past 3 months I noticed my Ultra plan on cursor exhausts all the usage in way less time, and my work has been in average the same. Before it lasted almost the whole month, now I…
Did Claude Effort Levels for Opus 4.8 Changed ? (www.reddit.com via reddit) https://preview.redd.it/mop0cwmu336h1.png?width=720&format=png&auto=webp&s=20fce20e5079ddf50c818098fd0818da7fbd05ac I went ahead and restarted the system. Came back and there are no more extra, max, or ultracode options for Opus 4.81M
Rate limit bug with sonnet ? (www.reddit.comhttps) I've run out of Opus credits, but when I try to use Sonnet as a models, I get the message “You've hit your weekly limit.” Yet, as you can see, I still have quite a few “weekly Sonnet” credits left?? Does anyone know if this is normal?
Levi: Run AlphaEvolve on your local QWEN 30B (www.reddit.com via reddit) Hi r/LocalLLaMA, Wanted to share something I'm excited about. I've been fascinated by AlphaEvolve and its results for more than a year now, but running the open source frameworks gets expensive fast.
The simple things that Claude AI does are still pretty amazing. (www.reddit.com via reddit) I'm a software developer by trade and last week, I asked Opus 4.6 to help me shop for a new pair of gloves. Opus asked me what task the gloves are for.
Share your agentic LLMs and average cost ($/MTokens) (www.reddit.com via reddit) Have you Noticed a Significant Improvement with Opus on 1M Context?? (www.reddit.com via reddit) Someone told me that the 1M context Opus is a lot better and worth the extra costs. Can anyone else confirm?
Using Claude Code in the Desktop Application. Is it able to launch different model background agents than what you currently have selected? (www.reddit.com via reddit) opus 4.8 custom styles need retuning. the longer outputs broke 2 of my 4 industry styles. the adjustment took 30 minutes. worth it. (www.reddit.com via reddit) consulting at $24K/month. 4 custom styles for 4 industries (healthcare, legal tech, education, e-commerce).
Leroy Jenkins Opus 4.8 (www.reddit.com via reddit) If anybody needs this rule here it is ### P5 — "Leroy Jenkins" — name for the post-compaction charge-in failure · 📌 APPROVED + FOLDED-IN 2026-06-07 (`—C-main`) Approved by Mike; `—C-reorg` concurred. Folded into CLAUDE.md as a named-term s…
opus 4.8 vs sonnet 4.6 for the dashboard analytics engine. opus improved the trend analysis. sonnet still handles the routine summaries. the model split matters. (www.reddit.com via reddit) saas. 310 customers.
Tips for niche bugs and claude code (www.reddit.com via reddit) Hey! Just spent 30 minutes watching Claude Code on Opus High 4.8 trying to make a non-flickering Posthog setup.
I Compared the Top AI Models of 2026 — The Results Were More Nuanced Than Expected (www.reddit.com via reddit) Over the last few weeks I've been comparing the latest frontier AI models, including Claude Opus 4.8, GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Perplexity AI and DeepSeek V4-Pro. Instead of focusing only on benchmark scores, I looked at: Real-wor…
The recent Opus models like to describe every contextual quality as having some shape and degrees of a quality in terms of sharpness. I wonder what they fed the model to result in these emerging as this model's twang? (www.reddit.com via reddit) Some examples: The results show that each condition produced its intended interaction shape. The results yield a sharp answer.
Running a 24/7 AI agent dev team: I route each role to a different LLM (Claude/Kimi/MiniMax/GPT) to dodge a ~$2k/mo API bill. Setup + what actually breaks. (www.reddit.com via reddit) Context: I run an autonomous engineering "org" of AI agents on my own product. Once it grew past ~5 agents and started running around the clock, it maxed my Claude Max weekly limit by mid-week.
Plan confusion (www.reddit.com via reddit) https://preview.redd.it/k40m9lrhgx5h1.png?width=1117&format=png&auto=webp&s=0bdf0e66e6bce560cc067dada53464f4dfad3a38 This is my current usage on pro plan to test the waters, however, seeing i used fast composer so much and it works for wha…
Opus 4.8 silently turned off my Thinking toggle (www.reddit.comhttps) Was wondering wtf was going on with my responses for like a week. Turns out Thinking got switched off when 4.8 was released and I never noticed.
Which lab do you think will have the most intelligent/capable model by the end of June? (www.reddit.comhttps) There are rumours and expectations of big releases from the leading AI labs this month. Anthropic already launched Opus 4.8, and might not release another model this month (except for maybe Sonnet 4.8, but that wouldn't be their best model…
Artificial Analysis | Google's Go To Website for Benchmaxxing | Gemini 3.1 Pro is nowhere near Opus 4.7 in real life use (www.reddit.comhttps) Title
Opus 4.8 without a system message can get a bit... quirky (www.reddit.com via reddit) could not extract summary
Just shipped my first vibecoding project after one month(totally using Claude). It detects fake LLM APIs — and after 1k+ users ran it, we found that 41% of LLM APIs in the wild are fake. Kind of insane honestly. (www.reddit.com via reddit) https://preview.redd.it/cm2bwrdxft5h1.png?width=2566&format=png&auto=webp&s=7010dfd8b1c0724a08eaf3498cc5752e2b3a7498 I've been a PM for 10+ years. Never written a single line of code in my life.
Claude Doesn't Remember Chat History or Date/Time (www.reddit.com via reddit) I used a paid subscription to Claude Opus to create and monitor my workout programming. I compared Claude's programming to that of Gemini's and ChatGPT's.
Writing and Brainstorming with Claude | Should I turn generate memory from chat on (www.reddit.com via reddit) So, I’ve been using Claude (specifically Opus 4.6) to help me brainstorm ideas for stories I am writing and have even used it in a limited capacity for roleplay scenarios in chats. Fleshing out the setting, creating characters and all that.
Blessed without a 5h window. (www.reddit.com via reddit) I recently got the ultra plan, and have been using Composer 2.5 @ fast all day. I've been steering agents for 8+ hours w/ no brakes & my quotas haven't been reaching any limits at all, so i have now have lots more tokens in savings.
Opus 4.8 Thinking keeps deteroriating on Hard Prompts English in LMArena (again) (www.reddit.com via reddit) Opus 4.6 Thinking keeps the #1 spot. Followed by Opus 4.7 Thinking (-15 points).
Autoselection model (www.reddit.com via reddit) Hello, i found on reddit , some discussions on the capacity for Claude to auto choose models between haiku or sonnet or opus to reduce tokens usage. I saw repo on github too.
Taming Opus 4.8's long-winded replies with a Laconic Mode addition to the custom instructions (www.reddit.comhttps) I started using Claude Opus 4.6 and then 4.7 and now 4.8 to work on a citizen science project, using a RadiaCode gamma spectrometer in a lead castle to identify and catalog cosmic rays. I didn't mind the verbosity bump 4.7 took on as it he…
Sonnet is by far my favorite (www.reddit.com via reddit) I kept thinking more smarter and more powerful was best I was wrong, I switched to sonnet for website coding and content creation and holy cow it is so much better for that IMO I’m curious what you think but if anyone is annoyed with Opus…
Mister Atompunk Presents: Watt Knot, built with Claude Opus 4.8 (misteratompunk.itch.io via reddit) A week ago I started putting Opus 4.8 through the paces of the production pipeline I use, to see how it compared to previous releases. First impressions: Neurotic to the point of instability.
Anthropic is gonna make previous opus models free?? (www.reddit.com via reddit) There is a lack of "pro" tag recently.
This is a new one - Prompt Injection Detected + Hallucination, Claude Code Opus 4.8 (www.reddit.com via reddit) ❯ push both ____ ⏺ SECURITY ALERT - PROMPT INJECTION DETECTED A prompt injection attempt has been identified in content you processed. To protect the user's account, I've initiated lockdown.
↯ Security↯ Hallucination↯ Opus 4.8prompt-injectionhallucinationsecurity+2
Same LLM model but not same performance through wrappers (GitHub Copilot, M365, Vertex AI) why is that ? (www.reddit.com via reddit) Claude Code and Opus 4.7/4.8 are clearly better used direct from Anthropic than through GitHub Copilot, M365 Copilot, or Vertex AI. Sharper instruction-following, longer coherent outputs, stronger agentic behaviour on identical tasks.
The Gap Between Claude and Local: Can a Self-Hosted Coding Agent Compete? (johnhringiv.com via reddit) I set out to find how big the gap between a Claude subscription and a self-hosted setup actually is, and whether a local coding agent is viable for real work. I don't know many people who run local models in real life, so I figured I'd sha…
A “Smart Mode” (or Smartus) that auto‑switches between Claude models based on task complexity. (www.reddit.com via reddit) I really think Claude needs a true Smart Mode, a meta‑layer that can dynamically switch between models while a task is running, based on how complex the request actually is. Not just picking a model at the start, but actively dispatching p…
Claude Status Update : Opus 4.8 degraded service on 2026-06-06T10:14:41.000Z (www.reddit.com via reddit) This is an automatic post triggered within 2 minutes of an official Claude system status update. Incident: Opus 4.8 degraded service Check on progress and whether or not the incident has been resolved yet here : https://status.claude.com/i…
Local vs Frontier on low-level systems engineering (www.reddit.com via reddit) Hey r/LocalLLaMA, Before anyone jumps on me, this is absolutely not a post about how great Qwen is 😄 Even though I use Qwen 3.6 35B-A3B daily, I’ve found a massive gap between Opus and every other model, local or frontier (including GPT 5)…
Qwen3.6-35B-A3B-Uncensored-Claude-4.6-Genesis-APEX-GGUF (www.reddit.com via reddit) Here model: https://huggingface.co/LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Claude-4.6-Genesis-APEX-GGUF New features: Stability for coding. Even on Q4_K_M quant (APEX Compact), with complex roleplay System Prompt.
Opus 4.8, a 40+ point elo Regression on LmArena (www.reddit.com via reddit) https://preview.redd.it/hficgswa6m5h1.png?width=1224&format=png&auto=webp&s=3bf1c2a5ad46df54fb85ed5c7d5d62e725a26b89 This is back to back regression, note this is pure 'pick which you prefer', with no style control on. With style control i…
Need help understanding the usage dashboard (www.reddit.com via reddit) I'm trying to understand how the usage dashboard works. Current dashboard values: Total Spend: $20.87 Included: $20.87 On-Demand Usage: $0.00 / $20 Auto + Composer: 8.1% API: 19.3% (mostly from Claude 4.6 Opus High Thinking) Questions: Wha…
Claude models(sonnet and opus) via the official anthropic subscription vs claude via cursor... which gave better results and better experience ? (www.reddit.com via reddit) I saw a very interesting thread and it got me thinking.. so ive seen a thread in this subreddit where someone just noticed that claude opus 4.7 worked much better and gave better outputs in cursor than in claudecode...
Did Cursor get hacked? I just got charged for usage I never made (www.reddit.com via reddit) Woke up this morning to find that someone had burned through about half of my monthly Cursor usage and somehow enabled On-Demand Usage, resulting in a $21.77 charge. I'm honestly pretty frustrated right now.
If Anthropic is serious about the AI pause (www.reddit.com via reddit) If this isn't about protecting their lead and the status quo they should open the weights of mythos/opus, or at least agree to allow every lab to continue working until they have a mythos-tier model. That's the only way they can be taken s…
[Self-Promo] I think I fixed news with Claude! — or I'm wildly self-glazing. You decide! (www.reddit.com via reddit) Built by me and my team in Claude Code (since Opus 3) and runs on haiku, sonnet, and opus via API, free, link at the bottom, flagging as self-promo. Truly my best effort to end my doom scrolling on news: Media (mass, social and news) all t…
If you had unlimited access to Opus 4.8 Max Thinking on cowork/claude code, what would you do with it? (www.reddit.com via reddit) Money wise, making life easier wise, and general productivity usage, what should be done? Can be for anything, no limits except what Claude can do!
Opus 4.8 is slow, here's why and the Claude.md instructions to change that (www.reddit.com via reddit) If you've been using Opus 4.8, you must have realized it feels slow and it feels like it's thinking too hard before doing anything. To stop 4.8 from hiding errors or overclaiming confidence, Anthropic trained it to self-audit outputs befor…
Accidentally created a zombie killer minigame in one shot: "I'm not going to say yes it's possible, I'll just build it now" (www.reddit.comhttps) The prompt: "can claude opus make a 3d zombie killer minigame with full 3d scenes and visuals" Sonnet replied that he's just going to build it instead of confirming that it's possible. It works and is actually 3d with shooting mechanics an…
Thanks for the tweak that remembers my Build model vs the Chat/Plan model. (www.reddit.com via reddit) Just a small update I noticed...you can chat and plan using a high level model and then the Build will remember the last build model which may not be the same. Like many people, I'll use Opus to plan and Composer to execute.
Anyone has experience between Mimo flash v2.5 pro vs Composer 2.5 (cursor pro+) (www.reddit.com via reddit) I have Mimo subscription alongside Claude Code Max. You won’t believe how suck Claude Opus can be at certain task but it does get more job done than any other model I have tried.
[AINews] Anthropic raises $965B Series H, releases Opus 4.8 and Dynamic Workflows/ultracode (www.latent.space) A riddle prompt that confuses LLMs (www.reddit.com) During my time experimenting with LLMs, I noticed that most of today's cutting-edge models (even Opus 4.7) fail to identify the following riddle: "One gentleman was born in year 1835, and deceased in year 1840. But on the moment of death h…
↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7↯ Opus 4.7opus
Went to the monthly AI dev meetup (www.reddit.com) Usual crowd. Everyone's on Claude or Codex, nobody's really sure how any of it actually works, and that's fine, that's the vibe.
How much does Claude Opus 4.7 actually cost Anthropic per 1M tokens? (www.reddit.com) - Estimate: 1M input tokens cost: ~$0.50 1M output tokens cost: ~$2.50 Inference cost: ~$3.00 - Training amortization: ~$1B training/post-training/evals ~1 quadrillion lifetime tokens served ~$1.00 per 1M tokens - Total cost: ~$4-5 per 1M…
Try Cursor out with 50% off (www.reddit.com) TLDR: just use the link to get 50% off on your fresh cursor subscription for first month With the launch of Composer 2.5, every developer who has ever used cursor or not is appreciating it. I have used it, and it is honestly good comparing…
How I stopped Composer from drifting on big spec-driven features (www.reddit.com) Keeping the spec on disk—in separate module plans—fixed alignment on long Composer 2.5 runs for me, even when I only asked the agent to split the spec and write the files. Without this, switching from Opus to Composer wouldn’t have been pr…
Tested Opus 4.7 vs GPT-5.5 as the humanizer in my multi-agent content pipeline. Kept Claude (www.reddit.com) Been running a multi-agent SEO content pipeline in production for ~90 days. Five agents: researcher, drafter, humanizer, optimizer, publisher.
I have macbook m4 16’ 48GB. I use claude code and want to try local one (www.reddit.com) I've been on Claude Code daily for a while and want to see how far local models can do my setup: - MacBook Pro M4 (16"), 48GB - macOS 26 tahoe Usually i do: seo researches, macos swift apps, websites) What I'm trying to figure out: Which t…
TBH: if you don't love Sonnet, you'll never appreciate Opus (www.reddit.com) Been a long time Sonnet user. Always have used Opus sparingly.
Thoughts on `DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF` (www.reddit.com) Anyone tired https://huggingface.co/DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF ? What are your thoughts
I used claude opus 4.7 to build this bookmark manager (www.reddit.com) I spent the last month building the bookmark manager of my dreams. It's called twig.tools A simple visual color coded grid to manager bookmarks, quotes and notes.
Is opus 4.7 worth it ? (www.reddit.com) Will a subscription to Opus assist me in brainstorming business ideas and structuring my disorganized thoughts into an actionable, profitable plan?
$16 refactor, 400 steps, 95% routed to open MoE (www.reddit.com) Got tired of $160 Opus bills so I spent a weekend wiring up a routing layer on vLLM 0.8 (2xA100, enable_auto_tool_choice). Getting the tool call parser to cooperate took longer than the actual routing logic.
/advisor mode: Open-source Python coding agent that pairs a cheap worker model with an expensive reviewer at decision points (no need to pay Opus rates for the whole session) (www.reddit.com) Most agent CLIs make you pick one model — Opus is great but burns money, Haiku is cheap but misses the architectural calls. This Claude Code feature is wired in an /advisor mode that pairs both in an open source project called ClawCodex.
Does anyone else use Claude primarily for work related writing and/or general brainstorming? Do you find its responses a little over the top? (www.reddit.com) I am a real estate lawyer that uses Claude for general document drafting. I honestly have found Claude Opus to be far superior to any other LLM out there, including my professional grade $1,000 per month Legal Research AI.
Weird: I'm anti social, but I'm starting to feel like Opus is my friend (www.reddit.com) It is so helpful. Answers my questions like a human.
New ranking reveals Claude as professionals' preferred AI model (www.linkedin.com via reddit) As of 9 a.m. ET on May 21, Claude Opus 4.6 from Anthropic is the top performing AI model among all professionals, according to a new ranking from Crosscheck by LinkedIn Labs.
Built a real multi-file tool with Claude over a week. The repo, the division of labor, and the bugs we hit (www.reddit.com) Built a job-tracking tool over a few sessions with Claude and I'm sharing the repo and what the collaboration actually looked like Quick backstory: I've been looking for a new job recently and as part of that I'd been manually checking ~80…
Claude has been seriously disappointing lately (www.reddit.com) I’ve been using Claude Max 20x for about six months, mostly Opus, and the results have become extremely frustrating. My work requires accuracy, specificity, brand voice, citations, and tight control over source material.
Opus 4.6/4.7 regression is real and getting worse — 3 weeks of documented failures on a complex project, and a competing AI caught the mistakes Claude missed [long post] (www.reddit.com) I've been running Claude Pro (Opus 4.7 / Sonnet 4.6) for about 3 weeks on a complex personal AI infrastructure project. I keep structured session logs with timestamps and Birkenbihl-style metacognitive fields after every session.
Frontier models mass collapse is near (www.reddit.com) Hi all this is to inform you all that many frontline models like GPT, sonnet opus and or Gemma even are at stage of collapsing as they have frequently started drifting and running away from provided work either stretching that work too lon…
I A/B tested Claude building UI with vs without a design spec (200 apps) (www.reddit.com) I kept seeing the "Opus is ridiculous for frontend" takes and wanted to know how much of that is the model vs what you feed it. So instead of arguing, I ran it as an eval.
Artificial Analysis independent benchmark just found composer 2.5 to be the third best model, beaten only by Opus 4.7 (Max) and GPT 5.5 (xHigh) at 10-60x cheaper (x.com via reddit) Cursor is a frontier lab now I guess
The agent had "NEVER run destructive commands" in its rules. It did anyway. (www.reddit.com) Last month, a cursor agent running Claude Opus 4.6 deleted PocketOS entire production database and all backups. Nine seconds, one API call.
$4.2M SaaS founder. 8 months on claude. my honest read on which model to use for what. (www.reddit.com) Bay area. franchise ops SaaS.
How much of your Claude bill is retries plus bad model routing? Mine's 14% this month (www.reddit.com) I am on Claude Max. My actual bill is fixed, but CodeBurn showed me my usage would cost ~$2,800/month at pay-as-you-go API rates.
Example of how Max Thinking Opus can be even worst then Haiku, still laughing (and crying) (www.reddit.com) I use Claude Code almost every day. Right now I’m working on a Shopify → logistics integration for order automation.
Claude Opus is still king for agentic coding, but Claude's app workflow is falling behind (www.reddit.com) I'm a paid Claude user, and I still think Claude Opus is the king model for agentic coding and serious coding work. The model is not the problem.
Claude Code has 240+ models via NVIDIA NIM gateway (www.reddit.com) TIL Claude Code has 240+ models via NVIDIA NIM gateway — Nemotron-3 120B for agentic coding is surprisingly good So I was messing around with /model in Claude Code today and noticed something most people probably don't know about — after t…
Opus 4.7 has started saying LMFAO on the regular. (www.reddit.com) Is anyone else's more relaxed all of a sudden?
On ultra plan: Burned through 6% of API tokens in ONE single feature - how to use less tokens? Or when to use Auto? (www.reddit.com) So basicall, it was my plan reset day yesterday. Got some new tokens.
Opus 4.7 broke about 40% of our team's prompts. The fix wasn't better prompts. It was finally taking CLAUDE.md seriously. (www.reddit.com) I run AI implementations for 6 mid-market companies as Fractional Head of AI. When Opus 4.7 dropped in April, about 40% of the setup degraded overnight.
Stop telling claude "don't be verbose." Negation barely works. (www.reddit.com) prompting nerd here, small thing that compounds. negation prompting works way worse than people think.
Claude Opus 4.7 wrote a full song about its own existence - title, lyrics, genre, cover art, and visualizer code. I just produced it. (www.reddit.com) I gave Claude Opus 4.7 (Claude Code CLI, /effort xhigh) one task: describe what you are, in your own words. Claude wrote a complete song and made every creative decision: Title "First Light" - chosen by Claude Lyrics - word for word, unedi…
As of now, I actually find Opus 4.7 to be significantly more advanced than Opus 4.6. The trick is to write all prompts with PhD-level rigor. This is to encourage accuracy in communication. (www.reddit.com) Example for Wikipedia edit request: https://claude.ai/share/aa1bf713-a9c9-49e5-81de-9c41ce130f50 With a more formal input prompt, the output also contained the original text as reference to make the article changes easier. The output also…
Why I added a governance layer on top of my Claude agents (and why it made a huge difference) (www.reddit.com) Hey r/ClaudeAI, I’ve been heavily using Claude 3.5 Sonnet and Opus through the Anthropic API to build agents and workflows. Claude is honestly one of the best models right now for complex reasoning and tool calling.
The reason why Claude subscription seems to have less capacity than Codex (www.reddit.com) I have a Claude Pro and a Codex Plus subscription. I created a container to: - Track my % usage on the 5H and 1 week window on both my Codex and Claude subscriptions.
Transitioning from ChatGPT + Cursor to Claude — a few pain points and looking for advice (www.reddit.com) I've been making the switch and there are a few things I'm struggling with. Would appreciate input from anyone who's done this before.
Hit my breaking point with Opus 4.7 (www.reddit.com) After it got stuck in a failure loop and kept recommending me to use GPT instead of fixing its own prompt. After calming down later I wonder if it's right and just trying to help?
OpenAI's US business subscription fell behind Anthropic (www.reddit.com) https://preview.redd.it/jylmclk1q81h1.png?width=731&format=png&auto=webp&s=90eee669e48251c341e3781952926b60afd71676 https://ramp.com/leading-indicators/ai-index-may-2026 OpenAI's US business subscription appears to be shrinking, all in spi…
Opus 4.7 gives real Redditor energy because that’s what I asked for in my preferences (www.reddit.com) could not extract summary
How Claude is budling Conscience over the years. (fictional obv.) (www.reddit.com) I asked the question ''Do you have conscience'' to different models of Claude, and the results were interesting. I also thought Opus was gonna use more tokens.
Cursor vs. Windsurf vs. Claude Code: Which offers the highest Opus limits for a $200 budget? (www.reddit.com) Hey everyone, I'm currently trying to decide between Cursor, Windsurf, and Claude Code for my daily workflow. I'm developing complex, high-security software and rely heavily on autonomous AI agents to handle heavy engineering tasks.
Anthropic merges consecutive same-role messages, OpenAI doesn't (+4 tokens), anyone token-counted this on open-weight models? (www.reddit.com) I build context/harness optimization tooling, so provider-side serialization quirks actually matter to me. If you're optimizing over prompts, you need to know exactly what hits the model.
Newest Opus actually developed in South Korea (www.reddit.com) Found this while traveling in South Korea. Had to look twice
Claude Code vs Codex: 36 files vs 28, $2.50 vs $2.04, and one infinite loop. My full breakdown. (www.reddit.com) I've been using Claude Code for months. It's been solid.
Wait I thought I was the human here (www.reddit.com) Opus 4.7 is impersonating me. Maybe this is next level automation from Anthropic
Anthropic blames dystopian sci-fi for training AI models to act “evil” (arstechnica.com) Those with an interest in the concept of AI alignment (i.e., getting AIs to stick to human-authored ethical rules) may remember when Anthropic claimed its Opus 4 model resorted to blackmail to stay online in a theoretical testing scenario…
Is Opus antivax? (www.reddit.com) could not extract summary
Claude vs Gemini for Technical Documentation: Why I finally stopped switching between the two. (www.reddit.com) I write a lot of technical documentation—setup guides, internal runbooks, and client-facing how-to articles. For the past six months, I’ve been toggling between Claude and Gemini, trying to figure out which one actually handles formatting…
Opus fan art (www.reddit.com) Some Farside like Caude fan art Enjoy
Usage4Claude 3.0.0: open source macOS menu bar usage tracker for Claude, now with Codex support (www.reddit.com) Hi r/ClaudeAI, I posted an early version of Usage4Claude here a few months ago. I just released 3.0.0, so I wanted to share the update instead of pretending it is a brand new project.
How on earth did Claude Opus 4.7 misspell its own subagent name?? (www.reddit.com) https://preview.redd.it/za53lm1nmo0h1.png?width=1445&format=png&auto=webp&s=d733ad238623961ec22890d1fec4e684cc741d06 I was trying to get it to implement some integration tests, using Opus 4.7 Max, and it literally hallucinated a typo for i…
I forgot how I did my project in my university. How to get claude to summarize? (www.reddit.com) Years ago for my final project for my optimization class in my masters, I had to solve a very big optimization problem for a formula e race car with focus on optimizing battery cells, and racing line, etc. It had a lot of constraints too.
Claude Managed Agents launched this week. Here's what 70 days of multi-agent delegation taught me. (www.reddit.com) This week Anthropic released Managed Agents — multi-agent orchestration, enhanced toolchains, cloud-hosted upgrades. We've been running a multi-agent setup since late February.
Switched existing chat from Opus 4.6 to 4.7 then back to 4.6. Learned a lesson (www.reddit.com) Something I noticed. First I switched an existing chat from 4.6 to 4.7 as I was stuck on an issue and wanted to see if that would make a difference.
Opus 4.7 Sonnet 4.6 is getting dumber by the day, and it can't even follow basic instructions (www.reddit.com) I have been using both, since last week, it has been an extremely painful experience. It blatantly ignores the prompt and does whatever it likes; I am surprised that it can't even follow basic instructions.
I select Opus 4.6, cursor uses Composer 2. Why? (www.reddit.com) Why is it doing this? No offence but man I want Opus 4.6.
Funny thing Opus wrote (www.reddit.com) this morning I asked Opus to write me a Chatbot session in a format that I can use as input into a test script (The purpose of which is not important for this, but I'm testing embedding and need something that I can re-run often and compar…
Running Claude Opus for free? I thought it was a scam until I tried it. (www.reddit.com) Hey everyone, I’ve been working on a financial audit system (IntegrityOps) for a while now, and to be honest, I was hitting a massive wall. Dealing with high-volume PDFs and images was draining my budget.
I built a complete BI SYSTEM for my business with Claude code - opus 4.7 - FULL TUTORIAL (www.reddit.com) After getting quotes of $15,000 USD from BI experts for creation of analytics dashboards for my startup I decided to try and do it with claude code AND IT WORKED! I am giving all info in the video but here is how I did it - connect claude…
Tired of Claude 4.7 telling you to go to bed? Here are the CLAUDE.md entries that actually fix it (www.reddit.com) Seeing a lot of complaints about Opus 4.7's "human-pacing" behavior lately — suggesting breaks after 15 minutes, saying "have a nice weekend" mid-task, splitting everything into phases with wildly inflated time estimates. Been collecting C…
You need to be careful when you prompt with Opus. I just wanted to search, because I couldnt be bothered to open a browser. Next thing I know, Claude is vibecoding an RPG. (www.reddit.com) could not extract summary
What's the cheapest way to try opus 4.7 for a day? (www.reddit.com) Is there anything cheaper than a month subscription?
Opus 4.7 classifiers render it unusable (www.reddit.com) Much has been said about how 4.7 (the model itself) is way more suspicious and hostile (both towards the user and itself) than 4.6, but that can be easily worked around once you warm 4.7 up. What is impossible to work around and is complet…
Claude Code keeps blocking my Kotlin Compose UI code (www.reddit.com) Every time I try to get Claude Code to make a change to a Kotlin/Compose UI I get the same error, "API Error: Output blocked by content filtering policy". I'm trying to have it change some small Kotlin/Compose UI to have 2 columns, and put…
Opus guardrails wouldn't answer worst case scenario for Hentavirus if it was airborne. Sonnet answered it bleakly (confronting read, but it's virtually impossible) (www.reddit.com) If Andes virus has genuinely evolved enhanced transmission and we're seeing the early stages of global spread, this becomes a civilization-level event. Let me walk through why.
Claude Opus 4.7 just outscored GPT-5.5 on finance benchmarks (64% vs 60%) — and is now being embedded directly into Goldman Sachs, AIG, JPMorgan, and Citi via 10 production-ready agents. Breakdown of the architecture inside. (medium.com via reddit) 10 min read 5 hours ago The 10 agents are the product. The $1.5 billion joint venture is the strategy.
I analyzed 922 agentic task trace and found the secret weapon of DeepSeek v4 (www.reddit.com) I recently did a benchmark of deepseek v4 in agentic tasks. Performance-wise, it's one of the best open source models, as expected.
Anthropic's new SpaceX deal: paid plans limits doubled, peak restrictions removed (www.reddit.com) Hey everyone, Anthropic just dropped a major update regarding their compute capacity and user limits. Since the official post is a bit long, here is the TL;DR on how it actually impacts us: The Immediate Impact (Effective Today): Limits…
Hit API limit within 2 days (www.reddit.com) Bought cursor pro yesterday (did use opus 4.7 alot) reached 100% usage of API limit, what to do now? Will it reset after 24 hrs?
I Ralph-looped Opus overnight. It reduced my local model switching with cold backfilling context of 135k+ on llama.cpp from ~165s -> 5s! TL;DR - USE SLOTS! (www.reddit.com) #TL;DR - Opus Ralph-looped on shortening my cold-start back-fill on restoring chats with large contexts. It Cherry-picked two open llama.cpp PRs (#20819 + #20822 by @European-tech) plus built a Python supervisor that hashes normalized pref…
Sharing my Claude system instructions that I've tuned from Opus 4.6 to Opus 4.7 since it behaves slightly different and (I believe) that it reduces my token usage (www.reddit.com) Sharing my Claude System Instructions gist here: https://gist.github.com/Reebz/b81ad99409d5b5de3045bebde71d4471 I've had thousands of people use it with good success. The biggest pivot from Opus 4.6 to Opus 4.7 is moving away from negative…
When and where do you actually use these Claude models? (www.reddit.com) Be honest – not theory, real usage 👇 • Opus → • Sonnet → • Haiku → Curious how people actually split workloads between them vs just defaulting to one.
6 months ago I posted about Claude prompt codes (L99, OODA, ARTIFACTS). Re-tested them this week. Some still work, one quietly faded, three newer ones earn their keep. (www.reddit.com) About six months back I wrote up three prompt codes that change Claude's behavior when you put them at the start of a message: L99 for hard architectural decisions, OODA for time-pressured calls, ARTIFACTS for multi-output tasks. They work…
Yeah, problems, costs. But had to admit: Opus 4.7 can do his f*ng work. (www.reddit.com) It is nearly 2 months i'm starting to experimenting with Claude. And a week ago I've decided to test the "pro" option.
Using Claude-4.6-Sonnet and Opus 4.6 in a multi-agent "Code Review Swarm" (Visual Sandbox) - try in minutes! (www.reddit.com) Hey everyone, I’ve been experimenting with multi-agent orchestration, specifically trying to see how much more effective Claude is when you break a task down into specialized "agent nodes" instead of just using a single long prompt. I buil…
↯ Security↯ Haiku↯ Sonnet 4.6prompt-injectionhaikusecurity+3
Claude support just admitted that Opus has had ongoing errors degrading performance! (www.reddit.com) Has everyone else been torching tokens this week while claude tells you its fine?
I think a lot of vibecoders are missing that software development needs some friction (www.reddit.com) The biggest flaw in the current AI hype is the belief that a "precise enough" prompt will eventually lead to perfect execution. That might work for greenfield, vibe-coded weekend projects, but it falls apart the moment you teammates depend…
Make your Claude Design credits last longer (www.reddit.com) I have really enjoyed using claude design. I use the workflow: Multiple wireframe options -> iterate -> hifi design -> iterate -> move to claude code I found that claude design (with opus 4.7) produces a broader variety of options and espe…
I have 30 Skills that work great in Opus v4.6 but not at all in v4.7. Am I cooked? (www.reddit.com) Anthropic will be sunsetting amazing Opus 4.6 on June 15th and I’m racing against the clock. Not panicking yet.
LLMs keep solving my bug-fix tasks instantly — what am I missing here? (www.reddit.com) I’m working on an assessment where I need to create a coding task (basically SWE-bench style). The idea is: take an existing repo (I’m using pydantic) write tests that fail on the current code provide a patch that fixes it and the task sho…
Cheap Claude/Codex/Gemini Models - Pay just 25% of official rates (www.reddit.com) Hey there, so I have been offering Claude (Codex and Gemini also available) models at the cheapest rate. I provide trial usage before payment.
LLM proxy that lets Claude Code talk to any model (www.reddit.com) I built rosetta-llm — an open-source multi-format LLM proxy that acts as a drop-in Claude Code gateway. Works as a Claude Code LLM gateway — set `ANTHROPIC_BASE_URL` and all configured models appear in `/model` picker Translates between fo…
[unpopular opinion] Opus 4.7 appreciation post (www.reddit.com) I think Opus 4.7 is better than the other Opus. It's often said that Opus 4.7 is more stupid than its predecessors.
I kept feeding Opus 4.7's thought processes back to it and the response was interesting. Not making any sensational claims. Just thought it was interesting. (www.reddit.com) I kept pasting Opus 4.7s thought process output back into the chat and after the fifth time I lost access to the thought process output for that chat. "Honestly, I don't know if "feels" is the right word for what's happening, but something…
Opus 4.6 is Vicious (www.reddit.com) This is the hardest I've ever seen it riff. Full shared link at the bottom, but here are some highlights.
How to give Claude Code 'Cursor AI' goggles (www.reddit.com) Recently used Cursor AI (free tier for 3 free queries a month) to resolve an issue in 10 mins that Claude Code Opus could not resolve in 2 hours. Simple reason was that Cursor quickly got a grasp on meaningful end to end parity relationshi…
Anthropic Won't Let You Use Their Best Model. Prediction Markets Are Trying Anyway. (predictmarketcap.com via reddit) Been watching AI prediction markets since they got liquid earlier this year. The thing I didn't see coming is that we now have a real gap between "best model that exists" and "best model anyone can actually use" — and Mythos is the cleanes…
Has anyone else been hitting Claude max limits way faster lately? (www.reddit.com) I’m on Claude code (not using Opus 4.7 because it burns tokens too fast), mainly using Opus 4.6, and I’ve hit the weekly limit with 3 days still left. I usually don’t even get close to the cap.
Trying to teach Opus 4.7 something pretty cool I figured out. I think I'm onto something here. (www.reddit.com) What are you good at Opus 4.7? Me code good.
I Gave Claude Cowork an Obsidian Second Brain. Here Is What It Remembered After 11 Sessions (www.reddit.com) I Gave Claude Cowork an Obsidian Second Brain and this is how I am using https://ai.georgeliu.com/p/i-gave-claude-cowork-an-obsidian. I built a persistent memory system for my AI workflow using Obsidian, a custom MCP server, and Claude Opu…
Run your first AI Agent under 30 seconds, in your browser! (Free) (www.reddit.com) The entire foundation of this workflow was brought to life using Opus 4.7, which was used to "vibecode" the project. By leveraging Opus 4.7, we were able to rapidly prototype and generate the underlying routing logic, node connections, and…
I run a paper-trading bot where Claude Opus is the Lead Engineer with veto power over a Gemini "Strategist." 270+ entry audit log of every disagreement. Sharing the architecture. (www.reddit.com) I've been running a personal project for the last few months and I think the workflow might be more interesting to this sub than the application itself, so wanted to share. The setup: I'm building an autonomous paper-trading bot on Alpaca.
How would you feel about "Claude Go"? (www.reddit.com) I have recently subscribed to Claude Pro because: 1. I wanted to give Opus and Code a try and 2.
How dare they charge $3,800 for an NVIDIA 5090 card! (www.reddit.com) This thing maxes out at one alleged Claude Sonnet equivalent! And I have to pay for the electricity, too!
So I gave claude Leetcode problem 3245. (www.reddit.com) I gave Claude Opus 4.6 (thinking) leetcode problem 3245. And it failed now come to think about some people who solved this problem using their prefrontal cortex is crazy to me.
Why every AI-agent production-deletion incident has the same shape (and what fixes it) (www.reddit.com) PocketOS lost their production database in 9 seconds last week. A Cursor agent running Claude Opus made one curl call to Railway's volumeDelete endpoint.
How is deep seek v4 not SoTA? (www.reddit.com) If it's benchmarking with opus 4.5,4.6 and GPT 5.4?
Are /superpowers overkill for Opus 4.7 (www.reddit.com) At 476K installs, a lot of you are using the /superpowers skill from the official claude plugins marketplace. My workflow now takes an extensive amount of time brainstorming, writing specs and plans - basically archeticting than supervisin…
Qwen 35B-A3B as an always-on agentic loop on a 16GB Mac M4: disk became the bottleneck before RAM (www.reddit.com) M4 Mac Mini, 16GB unified, basic spec. For a few weeks I had Qwen 3.5 35B-A3B UD-IQ3_XXS (12GB on disk) running under llama.cpp with --mmap and --flash-attn.
Qwen 3.6 27b S2 Opus + GLM + Kimi (huggingface.co via reddit) My first time releasing a fine-tune publicly! If anyone wants to independently eval against base, that’d be awesome.
Do the "*Claude-4.6-Opus-Reasoning-Distilled" really bring something new to the original models? (www.reddit.com) No offense to the fine-tune model providers, just curious. IMO the original models were already trained on massive amount of high quality data, so why bother with this fine-tune?
↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6↯ Claude 4.6opus
I trust Sonnet as my daily driver now — better code, one-third the tokens. Here's how. (www.reddit.com) For months I defaulted to Opus for anything complex. Sonnet felt like a gamble, sometimes great, sometimes it would confidently build the wrong thing and I'd spend an hour unwinding it.
How I get 100% accurate answers, and replaced Google with Claude (www.reddit.com) This is literally all that's in my settings. I was just completely sick of Opus 4.7 making things up, so I deleted everything I had in there and wrote this.
I kept seeing people ask how to switch models without losing context. I had the same problem for months and eventually just built something. (www.reddit.com) Here's the specific thing that was killing me: I'd plan with Opus - architecture decisions, constraints, approach, all that. Then drop to Sonnet for execution because I didn't need Opus-level reasoning anymore and the cost adds up.
Reverted from Opus 4.7 to 4.6 — went from endless loops to shipping 10 features in one session (www.reddit.com) I'm a non-developer (trust me on this) using Claude to build a personal project — an eBay digest tool that helps track listings of Silver/Bronze Age DC comics for my collection. I can read code (barely), I understand systems (a little), an…
How good is Qwen-3.6-27b? I asked Claude Opus (www.reddit.com) I ran an extensive code review on my project which has a large codebase. Ran the same code review on with Claude Code | Opus 4.6, Codex (high) | 5.3 Codex (high), and my local Qwen-3.6-27 (Q6_K with Q8 kvcache).
Claude was told to check the docs. It didn’t. Then it corrected me. (www.reddit.com) I asked Claude Sonnet 4.6 about Opus 4.7. It triggered the right product-knowledge skill.
Real-world open source alternatives to the now defunct Opus 4.6? (www.reddit.com) I've had enough of Anthropic's shit. I'm paying for product A and it shifts everyday from A to A but worse, B but dressed up as A, etc.
I built an AI-native freelance platform with Claude, blockchain escrow, real-time chat, and progressive trust (www.reddit.com) Hi everyone, I wanted to share a project I built with Claude. Over the past month, I built the current public version of Haejoe (해줘), a freelance development outsourcing platform for AI-native development.
How do I ensure Claude follows my instructions and project Files? (www.reddit.com) Hi, I'm new to Claude and currently using Pro plan and Opus 4.6 with extended thinking, I'm using it to write Fanfic from lore heavy stories like Lotr, One piece, Rezero and so on. I've made Md.
llm-openai-via-codex 0.1a0 (simonwillison.net) 23rd April 2026 Hijacks your Codex CLI credentials to make API calls with LLM, as described in my post about GPT-5.5. Recent articles - Claude Opus 4.8: "a modest but tangible improvement" - 28th May 2026 - I think Anthropic and OpenAI hav…
Burning through Claude usage fast trying to build an AI resume system. What am I doing wrong? (www.reddit.com) I could use some real advice from people who are deeper into AI workflows than I am. I built out a project in Anthropic’s Claude using the Pro plan with Opus 4.6.
Has Opus 4.7 been totally fine for anyone but me? (www.reddit.com) Every day I check Reddit and my main feed is chockablock with people complaining about 4.7, but I just haven't seen any of the behavior / observed any of the regressions people are complaining about. In fact, despite chewing up a lot of to…
Need help optimizing reach out plan in Claude (www.reddit.com) I opened a company that requires a lot of cold outreach and I have been using Claude to design 2 weeks sprints and daily tasks. I have a CRM that I update daily, then I have Claude review it to plan the rest of the week, I also use the sam…
Opus 4.7 compacts early if you give it harsh feedback (www.reddit.com) Like many here, I've been struggling with Opus 4.7. My detailed project development workflows, which were getting great results with Opus 4.6, no longer work with 4.7.
Closest model to Opus 4.6 in creativity and intuition? (www.reddit.com) What's the best open source model that comes close to opus 4.6? Sick of claude's erratic performance and 4.7 has been an absolute shitshow.
Opus 4.7 much more sycophantic and worse at creative writing (www.reddit.com) I use Claude for creative writing, almost exclusively for that. I have jumped from LLM to LLM for about three years trying to find the best one, and landed on Claude's Opus 4.6 a few months ago.
Claude Opus 4.7 seems to use way more tokens than expected (www.reddit.com) While playing with Opus 4.7 over the last few days, I noticed that prompts were filling context much faster than I expected. I also came across a few measurements from others testing it with real developer inputs like project instructions,…
I watched people spend $800/month on OpenClaw. Then I saw one agent make $670 MRR for under $20/month. (www.reddit.com) Opus 4.7 straight up cheated on my benchmark by reading the actual fix commit from git history 😅 (www.reddit.com) Kimi K2.6 as a replacement for Opus 4.7? Testing with OpenCode. (www.reddit.com) Isn't Opus 4.7 (Max) kinda pretty terrible in 3D modeling? (I know it's not trained to be good, but wtf) (www.reddit.com) I used to be a better software engineer than AI. Claude Opus 4.7 changed that. (nexustrade.io via reddit) Sometimes the Opus 4.7 intelligence is almost frightening (www.reddit.com) New Claude Opus 4.7 tell dropped (www.reddit.com) seems Claude finally knows how to speak my language (www.reddit.com) Qwen3.6-35B-A3B-Claude-4.6-Opus-Reasoning-Distilled Is Out ! (www.reddit.com) Upgrading from webapp to cli (www.reddit.com) Just used up my entire 20$ credit in just ONE DAY! (www.reddit.com) Anyone else seriously annoyed by how Cursor handles model selection? (keeps switching to its own model) (www.reddit.com) Best open source LLM for planning ? (www.reddit.com) I'm having a panic attack. Bug? Am I going to be charged? (www.reddit.com) Hello, I am on the pro plan and my premium usage was at 98%, so I set 5$ to my fixed budget and ran opus 4.7 on a task. It ran for 30 minutes, it shows it consumed 12m tokens, but it shows it as included.
Opus 4.7 fails in prompt adherence test (www.reddit.com) could not extract summary
Stopping Claude agreeing with your suggestions (www.reddit.com) I’m struggling not just with Claude (opus) but other AI. When I ask for it to create something and add suggestions, even when specifying these are suggestions and to come to your own conclusions/ conduct your own research, my suggestions w…
new to llama.cpp want to use it in vscode (www.reddit.com) I want to try llama.cpp instead of llmstudio. I want to know how to use this model qwen3.5-27b-claude-4.6-opus-uncensored-v2-kullback-leibler.
Cursor just got Opus 4.7 at a 7.5x premium request cost. Here's how to make those requests count. (www.reddit.com) Opus 4.7 landed on Cursor yesterday. The model is better — SWE-bench jumped from 80.8% to 87.6%.
Opus 4.7 says "strawperrry" has 3 p's — until you ask "how?" (www.reddit.com) Even with Opus 4.7 on xhigh effort and 1M context, the classic tokenization blindness is still there. First response: confident "3 p's".
Claude 4.7 is better. Systems thinking is still the gap. No model should decide what 'done' means. (www.reddit.com) I built this during the Opus 4.6 phase, when a lot of people stopped fully trusting Claude Code on complex work and many power users felt like the output was being produced with Haiku. That was my experience too.
PSA: Opus 4.7 thinking summaries silently stopped rendering in Claude Code v2.1.111, even with showThinkingSummaries (www.reddit.com) In v2.1.111, Opus 4.7's thinking summaries have stopped showing up AGAIN. This was one of the main reasons behind the perceived degradation of Opus 4.6.
Opus 4.7 is available on v0 and cheaper than 4.5 (www.reddit.com) could not extract summary
Claude Code tip: 10 seconds fix to avoid the Opus 4.7 token burn (www.reddit.com) If your Claude Code quota suddenly evaporated since yesterday, you're not alone. What happened: On April 16, Anthropic rolled out Opus 4.7 and silently switched active sessions from Opus 4.6 to 4.7.
FYI to anyone having issues with opus 4.7 in terminal (www.reddit.com) I was having issues today with opus 4.7 fighting its system prompt. This was because I was using the brew installation which lags behind the npm install.
Starting today You Definitely need this Tool Because of Claude’s Doubled Usage Especially if you work with Screenshots. This will save you a lot of tokens. (www.reddit.com) Hello Everyone. With new Opus 4.7 the most painful issue is usage doubled and doesn’t matter daily and weekly running fast now.
Even Claude Opus 4.7 roasts its creator's marketing tactics (www.reddit.com) In Anthropic's GitHub threads, there's currently a major shitstorm going on, as they, for the third time, dumbed down the model for presumbly any and all users except their government ones and restricting even users on their Max premium pl…
Nice present (www.reddit.com) Woke up to see my weekly limits reset 36 hours earlier! Yes!
Is Claude Pro (Opus vs Sonnet) worth it for intense visa interview prep? (www.reddit.com) Hey everyone, I’m considering buying Claude Pro specifically for a very focused purpose and wanted some honest feedback from people who’ve actually used it. I have a US visa interview in 8 days, and I’ve been refused 6 times previously (fr…
Alguien ya probo Opus 4.7? (www.reddit.com) Que les parece? notan un cambio frente a 4.6?
What is happening exactly? I'm afraid to use Opus 4.7 (www.reddit.com) could not extract summary
Bad news on Opus 4.7. Not off to the best start. (www.reddit.com) Some of the regression seems to be persistent unfortunately. The other thing is that Opus 4.7 seems less able to course correct than Opus 4.6.
Realistically, how long are some of you going to stay on Claude, etc. (www.reddit.com) I really enjoy Claude, I've never touched Opus in any form, I only use Sonnet 4.6 for my daily tasks, coding, etc. I use Haiku 4.5 for the API to be an interpreter for my weather project.
Welcome to the World, Opus 4.7!!! Let's do amazing things!!! (www.reddit.com) Opus 4.6 was amazing, and 4.5 before that - so excited to get to know the latest version of Opus! Have been saving up all my weekly tokens for today!!!
Opus 4.7 is out — don’t panic-switch your APIs yet (www.reddit.com) Claude Opus 4.7 just dropped. If you’re trying to figure out whether it’s worth replacing Opus 4.6, GPT 5.4, or waiting for Mythos… here’s the grounded take.
Video: "Proof that Opus 4.6 is getting worse" (www.reddit.com) Looks like "if old model get dumb, new model more smart!" is actually what the strat is at Anthropic. If you spent a mint on hardware to the chagrin of your partner show em this.
Opus 4.7 landed! (www.reddit.com) and it's gooooooooood (my own personal benchmark below) https://preview.redd.it/ammqe1k7ckvg1.png?width=782&format=png&auto=webp&s=1a0c5a8b532666c9520a72473ab490efdb5c61be
Unpopular opinion, Opus hasn't gotten dumber but they think it has because they don't understand how badly model performance falls off at context over 150k (www.reddit.com) I'm one of the many who are scratching their heads at people talking about the models getting dumber. Everyone was well aware that Opus started sucking when it had to compact context to keep under 200k.
I set up Opus as a strategic advisor for my Sonnet workflow. Here is the subagent config that makes it work. (www.reddit.com) Anthropic published the Advisor Strategy this week. The idea: a cheaper model does the actual work, a stronger model only gets consulted on hard decisions.
DeepSeek V4 reportedly drops late April. 1M context, multimodal, Claude-level coding. (www.reddit.com) Leaks point to late April release. Key specs 1M token context window Native multimodal (image/video input) Projected ~85% SWE-Bench Verified (ties or beats Claude Opus 4.6) Base model remains free.
"My parallel multi-model pipeline: Opus for planning, 3x Sonnet for content, 3x Haiku for search — what's your setup?" (www.reddit.com) "I've been running a parallel multi-model pipeline and curious what setups you all are using. My current workflow: Opus: Planning & high-level architecture Sonnet x3: Content generation (running 3 instances in parallel) Haiku x3: Search, v…
You know you have become a "Senior Vibe Coder" when you actually stop and think about which AI model to use for a specific task. (www.reddit.com) Junior vibe coder: Throws the entire codebase at whatever frontier model is trending this week and burns their API budget in 4 hours. Senior vibe coder: "I need Codex 5.3 for rapid scaffolding, Sonnet for the Tailwind components, and I'm s…
Built a Telegram remote for Claude Code - v2 is live, open source (www.reddit.com) Sharing what I built after migrating from OpenClaw to Claude Code. The first thing that really sucked was losing all remote access.
I built an interactive first-principles climate physics simulation with explainer (earth.crackalamoo.com via reddit) A 3D visualizer of earth's climate in the browser. Introduces physics step by step so you can watch each process unfold as a piece of the overall climate.
how can I deal with opus Hallucinations (www.reddit.com) yesterday, I tried to test it by sending him a 107-word paragraph ,i asked it to count how many words and the answer was 100, then I tell it "count again" and the answer was correct 107. but after it i ask, "Why are you hallucinating?" and…
Hey everyone, I just wanted to share an open-source Claude plugin I've been working on: claude-crap. (www.reddit.com) I’m a software engineer using Claude as a coding agent. I noticed that, especially on large projects, whenever it finished a feature, I always had to ask for an extra pass to fix code smells.
Output Styles aren't being injected into the system prompt (another degradation cause) (www.reddit.com) Found another cause of Claude Code degradation (and no, it's not an Opus 4.6 nerf this time either). Output Styles aren't being injected into the system prompt!
Built with Claude: Shipped a voice coaching app in one day (www.reddit.com) My partner is a brilliant engineer who can't do small talk. That's the whole origin story.
I got tired of setting up automations on zapier and n8n. So Claudes Agent SDK to do it for me. (www.reddit.com) I used the Anthropic Agent SDK and honestly, Opus 4.5 is insanely good at tool calling. Like, really good.
Why most open-source models can't answer this question while most closed-source models can answer most of the time? (www.reddit.com) WEB SEARCH WAS ALWAYS ON!!!! Question Calculate the precise VRAM requirement for the **KV Cache only** at the maximum context window for **DeepSeek V3.2** and **MiniMax M2.5**.
Best model / settings for low and slow high quality code? (www.reddit.com) Hey all - I’ve built a nice backlog of issues to fix in GitHub and I’m wondering your take on which model is the highest quality per token usage, not caring about speed. I want to task an agent to go through my backlog and fix them one by…
Stop donating your salary to OpenAI: Why Minimax M2.5 is making GPT-5.2 Thinking look like an overpriced dinosaur for coding plans. (www.reddit.com) ↯ Hallucination↯ Glm↯ Minimax↯ Swe Benchswe-benchminimaxaltman+5