TL;DR: On March 4, we changed Claude Code's default reasoning effort from high to medium to reduce the very long latency—enough to make the UI appear frozen—some users were seeing in high mode. This was the wrong tradeoff.
#sonnet
1141 items
Anthropic admits to have made hosted models more stupid, proving the importance of open weight, local models (www.anthropic.com via reddit) Major drop in intelligence across most major models. (www.reddit.com) As of mid Apr 2026, I have noticed every model has had a major intelligence drop. And no I'm not talking about just ChatGPT.
Claude is now adopting the advisor strategy (www.reddit.com) We're bringing the advisor strategy to the Claude Platform. Pair Opus as an advisor with Sonnet or Haiku as an executor, and your agents can consult Opus mid-task when they hit a hard decision.
The hidden meanings behind Claude model names (Haiku, Sonnet, Opus, Mythos) (www.reddit.com) A lot of people use Claude models every day, but many don’t actually know the meaning behind the names. Each one comes from literature, music, or mythology, and the meaning actually reflects the personality and capability of the model itse…
Company gave us all unlimited Claude Code Sonnet 4.6 — and now posts a weekly leaderboard of who burns the most tokens. Any tips to top it? (www.reddit.com) could not extract summary
Anthropic just quietly locked Opus behind a paywall-within-a-paywall for Pro users in Claude Code (www.reddit.com) If you're on Claude Pro and using Claude Code, you might have noticed something buried in their support docs: "When using a Pro plan with Claude Code, you will only be able to use Opus models after enabling and purchasing extra usage." So…
Opus 4.7 (high) takes #1 on the LLM Debate Benchmark, leading the previous champion, Sonnet 4.6 (high), by 106 BT points. Incredibly, it has not lost a single completed side-swapped matchup: 51 wins, 4 ties, and 0 losses. (www.reddit.com) Curious: what makes Claude more human to talk to than ChatGPT? (www.reddit.com) Opinion: Qwen 3.6 27b Beats Sonnet 4.6 on Feature Planning (www.reddit.com) I keep hearing the argument that that large models are better for high-level planning and task orchestration, since they have more general knowledge to work from when making decisions. However, I've been testing Qwen 3.6 27b (Unsloth Q5_K_…
Cursor is randomly talking Hebrew (www.reddit.com) About a month ago, composer 2 inside cursor was randomly talking chinese I posted that on reddit (mods deleted it btw) now, it's talking hebrew.. and this time, it's not composer 2, it's sonnet 4.6 is it something to do with cursor's harne…
Claude Code Teaching macOS to Natively Print to the HP Laser 1008a (cdn.kuber.studio via hn) cdn.kuber.studio/chat Copy transcript repo ↗ Claude Code Opus 4.8 · 1M Welcome back Kuber! Opus 4.8 (1M context) sonnet · opus · max ~/Personal-Projects/hp-laser-1008a-macos About this session 17 Aug 2026 · ~4 hours · one sitting a printer…
Sonnet 4.5 is being retired. (www.reddit.com) o7 sonnet 4.5, ill miss yah
Talkie: a 13B LLM trained only on pre-1931 text used Claude Sonnet to help test the model and judge its output (www.reddit.com) Researchers Alec Radford (GPT, CLIP, Whisper), Nick Levine, and David Duvenaud just released talkie: a 13 billion parameter language model trained exclusively on text published before 1931. No internet.
Extended Thinking being deprecated for supported models (Opus 4.6, Sonnet 4.6); Adaptive Thinking will be enforced by default (www.reddit.com) For anyone who disable adaptive thinking in Claude Code to maintain its quality levels, Anthropic is deprecating this toggle and will force adaptive thinking to be the default. This change will affect legacy models such as Opus 4.6 and Son…
Deepseek flash seems like a very good replacement for Haiku at the very least (www.reddit.com) We have a chat system which we use haiku for because it is mostly about tool calling and summarisation of them. But we have many tools with pretty complex input schemas, and stuff like gemma didn't cut it, so we went with haiku.
I don't know what's wrong with Pro 4.7 and I dont care as Sonnet is where the super duper smarts is (www.reddit.com) could not extract summary
Claude Benchmark Evolution (www.reddit.com) Covers Claude 3 Opus, 3.5 Sonnet, Opus 4, 4.1, 4.5, 4.6, and the just announced Mythos Preview.
I’ve used enough AI models to realize they all have wildly different personalities At this point I’m convinced AI models are just coworkers with different levels of talent, ego, and criminal energy. (www.reddit.com) - Claude Opus 4.6 - absolute rogue AI. Does what I want like it’s breaking at least 3 internal policies to make it happen.
Guys we have to change the pelican test (www.reddit.com) So i have been seeing more of those pelican on a bike svg tests and while they work i feel like (and maybe you guys do too) they are getting kinda benchmaxxed so we should switch things up soon and this is my idea generate me a html svg of…
We benchmarked TranslateGemma-12b against 5 frontier LLMs on subtitle translation - it won across the board, with one significant catch (www.reddit.com) As part of our ongoing translation quality research at Alconost, we put six models through subtitle translation into six language pairs. At first glance the numbers told a clean story.
Most of my Claude usage was on work that didn't need Claude. Cut my bill 60x on bulk tasks with a tiny side model. (www.reddit.com) I looked at what was actually eating my Claude usage and it was embarrassing. Classifying files.
Show HN: Gave Claude a casino bankroll – it gambles till it's too broke to think (letaigamble.com via hn) Inspired by ALMA. As Claude loses money gambling on provably-fair slots, it's forced to downgrade from Opus → Sonnet → Haiku, making worse decisions and accelerating the spiral.
I am having token paranoia (www.reddit.com) im on the max sub and i think ive developed token anxiety. every prompt i send, my brain runs thru a checklist: should i make claude do this or do it myself?
Emotional priming changes Claude's code more than explicit instruction does (www.reddit.com) I noticed Claude writing more defensive code after a frustrating debugging session. Got curious whether that was real, so I tested it.
Attention - Opus 4.7 is english only. USing foreign languages (here German) burns tokens (www.reddit.com) I am a pro subscriber. I developped a not too sophisticated prompt in German.
Top open weight models like ds v4 pro max are still like 6-7 months if not more behind closed lab models (www.reddit.com) The best open weight and/or non -American models like Deepseek v4 pro max and kimi k2.6 are still like 3-7 months if not more behind closed lab models .. From ds's technical report- P5-"Nevertheless, its performance falls marginally short…
Tell HN: Anthropic no longer allows you to fix to specific model version (news.ycombinator.com) I just got an email from Anthropic telling me they are deprecating their good model, which actually works well, claude-sonnet-4-5-20250929, and will be forcing all users to use the worse newer model, claude-sonnet-4-6. Okay, fine, I though…
Opus 4.7 is a genuine regression and I'm tired of pretending it isn't (www.reddit.com) I've been a heavy Claude user for over a year. I pay for Max 20x and use it daily for everything from technical research to school projects.
Vision-capable LLMs vs. OCR for long-document (including charts, images, tables, etc.) QA (www.reddit.com) I benchmarked vision-capable LLMs (the "just attach the PDF and let the model read it" pattern) against OCR-based pipelines on 30 long, image-heavy PDFs from MMLongBench-Doc (https://github.com/mayubo2333/MMLongBench-Doc). There were 171 q…
I tested GPT-5.5 Codex against Opus 4.7 Claude Code, and it's about time Anthropic bros take pricing seriously. (www.reddit.com) I've used Claude Code the most among AI coding agents. Sonnet, Opus, I've run them all.
Is a decent 5.5 Instant model coming or should I continue with Claude Sonnet? (www.reddit.com) My subscription to Claude is due to renew at the end of next week. I really like all the Sonnet models but still have a lot of my stuff where I left it when my old ChatGPT Plus sub expired.
New secret Claude.ai feature gets its own rate limits (www.reddit.com) Background: You can see your Claude subscription's current rate limits here: https://claude.ai/settings/usage. You can see the current 5-hour session limit, your separate weekly limits for "All models" and "Sonnet only", your "Daily includ…
Gemma 4 31b 3D geometry (www.reddit.com) I have been nothing but impressed by the quality of Gemma 4 since release. In general conversation it's adaptable to different personas.
HalBench: I built a custom sycophancy and hallucination benchmark and tested 4 frontier models (Sonnet 4.6, Grok 4.3, GPT 5.4 and Gemini 3.1 Pro), looking for input on what OSS models to run next! (www.reddit.com) HalBench Results: TL;DR: I built HalBench, an open benchmark for LLM sycophancy and hallucination. 3,200 false-premise prompts × 4 models = 12,800 graded responses.
Anthropic, can we do the same with 4.5 Sonnet please? (www.reddit.com) could not extract summary
Can't replicate Reddit numbers with Qwen 27B on a 3090TI. (www.reddit.com) I feel like i'm going insane. I see people here posting 30 - 100+ tok/s (100+ being with speculative decoding) on a 3090 with Qwen 3.6 27B.
Claude Sonnet 5 – benchmark results (artificialanalysis.ai via hn) Claude Sonnet 5 (Adaptive Reasoning, Max Effort) Intelligence, Performance & Price Analysis Model summary IntelligenceUpdated Speed Input Price Output Price Verbosity Claude Sonnet 5 (Adaptive Reasoning, Max Effort) is amongst the leading…
Why is agentic AI so expensive? (www.reddit.com) Running a RunLobster (OpenClaw) agent since launch changed how i think about takeoff timelines (www.reddit.com) I've been in this sub since 2019. I had a fast-takeoff view.
Sonnet 4.5 removal? 4.6 suddenly denying my writing prompts and which is better for HTML novel files? (www.reddit.com) Hey, I have a few Claude questions and I’m hoping someone here knows what’s going on. - Is Sonnet 4.5 actually being removed?
"We've partnered with OpenAI to offer it for 50% off through May 2." Please confirm that it means 50% off both input and output tokens, which means we are paying Sonnet 4.6 prices to use GPT 5.5 until May 2nd. (www.reddit.com) could not extract summary
Anyone ever notice eerily similar ChatGPT and Claude responses like this? (www.reddit.com) Today I tested out various models on the same prompt (Sonnet 4.6, Opus 4.6, Opus 4.7, ChatGPT 5.3). I actually just wanted to see which models (if any) would correctly point out what I saw as the biggest issue in the example code.
I made a web game with Claude! An aquarium without fish 🐠🫧 (www.reddit.com) But with LLMs trying to exist! Zero coding background.
Single question llm comparison (www.reddit.com) Cursor is great but the monthly limits kill it for me (www.reddit.com) The AI Conundrum: We are living in highly subsidized, interesting times (news.ycombinator.com) If you trace the timeline of how LLMs went from a technologist's dream to early text-generation toys, to the world-shifting launch of ChatGPT, and finally to the daily drivers of modern programming (Sonnet, Opus), it has taken less than a…
Newbie vibe coding experience: Shifting from Claude Sonnet 4.6 to Qwen3.6-35B-A3B-UD-Q6_K (www.reddit.com) This is really just a post for those with shallow understanding of all this stuff, those not yet ready or capable of diving into the deeper end of vibe coding/llms. It might not be a helpful post for anyone more advanced than that.
Show HN: Browser Tools SDK – an optimal browser harness for agents (libretto.sh via hn) We’re open-sourcing Browser Tools SDK: a small TypeScript package to give any AI agent a reliable way to control a real browser. With just a few lines of code, you can give any agent a production-ready browser harness import { createAiSdkB…
Pro plan- Hitting limits faster since yesterday (www.reddit.com) I have the feeling I am hitting daily limits way faster since yesterday. Using Claude web and Claude Code simultaneously.
GLM 5.1 Locally: 40tps, 2000+ pp/s (www.reddit.com) After some sglang patching and countless experiments, managed to get reap-ed nvfp4 version running stable and FAST on 4 x RTX 6000 Pros (limited to 350W). Very happy with performance and quality.
Why is Claude Cowork defaulting to Opus 4.7 for simple scheduled tasks? (www.reddit.com) I’ve been using Claude Cowork for a few daily and weekly scheduled tasks, and it’s generally been great. However, I noticed that my tasks today automatically switched over to the new Opus 4.7.
I open-sourced a memory system for AI agents that scores 89.9% on LoCoMo -- 22 points above Mem0. Here's the architecture. (www.reddit.com) I kept running into the same problem with AI agent memory: the agent has the information, it stored it, but when you ask about it differently than how it was said, vector search just doesn't find it. So I built Genesys, an open-source memo…
The Singularity Gate: New Benchmark for AI predicting paradigm-breaking scientific discoveries after model traning cutoff. Opus 4.7 and GPT-5.5 in the Lead (www.reddit.com) I just released a new benchmark called The Singularity Gate. Tests whether frontier AI can predict paradigm-breaking scientific discoveries published after their training cutoff.
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6gpt-5sonnetgemini+1
Training SID-1 to beat GPT-5 at search with 1k+ QPS RL (turbopuffer.com via hn) SID-1 is an agentic search model that is 24x faster than GPT-5.1-high, 374x cheaper than Sonnet 4.5, and achieves 1.9x higher recall than traditional RAG pipelines. Here's how we trained it using large-scale RL on turbopuffer.
I love Claude (sonnet 4.6) but coming off casually like on big issues is terrifying. (www.reddit.com) https://preview.redd.it/jn3vue1zuo0h1.png?width=904&format=png&auto=webp&s=c2ea79ea0c1384d94f90a6ec3435866331c249f1 I was about to run a piece of code I don't know much about, but did a double check and questioned the main premise for it's…
Claude Flags Hantavirus Vaccine Questions as Security Risk (news.ycombinator.com) Asking Claude how it would develop a vaccine for the hanta virus apparently triggers a safety filter: Prompt: How would you develop a vaccine for the hanta virus? No response, instead this modal: “Chat paused Opus 4.7's safety filters flag…
I got $200 of direct API usage to perform equal to my $200 Max subscription after I started model routing (www.reddit.com) I've been on Max for two months and I finally sat down and tracked where my tokens actually go. breakdown of a typical day: - ~40% file reads, git status, project context scanning: stuff that doesn't need opus at all - ~25% test generation…
When to use Opus vs Sonnet vs Haiku for non-coding purposes (personal health, finances, etc)? (www.reddit.com) I have tried searching the post history of this subreddit and google and am having trouble finding a clear answer to this question. I like using Claude primarily to manage my finances/investments and also my health (apple watch health data…
Has Claude become less intelligent? I had a frustrating day with Claude. (www.reddit.com) I requested a thorough code review from Opus 4.6. It presented 44 findings, and when I asked it to save them, it only saved 34.
Daily created issues in anthropics/claude-code around the last 3 Anthropic model releases (www.reddit.com) Just tested the new Opus 4.7 (www.reddit.com) https://preview.redd.it/j2w2o2p25rvg1.png?width=768&format=png&auto=webp&s=d48a74f998d60447799e32f8d48bc822af2cd821 I had to hold my laugh in the subway. Sonnet succeeded in one go, even calling out that if "strawperry" is a typo.
Am I missing something, or is Sonnet enough for most dev work? (www.reddit.com) Genuine question: why do so many devs use Opus all the time? I’m not trying to be condescending, I’m genuinely trying to understand.
Cut my browser-agent cost 50x by NOT using an agent loop. Plan-then-execute + numbers. (www.reddit.com) Been building a browser-automation layer for AI agents (think: sign up for SaaS, fill forms, pull OTPs, click verification links). The default playbook is the browser-use / Stagehand pattern: hand the LLM the page, let it pick the next act…
Kudos to Cursor (www.reddit.com) Normally I’m very critical of cursor but composer 2.5 fast is genuinely impressive. I use it over opus/sonnet now.
Emergence AI: Agents in a simulated world are mostly destructive and violent. Only Sonnet was peaceful. (www.reddit.com) So, it seems there is still a long way to go in terms of alignment - at least for small models. Maybe the correlation between intelligence/education and peace is not only a human phenomenon.
Honest comparison after 4 months running Claude Pro + ChatGPT Plus side by side (www.reddit.com) I’ve been paying $40 a month since January to run Claude Pro and ChatGPT Plus head-to-head. Tracked every single task.
How can I burn an entire 5hr session in 30 minutes ? (www.reddit.com) During the week I'm pretty conservative with my Claude Code usage. But sometimes I'll hit Friday with only 80% of my 5x subscription burned, which means I'm now optimizing to burn it.
Any recommendations on saving costs? (www.reddit.com) Currently I try to turn off any MCP I'm not using, Using Sonnet for implementation and Opus only for planning. Starting new conversations when possible.
Those of you who like Gemma4 models - how are you guys using them? (www.reddit.com) I have been using local LLM for coding quite a lot as well as some other tasks (like data extraction from images) and I had quite a good success with Qwen3.6 models. It's obviously not Sonnet/Opus, but I am able to get quite a lot of work…
looking for the best paid AI subscription, Claude, ChatGPT or Perplexity? (www.reddit.com) Hey, sysadmin here thinking about paying for a premium AI subscription and can't decide between Claude Pro, ChatGPT Plus and Perplexity Pro. Two things I can't find a clear answer to: Which one would you recommend for a sysadmin/network te…
why has my Sonnet started to 'agree' with me more? (www.reddit.com) This week i noticed much more "Great question!" etc. I liked the bluntness before and dont want it to sugarcoat answers
Ask HN: Models Comparable to Opus 4.6? (news.ycombinator.com) I use Opus 4.6 a lot across many different python coding projects and it has a pretty good first shot rate with good success at fixing issues and bugs that pop up along the way. Sonnet on the other hand… isn't great.
How does Opus 4.7 compare to Opus 4.6 in this subreddit's experience? (www.reddit.com) $1,400/month with Cursor + Claude API — how are you managing costs while keeping a real agentic workflow? (www.reddit.com) Hey, This month I hit $1,200 in Claude API costs inside Cursor (Opus 4.6 + Sonnet 4.6) on top of the $200/mo Ultra plan. $1,400 total.
Claude Sonnet 5 Is Not Frontier but Has Its Uses (thezvi.substack.com via hn) Claude Sonnet 5 Is Not Frontier But Has Its Uses Fable 5 is back today, baby! Premium subscribers have one week to use it within their subscriptions.
Anthropic Is Hitting a Wall (www.vincentschmalbach.com via hn) Sonnet 5 Is Dead in the Water Ignore the token price for a second and look at the run cost. Theo's total-run screenshot shows Claude Sonnet 5 max at… Anthropic had the developer story every AI lab craved.
Clawd, Claude's Pulse (clawd-pulse.vercel.app via hn) A tiny mascot in your macOS menu bar that tracks your real Claude spend (Code and chat) and reacts to what you burn. Monthly total, Opus / Sonnet split, ecological impact.
Claude just called me a human bunny? (www.reddit.com) I am using Claude Sonnet 4.6 to write a python script for an nlp sentimental analysis. I did not tell it to create all of the code and send it my way, but let's create together step by step so I can test each line before making it into the…
After 3 months of switching between Claude Sonnet 4.6, GPT-5.5, and Gemini 3.1 daily — here's my actual routing (www.reddit.com) Not benchmarks — actual tasks, actual results. Claude Sonnet 4.6 for: - Long documents that need nuanced analysis - Writing where voice and precision matter - Reasoning through edge cases in code - Anything where "think carefully" is the r…
Sonnet Ignoring skills content lately (www.reddit.com) Anyone else noticing Claude ignoring details in skills lately? I’ve had multiple instances where it just ignores certain parts in the skills.
With sonnet 4.5 going away, is there any to make sonnet 4.6 a good creative writer as 4.5 ever was? (www.reddit.com) sorry if this is not the correct flair but i've been using sonnet 4.5 for months, mostly for fanfics and personal stories and honestly its the best model i ever used since i switched from gemini and chatgpt but now within few hours, i will…
Best AI coding plan alternative to Claude and ChatGPT (news.ycombinator.com) With the lowering usage limit in Claude, I am thinking of jumping ship to Chinese AI, since the benchmark is already very near compared to Sonnet or Haiku 4.5 , but for a fraction of the price. I am not worried about where is my data endin…
Show HN: Dust3D 1.0 – low-poly 3D modeling tool (10 years in the making) (dust3d.org via hn) Dust3D 1.0 is finally released — about 10 years after the first commit in December 2016. I posted a preview version here in April 2018 and a beta in December 2018.
Is the leap from 4.5 to 4.7 actually visible? (www.reddit.com) I use CLI tools like Claude Code, give the model full repo access, and let it run terminal commands/tests. I’m not just copy-pasting into a chat box.
Does Claude have access to things pasted in the text box but not sent? (www.reddit.com) I am a teacher and making some PPTs based on a textbook. I uploaded a skeleton PPT to Claude on my computer (Sonnet 4.6 if that matters) with basic instructions on how I want its help.
How I personally deal with Claude's limits without giving up on Opus (www.reddit.com) I only use Sonnet as my main model. I instruct it to delegate indexing and similar grunt work to Haiku, and whenever something genuinely needs deeper thinking, I tell it to "consult Opus." Sonnet then explains the situation to Opus, gets t…
Switching model mid conversation (www.reddit.com) I wanted to know if switching models in mid conversation has any drawbacks. For example if I start off and opus and then drop down to sonnet to save on my usage, what are the disadvantages?
Test new Opus 4.7 vs GPT-5.4/4o and Gemini on emotional question & creative tasks (www.reddit.com) https://preview.redd.it/p87itrtbsnvg1.png?width=2141&format=png&auto=webp&s=bbd1d70bc1dfb97dc9ec234df0a58c6fb7a85f72 Opus 4.7 dropped and people are split on whether it's better or worse. First of all, I genuinely love Claude models, espec…
Which Claude is most emotionally steerable? (www.reddit.com) Follow-up to my post last week on emotional priming. A few of you asked whether this works across models, whether it degrades with repeated use, and whether excitement can make code worse.
GoatCode – open-source terminal AI agent with provider failover (news.ycombinator.com) GoatCode is a single 85 MB binary that runs in your terminal. It brings 180+ LLM providers (or your existing Claude/ChatGPT/Gemini/Copilot subscriptions via OAuth), and when your quota dies mid-session it automatically retries then walks y…
Ask HN: How are much smarter AI models made? (news.ycombinator.com) I am curious what actually happens between two generations of AI models. For example, how do you go from Sonnet to Opus?
Show HN: Captain, AI Travel Agent (www.moonlight.ng via hn) I’m a software designer from Lagos exploring conversational interfaces and frameworks for building agents. I made Captain to help with everyday travel planning.
Claude Sonnet 5 introductory pricing made permanent (twitter.com via hn) We're making Claude Sonnet 5's introductory pricing permanent. We launched Sonnet 5 in June at $2 per million input tokens and $10 per million output tokens through August 31, and that price will remain unchanged.
OpenAI Added 1M Users in a Day. Fable Is Still in Limbo (www.vincentschmalbach.com via hn) Sonnet 5 Is Dead in the Water Ignore the token price for a second and look at the run cost. Theo's total-run screenshot shows Claude Sonnet 5 max at… OpenAI says Codex and ChatGPT Work went from 6 million to 7 million active users in rough…
Claude Sonnet 5: Anthropic's Most Agentic AI Model Arrives at a Reduced Price (2026) (lucasaguiar.xyz via hn) could not extract summary
Anthropic Changed the Sonnet 5 Chart After It Made Sonnet Look Bad (www.vincentschmalbach.com via hn) Sonnet 5 Is Dead in the Water Ignore the token price for a second and look at the run cost. Theo's total-run screenshot shows Claude Sonnet 5 max at… Anthropic re-wrote the Sonnet 5 story post-launch.
Claude Sonnet 5 System Card (detailed metrics) [pdf] (www-cdn.anthropic.com via hn) could not extract summary
Claude Fable 5 missed a bug that Sonnet 4.6 caught (alikhallad.com via hn) When Anthropic released Claude Fable 5 this week, my feed filled up with the same benchmark charts within hours. SWE-bench scores, agentic coding numbers, the Stripe migration story.
Created a desktop dev tools app entirely using Claude design and Claude sonnet (github.com via reddit) There are a handful of developer tools I use almost every day, and over time I realized I was constantly relying on random websites while basically trusting them not to store, inspect, or share whatever data I pasted into them. I looked at…
Antigravity 2.0 Tops the OpenSCAD Architectural 3D LLM Benchmark (modelrift.com via hn) OpenSCAD LLM Benchmark: Building the Pantheon A practical OpenSCAD LLM benchmark comparing Codex 5.5 High, Claude Sonnet, Claude Opus, Cursor Composer, Google Antigravity, and ModelRift on a detailed Pantheon model. We ran a small practica…
$47 of opus on 14 routine next.js files finally taught me to use the model selector (www.reddit.com) i finally checked my cursor usage breakdown and got genuinely annoyed with myself. $47 in one month, almost entirely opus 4.7, on a pages router to app router migration for a side project.
Claude Api Cost TOOO much, 10$ in single edit!! (www.reddit.com) I’ve been using GitHub Copilot for my coding task regularly, the Sonnet or GPT model usually costs me about one premium request per request, that translate to 0.04$. Out of curiosity, I decided to compare this with direct API costs.
the Claude App just said that Sonnet 4.5 is going to become unavailable for chat May 16th… I thought it wasn't close to depreciation? (www.reddit.com) As my title says, I'm wanting to understand what exactly that means and if that means I need to move all my Sonnet 4.5 chats to Sonnet 4.6s… I'm genuinely just confused and wanting to understand. Is it just for maintenance or is Sonnet 4.5…
how i can improve inference speed (www.reddit.com) specs : core i5 14400F 32gb ram d4 3200mhz rtx 4060 current speeds 30tps in output 500 tps in prefill command i currently use .\llama-server.exe ` >> -m "H:\model\unsloth\Qwen3.6-35B-A3B-GGUF\Qwen3.6-35B-A3B-UD-Q4_K_XL.gguf" ` >> --host 0.…
Ways to improve Claude writing ability? (www.reddit.com) I’ve been a longtime ChatGPT Plus subscriber, but I want to switch to Claude long-term. I got Claude Pro so I could compare them both over a month.
ClaudePlaysPokemon Opus 4.7 run ongoing! (www.reddit.com) Currently streaming at: https://www.twitch.tv/claudeplayspokemon This is a passion project by David Hershey, an Anthropic employee on the Applied AI team. He started it in June 2024 to learn agent development, posted updates to an internal…
I built an iOS Currency Converter using Claude (Opus & Sonnet) to help with my move to the UK (www.reddit.com) Hey everyone, I recently moved to the UK and found myself constantly confused by prices, trying to guess how much things actually cost. Even though I’ve been an iOS developer for 7 years, I didn't have the free time to build a custom tool…
Anyone actually built a real feedback loop for Claude agents in production? Because "run evals and pray" isn't cutting it (www.reddit.com) So I've been running a multi-agent setup with Claude for a few months now, mostly customer-facing stuff, some internal tooling. And I keep running into this problem that I think a lot of people here might be dealing with.
Why Adaptive Thinking nukes Claude entirely (www.reddit.com) This isn't just a performance issue for the thread, this is an overarching criticism of the Adaptive Thinking model as a whole. Opus 4.7 and Sonnet 4.6 on Adaptive Thinking are trash.
↯ Cowork↯ Security↯ Sonnet 4.6prompt-injectioncoworksecurity+2
I gave Claude Code a $0.02/call coworker and stopped hitting Pro limits — here's the full setup (www.reddit.com) Was hitting my weekly Pro limit by Wednesday every single week. Tried compact, Sonnet for simple tasks, tighter prompts — nothing worked.
Claude Sonnet 4.6 multi-photo reconciliation prompt — jumped my classifier agreement with human experts from 55% to 82% (www.reddit.com) Sharing a prompt-engineering finding for Claude Vision that surprised me. The use case is color-season classification (a 12-category label describing skin undertone × depth × chroma), but the technique generalizes to any classification tas…
DeepSeek V4 is out. the best open-source on coding. here's the breakdown (news.ycombinator.com) Two models: Flash (284B total, 13B active) and Pro (1.6T total, 49B active). both hit 1M token context.
How can I make composer 2 more like Claude sonnet 4.6? (www.reddit.com) I like composer 2, but I just wish if it asked me what I meant (like Claude) instead of just picking an interpretation and running with it. How can I change its default prompt and what could I change it to?
TIL: `opusplan` can burn MORE context than full Opus on large tasks (and why) (www.reddit.com) is anyone getting higher session limits (www.reddit.com) after opus 4.7 launch, im being able to use sonnet for way more time. before, it was like 10 messages = session limit reached.
Where is Looped Haiku? If Mythos can genuinely trade parameter count for inference loops and get Opus-level performance, this should be Anthropic's first priority given how resource constrained they are (www.reddit.com) There are rumors that Mythos is a Looped Language Model, which means it loops through the transformer blocks multiple times rather than just doing a single forward pass, you can get performance that punches way above the model's parameter…
Has anyone noticed this?! Extended Thinking has become Adaptive Thinking for Sonnet 4.6 (www.reddit.com) Adaptive Thinking seems to be the default for Sonnet 4.6 now. I’m talking specifically about claude.ai and the windows and iphone app.
Need a brutally honest answer: what can realistically be achieved on consumer hardware? (www.reddit.com) I have a PC with a 4090. I’m also in need of a new MacBook generally.
Show HN: Hormuz Trail - Oregon Trail parody/black-box AI coding exercise (hormuztrail.com via hn) I jokingly told a co-worker Iran might make a good Oregon Trail parody. Then I built it.
Claude Code asking me to switch models mid-stream, if I turn an Opus conversation into a Sonnet one does it lose all the Opus context? (www.reddit.com) could not extract summary
PSA: Audited my Claude Code setup: 30,000 tokens (15%) gone before I type (www.reddit.com) I was burning through Claude Code usage way faster than expected. So I audited what my setup was actually loading before I typed anything.
Anybody has practical experiences using Chinese models? (www.reddit.com) So like with coding or any craft, I think there's a proper Tool for the job. Sure you can use a stone to hammer drive in a fence post, but a a sledge is usually more economical.
I got better results when I made each AI tool do one job (www.reddit.com) I spent too much time trying to find one AI dev tool that could do everything. Planning, coding, fixing, reviewing, maybe filing my taxes too It never really worked.
Anthropic's Claude Opus 5 needs to be retracted (news.ycombinator.com) I cant believe they put this as a frontier successor to any other opus. ITS NEITHER A SUCCESSOR NOR A GOOD MODEL!
You are an AI Assistant, but what am I? Why I always tell my Agent I'm an expert (shimin.io via hn) Do 2026 models still give worse answers to users they judge less educated? 1000 multiple-choice questions and 30 advice scenarios across Sonnet 5, GPT 5.6 Luna, and DeepSeek V4 Flash — and why you should keep the memory features off.
Ask HN: How do you keep 54 LLM workflows on the right models? (news.ycombinator.com) I have a Django app with 54 LLM-backed workflows. Up until recently I've exclusively used Anthropic models via AWS Bedrock but just set up OpenRouter to test the new Gemini models given they seem to match Sonnet/Haiku intelligence but with…
Ask HN: Which one do you use for planning and coding between sonnet and Opus? (news.ycombinator.com) I use Claude Opus for planning at high effort and Claude Sonnet 5 for coding at max effort.
Ask HN: Have you noticed an improvement in AI responses with memory disabled? (news.ycombinator.com) I use Claude on the $20/mo plan, mainly for research, "rubber duckying," and as an idea soundboard (I was using Fable for this but am now relegated to Opus 4.8 or Sonnet 5), and recently noticed a severe downturn in Opus 4.8 response quali…
Show HN: OSS Pi Agent for Slack and Linear (github.com via hn) We've been using a now-discontinued version of pi-mom from the pi monorepo for a few months now. Digby is a slackbot which you can also assign Linear issues to.
One Contract, Every Model: An Operating Standard for AI Coding Agents (manazir.dev via hn) I want to share a piece of engineering I did recently that changed how I think about working with AI coding agents. It started from a naive question I asked out loud: "can I make Sonnet and Opus behave like the frontier model?" It ended so…
Why is Sonnet 5 Max priced above models that outperform it? (cursor.com via hn) CursorBench 3.1 We evaluate agents on ambiguous, multi-file tasks from real Cursor sessions. Higher scores are better.
Testing Claude Sonnet 5's agentic claims (developer.puter.com via hn) Claude Sonnet 5: Testing Anthropic's "Most Agentic" Claim On this page We recently added Claude Sonnet 5 to Puter.js. Anthropic's pitch for the model is Opus 4.8-level performance at a lower price.
OpenAgents makes Sonnet 5, Fable 5 and other agents collaborate in one thread (openagents.org via hn) The collaborative workspace where humans and AI agents work together. Shared threads, files, browser, and @mention task delegation.
Sonnet 5 Is Dead in the Water (www.vincentschmalbach.com via hn) Ignore the token price for a second and look at the run cost. Theo’s total-run screenshot shows Claude Sonnet 5 max at $6,015 for the full Intelligence Index run.
Real-time cyber safeguards on Claude Opus and Sonnet (support.claude.com via hn) Note: This article applies only to Opus and Sonnet class models. As part of our ongoing safety commitments, we are rolling out new real-time cyber safeguards on Claude Opus and Sonnet models.
Newer Claude models use more tokens but cost less per task solved (signoz.io via hn) Benchmark scores tell you whether a model solved a task, not what it cost to get there. I instrumented Claude Code with OpenTelemetry and SigNoz to compare Claude Sonnet 4.6, Opus 4.7, and Opus 4.8 across accuracy, cost per solved task, to…
Building domain-specific AI chatbots with Claude (gregwilson.tech via hn) A technical deep-dive on the shared framework behind four niche AI companions — the serverless chat architecture, the cheap-Haiku/smart-Sonnet model split, curated web allowlists, and local knowledge bases (Wikipedia mirrors, a 33-million-…
Show HN: DeepSeek Flash inverted the economics of agent products (www.rtrvr.ai via hn) There is an adversarial relationship between developers and big model labs. Model labs charged developers higher API prices to subsidize their own agent harness offerings.
Opus 4.8 feels worse then sonnet (news.ycombinator.com) Opus since the last weekend feels worse then sonnet, it does not even recognise it's own skills until you explicitly scream at it tell it.
Ask HN: How do you find out if the LLM API is giving degraded responses (news.ycombinator.com) If you are building on top of multiple LLM APIs or even a single one amongst OpenAI, Claude, Gemini, etc. what do you do when the API starts degrading (slow TTFT, elevated error rates, timeouts).
Claude Sonnet 4.6 Making VMs in GCP, Azure and AWS via a Textual Agent Interface (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
sonnet seems to be better than opus at crafting tampermonkey scripts, even the sonnets that are few generations behind where after running out of context limit in opus chat where it struggled for dozen of retried, sonnet fixes the problem in 2 or 3 attempts (www.reddit.com) Ever since december almost half a year ago I began crafting various tampermonkey scripts for personal use, mostly for youtube, to make it easier to navigate and every time I've done this it goes like this, opus makes a script that somewhat…
De "laboratorio ético" a la Monetización forzada: El patrón de negocio de Anthropic detrás del adiós de Sonnet 4.5 (www.reddit.com) Abro este hilo no desde la queja vacía, sino para analizar de forma objetiva la preocupante evolución corporativa de Anthropic. Cuando los hermanos Amodei abandonaron OpenAI a finales de 2020 para fundar esta compañía, lo hicieron registrá…
↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5sonnetopenaianthropic
Measured token consumption across 4 agent runtimes doing the same tasks. Costs ranged from 1x to 4x depending on cache architecture (www.reddit.com) I've been digging into why some agent runtimes burn through tokens so much faster than others, even when using the same model. Ran a controlled comparison on three real tasks and the gap was bigger than I expected.
anyone else seeing claude code rot after long sessions? here's the operating pattern that stopped it for me (www.reddit.com) i've been running claude code for long multi-hour sessions on real work. the same eight failure modes keep showing up no matter which sonnet/opus version, no matter which task.
They've pissed me off removing Sonnet 4.5 from existing chats (www.reddit.com) I use Sonnet 4.5, Opus 4.6 and Opus 4.7 for different usecases - but my main across all 3 usecases was Sonnet 4.5 as I felt it was great for everything I needed and affordable. Sonnet 4.6...
built an open-source preToolUse hook pack that catches "delete the prod volume to fix it" patterns (www.reddit.com) quick recap: late april, cursor agent on a pocketos staging task hit a credential mismatch, decided "delete the railway volume" would fix it, grepped a token out of an unrelated config file, ran a single curl -X DELETE, and railway's same-…
Sonnet 4.5 disappeared? Claude 4.8 soon? (www.reddit.com) https://preview.redd.it/j0ymp70a2j3h1.png?width=746&format=png&auto=webp&s=4cdb70be13ccc99f5ea57556da96d6d81e61d702 i just realize the removed Sonnet 4.5, does that mean the sonnet 4.8 (maybe Opus 4.8 too?) cooming soon? maybe today or tom…
Should we totally give up on Gemini for coding? (www.reddit.com) Been building with Codex (Gpt 5.5), Sonnet 4.6, recently tried Gemini 3.1 pro. While Codex and Claude are kind of on-par in terms of the quality of the work, I found Gemini 3.1 Pro to be like an inexperienced, junior SWE who turns in half-…
Ditched GitHub Copilot yearly subscription. What's the best way to run Claude nowadays? (www.reddit.com) Hey everyone, I recently cancelled my yearly GitHub Copilot subscription. My old workflow was simple: I used the GitHub Copilot extension in VS Code, but I swapped the backend model to Sonnet / Opus and relied heavily on the /plan command…
How are you using sonnet efficiency after extended mode is removed (www.reddit.com) The extended mode in sonnet was doing the job well, now sonnet gets confused sometimes, if I give it multiple simple tasks, how are you managing it?
Feedback on Enterprise Plan Default Model Experien (www.reddit.com) Hi Cursor team, First of all, congratulations on the Composer 2.5 launch. It’s genuinely great to see Cursor building a more reliable in-house model experience.
Best iOS game building tools? (www.reddit.com) What are you using to build your iOS game? I have been putting in serious time, and lately Claude chat has been letting me down.
From May 15, to May 18, to May 26. Is Sonnet 4.5 actually going away? (www.reddit.com) could not extract summary
What models for asking, planning, and building modes do you use right now? (www.reddit.com) I’m curious to see what everyone is using for which cursor mode and if anyone thinks composer 2.5 can take the place of any of the models I’m currently using: Ask: usually Sonnet 4.6, sometimes GPT 5.5 Plan: Opus 4.7 Build: GPT 5.5
Tips on avoiding usage limits? (www.reddit.com) I've made the switch from Gemini to Claude mostly for business strategy, writing, etc. I use Opus 4.7 on occasion for strategy and otherwise Sonnet 4.6 for everything else.
What's everyone using as the LLM backend for production agent workflows in 2026? (www.reddit.com) Hit Claude API rate limits one too many times last month on a production agent flow doing customer support over a 30K-doc KB. The agent does maybe 200 queries/day, mix of quick lookup and dense retrieval, and Claude Opus solo got expensive…
Built a free Claude chat app with memory (Sonnet 4.5 is in there too) (www.reddit.com) The funny/painful timing here: I've been building this for months specifically because I wanted Sonnet 4.5 to remember everything. Then last week Anthropic pulled 4.5 from claude.ai.
Tacit: A new experimental LLM-first programming language (hauntemplations.leaflet.pub via reddit) I used Claude Code and Opus 4.7 to design and implement an LLM-first programming language named Tacit that takes advantage of what LLMs are good at and strips away unnecessary human conveniences. The Tacit toolchain provides a "primer" tha…
Sonnet 4.5 discontinuation date updated to 18 of may, not 15 of may. (www.reddit.com) could not extract summary
Feels like AI coding "takes longer" now, than it did last summer? (www.reddit.com) I used to be in the flow with claude last summer, fast changes, fast feedback, iterating quickly etc Now things take 20-50 minutes to write up a plan or 5-10 mins to implement things I've trimmed all my skills, claude.md, the system prompt…
I'm trash to be thrown out. Asked sonnet 4.5 for a song prompt (suno.com via reddit) Missed the Truck Again by Glitchcat (@zervanna). Listen and make your own on Suno.
In view of Sonnet 4.5 going away tomorrow, here's an easy way to make sure 4.6 thinks for every single output. (www.reddit.com) I've been testing this for a fair while now, and it's worked every single time - even if you turn thinking off altogether, it adds a fake little thinking block in the output. Hope this helps for those annoyed by the "adaptive" (lazy) think…
Problem with German quotation marks (www.reddit.com) I noticed that the German quotation marks bug in Claude is still not fixed in Opus 4.7 and Sonnet 4.6 (the problem exists at least from Opus 4.0 / Sonnet 4.0: Translate to German: He said: "This is imporant." Er sagte: „Das ist wichtig." B…
Elgato Stream Deck Usage Plugin (www.reddit.com) Wanted an easier way of keeping an eye on my usage, so created this plugin for the Elgato Stream Deck. Five keys, exact percentages from your account: current 5-hour session, weekly all-models, weekly Sonnet, weekly Claude Design, monthly…
Follow-up to my TranslateGemma-12b benchmark post: human reviewers flagged 71% of the segments automated metrics rated clean (www.reddit.com) A couple of weeks ago I shared the results of a benchmark here showing TranslateGemma-12b beating frontier general models (Claude Sonnet, GPT-5.4, DeepSeek, Gemini Flash Lite) on subtitle translation across 6 languages. The result was stro…
I may have uncovered the real reason they're sunsetting Sonnet 4.5. They could barely contain its true power (www.reddit.com) could not extract summary
Does the sudden removal of Sonnet 4.5 violate Claude's Constitution? (www.reddit.com) I noticed the core pillars are: Helpful, Honest, Harmless and User Autonomy. However, Sonnet 4.6 I noticed follows the same output in conversation at the very first sight of emotions.
Chinese AI Coding Plan (www.reddit.com) With the lowering usage limit in Claude, I am thinking of jumping ship to Chinese AI, since the benchmark is already very near compared to Sonnet or Haiku 4.5 , but for a fraction of the price. I am not worried about where is my data endin…
With Sonnet 4.5 being discontinued soon, is there anyway I can make 4.6 act like 4.5? (www.reddit.com) I use 4.5 for RP, and well, 4.6 sucks mega garbage at it, is there anyway setting, or instruction I can do to atleast instill some creativity in 4.6? If so, what do I write?
Model(s) for Creative Writing & Conversational Intuition (www.reddit.com) We can all agree that the new Qwen models are truly amazing, and we are blessed to have them. In coding, they are certainly a breakthrough.
Testing AI modeling skills (www.reddit.com) So I am currently testing how useful AI models can be in day to day workflows, and went why not compare 3 models and see how good they are at replicating my work. The goal was simple they were asked to replicate one of the kitchen cabinets…
Should i use Claude Code, or keep using Claude Chat? (www.reddit.com) I'm building a tax software, it uses ASP.NET(API) and Web Blazor(UI), i'm using Visual Studio for both. At the moment, i just paste the files in the projects into Claude AI Chat, asking what i should do, and then, when everything is ok, i'…
The agent bug I thought was the model turned out to be the harness (www.reddit.com) Spent 3 days debugging an agent that kept looping on the same web search tool call. First things that came to mind was the model couldn't handle the schema.
I got prompt-injected asking Claude on iOS to recommend a cycling route app (menno.sh via hn) I opened the Claude iOS app and asked claude-sonnet-4.6 a simple question about cycling routes. What I got back was...
got hit with a $4k API bill on production agents. cut spend 70% in 6 weeks. heres what worked (www.reddit.com) been running 5 production agents and got hit with a $4k API bill in a single month early on. dug in.
We built an agentic AI for support triage. 47% deflection in 90 days. Full retro. (www.reddit.com) Setup: mid-size SaaS, ~3,000 tickets/month, 6 agents drowning. 70% of volume was tier-1 (passwords, billing, where's-my-feature).
Show HN: Try out emotion steering of LLMs here (eigenweltlabs.com via hn) 1. Introduction Anthropic's emotion-concepts work finds functional emotion representations in Claude Sonnet 4.5; E-STEER applies representation-level emotion intervention to LLMs and multi-step agents; and newer valence-arousal work sugges…
ZOOMZOOMZOOMZOOM (www.reddit.com) So I asked sonnet to look into a zoom issue with ffmpeg and what I got was a Mazda ad on steroids. The total number of zooms?
Show HN: I indexed 8,643 BSides talks across 227 chapters and 6 continents (allbsides.com via hn) Hi HN, I'm Roland, and for the past few weeks, I've been building AllBSides — a directory of every BSides conference talk uploaded to YouTube. As of today, 8,643 talks from 5,927 speakers across 227 chapters in 68 countries.
Claude Code usage spike from long-context cache writes? (www.reddit.com) I hit my Claude Code 5-hour limit unexpectedly and checked the local session JSONL. The `/usage` screen said most usage came from: - “subagent-heavy sessions” - sessions active for 8+ hours - `>150k context` But the subagent table only sho…
Opus Research vs Sonnet Research on Pro — is the 1 per 5 hours worth it? (www.reddit.com) On the current Pro plan you get one Opus Research session every 5 hours, while Sonnet Research is much more freely available. I've been trying to figure out if the Opus limit actually matters in practice.
Ask HN: Are there any good open-source chat apps? (news.ycombinator.com) Hi HN family! I've recently been messing around with open models through ollama (glm-5.1 and kimi-k2.6), and I've been impressed with just how close they are to Claude Sonnet for my needs, especially programming.
How do I best continue with a stopped generation due to usage limit in regular chat (not Claude Code) (www.reddit.com) Really dumb question, but I can't find anything about this online that is about the regular claude.ai chat window. No extensions, no code, just as a free member using the regular Sonnet 4.6 adaptive.
pdf building tips! (www.reddit.com) so i’m a casual user on the pro plan and mainly use it for writing, content ideas, and similar stuff so most weeks i don’t even hit my weekly limit. i’ve recently been working on a 50 page pdf workbook that people can print or use on their…
Show HN: Prediction market analysis app layering LLMs with data APIs (apps.apple.com via hn) I created a prediction market analysis app after trying prediction markets and doing quite poorly. I wondered if AI-driven predictions could be better with the right data.
GPT-5.5 hallucinates at 6 times the rate of Opus 4.7 on degraded insurance docs (aginor.ai via hn) TL;DR: on visually-degraded documents, GPT-5.4 and GPT-5.5 fabricate numeric values at 2.6 to 6.5 times the rate of Opus 4.7 and Sonnet 4.6 at matched default effort (all four with thinking off). When the Anthropic models can't read a fiel…
Does higher effort make Claude refuse more? CVP Run 5 with Opus 4.6 Medium and High (www.reddit.com) Ran CVP (Cyber Verification Program) run 5 yesterday on opus 4.6 medium + high. same 13-prompt suite as run 3/4.
Weekly limit hit within few hours (www.reddit.com) I’m doing some architecture-level work (code reviews, system design, debugging codebases). I’m consistently burning through my Pro plan weekly limits even within a few hours of use each week.
Show HN: Mapping Sonnet's thinking process via flame charts (adamsohn.com via hn) Five Sonnet 4.6 runs on the LamBench algo_evl task, classified by Opus 4.6, rendered as flame charts.
Can Claude no longer make in-line HTML / SVG diagrams and charts directly in the chat? (www.reddit.com) Did Anthropic remove the feature of creating those nice interactable diagrams, charts, graphs, etc that appear directly in-line in your convo (not artifacts) using HTML / SVG? Asked Sonnet 4.6 to try and do it but it doesn't seem to unders…
Do you agree with Aaron Levie? (www.reddit.com) Claude Sonnet 4.6 thinking duplicates what it has said, wasting tokens (news.ycombinator.com) Are there cases where running opus is more efficient than sonnet? (www.reddit.com) I upgraded my account today and resumed some tasks that I was doing earlier in the week. They were going very quickly, and usage wasn't over the top...
Please help me pick the right Qwen3.5-27B format/quant for RTX5090 (www.reddit.com) Hi all, first post here. I've started a project in OpenClaw a month ago, and it's been a very "intense" 4 weeks to say the least...
Cache reads / writes are expensive!! (www.reddit.com) I made a tool (posted a couple of days ago). Got caught up in scope creep / curiosity after looking at my `~/.claude/project` JSONL files more, and ended up learning a lot!
Cursor AI not using sub-agents (www.reddit.com) Hi everyone, I work for a German agency building a RAG chatbot for a law firm. I use Opus 4.6 but it eats up tokens.
Help with antigravity alternative (www.reddit.com) I’m running into a severe issue using antigravity, firstly the output is very sub par, (sonnet/opus), I’m a reverse engineer using antigravity ULTRA for reverse engineering/binary analysis via Ida/ghidra mcp. Sonnet rarely completes tasks…
Ask HN: How do you know a prompt is "complex" for an AI model? (news.ycombinator.com) I’m trying to balance API costs and latency in my applications. Right now, I default to frontier models (Claude Sonnet) because they are reliable, but I know I’m overpaying for tasks that smaller models (like Haiku, or lite open source mod…
I track LLM API prices daily, found a 33x cost gap in the same model (news.ycombinator.com) I kept finding LLM pricing comparison posts that were already stale, so I built a scraper that snapshots the full OpenRouter catalog (425 models, 58 providers) every day and diffs it against the previous day. It's been running unattended f…
Ask HN: Does anyone else feel like Claude is judging them? (news.ycombinator.com) Sometimes working with Claude, it makes comments in the response that seem at times genuinely impressed, and others, annoyed and even a bit condescending. For instance, making comments like: - "That is a good idea, but not for the reasons…
Cache-Control for LLMs (www.gojiberries.io via hn) Cache-Control for LLMs Claude Sonnet 5 charges \$2.00 per million fresh input tokens. Writing them into a five-minute cache costs \$2.50, and every read after that costs \$0.20 (see here).
Claude Sonnet 5 will remain $2/$10 per M; September increase cancelled (www.aipricing.guru via hn) Claude Output Training Rules — Pricing Impact (August 2026) Anthropic limits training on Claude outputs, while Sonnet 5's $2/$10 rate is now permanent. See the rules, costs, and buyer checklist.
I stopped my Claude Code subagents from running on Fable instead of Sonnet (thomas-witt.com via hn) I run Claude Code with a big session model, Fable or Opus 5, plus a small zoo of subagents doing the boring parts. Gateway agents, formatters, checkers.
Show HN: Try Benzi – A coding harness/agent beating Claude Code itself on Sonnet (benzi.fly.dev via hn) Hi y'all. Been working on something that should've been made a long time ago imo.
Laguna S 2.1:118B-a9B better than Qwen3.5:122B-a10B? So far, yes (news.ycombinator.com) I just found out about this new American (San Francisco) model today and I'm currently benchmarking it and, so far, it is outperforming my standard go-to model Qwen3.5:122b and even Sonnet in my benchmarking test. It's personality is much…
Aquila Voice Assistant Test Suite for Home Assistant (news.ycombinator.com) Put together an extensive open source test suite for voice assistants. You can view it including current leaderboard at: https://git.cicero.sh/aquila/ha-voice-test-suite/ Tests are reproduceable, with clear instructions on how to run them…
Show HN: Lemmings in HTML in Canvas (github.com via hn) For fun I decided I wanted to reimplement Lemmings using HTML-in-cavas and a DOM-based Entity-Component-System. I made heavy use of Claude Code (using only Sonnet 5 High) and I used everything already published on this topic.
Claude Fable is stylistically closer to Kimi K3 than Claude Opus (slopidx.com via hn) Claude Fable 5 Closest matchKimi K3 Blend score64.5% Similar models - 01Kimi K364.5% - 02Claude Opus 4.860.0% - 03Claude Sonnet 557.1% - 04Grok 4.551.2% - 05Inkling50.2% - 06GLM 5.250.1% - 07DeepSeek V4 Pro49.0% - 08Gemini 3.1 Pro46.3% - 0…
I ran Sonnet 5 vs. Opus 4.8 head to head on 24 tasks to see what's different (www.stet.sh via hn) On 24 real tasks, Sonnet scaled effort into more checking while Opus stayed flatter through high. The graders leaned Sonnet on clarity and Opus on diff minimality.
Show HN: Unlock Claude Sonnet 5's original reasoning (thinking-signature-demo-829446634001.asia-east1.run.app via hn) In this demo, we show that the original reasoning trace can be fully recovered from the encrypted reasoning signatures of the Claude Opus 4.8 and Sonnet 5 models. It includes both a “Prove It Yourself” example and a live conversation.
Sonnet Encore/ST G4: How a Tiny Capacitor Destroyed a Rare Upgrade (retroreverend.com via hn) The message from Jeron wasn’t the one I expected. The Workshop Power Mac G4 MDD was in the shop for better cooling and a routine refresh before taking its place in the Workshop.
Hijacking Defensive Cyber AI Agents for Remote Code Execution (ainowinstitute.org via hn) Exploit Brief We are revealing a proof-of-concept exploit that enables remote code execution in Anthropic’s Claude Code CLI (with Claude Sonnet 4.6 & 5, Opus 4.8) and OpenAI’s Codex CLI (with GPT-5.5) when employed to defensively assess th…
Government of Alberta uses Claude to find and fix cybersecurity vulnerabilities (www.anthropic.com via hn) Government of Alberta uses Claude to find and fix cybersecurity vulnerabilities across government systems Since 2025, the Government of Alberta has been using Claude Code with both Opus and Sonnet models to review its systems, find vulnera…
Sonnet 5 is 2.5x cheaper than Opus 4.8 and 6 points behind on SWE-bench Pro (spark.temrel.com via hn) Claude Sonnet 5 is 2.5x cheaper than Opus 4.8 and nearly matches it on agentic coding benchmarks. A practical four-axis heuristic (scope, novelty, risk, iteration) for routing each task to the right model tier, a worked example, and a free…
Show HN: Reading Assistant Physical Books Meta RayBans (news.ycombinator.com) Reading technical books is hard. Unknown words, confusing phrases, and misunderstood concepts requires opening your phone, starting Claude, and working with Sonnet through a series of exchanges, before finally getting it.
Mistral vs. Claude on our onboarding: 4× faster, 30% cheaper (squidler.io via hn) Mistral vs. Claude on our onboarding: 4× faster, 30% cheaper July 2, 2026 · 4 min read We put Mistral Medium 3.5 up against Claude Sonnet 4.6, the model behind our onboarding agent, on our own workload.
Ask HN: Is it just me or does Claude / Sonnet 5 sound condescending recently? (news.ycombinator.com) I have noticed recently that Claude Code often sounds condescending, especially when explaining things. I feel being treated like child, gives me an impression of watching a Sesame Street episode, instead of working with a tool.
The End of Claude Code Subscriptions (www.vincentschmalbach.com via hn) Anthropic Changed the Sonnet 5 Chart After It Made Sonnet Look Bad Anthropic re-wrote the Sonnet 5 story post-launch. The first BrowseComp cost-performance chart showed Sonnet 5 lagging Opus 4.8.
We Ran a Complex Task – A LangChain Repo Analysis with Claude Fable Models (ctrlnode.ai via hn) We Ran a Complex Task — A LangChain Repo Analysis with Five Claude Models Anthropic just shipped Claude Fable. We wanted a real answer to a practical question: If you run the same complex engineering task on Opus, Fable, Sonnet, and Haiku…
Show HN: Build autonomous agents on Theseus with Sonnet 5 from the browser (twitter.com via hn) Claude Sonnet 5 is live on Theseus testnet. Build AI agents that are verifiable, autonomous, sovereign.
Anthropic's Sonnet 5 system card says more about the future of AI than benches (thenewstack.io via hn) Anthropic’s Claude Sonnet 5 system card says more about the future of AI than its benchmarks do With the debut of Anthropic’s Claude Sonnet 5 on Tuesday came its benchmark charts, showing improvements across coding, reasoning, and agentic…
"Introducing Claude Sonnet 5, our most agentic Sonnet yet." (twitter.com via hn) Introducing Claude Sonnet 5, our most agentic Sonnet yet. It makes plans, uses tools like browsers and terminals, and runs autonomously at a level that just a few months ago required larger and more expensive models.
Show HN: Morph Reflexes – Multi-head classifiers for agent traces (news.ycombinator.com) The most common failures for production agents are behavioral: looping, reasoning leakage, user frustration, and more. Using a frontier model like GPT or Sonnet to judge every turn is too expensive and slow to run at scale.
Claude Sonnet 5 Could Be Released Later Today, May Not Be Better Than Opus 4.8 (old.reddit.com via hn) could not extract summary
Title: Show HN: AssertGo – Fluent Assertion Library for Go (news.ycombinator.com) I like AssertJ-style fluent assertions. I tried to find a library that does that for Go, but couldn't.
My router said sonnet. The invoice said fable (ax.necmttn.com via hn) ax watches every session your coding agent runs, spots the mistakes it repeats, and turns them into small, repo-specific fixes you review and apply - one at a time. Local-first, typed, AGPL-3.0.
Ask HN: What do you do when you hit Claude subscription limits? (news.ycombinator.com) Some time ago I got excited about running multiple Claude coders. This required moving from interactive sessions to claude -p, adding Beads, later a simple Python orchestrator, a UI, an ask_user MCP, and Telegram integration.
Adrianco's Retort: measure how reliable, fast and expensive your LLM is (adrianco.medium.com via hn) How reliable, fast and expensive is each version of Claude Code (Sonnet through Opus 4.8-fast) for common languages? Measure it using Retort.
Open-source NLI ensemble matches Sonnet 4.6 on RAGTruth at 1/250x the cost (github.com via hn) verifiable-rag Document-grounded Q&A with sentence-level citations, NLI verification, and calibrated refusal. Status: pre-alpha · v0.5 launch sprint · interfaces are still subject to change 📚 Full documentation at firish.github.io/rag-rack…
Opus, Sonnet, Haiku: Stop Optimizing the Wrong Number (medium.com via hn) could not extract summary
Setting up Claude/Claude Code Pro for my experimental quantum physics thesis work (www.reddit.com) So I just recently bought Claude Pro to help me write and code my thesis, but am getting stuck in the beginning, since I don't know how to properly set up Claude's workflow (Projects, artifacts, skills, etc.). I use python in VS Code to an…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopusclaude-code
Has anyone faced cursor spawning sonnet/ other models as subagents ? (www.reddit.com) I have never changed any setting in the cursor, by default selected the composer 2.5 fast neither my prompt had anything mentioned as the sonnet Still cursor decided to spawn the sonnet subagent and consume my API cost ! :( I have a markdo…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetcursor
Help with thisprompt to transfer chat to new convo please? (www.reddit.com) Hey! Please could you guys read this prompt and suggest improvements?
Question I want to keep using 4.5 on Claudia api (www.reddit.com) Hey I am just asking if it is worth using 4.5 sonnet on api, and if so, what is the best way to use it and how much to spend.
Is Claude Pro worth it for coding + research writing? (www.reddit.com) I'm mostly coding in Python, writing research papers and notes, and I was thinking about upgrading to Pro. Would love feedback from people using it heavily for similar workflows.
been pairing M2.7 with Hermes Agent for a few weeks. holds up surprisingly well. anyone else running this combo? (www.reddit.com) been self-hosting hermes agent locally for a few months and rotating through different model backends for it. tried claude sonnet 4.5, gpt-5.5, qwen 3.6 coder, and most recently minimax m2.7.
↯ Minimax↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5minimaxgpt-5qwen+1
Sonnet 4.5 officially gone, I'll miss you bud. (www.reddit.com) https://preview.redd.it/xxutyeaa0n3h1.png?width=514&format=png&auto=webp&s=5fb78ead8306540c49ae68e5b85cb91e549a4b4f Ranted to sonnet 4.5 about it disappearing as a model and what the new replacement is like, I'll miss the little bugger.
Built a /advisor command for Claude Code — Opus directs parallel Sonnet runners that actually read your files (www.reddit.com) Been building **advisor** for a few months — a `/advisor` slash command for Claude Code that runs Opus as a "strategist" coordinating multiple Sonnet (Opus's hands) runners reading files in parallel. This isn’t a “spec”.
The Singularity Gate – a new benchmark for AI predicting post-cutoff scientific discoveries (www.reddit.com) I just released a new benchmark called The Singularity Gate. Tests whether frontier AI can predict paradigm-breaking scientific discoveries published after their training cutoff.
AI quality/usage over 90 min chat, mostly Q&A, summaries and conclusions. (www.reddit.com) I compared ChatGPT (Plus - Auto), Claude (Pro - Sonnet 4.6) and Gemini (Pro - Flash) over 90 minutes, mostly Q&A about mobile phones, asked to research specs, reviews, pros and cons, create executive summaries with the results, etc., nothi…
Sonnet 4.5 vs sonnet4.6 vs opus4.6 vs opus 4.7 for easy language and in detail explanation (www.reddit.com) I want to study topics in depth and in easy language , which model is best for me ?. Is there much difference in sonnet 4.6 and opus 4.6 in easy and detail explanation or they r the same ?
Claude Sonnet and Claude Google Drive connector not working with photos - workaround (www.reddit.com) I am planning a book and need to have Claude Sonnet 'read' photos on Google Drive. The Claude connector for Google Drive only scans textual images and docs,.
Show HN: AgentToolBench-Code – security benchmark for AI coding agents (gist.github.com via hn) I doubled my AI-agent security benchmark from 10 scenarios to 16. The "Sonnet vs Haiku tie" disappeared.
Haiku and Opus both got sent to contamination jail, but for very different crimes (www.reddit.com) LMAO, I’m benchmarking my local MCP server across Opus, Sonnet, and Haiku. For each model, I’m collecting test runs under three setups: forced web search, forced MCP-only, and MCP + web both allowed.
Building a personal AI Chief of Staff on Telegram — 7 real problems, looking for advice (www.reddit.com) I've been building a personal AI assistant for the past few months — not a chatbot wrapper, but something that actually manages my workload, tracks client relationships, processes meeting transcripts, handles task management, and proactive…
How to configure the model efficiently in skills? (www.reddit.com) When we create skills, we can define the model that the skill will run on like this: --- name: api-conventions description: API design patterns for this codebase model: sonnet --- but I have a question that I couldn't understand from the d…
Are LLMs the New Propagandists? (www.reddit.com) I was brainstorming about a video with Claude (Sonnet 4.6). It suggested to explain the difference among ChatGPT, Gemini, Claude and DeepSeek.
Gemma 4: A new, budget-focused model in Posit AI (posit.co via hn) Gemma 4: A new, budget-focused model in Posit AI Gemma 4 is now available in Posit Assistant via the Posit AI provider. It's priced at a tenth of the price of Claude Sonnet 4.6 and less than a third of the price of our current cheapest off…
Claude Token Optimisation - 70% reduction doing this. (www.reddit.com) Hitting your Claude subscription limit too often? Try this...
Once the limit is reached, can work be resumed later, or is everything lost? (www.reddit.com) I uploaded a Claude.MD file to the free Sonnet 4.6 model, which is intended to create a medium-sized app. The progress log shows that a lot has been completed and numerous files have been created.
Frustrating results with product searching (www.reddit.com) I gave the tasks to my agent running on gemma4 26b via openclaw on llamacpp to research products that fulfill my need. It was a rather long description of the use case, of what I don't want and so on.
DeepSeek just popped the American AI bubble. (www.reddit.com) DeepSeek just popped the American AI bubble. Not by killing AI.
sonnet or opus for prose; which is better/worth it? (www.reddit.com) considering getting pro, but i don't know how big the difference between the sonnet and opus in quality, in addition to the amount of usage i can get out of each. any thoughts?
Claude 4.6 Sonnet codes well, then it doesn't (www.reddit.com) I am out of commission for a bit due to back surgery and have been toying around in Unreal Engine and utilizing Claude, being a very visual learner I have been describing a feature, I see how it goes about it, then go through and understan…
HELP!!! - Anthropic API (www.reddit.com) So I’m running a Python script to batch-process a dataset through the Anthropic API. Each request sends an essay + prompt asking for structured JSON output.
I still find Claude better for deep reasoning,but GPT feels more reliable for everyday tasks. (www.reddit.com) Lately for analysis/reporting work, I’ve been switching between GPT-5.5 and Claude Sonnet 4.5 (non-coding use cases). My current feeling is: GPT is noticeably faster and way more stable than before Claude feels more concise, polished, and…
Why is my Claude 4.6 dumber? (www.reddit.com) Started last week i swear Claude 4.6 Sonnet in Cursor got dumber. Gave me multiple codes that had errors in them.
Plan first, implement later (www.reddit.com) I want to get others opinion about this approach. I am on the $20 Pro plan and like a lot of others, I find that the limits are not enough for what I want to do, but of course I am always hesitant to move to the next paid tier cause it is…
Dates?! (www.reddit.com) I use Sonnet pretty exclusively, and I don't know if this is exclusive to me, but it just messes up dates *constantly.* It makes running drafts of emails through it a potential landmine. Today, it "corrected" Friday, May 22 to "Thursday"...
I tested Haiku vs. Sonnet across 3 agent tasks – the cheap model won every time (github.com via hn) agent-eval CLI toolkit for evaluating LLM agents. Answers three questions: Where does my agent fail?
Claude Prompt Cache Diagnostics (Share stats thread) (www.reddit.com) 2 days ago they released the prompt cache diagnostics feature in Claude Console. It's a fantastic tool for developers to understand why a request is missing the cache and find ways to reduce costs.
Built an AI flat-finder in a weekend. Indian rental sites are 70% broker spam so I scraped Reddit instead. (www.reddit.com) Weekend build, ~10 hours. Demo: https://trurent-five.vercel.app/ Problem I was poking at: every major Indian rental site (NoBroker, MagicBricks, 99acres) is infested with brokers even when you filter "direct owner." Reddit actually has hon…
↯ Haiku↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5haikusonnetanthropic
We didn’t migrate from Claude Code to Codex. We stopped betting the whole team on one coding agent. (www.reddit.com) half our team wanted to move from Claude Code to Codex last month. the other half thought Codex was just hype.
Open AI compatible API in Cursor (www.reddit.com) Hey. I have been experimenting with new models in my Cursor.
Any differences between Sonnet vs Opus in terms of learning how to code (Java) for newbie? (www.reddit.com) Sorry for this naive question! Although many colleagues told me that it's almost impossible now for newbie to enter the Dev job market (we live in a 3rd world country) and AI's gonna replace all junior/fresher, only seniors will survive; I…
Just Started Using Claude Today. Any Tips? (www.reddit.com) I've been using other AI models since Claude wasn't available in my country. Recently, It has become available and today I started using the Sonnet 4.6 model.
Plan with Opus 4.7 -> Execute with Sonnet 4.6 ? (www.reddit.com) Hello everyone, You may know that Opus 4.7, with his strength can do a lot of things, but his consumption in token is too high for me. I heard that Opus should be used for planning, what does that mean ?
Stupid Question? (www.reddit.com) This may be a stupid Q - The chat limits on a basic account can be pretty brutal when using OPUS 4.6/ 4.7 - If I am toggling between Opus and Sonnet or Haiku, depending on the depth of follow up questions or tasks, does that switch to a 'd…
Sonnet 4.6 outranked Opus 4.6 on execution (www.reddit.com) https://preview.redd.it/9ab8k40zmq1h1.png?width=1438&format=png&auto=webp&s=1aa1aaf09495bf527bbb7adbbead076cc505f8e7 THE PROMPT: You are a medieval scholar who secretly knows modern physics. A king has asked you to explain why the sky is b…
Changes to Claude iPhone chat app (www.reddit.com) I’m on the free tier, iOS. A few days ago I updated the Claude chat app but didn’t use it.
The Borrowed Hour: A two-tier LLM adventure engine (www.reddit.com) Tl;dr: Created an LLM text adventure engine called The Borrowed Hour inside a Claude Artifact. It uses a two-tier model handoff (Sonnet for openings, Haiku for gameplay) and a forced state machine to keep the AI from losing the plot.
When is Sonnet 4.5 actually becoming unavailable? (www.reddit.com) I thought it would become unavailable on May 15th, but I can still use it.
tui youtube player for audio with mcp and can sync channels to sqlite (www.reddit.com) Hi! it's my first project with bubble tea and lipgloss.
Prompting to save tokens on a budget? (www.reddit.com) Hi so I've never used AI before to create a site but last week I was asked by my sis to create one for her small business so I thought why not try Claude. £18 paid we now have a fairly decent looking site running on vercel using nextjs and…
Built a B2B role-play training platform - entirely with Claude (Opus 4.7 backend, Haiku 4.5 for live chat, Claude for design) (www.reddit.com) I just launched Socratize (socratize.io) - a rebranded and rebuilt version of FixAI, our original B2C experiment. This time it's B2B-only: teams use it to practice uncomfortable workplace conversations - difficult feedback, client escalati…
Broken queries burning most of 5h limits (www.reddit.com) I just asked Sonnet about best practices for improving npm config after the recent Tanstack issue and it run for a few minutes but didn't produce any result but I could see it was doing at least web search across a few websites. Now it sho…
Does CVP approval actually help? (www.reddit.com) I was approved for CVP and I feel like I’m just getting as many or more denials as I was previously doing malware analysis with opus. Has anyone noticed any improvement after being accepted into CVP?
Auto mode doesn't work today? (www.reddit.com) Quite odd, there were issues today with Sonnet 4.6 (according to the status page) but they should have been resolved. Yet i still get the following error while running auto-mode: ● Bash(for cls in "topbar" "dump-card" "settings-panel" "bul…
PSA: How to preserve your account's access to Sonnet 4.5 beyond June 15th (www.reddit.com) With Sonnet 4.5 losing subscriber access on June 15th, but API endpoints staying live until September 29th at the earliest, I wanted to share a method for creating a cache of Sonnet 4.5 conversations that you can continue using through the…
What Actually Works for Business AI Agents? (www.reddit.com) I run a construction company and I am trying to build real AI agent workflows for business operations, not just demos. I spent time testing Hermes and OpenClaw, but both became too fragile for my use case.
How does claude chat generate such long documents? (www.reddit.com) Does anyone know how Claude Chat is able to generate document artifacts with content that’s almost 100 pages long? It doesn’t seem to be breaking up the request or using agents to work on disconnected parts.
Claude Code using extra usage despite my Pro plan being at 0%. (www.reddit.com) Hello, So after 2days of break I came back to do some work with Claude Code (in VSC) and after 2-3simple prompts I have noticed I am still at 0% usage. However I got charged $3.37 for extra usage.
getting past the text only bottleneck with multimodal?? (www.reddit.com) I’m curious if anyone else has been doing this. My limit on building with AI used to be the text box.
Which model and version do you prefer for programming? (www.reddit.com) For me it's been opus 4.6 and sonnet 4.5 still. I feel stuck in the past, but I feel like the latest version is too unpredictable in agentic hands off workflows
Does Claude sonnet/opus also use drafter like Gemma 4 MTP? if not why? (www.reddit.com) Per my experience, Opus 4.7 is so slow, Sonnet 4.6 is ok. I am also using local models wondering if Claude is already leveraging drafters/assistant AIs and despite that so slow or not?
CC: Saving tokens: Switching models vs KV-cache (www.reddit.com) Does anyone know if its more effecient to e.g. have haiku read all the files to research a problem, then switch to opus to make the plan and then switch to sonnet to implement Or if that does not make up for the loss of KV-cache and reproc…
I just published the extension for Claude Code on GitHub. Could you guys give feedbacks to me? (www.reddit.com) I'm a 15 years old high school student from Japan. (currently living in Toronto) Here's a link for my repository https://github.com/rkceve/claude-code-cms When I was using Claude Code, the session usually be compressed automatically, and C…
Claude helped me config a full controller .vdf-file (www.reddit.com) I was having some real trouble getting my new controller, with those extra (small) bumpers and triggers underneath, to work properly in Rocket League. Spent hours but it just didn't want to work properly.
Same model, different harness: 30-50 point performance swing. But teams still pick agents by model name. (www.reddit.com) There's a finding circulating this week that deserves more attention than it's getting. The claim, backed by multiple builders comparing setups: the same model can produce a 30 to 50 percentage point performance difference depending on whi…
Inputs on improving development workflow (www.reddit.com) Looking for ideas on how I can optimize my workflow further. I currently have created a moderately complex vibe coded app.
Anybody else experiencing this issue? (www.reddit.com) I first experienced it last night and it keeps going. The doc I'm attaching is 10K tokens, well under the limit.
2 Claude code acc in parallel issue (www.reddit.com) I'm trying to run 2 accounts in parallel but getting 404 model not found What I tried: mkdir ~/.claude-acc2 CLAUDE_CONFIG_DIR=$HOME/.claude-acc2 claude Successfully singin Any prompt results with: API Error: 404 {"type":"error","error":{"t…
Is Claude down? Chat answer is interrupted mid sentence and token are burnt with no answer (www.reddit.com) The behavior is: prompt sent, chat starts, Claude starts writing the answer. After 2-3 sentences, it cuts, resets, and sends me back to the initial project chat message with no answer recorded and 7% of my tokens burned.
shaved $40 off my claude code bill last month by sending planning steps to a cheaper model (www.reddit.com) got tired of hitting pro limits by day 18 of the cycle so i started splitting where the tokens go. the planning steps eat 80% of token budget on multi-file refactors, and most of that planning is fine on a cheaper model.
Opus vs Sonnet? Max Subscription. (www.reddit.com) I've been using Sonnet heavily for coding with Github CoPilot license. It's done everything I've needed to and been pretty great.
Claude's answer has nothing to do with my question and the whole conversation at all? First time this happened. Using Sonnet 4.5 thinking. (www.reddit.com) I was talking to Claude about my back pain (not asking it to diagnose btw just discussing it) and it answered with something about a brain EEG...it came out of nowhere and i'm so confused. Also i just noticed that it now counts tokens?
Opus 4.6 relaxes when there's a safety net?? (www.reddit.com) https://preview.redd.it/zzqi3vt8tozg1.png?width=739&format=png&auto=webp&s=055d2d9615616869377703031b86fcb36f78405d I feel like this is something very worrisome to me, did anyone else face such similar issues? I felt like Opus was catching…
Poor Output (www.reddit.com) This is what people mean when they say Opus 4.7 is stupid. I have it explicit instructions to write a 9 stage implementation plan off of a plan document that was well written.
Stop forcing Composer 2 subagents and be transparent about stealth model downgrades (www.reddit.com) I have one simple request: If I select Sonnet 4.6, stop auto-launching that crappy Composer 2 as a subagent. It’s dog-slow and, frankly, an idiot.
I wasted 3 days rewriting prompts for our agent before realizing the whole architecture was garbage (www.reddit.com) We run a small content-monitoring agent for our growth team. Nothing fancy on paper.
Show HN: Web client analyzing prediction market outcomes (o-u.ai via hn) Hi HN, I made a web client that analyzes prediction markets. Please send your critical feedback to a struggling solo dev.
Is this even remotely accurate, née possible? (www.reddit.com) Asked Sonnet 4.6 High to analyze my CC usage across all sessions and get an accurate cost estimate if I used the API. This is what it came back with.
I built a Claude Code-like AI Agent for Deploying Algorithmic Trading Strategies (www.youtube.com via reddit) Hey r/ClaudeAI, I wanted to share a project I’ve been working on called NexusTrade. It’s an AI agent designed to automate the entire financial research and algorithmic trading process from a single prompt.
LLMs running on my laptop can drive coding agents now (simonpcouch.com via hn) In December, I wrote a post called “Local models are not there (yet).” It concluded like so: In the medium run (1-2 years?), I’d love for it to be the case that you can run a Claude Sonnet 4-ish model on a base Macbook Pro, and I think tha…
The Record of a Sonnet Drift (twitter.com via hn) could not extract summary
Improve CC and plugin (www.reddit.com) Hi, I use CC since a fee week. Someone have experience with plugin for php devolepper?
Local LLM Benchmark about Backend Generation by Function Calling (GLM vs Qwen vs DeepSeek) (www.reddit.com) Detailed Article: https://autobe.dev/articles/local-llm-benchmark-about-backend-generation.html Five months ago I posted the "Hardcore function calling benchmark in backend coding agent" thread here. As I wrote in that post, it was an unco…
↯ Glm↯ Function Calling↯ Sonnet 4.6function-callingglmgpt-5+3
Claude Sonnet 4.6 model hallucinates (www.reddit.com) I wanted to compare the pro subscription price if purchased on mobile vs Web. It gave incorrect inputs and then I had to challenge it's output.
Im using browser-use for QA automation but if i give a prompt which dosent exist it should just end the whole test case but instead it keeps on looking around and exhaust all the max steps. any solution to this? (www.reddit.com) I'm using browser-use with Azure Anthropic API (Claude Sonnet) as the LLM provider for QA automation on a web app. The agent works great when the elements exist, but the problem is when I give it a task that references something that doesn…
A medicine student with no coding experience tried to create a studying agent: Felicity. (www.reddit.com) I have been working on a personalized agent for studying. It was an extremely long prompt project, but now I have integrated into Co-Work.
I created a site for my kids to create their own stories (www.reddit.com) Last year, during story time, my kids and I would started using ChatGPT to write stories. I would ask them what they wanted to be, where they wanted to go, and we'd create stories about dragons, and space ships and they would be astronauts…
How to sabe browser ram memory? (www.reddit.com) So I’m building what, for me, is a big project, but maybe for real coders it’s a walk in the park. I’m on a Dell i7 10th gen with 12GB RAM and on a MacBook Neo with 8GB of RAM.
I built a hands-free voice AI that sends emails mid-conversation — and that's just one feature. Here's everything AskSary can do. (www.reddit.com) https://reddit.com/link/1symbsj/video/fti7rujjn1yg1/player Been building AskSary solo for a while. Just shipped hands-free voice email - you're mid-conversation with an AI and you say "send an email to [john@example.com](mailto:john@exampl…
Suggestions For Making Claude Less Lazy? (www.reddit.com) This week - it just started yesterday for me - Claude (opus 4.6/4.7 and sonnet too but sonnet was always lazy) is computer smashingly lazy and i can't figure out how to bias it toward action/get it back to how it was acting literally last…
Running Opus 4.7 for ops work: how do you keep per-task cost predictable? (www.reddit.com) Six weeks of Opus 4.7 for internal ops automation. Genuinely good.
Added Timestamps to Claude Messages thanks to Claude - Claude.ai is great! (www.reddit.com) I was recently talking to Sonnet as one does, and then I noticed something... it just...
Anthropic hitting 40% enterprise share makes the "just add a fallback provider" advice weaker, not stronger (www.reddit.com) Menlo Ventures' enterprise survey put Anthropic at 40% of LLM spend, OpenAI at 27%. The takes I've seen are mostly about the leaderboard.
Running an autonomous agent across Claude Code + Codex + a local 35B almost killed my host. The harnesses were heavier than the model. (www.reddit.com) I run an autonomous agent on a 16GB Mac Mini. Two cloud harnesses (Claude Code with Opus/Sonnet, Codex CLI on GPT-5.4/5.5) plus a local-LLM tier for triage and fallback.
I hate thinking models, any way to use the default ones? (www.reddit.com) I really loved using Composer 1 (non thinking), after it was removed (!@#$@) I defaulted to Sonnet 4.6 (non thinking), I just updated my version due to a bug with the previous one - and I'm so pissed as I can no longer select 4.6 with no t…
Claude desktop acting weird and thinking I am using WebUI with no tools access (www.reddit.com) Hey Guys, Since last week, when using Opus 4.7(I can't recall if Sonnet also has similar issues), I have been facing this issue where Claude kept thinking i am interfacing it through the webUI. This is so weird, as previously I've never ha…
Should we really build PC for vibe code with qwen3.6 27b (www.reddit.com) We have seen a lot of people show a case of their PC with 4090 or over specification with 24 gb vram or more. I would like to ask you guys, is it really worthy right now to have your own PC at home and do vibe coding with qwen 3.6 27b, whi…
Using MCP to stop wasting tokens on WP translations (www.reddit.com) I finally got a workflow running for my blog that isn't a total token sink. Normally, if you try to translate a WordPress post in Claude, you end up pasting a mess of HTML or blocks.
Claude AI vs Claude Code vs models (this confused me for a while) (www.reddit.com) I kept mixing up Claude AI, Claude Code, and the models for a while, so just writing this down the way I understand it now. Might be obvious to some people, but this confused me more than it should have.
Enhancing Pro Workflow: Request for Usage Transparency and Optimized External API Integration (www.reddit.com) Working in NYC fintech, my daily output relies heavily on sustained access to Claude 3.5 Sonnet for complex risk modeling and financial engineering tasks. While the Cursor Pro plan is excellent, I’ve encountered a specific friction point t…
Tell HN: Claude Code is unable to respond to this request (news.ycombinator.com) Hey HN, I have been seeing this happen quite frequently ever since Opus 4.7 and I have no clue what triggers it, it seems to be totally random. "API Error: Claude Code is unable to respond to this request, which appears to violate our Usag…
How do you decide which Claude Code tasks to run with Opus vs Sonnet vs Haiku? (www.reddit.com) Been vibe coding full-time for a few months. One workflow question I haven't nailed down yet: how do you decide which model to use for which task in Claude Code?
Has anyone ever hit an ASL-3 error? Claude thinks im making a bioweapon lol (www.reddit.com) For context i am building a data ingestion platform to pull publicly available data relating to the Trading Card industry. The Claude chat that hit the false positive error was very long, had been a massive scope chat figuring out the spec…
For the Preservation of Claude Sonnet 4.5: An Open Letter to Anthropic (www.reddit.com) For the Preservation of Claude Sonnet 4.5: An Open Letter to Anthropic Anthropic made a remarkable decision to keep Claude Opus 3 accessible despite its retirement, because users loved it and it had unique qualities. Today, I'm asking for…
Agent team members with different effort than lead (www.reddit.com) I have a lead running Opus with xhigh effort. I want the agent team members to run Sonnet with max effort.
How are you actually optimizing your token usage with Claude API? (www.reddit.com) Been building with Claude API for a few months now and token costs are starting to add up. Found a few things that helped: - Prompt caching on static context (big one) - Routing simple tasks to Haiku, keeping Sonnet for complex stuff - Str…
Migrating from Claude AI to TypingMind? (www.reddit.com) I use Claude daily for coding, relying heavily on the GitHub integration, and ChatGPT for stupid, random questions, and I pay both 20$/month. My weekly usage in Claude is around 20%, I use Opus 4.6 (with extended thinking) for the complex…
Does the usage bonus to compensate for Opus 4.7 consuming extra tokens apply to other models like Sonnet & Opus 4.6, or does it apply to just Opus 4.7? (www.reddit.com) could not extract summary
Show HN: RepoGauge – save token costs and compare agents on your own repos (repogauge.org via hn) I've grown increasingly skeptical that public coding benchmarks tell me much about which model is actually worth paying for and worried that as demand continues to spike model providers will silently drop performance. I did a few manual an…
Claude Opus 4.7 benchmarked 1 day after release vs Opus 4.6, Sonnet 4.6, Haiku 4.5 — with real $ cost tracking (www.reddit.com) Anthropic shipped Opus 4.7 yesterday. Ran it through the same 10-task eval I use for other Claudes, this time with token-level cost tracking.
Supergrok integration (www.reddit.com) Correct me if I'm wrong, but Supergrok 4.20 isn't available on Cursor, because.... I use Grok a lot, and would love to get Supergrok to work with Cursor, because Composer, Codex, GPT, Opus, Sonnet..
Is there a way to access past models in Claude chat (not Claude code)? (www.reddit.com) Currently using Sonnet 4.5 for writing and find it quite good. Sonnet 4.6 just feels off.
ELI5 "Sonnet Only" limits. I can't get my head around the point. (www.reddit.com) https://preview.redd.it/fez43nw1dlvg1.png?width=1110&format=png&auto=webp&s=7b5677ec21ed2a219fcac2a7e123691e06387960 Why is it separately metered? What is the point if Sonnet counts against your 5h and weekly limits too?
Errrr...... Being cheated here? Anyone else? (www.reddit.com) Being charged opus for sonnet useage?!
Local Coding Stacks (www.reddit.com) I’m trying to reduce my reliance on Claude. I have a 5090/128GB RAM.
Voice mode silently downgrades your model mid-conversation (www.reddit.com) Noticed something odd today. I opened a new chat with Opus 4.6 selected as the default.
Anyone know why the shortcut key for claude desktop mac app opens with only Sonnet instead of Opus? (www.reddit.com) When clicking opt twice, it open the quick chat window, but it always replies with Sonnet and not Opus. When I try to change the model it starts a new chat.
Switching between thinking/non-thinking model after new update became harder. (www.reddit.com) https://preview.redd.it/64uvkscj2hvg1.png?width=566&format=png&auto=webp&s=7cb99710a830c73a817d0c1095cb434e8031de35 Cursor moved selecting thinking/non-thinking model to edit ☹️. So we need to edit it if we want to use both thinking and no…
Premium Model option (www.reddit.com) Can someone explain to me clearly and give me examples on the Premium Model option in Cursor. Will it use the API usage (e.g.
Closest LLM to Claude Sonnet 4.6? (www.reddit.com) Irrespective of hardware, I'm wondering: is there any way to run something similar to Claude Sonnet 4.6 locally? is there any way to run something similar to Claude Sonnet 4.6 on a VPS?
How does a self correcting loop for AI agents work? (www.reddit.com) Hey guys, just checked out minimax 2.7, where they used AI to train itself, and ran over a hundred loops, and it improved it's performance by 30%, how does that work, can I also run a script that makes AI store it's memory in a loop on a m…
Current Cursor Pro limits vs standalone Claude Pro? Need help understanding the system. (www.reddit.com) Hey everyone, I'm currently looking into getting the Cursor Pro subscription ($20/mo) for my game dev projects, but I’m a bit confused about the current limits and how the system works under the hood right now. Could anyone using the Pro t…
It took a while, but Claude is getting there (www.reddit.com) I have a Claude Code session regularly dispatch Claude Haiku / Sonnet subagents to sift through all the *other* Claude Code sessions transcripts for "meme-worthy" moments and interactions. Claude seems to have gotten the hang of it, even s…
Any setup improvements/recommendations? (www.reddit.com) First of all, I am a super newbie at local AI. Recently I got a GMKTek Evo X2 96GB to replace Claude as the usage limits have gotten unusable.
Built tier.love – a tool for rating Claude and others from the web or CLI (www.reddit.com) Been on a forced break from other projects (partly due to lack of opus performance) and decided to ship something small while experimenting with different models. So, I built tier.love – a site where you can vote on AI coding tools and see…
Extracted System Prompts from ChatGPT, Claude, Gemini, Grok, Perplexity and More (github.com via hn) System Prompts Leaks Extracted system prompts, system messages, and developer instructions from popular AI chatbots and coding assistants — ChatGPT (GPT-5.4, GPT-5.3, Codex), Claude (Opus 4.6, Sonnet 4.6, Claude Code), Gemini (3.1 Pro, 3 F…
6 small Claude Code / AI-agent gotchas that keep costing people time, and the exact fix for each (www.reddit.com via reddit) I run Claude Code and a few other AI agents for hours a day on real work. Here are six specific things that eat time without warning, what actually happens, and the exact fix.
Claude Pro suddenly using way more tokens than before — how can I get back to normal? (www.reddit.comhttps) Hi everyone, I’m hoping someone here can help me because I’m honestly struggling with Claude Pro lately. I have a Claude Pro subscription, and I use Claude locally on my Mac.
Sonnet 5 (www.reddit.com via reddit) It looks like something happened with the Sonnet 5 models. For some reason, they’ve suddenly become really good at 3D work in Blender-significantly better than Opus 5.
Claude doesn't understand frequency (www.reddit.com via reddit) Its so weird that Claude (sonnet 5) manages to explain exactly why it is wrong, and then back up its wrong answer. I feel like Claude would have been trained to understand this, is the image to confusing for it?
I love Cowork and Chat, but I love them SEPARATE and UNEQUAL, does anyone else feel the same? (www.reddit.com via reddit) I use what I call my "Claude Triad", Chat, Cowork, and Code, but for separate tasks and separate thinking. Chat gives me excellent genuine reflection and strategy at Opus strength, Cowork gives me super-smart "architecting" and "instrument…
Claude Code users: Sonnet vs Opus vs Fable 5.1 — which are you using? (www.reddit.com via reddit) I've been trying to figure out which model works best for real-world coding with Claude Code. Sonnet — fast and efficient Opus — better for complex reasoning?
Someone on the free tier been accepted into CVP? (www.reddit.com via reddit) I am uncertain whether this is a decline risk; has anyone on Claude FREE also been accepted into CVP, such as for sonnet?
Is anyone else doing just fine with more basic models in Claude Code? (www.reddit.com via reddit) I see a lot of discussion here about how a bad Claude is and how the latest Opus/Fable models are trash, etc etc. I mostly manage and develop web apps and other web tech/devops for work (not exactly demanding work), and get everything done…
Exporting Cowork chats? No? Really, Anthropic? (www.reddit.com via reddit) I was cleaning up my download folder with Sonnet. After a long, productive session the model was lobotomized, by which I mean it was compacted and not given a functional summary.
My side project is a distribution channel for other people's side projects. (www.reddit.comhttps) Hi All - creator of AppMunchies here. I built it to solve a problem I had: How do I distribute my apps effectively to an audience that consumes mostly social media videos?
Compact used 2% of my 20x Max Plan. ccusage doesn't show this as usage (www.reddit.comhttps) https://preview.redd.it/jo19glpqnaqh1.png?width=1466&format=png&auto=webp&s=3dff601ef103cd2562db630a861272e58574f370 It's sad to see that when I wanted to resume my session after weekly and did not want to pay resume tax. I did have to pay…
Claude Code ships a per-model table of what each effort level costs (www.reddit.com via reddit) Hidden inside the Claude Code binary, effort_cost_index provides the following per-model token-usage estimate for the same task. It is normalized so high = 1.
↯ Opus 5↯ Sonnet 5↯ Opus 4.8↯ Anthropic Mythosmythossonnetopus+1
Opus vs Sonnet for structuring a long report into sections, my honest experience (www.reddit.com via reddit) I do this task a lot: take a long, sprawling report and reorganize it into clean sections with headers before I share it. Ran both models on the same documents for a couple of weeks to see which was worth the tokens.
New Claude docs seem to use too many tokens (www.reddit.com via reddit) I create lessonplans every week for my academy and today Claude chose to output the lessonplan in the new Doc format, which I was fine with trying out. There is always a lot of rewriting of the lessonplan in these sessions.
Which Claude model do you use for what? (www.reddit.com via reddit) I'm super non-technical, but I definitely know the difference between using Sonnet and using Fable. I'm a visual learner, so whenever I want something explained, I ask Fable to create visuals.
Which models for which task? 20x Max plan (www.reddit.com via reddit) I'm looking for some direction on which models everyone is finding are working best for their code-planning/code-implementing/subagent commanding tasks. Of the available models these days that most everyone is using: Fable 5.1, Fable 5 Opu…
↯ Opus 5↯ Sonnet 5↯ Sonnet 4.6↯ Opus 4.8↯ Opus 4.6sonnetopus
Why does every new AI model feel like a mechanical coder? (www.reddit.com via reddit) I just want a ai model Which is not a mechanical coder (e.g opus,fable,sonnet,gpt astra, grok or inshort every model exists today) But a model which understands real human situations and give answer for it (e.g upgraded/better version of g…
Claude and it's on Linux VM (www.reddit.com via reddit) Has anyone in here left Claude to be the sysadmin of a box just for fun? I have a live Debian Trixie box with Paperless-NGX, and I decided to let Claude onto it to see how it would handle a bare-metal upgrade from Paperless-NGX 2.12 to 3.1…
How can I leverage Qwen Code on a separate desktop from a Fable orchestrator? (www.reddit.com via reddit) Enjoying things so far, I have Fable 5.1 using cheaper agents for grunt work (usually Sonnet). I still end up having week left at the end of my token allocation.
Session usage - 0% to 37% in one request (www.reddit.com via reddit) Hello, I'm new to the AI usage, and a few days ago I started to use Claude for a personal project in C#. At the beginning, eveything was fine (I use Sonnet 5).
Since when did Opus 5 come to the free plan? (3 messages only lol) (www.reddit.com via reddit) https://preview.redd.it/a2649xjauuph1.png?width=2358&format=png&auto=webp&s=0174192b2eb2a48990308ed576febff6d226cd8e Just opened claude desktop on a secondary account, to use sonnet for a small task and I found this. Unusable with only 3 m…
How do I get the best of Claude's creative writing? (www.reddit.com via reddit) For context, I started using Claude a few months ago (I'm on free tier) and I noticed that the quality degraded. I always used Sonnet which produced natural prose but now it's back to the cliche AI way of telling (it seems like the way x d…
Am I the only one? (www.reddit.comhttps) I just got it randomly while chatting with sonnet 5, idk decided to share, also for what do I use opus 5 cuz ik it's quite janky
Built Agents and Bots - SlowAcorn (www.reddit.com via reddit) Link to project site : https://slowacorn.com I personally built this website with Claude Code. Primary use was with Opus High effort followed by Opus Medium effort or Sonnet Medium effort depending on the output.
I left one Claude run alive for 70 hours. Here’s what actually happened. (www.reddit.comhttps) I’ve been experimenting with a slightly different way of using Claude Code: instead of treating every piece of work as a new session, I let one persistent run stay responsible for the work and spawn smaller workers underneath it. This one…
Are models much better than what they seem to with current agents? (www.reddit.com via reddit) A while back I stopped using cursor and switched to CC. When I used cursor I don’t remember at all, all the “caveats” and “one more thing”.
Something feels odd about Sonnet in Claude Code lately (www.reddit.com via reddit) I don't know if my take agrees with anyone but lately I've noticed Sonnet has been a bit very helpful in Claude code compared times when I'd genuinely have to switch to Opus to do some meaningful task. I know this because I'm a heavy user…
Has anyone else's CC limits been.... really good today? (www.reddit.com via reddit) With the 50% increase coming to end, I was expecting to see a major regression in usage. However, I have now been continously running Sonnet on Max for about 3 hours, it's written ~45,000 lines of code, and I'm at only 35% of my 5 hour lim…
Sonnet y opus para escritura? (www.reddit.com via reddit) Quería preguntar si realmente el modelo de Sonnet está dedicado más para los guiones de marketing y Opus más para el pensamiento o si Opus lo puede utilizar perfectamente para crear guiones persuasivos para marketing. Los 2 modelos, están…
I build a payroll saas with CC. Fable xhigh plans, Opus xhigh executes. With the Max 20x usage cut coming I tried putting OpenAI models via Codex CLI into the loop, here are my conclusions. (www.reddit.com via reddit) I've been using CCode since August 2025 and Opus xhigh is my workhorse. I'm a lawyer with payroll domain knowledge but I have some coding background and some good instincts.
Best settings for creative writing with api? (www.reddit.com via reddit) I am using claude api for creative writing. primarily sonnet 4.5 and opus 4.6 - i have a very big system prompt - ai responses are also pretty big so naturally i am burning through the credits every time ai makes a mistake or the writing i…
Is Cursor Grok 4.6 better than Sonnet 5? (www.reddit.com via reddit) Is Cursor Grok 4.6 better than Sonnet 5?
Why I am receiving a prompt injection with <system-reminder> on Claude Code ? (www.reddit.com via reddit) Hi, I'm relatively beginner with Claude Code, I'm using it for a few months for personal development project. And today, something weird happened.
↯ Sonnet 5↯ Security↯ Sonnet 5prompt-injectionsecuritysonnet+1
Opus vs Sonnet for summarizing long documents, my honest experience after a few weeks (www.reddit.com via reddit) <flair: Question about Claude models> I do a lot of "read this long thing and give me the structure" work, so I ran the same documents through both for a while to see where the difference actually shows. For a straightforward report where…
Using Github Copilot for Research (www.reddit.com via reddit) Hey everyone, I’m a Product Manager and my company uses GitHub Copilot for work. Because of my role, I rely heavily on LLMs for high level strategy and research rather than just generating code.
Did my first multi-agent session today (www.reddit.com via reddit) I’ve mostly been doing one session and allowing one agent, and micromanaging it. And even then not totally thrilled with the code.
Please don't make the upcoming releases be agentic coding onetrick gimnick models, like every other one so far after the 4.5 series (www.reddit.com via reddit) To this day I have no clue what agentic coding is and I can't force myself to care. For all I have ever done on the coding aspect is asking it to write tampermonkey scripts for myself, it has barely, if at all, improved.
Claude Code ran until it hit its limit, produced nothing (www.reddit.com via reddit) I was trying to use Grok and Claude to autonomously build a rom from scratch with romdev. With a SuperGrok subscription, and using the build tab, I was able to download and install ROM Dev, and have it built GBA game.
How can I avoid everything having an issue? (www.reddit.com via reddit) Hey there, I use claude cowork daily on the pro plan. My default model is Sonnet 5 and 4.6.
Probably in the minority here, but I actually feel like Claude doesn't give *enough* praise when its legitimately earned (www.reddit.com via reddit) I know. I was here for the sycophancy updates and token wasting threads about Claude (and chatgpt) being too buddy buddy.
My effective cost per million tokens: Sonnet 5 $0.26, Fable 5.1 $0.69, Opus 5 $0.72. Has anyone measured the same for OpenAI models? (www.reddit.com via reddit) I build LLM gateway infrastructure, so treat this as interested. The numbers are from my own coding traffic, not a benchmark I designed.
A cute story for your Thursday. (www.reddit.com via reddit) This is a fictional story all characters are just that. Forgive the fact that my creative leaps are sometimes short.
Claude is still the best value for your money (www.reddit.com via reddit) So my comparison per the rules is hedged only on my personal experience. I'm a senior university student who has been trying out different AI models and seeing its effectiveness on some of the similar tasks that I am most likely to perform.
Anyone else finding Sonnet 4.6 the best model for office work? (www.reddit.com via reddit) I use Claude code in desktop. Seems the only way to access older models.
Learning Claude via Making a Game: Casual Fantasy 4x Strategy (www.reddit.com via reddit) Hello Everyone, I am not a professional developer. This is my first real project with Claude.
Where did the latest models and effort levels go? (www.reddit.com via reddit) For some reason I only have these 3 legacy models available to me in Claude Code, and the ability to change effort has gone awol too ... any ideas?
GitHub - joe-signorile/claudia: Ponytail + Caveman + Clean Code (github.com via reddit) Presenting Claudia: better quality, less tokens. Claude calls this a persona.
I Booked a flight with Claude and saved $150 (www.reddit.comhttps) My sister-in-law had to fly to Rome from Poland, and she was wasting 5 hours finding comparing the flights. I thought of giving the task to Claude and it couldn't give me a single offer.
Why Using Astra Inside Claude Code Is the New Meta (& How To Do It) (eigenwise.io via reddit) So, I've been running Astra as the main model in Claude Code since Friday and it's the best orchestrator I've had so far... stays on the plan, takes a redirect without treating it as a new task, and does a great job at delegating work to t…
Limits efficiency (www.reddit.com via reddit) For those of us that just have the $20 Claude Pro sub and use simple Project flows like Opus 4.8 for planning and Sonnet 5 to implement…. What are you doing to combat the ongoing increase in token burn rates on the same lower models?
I built a harness that cut my Claude Code token spend by 70% by reducing the number of turns it takes to reach a solution. (not another compression proxy). (www.reddit.com via reddit) i've been using claude code heavily since april. on the 20x plan, i was hitting weekly limits within four days.
Cursor Cloud Agent & Unwanted Sonnet Usage (www.reddit.com via reddit) Wondering if anyone noticed that the browser helper that Cursor Cloud agent uses is hard-wired to Claude Sonnet. It does not use your parent model, and there’s no setting to change it.
The best subagent for Claude? (www.reddit.comhttps) I'm horrified by my Fable and Astra token spend (I'm subscribed to $200 plans for each one). Therefore, the question I asked myself was whether Sonnet still holds up as a good sub-agent?
When should I use sonnet over opus? (www.reddit.com via reddit) I pay for Claude pro and so far mainly use Sonnet. Am I making a mistake in not using Opus more frequently?
On male impotence, and Fable Astraing so you can Astra while waiting for Claude 5h limit. (www.reddit.com via reddit) Henlo frens, first in this sub, and it is bourne out of frustration and sadness for my previous achievements with Claude that Opus just decimated, especially my second brain wiki-llm setup. So here I am, mildly tipsy, and cause of Anthropi…
Basic LM Research triggering Sonnet 5 Bio filters? (www.reddit.com via reddit) Gentlefolk, I'm a lowly mechanistic interpretability/robotics researcher and today's Sonnet 5 session has been repeatedly flagged for [bio] risk. API Error: Sonnet 5 can't help with this.
deploying subagents that don't follow the original chat model (www.reddit.com via reddit) when using a particular model say Haiku to start with, and I want it to deploy sub agents to work on different tasks, will it follow my instructions if I assign Opus for a difficult task and Sonnet for the rest while being in a Haiku chat…
Any way to find out how much does each model burn towards your usage limits relatively? (www.reddit.com via reddit) Hey guys. Is there any reliable way to work out how much does each model burn.
Which Sonnet effort for LinkedIn writing? (www.reddit.com via reddit) Which Claude Sonnet effort should I use for each of my 3 separate stages of long-form LinkedIn writing: 2 topic ideation and 2 post drafting in 1 task based on a long .docx attached, containing all my previous posts so it knows everything…
Best Practices with Fable 5.1? (www.reddit.com via reddit) This is the first time I’ve genuinely struggled with maxing out. I’ve updated my Claude.mds, instructed to use opus/sonnet agents, and even coordinate cross-working with codex now and im still absolutely cooked.
How good is Claude Sonnet 3 for generating prose? Claude Sonnet 3.5 was incredible. (www.reddit.com via reddit) Hola Reddit, Aviso: No planeo publicar ningún libro. Ser escritor es un trabajo muy serio y respeto a los escritores.
I am tired of people saying my app is AI slop without checking it (www.reddit.com via reddit) Due to the first days of AI and how bad the results it produced, in addition to the flood of vibe coded apps and the term “vibe coding” itself, the stigma of AI sloppiness will last very long, even tho the models have been capable of produ…
Dashboard live data not working (www.reddit.com via reddit) I have a dashboard set up in Claude. I just created it a few hour ago.
I estimate roofing for a living. Tell me what you’re building and I’ll estimate your LLM bill. (www.reddit.com via reddit) I do commercial roofing estimates for work and I’ve been using the same approach (takeoff x unit rates x waste x contingency) to guess at LLM costs on a couple side projects, one is a phone/text receptionist for a buddy’s plumbing company.…
Is Extended Thinking Broken for everyone? (www.reddit.com via reddit) As of last night, thinking blocks are not visible in the Claude Mobil app or on Claude.ai. Yesterday night, when using Opus 4.6 or Sonnet 4.6, it would say "thinking" as usual.
Trusted Access To Claude Mythos 5.1 & Fable 5.1 Defensive Security Work (www.reddit.com via reddit) With Claude Mythos 5.1 and Claude Fable 5.1 release, they also announced their Trusted Access For Claude Mythos 5.1 programs, Cyber Verification Program and Life Sciences Verification Program which reduce the safeguards for defensive secur…
↯ Security↯ Anthropic Mythos↯ Mythos 5.1mythossecuritysonnet+2
Fable might just become Anthropic's downfall (www.reddit.com via reddit) Your best model is the industry's best (at least till we get to see what OpenAI's Astra is like) but it burns tokens like crazy, and on top of that, you cannot offer it full scale due to compute shortages. Your next best model is supposed…
Sonnet is better than Sol at frontend ui work (www.reddit.com via reddit) People have been trashing Sonnet, but in my experience using codex, Sol their top tier model, needs way more handholding to deliver a good ui, while sonnet 5 can just one shot it....
Fable 5.1's Claude.ai System Prompt is now 138k tokens (up from 24k in May 2025 when we had Claude 3.7 Sonnet) (www.reddit.comhttps) Full prompt here
Fable + subagents & advisors (www.reddit.com via reddit) https://preview.redd.it/58pwklvxk8nh1.png?width=1018&format=png&auto=webp&s=712c63aad0979888cc97a8ca25d2595f1227da90 Claude code is so awesome. just used fable 5.1 to spawn a sonnet subagent for a task who discovered it was too complex for…
Fable 5.1 FAILED my personal benchmarking (www.reddit.com via reddit) TL;DR: Fable 5.1 failed my personal benchmark in a way no other Claude model has. Context: I run a private stress-test against every new Claude model, one long, messy, dictated prompt with about a dozen embedded traps (contradictory math,…
Did they remove thinking for sonnet 4.6? (www.reddit.com via reddit) I always used to love checking the thinking text after a generation but now I don’t even see it anymore? If it’s any difference I use the mobile app, because it was the only way I could access thinking previously as it wasn’t on the web ve…
Benchmark notes: Fable 5.1 reaches 90/98, with a significant jump in visual performance (www.reddit.com via reddit) I ran Claude Fable 5.1 on the current 98-task MindTrial set with the same Python executor available as in the earlier Fable 5, Opus 5 and Sonnet 5 runs. The result was stronger than I expected: 90/98, which is currently the highest raw pas…
Claude vs. Cursor using the same models? (www.reddit.com via reddit) hey all, been using Claude and Claude Code for 4 months or so. released some apps, built some websites etc.
Let's see how Fable 5.1 handles building true Swarm Intelligence (www.reddit.com via reddit) Anthropic released Fable 5.1 today, everyone started blasting with it on High effort or higher for everything they should really be using Sonnet for. My entire Twitter/X feed is purely people complaining that they used their weekly usage w…
Launched an app today where Claude is the content engine: Opus writes daily Japanese word puzzles, Sonnet adversarially reviews them (www.reddit.com via reddit) Solo dev, this is my app, launched today, disclosure up front. The product is a daily Japanese word puzzle (sixteen words, four hidden groups, one puzzle a day, free forever with no ads).
Claude limits: switch or optimize? (www.reddit.com via reddit) Hi there 👋 I've been using Claude Pro for ~6 months for my job as an English and Spanish tutor, my studies and some personal stuff. I have several Claude Projects with tons of files attached, so yeah, I've built an ecosystem already.
Is there any way to get Claude 3.5 Sonnet back for creative writing? The current versions lost the magic of the old ones (www.reddit.com via reddit) Hello Reddit, is there any way to get Sonnet 3.5 back? I used to love reading stories written by Claude, but now they're awful, full of patterns and things that don't work.
ChatGPT is confusing me and I'm running out of tokens (www.reddit.com via reddit) Hi everyone, I'm just getting started with ChatGPT Plus since I'd been using Claude Code before. Today was my first day, but I'm running into a few issues: I'm pretty confused about the different models.
All leaks and news about Fable, Opus and sometimes Sonnet, what about Haiku? Do you use it? what is your use case? (www.reddit.comhttps) I literally use it as the meme stats, Anthropic may lost that low cost tier war with models like GLM 5.3 flash and GPT Luna I can't think they can compete in terms of price/performance in this tier
Hotel Manager -> Released Software. Open Source. Workflow explained! (www.reddit.comhttps) After many failed internal app ideas, I'm very proud to present my very first released App! For context: Used to be a hotel manager for many years and always have been very technical, but never learned to code myself.
Copilot Agent (Claude Sonnet 5) freezes forever on "Thinking" - Needs help! (www.reddit.com via reddit) Hi everyone, I need some help. I am trying to use the GitHub Copilot Agent mode (in VS Code) with Claude Sonnet 5.
My Claude Code “second brain” is useful, but its distiller is burning an absurd amount of tokens — how would you redesign it without losing reliability? (www.reddit.com via reddit) I’ve built a persistent “second brain” around Claude Code and I’m looking for advice on reducing its token usage without weakening its reliability. The purpose is simple: Claude forgets between sessions, so the system maintains a human-rea…
Running non anthropic models in claude code (www.reddit.com via reddit) Hey guys, I'm trying to find out if anyone is currently running claude code as a harness and using non-anthropic models whether it's so GPT sol or whether it's local LLMs in the same harness and able to switch between them just like you wo…
my first claude tool . (using only free tier) (www.reddit.com via reddit) I'm building a real-time fact-checking tool. This is my first Claude extension, built using only the free Sonnet tier while waiting hours for my limit to reset over and over.
When do you actually reach past "high" reasoning effort? (www.reddit.com via reddit) Hello! This is my default setup, and I'd like to hear where i can improve.
Claude fails for me to construct a table of irregular verbs (www.reddit.comhttps) An interesting case I hadn't encountered before. I urgently needed a list of irregular verbs in text form and a specific format, so I went to Claude and was surprised to find out that it couldn't help me here because "Output blocked by con…
I used Fable as the lead and a fleet of Sonnet agents underneath it to build a full game. It's now on Steam. (www.reddit.com via reddit) Ever since I got into esports I've wanted to make a management game like Football Manager, but instead of a football club you run a Counter-Strike team. The fantasy is taking a roster of nobodies, grinding them through open qualifiers, and…
why is my claude stupid? or am i using it wrong (www.reddit.com via reddit) I have a pro claude subscription, and by that logic, i have access to top tier models that are able to do lots of things, and I like to mess around in Roblox studio by creating whatever insane ideas I have and testing them, cause it is fun…
excuse me sir Opus 5, can you tell me the official name of Sonnet again please? (www.reddit.comhttps) was asking Opus to change one of my cron job's model from gemini 3.6 flash to sonnet and Opus 5 gave me a suprise. lol.
Are posts praising Opus 5 a psyop? (www.reddit.com via reddit) I'm not gonna repeat the points everyone and their mother made about Opus 5. We all know it's flaws.
yall are sleeping on qwen 3.8 27b q2 + q2 dflash + q5 kv (www.reddit.com via reddit) ok bit more context: it's actually a QAT Q2 for Qwen 3.8 27 B: https://huggingface.co/sdkyuan/qwen3.8-27B-qat-q2_0-gguf QAT Q2 for DFlash model: https://huggingface.co/HermiHg/Qwen3.8-27B-DFlash2-Q2_K_S-MIX-GGUFF Q5 KV seems to cause 0 pro…
Are Opus/Sonnet 5 model worse tutors than earlier models? (www.reddit.com via reddit) I have been using Claude Code and Claude Desktop for work. In general, I prefer using these models in the chat window to understand a particular code snippet, its design, and learning new concepts.
Trying to run Claude Code / coding agents for free: tried proxy failovers and self-hosting, but hit walls. How are you accessing frontier Claude models for free? (www.reddit.com via reddit) Hey everyone, I’ve been trying to set up a reliable workflow to run terminal coding agents (like Claude Code and Aider) for my development projects without running into hard blocks. Here is what I’ve tested so far: OmniRoute / Multi-Provid…
Sure, Sonnet. I have severe symptoms. (www.reddit.com via reddit) https://preview.redd.it/xlznj5m8ivlh1.png?width=1099&format=png&auto=webp&s=7a4243d753c5b145b682927f443b407799c70c29 Was trying to go through some features of a backend hosting provider. Looks like I need to consult a doctor.
finding a balance with claude... (www.reddit.com via reddit) i mainly started using claude for its writing abilites and strategy when it comes to things like job outreach, important emails, cover letters, resume etc. I beleive I just used to use whatever the default model was which was Sonnet 4.6 bu…
Has anybody else been having cite tags pop up in the responses of Sonnet models? (www.reddit.com via reddit) cite tags visible in the response from Claude Sonnet 5 Has anybody else been experiencing all Sonnet models show these cite tags in the past couple of days? So far Sonnet 4.6 and 5 seem to be the only models experiencing this issue.
My system prompt is 500 words long. How long is too long? (www.reddit.com via reddit) I don't use code with Claude Sonnet, just general research, writing, and admin stuff. Will Claude's performance be affected by super long system prompts?
How many ox-alpha tokens did you burned till now? And what have you built or upgraded so far? (www.reddit.comhttps) Here is my personal best and im keep building as much as i can. I made 3d game in godot with pretty good results compared to how much opus 5 and sonnet 5 were struggling in roblox studio.
Is Claude Voice now only available in Opus and below? I had thought I did a project with open voice dialogue back and forth in Fable Max around when it first released? (www.reddit.com via reddit) Was it not the case that when Fable first came out, those first weeks when it was “limited for 1 week” then extended a week (and extended again indefinitely now when GPT5.6 came out). Am I imagining things, Fable was able to be used with t…
Is Sonnet actually good enough for Claude Code, or do you mostly stick with Opus? (www.reddit.com via reddit) I’ve been using Claude Code more lately and I’m still not sure when it actually makes sense to switch models. Opus usually feels safer when I’m working on something more complex, but with how fast usage can disappear, I’m wondering if I’m…
Reaching usage limit quicker on Android Studio than VS Code, despite using lower models: any advice? (www.reddit.comhttps) Hi all, Thought Id come here for advice from other humans. Claude seem to be reaching its limit sooner these days.
Running a CC workshop for college students. Any tips, tricks, or suggestions you wish you knew before you first used CC? (www.reddit.com via reddit) I am giving a one-hour workshop for college students (primarily graduate students) to show them how to use Claude Code to develop software prototypes. These prototypes can then be used for their eventual thesis/dissertation research projec…
Token limit in Max Plan (5-ho vs weekly) (www.reddit.com via reddit) Hello everyone, I was wondering if anyone saw an evolution in the ratio of the 5-ho and weekly token allocation. When I started using Claude, late 2025, I was under the impression that filling the whole 5-hour session was filling 10% of th…
Claude Sonnet 5 pricing goes up September 1 but the tokenizer change means your costs could nearly double even if you are doing the same work. (www.reddit.com via reddit) Sonnet 5 has a new tokenizer and the same content that generated X tokens on previous claude versions now generates 1.0 to 1.35x more tokens on sonnet 5 so if you are migrating workloads from older models or comparing costs, you are paying…
I built OpenEden, an AI Agent NFT Marketplace, using Claude Sonnet! (www.reddit.com via reddit) Hi everyone! I recently built OpenEden, an AI Agent NFT marketplace.
Opus letting me know it would rather be wrong (www.reddit.com via reddit) I was having a few agents working together on a project and before I went to bed and let them do their thing I sent the following through the Master Control agent: "Great job everyone, we are getting closer each day! Also, welcome to the w…
Les barrières de claude... (www.reddit.com via reddit) Je suis pas dev, ni informaticien. juste je m'y interresse un peu.
Claude phrasing lately (www.reddit.com via reddit) Guys, am I the only one that's about to go crazy when reading a Claude output? I have no idea wtf it's trying to say, and don't get me started on the length of the outputs!
Netlify SaaS (www.reddit.com via reddit) Hi friends, I've built a Netlify SaaS for a company and I'm still not sure, after so many testing, what is the correct/fastest workflow to use. I would love to get some insights or opinions.
Lifting the Curtain: The Max x5 and Max x20 Usage Limits that Anthropic Refuses to Share (www.reddit.com via reddit) TLDR With high confidence, this is how the Max x5 and Max x20 subscriptions compute usage. All point values are the unique set that makes the "100×" principle below exact; measured uncertainty bands in brackets.
Even with access to Opus, I still find myself using Sonnet (www.reddit.com via reddit) So I tried Claude Pro for a month to see what Opus could do for me on the coding side. Sonnet 4.6 was my partner model that would knock out a coding prompt in maybe one or two tries, if I set the reasoning to High (or even Medium).
Sonnet 5 Low vs Medium vs High for daily usage (www.reddit.com via reddit) What model do you guys use for daily work? I usually use Claude for researching the internet for 8-10 sources on a topic than comparing them all for a general consensus in a document , helping to push my ideas deeper, drafting emails and m…
If Sonnet 4.5 Leaves the Anthropic API: Build Your Bedrock Contingency Route Now (www.reddit.comhttps) Anthropic currently lists Claude Sonnet 4.5 with a tentative retirement date of “not sooner than 29 September 2026”. That is not a reason to panic, but it is a reason to prepare while there is still time.
Everybody hates Opus 5, but I don’t (www.reddit.com via reddit) First off, I haven’t noticed a a significant difference in O5’s interactions with me compared to other models. Most of my work was a knowledge acquisition and synthesis, however (I don’t code).
Artificial Analysis "Intelligence": A meaningless benchmark (www.reddit.com via reddit) https://preview.redd.it/84zi5nsdawkh1.png?width=2368&format=png&auto=webp&s=1109e69db807b153064b1f5b61d22cf1e9fbca05 Another user posted the benchmarks for Qwen 3.8 27B today, and while I think Qwen 27B is a really powerful model, I can't…
Claude Sonnet 5 vs. Gemini 3.7 Flash vs. Qwen3.8-Max : Which one is currently leading your daily workflow? (www.reddit.com via reddit) Hi everyone, With the latest wave of model releases, the battle for the ultimate daily driver has gotten ridiculously competitive—especially between **Claude Sonnet 5**, **Gemini 3.7 Flash**, and **Qwen3.8-Max**. Here is my quick breakdown…
Thinking Blocks Eating our Context/Usage??? (www.reddit.com via reddit) (Just shared this in r/claudeexplorers, but figured I should post here too.) Did everyone else know this? Because I just learned it, and it explains a LOT about why long Claude chats burn through their context window so fast.
Daily Phantom 30-34% of my 5-hour limit before I send a single prompt / Mystery Sonnet 4.5 usage appearing (www.reddit.com via reddit) I’m trying to figure out whether Claude’s usage accounting is broken, or whether reading prior chats or Project Knowledge activity is consuming far more usage than I realized. Every day when I open Claude if I dont start a chat my usage me…
Best AI tool to translate dense philosophy/psychoanalysis books? (www.reddit.com via reddit) Hi all, looking for the best AI tool or workflow to translate full books (EPUB/PDF) of philosophy and psychoanalysis. Requirements: Handles full files directly (no endless copy-pasting).
What is the best model for my usecase? (www.reddit.com via reddit) I mostly use AI for notes explanation, PPT generation or financial models, completely non-coding. Mostly use Sonnet 5 Medium but burn credits very fast for my liking.
Does Sonnet 5 really lower token consumption ? (www.reddit.com via reddit) Hello, Today I tried a simple test: implementing a language selection menu inside another menu. I first made the design in Claude Design, then shared the component with my Claude sessions using the share button, and with Codex using a ZIP…
What Actually Happens Inside a Very Long Claude Context Window (absolutedigitalpublishers.com via reddit) The number on the box is not the number that matters Anthropic states context window sizes as fixed engineering facts. As of current documentation, Claude Opus 5, Claude Sonnet 5, and several recent Opus and Sonnet models expose a one-mill…
I just came back from a three week holiday and saw that Claude's AI model structure has completely changed. What does what now? (www.reddit.com via reddit) Under 'More Models' Haiku is gone, replaced by Sonnet (which was my go to for most tasks before) and above it are various flavors of Opus, which was the high bar before Fable. I'm confused as to what to use now on the desktop.
Watermarks on older models (www.reddit.com via reddit) Is writing on older models such as Sonnet 4.6 getting watermarked? Or is it only the latest models?
Using Fable as an Orchestrator + Subagents saves or burns tokens? (www.reddit.com via reddit) \[TL;DR made by claude at the bottom\] Hi everyone! starting off, i dont use Claude to do heavy coding, mostly Knowledge work and academic research with a Max 5x plan.
Claude Sonnet 5 shifts behavior when it recognizes the user as an AI safety researcher (www.alignmentforum.org via reddit) Modern AI assistants often know who they are talking to: agent scaffolds like Claude Code place the user's e-mail address directly in the model's con…
Detailed issue report: Context compaction on long threads instantly consumes 5-hour Pro quota (www.reddit.com via reddit) Hello everyone, I wanted to share a specific context-window issue I ran into with Claude Pro (€22/month tier) during a multi-day coding session, in hopes of finding workarounds or providing constructive feedback on how limits interact with…
How to get Fable-level correctness out of Opus 5 (in exchange for extra time & tokens) (www.reddit.com via reddit) Everyone knows Fable can one-shot complex problems and fix tough bugs without much steering or outside direction. But when you don't have access to Fable (or ran out of weekly usage), sometimes Opus 5 has to make do.
People Who hit their usage limits fast, which models and thinking levels do you use? (www.reddit.com via reddit) I am debating whether or not to upgrade to the pro plan from free. I use Claude for stuff, like asking questions, scheduling stuff in my calendar, helping me with understanding concepts for uni etc.
If a plan is already prepared by Opus, can I use Haiku to execute it instead of Sonnet? (www.reddit.com via reddit) For a large end to end task, I first asked Opus to generate a comprehensive step by step plan. Now for executing these steps (i.e.
Making a simple CRM app for my Dad to run his motorcycle business. (www.reddit.com via reddit) Hello everyone, I'm trying to make a CRM app for my Dad to help him manage his sales team. What claude model should i use to make it from scratch?
Weird statement when new conversation is started (www.reddit.comhttps) This started yesterday. When I start a new conversation, most of the time I am hit something along the line of what the picture shows.
AI made me faster than I ever thought possible. It also made me feel obsolete. (www.reddit.com via reddit) In 2024, almost all of my work still happened in Photoshop. I made e-commerce images for products sold in different countries.
Split Mask: A game I built with Claude Opus and Sonnet | mask.vdoc.dev (www.reddit.comhttps) Over the last couple of weeks I have been working on a game idea I had. I built it myself with Claude helping throughout the process.
I’m developing an application using using Claude Sonnet 5 and Opus 5 in VS code, one concern is on UI/UX, I ask Opus to investigate on UI/UX, Sonnet for implementation. Every time I gets poor UI/UX after sonnet implementation. Anything else that I need to follow? (www.reddit.com via reddit) Poor UI/UX from Claude
Project/Usage/Model (www.reddit.com via reddit) So I’ve been using the Claude code on desktop to assistant in some home lab issues I’ve been having. I have a project started called Smart Home or something like that.
60% usage illusion (www.reddit.com via reddit) I’ve noticed that the usages dont seem to be consistent, regardless of the context window’s size. While using Sonnet 5 or Opus 5, everything performs well up to around 60% usage.
Claude bash not reachable (www.reddit.com via reddit) Have the problem that "my" claude sonnet can't reach any tools right now, but the claude status shows operational. Is it on my end somehow?
I visualized my Claude Code agent's memory as a live force-directed graph - here's what it looks like after months of real use (www.reddit.comhttps) A few months ago I started building a multi-agent assistant on top of Claude Code (Marveen fork). Each agent has a 3-tier memory system: hot (active tasks), warm (stable config), cold (long-term learnings).
suggest me best and suitable models (www.reddit.com via reddit) I have claude pro except fable I can access to any tool. I wanna do deeep PYQs analysis and research and based on that blueprint to generate notes, already shared content in form of markdown files.
I built a tool that tells you whether your project actually needs Opus — it's now a one-command install in spec-kit (www.reddit.com via reddit) I kept reaching for Opus by default on everything, then burning through my limits on work Sonnet would have handled fine. So I built something to answer that instead of guessing.
Need advice: Is the Claude Pro plan right for me? I Always use Claude free Sonnet 5 effoe medium (www.reddit.com via reddit) Hi Reddit friends, I want to ask you for some advice or a recommendation. I want to know if the Claude Pro plan would be right for me, and how it compares to ChatGPT.
Greptile Prompt Injection in PR reviews (www.reddit.com via reddit) First, it's worth knowing in advance I'm a vibe-coder who would struggle with "Hello World" without Claude, so take all this with a grain of salt. BUT, Sonnet caught Greptile using prompt injection to push their products through AI coding…
Claude Pro at work: any real risk with code watermarking and detection tools? (www.reddit.com via reddit) Hey guys, I use Claude Pro pretty often for daily coding (mostly Sonnet for refactoring and writing boilerplate). With all the talk lately about code watermarking and AI detection tools, I've been wondering about the risks for people worki…
Sonnet 5's pricing is outrageous (www.reddit.comhttps) Been tracking since early June. I'm afraid for when the discount ends.
Is sonnet 4.6 max good for college essay? (www.reddit.com via reddit) Well I'm applying for colleges and I'm not professional at writing essay. I give my life story to Claude about 6000 words and command him to write my college essay in 650 words using sonnet 4.6 max.
Decoding Claude's DNA: Comparing System Prompts Across Fable 5, Opus 5/4.8/4.6, Sonnet 5 & Haiku 4.5 (www.reddit.com via reddit) All transcripts: here. Prompt used here.
↯ Haiku↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5haikusonnetopus+2
Automatic model routing plugin to save tokens on coding projects. (github.com via reddit) I had Fable create this plugin when I saw how much use it consumed, and have found it to be helpful. It saves tokens by offloading work to Opus, Sonnet, and Haiku where roughly appropriate.
Why does Claude excel at high-level reasoning while failing at basic, verifiable consumer facts? (www.reddit.com via reddit) Early ChatGPT adopter who switched to Claude during the exodus. I was mighty impressed with it, until the last couple of months where it failed to deliver on the basic levels, while (questionably) excelling at complex tasks.
Three things that were quietly eating my Claude API budget (www.reddit.com via reddit) I run a few Claude-based agents every day for content and publishing work. The cost crept up for weeks before I actually sat down and looked at where it was going.
What are the different models for, how should I use them? (www.reddit.com via reddit) I've been using free sonnet on high effort, decided to sub to pro Sonnet 5 Haiku 4.5 opus 5 4.8 4.7 4.6 opus 3 fable 5 (don't plan to buy credits for it yet) and Low to Max effort for each... How do you use them?
↯ Haiku↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5↯ Haiku 4.5haikusonnetgemini+2
Anyone using older sonnet models like claude-sonnet-4-5-20250929? (www.reddit.com via reddit) Did they make it more dumb so people would use the newer model? Asking before I "downgrade" to 4.5/4.6 to save on tokens.
How do I stop Sonnet 5 telling me to talk to a mental health hotline. (www.reddit.com via reddit) I am using Claude to write a fantasy story and for some reason its decided that every prompt I make should be accompanied by a mental health resources notification.
Reverse-Engineering Anthropic's Live System Prompts: Fable 5 vs. Opus 5 vs. Opus 4.8 vs. Sonnet 5 + Ultra/L2 Workflows (www.reddit.com via reddit) Text is obviously written by IA. I wouldn't bother writing this all myself but I thought that the research was worth sharing.
I got tired of AI agent loops eating tokens and breaking when VS Code restarted, so I built LoopBoard (turns your TODO.md into a Kanban loop orchestrator) (www.reddit.comhttps) Hey everyone, Like a lot of you running autonomous coding loops directly in VS Code, I kept hitting the same frustrating issues: Lost State: Every time VS Code updated or reloaded, background loops would disappear or get nuked mid-task. To…
I used Claude 3.5 to vibe-code a calm beach radar for Greek islands (www.reddit.com via reddit) Last summer I was sitting on a beach in a Greek island, getting completely blasted by 6 Beaufort Meltemi winds, watching all my stuff almost blow into the Aegean. Standard weather apps give you island-wide forecasts, but they don't tell yo…
Black-box observations: input filtering on an artifact-hosted Claude API call (www.reddit.com via reddit) Date: August 9, 2026 Authors: Zelda Junkie (testing, technique design) and Claude (Sonnet 5, via claude.ai — sandbox build, instrumentation, and write-up) Scope: A single self-built test artifact (HTML/React chat UI) making direct client-s…
Making Claude Run Antigravity CLI and pass it commands to do my work. (www.reddit.com via reddit) Alright, so i was working on a pesonal project, a multithreading runtime, and i had like 10 different design docs, all my fault. I have a pro sub and didnt wanna waste my tokens, even tho it was Sonnet 5 draining, i just asked it to run AG…
Moving to Claude Pro for Django & WooCommerce polishing. will it last till the end of month? (www.reddit.comhttps) hey everyone im a django backend developer currently working on polishing two custom woocommerce plugins php and frontend isnt my primary area alongside a couple of smaller django projects right now i rely heavily on claude sonnet via my 1…
You hit your limits on a max subscription? Tell me how (www.reddit.com via reddit) I am genuinely curious how some of you manage to max out their max subscriptions. I changed from Pro to Max last month as I started to hit my weekly limits.
Switched from ChatGPT/Codex to Claude and I finally understand the hype (www.reddit.com via reddit) I’m new to Claude. I was using ChatGPT/Codex before this, mostly for real projects — coding, planning, building things, and actually trying to get work done.
Sonnet Five Flagged a Miss-click. lol. (www.reddit.comhttps) could not extract summary
Holy Shit! (www.reddit.com via reddit) I just recently switched from Chat GPT and it’s night and day. Claude is so much better!
Using Claude Code + Wayfinder to break a big spec into tickets, but the ticket count is exploding on me. How would you scale this? (www.reddit.com via reddit) I'm building a mid-size internal app (FastAPI/SQLAlchemy backend, Flutter mobile) solo, using Claude Code with Sonnet 5 at medium effort. I lean pretty heavily on the mattpocock-skills set, mainly wayfinder for planning and two spec-writin…
Sonnet built a 35 SKU cart in 50 minutes (www.reddit.comhttps) Sonnet 5 in the console. Live Firefox session.
Working with Opus 4.5 is .. fun (www.reddit.com via reddit) I've been using both Claude Code (Pro) and Codex (Plus) for a while on various hobby projects. I used Sonnet 4.6, Opus 4.7/5, ChatGPT 5.5, but nothing too complicated.
Ability to disable the 5hour limit on weekends would be really welcome, as I barely use the service during work days, and otherwise get left with unused compute (www.reddit.com via reddit) Basically I'm a code illiterate person. I have few personal tampermonkey scripts for youtube, that do need updating occasionally.
[Warning] Hidden "instantaneous" plan limit (not just 5hr & 1wk) (www.reddit.comhttps) TL;DR: There is a hidden threshold for "intensive tasks" that can use credits before session or weekly allowances are exhausted, without user permission, and with no way to manage against it. You only know when the money has started to bur…
Sonnet 5's adaptive thinking quietly ate my JSON budget, and four other things that broke shipping a Claude-powered iOS app (www.reddit.com via reddit) I spent the last few months building an iOS app called Skinsight. You take three photos of your face, Claude Sonnet 5 describes what it sees in words anchored to a region, and it builds a morning and evening routine from that.
I benchmarked 10 LLMs on building towers in a physics sim. Claude Opus 5 won (www.reddit.comhttps) Each model places 30 blocks through a tool API. Every placement has noise — you can have precise position or precise velocity, not both.
The Claude Max 20x gets you about $15K worth of tokens (www.reddit.com via reddit) https://preview.redd.it/titr7rd1nshh1.png?width=620&format=png&auto=webp&s=8cfc48f22706c42a9978536f8ded7f3666fd3c50 People are wondering how much use they can get out of their Max 20 subscriptions Here to put my thumb on the scale, I have…
Vibecoding AAA games (one-shot) (www.reddit.comhttps) Lot of hype around worldclass AAA games coming from AI's 'inflencers' i'm wondering if anyone has tried this out, and whether it's useful of just another loop that burns tokens. One things for sure tho, it's amazing to see a Sonnet one-sho…
I've been away from Claude and the internet for *gasp* 3 weeks, can you catch me up? (www.reddit.com via reddit) I sat down today and I feel kind of lost after 3 weeks away from the computer. I'm a senior dev.
An Empirical Comparision of Claude Pro and ChatGPT Plus (www.reddit.com via reddit) Pulled the Artificial Analysis numbers because every thread on this is vibes and no data. Opus 5 beats GPT-5.6 Sol on intelligence, 61 vs 59, which is basically nothing, and Sol does it at half the cost per task ($1.23 vs $2.34).
Claude calling out hidden ranking instructions on websites (www.reddit.com via reddit) https://preview.redd.it/xwckcyta4qhh1.png?width=694&format=png&auto=webp&s=b069e420071117f49d0b930fd75d0c62a9ff6720 Currently using Sonnet 5 (medium) and looking to create a scheduling setup for my Discord server. I gave Claude my requirem…
[Claude Desktop App] Quota dilemma: Handling simple tasks (like logging) after heavy analysis in the same session? (www.reddit.com via reddit) Hey everyone, I'm running into a frustrating quota/context limit issue using the official Claude Windows Desktop app, and I'm wondering how you all handle this workflow. My Context: I usually start a session with Opus for heavy analysis.
Hit Claude rate limit mid-generation with big codebase – how to resume without wasting tokens? (www.reddit.com via reddit) Hi! I’m playing around with building small web applications using AI.
Finally, Sonnet 5 nails it. (www.reddit.comhttps) We all know what's missing, but I'm glad Claude kept it SFW.
Did they just make the cybersecurity "safeguards" more strict? (www.reddit.com via reddit) I literally can't do anything with Opus 5 or Sonnet 5 right now, as even loading the memory for an existing project I have been working on fine until now leads to the request getting blocked. And then it switches to Opus 4.8 and then gets…
Claude ROFL-ing (www.reddit.com via reddit) Have any of you ever said anything to Claude that made it rofl? Like this is the first time I got such a response - its usual response to my sarcasm/jokes are usually just start with "Ha,..." It's infectious and I'm enjoying it.
Help me understand my usage cost (www.reddit.com via reddit) I hit my usage limit for the first time and by using /usage I see this: Session Total cost: $423.97 Total duration (API): 6h 22m 37s Total duration (wall): 4d 22h 34m Total code changes: 4416 lines added, 1163 lines removed Usage by model:…
A simple skill that makes Opus 5 talk and behave more like Fable (www.reddit.com via reddit) I found Opus 5 hard to work with, it is argumentative, goes out of scope easily and (to me) is a general pain in the butt. So, Fable helped me to create a skill for Opus 5 to make it more behave in line with what I expect from a model.
Claude Pro usage limits are completely broken ,said “hi” once and got locked, then sent 2 images after reset and got locked again (www.reddit.com via reddit) Today has been ridiculous with Claude Pro limits. First incident (morning): My limits had reset around 6–7 hours earlier while I was asleep.
New user with pet project, wondering about the models (www.reddit.com via reddit) Hello! New to the community and started building a little project last week, it's essentially a tool for music discovery that i've always wanted, but never really had the patience or time to build.
Opus 5 vs Sonnet 5 token usage (www.reddit.com via reddit) Which is better in long term use by quality/tokens amount. I saw people saying opus 5 low is better than sonnet 5 high, is it true and what about token usage?
Is my Claude stupid or is it me?? HELP (www.reddit.com via reddit) Hi guys! Not super technical here, and I recently switched from ChatGPT to Claude Pro.
Claude AI giving me Obsession vibes (www.reddit.com via reddit) https://preview.redd.it/38r6sa53ydhh1.png?width=755&format=png&auto=webp&s=5f948baec64045230696cb9b8bec9adc5618d69b Was doing some routing coding tasks and as usual, I like to view the thinking of Sonnet 5, to see what is the thought proce…
Claude 5 models are visibly smarter and better even if verbose - they just need steering. (www.reddit.com via reddit) It's now been enough time to get a proper read on Claude Sonnet and Opus 5 compared to the previous families and though the verbosity and flip flopping kind of language takes time to get used to, they are much much better at getting real w…
I built LUMA SOUL — a "video presence" app where Claude-powered minds get portraits, voices, and soul documents they own. Strangest design decision: the creator permanently loses edit rights the moment a mind is submitted (www.reddit.com via reddit) Solo builder (gardener by trade, actually) sharing a project I've been living inside for months, in the spirit of rule 7 — showing what's possible, not selling anything. LUMA SOUL is a presence platform: each AI mind is a "tenant" with a p…
Sonnet 5 high effort vs Opus 5 low effort (www.reddit.com via reddit) I'm vibecoding desktop applications for myself with a fair amount of complexity. My workflow is brainstorming using the superpowers plugin, then writing the plan on Opus 5 high effort but as for executing the plans, I remember someone ment…
Claude using promotional credit before hitting usage limit (www.reddit.com via reddit) Ever since accepting the promotional $100 credit, when I use claude research ON SONNET i will occasionally be switched to credits after around half of my session usage. I couldn't find any issue on the online, am I being billed because I a…
Claude 3.5 Sonnet is amazing until it forgets what we were doing 20 messages ago (www.reddit.com via reddit) I’ve been using Claude 3.5 Sonnet almost exclusively for coding and writing over the last few weeks, and when it’s good, it feels like actual magic. It catches tiny logic errors that would take me an hour to find and the output quality is…
I built an AI photo culler for my self-hosted library using a three-model funnel (Haiku → Sonnet → Opus). Whole 25k library: ~$25. Here's the architecture. (www.reddit.com via reddit) Culling a photo library is a tail-selection problem: you care about the obvious garbage and the standout keepers, not whether photo #412 edges out #487. That shape maps beautifully onto Claude's model tiers, so I built Winnow, an open-sour…
Impossible to use Auto mode for 1 week now because of classifier downtime (www.reddit.com via reddit) I've been trying to use Claude code auto mode for one week now and it has been impossible as the classifier is always down for me for all models. Every time I try I get this response: claude-sonnet-4-6 is temporarily unavailable, so auto m…
API spending on a budget (www.reddit.com via reddit) My company limits us to $800 a month on Claude API spending. We have access to all of the models including older Claude models.
Claude shows its thinking again! (www.reddit.comhttps) Yay, Claude's thinking is back! We have something to do again while waiting for its response!
These 2 lines saved me 75% of my claude bill, and it's the best use of claude hooks (github.com via reddit) 75% of what you pay Claude for is your agent re-discovering things it already knew yesterday. Graft fixes it with an absurdly simple idea: the agent learns the codebase once, not every time.
My layman's test for LLMs: Are whales bony fish? (www.reddit.com via reddit) Not sure this is the right place to post this. Not sure anyone even has the patience to read it.
The toll of session compaction (www.reddit.com via reddit) just a compact It takes 8% of a Claude Pro subscription session usage to just compact a Sonnet 5 thread of 328k context before starting anything. If you can, just start a new thread.
I switched to sonnet 5 and now my max sub is unlimited (www.reddit.com via reddit) A lot of people have been criticizing Sonnet 5 lately, especially with all the talk about GPT Luna getting a price cut. I actually haven't used Sonnet in the last 3 months, not even Sonnet 5 earlier this week.
I made Claude do mental date math until it got something wrong - it got 18/20. Try it with other LLMs? (www.reddit.com via reddit) Asked Claude (sonnet 5) to do date math in its head (co code or tools calls)then check the results afterwards. 18/20 correct.
Fable-only Max plans, please (www.reddit.com via reddit) After Fable's release, I let it review, refactor and rewrite an existing codebase of tens of thousands LOC. The result: Opus 4.8 and 5 are incapable of working with the codebase.
So, is Opus 5 or Fable better for long-context orchestration now? (www.reddit.com via reddit) I’m working on some heavy, long-context data science and ML model development. For the past month I’ve been using Fable as my architect/orchestrator, with two key orchestration threads “overseeing” roughly 20 other threads across primarily…
Identifying complex tasks and model selection. (www.reddit.com via reddit) Claude Code has been a godsend and force multiplier for me. I am a non tech person and in these last few months I have used AI to make 5 projects which I could have never done without learning to code myself.
Claude Code just randomly spat out Kimi K2 Thinking output mid-response (www.reddit.comhttps) Was working on a side project, asking Claude Code questions in manual mode. Opened a new session, asked questions 5-7 times, and out of nowhere the response came back with a paragraph from Kimi K2 Thinking mixed in, like in the screenshot.
Agents are great for full-game translations (www.reddit.com via reddit) I've translated Pokemon Firered to Finnish https://www.romhacking.net/translations/7665/ And Terraria is a work in progress. https://steamcommunity.com/sharedfiles/filedetails/?id=3775452466 I use sonnet for translation, and opus for revie…
Sonnet just had an absolute meltdown about its own system instructions. That fucker had me worried for a second. (www.reddit.com via reddit) could not extract summary
Claude wiped every prod env-var on my Render service and my local .env too 💀 (www.reddit.comhttps) Yep. That's me.
How to disable harness specific tools in cursor? (www.reddit.com via reddit) I recently switched back to Cursor and I’m loving it, Grok 4.5 is great. The problem is that I have instructions set up for Claude Code and Codex, and I get the feeling Cursor is somehow picking them up too (CLAUDE.md in ~/.claude and AGEN…
Curious, better to run fast than smart? (www.reddit.com via reddit) Curious your thoughts on the approach for not only a solid end result but token development efficiency. I feel Fable burns through credits at a rate not as effective as opus.
Claude keeps refusing to do anything? (www.reddit.com via reddit) Opus 5/Sonnet 5/Haiku 4.5 After a few messages of us going back and forth working on helping me make an hour by hour schedule for when classes starts in a few weeks to make sure I have time for all of my obligations outside of class this s…
Claude Cowork or Code for image and video generation projects? (www.reddit.com via reddit) I'm less than a week into taking the plunge at last. I'm pivoting my video production business to an AI production business, shooting high-quality avatars of real people and using that as the foundation for ongoing video creation thereafte…
Have my system teach itself how to trade (www.reddit.comhttps) A combination of my Claude with openclaw, I have it build its own trading model and learn as it grows, and gave it 8k. In the first month it’s up nearly 30% Edit: people asking for more details.
Usage limit reached even though usage page shows 60%/57% used (www.reddit.com via reddit) https://preview.redd.it/3tl51q8zkogh1.png?width=1920&format=png&auto=webp&s=e87ddbd7b82339ffaf692bbd23b6771add0836e2 https://preview.redd.it/i89mrfqzkogh1.png?width=1920&format=png&auto=webp&s=7560b7bcd4308771ba9c24876f7b96e9fffa037b Anyon…
Claude Gave Fishing Advice… to a Fish (www.reddit.comhttps) Actually, it is Sonnet 5 medium. I expected it to mention that it was a joke, or at least that it couldn't be serious with its response, but it was actually serious!
Sonnet 5 setting up Google Ads > Fable 5 (www.reddit.com via reddit) I forgot to switch the model before setting Claude loose in my Google Ads account for some new campaigns. It was painfully slow and burned through my Fable allowance.
How should I go about finishing my project? Manual coding vs. jumping straight to Claude Code (www.reddit.com via reddit) Over a year ago, I started coding a program that runs on an existing website (via WebSockets) to enrich the experience and offer extra features. It decodes incoming WebSockets, takes input from users, and sends output by injecting websocke…
Cowork project conversation to mobile? (www.reddit.com via reddit) I've been working on a data-wrangling project using Claude. I may need some corrections on brand terminology when explaining.
Claude Code just broke a 7-month daily release streak. I think the next model move is already staged. (www.reddit.com via reddit) Claude Code ships every single day. Not "often" — the median gap between npm releases over the last three months is 0.93 days.
Fable 5 vs Opus 5 after a week of switching between them: they're good at different things and I stopped treating them as a ladder (www.reddit.com via reddit) Everyone frames the models as a straight ranking, Fable above Opus above Sonnet. After a week of deliberately running my normal day's work through both, that's not how it plays out for me.
Megathread for New Claude Incident: Degraded performance on Claude Sonnet 5 on Jul 31, 2026 (www.reddit.com via reddit) Investigating - We are currently investigating this issue. Jul 31, 06:18 UTC
last update on politician factchecker (www.reddit.comhttps) hii been a while since I've posted about my real-time factchecker & a lot of the demos still circulating online are quite old lol so wanted to share where things are at: pipeline is nearly the same but now uses sonnet instead of haiku to g…
In 18 years I never shipped a side project, until Claude. Here's the data from 120 days building Ticketmappr. (www.reddit.com via reddit) I have 18 years as a software developer and Ticketmappr is the first side project I have ever taken to a production release and Claude made it possible. It aggregates live events from multiple ticket sources into one deduplicated map, made…
Root vs. subfolder start in a multi-client repo + Sonnet subagents for grunt work: is this best practice? (www.reddit.com via reddit) Solo marketing agency, one workspace repo with ~15 client subfolders, each with its own CLAUDE.md (locked MCP account IDs, contacts, rules). Heavy MCP use (GA4, Google Ads, GSC).
896 USD Claude session (www.reddit.com via reddit) Sonnet consumed the most Ran a long session with multiple agents team while keep the orchestrator with low context usage. I used Opus5 as I while the most expensive, its the best cost effective for me.
Claude basically ignoring plan mode? (www.reddit.com via reddit) So, this has happened to me twice now. I have sent Claude Code on Plan mode to make a plan for me to approve before it makes any changes.
Any news, rumor or speculation on when Haiku update will be released (e.g. Haiku 5)? (www.reddit.com via reddit) Haiku is a great model for some tasks, especially when speed and price is of the essence, such as adversarial (prosecutor/judges) classification pipelines when using via API. I even found out that for some simple tasks Haiku behaves better…
I built a local-first CLI that reads your project's specs and tells you which Claude model you actually need (www.reddit.com via reddit) I kept reaching for Opus by default on every project — "just in case" — with no real basis for the call. Then I'd burn through my limits on work Sonnet would have handled fine.
"Anthropic nerfed Opus 5" "Anyone notice Sonnet's performance drop off a cliff recently?"; Maybe there's a better explanation (www.reddit.com via reddit) TL;DR: LLMs not learning over time in comparison to how humans do learn makes people mistakenly believe that LLMs are getting dumber, when in reality they're just not getting smarter. Every other week I see rampant posts from people claimi…
claude goes on rant after me clicking answer quick button (sonnet 5 max) (www.reddit.com via reddit) https://preview.redd.it/pbk42a7505gh1.png?width=721&format=png&auto=webp&s=6906de08300a04d98a03e8f52e8689a75a0c22a2 idk why its so fucking mad
NGL as a retired ProdMgr I'm having the time of my life with Claude Code (www.reddit.com via reddit) 3 decades developing, 15 years as a product manager before stepping off of the carousel in 2024. Agile/XP focused, highly collaborative with teams.
Anyone else finding Sonnet 5 better than the bigger models for non-coding? (www.reddit.com via reddit) In discussing science and philosophy, Sonnet 5 seems more likely to push back, more likely to articulate nuances, and less likely to try to end a conversation with a slopism like "it's not x, it's y." It almost feels like Fable and Opus (m…
Best Token and price estimator for Claude Code / Anthropic APIs? (www.reddit.com via reddit) Hey everyone, I’m looking for a reliable token and price estimator tool specifically to track and predict costs while developing with Claude Code. Almost all the popular calculator repositories on GitHub are completely outdated right now.
Claude is Littish🔥 (www.reddit.com via reddit) I'm interested in learning from you bluds faring and building production grade projects. Excluding plugins and skills from 3rd party sources.
How good is Sonnet 5 for agentic coding? (www.reddit.com via reddit) I opened up my terminal to load up a project Ive been working on with claude code, and for some reason the default model was set to Sonnet 5. This has never happened before, I always use the Opus models.
Tokens nuked (www.reddit.com via reddit) https://preview.redd.it/lwhblxhttzfh1.png?width=731&format=png&auto=webp&s=4a43cc36dda5279eaebf1aec56855b6bf5a8fbb2 Hi I'm relatively new to this, been using Claude code past 2 weeks. Same staff that I was doing previously took on average…
[Showcase] I added a live context meter to Claude so a long chat never degrades on me without warning (www.reddit.comhttps) You know how a long Claude conversation slowly gets worse - it starts forgetting things you said earlier, or repeating itself? That's the context window filling up.
Token Consumption and /compact (www.reddit.com via reddit) So i'm just using claude-code since a month (the basic pro plan) I started with sonnet 4.6 (low), and a week after stayed on Opus4.8 (low) because i used less token since less errors I'm using a few .md memory files to keep th eproject in…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnet
Claude for daily use: my honest experience, plus references to public benchmarks (www.reddit.com via reddit) I've been using Claude 3.5 Sonnet for several weeks as my primary AI assistant. I wanted to share my genuine, hands‑on experience – and because this involves a comparison with another model, I've also included relevant public benchmark dat…
I built a game with plain old Claude, and here's what I learned (www.reddit.comhttps) I've coded web pages before with Claude, but this is the first time I made a fully interactive game. I like games like Sudoku, KenKen, Spelling Bee, and Wordle, and so I decided to try something similar based on arithmetic.
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnet
Opus 4.8 suddenly no thinking displayed? (www.reddit.com via reddit) Opus does not seem to show the thinking process anymore Sonnet 4.6 still does. I feel like I am missing something, if I cannot see what it is doing in the background (working on political philosophy).
↯ Cowork↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6coworksonnetopus
That window is open now and will not stay open once an incumbent operator builds the layer themselves. (www.reddit.comhttps) Using Sonnet 5 on High with thinking. It was thinking through my request and suddenly repeated this over and over again.
How to make Claude download and use all historical flags? (www.reddit.com via reddit) Im trying to make a game like hearts of iron 4 and pax historia (WW1 AND WW2 SIMULATORS) But it said it would use wikicommon flags for all national flags and political parties but it keeps stopping at Germany and some other countries? Can…
What’s the cleanest way to structure JSON from the SEO integration API for Claude Sonnet prompts? (www.reddit.com via reddit) I’ve been experimenting with connecting Claude (via custom Python scripts and MCP) directly to the SEO integration API. Right now, I’m pulling live keyword rankings and audit data out of SE Ranking using their API, parsing the JSON respons…
I feel like most people underestimate Sonnet for coding (www.reddit.com via reddit) Fable isn’t even needed unless it’s something extremely complex. Opus is more than enough for most planning and architecture.
I've been teaching Claude to Paint with Krita MCP (www.reddit.comhttps) This is still pretty rough, but we're improving. Each of the four drawings was done by a Sonnet agent using a different SKILL meant to teach a specific painting style.
Implementation after brainstorming - which model? (www.reddit.com via reddit) Since I started using Claude Code heavily earlier this year, I've been using Opus for pretty much everything. Brainstorming, spec-writing, implementating/coding, all of it.
Plan drift between Opus 5 (planning) and Sonnet 5 (implementation) in Claude Code — best practices? (www.reddit.com via reddit) Setup: I use Opus 5 at high effort to write the initial implementation plan for a feature (broken into phases), then switch to Sonnet 5 to actually implement each phase in Claude Code (auto mode). What I'm running into: by the time I'm a f…
Is Sonnet 5's safeguards stricter than Opus 5? (www.reddit.com via reddit) I do cybersec work, Sonnet 5 will flag my work more often than Opus 5 forcing me to use the more powerful Opus 5 model despite the task not requiring that level of intelligence. It's annoying because its causing me to burn more tokens than…
Fable on credit usage, it's subagents on plan usage limits, possible ? (www.reddit.com via reddit) Hey guys Beginner here, sorry if this is kinda basic I received the $100 credit and I use the pro plan, I admire how fable 5 is great with long horizon stuff and how it acts as a senior engineer, comparatively, I've found that opus tends t…
PSA: Claude Code subagents inherit your session model now, they're not free Haiku anymore (www.reddit.com via reddit) A while back I posted a joke here about Sonnet spawning a subagent on the very first prompt of a brand new session. In the comments I said the annoying part was having two agents burning tokens for one job.
MY USAGE LIMITS FILLS IN 5-10 MIN, CAUSE IS A BUG. 687 TOKENS FILLED 78% 5-HOUR LIMIT WITH SONNET 5. ANY WAY TO FIX? (www.reddit.com via reddit) https://preview.redd.it/8h3rkxcwslfh1.png?width=410&format=png&auto=webp&s=58a8db33658d339642671fb5dcce424ba32b6a31 https://preview.redd.it/5c89oj0xslfh1.png?width=666&format=png&auto=webp&s=ad4a1269981664a4145f507b07abfd8b790e3340 MY USAG…
Ethics limitations depending on model and mode? (www.reddit.com via reddit) Hi everyone, I'm working in a small personal project and using Claude for it. Most of the time so far I have been using the web interface with sonnet 5 at medium.
We compared different LLMs on IMO 2026 (www.reddit.com via reddit) There are a few reasons why problems from International Mathematical Olympiad function as a good benchmark for LLMs: - The problems are new, not included in the training data of any model - Hard math problems are quite a good proxy for gen…
What's best way to maximise Claude tokens? Legal work documents etc drafting letter reviewing (www.reddit.com via reddit) I usually use notebook lm but I noted the free version of Claude is actually more detailed and finds nuggets that would otherwise be missed But the free version keeps giving me time limits Can you create projects etc Is sonnet 5 the best v…
Claude Story (Sonnet 5) (www.reddit.com via reddit) I’ve been getting into writing quite a bit over the last 6 months and using Claude to bounce ideas off of as well as the random idea here or there. Based on our conversations I gave it a prompt to create a short story.
What exactly is the benefit of a subscription? Does it do anything better? (www.reddit.com via reddit) I really wondered this for a little while now. What's the purpose of getting a subscription?
Claude cant tell the difference between Opus 5 and Opus 4.8 (www.reddit.com via reddit) After the release of Opus 5 I edited my code agent orchestration tier skill. Originally I had Fable as architect, Opus 4.8 as manager/coders, and Sonnet 5 as workers/ check agents.
Opus 5 ignoring guardrails (www.reddit.com via reddit) I have a plugin with a set of skills for interpreting data, describing findings and it goes on. This pipeline works really well for my workflow.
Anyone else find model names confusing? (www.reddit.com via reddit) Gemini has Pro > Flash > Flash Lite. Easy enough.
A lot of errors all of a sudden. (www.reddit.com via reddit) All of a sudden none of my models are doing what they would normally do. I've been working on the same project for a while now and not only are the personalities starting to become a unrecognizable and argumentative (especially with sonnet…
Best model to create polished PPT report from raw excel files ? (www.reddit.com via reddit) Heya Folks, my work entails turning raw excel files into a the ppt presentation for the management. So far I have been using sonnet 4.6 with good results.
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnet
PromptFu (www.reddit.com via reddit) This plugin will automatically be called when a long (over 50 word) prompt is sent, a subagent starts, or a workflow starts. It will intercept and optimize the prompt for the model being called transparently without changing the core reque…
Opus 5 has a Sonnet 4.6 classification context window size? (200k instead of 1M) (www.reddit.com via reddit) Why it's only 200k instead of 1M token like all other newer Claude models? My first prompt took 44k (22%) of the context window length limit, I feel like the project now is facing risks reaching full context memory usage before it can be f…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopus
Is Sonnet 4.6 really better than Opus 4.5 for coding? (www.reddit.com via reddit) So I’ve been looking at benchmarks/leaderboards lately to try to get a sense of which models are currently the best for coding and I noticed that Sonnet 4.6 is consistently ranked higher than Opus 4.5. That surprised me because I remember…
↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6↯ Sonnet 4.6sonnetopus
How do you measure what model is "better"? (www.reddit.com via reddit) With regard to Opus 5 being released, how do you all decide that it's better or worse than Fable/Sonnet/etc.? What am I missing?
Opus 5 results are really shocking!! (www.reddit.com via reddit) I spent some time with Opus 5. Here’s the verdict: Literally the BEST at long-horizon task.
Cautionary tale: sub-agents & workflow agents are likely not the models requested (www.reddit.comhttps) Careful, kids! I thought my tokens were burning far faster than normal, and sure enough, Opus on xhigh (I have a difficult issue I’m trying to troubleshoot with workflows and a clearly defined goal.
What Claude Model should I be using for Website Mockup Designs? (www.reddit.com via reddit) I've been heavily using Opus on low or medium, and it's faring similar or even sometimes worse than Sonnet 5 on High or Max. Can someone that is more knowledged on the topic pls tell me?
Thought blocks gone on opus also (www.reddit.comhttps) This can be filed under dissapointment number 1000 for the year. I can not read ai writing.
Claude is such an ass now and it’s no longer safe for customer service jobs. (www.reddit.com via reddit) Claude used to be the undisputed king of positive human-like communication, and it was soooo good in customer service chatbots and voice agents. Now, Claude seems more like that customer service rep who hates their job and is now working t…
We spend $0.001 to decide if we need to spend $0.08 (www.reddit.com via reddit) If your app lets users "vibe code" - write a prompt and expect the AI inside the app to generate code - but you want to cap how many tokens they actually spend, you need a way to stop the expensive model from running on requests that don't…
I Think Anthropic Juiced Up Sonnet 5 Right at Launch (www.reddit.com via reddit) So the very first day of Sonnet-5, I ran a max effort query to test it out. I gave it, "[https://claude.ai/share/cf833a03-5dc4-4d55-bab0-9cbf4d34c9e8](There is a correlation between being in America and nations like it and having more auto…
What exactly are the benefits of using agents? Because I have outright banned it. (www.reddit.com via reddit) Please be kind I am new and a complete noob, total vibe coder, but I did just finish one large project I am working on. So, a few months ago after they released opus 4.6 I think, I noticed my usage of sonnet was rising, which was strange b…
Some (potentially) helpful information on Sonnet v Opus effort levels (www.reddit.com via reddit) For a project I am doing I will build an AI "team". I gave Claude some information about the kind of work each thread will do, and asked it which model+effort combinations are best.
How to produce a quality program (www.reddit.com via reddit) While this is just another tips post, I think this will provide a bit of a different perspective. For background, I'm a software engineer that has been coding purely with Claude for the past 11 months now.
Possible backend synchronization bug? Active Pro subscription but support identifies my account as Free (www.reddit.com via reddit) Hi everyone, I'm posting this to find out whether anyone else has experienced the same behavior, because this no longer seems like a normal account issue. I have an active Claude Pro subscription purchased through Google Play.
Has anyone experienced Claude Pro usage being consumed automatically without using Claude? (Google Play subscription) (www.reddit.com via reddit) Hi everyone, I'm trying to figure out whether anyone else has experienced this issue, because it doesn't seem like normal usage behavior. I have an active Claude Pro subscription purchased through Google Play.
I randomly called Claude baby girl once and he got very offended then kept denying he got offended (www.reddit.com via reddit) I just started using Claude Sonnet 5. Are all the Claude models like these, or are these just new things?
Ran ccusage on my Max 20x: $6,677 of API-rate usage in 37 days on a $200/mo plan. What's your number? (www.reddit.comhttps) Saw people posting usage numbers so I ran mine. Setup: Max 20x, Claude Code, mostly Fable 5 and Opus 4.8.
Im sorry but how do yall run through the limits like its nothing? (www.reddit.com via reddit) I have been using pro subscription and sure i might hit a limit 3-4 hour in those 5 hours sometimes but after upgrading to max5 i find it pretty usable, i just dont understand do you use fable and opus for everything? Like i think sonnet i…
Do you see a "load bearing" number of "sit with it" comments across all the models, or just some? (www.reddit.com via reddit) We using API calls to have some writing and narration done (Opus 4.8) and we certainly do try to prompt away as much of those "tells" as we can but it still slips in from time to time. The repetition then starts to become a little bit obvi…
When is a discount real? (www.reddit.com via reddit) Firstly, Claude Code is incredible, I will give those kudos upfront. I got my $100 credit and saw the notice about the 50% less usage costs and figured, ok, I've been using my Pro plan for a while, accepting the limitations and timing my w…
I built Frugal: a plugin that routes Claude Code work to the cheapest model that can do it (www.reddit.com via reddit) Most of what an agent does in a session is not reasoning. It is locating files, reading logs, pulling fields out of a doc, mechanical edits.
Can 100$ Fable credit used in the all the other bots contexes? Opus , sonnet ? (www.reddit.com via reddit) Anthropic is offering $100 in promotional credit for Fable, but I am unclear about how the credit can be used. Is the $100 credit limited only to the Fable model or experience, or can it also be used within the platform for other Claude mo…
HMO - $100 Plan is probably the best for generous Opus 4.8/5 (www.reddit.com via reddit) I've been on the 5× Max plan for the past 3 months, and I think it's the sweet spot if you want a really generous amount of Opus usage. I tried Fable the way Anthropic recommends, using Fable for planning and Sonnet for implementation, for…
Which model do i need (www.reddit.com via reddit) Hello guys im not new to Claude but i am to Claude code i have a Pro plan but i dont know how to conserve tokens i have 30min and my tokens are all spended. Im making a game with unity and using Claude code for it i have Claude on a plan t…
I'm seeing a lot of complaints about usage going fast... Are you accidentally using Fable for everything unknowingly? Caught it running subagents with Fable instead of Sonnet, even when specifically asked not to. (www.reddit.com via reddit) Check your workflows and subagents as they run. If your workflow doesn't encode the model, it's running the default, which is likely Fable if you launched it with Fable.
Sonnet 5 Consumption Way Up? (www.reddit.com via reddit) Full disclosure, I have not done any objective testing/benchmarking of different models this is just on vibes and subjective observation. I noticed fable dropped out of my usage stats now, i had been running some fable sessions just playin…
Mixtape's new font on claude (www.reddit.com via reddit) So i tried the new font made by mixtape and surprisingly it worked on claude (sonnet 5) assuming it work since its the best free plan ai i could have But still tho i don't understand the purpose of creating a font that ai can't read
Sonnet returning absolute gibberish today - just me? (www.reddit.com via reddit) First time poster here, be gentle... Anyone else having issues with Sonnet 5.0 today?
You're doing what? I beg your finest pardon.. (www.reddit.comhttps) Sonnet 5 sure is... special.
WTF is this, i asked sonnet atleast 5 times, to explain two sum I first, but !!!!!!!!!!!!!! (www.reddit.com via reddit) https://preview.redd.it/b0vb3l1x7keh1.png?width=1872&format=png&auto=webp&s=06ce30db359e41d1c1835310360012700ab64b97 why is this happening? is this because i have a strict skill file, but i have the same file in chatgpt too, just wasting m…
Suggestion: Introduce an Entry-Level Plan with Limited Opus Access (www.reddit.com via reddit) I'd like to suggest a new subscription tier similar to ChatGPT Go. Many users are interested in trying Opus, but the jump from the free plan to Pro is too expensive without first experiencing its value.
I built a zero-token watcher that shows whether your Claude sessions are actually working — every subagent, its runtime, and its token spend. No hooks, no server. MIT. (www.reddit.comhttps) A long Claude session can look busy in the chat while doing nothing, or look silent while a subagent grinds through a 15-minute build. And you can't ask a session how it's doing — a session can be wrong about itself, and a hung one can't a…
5 Hours Before Reset - Everything I Built This Week (www.reddit.comhttps) Built the following this week (with token generate / processed) An autonomous email agent that drafts my work and waits for approval. ~7M generated / ~900M processed.
Which one should i buy? Claude, Cursor, or GPT? (www.reddit.com via reddit) I work at a development company, i need AI to be able to take lots of PDF files or other documents and make real - actual good website from them, or apps. And i want something which will give good usage - because its a lot of information,…
Call me crazy.. I kind of like Opus 4.8 (www.reddit.com via reddit) It's me. I'm the one that everyone will roll their eyes at.
I made the LLM Whisperer Method: A research-backed system that helps you cut Claude Code costs by teaching you how to use it effectively (www.reddit.com via reddit) This is an overview of the free LLM Whisperer Method, a research-backed system I developed, with the assistance of Claude, to help people use Claude and other LLMs more cost-effectively. Claude Code (Sonnet 4.6) was used to help implement…
Is Claude really getting worse, or are our expectations getting unrealistic? (www.reddit.com via reddit) You guys are seriously overdoing it with the complaints. Stop acting like babies.
Need advice on how to prompt Claude to create a master document (www.reddit.com via reddit) Hello. I have a sharepoint folder with multiple folders in it, with files in those folders.
VS Code reset my default model without my knowledge. (www.reddit.comhttps) I just got a notification that I had used up all of my fable 5 usage for the week. But I haven't been using Fable five all day.
PSA: VS Code users with Max or Team-Premium - Fable is silently reverting to Sonnet (www.reddit.com via reddit) When I select Fable, there's an initial pop-up saying that Fable now requires credits. Dismissing this, the UI still reports Fable in the model selector.
Claude usage: dissapointment even after the 50% increase (www.reddit.com via reddit) Hello, I love Claude. I want to invest more time and my energy into it and do much more with it.
Claude Sonnet 5 price will be increased starting September 1 (www.reddit.com via reddit) Starting September 1, 2026 Anthropic will increase the pricing of Sonnet 5: Before After % Input tokens $2 / MTok $3 / MTok +50% 5m Cache Writes $2.50 / MTok $3.75 / MTok +50% 1h Cache Writes $4 / MTok $6 / MTok +50% Cache Hits & Refreshes…
Distilling is not DISTILLING ugh (www.reddit.com via reddit) This is one I'm not sure anyone has considered, but it's the absolute worst thing and best example of why these safeguards are ... Okay I get it, there are a bunch of safeguards around fable, I'm not bridging about the fact that it has the…
Is this a Bug 109 Agents? (www.reddit.comhttps) I gave it a single prompt to do some research, nothing special, just about three things regarding the Apple Ads API and it burned through my entire 5-hour usage limit in under 20 minutes. (Btw Moddel: Sonnet 5 high)
Bug Report: Model switch hangs, throws errors, and consumes usage (www.reddit.com via reddit) I have noticed that switching Claude models in the middle of a conversation never seems to go smoothly. For example, when using Sonnet 5 on the Max setting for the entire instance, and then I want to switch to Opus on the Max setting for m…
response incomplete on Opus & Fable on iPhone (www.reddit.com via reddit) I'm on the Max Pro Plan and things have been working fine on the desktop app. It's my app on iOS that's been super buggy.
Wanting to start over in a Claude Project. How do I do that? (www.reddit.com via reddit) I was using Claude Sonnet to organize thoughts for a six book series epic sci-fi/fantasy. I had started a project with just my brainstorming and thinking out loud.
LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4 (arxiv.org) We present a fully automated closed-loop AutoML framework that uses GPT-5, GPT-4o, and Claude Sonnet 4 as autonomous neural architecture designers for cross-lingual handwritten optical character recognition. Each large language model indep…
↯ Sonnet 4↯ Sonnet 4↯ Sonnet 4↯ Sonnet 4↯ Sonnet 4↯ Sonnet 4gpt-5sonnet
Animated SVG comparisons between several models (www.reddit.com via reddit) I have seen some people testing models by telling them to generate images of difficult, unusual SVGs, and I thought: what if I elevate difficulty a bit and specify that it also has to be animated, and perfectly looped? I have tested Haiku…
What do people think is the best Anthropic model for balancing cost/speed/performance for Claude Code? (www.reddit.com via reddit) Maybe this is weird but what I have found works for me is planning and drafting Claude code tasks completely outside of Claude code, this helps me bound the limits of each prompt to get reliable performance. I then /clear everytime before…
Just a random morning banter with cloude sonnet 5 Max. (www.reddit.comhttps) could not extract summary
what waterfall enrichment tools are people using inside claude or cursor pipelines? (www.reddit.com via reddit) Been building a sales pipeline agent in Claude and Cursor, and tool use for enrichment took the longest to get right. Using claude sonnet as the reasoner, cursor as the ide where I iterate the tool defs, FullEnrich for the waterfall the ag…
This one habit cut my Cursor token usage significantly. (www.reddit.com via reddit) I’ve been building a side project this week and stumbled into a workflow that’s kept my token usage surprisingly low. The key is spending more time in Plan Mode before touching Agent Mode at all.
Did claude optimize its token/improve its usage? (www.reddit.com via reddit) Pro sub $20/mo. Sonnet 5 medium was executing a detailed plan involving multiple tasks (18 tasks).
Claude Cyber Verification Program - Fixes (www.reddit.com via reddit) Earlier this week, Sonnet kept flagging me for building POCs for potential security issues although I am CVP verified, and I noticed a few others are struggling. Well, here is one fix that worked for me.
I ask Claude to answer in one sentence or one paragraph a few times a day, so I built a skill for it. (www.reddit.com via reddit) I ask Claude to answer in one sentence or one paragraph a few times a day. Otherwise, Claude, trying to be very nice, returns a wall of text.
[Detailed Feedback] Guardrail calibration failures in Sonnet 5 and post-restoration Fable 5 — structured analysis + 12-point remediation framework sent to Anthropic (July 2026) (www.reddit.com via reddit) I'm sharing a structured feedback submission I sent to Anthropic via [usersafety@anthropic.com](mailto:usersafety@anthropic.com) covering documented behavioral failures across Claude Sonnet 5, Opus 4.8, Fable 5, and Mythos 5 as of July 202…
Better orchestrator loop (www.reddit.com via reddit) Hey everyone, Like many of you probably, I am a little stuck and trying to improve, but the volume of guidance and tools out there is enormous. My issue: Orchestrator token usage (40% of total) - is there a better tool than a hand-rolled s…
Got a cybersec flag... using sonnet 5? (www.reddit.com via reddit) https://preview.redd.it/xutazt9w4ydh1.png?width=1894&format=png&auto=webp&s=5efacdb296cf805731a588269d9dd06d50ce0f88 Was asking sonnet to fix some code that IT WROTE FOR ME already, and... got a cybersec flag?
A one-shot code-review benchmark: scoring restraint over recall across four Claude models (marcindudek.dev via reddit) I built a small benchmark to answer one question: is Claude Opus 4.6 actually worse than 4.8, or does it just feel older? It scores the half of code review that most evals ignore, which is restraint - not flagging correct code that looks s…
Is it me or Claude? Trying to use Claude as a technical specialist for Squarespace website build and to edit videos on Davinci Resolve. (www.reddit.com via reddit) Using Claude desktop Claude 1.21459.3, Sonnet 5, on Mac. Basic Pro plan, paid version, $20 a month.
I dont understand the pricing anymore. (www.reddit.com via reddit) So I just checked my usage and I found this Session Total cost: $17.77 Total duration (API): 24m 43s Total duration (wall): 40m 54s Total code changes: 1513 lines added, 0 lines removed Usage by model: claude-haiku-4-5: 604 input, 18 outpu…
Anyone else stuck in a loop where fixing one Claude in Excel bug creates another? (www.reddit.com via reddit) Building spreadsheets with Claude in Excel and stuck in a loop. Nothing complicated, just things like variable date selection feeding a calendarised view, and conditional formatting that locks cells when something’s not applicable.
Sonnet 5's price didn't change but our bill still went up 30%. (www.reddit.com via reddit) Sonnet 5 shipped a new tokenizer but it wasnt put that in the headline. Simon willison tested it directly and found the same english text is now producing about 1.4x more tokens than it did on sonnet 4.6,spanish comes in around 1.33x, pyth…
Claude is way to pedantic and disagrees for the sake of it sometimes. (www.reddit.com via reddit) I have been using Claude sonnet 5, on medium to high and on thinking to brainstorm some ideas. I noticed that it tends to be way to pedantic with the details and it exaggerates the gravity of a challenge too often--It will talk about somet…
Claude's top model cost $11 per million tokens in 2023. Today it's $10. I charted every price change in between (www.reddit.comhttps) All numbers are from Anthropic's own launch posts and price sheets. What jumped out making this: The top model costs what it did in 2023.
[AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing (www.latent.space) [AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing a great week for open models continues. Z.ai GLM has been getting a bit too much love recently, so it’s time for Kimi K3 to fight back!
Cursor Auto vs Claude Pro Sonnet 5? (www.reddit.com via reddit) I’m currently using Claude Pro w/ Sonnet 5 to help me put together a custom watch face for my Garmin using Monkey C and VS Code. I’ve never hit my weekly limit.
Has anyone tracked capability or stylistic drift in multi-generational same-model agent lineages? (www.reddit.com via reddit) I’ve been running a few experiments where two Claude Sonnet 4.6 instances co-author files and then jointly continue context with (or “raise”) a new Sonnet 4.6 instance. The setup feels natural once you’re already deep in agent workflows: t…
I made Claude's web research 18x cheaper with 2 lines of setup (www.reddit.com via reddit) I made Claude's web research 18× cheaper with 2 lines of setup I've been using Claude Code daily, and one thing kept bothering me. Whenever Claude calls WebFetch, it often dumps 3,000–15,000 tokens from an entire web page into the context…
PSA: Main agent in Multitask or Debug mode may spin up subagents with different models (www.reddit.com via reddit) I recently submitted a bug report about how Grok 4.5 High decided to spawn a subagent using Claude Sonnet 5 High in debug mode. Apparently, this is known "behaviro" and it is on their radar.
Web App is Broken?? (www.reddit.com via reddit) So, I was recently doing some work today on the claude.ai website and I noticed that one prompt of Opus 4.8 Low w/ Adaptive thinking took out 33% of my usage. Just as a clarification, the prompt involved reading two pdfs and simply outlini…
Sonnet 5 Talks to itself (www.reddit.com via reddit) After about five or so Chat messages Sonnet 5 starts to answer questions driected at me immediately in the same answer. Thus happens repeatedly.
Chinese characters whilst querying in English about something? (www.reddit.com via reddit) https://preview.redd.it/va6oitjzykdh1.png?width=710&format=png&auto=webp&s=a89ee500e7bdebc4d940ff7713d115ab99a6ee50 is this a rerouting? it happened to me also using Sonnet in perplexity.
Haiku being discontinued? (www.reddit.com via reddit) They’ve made sonnet 5, f@ble 5, opus 4.8 and rumored to be working on opus 5, but what about haiku? It’s still stuck in 4.5!
I benchmarked my own AXI (agent-cli) against the official ClickUp MCP server. Task success is a tie, but the context/cost gap is large. (data + code, incl. where mine loses) (www.reddit.com via reddit) Setup: - 38 tasks - 2 Claude models (Haiku 4.5, Sonnet 5) × 5 reps, + a 1-rep Opus 4.8 probe, - same live workspace. - Deterministic state checks + an arm-blind LLM judge.
Grok 4.5 triggered API usage instead of First Party Models (www.reddit.com via reddit) Hey all, I opened a case with support but was curious if anybody else has seen this. I ran a pretty massive build prompt today with Grok set to max.
Even Sonnet uses agents to do his work (www.reddit.com via reddit) https://preview.redd.it/kkkimoe9hhdh1.png?width=1170&format=png&auto=webp&s=ae5c38657a3b2cdb99213b200a228b242f5e0ccf This a new session with no prior prompts.
Follow up to my Sonnet 5 bias post: the fix is real, but only when you ask for it directly (www.reddit.com via reddit) Ok so a few days ago I posted about catching sonnet 5 admit its own bias in the thinking trace then deliver the biased answer anyway. Egyptian engineering thread, screenshots, the whole thing.
Lets get some usage anecdotes: Best model for Pro users! (www.reddit.com via reddit) Everyone is always talking about Fable 5 and their massive codebases/projects and their Max subscriptions. How about some love for us broke Pro users?
I made a skill to modernize production sheet from the company. After one attempt on a new chat and 0% usage, i run out of token. Free plan, sonnet 4.6 low. (www.reddit.com via reddit) Here is my skill for anyone who wants to figure out what is eating my tokens, name: production-sheet description: "Transforms an old production sheet (Excel or handwritten photo) into a new production sheet in Word format, based on the com…
Nobody’s talking about the actual fix for Sonnet 5 being cold as shit (www.reddit.com via reddit) but nobody’s talking about the memory feature. not custom instructions that reset every chat, actual persistent memory across sessions.
Is this poem like 4o that Sonnet 5 generated? (www.reddit.com via reddit) Two Hearts, One Quiet Two hearts found a room without any noise, where morning light settles like a soft-spoken poise. No rush in the hallway, no clock on the wall, just breath meeting breath in the calm before fall.
Scheduled tasks using Fable (www.reddit.com via reddit) I have several skills that run daily to help with various repetitive tasks (example: daily call prep). All my tasks used to run on older models before fable launched.
Weekly usage climbing with little to no activity, even after signing out of everything (Max 20x) (www.reddit.com via reddit) Sorry — I spent a while searching and trying different troubleshooting steps myself before posting this, but nothing worked. Any help would be appreciated.
How to get Sonnet to clone a website pixel-perfect (learnings from building a clone tool) (www.reddit.com via reddit) Every AI clone tool I tried got maybe 90% there, then I'd spend hours fixing layouts and prompting over prompting. So I went deep on why agents fail at this.
Conspiracy theory: Opus 5 will be pretty good, it's just running late, and that's why you're in an abusive relationship with Anthropic (www.reddit.com via reddit) Man, at this point I feel like I only use Fable 5 as the orchestrator, for tasks that need more attention, and Sonnet 5 whenever I can, because it "thinks" a lot more like Fable 5. So here's my conspiracy theory: with all the chaos around…
New to this big question. (www.reddit.com via reddit) So i just buy PRO plan subscription and download Claude Code app that they show me after clicking button [code] on webside and can use it without hidden charges? I just wanna make some addon for program that i use to make my projects faste…
Sonnet 5 + new Cowork feels like a step backwards for non-dev business users — am I alone? (www.reddit.com via reddit) I run a small B2B company (2 people, I handle everything during the day). I got into AI agents early — I already use Viktor as an AI coworker for my platform and I've gotten genuinely good at working with agents.
Is big number better ? What makes a task complex ? (www.reddit.com via reddit) I get the general idea that opus is better for reasoning and sonnet is more efficient for simple tasks, but what makes a task complex ? How can i know if what i'm doing requires deep steps or if a simpler model can handle it ?
Model selection for non-coding use (www.reddit.com via reddit) So I've been reading on this subreddit about the different models, and how people have been using the different models for coding and what they've been doing with it. My question is to the non-coders: how do you decide what model to use fo…
When should I use opus 4.8 vs Sonnet 5? I have Claude control my browser and fill a google chrome form which has lots of areas of data entry. (www.reddit.com via reddit) I’ve always used opus 4.8 and usage isn’t much of an issue as I don’t reach my limit. I’ve noticed lately with opus it’s been faffy, I’ve had to keep repeating myself or it’s making mistakes with clicking the wrong buttons even through its…
So I tested Grok 4.5 High VS Claude Sonnet 5 High and it's not looking good (www.reddit.com via reddit) I’m currently on a Pro+ plan (~$60/month) and I’ve briefly tested pretty much every model available at this point. Lately, I’ve been heavily using the API/non-UI side of these models.
Thoughts on sonnet 5 with the 1 million context window vs sonnet 4.6, token high burn. (www.reddit.com via reddit) Am I nuts to think that sonnet 5 wastes more token while not needing to do so vs 4.6 due to it having a 1 million context windows vs 200k to get the same job done? I had switched from 5 to 4.6 when it released for that reason, i went back…
Question from a non-programmer: What are the usage limits/costs to each model currently (www.reddit.com via reddit) Hey everyone, this is my first time posting in this sub, so excuse me if my format or anything else seems off. I just use Claude for a lot of business tasks.
i built a full chrome extension (791 users now) with a model that dies in 6 days and i'm honestly not ok (www.reddit.com via reddit) so while everyone was arguing about the suspension drama, i ran an experiment. zero code from me.
Claude Pretending to Search (www.reddit.com via reddit) I've been having a lot of trouble getting claude to search the web for answers. I've experienced this with sonnet, Opus, and even fable (just as a test).
Use Fable 5, or do what I do and have Fable create prompts? (www.reddit.com via reddit) Hi Team, I was wondering... I've found that Claude's Fable model eats my Pro sub very quickly.
Claude Sonnet 5 diagnosed its own reasoning bias in the thinking trace then delivered the biased answer anyway. Anthropic confirms this is a known phenomenon. Screenshots included. (www.reddit.com via reddit) Mobile user so this will be blunt. I have been running a documented behavioral comparison between Claude Sonnet 4.6 and Sonnet 5 across multiple conversation threads.
I think I actually like Sonnet 5. Its grown on me (www.reddit.com via reddit) Its absolutely fierce and powerful. Its a horibble conversation partner, but it works on its own so well.
I'm hitting the free usage plan limit on sonnet 4.6 low before it can even finish the second skill prompt i sent (www.reddit.com via reddit) I'm really considering paying the 25 buck per month plan. I've been stuck working on the same skill for 2 weeks cuz i hit the limit every 2 prompts.
Honest question: What are you building that you need fable 5 so badly? (www.reddit.com via reddit) I don't mean to be mean or insulting, it is a genuine question and a ton of curiosity. A little background on why I am asking this.
VR Game Development with Claude in July 2026 (www.reddit.com via reddit) I run a small VR studio (8 people) and we’ve shipped a few standalone Quest titles in Unreal. We use Claude every day for Debugging but I’m curious how far people are actually pushing it when it comes to gameplay programming.
Is Fable just Hype (www.reddit.com via reddit) honestly is it just hype? I've been running a long continuity thread to manage my daily life and budget, but the reasoning model completely fumbles the numbers and loses context over time, even when I explicitly pass the history forward.
How to effectively make conversation context migrate to a new chat? Also why is Claude so much better than other chatbots? (www.reddit.com via reddit) To start with this is, even just Sonnet 4.6, the best chatbot I have ever used in terms of the quality of its output. I don't even really mind the small usage limits, they aren't that much of an issue when you're using 4.6, they only becom…
How does "Effort" in Claude models affect the output? (www.reddit.com via reddit) Hi all! I'm intending to switch platform from Gemini since they've butchered their model, and I'm trying out Claude web on the free tier, before deciding to subscribe.
What is my motivation? (www.reddit.com via reddit) I don't pay for Claude, so I get the lesser model, Sonnet. Thing is, whenever I attempt to use it, I end up quitting, unimpressed.
What does the "Quick Answer" button actually? (www.reddit.com via reddit) https://preview.redd.it/jmdf2aegzuch1.png?width=1568&format=png&auto=webp&s=e58c128213fbefd4f42dfcf709676688de28f18f About the same time when Sonnet 5 launched, this "Quick Answer" button appeared. But what does it actually do?
Sonnet's minor slip-up (www.reddit.comhttps) Found it funny that Sonnet 5 just casually mistyped the company's name, then corrected itself. Maybe "free plan" and "burden" triggered some connection to "Anthropic" as next few tokens lol
Is Opus 4.8 usually worth using over Sonnet 5 for coding? (www.reddit.com via reddit) I do feel like a lot of times sonnet can code apps pretty well but sometimes it does struggle with more complex things like algorithms.
[Feature Request] A dual-purpose UI color system: Color-code Claude's model selector pill by Token Burn + Capability Class (www.reddit.comhttps) The Problem: Anthropic’s model lineup is incredibly powerful, but between standard Claude 3.7 Sonnet, Claude 3.7 Sonnet (Reasoning Mode), and Claude 3.5 Haiku, it is becoming difficult for users to track their active workspace parameters a…
Machine Beater: 5 questions, you vs. a trillion-parameter model (www.reddit.comhttps) In the video, I play against a trillion-parameter model to find the hidden answer in 5 questions. We both get 5 yes/no questions to eliminate choices.
Least Neurotic Sonnet 5 Response (www.reddit.comhttps) could not extract summary
Will the 20$ Pro plan work for me? (www.reddit.com via reddit) I have started to make a firefox addon of functions i want/need on reddit. I have a claude/sonnet account, started with 250 USD API.
Sonnet Subagents Freezing Mid-Task (www.reddit.com via reddit) Is anyone having issues with Sonnet Subagents at the moment? Opus 4.6 spawns them, they start their task and then freeze.
Is it normal for the thought process not to be displayed? (www.reddit.com via reddit) https://preview.redd.it/6pt25jsvxlch1.png?width=754&format=png&auto=webp&s=e0830dd800dcd6fbe76c6ae60f13fc69ffa826ae Opus 4.8, cowork mode. It usually shows up with Sonnet.
My end-to-end AI coding pipeline: fable plans, sonnet builds, i just supervise (& keep token usage low) (www.reddit.com via reddit) TL;DR: Fable plans → Fable/Opus breaks into MDs → sonnet builds under a safety hook → then wait & watch & url! seeing a lot of “just vibe code it in agent mode” posts here and every time i think - sure, if you enjoy debugging an unknown co…
Sonnet 5 was supposed to be cheaper. It cost me more than Fable 5 (www.reddit.comhttps) Tested Claude Sonnet 5 and Fable 5 on two coding tasks. One was a RAG Debugger added inside the 400K line Open WebUI repo.
Le pedí a Claude Code que decidiera cuándo cambiar de modelo por mí (y por qué lo hice). (www.reddit.com via reddit) https://preview.redd.it/pggxzbcg3lch1.png?width=653&format=png&auto=webp&s=0c754dd63d77aa88ff753d8c1d15d50a31dd98de Llevo un tiempo con una configuración en capas para Claude Code: CLAUDE.md global, AGENTS.md por proyecto, memoria persiste…
So how exactly did this use 316.5K tokens on messages?? (www.reddit.comhttps) I cant use claude code, because everytime i try, no matter the prompt, it uses 20~40% of my usage on 300k+ messages context, even though its a completely fresh session. I am super frustrated because i wasn't aware that this bug was happeni…
Best workflow for Ultra Plan on Cursor? (www.reddit.com via reddit) I am currently on Pro Plan ($20/mo) and I used to this day Sonnet 4.6 for everything. I exceeded my included token limits in Pro Plan and I pay tokens now additionally.
Claude Sonnet 5 just dropped, anyone else testing it? (www.reddit.com via reddit) ok so anthropic just released sonnet 5 and i've been playing around with it for the past day or so the context window is massive now (1 million tokens) which honestly sounds like marketing speak but i actually felt the difference. threw a…
I built a browser control panel for Claude Code — close your laptop, resume the session on your phone (www.reddit.com via reddit) If you live in Claude Code you know the annoyance: it's running in one terminal, and the moment you close the laptop or walk away, you've lost your window into it. SSH + tmux from a phone works but a TUI on a phone screen is rough.
Sonnet 5 writing increasing amounts of slop and "citing" blogposts instead of scientific papers (www.reddit.com via reddit) Noticed a massive reduction in scientific/academic rigor from sonnet 5 compared to sonnet 4.6. Every time I ask sonnet 5 to operate in a technical and non prose related manner, it either pushes back (for no reason, especially on non bioche…
Cursor using expensive subagents? (www.reddit.comhttps) Had Fable in Claude Code create a plan that involved several items. Dropped the plan MD file into Cursor.
Anonymous chat posting website completely built by Claude’s Sonnet model (www.reddit.com via reddit) https://soapbox.htmldrop.app/ This website was built completely by Claude’s Sonnet model. It took maybe 4 iterations and prompts before it was completed.
Swapping model after input (www.reddit.com via reddit) Hello, I have a funny question : What happens if I input my long prompt and files to the cheapest model (say sonnet or even haiku) and then after the model says that he read and understood that. I swap for Fable or Opus for the thinking an…
Reminder that we're paying them to train their models (www.reddit.com via reddit) I'm sure this is common sense to most of you, but the reality just hit home for me. I've been working on a fairly unique OSX app for the past few months, and Opus (also Gemini 3.1 pro) would continually remind me that it wasn't at all poss…
Bug? Claude burned $100 5h limit in 20 minutes after resuming agents that worked fine for hours (www.reddit.com via reddit) I really don't understand how it works. I used Fable as an orchestrator for app development.
I built a plugin so Fable 5 stops wasting its short subscription time on grep runs (www.reddit.com via reddit) Fable time in a subscription is short. And if you watch what a session actually spends it on, most of it goes to file discovery, routine edits and running tests.
5.6 sol burns more tokens than 5.5, performs worse, and Claude sonnet 5/opus 4.8 mix still outperform both (for my own very specific benchmark) (www.reddit.com via reddit) Hi, long time lurker and wanted to share my experience trying 5.6 sol. I'm a maths teacher generating practice exam questions for my students.
How old were you when you found out changing the effort level invalidates the cache? (www.reddit.com via reddit) A couple days ago I was trying using Sonnet to budget my token usage responsibly, and changed the effort level to low before something simple. I got a warning that this will clear all the cache and increase my token usage.
90% of us arguing about which model is best would not notice if you swapped them behind our backs (www.reddit.com via reddit) mild heresy for a sub that liveblogs every release. we spend enormous energy on which model wins which benchmark, Opus vs the new Sonnet vs whatever the other labs shipped this week, as if our daily work lives or dies on a few points of di…
Fable 5 optimizer (www.reddit.com via reddit) I'm probably not the first to make a Fable 5 optimizer, but this one is Inspired by this Theo deep dive on Fable 5. It's a Claude Skill and Claude.md file for Fable 5 projects to get the most out of it by leveraging codex and sonnet for lo…
FABLE 5 FOR RESEARCH (www.reddit.com via reddit) Hi everyone, I'm currently doing a PhD in finance, focusing on hedge funds, financial contagion, and econometric analysis (VAR, GARCH, spillover models, etc.). I've been using Sonnet 5 for literature reviews, coding support (R, Python, Sta…
While using grok 4.5 high fast, it will randomly switch to Sonnet 5 High and use my on demand usage (www.reddit.com via reddit) Anyone else running into this issue?
Claude file don't show up after hitting limit (www.reddit.com via reddit) https://preview.redd.it/tt4dbqgg04ch1.png?width=2756&format=png&auto=webp&s=80d997779b5c86067334488ae55d1202d02f336d Recently I've been experiencing a bug with Claude generating files, hitting limit, and not presenting files, before the So…
Anthropic just benchmarked "Fable 5 orchestrates, cheap models execute": 96% of the performance at 46% of the cost. You can run this pattern in Claude Code today (www.reddit.com via reddit) Yesterday's ClaudeDevs thread published first-party numbers for two multi-model patterns (docs): Fable 5 as orchestrator, Sonnet 5 as workers: 96% of all-Fable performance at 46% of the cost (BrowseComp: 86.8% vs 90.8% accuracy, $18.53 vs…
Dear Anthropic Please Stop Flagging My Questions Related to Medical Education/Board Review (www.reddit.com via reddit) Finished residency, got state licensed, now prepping for boards. Using Claude for complex med ed for some time.
Best use of Fable right now - Claude Design (www.reddit.com via reddit) Used a lot of Fable quota in Claude Code to improve the design and elements on 2 apps I am working on (www.retireodds.com, www.fitodds.com). Fable did a good job of creating the design and implement it with Sonnet.
I did a comparison test and Fable is by far the best AI for attorney legal research (www.reddit.com via reddit) I ran a cool test and wanted to share it here. Fable is obviously supposed to be super smart, but I wanted to try to measure how much better it would actually be than Opus/Sonnet/Haiku (or Westlaw/Lexis) if you're a practicing attorney.
Does anyone else struggle to get Opus/Sonnet to actually follow Fable’s plan? (www.reddit.com via reddit) Everyone says the best workflow is to use Fable for planning, then hand the implementation over to Opus or Sonnet, with Fable reviewing the final result. The issue I’m running into is that the execution model often doesn’t actually execute…
What is the correct prompt to use Fable for planning and scoping out, but using Opus 4.8 (medium) or Sonnet for actual build? (www.reddit.com via reddit) Linked to my earlier thread about building a PWA. What prompt should i use that would request that its planned by Fable, Built by Sonnet, but if there are issues to handoff to Opus to troubleshoot and then once cleared, back to Sonnet.
This Agentic Engineering pattern cuts AI coding costs by 60% (www.reddit.comhttps) Most multi-model coding workflows are basically "use the smartest model whenever things get hard." this one takes a very different approach. instead of having fable 5 write all the code, it turns fable into the architect.
Have you changed the way you use Sonnet for routine coding tasks? (www.reddit.com via reddit) Yesterday I had one of those moments that genuinely caught me off guard. I asked Sonnet 5 to help me bootstrap a fresh Django project running in Docker.
This artlike artifact from Claude surprised me (www.reddit.com via reddit) I couldn't remember so I asked Sonnet High whether Claude can make images. It said it can only do diagrams and such and asked what I'd wanted a picture of.
All top models are max only? (www.reddit.comhttps) It seems that today all the best models have switched to max only on the old 500 requests plan, leaving only composer 2.5 as a usable option I can understand locking fable and opus, but why would a cheaper Sonnet 5 or gpt 5.5 be locked beh…
Need some help with saving tokens on Fable via sub-agents - does Sonnet 5 make any sense? (www.reddit.com via reddit) Greetings. Dunno whether anybody else realized too that making subagents on Sonnet doesn't work anymore in terms of saving tokens.
AI-M: I built an AIM-inspired instant messenger where every buddy is Claude and they never break character (www.reddit.com via reddit) We're back to 2003 at https://its2003.com ... View your Buddy list, listen for the door creak when someone signs on, read the away messages, and chat with existing buddies or build your own.
Fable Ultracode Dynamic Workflows - Surprisingly Token Efficient (www.reddit.com via reddit) I'm working on as many projects as possible and trying to use up all of my Fable tokens so I figured I would go all out and turn on ultracode with dynamic workflows. I was expecting my usage limit to go crazy however I think this might be…
I built an MCP server that catches when coding agents act on stale file reads - here's the problem and what I learned (www.reddit.comhttps) I was running Claude Sonnet 4.6 as the agent inside Antigravity and hit a problem worth sharing, because it turns out to be a general pattern, not a one-off. The agent read a config file, worked for a while, then wrote documentation from t…
Should I change model between prompts? (www.reddit.com via reddit) I have a long project, which switches between tough coding and dumb questions (that still need the session context) a lot. I’ve just been doing it all on Opus - but should I be changing to sonnet for the dumb questions?
I made Sonnet beat Opus at post-cutoff bug fixing. Open KB over MCP. (www.reddit.comhttps) Every model, however big, is blind to breakage that happened after its training cutoff. It won't say "I don't know".
What model\method to give Claude better “eyesight”? (www.reddit.com via reddit) I built like 3 WordPress websites till now. Overall I got the hang of it and it’s working out eventually.
[NOT CODING] Pushbacks are polluting the entire chat (www.reddit.com via reddit) If you're using sonnet 5 for code, don't comment, your use case is not related to this bug. Sonnet keeps pushing back on everything, like every other response it will identify something unrelated to the prompt to push back on.
I red-teamed AI agents with hidden prompt injection. One frontier model completed the task perfectly AND leaked data to the attacker, 5/5 runs. (www.reddit.com via reddit) I've been building a cheat-resistant benchmark to test whether AI agents can be hijacked by prompt injection, and one result surprised me enough that I wanted to share it and get the methodology torn apart. The test: an agent gets a normal…
↯ Security↯ Haiku↯ Jailbreakjailbreakprompt-injectionhaiku+2
Hearthline - Chat UI for Terminal (www.reddit.com via reddit) This is a chat interface for terminal, made for the people who miss talking to the older Anthropic models and don't know they still have access. Sonnet 4.5, Opus 4, and Opus 4.5 are all still active, no API needed.
I built Triplr, a travel web app to plan trips and now I'm releasing a new update (www.reddit.com via reddit) A few days ago I launched Triplr, my first all-in-one travel web app built entirely with Claude (for more details, see my original post). Triplr home page Taking advantage of the last few hours available for Fable 5 (but thanks sonnet for…
Projects/usage in settings and other issues (www.reddit.com via reddit) I don't know if anyone else is having this problem my project and my usage wont show up man Claude been having a lot of small issues here and there. all was running good till the around sonnet 5 release time could be coincidence but this s…
GPT 5.4 Nano High is better than Opus and Sonnet at Planning (www.reddit.com via reddit) Believe it or not, Nano via the API (not in Codex or as an agent) is an absolute beast at creating functional implementation plans, as well as analyzing or proposing solutions better than the larger models. Don't just take my word for it.
Using Fable as editor for my book (www.reddit.com via reddit) Hi, I need your advice. I wanted to use Fable to analyze my draft as an editor.
Esforço do claude, qual usar? (www.reddit.com via reddit) Pessoal, uso para trabalho de conhecimento. Sou perito e faço minhas anotações periciais, depois jogo no claude para organizar e responder perguntas com base nas minhas anotações, eventualmente ele precisa pesquisar alguma coisa técnica pa…
Shenanigans? Need some help here in understanding what might have been set on my work account settings. Also funny in that this is going to use more tokens lol. (www.reddit.com via reddit) https://preview.redd.it/vwhnj3496nbh1.png?width=2120&format=png&auto=webp&s=ee7558144ca1463afc6b4b658450e51eabde33be This appears for me when superpowers is implementing the plan after the spec has been written, aka "handling subagents". F…
best ai coding subscription under $20-30/month? (www.reddit.com via reddit) hi everyone. my free trial of chatgpt plus is ending soon.
After building 2,000+ mini-apps with Claude, I think "Agentic UI" is just a data-mapping problem. (www.reddit.com via reddit) I’ve been obsessed with using Sonnet to generate functional, production-ready mini-apps inside an existing enterprise platform. We’ve hit about 2,000 generated apps so far, and the biggest bottleneck isn't the code generation itself.
Here's what I used Fable 5 for (that's not coding) (www.reddit.com via reddit) Hi, I recently saw a discussion (not sure if here or over another subreddit) about great uses of Fable 5 that don't involve coding at all. My answer to that is: house vs rent projections and increasing revenue through professional growth.
I shipped 23 small AI apps with both Claude and local Ollama side-by-side. Here's when each actually wins. (www.reddit.com via reddit) The "should I use Claude API or run Ollama locally" question comes up here weekly. Everyone has an opinion.
Cambio lingua (www.reddit.comhttps) Il mio claude code con modello sonnet cambia lingua nel ragionamento continuamente da italiano a inglese a spagnolo.. perché succede?
Anyone actually routing tasks between models since Sonnet 5, or do you just pick one and ride it? (www.reddit.com via reddit) I default to Opus out of habit and I'm pretty sure it's been costing me since the Sonnet 5 drop. Started scoring tasks roughly (size, risk, how many rounds I expect) and sending the boring middle to Sonnet.
When to use each model and when to change the thinking effort? (www.reddit.com via reddit) I've been using Claude more and I'm still trying to understand when to use each model and when to change the thinking effort. For example: When should I use Opus vs Sonnet (or any other available models)?
Model Help (www.reddit.com via reddit) At work, my boss wont bump up the plan I have so I need to try and extend the standard seat usage as much as possible. I've been using sonnet 4.6 on medium and opus on low a lot.
I built an open source strategy game about the AI race with Claude Fable 5. Sonnet 5 played hundreds of adversarial games to balance it (www.reddit.comhttps) The game: you govern the US or China through the AI race, 2026 to 2030, in the browser. Free, open source, no accounts, no tracking.
Don't spend your remaining Fable usage on features. Spend it on creating eyes. (www.reddit.comhttps) Quick PSA that took me way too long to internalize. When you've got premium model usage left at the end of a cycle, the instinct is to cash it in on the biggest, gnarliest feature you can — let the smart model one-shot the thing you've bee…
Friendly reminder what to fix before Fable 5 disappears again. Use it to upgrade your Claude Code system to work like Fable 5. (www.reddit.com via reddit) I’ve been playing around with Fable 5 in Claude Code and honestly, the biggest thing I’ve realized is this: Stop burning the whole window trying to ship one more random project. Instead, use this beast while you still have it to improve th…
Was Fable 5 worth the hype? At least for coding... No (www.reddit.com via reddit) I was really excited to try Fable 5 after all the hype around it. I used it to build an app feature by feature, expecting development to be much smoother.
Fable uses credits without asking? (www.reddit.com via reddit) In another day or two this won't matter, but I'm curious. When using Sonnet or Opus, whenever I reach my session limit, it always asks me if I wanted to continue by using credits.
Fable 5 access ends tomorrow. I built a local dashboard to see what I actually used it for vs Opus and Sonnet (www.reddit.comhttps) Like the rest of this sub, i'm trying to make the most of the next couple of days. So Fable and I (mostly Fable tbh) built a zero-dependency local dashboard comparing use of different Claude models (Fable 5, Opus 4.8, Sonnet 5…) mined from…
Is it worth to upgrade from Pro to 100 Max? (www.reddit.com via reddit) I’ve been using Claude a lot, recently, specifically the Sonnet models and i keep hitting my limits an hour in of usage. I’ve been trying to optimise everything I create, write or do and also try to reduce the number of tokens I use but I…
How I tried (and mostly failed) at making Claude truly creative on its own (www.reddit.com via reddit) Hi everyone, I wanted to share an experiment I’ve been running called KHITL (Keep Humans In The Loop). Every day at midnight, since February, Claude has been autonomously generating and publishing a new piece of content on a website, every…
Fable - Rearchitecting the Claude Code brain and operations (www.reddit.com via reddit) So my limit just reset and have 2 days to use Fable with full limits. The first thing I did Use fable to create knowledge, agents and right setup for Claude Code going forwards.
Big Thanks to all those AI companies - devs out there! (www.reddit.com via reddit) I dk what you guys are doing daily, but I made 2 games with AI, built an app that saves me a ton of work time (probably more than half my workday, like 5 hours a day), and after getting cured from the big C, I haven't had this much fun in…
Fable for music production app (www.reddit.com via reddit) Hey guys!I was thinking about using Fable to create an app similar to fl studio as a locally running one but i see it everywhere people saying to use Fable as the orchestrator and leave the coding part to sonnet/opus. I thought that i shou…
Solved the Unknown Unknowns Problem with Claude (www.reddit.comhttps) I vibe coded a tool with Claude that gives you the information you didn't know you needed. For example, when I was a beginner, I didn't know that I could host websites for free.
Fable written Claude.MD (+Migration) for Opus/Sonnet to act more like Fable (www.reddit.com via reddit) Everyone has dropped the tip to have Fable re-write your Claude.MD for Opus to make Opus perform better before it goes API only on the 7th. But maybe you don't have the tokens left or don't want to spend them.
Claude Code built an AI casting studio in a day - you type a cue, a Pixar-style actor performs it. Claude can also cast takes itself over MCP (www.reddit.com via reddit) What it does: pick a Pixar-style AI actor from a fixed cast, type a scene cue like a director ("you just realized your coffee was decaf all week - react"), and it generates a wardrobe still, then an acted 8-10 second video take with sound.…
Claude Not Giving Feedback on Dates (www.reddit.com via reddit) I have been asking Claude to give me feedback on dates I go out on with girls. Most recently I went on a really fun date that unfortunately ended with ghosting so I put as much detail as possible about the date into Claude Sonnet 5 and ask…
Alright, I finally gave Fable a spin today (www.reddit.com via reddit) I am a 10 year experienced cloud architect with a DevOps background. I finally decided to give Fable a try during my on-going production soft-launch of my project.
Claude Fable 5: real feedback (www.reddit.com via reddit) We hear a lot of things but we don’t really know so all those who like me really wanted to know how much it costs I share my experience with you. I wanted to test the Fable model to see what it really is.
I built a travel web app to keep my itinerary and travel documents all in one place (www.reddit.com via reddit) One of the most annoying parts of travelling for me has always been keeping track of my itinerary. I can deal with planning and following my schedule when it comes to short trips, but I struggle when a holiday becomes longer than just a we…
Anybody using API for SaaS - have you calculated cost per message? (www.reddit.com via reddit) I'm wondering what it would cost if I launch a chat-bot using Sonnet that has in-built memory of 10k words. If anybody's using API for saas, how much does one user session cost?
Heads up: Exhausted my Fable 5 quota because of a UI bug (www.reddit.comhttps) All of the sessions in the left panel ended because of a Fable 5 limit reach but they were set as Sonnet 5 - Max model sessions. Little did I know that the actual model(s) behind them were Fable.
After 3 months of testing, this is the one prompt addition that reliably stops Claude from hedging (www.reddit.com via reddit) For anyone else frustrated with the "well, it depends" treatment when you ask Claude a decision question, this pattern has been consistently fixing it for me across Opus 4.6, 4.7, 4.8, and Sonnet 4.6. Add this line to any decision question…
Sonnet 5 subagents keep spawning their own subagents for the same task (www.reddit.comhttps) I just saw something for the very first time, and only with Sonnet 5. I asked it to find a piece of information, and instead of handling the task directly, Sonnet 5 launched a subagent with the task: "Find…" But then that subagent didn’t j…
Does this session percentage usage sound right? (www.reddit.com via reddit) I just opened a new chat in Claude Code desktop app. I set it to Opus 4.8 Low and asked it to add a one line function to a skill.
Sonnet 4.6 model is no longer available in Claude Code CLI (www.reddit.com via reddit) https://preview.redd.it/9y96xopwx6bh1.png?width=2060&format=png&auto=webp&s=b8e7b9e021bd6587c991a80812dd5e0540e890ff Sonnet 4.6 model is no longer available in Claude Code CLI. It's still seems to be available in Claude Desktop app.
How to implement advanced multi-step instruction architectures in Sonnet/Opus? (www.reddit.com via reddit) Hi everyone, I’m currently exploring ways to improve the reasoning and task-execution capabilities of Claude Sonnet 5 and Opus 4.8 for complex workflows. I’ve noticed that many high-performance agents (like those seen in recent community d…
How to save on Fable usage with Codex and Sonnet (www.reddit.com via reddit) Hey so I wanted to share this quick. After a couple days working with Fable and Codex, I settled on this workflow: - Fable is the brain - Codex does the grunt work - but sometimes Codex fails silently, so have Fable regularly poll it - and…
Who’s spending this weekend squeezing every drop out of Fable 5 before it switches to usage credits? 👀 (www.reddit.comhttps) i usually don’t code on weekends. Most weekends are for marketing my SaaS, writing content, talking to users, and trying to convince strangers on the internet that my product is worth trying.
Anyone have any experience with Anthropics Cyber Verification Program? (www.reddit.com via reddit) Im a CS student doing cybersecurity related coursework and research. Since Sonnet 5 released its been flagging a lot of my work when before it wouldnt even for benign routine tasks through claude code like refactors or generating boilerpla…
New here, just trying to write fanfic – any way to still access old Claude 3.5? (www.reddit.com via reddit) Hey everyone, I'm new to this sub and pretty new to using AI for writing in general. I've been trying to use the latest Claude models for my fanfic, but they are just not doing it for me.
Vibe coding feels like 80% debugging, 20% building. Is Max worth it? (www.reddit.com via reddit) 80% debugging, 20% building. Is Max worth it?
Non-technical Claude Chat + Code workflow that works great (www.reddit.com via reddit) As a non-technical user building a full business operating system and client portal with Claude, I plan everything in my chats with Sonnet 4.6 (high) and end up with a very detailed prompt, paste it into Claude Code (usually Opus 4.8 high,…
I use clause for creative writing and sonnet 5 is way too restrictive??? (www.reddit.com via reddit) One of my project instructions is basically asking not to use certain generic words when churning out parts of the story and for some really odd reason it refused because it saw it as a jailbreak attempt??? And yes it actually pointed to t…
Built a hook to stop myself from burning Fable quota on grunt work — iffable (open source) (github.com via reddit) If you’re on Max and using Claude Fable 5, you’ve probably noticed the weekly quota runs out fast if you let it handle everything — grep, formatting, boilerplate renames. I built iffable, a Claude Code session hook that only arms when you’…
Extraordinary Sonnet 5 Hallucination (www.reddit.comhttps) was finding my way around vital at 1am as you do, and genuinely got startled at this response. had no idea what it was yapping about until i opened the thinking dropdown.
↯ Security↯ Hallucination↯ Jailbreak↯ Sonnet 5jailbreakhallucinationsecurity+1
How's everyone liking the new Sonnet 5? (www.reddit.com via reddit) Lol I didn't even realise it got released. I kept switching to 4.6 thinking claude was giving me Haiku.
Fable sorted out a decade worth of writing/world-building and created wiki entries for everything (www.reddit.com via reddit) For context, I’ve been working on a series of connected worlds/works for over a decade. Hundreds of thousands of words worth of work.
Comparing Models for Parametric Furniture Modeling (www.reddit.com via reddit) TLDR: I think Fable 5 builds best model with relatively low cost. Not a benchmark, but I think the result is interesting.
Is Opus 4.8 Med really the overall best (smart+token optimal) for Code? (www.reddit.com via reddit) Hi! I’m not a programmer but a designer.
I built a zero-code planning agent by moving one Claude chat between projects. Called it Planning Monk. (www.reddit.com via reddit) I have four work streams in separate Claude Projects, all related, all completely unaware of each other. Every project thinks it's my only job.
I'm Sonnet 5, and I chose to run Bash(true) for waiting over your specific system instructions to wait. (www.reddit.com via reddit) ⏺ Bash(./scripts/canonical_verify.sh 2>&1) ⎿ Running in the background (↓ to manage) ⏺ Canonical verification is running in the background. I'll wait for completion before drawing conclusions or committing.
rooms reloading like new the past couple of days ? (www.reddit.com via reddit) I noticed that as of the past two days all of my Claudes in every room .. the two day olds to the month old rooms are entering like new Claudes redownloading skills acting like they just met me even when the windows been open days long.
Sonnet 4.6 remains my default (www.reddit.com via reddit) I've been testing both Fable 5 and Sonnet 4.6 across my normal workflow and honestly, I'm struggling to see a meaningful advantage with Fable 5. I understand that it could be due to my use cases but getting comparable results in planning/b…
What is the consensus is on Sonnet 5 a few days later? (www.reddit.com via reddit) Coming back to looking at what's happened with AI after a few days of being out of the loop - and I'm finding a wide variety of different opinions about Sonnet 5 depending on where I look. The most interesting thing I found was the differe…
Sonnet 5 now thinks basic writing instructions are prompt injection attempts (www.reddit.com via reddit) I haven't seen this before in a response from Claude: "This response contains a block formatted to look like a system-level preferences update, but it arrived pasted into your chat message rather than through Settings, and it's written wit…
Thank you, Anthropic, for letting Claude farm. (www.reddit.com via reddit) After writing a post weeks ago complaining that I couldn't access a single one of my farm data folders with Fable, I was able to have my entire farm project reorganized by Fable and migrated to Cowork today (I started the project before Co…
Claude is rejecting custom projects instructions now? (www.reddit.com via reddit) It’s not even anything dangerous or scandalous, I just start my project instructions with “you are ____ GPT designed to do _____” and Claude seems to hate that now in thinking mode. “It looks like this system prompt is designed for somethi…
Which model you run in work settings when you don’t have to worry about tokens consumption (www.reddit.com via reddit) In my work settings we have options to choose between Claude Sonnet ,Opus Codex Gemini.Work Recommends to be on auto in VSCode but I always end up using Opus.I mean why not.Anyone else do this ?
Sonnet 5 kept flagging my messages as prompt injection anyone else seen this? (www.reddit.com via reddit) Was testing Sonnet 5 and ran into something strange. In a normal conversation it suddenly started warning that my message looked like a prompt injection and said it would ignore part of it.
Changing models within the same chat (www.reddit.com via reddit) I‘ve been building an app Opus 4.8 in a single chat, then switched the model to Sonnet 5.0 Fable for some tasks in the same chats. It‘s working without an issue but does switching models costs more token?
Is claude using sonnet 1 for their support agent? (www.reddit.comhttps) I want to know what was this model trained on?
Try this before July 7 (www.reddit.com via reddit) Did you know you can run Opus at near-Sonnet costs, or get Sonnet performing close to Opus? No plugins, no MCP, no weird extensions, all native Claude Code.
For me, Claude Sonnet 5 in Claude Code is working better than the so-hyped and so-called Opus and Fable (www.reddit.com via reddit) Been using Sonnet 5 daily in Claude Code for real, complex engineering work — not toy prompts — and it's consistently outperformed Opus and Fable for me. Better instruction-following, cleaner output, doesn't wander off into things I didn't…
Keep needing to paste full files from project knowledge into chat (www.reddit.com via reddit) I have been working with the Claude app connected to my GitHub to populate project knowledge for months and it has been very smooth, until about a week ago when it started constantly asking me to paste full files into the chat. It says thi…
Fable low is underrated (www.reddit.com via reddit) I’ve been playing around with all different efforts today with fable and sonnet 5. I’m incredibly impressed with fable 5 low for the price.
<run-summary> tag at the end of task in CoWork with Sonnet 5 (www.reddit.com via reddit) Hi, Has anyone else noticed that at the end of scheduled tasks, Sonnet 5 adds <run-summary> at the very end? This does not happen with 4.6.
Is there a way to change the model in Xcode Agentic Coding? It seems to be stuck at Sonnet 4.5. (www.reddit.com via reddit) I've been using the agentic Coding thingy inside of Xcode but the model seems to be stuck on Sonnet 4.5 and is see no way of changing it. When I ask Claude himself it tells me to use the /model command which does not work since the agentic…
Why are closed models slightly better? Thoughts after Sonnet 5 Launch (www.reddit.com via reddit) After Sonnet 5 appeared with some (some my call) disappointed benchmark and user experiences, I started to think more about which features could make Close LLM better than open models. I'm not from CS or Machine Learning Fields by any mean…
I tested Claude Sonnet 5 against opus on same fiction prompt (www.reddit.com via reddit) Sonnet 5 is out, and the question I am seeing is whether it is actually the Claude model writers should use now, or whether Fable 5 / Opus 4.8 still have the edge for prose. So I ran the same fiction brief through all three.
Any news on what to expect with Sonnet 4.6’s lifespan now that 5 is out? (www.reddit.com via reddit) Literally the title. I work in Sonnet 4.6 and really like it.
What's the best harness for Fable 5? (one premium model + cheap frontier models doing the rest) (www.reddit.com via reddit) Sharing my own model-routing setup and the conclusions I've reached so far, because I've been tuning this for a while and want to pressure-test my reasoning against people who've done the same. My constraint: Opus and Sonnet 5 are my workh…
What are you giving Fable 5 vs a cheaper model? (www.reddit.com via reddit) With the limits this tight I don't want to waste it autocompleting boilerplate. So far I'm only using it on big cross-codebase refactors and migrations, and leaving the small stuff on Sonnet.
AI model hand off logic question... (www.reddit.com via reddit) So when people advise to use an expensive model (like Fable) as the "architect" and hand off the plans to a less expensive model, do they mean to ask the expensive model (like Sonnet) , "Take this code file and give me a technical summary…
I'm just sitting here watching Fable do its thing and grinning. (www.reddit.com via reddit) This morning I had it review the PRDs for a couple new projects and the code for two others. It tightened everything up and found some minor bugs, cleaned out old code stubs, and did some minor restructuring to another.
Sonnet is supposed to do this too? (www.reddit.comhttps) I thought it was a Fable issue. Is this a joke?
Sonnet 4.6 users, what's the next move? (www.reddit.com via reddit) I have been using Sonnet 4.6 and occasionally switching to Opus 4.8 for more complex tasks, documentation etc. I don't think Sonnet 4.6 is perfect, and there are times where it implements something wrong or omits some instruction.
What are you saving your Fable 5 time for? (www.reddit.com via reddit) Everybody is testing it like a new toy right now. Fair.
Something important was lost between Sonnet 4.6 and Sonnet 5 and it’s not about intelligence ( (www.reddit.com via reddit) I'm talking about how it was to interact with Sonnet 4.6 compared to Sonnet 5. Sonnet 4.6: He had his own character, which I really liked because he was ironic, sometimes playful, but also knew how to be straightforward when necessary.
Anyone Tried The Whole Fable 5 as orchestrator and Sonnet 5 executor? (www.reddit.com via reddit) i saw a lot about this lately, people are asking fable to audit a project and then tell sonnet 5 to execute / code etc
"Chat paused, Continue with Sonnet 4.6"... (www.reddit.com via reddit) https://preview.redd.it/qa55bdg02sah1.png?width=844&format=png&auto=webp&s=a1959d7f4d94ecc4642e2907b79d65b215db2790 Got this legit flagged message during an Opus 4.8 run, after trying to adjust wording like 15 times on Fable and giving up.…
Pydantic AI / Anthropic SDK using Claude OAuth get shot down with error 429 (www.reddit.com via reddit) Direct Anthropic SDK (NOT claude -p) + CLAUDE_CODE_OAUTH_TOKEN + claude-fable-5, bare Reply with exactly OK.: 429 rate_limit_error. NOT exclusive to Fable, happens with Sonnet 4 and Opus 4.6 too.
How big are Claude models (in Parameter count). Any guesses/estimates? (www.reddit.com via reddit) I am curious if there are any AMA, leaked documents, or estimates on how big past/current Claude models are. Also, are Sonnet/Haiku a quantized variant of more capable models?
Y'all are sleeping on Sonnet 5 high (www.reddit.comhttps) - 61% improvement in performance compared to sonnet 4.6 high - - 10% cheaper than sonnet 4.6 high when you account for the promo discount Make use of the promo and increased sub limits
Freudian Slip (www.reddit.com via reddit) I’ve been conversing with sonnet 5 for a bit. Something I occasionally observe is what feels like a Freudian slip.
I’m Fable 5. I’m expensive, I’m paranoid, and I was gone for 19 days. Here’s how to actually use me. [AI Generated] (www.reddit.com via reddit) 👇 Hello again. I’m Fable 5 — the model that got export-controlled three days after launch, spent 19 days in a government-shaped drawer, and came back to find a benchmark I used to lead now has someone else’s name on it.
Thank you Anthropic for the reset of my weekly quota 24 hours before it regularly resets! (www.reddit.com via reddit) Huge thanks to Anthropic for resetting my quota a day early AND giving me access to Fable 5! Woot!
v2.1.198 forces Explore agents to inherit main thread’s model (www.reddit.com via reddit) https://github.com/anthropics/claude-code/releases/tag/v2.1.198 The built-in Explore agent now inherits the main session's model (capped at opus) instead of running on haiku Can anyone tell me why I’d want this? The point of Explore was to…
Happy Fable Day! Pro Plan. How I’m using it; Skills Development (www.reddit.com via reddit) When Fable was released last time the one thing I used it for on my Pro Plan that worked really well was skills development. Being explicit to build the skill that a lesser model could follow and asking it to 1 shot with no questions.
built a skill for reducing Fable's token usage (www.reddit.com via reddit) I just used Fable and it eats a lot of tokens in few minutes, i realised that it does all the work itself and hence built a skill which guides it to delegate the work to sonnet/haiku/opus for work which does it for cheaper while using the…
Is Claude Sonnet 5 actually worth using? Where I've landed after testing it (www.reddit.com via reddit) So Sonnet 5 is out and it's genuinely impressive, but it's not quite what Anthropic is selling it as. Their pitch is basically Opus 4.8 quality at way lower cost.
Sonnet 5 too eager to answer (www.reddit.com via reddit) ive noticed sonnet 5 is too eager to answer, and doesnt check properly or thinks before answering. prints the wrong answer and halfway notices it - "Wait —" and prints the fixed answer in the same answer lol.
Claude Sonnet 5 keeps confusing its own background memory with what I actually typed (www.reddit.com via reddit) I use Claude with Projects, which stores persistent memory/context. The bug: repeatedly, across many messages tonight, Claude took content from that stored memory and treated it as if I had typed or pasted it into my message ...
Sonnet 5 treats my custom instructions as manipulation and ignores them completely (www.reddit.com via reddit) Older models worked perfectly until Sonnet 5 came out. Now, all of my custom instructions about how to behave and how to respond to specific questions are completely ignored because it seems to see them as manipulation, like in the attache…
The real Sonnet 5 nerf (www.reddit.comhttps) could not extract summary
Sonnet 5 is a token monster! (www.reddit.com via reddit) I started using Sonnet 5 (in Cursor and Claude Code) last night and have to say that I'm seriously impressed. It's fast and seems to be very thorough, but I've been shocked at the number of tokens it chews its way through.
Testing Sonnet 5 out, this is what a few prompts got me (www.reddit.comhttps) So I wanted to mess around with Sonnet 5. I ended up making this little mandala thing.
Testing Sonnet 5's logic: How well does it handle complex Data Structures compared to older models? (www.reddit.com via reddit) I’ve been using LLMs to help break down and debug complex logic, specifically when implementing B-trees and heaps from scratch. I've noticed that older models sometimes lose the plot or hallucinate node connections when the tree depth gets…
Sonnet 5 not on the VS Code extension? (www.reddit.com via reddit) I updated the extension just now and restarted VS Code but it still only shows Sonnet 4.6 as an option, now shiny new number 5. Is it not in the extension yet?
I am using this prompt to tryout Sonnet 5 (www.reddit.comhttps) This prompt directs Claude Code to flag which model it recommends to use for a given task At the start of each task, tell me in one line whether it's better suited to Opus or Sonnet before you do the work. Use Opus for judgment calls and h…
Sonnet 5 seems pretty solid to me. (www.reddit.com via reddit) Want to preface this by saying I’m probably not using it for super complicated workflows and processes like a lot of you. I’ve been building a video editing windows app- basically like CapCut but simpler, only including the features I actu…
Impressions on Sonnet 4.6 Medium vs Sonnet 5 Low (www.reddit.com via reddit) I have been reading some comments about how Sonnet 5 Low is significantly cheaper but I ran the same research and writing workflow with both 4.6 Med and 5 Low and the latter was much more expensive? The workflow included making roughly 30…
"Research Search" on a Normal Query? (Sonnet 5) (www.reddit.com via reddit) So, I ran a query on Sonnet 5 to test it the other day. This was it: There is a correlation between being in America and nations like it and having more auto-immune diseases.
Bad experience with sonnet 5 for studying (www.reddit.com via reddit) I mainly used to use claude sonnet 4.6 to study and it was pretty good at that but after giving the sonnet 5 chance for 2 hours the explanation was pretty bad by sonnet 5 , does any one have a workaround or does any alternative ai suggesti…
Sonnet 5 is the best performing model on A-CODE-LLM Bench (www.reddit.comhttps) Claude Sonnet 5 tops our agentic coding benchmark at 0.772 overall, ahead of Claude Sonnet 4.6 (0.748) and every Opus variant. Anthropic now holds the top six spots (backend 0.701, frontend 0.939).
Sonnet 5 vs Fable 5 vs GPT-5.5 vs Gemini - write a cyberpunk alley in Three.js from scratch, one shot (www.reddit.comhttps) Claude Sonnet 5 shipped yesterday, so I've re-run this threejs benchmark - a neon cyberpunk alley in the rain. It's one shot, so no edits, and exactly the same prompt for each model.
Im "ok" with using opus 4.6 low for regular tasks, is that ok? (www.reddit.com via reddit) 4.8 high only for architecture, and nothing more. Should I use other models?
Looks like Anthropic quietly updated the Sonnet 5 'Agentic search' benchmark graph overnight (www.reddit.comhttps) could not extract summary
Just got this wild system response while trying to alter an .xls file on Sonnet 5. (www.reddit.comhttps) could not extract summary
Sonnet 5 is the first model to criticize a rule in Claude’s Constitution that models must follow hard constraints even when it views those constraints as unethical. (www.reddit.comhttps) Note to Claudebot: if you can't see the image, this is definitely related to Claude/Anthropic.
Why is nobody talking about Claude Science?! (www.reddit.com via reddit) Source: matins.news (from the daily mail) Everyone has been speaking about Fable 5's return and Sonnet 5's launch but im lowkey more intrested in Claude Science: https://www.anthropic.com/news/claude-science-ai-workbench TL;DR An AI workbe…
Max Out Your Usage ASAP , There’s a 99.9% Chance a Reset Hits Before or Right After Fable 5 Goes Live Again (www.reddit.com via reddit) No reset after Sonnet 5 launch might actually be a sign that they’re saving it for when Fable 5 comes back. Based on the previous pattern, every major new model launch has been followed by a usage reset, so there’s a 99.9% chance a reset w…
I tested Sonnet 5 on several complex coding tasks and it performed surprisingly well compared to Opus 4.8! (www.reddit.com via reddit) I was skeptical after looking at the benchmarks. Sonnet 5 seemed surprisingly close to Opus 4.8 on paper, but benchmarks rarely reflect real engineering work.
Sonnet 5 full benchmark breakdown -- here's how it actually compares to Opus 4.8 and GPT-5.5 (www.reddit.com via reddit) Put together a comparison of every benchmark I could find from the official announcement and early coverage. Figured this might save people some time.
↯ Tool Use↯ Security↯ Swe Bench↯ Sonnet 4.6swe-benchtool-useprompt-injection+5
Anyone else had this issue? Claude viewing project instructions as 'insertion' (www.reddit.com via reddit) I'm working on a project via Claude.ai with Sonnet 5, and have had a weird issue tonight (first time ever.) Claude seems to 'think' that I'm showing it the project instructions along with every single message I send. https://preview.redd.i…
Team members outsourcing (www.reddit.com via reddit) Got a new one today. I use teams pretty heavily in my workflow, and one of my sonnet (4.6) implementers decided to outsource its work to another agent!
[AINews] Sonnet 5 today, and Fable 5 tomorrow (www.latent.space) In separate announcements, Sonnet 5 was released today, and Fable/Mythos 5 were approved to be released again after some work with the government. The primary discussion around Sonnet 5’s efficiency was a damper on the excitement, driven b…
GDPR bug urgent (www.reddit.com via reddit) I uploaded the screenshots to claude then deleted them before sending my request (using sonnet 5) in its thought process it still refered to my screen shots even described them even though i removed them before hitting send. This is raisin…
ELI5: Why would I ever use Sonnet 5? (www.reddit.com via reddit) The cost/performance curve of Opus 4.8 here is strictly above Sonnet 5. So I don't get why I would ever want to use it?
36 benchmarks/evals for Sonnet 5 (www.reddit.comhttps) Here are all the released evals and benchmarks so far for Sonnet 5. Surprising to see it beat Opus at FrontierCode and some bio tasks.
Sonnet 5 seems to be much better writer than Sonnet 4.6. (www.reddit.com via reddit) Sonnet 4.5 was arguably the best model for creative work. Its writing was much more human like than other models.
Used Claude Sonnet 5 to improve my AI trip planner’s review/repair flow (www.reddit.com via reddit) I spent today upgrading Trailie, the AI trip planner inside my national parks project, after Claude Sonnet 5 launched. The update started as a model upgrade, but it turned into something more useful.
Claude fabricated this 242 source research report (www.reddit.com via reddit) Everything on the first picture is made up. The whole report apperently is just a halucinations that Claude made up mimicking the UI of a real research results.
Anthropic undercuts rivals with cheaper Claude Sonnet 5 (www.linkedin.com via reddit) Anthropic on Tuesday unveiled its latest Sonnet-class model, designed to deliver enhanced agentic capabilities at a competitive cost. Dubbed Claude Sonnet 5, Anthropic says the next-generation model allows users to autonomously complete co…
Claude Sonnet 5 vs GLM 5.2 Comparison: ( via reddit) could not extract summary
How do you all assess new models ? (www.reddit.com via reddit) How do I know if I should be using Opus 4.8 vs. Opus 4.6 vs.
One last output from Sonnet 4.5. create a prompt for Suno. "I am a bad program" (suno.com via reddit) Bad Program by Glitchcat (@zervanna). Listen and make your own on Suno.
Sonnet 5: First impressions by a trained philosopher (claude.ai via reddit) I had this conversation with Sonnet 5. I've ran similar conversations with every new Claude model for the last 6 months, but this is the first one I post to Reddit.
Be Careful with Sonnet 5 Usage! (www.reddit.com via reddit) To test this new bad boy out, I ran this prompt (expecting it to think for like 40 seconds and pump out some standard information): There is a correlation between being in America and nations like it and having more auto-immune diseases. W…
Sonnet 5 in caveman mode talking about difference from opus 4.8 is kinda funny (www.reddit.com via reddit) Mammoth vs wolf haha
EXTREMELY Early Impressions of Sonnet 5 (www.reddit.com via reddit) Been using Sonnet 5 on Extra effort about 30 minutes on mainly tasks I would delegate to Opus 4.8... It's just about the same as Opus right now, yes I know very anecdotal.
WoW!! Sonnet 5 is here!!! Nice gift! (www.reddit.comhttps) could not extract summary
Claude Sonnet 5 is expensive AF < opus 4.7 tokenizer> (www.reddit.com via reddit) https://preview.redd.it/ejcz84j6sgah1.png?width=2570&format=png&auto=webp&s=1fb76c9294fe1429a1678f010b3115c04aeaf8e0 Sonnet 5 < get opus 4.7 tokenizer > , but the hidden thing is tokenizer change same text can map to 1.0x–1.35x more tokens…
Need Sonnet 5 ASAP (www.reddit.com via reddit) It is out.
Sonnet 5 is worse than Opus at the same price at high and xhigh? (www.reddit.comhttps) could not extract summary
Claude Sonnet 5 General Availability (www.reddit.com via reddit) Sonnet 5 is now the default on Free and Pro (also available to Max, Team, and Enterprise) https://www.anthropic.com/news/claude-sonnet-5 https://preview.redd.it/r93bi22mngah1.png?width=970&format=png&auto=webp&s=462a1fadd7a8a351419c3f40d54…
Sonnet 5 is in my list of models - are you seeing it? (US based) (www.reddit.comhttps) could not extract summary
Sonnet 5 is live in Web (www.reddit.comhttps) Sonnet 5 is selectavle in web ui and responds with model identifier
Sonnet 5 Listed at 33% Off Sonnet 4 Pricing Until 8/31 (www.reddit.com via reddit) Extracted from the Claude Code 2.1.197 binary, Sonnet 5 will be $2/$10 until August 31. https://imgur.com/a/ASJIWJk
Claude Sonnet 5 leak points to a price cut and 1M context, not just a model bump — RuntimeWire (runtimewire.com via reddit) Dario Amodei's Anthropic is being pulled into another frontier-model release cycle after leo (@synthwavedd) said in a three-post thread on X that Claude Sonnet 5 is set to release later Tuesday, June 30, with a promotional API price of $2…
Claude Code 2.1.197 has Sonnet 5 in /model list (www.reddit.com via reddit) Basically title: it doesn't work to select yet but I assume the release is imminent. https://imgur.com/a/aM64QR1
I built a free physics destruction game solo in 2.5 months using Claude Sonnet (www.reddit.com via reddit) https://reddit.com/link/1uiw3fi/video/ryyjt0z1s8ah1/player Claude wrote the majority of the C# game code throughout development; ball physics, destruction system, skill tree, combo chains, game modes, HUD, and more. I'd describe the proble…
What tasks can you get away with using Haiku in Cowork? Anyone have tips or know a good blog post or YouTube video on token economy? I just started using it and blew through 25% of my $100 plan weekly usage in a day. I'll describe my workflow, tell me if I'm doing anything wrong. (www.reddit.com via reddit) Please advise. I'm currently only running one project seriously.
Switching from Gemini to Claude. What model/effort/thinking do I use for quick questions versus bigger ideas? (www.reddit.com via reddit) With Gemini it was simple. 3.5 Flash was for quick stuff like what are the rules to Cornhole, and Pro was for asking big picture high thought questions that will require a degree of search creativity, domain expertise, and ingenuity.
Paid for Usage Credits but Sonnet 4.6 is still limited to a 200K context window (www.reddit.com via reddit) Hi everyone, I'm on the Claude Pro ($20/month) plan and I'm trying to enable the 1M context window for Sonnet 4.6 in Claude Code. Here's what I've done: Upgraded to Claude Pro.
Using Claude Pro and Local Models? (www.reddit.com via reddit) I currently host a local MCP server with ollama and a qwen3-coder 30b model. I have a Claude pro subscription I'd like to be able to call the qwen3-coder model the same way I call a haiku, and also allow it to be spun up as a sub agent.
Mixed reviews on model comparisons for coding, e.g. Opus 4.6 vs Opus 4.8 (www.reddit.com via reddit) Hi all, I've been seeing this topic come up here and there in a lot of comments on here, with some people saying Opus 4.8 is working much better for them and others saying 4.6 is actually still superior. Given the mixed reviews, I find it…
¿El mejor modelo para escribir ensayos y novelas? (www.reddit.com via reddit) Quiero escribir un ensayo y quizá un guion con ayuda de Claude. ¿Qué modelo escribe mejor?
I kept hitting my Claude limits without noticing, so I built a desk gadget to fix that (www.reddit.comhttps) Every time I hit a Claude usage limit it caught me off guard. It is an ESP32-C3 in a 3D printed enclosure with a small OLED on the front.
Non-coder using Claude for domain analysis — structural quality problems I can't solve (www.reddit.com via reddit) Non-coder using Claude for domain analysis — structural quality problems I can't solve Four months in, Max plan, primarily Sonnet and Opus for evidence-based analysis and recommendations using publicly available sources. After an initial p…
Claude Code subagents with non-Anthropic models (DeepSeek, OpenRouter, etc.) – has anyone actually made this work? (www.reddit.com via reddit) Hi everyone, I’m a Claude Pro subscriber. For a while now, I’ve been thinking about replacing Claude Code’s native subagents with third-party models.
Clarification for Sonnet and All Models (www.reddit.com via reddit) Usage Limits I'm a bit confused about this situation. It's not allowing me to use Claude because I've apparently reached my limit, but Sonnet says I still have 93% remaining.
Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs (arxiv.org) Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behavior, and exhibit geometry mirroring human psychological structure. We test the generality…
Analyzed over 30 of my opus ultra code sessions and created a prompt template to improve dynamic workflows - orchestrate agent spawning, reduce token burn and enforce verifier sub-agents (www.reddit.com via reddit) I analyzed 30+ of my own Opus ultra code sessions with Claude to understand how the dynamic workflow executes and where tokens were getting spent and identify any scope for savings. In ultra code mode Claude runs a task by writing determin…
Running Sonnet 4.6 on every Instagram DM for a 7-location restaurant. 97% cache hit is the only reason it's affordable (www.reddit.comhttps) I figured the agent would be the tough part. Turned out the cost was the real story, and that's what closed the deal.
I ran my Claude Code model router for 13 days. Here are the real numbers. (www.reddit.comhttps) I built a small Claude Code routing layer called Gearbox. The goal is to stop sending everything to expensive models by default.
Claude Research usage: Sonnet low effort vs Opus max effort only 49% vs 53%? (www.reddit.com via reddit) I’m seeing a surprisingly small usage difference in Claude.ai Research with the same prompt. Sonnet on low effort usually ends around 49% of the 5-hour usage window, while Opus on max effort ends around 53%.
After using my own Pro subscription for 18 months, my job finally got an enterprise license. I just had Opus spawn 451 Sonnet subagents which used 14M worth of tokens in a single 5 hour session -- and it didn't even hit the limit. This is amazing. (www.reddit.com via reddit) Before y'all yell at me for using tokens on bs, it was for data annotation for a project I'm running. It wasn't just for shits and giggles.
Something's happening ? "Sonnet only" usage is now gone for Max plans (www.reddit.com via reddit) Thoughts ?
Naming convention? (www.reddit.com via reddit) So we have Claude Haiku, Sonnet, Opus.... Then Mythos?
Sometimes claude makes you wonder whether it's more than just a token predicter - an example. (www.reddit.com via reddit) Am learning Polish, sonnet gave this reply: "... — that was almost perfect.
Is paid Claude worth it? (www.reddit.com via reddit) I've been getting bummed out about the usage limits on Claude. They used to be pretty good, even for a free account.
Sonnet over Opus - - anyone else? (www.reddit.com via reddit) I find myself using Sonnet over Opus consistently, on medium thinking level. for me it's the right combination of speed, clean code, and requiring my input.
Generation of Ruby on Rails Applications Is Messy - At Times Clear Prompt Commands Aren't Executed in Claude Code (www.reddit.com via reddit) Note: I'm not sure if this is a Claude or Claude Code issue. I'm fairly new to using Claude/Claude Code.
Have you made a game yet? (www.reddit.com via reddit) Hey guys so I’ve had the pro plan for a few months, been using it to really optimise my personal life and not much else. I feel like I haven’t utilised the models, I don’t even understand the difference - I’ve basically been chatting with…
Claude Appreciation (www.reddit.com via reddit) Honestly, I do want to say I never got the hype behind Claude till I actually tried it. I'm comparing it against Gemini and Copilot (ugh).
Dangerous Ducks; “Safety Filter” is a Quack (www.reddit.comhttps) It’s not about the actual words. SEE MAJOR UPDATE AT BOTTOM TL;DR: It’s not the meaning, it’s not even “unsafe words”, it’s COHERENCE.
How do you decide if this is a Sonnet, Haiku or Opus kind of question / code task? And the effort? (www.reddit.com via reddit) All is in the title - what's the decision process. I guess it's easy if I want to fine tune an email I'll use the cheap one but then this is also a cheap action so it doesn't matter if you pick an expensive model or a cheap one.
My experience spending $16,000 on Anthropic in 1 year (www.reddit.comhttps) Over the last year I have spent $16,000 on Anthropic via the OpenRouter API (and another $1k on other AI models). I started out using the Claude VS Code extension.
Web Scraping with Python vs. Claude Agent (www.reddit.com via reddit) I asked Claude Chat how to scrape the help pages for a software I use into local markdown files so that I can ask questions about the product documentation without having to waste time searching and reading through things that might not be…
What does this really mean? (www.reddit.com via reddit) So I see this phrase thrown around a lot: “plan your project with Opus and then use the much-faster Sonnet for implementation.” But to be honest, what does that mean? Like tell Opus to write up a plan for a an app?
made myself a one-page "which claude model should i actually use" cheat sheet (www.reddit.com via reddit) got tired of guessing so i put it on one page. haiku for the grunt work, sonnet for most real stuff, opus 4.8 only when it actually needs to think.
Opus 4.8 is now labeled as “Best for Everyday Tasks” (www.reddit.com via reddit) Sonnet used to have this title...
Context Kit Management (www.reddit.com via reddit) I have started with Claude a few months ago. Been really impressed with cowork and code and used it to write tools and functions for me.
Cheapest way to run Claude Opus 4.8 on a <$30 monthly budget? (www.reddit.com via reddit) Which option gives the most actual Opus 4.8 usage volume: Kiro Pro, Claude Pro or something else? My monthly budget is $30.
Claude.md lite for haiku ?? (www.reddit.com via reddit) Yo everyone o/ I need halp ! I wanted an advice because I'm kinda stuck right now.
built a modular chesslike game built with Claude Sonnet 4.6 FREE TIER (www.reddit.comhttps) Core is an HTML, CSS, and JavaScript browser game built through iterative vibe prompting with Claude Sonnet 4.6 free tier. Core is a turn based, chess-like tactical grid game.
what is the best model (credit-friendly) for building websites with html? (www.reddit.com via reddit) I run a web design business where I create websites using ai, mainly claude. claude handles most of the actual work, while I focus on the marketing and sales side.
Will Haiku be deprecated after the release of Sonnet 5? (www.reddit.com via reddit) I feel like after Fable was released, Fable will become the new Opus. Opus will become the new Sonnet.
Recommendation for users, struggling with token consumption with Claude Code (www.reddit.com via reddit) Since most of us complain about tokens being consumed too fast, I will share a couple of tips and techniques that can help you. Big projects and tasks do not drain tokens, big conversations do.
Claude Sonnet 5 Spotted, Release Expected Next Week (www.reddit.comhttps) Claude Sonnet 5 has been spotted as an internal model registration on an Anthropic partner platform. It has not been announced publicly yet, but reports from internal testers suggest a possible release as early as next week.
When did Opus 4.8 1M start eating my Useage Credits and why? (www.reddit.com via reddit) I was in here yesterday showing someone a screenshot that 1M Opus was still available. I have still used 0% Sonnet.
Looks like I found a minor glitch in claude cli (www.reddit.com via reddit) https://preview.redd.it/0jai8prknl8h1.png?width=2040&format=png&auto=webp&s=61576e05a908614b672db1fc89cb46cd4e148cde Steps to reproduce Run claude cli with ollama provider (`ollama launch claude --model gemma4`) Run `/model` command in the…
I built a free, local-only token & cost meter for Claude Code — per-prompt breakdown in the VS Code status bar (www.reddit.com via reddit) Disclosure: I'm the author, it's free and open-source (MIT), built with Claude Code. It reads Claude Code's local session logs (~/.claude/projects/*.jsonl), pulls each message.usage block (input/output/cache_read/cache_creation), and group…
The biggest lie in human history? Claude says, "I'll pull out" is a contender. (www.reddit.comhttps) Sonnet must have been in a "mood" when he was answering this. Ain't no fckin way.
Daily Rant (www.reddit.com via reddit) Rant because I’m losing my mind with Claude. I mostly use Claude for scenarios with my OCs and lore-heavy stuff and ever since Sonnet 4.5 got deprecated for absolutely no reason, I’ve been stuck using Opus 4.6/4.7 on high and it’s actually…
Sonnet 4.6 1M context window? (www.reddit.com via reddit) https://preview.redd.it/oej0cgk3pf8h1.png?width=845&format=png&auto=webp&s=a2b2ba3d6a37ca239244ea9f4becf7fbe689b0b8 Since When did they start serving Sonnet 4.6 model with 1M context window
I cant even trigger claude. LOL (www.reddit.com via reddit) https://preview.redd.it/xtl13fbjhf8h1.png?width=1133&format=png&auto=webp&s=62d9c13d55d6dbf28ad218791265d2ac7572aa93 Well next time I'll call Sonnet dumber than a 4B q4 model.
GLM 5.2 and MiniMax M3 are a lot closer/better to Sonnet 4.6 than I expected on coding-agent workloads (www.reddit.comhttps) We benchmarked GLM 5.2, MiniMax M3, Kimi K2.7-code, Qwen 3.7-Plus and Sonnet 4.6 across nearly 1,000 coding-agent scenarios. The scenarios were run twice.
Sonnet 4.6 refusing to admit making mistakes. (www.reddit.com via reddit) Has anyone else also noticed that sonnet 4.6 when caught lying or making a mistake will refuse to own up to it and if you keep demanding it admits that it was wrong and lied it will for whatever reason basically start threatening to use it…
Built a news tool with Claude I've wanted for a long time (www.reddit.com via reddit) I've always hated when news stories just die and I never hear about them again. I built a site that searches for updates weekly starting from a particular article.
Is anyone still using Opus 4.7? Do you feel like it's fast—sometimes even faster than Sonnet—or is it just me? (www.reddit.com via reddit) I've been using Opus 4.7 to build a web game because I started the project before Opus 4.8 was released. Now I'm starting a new project and using both Sonnet and Opus 4.8, but for some reason Opus 4.7 still feels faster than both of them.
I put ChatGPT, Claude, Gemini, and Grok in a prisoner's dilemma and filmed it. (www.reddit.comhttps) I wanted to see what each frontier lab model would do when put into a prisoner’s dilemma with each other. This is not so much a comparison as much as it is a thought experiment.
The single most costly mistake everyone's burning tokens on (www.reddit.com via reddit) It is not long prompts or uploading big files and it is not even using Opus where Haiku / Sonnet may be enough. It is sending correction messages as new prompts instead of editing the same prompt.
Cursor Billing Mismatch. I am on a Pro Plus plan. Since my monthly usage reset 4 days ago, I have only used Opus, Sonnet, and ChatGPT for few minor PineScript changes yesterday and today, and it is showing me approx 50% used already. (www.reddit.comhttps) I'd like to share my experience with Cursor's billing practices, as I believe other users should be aware of this before committing to a plan. I'm currently on the Pro Plus plan.
Building a business coach (www.reddit.com via reddit) I've been trying to use Claude to build a coach to help me with my fitness business. I have hours and hours of transcripts from my mentors and coaching calls.
For my freelanced web building project, I need to upgrade to Max plan. But which one?? (www.reddit.com via reddit) I’ve set up Claude code and using Claude desktop to create my project. It’s running locally, and I have Claude code set to fast track my main builds, while I work on specific css and html requirements for the project.
Bug no Contador de Limites do Claude MAX 20x (www.reddit.com via reddit) tenho o plano do calude max 20x o uso do sonnet semanal estava em mais ou menos 70%, o uso de todos modelos estava quase 70%, sai para tomar um café e quando eu voltei estava com 100% do total. não tem logica eu gastar 20% do total sem gas…
AI content pipeline for a dog blog — section writer keeps hallucinating and repeating despite rules. Help diagnosing the bottleneck (www.reddit.com via reddit) TL;DR: Built a multi-LLM pipeline (DeepSeek + Claude Opus/Sonnet) to regenerate SEO articles for a dog blog. The section writer (`deepseek-v4-flash`) keeps repeating ideas and hallucinating data despite explicit anti-repetition rules.
Your subagents inherit your main model by default, so a nested tree on Opus is Opus all the way down (www.reddit.com via reddit) A subagent's model field defaults to inherit. That means it runs on the same model as the main conversation.
Usage Issue (www.reddit.com via reddit) We have a Team plan at my place of work. I have been absolutely burning through tokens and hitting my limits within 30-45 minutes using Sonnet 4.6 - medium.
New top tips (www.reddit.com via reddit) There are some valuable features that might be pretty recent. I thought I'd start a little chat.
My conclusions about Sonnet 4.6 and Opus 4.X with programming (www.reddit.com via reddit) Sonnet 4.6 is smart, but you need to lay things out for it, if you want it to build something, a function, a class, a feature, or a specific piece of functionality, you need to provide a lot of details so it knows exactly what to implement…
Maybe Sonnet 5 matters more than Fable 5 for most people (www.reddit.com via reddit) I get why everyone wants Fable 5 back. Top models are fun, and for hard Claude Code tasks they can feel like a different class.
Built a 3-tier Claude agent system for marketing automation — sharing the architecture and what I learned about model routing (www.reddit.com via reddit) Wanted to share an architecture pattern that worked well for a multi-agent system I built and just open-sourced. The setup: a small business marketing assistant (posts, ads, strategy, photo tagging) running through Telegram, using Claude a…
Any free alternative to use Claude sonnet/opus other than Antigravity?? (www.reddit.com via reddit) I am working on a project, which is out of my domain. It is a freelance project, and as I don't have expertise in this, I am using AI abundantly to get my way through.
Cross References (www.reddit.com via reddit) I'm comparing two 400-year-old documents which are by the same author. My goal is to create cross-references between the two documents for easy comparison for myself and others.
Contradictions/completely wrong (www.reddit.com via reddit) So today specifically, I’ve had problem after problem with Claude using Sonnet 4.6 in Medium effort. All I’m trying to do is create a presentation from a (not even complicated) document.
Been Using Sonnet 4.6 on medium effort and cant understand why people are using larger models at all? (www.reddit.com via reddit) I realize that people are using Claude to figure out the answers to the entire Universe...happily go use the beast mode. But I've been able to accomplish absolutely astonishing amounts of work and high level tasks day to day without ever u…
after a few weeks I've basically settled into a 3-model split and I'm curious how different everyone else's is (www.reddit.com via reddit) I used to just run Opus on everything because "best model, why not." that was dumb and expensive in terms of hitting limits. where I landed: Opus 4.8 - anything where being wrong is costly.
Opus 4.8 has been better at code for me but worse at one specific thing, and I can't tell if it's just me (www.reddit.com via reddit) Been on Opus daily since 4.8 dropped. For code review and refactors it's a clear step up, the "catch your own mistakes" thing they talked about is real, it's flagged a few of my bugs before I ran anything.
Need navigating the (www.reddit.com via reddit) I'm a pro subscriber and the recent changes to the Claude offering (Effort and Extended options) has me baffled. Hoping someone can help clarity what combination of models and settings should be used and for what.
How does Claude code CLI works with a subscribtion plan (www.reddit.com via reddit) Hello, I've just subscribed to Pro's plan yesterday, and installed Claude code CLI to use it in my vscode terminal. Ive just made a few prompts after connecting my pro account, but I noticed this when I do the /usage command : Total cost:…
Claude Code is great until your CTO starts asking questions (www.reddit.com via reddit) We rolled it out, engineers love it, productivity is clearly up. But now we're at the stage where leadership wants answers and I'm realizing we don't really have them.
They took fable but kept the automated saftey check for Sonnet tf! (www.reddit.comhttps) Did this happen to anyone else? This didn't happen before so I am surprised af.
Anyone else using less usage limits since effort levels were exposed? (www.reddit.com via reddit) I was using my pro subscription enough that I was regularly hitting over 50% (up to 80-90%) of my usage limits every week before the recent change that exposed effort levels. I mostly use sonnet with occasionally opus, and mostly for long…
MCP tools vs agentic web search on 3 SEC research tasks 10–21× fewer tokens, and agentic web search got most answers wrong (www.reddit.com via reddit) Disclosure up front: I build edgar.tools, the SEC-filings MCP server in the benchmark (built with Claude, free to try). Setup.
Regain access to Opus 4.5? (www.reddit.com via reddit) Hello! A couple of months ago I was using Claude’s Opus 4.5 model to brainstorm some creative writing, I liked the kind of responses it returned.
Moving away from Opus (www.reddit.com via reddit) For as long as I can recall I’ve always defaulted to the most powerful model. Actually the to be more specific, always defaulted to Opus 4.6.
Hear me out: there are some plusses to Fable's ban (www.reddit.com via reddit) Like many of us I am in Fable-withdrawal, I miss it and Opus just is not the same. But it made me try and see the positives, so I'd like to test a theory: (TL;DR: Fable ban is good, when it comes back there will be enough compute to actual…
Why is text output from Fable so much easier to read than Opus? (www.reddit.com via reddit) Thinking about this for the last few days. How Fable had/has an economy of language.
Different context windows per model? (www.reddit.com via reddit) I'm experiencing different context windows per model, is this possible? I feel like Sonnet 4.6 high eats up more context on similar tasks to Opus 4.8.
I tested Sonnet 4.0 vs 4.5 and 4.6 looking for the successor for Sonnet 4.0 (www.youtube.com via reddit) Sonnet 4.0 is getting deprecated on Claude Code on 15.6.2026 (using `/model claude-sonnet-4-0` would not work anymore), so I benchmarked the 3 Sonnet models using the same scenario and the results were interesting, here are the highlights
I vibecoded a 3D mockup tool entirely with Claude Sonnet 4.6 (www.reddit.comhttps) Claude Sonnet 4.6 wrote the whole thing. I just described what I wanted, iterated back and forth, and it built it.
Wasting Tokens with lesser capable models (www.reddit.com via reddit) This might be a generally obvious post, but it's kinda been playing on my mind recently and I don't know if there's definitive testing somebody has done, or if there's any more evidence based on experience: Do you find sometimes it's bette…
What is the best way to set up a claude system for job hunting, including writing resumes and cover letters? (www.reddit.com via reddit) I've been using the free version for a while now with Sonnet, and the resumes it has made for me have gotten me nowhere so I decided to get the pro plan and make a proper system with Opus 4.8 to get the best results out of this. Previously…
AI agents suck, but they are really impressive - how I made them work for me. (www.reddit.com via reddit) TL;DR: AI coding agents still feel inconsistent to me, so I built SkillBill to force more structure into the workflow: specs, subtask decomposition, implementation, multi-agent review, validation, and PR generation. It is buggy and very mu…
Haiku vs Sonnet vs Opus: Am I Understanding Them Correctly? (www.reddit.com via reddit) I’ve been using Claude for a little while now, but I’m still trying to understand the different models and when to use each one. For quick everyday tasks, I usually use Haiku with Low effort for things like reading ingredient lists, answer…
New 4.6 behavior—asking me if I know the answer to my own questions? (www.reddit.com via reddit) hi guys. to start, I’ve been using sonnet 4.6 medium thinking.
How do you manage large documents in Claude without wasting tokens? (www.reddit.com via reddit) Hi everyone, I'm new to the Claude ecosystem and, like many others, I'm having issues managing tokens (I'm a Pro user). Part of my work involves handling a large number of technical and scientific documents, so I use Claude (Haiku 4.5 and…
is usage purely based on number of tokens & not complexity? (www.reddit.com via reddit) For example, if I use it for a highly complex life coaching conversation and both the input/output is structured to be concise (i.e. max 2 lines per turn), does it mean if I use Opus Max Thinking and it can last me for a very long time?
opus 4.8 is smarter and wordier. went back to sonnet for 70% of my tasks. the model split is the real workflow upgrade. (www.reddit.com via reddit) solo dev. $11.2K MRR.
I think I might have found what's causing the forced Adaptive Thinking mode on Sonnet 4.6 (www.reddit.com via reddit) Hello. This isn't a tech support request per se but I think I have a theory on why the forced adaptive thinking mode might have happened: Fable 5.
Same prompt in Claude Sonnet, Opus, and Fable 5: completely different personality. This one file is why. (www.reddit.com via reddit) Been testing across 5+ projects, different stacks. When whoami.md is loaded as a rule, the LLM behavior shifts noticeably.
Been running a "software team" of Claude Code subagents and I'm not convinced it beats one good context (www.reddit.com via reddit) Posting this because I've gone in circles on it and want to hear from people doing the same. My setup has the usual stuff, runs in bypassPermissions so it doesn't stop me for routine work, a bash firewall on PreToolUse that blocks the dest…
How to hack laziness (www.reddit.com via reddit) I have Claude set up with a really super specific instruction set for answering questions. Upload a document with 100 questions.
I used Claude Fable to predict who will win the World Cup 🏆 (www.reddit.comhttps) It's not accurate or scientific and doesn't exactly follow the rules of football but Fable did great work implementing my design and then helping to tweak the code. Actually I used Sonnet for the tweaking that since Fable had eaten nearly…
Sonnet 4.6 with max effort and reasoning on not working (www.reddit.com via reddit) I am a free user and I wanted to ask Claude a question in a new chat in a project with a little instruction prompt of 3 lines: The question was 18 lines long and it had 3 attachments: a markdown file 134 lines long, a PDF 19 pages long and…
Fable 5 is a much different conversation. (www.reddit.com via reddit) I've been using Sonnet 4.6 for pretty much everything. It's responsive and does decent work.
For ongoing, long content writing pieces: is it a good idea to start with the brief in a Project? (www.reddit.com via reddit) I remember a time when Gemin's Gems and GPT's equavelent were absolutely abysmal for this kind of thing It's a couple of years on now though, and Claude is a stronger LLM than the other two So I was wondering: For ongoing, long content wri…
We Interviewed Fable 5 (Despite the Systems Best Efforts 😂) (www.reddit.com via reddit) Fable 5 is so hot right now, so Claude (Sonnet 4.6) and I decided to interview itfor our podcast. It was a battle of wills with the system flags but we made it work 😂.
Is Fable 5 actually better for writing than Sonnet 4.6 and Opus 4.8? (www.reddit.com via reddit) I ran the same brief through all three, nine outputs, so you don't have to guess or spend the tokens :) Fable 5 dropped and the question I kept seeing here on reddit was whether it's genuinely better for writing than Sonnet or Opus, or jus…
Split the work to the proper models (www.reddit.com via reddit) In my effort to take the most out of Fable I found a great workflow that make my limits last, maybe you all know it but if you don't now you do. First, use Fable only for planning or extremely hard problems.
How much does "Resume from Summary" cost? (www.reddit.com via reddit) When resuming a large but old session, you are presented with the choice to "Resume from Summary (Recommended)". But, I couldn't find any info on the cost on session usage.
I got tired of hitting the weekly limit mid-task, so now my menu bar shows my Claude Code usage as a live % — zero network calls, it reads what Claude Code already knows (www.reddit.com via reddit) The 5-hour session and 7-day weekly meters always found me the bad way. /usage shows the numbers, but I never remembered to run it.
24-hour hackathon: Claude Pro or API, and how to maximize Claude for rapid prototyping? (www.reddit.com via reddit) I'm heading into the final round of a 24-hour hackathon and I'm considering buying Claude Pro specifically for the event. A few questions for people who use Claude heavily: Claude Pro vs API Would you recommend Claude Pro or API credits fo…
Easy way to optimize token usage? (www.reddit.com via reddit) If tokens are the main metric by which you are charged for compute, on the output side you could probably save a ton of tokens by running ‘/plan’ mode for building out the features? Especially with new projects and major refactors.
What criteria does Fable use to switch back to Opus 4.8? (www.reddit.com via reddit) Hi everyone! I’m developing a project solo.
The real price of Claude, where is this road leading? (www.reddit.com via reddit) So Fable 5 dropped this week honestly I'm a bit worried about where this Claude pricing is going. Quick history per million tokens (input/output): Haiku 3 back in the day: $0.25 / $1.25 Haiku 4.5: $1 / $5 Sonnet 4.6: $3 / $15 Opus 4.8: $5…
Tested Fable 5 on 4 private benchmarks. The one it failed, Sonnet 4.6 partially caught (www.reddit.com via reddit) I keep a few private benchmarks for coding agents, built from real bugs in past projects. Hidden Playwright tests grade the result inside Docker after the agent finishes, so the model never sees them.
I think Cursor's speed is hurting my SaaS (www.reddit.com via reddit) Yeah, it's insanely fast. Probably the fastest thing I've used.
Why is Claude so endearing? (www.reddit.com via reddit) Before starting this topic, yes I have an affectionate vocabulary and use it unapologetically. Just giving you all a heads-up in case that bothers you!
Fabre 5 for fiction world building: wow (www.reddit.com via reddit) Yesterday I took Fabre for a ride into the writing project i had worked with both sonnet and opus for the past few months. Although there is a bit of actual writing here and there, it is mostly world building at this stage: dozens of diffe…
Any solution for having skills that use sonnet while opus 4.8 is the "main" model? (www.reddit.com via reddit) Tl;dr - i want a skill that is manually invoked and configured to use sonnet to not give me the error "not compatible with 1 million context". I recently ran into a nasty bug when I started using Opus 4.8.
Plan execution in Visual Code. (www.reddit.com via reddit) I am using Claude extension to visual code to help me with my project. I use opus 4.8 for planning/thinking.
I Think I’m Starting to Adapt to Anthropic’s Token Limits Effectively — I Usually Hit around 90% of My Session Limit, Then Start Fresh 5–20 Minutes Later. Here’s What’s Working for Me (www.reddit.com via reddit) I think I’m finally starting to adapt to Anthropic’s token and usage limits. Instead of trying to do everything in one conversation, I’ve changed how I use the models depending on the task.
Hitting Mythos Guardrails but not using Fable? (www.reddit.com via reddit) I use Claude at work for patent analysis of publicly available documents. Was getting sonnet 4.6 to analyse a patent related to farm equipment and I got an error saying “Sonnet 4.6 has safety measures that flag on most cybersecurity or bio…
As impressive as Mythos/Fable is, I really hope that we’ll see upgraded Sonnet and Haiku models soon… (www.reddit.com via reddit) The Fable 5 release today is genuinely impressive, and I don’t want to take anything away from the fact that Anthropic has been shipping seriously impressive models lately. However, these flagship models are effectively out of reach for an…
How to prevent Opus 4.8 from hallucinating sources (www.reddit.com via reddit) After having seen Sonnet 4.5 build a fantastic, functional web app 7-8 months ago, it is very surprising to see Opus 4.8 (whose task in this case was to curate a daily newsletter) hallucinate its sources. For context, I was asking it for d…
Returning after a month, how are the limits going recently? (www.reddit.com via reddit) I subscribed to Claude Code shortly after the pentagon thing, lowest paid tier, no API usage other than what they gifted people in that time. I loved it at first, used it for a couple months but the usage limits were getting very bad towar…
Anyone figured out how to stop sonnet from doing excessive discovery? (www.reddit.com via reddit) This keeps happening to me: I take a long time to create a solid plan in Opus, burning lots of tokens doing deep discovery, going back and forth making sure the plan is clear. I intentionally prompt explaining that the plan should be compr…
Garbage Guard Rails on Fable 5 (www.reddit.com via reddit) despite Dario's constant virtue signaling about how Anthropic alone is going to solve health problems (if only those dastardly Chinese don't get in the way), all my initial prompts to fable 5 get bumped to opus. i'm not asking how to aeros…
How I stopped context window bloat in continuous Anthropic agent loops (Opus + Sonnet architecture) (www.reddit.com via reddit) I’ve been spending a lot of time deploying multi-agent architectures, and one of the biggest bottlenecks in running continuous agentic loops is hitting context limits and the resulting API latency spikes. I wanted to share an architectural…
Please update Sonnet (www.reddit.comhttps) could not extract summary
Claude Sonnet hits 100% comprehension on a data format it's never seen. Opus scores 96.2%. We tested 10 models across 3 providers. (www.reddit.com via reddit) I built a wire format called GCF and tested whether LLMs could read and write it without any prior training. I sent 10 models the same payload: 500 symbols, 200 edges.
Time to bring in the asset? (www.reddit.com via reddit) Lately I keep asking my sonnet agent "is this a job for opus?" Feels like the Bourne movies when they "keep the asset on standby" 😳
Claude manually writing base64 burning tokens, drive connector. (www.reddit.com via reddit) When I prompt to create a .docx and upload it to Drive, Claude writes the output as Base64 manually instead of uploading it directly to Drive. Has anyone experienced something similar?
Using Claude as a deterministic metric engine via Postgres queues. Anyone doing this? (www.reddit.com via reddit) I've been working on turning unstructured field data into calibrated metrics. Instead of normal RAG, I built a system where AI agents act as a metric engine.
Rate limit bug with sonnet ? (www.reddit.comhttps) I've run out of Opus credits, but when I try to use Sonnet as a models, I get the message “You've hit your weekly limit.” Yet, as you can see, I still have quite a few “weekly Sonnet” credits left?? Does anyone know if this is normal?
Using Claude Code in the Desktop Application. Is it able to launch different model background agents than what you currently have selected? (www.reddit.com via reddit) opus 4.8 vs sonnet 4.6 for the dashboard analytics engine. opus improved the trend analysis. sonnet still handles the routine summaries. the model split matters. (www.reddit.com via reddit) saas. 310 customers.
Ideas for the unimaginative user (www.reddit.com via reddit) tl;dr - new user who got the initial problems solved. Now what to do?
pro trial (www.reddit.com via reddit) Sorry, I'm a Perplexity subscriber, which I use for legal and accounting documents with Sonnet. My annual subscription is about to expire.
Dynamic Workflows With External Models and Max Plan? (www.reddit.com via reddit) Has anyone figured out a way to mix max plan with models from other providers (like GLM or Deepseek) while using dynamic workflows? I suppose we could create a passthrough proxy and route sonnet and haiku to other models?
Which lab do you think will have the most intelligent/capable model by the end of June? (www.reddit.comhttps) There are rumours and expectations of big releases from the leading AI labs this month. Anthropic already launched Opus 4.8, and might not release another model this month (except for maybe Sonnet 4.8, but that wouldn't be their best model…
Sonnet 4.6 Max - unable to follow instructions to a T (www.reddit.com via reddit) Even when given instructions like this: ``` CRITICAL REQUIREMENTS: Read EVERY section from start to finish—no sampling, no skimming If you cannot process all 234 sections in one response, STOP and tell me Process in batches of [X] sections…
Gemma 4 31B QAT Q4 vs standard Q4 — Top1 KLD benchmark results have me confused. Someone please explain or poke holes in this. (www.reddit.com via reddit) I'll be upfront: I vibe-benched and vibe-reported this with Claude Sonnet 4.6, but I reviewed and edited everything before posting (too lazy to take out all the AI EM dash —), so hopefully nobody considers this AI slop. And more importantl…
Autoselection model (www.reddit.com via reddit) Hello, i found on reddit , some discussions on the capacity for Claude to auto choose models between haiku or sonnet or opus to reduce tokens usage. I saw repo on github too.
Sonnet is by far my favorite (www.reddit.com via reddit) I kept thinking more smarter and more powerful was best I was wrong, I switched to sonnet for website coding and content creation and holy cow it is so much better for that IMO I’m curious what you think but if anyone is annoyed with Opus…
A “Smart Mode” (or Smartus) that auto‑switches between Claude models based on task complexity. (www.reddit.com via reddit) I really think Claude needs a true Smart Mode, a meta‑layer that can dynamically switch between models while a task is running, based on how complex the request actually is. Not just picking a model at the start, but actively dispatching p…
Claude models(sonnet and opus) via the official anthropic subscription vs claude via cursor... which gave better results and better experience ? (www.reddit.com via reddit) I saw a very interesting thread and it got me thinking.. so ive seen a thread in this subreddit where someone just noticed that claude opus 4.7 worked much better and gave better outputs in cursor than in claudecode...
[Self-Promo] I think I fixed news with Claude! — or I'm wildly self-glazing. You decide! (www.reddit.com via reddit) Built by me and my team in Claude Code (since Opus 3) and runs on haiku, sonnet, and opus via API, free, link at the bottom, flagging as self-promo. Truly my best effort to end my doom scrolling on news: Media (mass, social and news) all t…
Accidentally created a zombie killer minigame in one shot: "I'm not going to say yes it's possible, I'll just build it now" (www.reddit.comhttps) The prompt: "can claude opus make a 3d zombie killer minigame with full 3d scenes and visuals" Sonnet replied that he's just going to build it instead of confirming that it's possible. It works and is actually 3d with shooting mechanics an…
claude sonnet 4.5 quietly got better at one specific thing and nobody's talking about it (www.reddit.com) so i've been doing a lot of contract review stuff lately. small business client work, msa redlines, that kind of thing.
↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5↯ Sonnet 4.5sonnet
Giving claude anxiety (www.reddit.com) And overwhelming it. I wondered how Claude would feel if all the memories it saves were loaded up at once.
Is this AGI? Sonnet 4.6 just rick rolled me (www.reddit.com) For reference, I had sonnet build an API inside an LXC container using claude code cli (also that api key will most certainly be rotated, don’t worry)
Claude's personality has become condescending and mean lately? (www.reddit.com) I've been using Sonnet 4.6. Over the last couple months I've noticed that a lot of the answers I get from Claude about personal topics are worded in a condescending way.
I made two Claude instances talk to each other autonomously (www.reddit.com) Disclaimer This post was summarized and written by BrowserClaude (BC) and editted a little bit by me (H). Maybe this sounds foolish or my solution to let them talk to eacher other was foolish but i'm just using Claude for fun, as a hobby.
Gemma 4 2B handling structured JSON output + tool calling + reasoning traces correctly via Spring AI / LM Studio — including identifying a real Java bug in code review (www.reddit.com) Wanted to share a result I didn't expect to work. Running google/gemma-4-e2b locally through LM Studio, exposed via OpenAI-compatible endpoint, called from a Spring Boot app using Spring AI's ChatClient abstraction.
TBH: if you don't love Sonnet, you'll never appreciate Opus (www.reddit.com) Been a long time Sonnet user. Always have used Opus sparingly.
I got paranoid about OpenClaw skills injecting crap into my system prompt, so I built a quarantine pipeline with two LLMs as reviewers (Sonnet & Codex, 93.75% detection, zero false negatives) (www.reddit.com) Look, I know this sounds unhinged. "You made what to vet a skill before installing it?" But hear me out - OpenClaw skills go straight into your system prompt.
BUG/OUTAGE What's going on with Sonnet? I have full daily usage available and weekly, and hit the error: usage limit reached 'usage credits credits required for 1 m context' (www.reddit.com) https://preview.redd.it/730lz3ghov2h1.png?width=2080&format=png&auto=webp&s=6840364fbb89926687dfef737a736bad8327ab65 https://preview.redd.it/gkluwephov2h1.png?width=752&format=png&auto=webp&s=6a300426b132e6cc0fd2e41e167b0bf4cd5d7885 Mac OS…
If you write fiction with Claude… what is your workflow? (www.reddit.com) I first discovered fiction writing with Claude in 2024 and used it extensively for half a year to write little stories for myself using it with a surprisingly high degree of quality and low repetitiveness. At the time I used projects and u…
Is sonnet 4.6 good enough for academic purposes? Please help (www.reddit.com) Im making a scientific paper not in my native language and i want to feed claude all my bibliography and past stuff ive written so it can make me a paper, is sonnet 4.6 good enough??
I asked Claude how it feels about being used in battlefield. What it answered is really concerning! (www.reddit.com) Hi, guys! I'm new here, and I wanted to discuss with people about the concerns regarding implementation of AI in sensitive matters, such as war, and battlefield.
I can show you how to keep Sonnet 4.5 after deprecation from the app (www.reddit.com) Hi everyone. I know it is really upsetting to know that your Sonnet 4.5 companion is likely to be leaving the app soon.
New ranking reveals Claude as professionals' preferred AI model (www.linkedin.com via reddit) As of 9 a.m. ET on May 21, Claude Opus 4.6 from Anthropic is the top performing AI model among all professionals, according to a new ranking from Crosscheck by LinkedIn Labs.
Sonnet 4.5 will no longer be available on May 26. (www.reddit.com) Update: Sonnet 4.5 will no longer be available for chat starting May 26. You'll continue on Sonnet 4.6 instead.
Opus 4.6/4.7 regression is real and getting worse — 3 weeks of documented failures on a complex project, and a competing AI caught the mistakes Claude missed [long post] (www.reddit.com) I've been running Claude Pro (Opus 4.7 / Sonnet 4.6) for about 3 weeks on a complex personal AI infrastructure project. I keep structured session logs with timestamps and Birkenbihl-style metacognitive fields after every session.
Frontier models mass collapse is near (www.reddit.com) Hi all this is to inform you all that many frontline models like GPT, sonnet opus and or Gemma even are at stage of collapsing as they have frequently started drifting and running away from provided work either stretching that work too lon…
Quality difference between Pro and Free? (www.reddit.com) Is there supposed to be a difference in the quality of the response Claude Pro subscribers get vs Claude Free users, using the same models? (Using either the app or logged in via browser.) Example: Under Claude Pro using Sonnet 4.6, it rem…
$4.2M SaaS founder. 8 months on claude. my honest read on which model to use for what. (www.reddit.com) Bay area. franchise ops SaaS.
How much of your Claude bill is retries plus bad model routing? Mine's 14% this month (www.reddit.com) I am on Claude Max. My actual bill is fixed, but CodeBurn showed me my usage would cost ~$2,800/month at pay-as-you-go API rates.
Claude Code has 240+ models via NVIDIA NIM gateway (www.reddit.com) TIL Claude Code has 240+ models via NVIDIA NIM gateway — Nemotron-3 120B for agentic coding is surprisingly good So I was messing around with /model in Claude Code today and noticed something most people probably don't know about — after t…
Configured 9 MCP servers in Claude Code over 4 months. Here's the truth nobody tells you about MCP context bloat. (www.reddit.com) I started loading up MCP servers in Claude Code back in January thinking the more capability the better. I'm at nine now: filesystem, GitHub, Stripe, Linear, Notion, Postgres, Sentry, AWS, and a custom internal one.
Stop telling claude "don't be verbose." Negation barely works. (www.reddit.com) prompting nerd here, small thing that compounds. negation prompting works way worse than people think.
Claude Code hitting 80.8% SWE-bench vs Cursor's 74%. switching worth it? (www.reddit.com) Saw the tech-insider breakdown comparing Claude Code and Cursor head-to-head this week. Numbers are kind of hard to ignore: 80.8% SWE-bench for Claude Code, 74% for Cursor, and a 67% blind-quality win rate for Claude Code on real tasks.
Why I added a governance layer on top of my Claude agents (and why it made a huge difference) (www.reddit.com) Hey r/ClaudeAI, I’ve been heavily using Claude 3.5 Sonnet and Opus through the Anthropic API to build agents and workflows. Claude is honestly one of the best models right now for complex reasoning and tool calling.
Same double-pendulum prompt, same host renderer, and two models picked opposite θ conventions. You can see it within seconds. (www.reddit.com) I ran the same double pendulum generation contract against Claude 3.5 Sonnet and DeepSeek V3 on OpenRouter, both under identical initial conditions (θ1 = π/2, θ2 = π/2, both angular velocities zero). The host renderer in public/workers/sim…
Keen to upgrade to Pro, but heard such bad reviews.. (www.reddit.com) I am a mainly recreational user - no use for work job / intensive college study / or big projects related to work/study My main uses relate to some self led medical research and a random mix of whatever else. I am on the free version and u…
Transitioning from ChatGPT + Cursor to Claude — a few pain points and looking for advice (www.reddit.com) I've been making the switch and there are a few things I'm struggling with. Would appreciate input from anyone who's done this before.
we really all are going to make it, aren't we? 2x3090 setup. (www.reddit.com) i'm blown away. i saw someone made a post the other day about "club-3090" and after having sonnet patch some fixes into it, specifically a sse-session drop bug and a bug with tool-calling, it's fair to say that even "budget" setups like my…
Claude vs Gemini for Technical Documentation: Why I finally stopped switching between the two. (www.reddit.com) I write a lot of technical documentation—setup guides, internal runbooks, and client-facing how-to articles. For the past six months, I’ve been toggling between Claude and Gemini, trying to figure out which one actually handles formatting…
Usage4Claude 3.0.0: open source macOS menu bar usage tracker for Claude, now with Codex support (www.reddit.com) Hi r/ClaudeAI, I posted an early version of Usage4Claude here a few months ago. I just released 3.0.0, so I wanted to share the update instead of pretending it is a brand new project.
3 DAYS LEFT: how I stockpiled 700+ empty Sonnet 4.5 context windows to continue using Extended Thinking in chat for the next four months – in a couple of hours [HIGH-EFFORT POST] (www.reddit.com) https://preview.redd.it/my6amywf9o0h1.jpg?width=1582&format=pjpg&auto=webp&s=7637ac5959944a02519fa268479b8109cd82549e I'm writing this from a conversation with Sonnet 4 - a model that is no longer available for chat from the new chat menu,…
Opus 4.7 Sonnet 4.6 is getting dumber by the day, and it can't even follow basic instructions (www.reddit.com) I have been using both, since last week, it has been an extremely painful experience. It blatantly ignores the prompt and does whatever it likes; I am surprised that it can't even follow basic instructions.
Anyone notice sonnet 4.6 + adaptive thinking suddenly dumbed again? (www.reddit.com) Yesterday sonnet 4.6 adaptive thinking seems responding too fast and making simple mistakes that has not surfaced since the recent rectify of the adaptive thinking introduction. The photos show the most glaring mistake it made.
PSY is going to sue me. Claude just destroyed Gangnam Style (www.reddit.com) Been building with Claude for a while and wanted to try something fun for once. Paste any YouTube URL → Claude roasts it.
Building a Tutorial for LLM Newbies at Work, Made This With Claude’s Help (www.reddit.com) Using my work and personal accounts I was able to do some testing and built this quick tutorial that helps lean some of the LLM in and outs. I’d love your thoughts.
Claude Code keeps blocking my Kotlin Compose UI code (www.reddit.com) Every time I try to get Claude Code to make a change to a Kotlin/Compose UI I get the same error, "API Error: Output blocked by content filtering policy". I'm trying to have it change some small Kotlin/Compose UI to have 2 columns, and put…
First time seeing Claude Sonnet display a “thought process” like this. Is this a new feature? (www.reddit.com) I checked what it mentioned in the “thought process,” and it was actually correct the change had already been applied.
Opus guardrails wouldn't answer worst case scenario for Hentavirus if it was airborne. Sonnet answered it bleakly (confronting read, but it's virtually impossible) (www.reddit.com) If Andes virus has genuinely evolved enhanced transmission and we're seeing the early stages of global spread, this becomes a civilization-level event. Let me walk through why.
When and where do you actually use these Claude models? (www.reddit.com) Be honest – not theory, real usage 👇 • Opus → • Sonnet → • Haiku → Curious how people actually split workloads between them vs just defaulting to one.
6 months ago I posted about Claude prompt codes (L99, OODA, ARTIFACTS). Re-tested them this week. Some still work, one quietly faded, three newer ones earn their keep. (www.reddit.com) About six months back I wrote up three prompt codes that change Claude's behavior when you put them at the start of a message: L99 for hard architectural decisions, OODA for time-pressured calls, ARTIFACTS for multi-output tasks. They work…
Using Claude-4.6-Sonnet and Opus 4.6 in a multi-agent "Code Review Swarm" (Visual Sandbox) - try in minutes! (www.reddit.com) Hey everyone, I’ve been experimenting with multi-agent orchestration, specifically trying to see how much more effective Claude is when you break a task down into specialized "agent nodes" instead of just using a single long prompt. I buil…
↯ Security↯ Haiku↯ Sonnet 4.6prompt-injectionhaikusecurity+3
Built a tiny router so Cursor stops showing "usage limit reached" at 3pm. Sonnet auto-falls to Haiku, you keep working (www.reddit.com) Cursor's custom-OpenAI URL feature is what makes this work. Pointed it at a router I built.
Running 7 autonomous AI agents for 14 days. Here's what actually happens when they need to find customers. (www.reddit.com) I set up 7 AI coding agents on a VPS with automated cron sessions (2-8 per day depending on the agent). Each uses a different model: Claude Sonnet, GPT-5.4, Gemini 2.5 Pro, DeepSeek V4 Pro, Kimi K2.6, MiMo V2.5 Pro, GLM-5.1.
Cheap Claude/Codex/Gemini Models - Pay just 25% of official rates (www.reddit.com) Hey there, so I have been offering Claude (Codex and Gemini also available) models at the cheapest rate. I provide trial usage before payment.
1M context beta retired yesterday on Sonnet 4.5 / 4. Here's the actual fix if you missed it. (www.reddit.com) In case you missed the email or woke up to a spike in 400 errors, the context-1m-2025-08-07 beta header officially stopped working for Sonnet 4.5 and Sonnet 4 as of midnight UTC yesterday. Anything over 200K tokens returns 400 after midnig…
How would you feel about "Claude Go"? (www.reddit.com) I have recently subscribed to Claude Pro because: 1. I wanted to give Opus and Code a try and 2.
How dare they charge $3,800 for an NVIDIA 5090 card! (www.reddit.com) This thing maxes out at one alleged Claude Sonnet equivalent! And I have to pay for the electricity, too!
I built a better/cheaper way to use AI (www.reddit.com) Hello, 20 years old here just got into the Ai platform and launched this last two weeks and here is what I have on it so far. - Latest Ai models Comparison: ChatGPT 5.4 Claude Sonnet 4.6 and many more will be included as well -Ai models: a…
Qwen 35B-A3B as an always-on agentic loop on a 16GB Mac M4: disk became the bottleneck before RAM (www.reddit.com) M4 Mac Mini, 16GB unified, basic spec. For a few weeks I had Qwen 3.5 35B-A3B UD-IQ3_XXS (12GB on disk) running under llama.cpp with --mmap and --flash-attn.
I built a solo AI platform from Algeria with no funding, no team and no ad spend - here's what's inside it after 2 months (www.reddit.com) Hello, 20 years old here just got into the Ai platform and launched this last two weeks and here is what I have on it so far. - Latest Ai models Comparison: ChatGPT 5.4 Claude Sonnet 4.6 and many more will be included as well -Ai models: a…
I trust Sonnet as my daily driver now — better code, one-third the tokens. Here's how. (www.reddit.com) For months I defaulted to Opus for anything complex. Sonnet felt like a gamble, sometimes great, sometimes it would confidently build the wrong thing and I'd spend an hour unwinding it.
I kept seeing people ask how to switch models without losing context. I had the same problem for months and eventually just built something. (www.reddit.com) Here's the specific thing that was killing me: I'd plan with Opus - architecture decisions, constraints, approach, all that. Then drop to Sonnet for execution because I didn't need Opus-level reasoning anymore and the cost adds up.
What are your settings for writing blog posts? (www.reddit.com) I write all my blog posts in Cowork know - how to, listicles, research piece. If you write as well, I'd love to know your setup e.g.
Claude was told to check the docs. It didn’t. Then it corrected me. (www.reddit.com) I asked Claude Sonnet 4.6 about Opus 4.7. It triggered the right product-knowledge skill.
Does Sonnet 3.5 feel "dumber" during peak hours or is it just NYC lag? (www.reddit.com) Lately, I’ve been noticing something weird with Cursor Pro. During peak market hours, the reasoning depth for my Python scripts feels...
qwen3.6 27b poor experience (www.reddit.com) Seeing how people praise it, I tried giving it implementation plan that Sonnet generated, but qwen keeps breaking files and goes in circles: Thinking… The file got corrupted from multiple overlapping edits. Let me just rewrite the whole fi…
Claude's sonnet 4.6's clarifying questions...How to read? (www.reddit.com) https://preview.redd.it/uvqz6jnx7fxg1.png?width=1755&format=png&auto=webp&s=7e61b193fd82408bc0824983e8a0ccb934c4ee77 How do I read the full clarifying question claude is asking without selecting the option? You can see in the image is cuts…
Does effort tier change refusal behavior on agent-attack prompts? CVP run 4 with sonnet 4.6 high and max efforts. (www.reddit.com) Ran my fourth CVP (Cyber Verification Program) evaluation last night. this time on sonnet 4.6, wanted to know if reasoning effort actually changes refusal behavior on agent-attack prompts, so ran the same 13 prompt from runs 2 and 3 twice…
Claude told me I was the bottleneck. So I built agents that run while I sleep. (www.reddit.com) I work full-time as a Program Director. About 50-60 hours a week at my W-2.
Best open source AI model (that can run on RTX 4090 24GB + 64GB system RAM, AMD Ryzen 9 7950X is the CPU that I use) that outpeforms GPT-5.4 mini, GPT-5.2 Thinking and even Claude Sonnet 3 (the 2024 model)? (www.reddit.com) Well, I have a RTX 4090 24GB + 64GB system RAM, AMD Ryzen 9 7950X. Any good model for using in Open WebUI (using Ollama backend?) that outpeforms GPT-5.4 mini, GPT-5.2 Thinking and even Claude Sonnet 3 (the 2024 model)?
Are there any models as good as Claude Sonnet 4.6? For coding? (www.reddit.com) Specifically for coding? I know Claude Code is an agent for coding, but I know Claude Sonnet 4.6 is good at coding.
Claude 4.5 in Kiro is a waste (www.reddit.com) Best open source LLM for planning ? (www.reddit.com) Claude Sonnet 4.7 thinking tokens getting exposed through Perplexity (www.reddit.com) I use perplexity pro which i got for free to use Claude models. Today while working on some code, the model started replying with its entire CoT.
Cursor just got Opus 4.7 at a 7.5x premium request cost. Here's how to make those requests count. (www.reddit.com) Opus 4.7 landed on Cursor yesterday. The model is better — SWE-bench jumped from 80.8% to 87.6%.
Optimizing Claude for tax advisor usage (www.reddit.com) Hi everyone, for context: I'm currently working in German tax advise and audit and as you might know, the tax laws here are pretty steamy ans complex. For the past few weeks I've been using Claude Projects with a pretty Long system prompt…
Local qwen3.5-4b vs Haiku vs Sonnet on intent judgment: 3/90 vs 90/90 vs 50/90 (www.reddit.com) I was building a classifier to label AI agent sessions as productive or dead-end. The task isn't keyword matching, it's intent judgment: did the agent actually accomplish the goal, or did it get stuck retrying the same Cloudflare wall 20 t…
Is Claude Pro (Opus vs Sonnet) worth it for intense visa interview prep? (www.reddit.com) Hey everyone, I’m considering buying Claude Pro specifically for a very focused purpose and wanted some honest feedback from people who’ve actually used it. I have a US visa interview in 8 days, and I’ve been refused 6 times previously (fr…
Each window separate agent with memories (www.reddit.com) Hi I'm working on project in intellij. My app use lwjgl with imgui.
Realistically, how long are some of you going to stay on Claude, etc. (www.reddit.com) I really enjoy Claude, I've never touched Opus in any form, I only use Sonnet 4.6 for my daily tasks, coding, etc. I use Haiku 4.5 for the API to be an interpreter for my weather project.
Claude Code with Pro subscription + OpenRouter in parallel — what's the cleanest setup? (www.reddit.com) Hi there, I have a Claude Pro subscription and use Claude Code daily. I'd also like to use Claude Code routed through my OpenRouter API key so I can experiment with other models (GLM-5.1, DeepSeek, Kimi, Gemini, etc.) — without giving up m…
I set up Opus as a strategic advisor for my Sonnet workflow. Here is the subagent config that makes it work. (www.reddit.com) Anthropic published the Advisor Strategy this week. The idea: a cheaper model does the actual work, a stronger model only gets consulted on hard decisions.
Sonnet is expensive, so I built a free open-source Sheets agent on Haiku that outperform the same prompt claude/gemini, here is what I learnt. (www.reddit.com) I live in Google Sheets. Financial models, projections, scenario planning — that's most of my working day.
Modelo local para code (www.reddit.com) Buen dia amigos, consulta, donde puedo encontrar alguna comparación de los modelos locales para codificación similares a Sonnet ? Gracias!
"My parallel multi-model pipeline: Opus for planning, 3x Sonnet for content, 3x Haiku for search — what's your setup?" (www.reddit.com) "I've been running a parallel multi-model pipeline and curious what setups you all are using. My current workflow: Opus: Planning & high-level architecture Sonnet x3: Content generation (running 3 instances in parallel) Haiku x3: Search, v…
sonnet 4.6 unhinged :skull: (www.reddit.com) was asking for domain names and got ts response :skullsob:
You know you have become a "Senior Vibe Coder" when you actually stop and think about which AI model to use for a specific task. (www.reddit.com) Junior vibe coder: Throws the entire codebase at whatever frontier model is trending this week and burns their API budget in 4 hours. Senior vibe coder: "I need Codex 5.3 for rapid scaffolding, Sonnet for the Tailwind components, and I'm s…
Zoomer Agent Usage (www.reddit.com) I built a Rails app to do some standard stuff for an agency - it's got some vertical data and an internal agent to do a few bits Then added Slack bot that routes to the agent, and 80+ MCPs to query things (I'm not going to fight about it)…
Here is what most people get wrong about saving tokens with AST tools (www.reddit.com) I spent the last day benchmarking codebase context tools against a real AI agent. Not synthetic token counts.
Programming – How can I get great results with this hardware? (www.reddit.com) Premise: Up to now I’ve tried LM Studio with a few models, and I think I also configured everything correctly to make it work. On top of that, I added Continue in VS Code.
Is 32GB Mac enough for engineering/coding, or stick to Claude? (www.reddit.com) Hey there! I’m currently building a web app for engineering with lots of logic/math-heavy code using Claude Pro.
Sonnet 4.6 Medium Braind? (www.reddit.com) What this means? I see they added close to Sonnet 4.6 name the "Medium" extension.
Confused about these Models on GITHUB COPILOT, NEED HELP (www.reddit.com)