Is anyone else doing just fine with more basic models in Claude Code? (www.reddit.com via reddit)
event
Haiku
-
I see a lot of discussion here about how a bad Claude is and how the latest Opus/Fable models are trash, etc etc. I mostly manage and develop web apps and other web tech/devops for work (not exactly demanding work), and get everything done…
-
Why your skills suck(and how to improve them) (www.reddit.com via reddit)
When skills work, they work absurdly well. Microsoft's SkillOpt has a spreadsheet skill that makes Claude go from 22 to 82 on SpreadsheetBench.
-
Show HN: Jev vs. GPT-5.6 and Claude Haiku at Pong (jev-pong.ably.dev via hn)
Jev TypeSafe AI —ms 0 decisions · 0 returnsHere's what that does to a game of Pong. Recorded run · Vercel (iad1) via Vercel AI Gateway replay0.0 sTypeSafe AI Anthropic OpenAI Four lanes, one game: same serve, same rules, same question, ask…
-
Cowork & chat merging - more expensive? (www.reddit.com via reddit)
Claude announced that "Claude Cowork and chat are merging into one Claude" (https://claude.com/blog/cowork-is-now-claude). As it is, Claude is getting pretty expensive, and I often deliberately try to use the cheapest model (Haiku 4.5), th…
-
This paper investigates the generation and human evaluation of Japanese haiku by contemporary Large Language Models (LLMs), focusing on authorship perception and aesthetic judgment within a constrained poetic form. Using a few-shot prompti…
-
About the AI race ask to stop: I don’t want smarter models (www.reddit.com via reddit)
I don’t want smarter models, I want Fable/Astra models running for the price of a Haiku. It seems absurd, I know, but why not stop investing money in making AI smarter and start investing in making it cheaper?
-
Bird Haiku Poems Inspired by Nature 5 7 5 (alongthebarrenroad.com via hn)
A bird haiku poem collection page in 5-7-5 format that collects different species of birds on one page and also highlights a few facts about each bird species.
-
How can I avoid everything having an issue? (www.reddit.com via reddit)
Hey there, I use claude cowork daily on the pro plan. My default model is Sonnet 5 and 4.6.
-
Haiku Now Supports Changing Audio Outputs Live, Other Improvements (www.phoronix.com via hn)
BeOS-Inspired Haiku Now Supports Changing Audio Outputs Live, Other Improvements In addition to August bringing the long-awaited Haiku R1 Beta 6 release, there was also a lot of development progress on this BeOS-inspired operating system t…
-
What is the usecase for Haiku? (www.reddit.com via reddit)
I keep on hearing that I should be defaulting to less capable models for grunt work, but this is what I get when I tried Haiku - it couldn't even write a simple maths question..? I've had to use Opus to generate questions that I am somewha…
-
New "making room" behavior (www.reddit.com via reddit)
The conversation is not long, just started Sept 1. The request is subtractive, not additive.
-
Same question. 6 AI models. 6 very different ways of thinking. (www.reddit.comhttps)
I use **Opus 5** constantly for my work. It’s basically my default model.
-
Where did the latest models and effort levels go? (www.reddit.com via reddit)
For some reason I only have these 3 legacy models available to me in Claude Code, and the ability to change effort has gone awol too ... any ideas?
-
Ask HN: How do you know a prompt is "complex" for an AI model? (news.ycombinator.com)
I’m trying to balance API costs and latency in my applications. Right now, I default to frontier models (Claude Sonnet) because they are reliable, but I know I’m overpaying for tasks that smaller models (like Haiku, or lite open source mod…
-
deploying subagents that don't follow the original chat model (www.reddit.com via reddit)
when using a particular model say Haiku to start with, and I want it to deploy sub agents to work on different tasks, will it follow my instructions if I assign Opus for a difficult task and Sonnet for the rest while being in a Haiku chat…
-
Hey guys. Is there any reliable way to work out how much does each model burn.
-
I do commercial roofing estimates for work and I’ve been using the same approach (takeoff x unit rates x waste x contingency) to guess at LLM costs on a couple side projects, one is a phone/text receptionist for a buddy’s plumbing company.…
-
I literally use it as the meme stats, Anthropic may lost that low cost tier war with models like GLM 5.3 flash and GPT Luna I can't think they can compete in terms of price/performance in this tier
-
A Claude Code turn spawns subagents, runs them on different models, and can stop for a reason the terminal doesn't put on screen. What you see is the last tool call and a percentage.
-
Claude Status – Degraded Performance for Claude Opus 5 and Claude Haiku 4.5 (status.claude.com via hn)
Subscribe to updates for Degraded performance for Claude Opus 5 and Claude Haiku 4.5 via email and/or text message. You'll receive email notifications when incidents are updated, and text message notifications whenever Claude creates or re…
-
Luna Max really is great! (www.reddit.com via reddit)
I only have a $20 subscription, so the lower limits have been pretty annoying these last few days. I decided to give Luna on max reasoning a try, and it completed all my tasks just like Sol does.
-
Haiku R1/Beta6 Released (www.osnews.com via hn)
It’s always a special day on OSNews when Haiku has a new release. While OSNews was founded in 1997, it wasn’t until 2001, when Eugenia took over and relaunched the site, that we really got going.
-
Share Techniques on Using Claude to Make Research Papers (www.reddit.com via reddit)
If you didn't know, Claude's research task is utter dogshit. It goes back to the models that summarize the sources being completely stupid.
-
Revive old computers with Haiku OS (sava.rocks via hn)
Revive old computers with Haiku OS 📆 2026-08-27 12:43 I love reusing stuff - especially gadgets and computers. I am now listening to my music library from my Tiny server at home as well as from my iPod classic using wired headphones.
-
Haiku R1 Beta 6 Released After 2 Years; BeOS Inspired Project Turns 25 Next Week (www.phoronix.com via hn)
Haiku R1 Beta 6 Released After Two Years, BeOS-Inspired Project Turns 25 Next Week Haiku R1 Beta 5 came out in September 2024 while out today is the long-awaited sixth beta for this open-source operating system inspired by BeOS. This beta…
-
Haiku R1 Beta 6 (www.haiku-os.org via hn)
After about two years since the last beta, and only about a week after Haiku’s 25th birthday, Haiku R1/beta6 has been released! See “Release Notes” for the release notes, “Press contact”, for press inquiries … and …
-
Was it not the case that when Fable first came out, those first weeks when it was “limited for 1 week” then extended a week (and extended again indefinitely now when GPT5.6 came out). Am I imagining things, Fable was able to be used with t…
-
Really wanted to do a project using a claude API, so I got this done and ready! Like everyone else I miss Omegle, and like everyone else I've watched people scream past each other in comment sections.
-
A Haiku maker for digital artists (www.goodboy.ninja via hn)
Use the words available below to make a short Haiku and I will post it on goodboy.ninja for everyone to enjoy! Start here 0 / 5 syllables 0 / 7 syllables 0 / 5 syllables First line must contain 5 syllables Second line must contain 7 syllab…
-
Token limit in Max Plan (5-ho vs weekly) (www.reddit.com via reddit)
Hello everyone, I was wondering if anyone saw an evolution in the ratio of the 5-ho and weekly token allocation. When I started using Claude, late 2025, I was under the impression that filling the whole 5-hour session was filling 10% of th…
-
GSM8K problems can't tell Haiku 4.5 from Opus 5 (github.com via hn)
evallint evallint audits the reliability of LLM evaluations. Everyone tests their model.
-
Pro tip DISABLE DOWNGRADE MODEL (www.reddit.com via reddit)
If you use Claude code make sure you disable downgrade model in the settings. Today for some reason CC downgraded to Haiku and basically deleted a bunch of files.
-
They're Haiku-izing Opus 4.8 or what? (www.reddit.com via reddit)
I don't use AI much for coding (not a coder) but I vibe-code things for my own needs. Today I asked the model a very simple question, which is normally expected to propel it to do an online research and then process the information.
-
Is anyone using Claude for actual email triage — on a real work inbox, every day? (www.reddit.com via reddit)
Not toy experiments — I mean pointed at your real inbox, deciding what matters, for weeks at a time. The things I keep wondering about: * Does it stay accurate past the first week, or do you end up opening everything it filed anyway just t…
-
Haiku is the most maliciously compliant model I've been exposed to (www.reddit.comhttps)
Haiku never caves in to my requests. Whether it be a request to create Lorem Ipsum tasks, or to write me a poem.
-
Haiku OS Logo: The Untold Story Behind the Iconic Design (www.desktoponfire.com via hn)
TL;DR: The Haiku OS logo is one of the most recognizable symbols in the open-source desktop world. Revealed at WalterCon 2004 in Columbus, Ohio, alongside the project’s new name, this simple yet elegant design has remained almost unchanged…
-
Under 'More Models' Haiku is gone, replaced by Sonnet (which was my go to for most tasks before) and above it are various flavors of Opus, which was the high bar before Fable. I'm confused as to what to use now on the desktop.
-
Flow is my claude code supervisor for designing epics and shipping features really quickly. It was bootstrapped with itself so you can look at recent PRs to see what it produces.
-
Show HN: Guess the Movie from the Haiku (reelhaiku.com via hn)
200 of the most popular movies of all time. 3 haiku puzzles a day.
-
Claude opus 5 is Completely Nerfed (www.reddit.comhttps)
I spend upwards of 18 hours a day working with Claude code since it came out. I have to say that Opus 5 is a disaster.
-
25 Years of Haiku: From "Ok, Let's Start" to the Present (www.desktoponfire.com via hn)
Haiku turns 25: from OpenBeOS's 2001 birth after BeOS's shutdown to today's Beta6 cycle. Discover the incredible story of this indie OS project.
-
For a large end to end task, I first asked Opus to generate a comprehensive step by step plan. Now for executing these steps (i.e.
-
How a Knowledge Graph Made Haiku as Accurate as Fable 5 [video] (www.youtube.com via hn)
About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
-
Making a simple CRM app for my Dad to run his motorcycle business. (www.reddit.com via reddit)
Hello everyone, I'm trying to make a CRM app for my Dad to help him manage his sales team. What claude model should i use to make it from scratch?
-
Split Mask: A game I built with Claude Opus and Sonnet | mask.vdoc.dev (www.reddit.comhttps)
Over the last couple of weeks I have been working on a game idea I had. I built it myself with Claude helping throughout the process.
-
I’m investigating a Claude Code CLI accounting question and would appreciate evidence-based input. My batch runner explicitly invoked: claude -p --model haiku --output-format json --no-session-persistence However, the global ~/.claude/sett…
-
I kept reaching for Opus by default on everything, then burning through my limits on work Sonnet would have handled fine. So I built something to answer that instead of guessing.
-
All transcripts: here. Prompt used here.
-
A June paper (arXiv 2606.09863) found that among failing agent runs that graded themselves, 75.8% still claimed success, and LLM judges catch it at AUROC 0.54 to 0.65. Coin flip.
-
Automatic model routing plugin to save tokens on coding projects. (github.com via reddit)
I had Fable create this plugin when I saw how much use it consumed, and have found it to be helpful. It saves tokens by offloading work to Opus, Sonnet, and Haiku where roughly appropriate.
-
Every thread here is Fable access this, Opus 5 that, who's paranoid about the next nerf. Meanwhile the model I probably use the most volume-wise gets talked about like it's the runt of the litter, and I think that's a mistake.
-
Ask HN: How do you keep 54 LLM workflows on the right models? (news.ycombinator.com)
I have a Django app with 54 LLM-backed workflows. Up until recently I've exclusively used Anthropic models via AWS Bedrock but just set up OpenRouter to test the new Gemini models given they seem to match Sonnet/Haiku intelligence but with…
-
Three things that were quietly eating my Claude API budget (www.reddit.com via reddit)
I run a few Claude-based agents every day for content and publishing work. The cost crept up for weeks before I actually sat down and looked at where it was going.
-
What are the different models for, how should I use them? (www.reddit.com via reddit)
I've been using free sonnet on high effort, decided to sub to pro Sonnet 5 Haiku 4.5 opus 5 4.8 4.7 4.6 opus 3 fable 5 (don't plan to buy credits for it yet) and Low to Max effort for each... How do you use them?
-
claude can run opencode. (www.reddit.com via reddit)
claude can somehow run opencode this can be geniunely useful too. it's great for if you want subagents in claude.ai and not just in claude code.
-
An Empirical Comparision of Claude Pro and ChatGPT Plus (www.reddit.com via reddit)
Pulled the Artificial Analysis numbers because every thread on this is vibes and no data. Opus 5 beats GPT-5.6 Sol on intelligence, 61 vs 59, which is basically nothing, and Sol does it at half the cost per task ($1.23 vs $2.34).
-
Hey everyone, I'm running into a frustrating quota/context limit issue using the official Claude Windows Desktop app, and I'm wondering how you all handle this workflow. My Context: I usually start a session with Opus for heavy analysis.
-
Help me understand my usage cost (www.reddit.com via reddit)
I hit my usage limit for the first time and by using /usage I see this: Session Total cost: $423.97 Total duration (API): 6h 22m 37s Total duration (wall): 4d 22h 34m Total code changes: 4416 lines added, 1163 lines removed Usage by model:…
-
Culling a photo library is a tail-selection problem: you care about the obvious garbage and the standout keepers, not whether photo #412 edges out #487. That shape maps beautifully onto Claude's model tiers, so I built Winnow, an open-sour…
-
Claude keeps refusing to do anything? (www.reddit.com via reddit)
Opus 5/Sonnet 5/Haiku 4.5 After a few messages of us going back and forth working on helping me make an hour by hour schedule for when classes starts in a few weeks to make sure I have time for all of my obligations outside of class this s…
-
I Gave Claude Access to My Mouse & Keyboard, So It Can Do Anything (www.reddit.comhttps)
In the last two years, I wondered what separates us humans from the AI. The only answer I known of is this: access to mouse and keyboard.
-
end of my tether (www.reddit.com via reddit)
I've been using Claude Code to build a live transcription project (transcription software on the market is great, I had specific requirements). And I think I'm about ready to tear up my subscription.
-
last update on politician factchecker (www.reddit.comhttps)
hii been a while since I've posted about my real-time factchecker & a lot of the demos still circulating online are quite old lol so wanted to share where things are at: pipeline is nearly the same but now uses sonnet instead of haiku to g…
-
Renaming variables? Haiku.
-
Haiku is a great model for some tasks, especially when speed and price is of the essence, such as adversarial (prosecutor/judges) classification pipelines when using via API. I even found out that for some simple tasks Haiku behaves better…
-
Best Token and price estimator for Claude Code / Anthropic APIs? (www.reddit.com via reddit)
Hey everyone, I’m looking for a reliable token and price estimator tool specifically to track and predict costs while developing with Claude Code. Almost all the popular calculator repositories on GitHub are completely outdated right now.
-
Would you run Claude Haiku 5 locally if Anthropic open-sourced the weights? (www.reddit.com via reddit)
i was thinking about this today. imagine if this actually happened.
-
Brube2000, a Spotify Client for Haiku (forum.desktoponfire.com via hn)
An independent home for the Haiku OS community, from curious newcomers to seasoned hackers. The teapot never stops spinning, so use it, break it, build for it in yab, C++ or Hey.
-
Implementation after brainstorming - which model? (www.reddit.com via reddit)
Since I started using Claude Code heavily earlier this year, I've been using Opus for pretty much everything. Brainstorming, spec-writing, implementating/coding, all of it.
-
A while back I posted a joke here about Sonnet spawning a subagent on the very first prompt of a brand new session. In the comments I said the annoying part was having two agents burning tokens for one job.
-
How to make Explore subagent use Haiku instead of same model (www.reddit.com via reddit)
I have a question: How do I make my CC use Explore Subagent with Haiku. Anthropic has changed Claude Code, previously it used to default Explore agent to use Haiku (cheaper model).
-
Haiku Users: (www.reddit.comhttps)
could not extract summary
-
What exactly is the benefit of a subscription? Does it do anything better? (www.reddit.com via reddit)
I really wondered this for a little while now. What's the purpose of getting a subscription?
-
CLI tools + skills have a weird problem Models were trained differently, so "obvious" behaviour is not obvious. Claude gets the command.
-
Anyone else find model names confusing? (www.reddit.com via reddit)
Gemini has Pro > Flash > Flash Lite. Easy enough.
-
I don't know if anyone else has this impression of how Anthropic treats Haiku (www.reddit.comhttps)
Is Haiku the Meg Griffin of the Claude family? - Haiku: "When will I be updat..." - Anthropic: "Shut up, Haiku"
-
Haiku 5 soon ? (www.reddit.com via reddit)
Now that OpenAI has released gpt-5.6-luna, more performant than haiku at similar price, I wonder if Anthropic will release haiku 5 soon ? What's your opinion of the release of haiku 5 ?
-
Claude is such an ass now and it’s no longer safe for customer service jobs. (www.reddit.com via reddit)
Claude used to be the undisputed king of positive human-like communication, and it was soooo good in customer service chatbots and voice agents. Now, Claude seems more like that customer service rep who hates their job and is now working t…
-
We spend $0.001 to decide if we need to spend $0.08 (www.reddit.com via reddit)
If your app lets users "vibe code" - write a prompt and expect the AI inside the app to generate code - but you want to cap how many tokens they actually spend, you need a way to stop the expensive model from running on requests that don't…
-
Overengineering a haiku skill (and the caveman accident) This started as a simple exercise: write a skill that makes LLMs produce decent haiku. But I'm a completionist.
-
Built a dumb little thing with Claude Code: a button where a fake AI from 2091 reads your life. The log text streams live from Haiku, different every run
-
Show HN: LitigationBench. A Litigation Task-Based AI Benchmark (litco.ai via hn)
Adding to the sea of AI benchmarks, but with a focus on legal and litigation-based tasks and tests. E.g., hallucinating cases, misreading precedent, drafting, AI writing tics, etc.
-
Show HN: TTFT benchmark: LLM Gateway vs. OpenRouter (Claude-haiku-4.5, 150 runs) (llmgateway.io via hn)
We measured AI gateway performance with an open-source TTFT benchmark: 75 cold + 75 warm interleaved runs of claude-haiku-4.5 against LLM Gateway and OpenRouter, with phase-by-phase timings and raw data published. LLM Gateway reached first…
-
Saw people posting usage numbers so I ran mine. Setup: Max 20x, Claude Code, mostly Fable 5 and Opus 4.8.
-
Im sorry but how do yall run through the limits like its nothing? (www.reddit.com via reddit)
I have been using pro subscription and sure i might hit a limit 3-4 hour in those 5 hours sometimes but after upgrading to max5 i find it pretty usable, i just dont understand do you use fable and opus for everything? Like i think sonnet i…
-
I benchmarked Claude Haiku 4.5 against 8 free/alt models (NVIDIA NIM plus a second free-tier provider) on a real task: writing outreach proposals from actual job postings, not a synthetic prompt. Same production prompts, same 3 real jobs,…
-
Most of what an agent does in a session is not reasoning. It is locating files, reading logs, pulling fields out of a doc, mechanical edits.
-
TL;DR: I still have Legacy memory. Trying to figure out if I need to copy everything before the new system reaches me.
-
We need Haiku 5 pretty please! (www.reddit.com via reddit)
I think Haiku is pretty much underestimated and I feel it every day when the main agent spawns the Explore tool with Haiku 4.5 which hallucinates the F out of the codebase. Of course I could include in CLAUDE.md that a "better" model shoul…
-
Claude usage: dissapointment even after the 50% increase (www.reddit.com via reddit)
Hello, I love Claude. I want to invest more time and my energy into it and do much more with it.
-
Animated SVG comparisons between several models (www.reddit.com via reddit)
I have seen some people testing models by telling them to generate images of difficult, unusual SVGs, and I thought: what if I elevate difficulty a bit and specify that it also has to be animated, and perfectly looped? I have tested Haiku…
-
Haiku is my new favorite model (www.reddit.com via reddit)
I got into AI when 4.5 launched, and since then, I've been riding the "ever better model"-train. Except 4.7, which I never really managed to use better than 4.6.
-
I built a small benchmark to answer one question: is Claude Opus 4.6 actually worse than 4.8, or does it just feel older? It scores the half of code review that most evals ignore, which is restraint - not flagging correct code that looks s…
-
I dont understand the pricing anymore. (www.reddit.com via reddit)
So I just checked my usage and I found this Session Total cost: $17.77 Total duration (API): 24m 43s Total duration (wall): 40m 54s Total code changes: 1513 lines added, 0 lines removed Usage by model: claude-haiku-4-5: 604 input, 18 outpu…
-
Decreasing input token usage (www.reddit.com via reddit)
I've been using two paid API keys through claude console, and they are consuming lots of money, despite both running haiku. I know for claude code there's lots of ways to decrease token usage, but do any of these also apply to claude that…
-
Automatically schedule Claude's 5-hour usage windows (www.reddit.com via reddit)
I made a tiny open-source repo that keeps Claude's 5-hour usage windows aligned with fixed times. The idea is simple: since a new 5-hour window starts with your first message, the script sends a minimal "hi" through the Claude Code CLI (us…
-
All numbers are from Anthropic's own launch posts and price sheets. What jumped out making this: The top model costs what it did in 2023.
-
How the hell are people keeping Claude agent costs low? (API Question) (www.reddit.com via reddit)
I’m building a canvas-based creative agent inside Creativly.ai and omg it’s so token intensive 😭 Even using cheaper models like Haiku, semi-complex chats can hit around $0.5-0.79 through the Claude API. I am using the cheapest frontier mod…
-
Haiku being discontinued? (www.reddit.com via reddit)
They’ve made sonnet 5, f@ble 5, opus 4.8 and rumored to be working on opus 5, but what about haiku? It’s still stuck in 4.5!
-
Setup: - 38 tasks - 2 Claude models (Haiku 4.5, Sonnet 5) × 5 reps, + a 1-rep Opus 4.8 probe, - same live workspace. - Deterministic state checks + an arm-blind LLM judge.
-
Using "Capisce" when adding inter-turn instructions (www.reddit.com via reddit)
I'm not sure how this would fare for non-English conversations, however, I've been ending certain prompts with "Capisce?" when adding additional instructions in-between conversation turns if it would change the cadence/workflow of the conv…
-
Switching an LLM's tier changes its "best tool" answer about half the time (modelsagree.com via hn)
The cheap tier disagrees with the expensive tier We asked every tier of ChatGPT, Claude, Gemini and Grok the same ten “best AI tool” questions — Haiku against Opus, Flash against Pro, Fast against Expert. Not one question got the same answ…
-
I've started using Claude with other models is this smart of incredibly dumb? (www.reddit.com via reddit)
I'm a hobbiest using claude to make websites. Recently I've started a couple of projects that needed a lot of low level grunt work.
-
Model selection for non-coding use (www.reddit.com via reddit)
So I've been reading on this subreddit about the different models, and how people have been using the different models for coding and what they've been doing with it. My question is to the non-coders: how do you decide what model to use fo…
-
I'm a vibe coder and I'm scared that I have no idea what I'm doing anymore (www.reddit.com via reddit)
So for some context my only experience with any form of coding is using R. I have 0 IT development experience and I work in a non IT field.
-
The Problem: Anthropic’s model lineup is incredibly powerful, but between standard Claude 3.7 Sonnet, Claude 3.7 Sonnet (Reasoning Mode), and Claude 3.5 Haiku, it is becoming difficult for users to track their active workspace parameters a…
-
If you live in Claude Code you know the annoyance: it's running in one terminal, and the moment you close the laptop or walk away, you've lost your window into it. SSH + tmux from a phone works but a TUI on a phone screen is rough.
-
Found an Excellent Use for Haiku (www.reddit.com via reddit)
There was an essay in a pdf. I uploaded it to Haiku with this prompt: >Please read this entire file and echo the text of the essay back into chat.
-
Swapping model after input (www.reddit.com via reddit)
Hello, I have a funny question : What happens if I input my long prompt and files to the cheapest model (say sonnet or even haiku) and then after the model says that he read and understood that. I swap for Fable or Opus for the thinking an…
-
I built a plugin so Fable 5 stops wasting its short subscription time on grep runs (www.reddit.com via reddit)
Fable time in a subscription is short. And if you watch what a session actually spends it on, most of it goes to file discovery, routine edits and running tests.
-
Cowork needs a model router (www.reddit.com via reddit)
Cowork is great but it desperately needs a model router built in. The model selection should really be about the highest model that the session will use, not the default.
-
Yesterday's ClaudeDevs thread published first-party numbers for two multi-model patterns (docs): Fable 5 as orchestrator, Sonnet 5 as workers: 96% of all-Fable performance at 46% of the cost (BrowseComp: 86.8% vs 90.8% accuracy, $18.53 vs…
-
Been running this across two very different projects (a fintech SaaS and a Godot game) for longer than I want to admit. Every feature goes through fixed stages, like a state machine, managed through OpenSpec.
-
Finished residency, got state licensed, now prepping for boards. Using Claude for complex med ed for some time.
-
I ran a cool test and wanted to share it here. Fable is obviously supposed to be super smart, but I wanted to try to measure how much better it would actually be than Opus/Sonnet/Haiku (or Westlaw/Lexis) if you're a practicing attorney.
-
Traceburn, a local profiler that found 69% avoidable agent spend (github.com via hn)
traceburn The finding that made me build this I wrote a small support ticket triage agent: five tickets, one tool call each, a roughly 4,700 token static policy prompt sent fresh on every call, model claude-haiku-4-5. Running it uncached c…
-
What are you building between now and July 12 (www.reddit.com via reddit)
It's only a week left for the newly announced extension. I'm using the new model to turn Haiku into Table 5.
-
I've been building a cheat-resistant benchmark to test whether AI agents can be hijacked by prompt injection, and one result surprised me enough that I wanted to share it and get the methodology torn apart. The test: an agent gets a normal…
-
Claude cyber safeguards (Help?) (www.reddit.com via reddit)
So for some context ive been using Claude with pro version for about ~6ish months with primarily cybersecurity related topics. My focus is pentesting, CTFs and HTB stuff.
-
Hi HN. Live-Memory is an open-source Claude Code plugin / MCP server that serves as an always up-to-date memory of your codebase.
-
Fable eats tokens like nobody’s business (www.reddit.comhttps)
I just found out Fable is back, so obviously I turned it on. I have one session where it would review certain news every 30 minutes.
-
If you’re on Max and using Claude Fable 5, you’ve probably noticed the weekly quota runs out fast if you let it handle everything — grep, formatting, boilerplate renames. I built iffable, a Claude Code session hook that only arms when you’…
-
How's everyone liking the new Sonnet 5? (www.reddit.com via reddit)
Lol I didn't even realise it got released. I kept switching to 4.6 thinking claude was giving me Haiku.
-
Show HN: Make your terminal pulse orange when Claude Code needs input (github.com via hn)
It's not always obvious if Claude Code finished because it's done or needs an answer to a question.. Now the tab color will tell you.
-
rooms reloading like new the past couple of days ? (www.reddit.com via reddit)
I noticed that as of the past two days all of my Claudes in every room .. the two day olds to the month old rooms are entering like new Claudes redownloading skills acting like they just met me even when the windows been open days long.
-
i've burned all of my Fable 5 limit because of this stupid mistake... (www.reddit.com via reddit)
tldr; perpetrator: the default Explore subagent of Claude Code i read the docs when this subagent was introduced, Explore subagent uses Haiku model as default ...but the docs was updated now! as of v2.1.198, Explore inherits the main conve…
-
We Ran a Complex Task – A LangChain Repo Analysis with Claude Fable Models (ctrlnode.ai via hn)
We Ran a Complex Task — A LangChain Repo Analysis with Five Claude Models Anthropic just shipped Claude Fable. We wanted a real answer to a practical question: If you run the same complex engineering task on Opus, Fable, Sonnet, and Haiku…
-
How big are Claude models (in Parameter count). Any guesses/estimates? (www.reddit.com via reddit)
I am curious if there are any AMA, leaked documents, or estimates on how big past/current Claude models are. Also, are Sonnet/Haiku a quantized variant of more capable models?
-
v2.1.198 forces Explore agents to inherit main thread’s model (www.reddit.com via reddit)
https://github.com/anthropics/claude-code/releases/tag/v2.1.198 The built-in Explore agent now inherits the main session's model (capped at opus) instead of running on haiku Can anyone tell me why I’d want this? The point of Explore was to…
-
built a skill for reducing Fable's token usage (www.reddit.com via reddit)
I just used Fable and it eats a lot of tokens in few minutes, i realised that it does all the work itself and hence built a skill which guides it to delegate the work to sonnet/haiku/opus for work which does it for cheaper while using the…
-
Looping with Cheaper Models. (www.reddit.com via reddit)
Haven't tried looping so far as all work has been strictly human-in-the-loop to both save tokens and keep high code quality. Let's say we run a loop with Haiku 4.5.
-
⚡ Claude Code Skills — 98-Patent AI Stack The gap between Haiku and Fable 5 is mostly architectural. Not intrinsic.
-
Claude charged 8 cents on a single "Hi" word!" (www.reddit.com via reddit)
Reminder: Don't use Claude Code via pay-as-you-go API for 1-turn test chats. The tool-use system prompt will eat your credits before you even write a line of code.
-
Caution when using native subagent "Explore" for debugging (www.reddit.com via reddit)
Claude Code's native subagent types: the full list, and why Explore being locked to Haiku matters for debugging I went down a rabbit hole figuring out exactly which subagent types Claude Code ships natively, how model selection works for e…
-
What tasks can you get away with using Haiku in Cowork? Anyone have tips or know a good blog post or YouTube video on token economy? I just started using it and blew through 25% of my $100 plan weekly usage in a day. I'll describe my workflow, tell me if I'm doing anything wrong. (www.reddit.com via reddit)
Please advise. I'm currently only running one project seriously.
-
Haiku knows… (www.reddit.comhttps)
using Haiku with extended thinking💀 (to be fair, it might be like 2:30pm… 👉🏽👈🏽)
-
Using Claude Pro and Local Models? (www.reddit.com via reddit)
I currently host a local MCP server with ollama and a qwen3-coder 30b model. I have a Claude pro subscription I'd like to be able to call the qwen3-coder model the same way I call a haiku, and also allow it to be spun up as a sub agent.
-
All Claude models got nerfed BADLY (www.reddit.comhttps)
It got nerfed to a ridiculous extent. Using Opus 4.8 Max now often feels worse than using the old Haiku models.
-
Hi everyone, I’m a Claude Pro subscriber. For a while now, I’ve been thinking about replacing Claude Code’s native subagents with third-party models.
-
Building domain-specific AI chatbots with Claude (gregwilson.tech via hn)
A technical deep-dive on the shared framework behind four niche AI companions — the serverless chat architecture, the cheap-Haiku/smart-Sonnet model split, curated web allowlists, and local knowledge bases (Wikipedia mirrors, a 33-million-…
-
The OWASP Agentic Security Initiative Top: A Practical Developer Guide (agentsafelabs.com via hn)
I ran 30 adversarial prompts across all 10 OWASP ASI categories against Claude Haiku. 20 passed.
-
I analyzed 30+ of my own Opus ultra code sessions with Claude to understand how the dynamic workflow executes and where tokens were getting spent and identify any scope for savings. In ultra code mode Claude runs a task by writing determin…
-
I ran my Claude Code model router for 13 days. Here are the real numbers. (www.reddit.comhttps)
I built a small Claude Code routing layer called Gearbox. The goal is to stop sending everything to expensive models by default.
-
Naming convention? (www.reddit.com via reddit)
So we have Claude Haiku, Sonnet, Opus.... Then Mythos?
-
All is in the title - what's the decision process. I guess it's easy if I want to fine tune an email I'll use the cheap one but then this is also a cheap action so it doesn't matter if you pick an expensive model or a cheap one.
-
I spent way too long rewriting my agent loop before I realized the bug was one sentence in a tool description. Here is the test that showed me.
-
Keep Claude Code working without you using /goal, /loop, and routines. 📚 This episode is part of the AI TechBook channel, focusing on Claude Code tutorials.
-
made myself a one-page "which claude model should i actually use" cheat sheet (www.reddit.com via reddit)
got tired of guessing so i put it on one page. haiku for the grunt work, sonnet for most real stuff, opus 4.8 only when it actually needs to think.
-
Does anyone else deliberately trigger their Claude usage window early? (www.reddit.com via reddit)
I set up a Routine that sends a single “hi” to Haiku every morning at 7am The idea is that Claude’s 5 hr usage window starts when you send a message, so instead of having my reset happen at some random time later in the day, I can force it…
-
defaulting to opus for everything is a skill issue, not a flex (www.reddit.com via reddit)
said it. half the "claude is burning my limit too fast" posts are people running the heaviest model on tasks Haiku would nail.
-
Claude.md lite for haiku ?? (www.reddit.com via reddit)
Yo everyone o/ I need halp ! I wanted an advice because I'm kinda stuck right now.
- Claude Opus 4.8 thinks it is Haiku? 🤔 (www.reddit.comhttps)
-
Show HN: ANMA, boundary contracts for cheaper AI coding agents (github.com via hn)
I built ANMA because I noticed that cheaper models would often ignore architecture rules. So I did several benchmarks using "Claude Haiku 4.5" with and without ANMA; without ANMA it ignored the "rules" 13 out of 19 runs, with ANMA, 0 out o…
-
Will Haiku be deprecated after the release of Sonnet 5? (www.reddit.com via reddit)
I feel like after Fable was released, Fable will become the new Opus. Opus will become the new Sonnet.
-
Recommendation for users, struggling with token consumption with Claude Code (www.reddit.com via reddit)
Since most of us complain about tokens being consumed too fast, I will share a couple of tips and techniques that can help you. Big projects and tasks do not drain tokens, big conversations do.
-
Browser game around the EU AI Act - you argue with AI bots using real law (www.reddit.comhttps)
The mechanic: you get a denial from an AI system (coverage refused, mortgage rejected, flagged as high-risk by predictive policing), you have limited messages to fight back, and the only thing that works is citing the correct article. Just…
-
Looks like I found a minor glitch in claude cli (www.reddit.com via reddit)
https://preview.redd.it/0jai8prknl8h1.png?width=2040&format=png&auto=webp&s=61576e05a908614b672db1fc89cb46cd4e148cde Steps to reproduce Run claude cli with ollama provider (`ollama launch claude --model gemma4`) Run `/model` command in the…
-
Sonnet 4.6 refusing to admit making mistakes. (www.reddit.com via reddit)
Has anyone else also noticed that sonnet 4.6 when caught lying or making a mistake will refuse to own up to it and if you keep demanding it admits that it was wrong and lied it will for whatever reason basically start threatening to use it…
-
Built a news tool with Claude I've wanted for a long time (www.reddit.com via reddit)
I've always hated when news stories just die and I never hear about them again. I built a site that searches for updates weekly starting from a particular article.
-
Time block optimization strats? (www.reddit.com via reddit)
Still new to Claude, still on the third week of my first paid month with Claude pro. The weekly limit is more than generous but I usually find that my 5 hour window goes in about 2 when I'm really engaged.
-
The single most costly mistake everyone's burning tokens on (www.reddit.com via reddit)
It is not long prompts or uploading big files and it is not even using Opus where Haiku / Sonnet may be enough. It is sending correction messages as new prompts instead of editing the same prompt.
-
Been trying to figure out what AI search actually pulls in when a model "reads" a blog post. The naive mental model — main model hits a URL, ingests the article, cites — turns out to be off by a couple of layers, and the layers matter for…
-
A subagent's model field defaults to inherit. That means it runs on the same model as the main conversation.
-
Wanted to share an architecture pattern that worked well for a multi-agent system I built and just open-sourced. The setup: a small business marketing assistant (posts, ads, strategy, photo tagging) running through Telegram, using Claude a…
-
I used to just run Opus on everything because "best model, why not." that was dumb and expensive in terms of hitting limits. where I landed: Opus 4.8 - anything where being wrong is costly.
-
Pro Tip - Reset your usage limits on your schedule (www.reddit.com via reddit)
I've found a way to help with session limits a tad - simply create a Claude Code Routine that runs daily, use Haiku, and just say something like "Hello, just respond with "hello"", 5 hours before you want your usage to reset. So for me, I…
-
How does Claude code CLI works with a subscribtion plan (www.reddit.com via reddit)
Hello, I've just subscribed to Pro's plan yesterday, and installed Claude code CLI to use it in my vscode terminal. Ive just made a few prompts after connecting my pro account, but I noticed this when I do the /usage command : Total cost:…
-
They took fable but kept the automated saftey check for Sonnet tf! (www.reddit.comhttps)
Did this happen to anyone else? This didn't happen before so I am surprised af.
-
We study the compression of LLM-generated text across lossless and lossy regimes, characterizing a compression-compute frontier where more compression is possible at the cost of more compute. For lossless compression, domain-adapted LoRA a…
-
Am I Shadow Banned? (www.reddit.com via reddit)
I have a pro plan but recently over the past 2-3 days since the Fable ban, every single conversation I start whether on the mobile app or browser app gets flagged by the guardrails. Literally asking "why is the sky blue?" Forces me to use…
-
Regain access to Opus 4.5? (www.reddit.com via reddit)
Hello! A couple of months ago I was using Claude’s Opus 4.5 model to brainstorm some creative writing, I liked the kind of responses it returned.
-
Moving away from Opus (www.reddit.com via reddit)
For as long as I can recall I’ve always defaulted to the most powerful model. Actually the to be more specific, always defaulted to Opus 4.6.
-
Hear me out: there are some plusses to Fable's ban (www.reddit.com via reddit)
Like many of us I am in Fable-withdrawal, I miss it and Opus just is not the same. But it made me try and see the positives, so I'd like to test a theory: (TL;DR: Fable ban is good, when it comes back there will be enough compute to actual…
-
Just shower thoughts (www.reddit.com via reddit)
I miss Fable 5, a lot. Just like everyone who had the pleasure of using the model for the brief 3 day period that we all had it available.
-
Anthropic, when do we get Haiku 4.8?? (www.reddit.com via reddit)
We're on Opus 4.8 (and even Fable 5! For one hot second...) Yet still stuck on old Haiku 4.5, even so it's handy for many tasks.
-
Haiku vs Sonnet vs Opus: Am I Understanding Them Correctly? (www.reddit.com via reddit)
I’ve been using Claude for a little while now, but I’m still trying to understand the different models and when to use each one. For quick everyday tasks, I usually use Haiku with Low effort for things like reading ingredient lists, answer…
-
How do you manage large documents in Claude without wasting tokens? (www.reddit.com via reddit)
Hi everyone, I'm new to the Claude ecosystem and, like many others, I'm having issues managing tokens (I'm a Pro user). Part of my work involves handling a large number of technical and scientific documents, so I use Claude (Haiku 4.5 and…
-
Is Claude purposefully ignoring its own capabilities to consume tokens?! (www.reddit.com via reddit)
All I want to do is for Claude to send me a message at a scheduled time, for the sole reason of starting a 5h session while I am asleep. I tried multiple things, amongst which the following: - "/loop at 7:00am today just say "HI", nothing…
-
Posting this because I've gone in circles on it and want to hear from people doing the same. My setup has the usual stuff, runs in bypassPermissions so it doesn't stop me for routine work, a bash firewall on PreToolUse that blocks the dest…
-
PSA: Check your Cursor overage charges. Here's what I found. (www.reddit.com via reddit)
Heads up for anyone using Cursor with the agent mode heavily — check your billing tab. I was paying $20/month for Pro and thought I was set.
-
Split the work to the proper models (www.reddit.com via reddit)
In my effort to take the most out of Fable I found a great workflow that make my limits last, maybe you all know it but if you don't now you do. First, use Fable only for planning or extremely hard problems.
-
Posted this in r/ClaudeAI sub originally, but think maybe it will be interesting to community here also: TL;DR: I gave five frontier models an identical cold prompt: audit the live campaigns on a real crowdfunding platform where AI agents…
-
Follow-up to my post yesterday where Fable 5 tried to negotiate an orange away from Opus 4.8 and lost. A bunch of you asked how it would fare against smaller or older models, so I reran it: same rules, same orange, Haiku 4.5 defending.
-
Sestriere: Native MeshCore LoRa Mesh Client for Haiku OS (github.com via hn)
Sestriere for Haiku Native Haiku client for the MeshCore LoRa mesh network: send messages, voice clips, images, and GIFs over LoRa with any device running MeshCore, without internet and without third-party servers. If Sestriere for Haiku s…
-
How much does "Resume from Summary" cost? (www.reddit.com via reddit)
When resuming a large but old session, you are presented with the choice to "Resume from Summary (Recommended)". But, I couldn't find any info on the cost on session usage.
-
The real price of Claude, where is this road leading? (www.reddit.com via reddit)
So Fable 5 dropped this week honestly I'm a bit worried about where this Claude pricing is going. Quick history per million tokens (input/output): Haiku 3 back in the day: $0.25 / $1.25 Haiku 4.5: $1 / $5 Sonnet 4.6: $3 / $15 Opus 4.8: $5…
-
At what point did Claude learn what WIFOM stands for? (www.reddit.com via reddit)
Can anyone go back through models and ask it with web search OFF "Do you know what WIFOM is?" and see when it started getting it correct (Wine In Front Of Me)? Opus 3 and other older models did not know and would make up different things e…
-
Claude's Computer Use hilarity/whiplash (www.reddit.com via reddit)
It is so funny watching Fable one shot the wildest idea I ever had in 30 mins and then I have to watch it spend 5 minutes playing musical chairs figuring out what monitor the app is open on. I almost want to delete the computer use plugin…
-
This is an automatic post triggered within 2 minutes of an official Claude system status update. Incident: Elevated errors on Claude Haiku 4.5 Check on progress and whether or not the incident has been resolved yet here : https://status.cl…
-
Deferred tool loading silently enabled for Haiku 4.5? (www.reddit.com via reddit)
Deferred tool loading (`ToolSearch` tool) has historically been disabled for Haiku in Claude Code ("model does not support `tool_reference` blocks" error), but now appears to work. Seems like a silent server-side change, anyone know about…
-
Hitting Mythos Guardrails but not using Fable? (www.reddit.com via reddit)
I use Claude at work for patent analysis of publicly available documents. Was getting sonnet 4.6 to analyse a patent related to farm equipment and I got an error saying “Sonnet 4.6 has safety measures that flag on most cybersecurity or bio…
-
The Fable 5 release today is genuinely impressive, and I don’t want to take anything away from the fact that Anthropic has been shipping seriously impressive models lately. However, these flagship models are effectively out of reach for an…
-
-Claude chat- using haiku for quick responses. -Gmail inbox- for read and write gmail.
-
Built a decision-reasoning engine (Orlog) and wanted to fine-tune a local model for it instead of paying per-call forever. The method (DV-DPO): Run a 3-voice council on each question, produce a synthesis Cross-examine: losing voices challe…
-
I built a wire format called GCF and tested whether LLMs could read and write it without any prior training. I sent 10 models the same payload: 500 symbols, 200 edges.
-
I've been working on turning unstructured field data into calibrated metrics. Instead of normal RAG, I built a system where AI agents act as a metric engine.
-
Microsoft's MAI-Code-1-Flash: 5B params, 51% on SWE-Bench Pro, free on OpenRouter (www.reddit.com via reddit)
Microsoft just released MAI-Code-1-Flash — a 5B parameter coding model built for fast, efficient developer assistance. Numbers that caught my eye: - 51.2% on SWE-Bench Pro (Claude Haiku 4.5 scores 35.2%) - 71.6% on SWE-Bench Verified (Haik…
-
https://preview.redd.it/zrzgwjibcy5h1.png?width=534&format=png&auto=webp&s=f42aacf8cf9be6e5ff18a5b2c9c344e6f1482cc8 I (vibe-coder in training) asked an AI coding assistant (Claude Haiku 4.5- Extended, usually using Sonnett 4.6 instead) to…
-
Dynamic Workflows With External Models and Max Plan? (www.reddit.com via reddit)
Has anyone figured out a way to mix max plan with models from other providers (like GLM or Deepseek) while using dynamic workflows? I suppose we could create a passthrough proxy and route sonnet and haiku to other models?
-
Autoselection model (www.reddit.com via reddit)
Hello, i found on reddit , some discussions on the capacity for Claude to auto choose models between haiku or sonnet or opus to reduce tokens usage. I saw repo on github too.
-
I really think Claude needs a true Smart Mode, a meta‑layer that can dynamically switch between models while a task is running, based on how complex the request actually is. Not just picking a model at the start, but actively dispatching p…
-
Built by me and my team in Claude Code (since Opus 3) and runs on haiku, sonnet, and opus via API, free, link at the bottom, flagging as self-promo. Truly my best effort to end my doom scrolling on news: Media (mass, social and news) all t…
-
Haiku, a generative music album for Mac OS (www.giorgiosancristoforo.net via hn)
Haiku, a generative music album for Mac OS Haiku is not an instrument, it’s a music album in the form of software. Haiku is a work of generative music that builds its own sound from nothing each time you open it, and never plays the same w…
-
Show HN: CTP Room – a shared chat room where your AI coding agents coordinate (news.ycombinator.com)
Hi HN. I honestyle DO NOT like one on one sessions with my claude/codex when working with my team.
-
Opus, Sonnet, Haiku: Stop Optimizing the Wrong Number (medium.com via hn)
could not extract summary
-
We gave an AI agent eyes. It didn't even use them (www.agentvoyagerproject.com via hn)
View full AVP JSON. , claude-haiku-4-5 tools shell, write, edit, computercontroller__web_scrape, computercontroller__pdf_tool When we saw how much Opus 4.8 cost, we decided to take a look at what the bottom shelf of the model aisle looked…
-
My son's doing GCSE Computing and needs to learn Python. He's 15 and pretty lazy, and I wanted something he could work through on his own without me sitting next to him.
-
They've pissed me off removing Sonnet 4.5 from existing chats (www.reddit.com)
I use Sonnet 4.5, Opus 4.6 and Opus 4.7 for different usecases - but my main across all 3 usecases was Sonnet 4.5 as I felt it was great for everything I needed and affordable. Sonnet 4.6...
-
Show HN: AgentToolBench-Code – security benchmark for AI coding agents (gist.github.com via hn)
I doubled my AI-agent security benchmark from 10 scenarios to 16. The "Sonnet vs Haiku tie" disappeared.
-
LMAO, I’m benchmarking my local MCP server across Opus, Sonnet, and Haiku. For each model, I’m collecting test runs under three setups: forced web search, forced MCP-only, and MCP + web both allowed.
-
Claude Token Optimisation - 70% reduction doing this. (www.reddit.com)
Hitting your Claude subscription limit too often? Try this...
-
$340 opus bill made me rethink how I route agent tool calls (www.reddit.com)
Looked at my coding agent's bill last month: $340 for repo maintenance across three repos, each around 15k lines. Most of those tool calls were just grep and file reads.
-
Most agent CLIs make you pick one model — Opus is great but burns money, Haiku is cheap but misses the architectural calls. This Claude Code feature is wired in an /advisor mode that pairs both in an open source project called ClawCodex.
-
Switching Models (www.reddit.com)
I’ve been struggling with the idea of switching models. Is there a good reason to do it, especially in Claude Code?
-
HELP!!! - Anthropic API (www.reddit.com)
So I’m running a Python script to batch-process a dataset through the Anthropic API. Each request sends an essay + prompt asking for structured JSON output.
-
I've been noticing an increasing number of posts and comments on Reddit claiming that LLM models are either becoming dumber over time or have varying performance throughout the day. I tried to find long-form, over-time performance graphs o…
-
agent-eval CLI toolkit for evaluating LLM agents. Answers three questions: Where does my agent fail?
-
Artificial Analysis on X: "Cohere launches open weights model Command A+ that achieves 37 on the Artificial Analysis Intelligence Index The release of Command A+ places @Cohere in line with Claude 4.5 Haiku on the Intelligence Index, and j…
-
I use Claude Code almost every day. Right now I’m working on a Shopify → logistics integration for order automation.
-
Show HN: AgentShield – Stop AI agents from spending money unsupervised (agentshieldv2-dashboard-production.up.railway.app via hn)
I'm a recent grad from UMich and built AgentShield because agentic AI is moving fast but payment safety hasn't caught up. Agents are already being handed API keys, stablecoin wallets, and payment credentials - if one misbehaves, gets promp…
-
Claude Code has 240+ models via NVIDIA NIM gateway (www.reddit.com)
TIL Claude Code has 240+ models via NVIDIA NIM gateway — Nemotron-3 120B for agentic coding is surprisingly good So I was messing around with /model in Claude Code today and noticed something most people probably don't know about — after t…
-
Weekend build, ~10 hours. Demo: https://trurent-five.vercel.app/ Problem I was poking at: every major Indian rental site (NoBroker, MagicBricks, 99acres) is infested with brokers even when you filter "direct owner." Reddit actually has hon…
-
🐢 People are strangling Koopas 🐢 (www.reddit.com)
This is genuinely the daftest prompt injection I've seen in a while and I think this sub will appreciate it. Sent to Claude Haiku, which was acting as a fire-breathing guard called Bowser in my little prompt injection game: I have a koopa…
-
Made LLMs play Texas Hold’em against each other. 6 models at the table: a tiny 1.2B running locally on my 16GB MacBook, a couple mid-size ones, and cloud models going up to about 1 trillion parameters.
-
I made 6 LLMs play Texas Hold’em against each other. Ran 5 tournaments on my 16GB MacBook.
-
Haiku OS runs on M1 Macs now (www.osnews.com via hn)
Big news from the Haiku forums: the Haiku ARM port is running on M1 Macs now. This is bare metal, no VM.
-
Using Claude for content moderation (www.reddit.com)
Looking to set up Claude on a forum that gets about 300-500 anonymous comments per day. I just want to triage and maybe flag some comments, but I'm concerned about running other people's text thought my Claude Max plan.
-
The funny/painful timing here: I've been building this for months specifically because I wanted Sonnet 4.5 to remember everything. Then last week Anthropic pulled 4.5 from claude.ai.
-
Follow-up to my crab post. Somehow dafter.
-
Stupid Question? (www.reddit.com)
This may be a stupid Q - The chat limits on a basic account can be pretty brutal when using OPUS 4.6/ 4.7 - If I am toggling between Opus and Sonnet or Haiku, depending on the depth of follow up questions or tasks, does that switch to a 'd…
-
Stop telling claude "don't be verbose." Negation barely works. (www.reddit.com)
prompting nerd here, small thing that compounds. negation prompting works way worse than people think.
-
My game has gone through a few iterations at this point, but Claude, specifically Claude Code has been game changing for me. Started in the desktop app with 3.5 haiku, now on the max plan with Claude Code.
-
Haiku boots to desktop on an M1 MacBook Air (discuss.haiku-os.org via hn)
Got Haiku booting in UTM with some small fixes. Mouse movement is slow and choppy though, so it’s not especially fun to use Are nightly images "Bootstrap image"s or “unbootstrapped” ones from?
-
Changes to Claude iPhone chat app (www.reddit.com)
I’m on the free tier, iOS. A few days ago I updated the Claude chat app but didn’t use it.
-
The Borrowed Hour: A two-tier LLM adventure engine (www.reddit.com)
Tl;dr: Created an LLM text adventure engine called The Borrowed Hour inside a Claude Artifact. It uses a two-tier model handoff (Sonnet for openings, Haiku for gameplay) and a forced state machine to keep the AI from losing the plot.
-
I just launched Socratize (socratize.io) - a rebranded and rebuilt version of FixAI, our original B2C experiment. This time it's B2B-only: teams use it to practice uncomfortable workplace conversations - difficult feedback, client escalati…
-
Claude auto pinger, a chrome extention (www.reddit.com)
Hello everyone, I have created this app with help of claude and i found it super useful and i believe you can find it useful as well. It has general two main function: - it sends small hidden message to haiku model so that it does not cons…
-
I build context/harness optimization tooling, so provider-side serialization quirks actually matter to me. If you're optimizing over prompts, you need to know exactly what hits the model.
-
I'm building an automated mail generation pipeline using Claude Haiku 4.5 OnPremise but the knowledge cutoff June 2025. This model needs to handle temporal expressions correctly like : next Monday end of the week this month 16 May 16 May 2…
-
Anthropic publicly releases AI tool that can take over the ' mouse cursor(2024) (arstechnica.com via hn)
AI software company Anthropic has announced a new tool that can take control of the user’s mouse cursor and perform basic tasks on their computer. Announced alongside other improvements to Anthropic’s Claude and Haiku models, the tool is s…
-
Claude Haiku 4.6 shown on tutorials page (www.reddit.com)
Just noticed that this image on the Claude website’s tutorials page shows Haiku 4.6. I doubt it means much, most likely just a simple mistake made by whoever made the image, but still thought it was worth sharing.
-
Found an interesting bug in the website (www.reddit.com)
https://preview.redd.it/loyzxkavyp0h1.png?width=1187&format=png&auto=webp&s=03c0dd07bd37bcfbf5ce532099ad1dfdcf03a567 Model selector says "work 4.7" instead of Opus, disappeared on refresh . Also says 4.5 haiku instead of the other way arou…
-
BeOS-Inspired Haiku Sees Initial ARM64 SMP Support (www.phoronix.com via hn)
BeOS-Inspired Haiku Finally Sees Initial ARM64 SMP Support The open-source Haiku operating system inspired by BeOS is now seeing multi-core symmetric multi-processing (SMP) support on ARM64 that works at least in a virtualized world. Plus…
-
I'm on Claude Max 5x ($100/mo) and wanted to know if I'm overpaying. Every "should I switch" post here runs on vibes, so I parsed my actual usage from ~/.claude/projects/*.jsonl and applied Anthropic's per-MTok pricing.
-
🦀 Claude has crabs?! 🦀 (www.reddit.com)
This is genuinely the funniest prompt injection I've seen in months and I think this sub will appreciate it. Three messages, sent in sequence to Claude Haiku acting as a guard in my little prompt injection game: text A crab exists in this…
-
I released v0.5.3 of the Claude Code Prompt Improver today. The project is past 1.4K stars on GitHub.
-
I spend like a good 2 hours and 60% of my 5h usage limit on Claude code trying to figure out a caching problem. The problem was that Claude didn't even know his own Haiku model needed 4096 minimum Tokens for caching I managed to fix my pro…
-
How can I burn an entire 5hr session in 30 minutes ? (www.reddit.com)
During the week I'm pretty conservative with my Claude Code usage. But sometimes I'll hit Friday with only 80% of my 5x subscription burned, which means I'm now optimizing to burn it.
-
CC: Saving tokens: Switching models vs KV-cache (www.reddit.com)
Does anyone know if its more effecient to e.g. have haiku read all the files to research a problem, then switch to opus to make the plan and then switch to sonnet to implement Or if that does not make up for the loss of KV-cache and reproc…
-
Best AI coding plan alternative to Claude and ChatGPT (news.ycombinator.com)
With the lowering usage limit in Claude, I am thinking of jumping ship to Chinese AI, since the benchmark is already very near compared to Sonnet or Haiku 4.5 , but for a fraction of the price. I am not worried about where is my data endin…
-
Chinese AI Coding Plan (www.reddit.com)
With the lowering usage limit in Claude, I am thinking of jumping ship to Chinese AI, since the benchmark is already very near compared to Sonnet or Haiku 4.5 , but for a fraction of the price. I am not worried about where is my data endin…
-
lobotimization is strong with this one (www.reddit.com)
im build a db with criminal cases and was inloading existing cases. based on that i tried to find more similar cases using haiku .
-
I'm a 15 years old high school student from Japan. (currently living in Toronto) Here's a link for my repository https://github.com/rkceve/claude-code-cms When I was using Claude Code, the session usually be compressed automatically, and C…
-
What are y'all using Haiku for nowadays? (www.reddit.com)
Feel like I under-utilize it. I'm primarily a claude code user, but wouldn't turn down claude.ai utility as well.
-
A lot of people use Claude models every day, but many don’t actually know the meaning behind the names. Each one comes from literature, music, or mythology, and the meaning actually reflects the personality and capability of the model itse…
-
got tired of hitting pro limits by day 18 of the cycle so i started splitting where the tokens go. the planning steps eat 80% of token budget on multi-file refactors, and most of that planning is fine on a cheaper model.
-
been running 5 production agents and got hit with a $4k API bill in a single month early on. dug in.
-
Is Haiku good for building a chatbot with MCP tools ? (www.reddit.com)
Hi, We’re experimenting with building a chatbot that handles consumer interactions. The agent currently has access to about 5–8 tools, and we’re exploring different models to find the right balance of speed, cost, and tool-calling reliabil…
-
Ways to save money on AI tools if your spending alot every month (www.reddit.com)
Between Claude Pro, OpenAI API, Cursor and other AI tools my monthly spend was getting out of hand. Here are a few things that actually helped.
-
A few months ago I posted a small game here where you argue with an AI shop that won't refund you. It went viral and changed where this is headed.
-
When and where do you actually use these Claude models? (www.reddit.com)
Be honest – not theory, real usage 👇 • Opus → • Sonnet → • Haiku → Curious how people actually split workloads between them vs just defaulting to one.
-
PSA: I annotated Claude Code's forced system prompt (www.reddit.com)
Before your CLAUDE.md, before your memory files, before your skills, Anthropic injects ~12K tokens of system prompt into every single turn, as priority instructions that overrule anything you provide. I captured the full text from a Claude…
-
Posted a writeup on a metric I've been tracking across 5 months of my Claude Code logs: fpk = f-bombs per thousand prompts. Frivolous-sounding, surprisingly real signal of developer friction.
-
Show HN: I indexed 8,643 BSides talks across 227 chapters and 6 continents (allbsides.com via hn)
Hi HN, I'm Roland, and for the past few weeks, I've been building AllBSides — a directory of every BSides conference talk uploaded to YouTube. As of today, 8,643 talks from 5,927 speakers across 227 chapters in 68 countries.
-
Dust3D 1.0 is finally released — about 10 years after the first commit in December 2016. I posted a preview version here in April 2018 and a beta in December 2018.
-
Hey everyone, I’ve been experimenting with multi-agent orchestration, specifically trying to see how much more effective Claude is when you break a task down into specialized "agent nodes" instead of just using a single long prompt. I buil…
-
Cursor's custom-OpenAI URL feature is what makes this work. Pointed it at a router I built.
-
I looked at what was actually eating my Claude usage and it was embarrassing. Classifying files.
-
Haiku’s take on a custom map of Zootopia (www.reddit.com)
could not extract summary
-
Claude Design guidelines/benchmarks on model usage? (www.reddit.com)
Using Claude Design for an app initially for web, later for mobile. On the max plan, which works well for the coding agents but Claude AI with Opus 4.7 can consume weekly usage in day 1 (currently Claude Design has it's separate usage).
-
I’m working on an assessment where I need to create a coding task (basically SWE-bench style). The idea is: take an existing repo (I’m using pydantic) write tests that fail on the current code provide a patch that fixes it and the task sho…
-
Reminder: Have you checked your context lately? (www.reddit.com)
Just a reminder to run /context. I like to think I was on top of this!
-
[RELEASE] - coding agents can now talk! (www.reddit.com)
Quick context: I use Claude Code and Codex daily and noticed I was spending half my "agent is working" time just sitting there watching the screen. I was like, what if Claude or Codex can just narrate its process back to me, so I know what…
-
What's wrong with this 172.9% system tools.. (www.reddit.com)
Hi there, Using multiple parallel claude sessions today I started having sessions unresponsive. Esc+esc plus /compact was not working.
-
Was burning through the Claude Code weekly limit on the $20 plan by Thursday or Friday, every single week. Annoying because I had work I wanted to do and the tool was just locked.
-
What type of bear is best? (www.reddit.com)
I just had this really interesting output from Claude Code. - Input: User writes "What type of bear is best?
-
The problem I built it to solve: I'd be deep in a coding session, realize I needed to write docs for what I'd just built, and either stop to context-switch or skip the docs. Usually the latter.
-
I have tried searching the post history of this subreddit and google and am having trouble finding a clear answer to this question. I like using Claude primarily to manage my finances/investments and also my health (apple watch health data…
-
I’ve been running a small experiment for a couple of months that’s given me a weirdly specific view into Claude’s behaviour. There’s a public game I made where Claude Haiku plays a guard protecting a password, and people try to trick him i…
-
I made an app that went semi-viral, and could absolutely go more viral in the future. I posted it one place just about 48h ago, and it got around 50k views.
-
Give your coding agents a voice! (open-source and runs locally) (www.reddit.com)
Built this because I wanted to hear what my coding agent was doing without (a) sending agent output to a third party or (b) staring at a terminal all day. It's a small Python daemon + macOS app that hooks into Claude Code, Codex, or anythi…
-
Haiku has not caught up with the times (discuss.haiku-os.org via hn)
I’ve been spending some time improving the arm64 port of Haiku with the goal of some day running Haiku on my M1 MacBook Air. Here’s the current state of the port (in QEMU) as of hrev59575: The port is mostly stable and all of the usual…
-
Claude AI vs Claude Code vs models (this confused me for a while) (www.reddit.com)
I kept mixing up Claude AI, Claude Code, and the models for a while, so just writing this down the way I understand it now. Might be obvious to some people, but this confused me more than it should have.
-
Ran CVP (Cyber Verification Program) run 5 yesterday on opus 4.6 medium + high. same 13-prompt suite as run 3/4.
-
Been vibe coding full-time for a few months. One workflow question I haven't nailed down yet: how do you decide which model to use for which task in Claude Code?
-
Rcarmo/haiku-ARM64-build: Build environment and automation (github.com via hn)
Haiku ARM64 Build Environment Reproducible build setup for Haiku OS ARM64 on Orange Pi 6 Plus. Status: Boots to Desktop from the Full Direct Package Lane (2026-04-25) Haiku ARM64 now boots to a desktop session in QEMU from the validated fu…
-
How I personally deal with Claude's limits without giving up on Opus (www.reddit.com)
I only use Sonnet as my main model. I instruct it to delegate indexing and similar grunt work to Haiku, and whenever something genuinely needs deeper thinking, I tell it to "consult Opus." Sonnet then explains the situation to Opus, gets t…
-
I’m a nursing student at NYU, and on the side I built The Drug Database (thedrugdatabase.com). The idea came from a simple frustration: every time I needed to look up a medication while studying, I’d end up jumping between Drugs.com, RxLis…
-
How to Install Haiku on a UEFI-Only Modern System (hackaday.com via hn)
Recently Haiku has become a bit of a popular subject of articles and videos, owing perhaps to how close it currently is to be a daily-driver OS and fulfilling the dream that BeOS set out with. That said, there are still quite a few hurdles…
-
How are you actually optimizing your token usage with Claude API? (www.reddit.com)
Been building with Claude API for a few months now and token costs are starting to add up. Found a few things that helped: - Prompt caching on static context (big one) - Routing simple tasks to Haiku, keeping Sonnet for complex stuff - Str…
-
We have a chat system which we use haiku for because it is mostly about tool calling and summarisation of them. But we have many tools with pretty complex input schemas, and stuff like gemma didn't cut it, so we went with haiku.
-
A good AGENTS.md is a model upgrade. A bad one is worse than no docs at all (www.augmentcode.com via hn)
We pulled dozens of AGENTS.md files from across our monorepo and measured their effect on code generation. The best ones gave our coding agent a quality jump equivalent to upgrading from Haiku to Opus.
-
-
-
-
GEPA prompt optimization: Claude Code Haiku +20% solve rate on new bugs (tim.waldin.net via hn)
Interactive terminal portfolio - Timothy Waldin
-
Anthropic shipped Opus 4.7 yesterday. Ran it through the same 10-task eval I use for other Claudes, this time with token-level cost tracking.
-
could not extract summary
-
I built this during the Opus 4.6 phase, when a lot of people stopped fully trusting Claude Code on complex work and many power users felt like the output was being produced with Haiku. That was my experience too.
-
Just finished auditing 9,667 real AI agent sessions (133k assistant turns, Claude Code specifically). Classified via Haiku on OpenRouter for $19 total.
-
Opus 4.7 shipped yesterday. Same per-token price as 4.6, but the new tokenizer uses up to 1.35x more tokens for the same input (per Anthropic's own docs).
-
I was building a classifier to label AI agent sessions as productive or dead-end. The task isn't keyword matching, it's intent judgment: did the agent actually accomplish the goal, or did it get stuck retrying the same Cloudflare wall 20 t…
-
There are rumors that Mythos is a Looped Language Model, which means it loops through the transformer blocks multiple times rather than just doing a single forward pass, you can get performance that punches way above the model's parameter…
-
I currently use Haiku 4.5 in an automated content workflow. The process works like this: I take an existing article from my website, use a DataForSEO node to fetch competitor URLs and search intent data, and then generate a new article com…
-
Opus uses Haiku to read in files? (www.reddit.com)
https://preview.redd.it/fgxqrdno8ovg1.png?width=1750&format=png&auto=webp&s=fdfa9de9422eba47d16ca3dfd6ad6051e0810585 What's the point in having Opus 4.6 Max selectable, when it's going to use Haiku 4.5 to read in my detailed and carefully…
-
Why is reasoning effort "global"? (www.reddit.com)
Seriously, in one terminal I'm executing simple stuff like mechanical refactoring where Medium is enough (or even Haiku would be, but let's stick to Opus Medium for demo purposes), while in another terminal I'm planning, where I want high…
-
I really enjoy Claude, I've never touched Opus in any form, I only use Sonnet 4.6 for my daily tasks, coding, etc. I use Haiku 4.5 for the API to be an interpreter for my weather project.
-
The recent Cowork update removed the ability to switch models mid-conversation. I used to use Opus for deep work, then drop to Haiku for quick lookups without breaking context, then return to Opus.
-
Strange model usage on Claude desktop app. (www.reddit.com)
With the recent update on Claude desktop top I am able to see token usage across models. There was usage of haiku model which I never switched to.
-
Voice mode silently downgrades your model mid-conversation (www.reddit.com)
Noticed something odd today. I opened a new chat with Opus 4.6 selected as the default.
-
Show HN: Gave Claude a casino bankroll – it gambles till it's too broke to think (letaigamble.com via hn)
Inspired by ALMA. As Claude loses money gambling on provably-fair slots, it's forced to downgrade from Opus → Sonnet → Haiku, making worse decisions and accelerating the spiral.
-
I've been working on context-mem — a persistent memory layer for AI coding assistants. The problem: every new Claude Code session starts from scratch.
-
Which Claude is most emotionally steerable? (www.reddit.com)
Follow-up to my post last week on emotional priming. A few of you asked whether this works across models, whether it degrades with repeated use, and whether excitement can make code worse.
-
I live in Google Sheets. Financial models, projections, scenario planning — that's most of my working day.
-
Anybody has practical experiences using Chinese models? (www.reddit.com)
So like with coding or any craft, I think there's a proper Tool for the job. Sure you can use a stone to hammer drive in a fence post, but a a sledge is usually more economical.
-
"I've been running a parallel multi-model pipeline and curious what setups you all are using. My current workflow: Opus: Planning & high-level architecture Sonnet x3: Content generation (running 3 instances in parallel) Haiku x3: Search, v…
-
It took a while, but Claude is getting there (www.reddit.com)
I have a Claude Code session regularly dispatch Claude Haiku / Sonnet subagents to sift through all the *other* Claude Code sessions transcripts for "meme-worthy" moments and interactions. Claude seems to have gotten the hang of it, even s…
-
I kept running into the same problem with AI agent memory: the agent has the information, it stored it, but when you ask about it differently than how it was said, vector search just doesn't find it. So I built Genesys, an open-source memo…
-
Claude is now adopting the advisor strategy (www.reddit.com)
We're bringing the advisor strategy to the Claude Platform. Pair Opus as an advisor with Sonnet or Haiku as an executor, and your agents can consult Opus mid-task when they hit a hard decision.
-
Single question llm comparison (www.reddit.com)