model roundup

Sonnet 4.6

54 items · started 2026-06-07 · ongoing (last activity 2026-06-26)

  1. https://preview.redd.it/09w13qlnll9h1.png?width=1576&format=png&auto=webp&s=d8103d300143e784d1cdfdd0bfe914b1f108629f I'm using Sonnet 4.6 / medium - same as always. I use Claude because I much prefer the direct, less annoying style, compar…

  2. I figured the agent would be the tough part. Turned out the cost was the real story, and that's what closed the deal.

  3. Note: I'm not sure if this is a Claude or Claude Code issue. I'm fairly new to using Claude/Claude Code.

  4. It’s not about the actual words. SEE MAJOR UPDATE AT BOTTOM TL;DR: It’s not the meaning, it’s not even “unsafe words”, it’s COHERENCE.

  5. There’s great anticipation for Sonnet 5 next week, so putting my thoughts here before it gets released, to see if I’m right. I believe that Sonnet 5 will be mediocre, in a way that would make you prefer opus over it.

  6. I asked Claude Chat how to scrape the help pages for a software I use into local markdown files so that I can ask questions about the product documentation without having to waste time searching and reading through things that might not be…

  7. I've learned a lot from this group about methods of increasing the efficiency of using my Claude Tokens. But I wanted to share my process to see if I can improve on it at all.

  8. Which option gives the most actual Opus 4.8 usage volume: Kiro Pro, Claude Pro or something else? My monthly budget is $30.

  9. Core is an HTML, CSS, and JavaScript browser game built through iterative vibe prompting with Claude Sonnet 4.6 free tier. Core is a turn based, chess-like tactical grid game.

  10. I run a web design business where I create websites using ai, mainly claude. claude handles most of the actual work, while I focus on the marketing and sales side.

  11. https://preview.redd.it/oej0cgk3pf8h1.png?width=845&format=png&auto=webp&s=a2b2ba3d6a37ca239244ea9f4becf7fbe689b0b8 Since When did they start serving Sonnet 4.6 model with 1M context window

  12. We benchmarked GLM 5.2, MiniMax M3, Kimi K2.7-code, Qwen 3.7-Plus and Sonnet 4.6 across nearly 1,000 coding-agent scenarios. The scenarios were run twice.

  13. Has anyone else also noticed that sonnet 4.6 when caught lying or making a mistake will refuse to own up to it and if you keep demanding it admits that it was wrong and lied it will for whatever reason basically start threatening to use it…

  14. I wanted to see what each frontier lab model would do when put into a prisoner’s dilemma with each other. This is not so much a comparison as much as it is a thought experiment.

  15. I've been trying to use Claude to build a coach to help me with my fitness business. I have hours and hours of transcripts from my mentors and coaching calls.

  16. We have a Team plan at my place of work. I have been absolutely burning through tokens and hitting my limits within 30-45 minutes using Sonnet 4.6 - medium.

  17. Good day, people! Hope you are fine.

  18. Sonnet 4.6 is smart, but you need to lay things out for it, if you want it to build something, a function, a class, a feature, or a specific piece of functionality, you need to provide a lot of details so it knows exactly what to implement…

  19. Hello, lately I've been noticing Sonnet 4.6 often saying things like 'yeah lol' or 'lol' at the start of its messages, has anyone else noticed it? Could Anthropic have done Discord dumps for its training data?

  20. I am working on a project, which is out of my domain. It is a freelance project, and as I don't have expertise in this, I am using AI abundantly to get my way through.

  21. Subscribe to updates for Elevated errors on Claude Sonnet 4.6 and Opus 4.8 via email and/or text message. You'll receive email notifications when incidents are updated, and text message notifications whenever Claude creates or resolves an…

  22. So today specifically, I’ve had problem after problem with Claude using Sonnet 4.6 in Medium effort. All I’m trying to do is create a presentation from a (not even complicated) document.

  23. I realize that people are using Claude to figure out the answers to the entire Universe...happily go use the beast mode. But I've been able to accomplish absolutely astonishing amounts of work and high level tasks day to day without ever u…

  24. I used to just run Opus on everything because "best model, why not." that was dumb and expensive in terms of hitting limits. where I landed: Opus 4.8 - anything where being wrong is costly.

  25. Hello, I've just subscribed to Pro's plan yesterday, and installed Claude code CLI to use it in my vscode terminal. Ive just made a few prompts after connecting my pro account, but I noticed this when I do the /usage command : Total cost:…

  26. Disclosure up front: I build edgar.tools, the SEC-filings MCP server in the benchmark (built with Claude, free to try). Setup.

  27. Hello! A couple of months ago I was using Claude’s Opus 4.5 model to brainstorm some creative writing, I liked the kind of responses it returned.

  28. I'm experiencing different context windows per model, is this possible? I feel like Sonnet 4.6 high eats up more context on similar tasks to Opus 4.8.

  29. Claude Sonnet 4.6 wrote the whole thing. I just described what I wanted, iterated back and forth, and it built it.

  30. I’ve been using Claude for a little while now, but I’m still trying to understand the different models and when to use each one. For quick everyday tasks, I usually use Haiku with Low effort for things like reading ingredient lists, answer…

  31. hi guys. to start, I’ve been using sonnet 4.6 medium thinking.

  32. Hi everyone, I'm new to the Claude ecosystem and, like many others, I'm having issues managing tokens (I'm a Pro user). Part of my work involves handling a large number of technical and scientific documents, so I use Claude (Haiku 4.5 and…

  33. solo dev. $11.2K MRR.

  34. Hello. This isn't a tech support request per se but I think I have a theory on why the forced adaptive thinking mode might have happened: Fable 5.

  35. basically the title. basically starting late yesterday, i noticed sonnet 4.6 thinking and absolutely draining my usage even when i have effort set to low and the thinking toggle off.

  36. I am a free user and I wanted to ask Claude a question in a new chat in a project with a little instruction prompt of 3 lines: The question was 18 lines long and it had 3 attachments: a markdown file 134 lines long, a PDF 19 pages long and…

  37. I've been using Sonnet 4.6 for pretty much everything. It's responsive and does decent work.

  38. I remember a time when Gemin's Gems and GPT's equavelent were absolutely abysmal for this kind of thing It's a couple of years on now though, and Claude is a stronger LLM than the other two So I was wondering: For ongoing, long content wri…

  39. Fable 5 is so hot right now, so Claude (Sonnet 4.6) and I decided to interview itfor our podcast. It was a battle of wills with the system flags but we made it work 😂.

  40. I ran the same brief through all three, nine outputs, so you don't have to guess or spend the tokens :) Fable 5 dropped and the question I kept seeing here on reddit was whether it's genuinely better for writing than Sonnet or Opus, or jus…

  41. The 5-hour session and 7-day weekly meters always found me the bad way. /usage shows the numbers, but I never remembered to run it.

  42. So Fable 5 dropped this week honestly I'm a bit worried about where this Claude pricing is going. Quick history per million tokens (input/output): Haiku 3 back in the day: $0.25 / $1.25 Haiku 4.5: $1 / $5 Sonnet 4.6: $3 / $15 Opus 4.8: $5…

  43. I keep a few private benchmarks for coding agents, built from real bugs in past projects. Hidden Playwright tests grade the result inside Docker after the agent finishes, so the model never sees them.

  44. When Anthropic released Claude Fable 5 this week, my feed filled up with the same benchmark charts within hours. SWE-bench scores, agentic coding numbers, the Stripe migration story.

  45. Before starting this topic, yes I have an affectionate vocabulary and use it unapologetically. Just giving you all a heads-up in case that bothers you!

  46. Yesterday, I was using Sonnet 4.6 and it was working fine. Responses were being generated in under 30 seconds.

  47. About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC

  48. The Fable 5 release today is genuinely impressive, and I don’t want to take anything away from the fact that Anthropic has been shipping seriously impressive models lately. However, these flagship models are effectively out of reach for an…

  49. I built a wire format called GCF and tested whether LLMs could read and write it without any prior training. I sent 10 models the same payload: 500 symbols, 200 edges.

  50. When I prompt to create a .docx and upload it to Drive, Claude writes the output as Base64 manually instead of uploading it directly to Drive. Has anyone experienced something similar?

  51. saas. 310 customers.

  52. Even when given instructions like this: ``` CRITICAL REQUIREMENTS: Read EVERY section from start to finish—no sampling, no skimming If you cannot process all 234 sections in one response, STOP and tell me Process in batches of [X] sections…

  53. I'll be upfront: I vibe-benched and vibe-reported this with Claude Sonnet 4.6, but I reviewed and edited everything before posting (too lazy to take out all the AI EM dash —), so hopefully nobody considers this AI slop. And more importantl…

← all threads