A sweet treat (www.reddit.com via reddit)
model roundup
Opus 4.6
-
Finally had the privilege of fable spinning up 70 subs agents My max 5x plan stood no chance, didn’t even finish the first prompt and only lasted 20 minutes into the 5 hour session Granted it was doing a heavy research pass for my hat comp…
-
Ask HN: Dear Anthropic, can we please have thought traces back? (news.ycombinator.com)
Dear Anthropic, can we please have thought traces back? I can't verify whether or not the LLM is arriving at the conclusion from cheating, or if it's fudging or making stuff up.
-
Which will be best in $20 between Cursor Pro and Claude Code. (www.reddit.com via reddit)
Suggest me one. My priority is to do advance audit, bug finding and resolve them in my Flutter code.
-
Be careful running Claude Code subagents (www.reddit.com via reddit)
TLDR: Be careful with the use of subagents by actively limiting the number that can be created and don't allow them to spawn their own. Today, I ran into an issue with a prompt that I run frequently with Opus 4.6, 4.7, 4.8 with subagents t…
-
Why is Claude using usage credits when I haven't hit session or weekly limits? (www.reddit.comhttps)
I'm in Claude Code on Opus 4.6 at 25% of my session limit, 44% of my weekly limit, and it just charged me $4.60 in usage credits for a small prompt and won't work if I turn usage credits off. Wtf?
-
For content writing with natural tone: Opus 4.6 vs Opus 5 vs Fable? (www.reddit.com via reddit)
I'm wondering which model is currently the best for content writing that follows large instructions and produces natural tone, that is easy to read. My experience says it's Opus 4.6, but then it does not follow all the instructions.
-
I'm curious to know which models everyone is mainly using. (www.reddit.com via reddit)
I still can't bring myself to move away from Opus 4.6. I feel like it's more than enough to use, and the UI is good enough that I don't see any real need to upgrade.
-
I compared Opus 5, Fable, Sol, Qwen, and K3 on one strategy task (www.reddit.com via reddit)
I gave eight model and effort configurations the same prompt: design when a manager should use zero, one, or several AI advisers for an important decision without creating a permanent committee. This was one judged strategy sample, not a g…
-
Claude is Littish🔥 (www.reddit.com via reddit)
I'm interested in learning from you bluds faring and building production grade projects. Excluding plugins and skills from 3rd party sources.
-
Opus 5 for writing partner? (www.reddit.com via reddit)
I like using Claude as a talking partner to talk about my writing ideas and DND campaign planning with, simple stuff, mostly really character focus as I end up finding mapping out exact mindsets really fun. I have a pro subscription for it.
-
Claude Document Revision Error - A warning (www.reddit.com via reddit)
I was having Claude review and revise a document using Opus 4.6 Medium and a skill we created. I noticed that it accidentally deleted chunks of text without noticing.
-
Hi All!. :) Over the last few weeks, I’ve been experimenting with how far AI-assisted development can go beyond the usual web applications and automation scripts.
-
Claude Opus 5 takes second place on SimpleBench (www.reddit.com via reddit)
Opus 5 scores just below Fable 5 (1.3 percentage points lower), but vastly outperforms Opus 4.6, 4.7 and 4.8, and all other models tested. > SimpleBench includes over 200 multiple-choice questions covering spatio-temporal reasoning, social…