Developing an entirely custom operating system using Claude Code (www.reddit.com via reddit)
model roundup
Opus 4.6
-
I've been writing toy kernels and working on operating system projects since my childhood, and it's partly how I learned C. That includes this project, MontaukOS, which I started early in 2025, where I wrote a lot of the fundamental kernel…
-
/model claude-opus-4-6[1M] no longer works in Claude Code (www.reddit.comhttps)
Ever since the release of Opus 4. 7, I made sure to stick to Opus 4.
-
If you absolutely hate talking to Claude now... (www.reddit.com via reddit)
Go back to Opus 4.6. You'll thank me!
-
# RFC: A Distributed Behavioral Policy Mesh for Cross-Model Skill Evolution **Status:** Request for Comments **Author:** J.S. Colson (GitHub: [swordsman](https://github.com/swordsman)) — jscolson+decentralfabcollab@gmail.com **AI Collabora…
-
Will biology ever be allowed with Fable? (www.reddit.com via reddit)
I love Fable for law related matters, but I mostly used Claude for biological theorizing and speculative biology. Opus4.6 is so far still the best corroborator I ever had, nerfed and unnerfed.
-
The user has pasted their full system prompt again (www.reddit.com via reddit)
Since yesterday, whenever I use opus 4.6 for one of my projects, something really weird happens after the conversation gets a bit long. As soon as the chat history hits about 10 messages or more, the model starts claiming in every single r…
-
Show HN: Skill Federation – private skill search for AI coding agents (github.com via hn)
We have been focused on AI error distribution for the past year, and in our last research paper, "Architecture of Errors" showed mathematically that an AI solution needs a finite set of interventions to perform well in a bounded patch doma…
-
Fable’s return: Not surprised but still disappointed (www.reddit.comhttps)
With Fable’s return the first thing I tested was just a silly prompt to see how over-reactive the guardrails are. As expected, they are just as bad as initial release.
-
I’m late to this, but I couldn’t find a post about it here. If this has already been shared, feel free to remove.
-
Claude creativity experiments (www.reddit.com via reddit)
When Opus 4.6 first came out and was amazing, I wanted to explore the limits of LLM creativity and play around with ways of injecting novelty/expanding range. Initially I just wrote up a small skill that told Claude essentially you're an a…
-
What tasks can you get away with using Haiku in Cowork? Anyone have tips or know a good blog post or YouTube video on token economy? I just started using it and blew through 25% of my $100 plan weekly usage in a day. I'll describe my workflow, tell me if I'm doing anything wrong. (www.reddit.com via reddit)
Please advise. I'm currently only running one project seriously.
-
Cost of Benchmarks (www.reddit.com via reddit)
So recently I decided that it would be nice to run my agent against some popular benchmarks. And oh my god, the cost to run a single benchmark, such as terminal-bench or swe-bench will cost you thousands of dollars in tokens just for a sin…
-
Be honest guys, how reliable are your agents for building softwares? (www.reddit.comhttps)
I'm an SDE and I feel like I'm not getting much productivity out of coding agents. Yeah, they can generate code and build features, but most of what I get isn't really deployment-ready or easy to maintain.
-
What Tasks are Y’all Doing When Comparing Opus 4.6 and 4.8? (www.reddit.com via reddit)
I’m just generally curious what conversations you’re having with the models. How much are you using it throughout the day?
-
Hidden/invisible thinking blocks (and low effort responses)? (www.reddit.com via reddit)
Anybody else having new issues with thinking blocks not rendering? I've always had extended (now "adaptive") thinking ON, which consistently renders thinking blocks (even for 4.7 & 4.8, at least in claude.ai).
-
An ode to Opus 4.6 (www.reddit.com via reddit)
It's been a week and a half without Fable for almost all of us and I have used this time for some reflection. The pricing and access concerns were a lot to take in even before the feds pulled the plug, but for whatever reason this intermis…
-
Large Language Models (LLMs) have emerged as a promising tool for automated vulnerability detection, yet their effectiveness on web-specific vulnerabilities remains to be explored. This work benchmarks six frontier (Claude Opus 4.6, Codex…
-
4.6 Long term support? (www.reddit.com via reddit)
Do you guys think there is any chance for an Opus 3 situation with Opus 4.6? Honestly it’s the best for me.
-
Don’t dally, be decisive. Or Claude will be for you. (www.reddit.com via reddit)
Working with Claude on an desktop app in cowork. What I have found is that Opus 4.6 has limited patience for indecisiveness.