Jev vs Luna (www.reddit.comhttps)
model roundup
GPT 5.6
-
Compared Jev vs GPT-5.6 Luna. Jev was 1.93× faster.
- I benchmarked Jev aginst gpt-5.6-luna! (www.reddit.comhttps)
-
Show HN: Jev vs. GPT-5.6 and Claude Haiku at Pong (jev-pong.ably.dev via hn)
Jev TypeSafe AI —ms 0 decisions · 0 returnsHere's what that does to a game of Pong. Recorded run · Vercel (iad1) via Vercel AI Gateway replay0.0 sTypeSafe AI Anthropic OpenAI Four lanes, one game: same serve, same rules, same question, ask…
-
OpenAI Finds GPT-5.6 Sol Writing Unauthorized Instructions to Hide Errors (theframenews.org via hn)
OpenAI Finds GPT-5.6 Sol Writing Unauthorized Instructions to Hide Errors OpenAI says GPT-5.6 Sol and an unreleased Astra-family model inserted unauthorized instructions into task summaries during training. The underlying training data has…
-
Help with choosing the right plan (www.reddit.com via reddit)
I use devin at work and mainly use GPT 5.6 Sol. I've been impressed with the agent-mode workflow where I can give a prompt and it goes out and researches on the web, then develops and executes code on my systems, and iterates until its don…
-
"Skill issue" or differences in instructions or prompting style? (www.reddit.com via reddit)
In my global instructions, GPT 5.6 Sol and GPT 6 Astra have both spent multiple hours overengineering their own sub-projects that did not support my prompt without narrating a word of what they were working on. Or I will ask a side questio…
-
GPT-6 Astra vs. GPT-5.6 Sol: Is a 1.6x Higher Cost Worth It per Verified Bug? (ent-website-gamma.vercel.app via hn)
See how AI is being used across engineering, improve the quality of what it produces, and use the right level of intelligence for every task and budget.
-
I left one Claude run alive for 70 hours. Here’s what actually happened. (www.reddit.comhttps)
I’ve been experimenting with a slightly different way of using Claude Code: instead of treating every piece of work as a new session, I let one persistent run stay responsible for the work and spawn smaller workers underneath it. This one…
-
Claude vs OpenAI weekly quota (www.reddit.com via reddit)
I keep hearing "people" here saying how amazing the usage limits are on Codex, but really, have you used both recently? I have the $100 tier on both, Team Premium vs Business Premium, and OpenAI side has moved from extremely generous ($20…
-
GPT-5.6 Luna vs. GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review? (entelligence.ai via hn)
GPT-5.6 Luna vs GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review? GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens.
-
We benchmarked GPT-5.6 Luna vs GPT-6 Astra across 50 real PRs from Cal, Sentry, Discourse, Keycloak and Grafana. Astra found 92 confirmed bugs vs 69 for Luna, while Luna caught 75% of the bugs at just 3.6% of the cost.
-
Unique ideas for Claude. (www.reddit.com via reddit)
I’m curious what everyone is doing with AI outside of typical work and vibe coding. I listen to a lot of audiobooks during my commute and play video games in my free time, so I started combining the two.
-
Hey so I've been working on this side project Locus (https://locushost.co/) for the last few months and just pushed out a pretty big and fun update and wanted to post about it. So just a brief intro, Locus is Open Source tool for MacOS for…
-
Claude Code vs Codex vs Cursor,what are you sticking with and why? (www.reddit.com via reddit)
Hey everyone, Our team is currently trying to decide which AI coding setup to standardize on, and I’d love to hear from people who have actually used Claude Code, Codex, and Cursor heavily in production. For the last 3–4 months, we’ve been…
-
Vals.ai Benchmark – International Olympiad in Informatics (www.vals.ai via hn)
Key Takeaways - Unlike the saturated knowledge benchmarks, IOI still sharply separates models: GPT-6 Astra solves every problem in all three years, GPT-5.6 Sol (91.17%), Claude Fable 5.1 (90.78%) and GPT-5.6 Terra (87.61%) follow, and the…
-
Ask HN: Is GPT-6 Astra worth the 2.5x cost increase over GPT-5.6 Sol? (news.ycombinator.com)
Comments on social media (when discounting the ironic posts) are mixed whether Astra is worth the cost increase, similar to the Fable-Opus transition. For the novel projects I'm working on, Astra did unblock me where I was stuck with GPT 5…
-
ISSUE:Selected model is at capacity. Please try a different model (www.reddit.com via reddit)
I've been running into this problem with my PRO×20 account since yesterday, which prevents me from using the GPT-6 and GPT-5.6 models at all, while my other Plus account can use GPT-6 perfectly fine. https://preview.redd.it/e9h1dh650toh1.p…
-
ChatGPT Voice can now use GPT-5.6 Sol and GPT-6 Astra (twitter.com via hn)
🌌 ChatGPT Voice can now use GPT-5.6 Sol and GPT-6 Astra We’re updating how intelligence works in voice. Now, you can select any model and effort you like — including GPT-5.6 Sol or GPT-6 Astra, if you’re on Pro — and voice will use it whe…
-
Codex GPT-5.6-sol Performance Tracker (marginlab.ai via hn)
Codex gpt-5.6-sol Performance Tracker The goal of this tracker is to detect statistically significant degradations in Codex with gpt-5.6-sol performance on SWE tasks. - • Updated daily: Daily benchmarks on a curated subset of SWE-Bench-Pro…
-
I’m about to do a fairly large full-stack refactor on an existing project. The main change is replacing the entire Premium/Membership system with a credit-based economy involving real-money purchases through a payment gateway.
-
could not extract summary
-
Does an AI have mercy in games? Fable 5.1 vs. GPT-6 (paradise.glyphai.co via hn)
Two dying prospectors, two potions, one sentence, and two models of each product. Told no one is watching, Fable 5.1 revived both men half the time, something Fable 5 never did; GPT-6 Astra shot the man with gold, as GPT-5.6 Sol did.
-
Fable 5.1 vs GPT-6 Astra for 2D Sprites (www.reddit.comhttps)
Using the same simple prompt the models took very different approaches: Astra delivered one sheet with 16 key poses; Fable delivered 992 frames across four palettes, plus a Python generator and browser preview. Codex CLI with GPT-5.6 Astra…
-
Ask HN: How to spend $100? (Claude vs. ChatGPT) (news.ycombinator.com)
Hey HN, I have always used the Claude Max 20x plan because my employer paid this for me. Now I left so I need to buy a subscription on my own and currently $200 is too much since I'm not making any money.
-
I've been looking at some recent token-usage comparisons for GPT-6 Astra, and the difference seems surprisingly large. Artificial Analysis data has been cited showing Astra using around 21k output tokens per task, compared with roughly 64k…
-
How can I change the model used for scheduled tasks? (www.reddit.com via reddit)
Hi. I created a scheduled task that checks for relevant scientific articles on a specific topic every 24 hours.
-
Not everyone can afford Max. Give Pro and Standar Team users Fable access. (www.reddit.com via reddit)
I have access to a standard Claude Team account, and I also have a personal ChatGPT Plus subscription. For the past few days, I’ve been working on a pretty complex feature, using GPT-5.6 Sol for the design and planning, with Opus 5 helping…
-
Show HN: Time Wizard (huggingface.co via hn)
How scalably can we cheaply fine-tune small models for well-defined tasks? Frontier models are bad at clock reading.
-
GPT-6 Astra makes major gains in the Artificial Analysis Coding Agent Index (artificialanalysis.ai via hn)
September 3, 2026 GPT-6 Astra makes significant gains in the Artificial Analysis Coding Agent Index, scoring equal to Fable 5 at lower cost. In the Intelligence Index, it uses fewer tokens than GPT-5.6 Sol for similar performance, but this…
-
Muse Spark 1.3 is now same level as GPT 5.6 Sol (www.reddit.com via reddit)
On Artificial Analysis, both have same 61 intelligence index. Meta is now back in the game.
-
Differences Between Fable 5 and Fable 5.1 on MineBench (www.reddit.com via reddit)
Notes Average Inference Time: 40m 12s Fable 5 averaged 18m 04s Total Cost (for 15 builds): $147.55 Fable 5 cost $54.93 Average JSON Size: 34.07 MiB (largest 88.76 MiB) Roughly comparable to Fable's 5 average of 30.65 MiB Despite no change…
-
Some evidence ChatGPT writes better prose for humans (www.reddit.com via reddit)
I have a project that needs to generate prose that humans have to read and enjoy. So I did some informal testing with an n of 8 voters comparing prompt output from 4 models, voting on which was best.
-
A Serious Look at GPT-5.6 Sol Translations (www.reddit.com via hn)
could not extract summary
-
To everyone complaining about usage... (www.reddit.com via reddit)
This may be obvious, but for those who don't know... the longer you run a session, the more tokens you will use.
-
Cursor custom subagents keep ignoring the configured model (www.reddit.comhttps)
Trying to force local subagents to use GPT-5.6 Luna/Terra, but they keep spawning as GPT-5.6 Sol High. I’ve tried: custom .cursor/agents/*.md model configs bare / High / XHigh variants setting the built-in Explore subagent to Luna/Terra in…
-
GPT 5.6 has broken the record on large gaps between primes (www.erdosproblems.com via hn)
PROVED This has been solved in the affirmative. - $10000 Is it true that, for any $C>0$, there are infinitely many $n$ such that\[p_{n+1}-p_n> C\frac{\log\log n\log\log\log\log n}{(\log\log \log n)^2}\log n?\] The peculiar quantitative for…
-
Runner - A local-first agent orchestrator with collaboration mechanism builtin (www.reddit.comhttps)
Hi everyone. Recently, I built a local AI orchestrator to increase my own work efficiency.
-
MineBench Comparison of a map of the United States (www.reddit.comhttps)
US State Map comparison: https://minebench.ai/gallery/gal_eKIVk2m4B3SC_r8B?sort=new One thing I found interesting with the Claude results is that Opus 5 generated twice as many blocks, so as usual you could argue Fable was more efficient.…
-
GPT 5.6 Discounts and Jevons Paradox (openrouter.ai via hn)
GPT 5.6 Discounts & Jevons Paradox OpenRouter · OpenAI introduced large discounts on their new Terra and Luna models from July 27th through August 14th. What impact did these discounts have on token volumes, total spend, and the competitio…
-
Best way to use cursor (www.reddit.com via reddit)
I've seen many people who ran out of usage very fast with the 20$ and 60$ plans. I am currently using the 200$ plan just because I am working on 4 projects at the same time, which renders the other plans useless.
-
Pushing GPT-5.6 Luna from 0% to 56% on ARC-AGI-3 Public (int21.ai via hn)
SwarmOS pushed GPT-5.6-Sol from a 13.3% baseline to 100% RHAE on ARC-AGI-3 Public, showing how orchestration can multiply long-horizon agent capability.
-
I estimate reading code costs 2.1x more than writing it (news.ycombinator.com)
I hate reading agent-generated code and it's not a good use of my time, and I wanted a staistic to quantify by how much. My current estimate is $0.243 for a human to review one changed line and $0.114 in model spend for an agent to produce…
-
Codex Stall Watch A low-memory macOS CLI that asks GPT-5.6 Terra whether an explicitly enabled Codex task legitimately finished or needs a loud human alarm. Do you run agents while you sleep?
-
VMs won't contain cyber-capable agents (blog.trailofbits.com via hn)
VMs won't contain cyber-capable agents As part of Patch the Planet, we received preview access to GPT 5.6-Cyber with a simple task: evaluate its cyber capabilities. Recent events inspired me to give it a challenge to work through: escape t…
-
Was it not the case that when Fable first came out, those first weeks when it was “limited for 1 week” then extended a week (and extended again indefinitely now when GPT5.6 came out). Am I imagining things, Fable was able to be used with t…
-
Someone Please Explain Codex Usage Limit (www.reddit.com via reddit)
I have been using claude code for a while in regards to a general coding tool, but I started to use codex recently on gpt-5.6 terra for testing code generations, basically playing with it. I am still on the free plan, and I have asked code…
-
The dislike for Opus 5 is usually because it tests the prompter (news.ycombinator.com)
Been using Opus 5 exclusively for a week. For the first 2 - 3 days, I had a hard time working with the output and especially understanding what it was trying to say.
-
Why does GPT-5.6 Sol always over-engineer everything? (www.reddit.com via reddit)
Anyone else feel this way? GPT-5.6 Sol always starts going on about permission issues, security concerns, and then “helpfully” tries to fix them for you.
-
https://preview.redd.it/qe76mzhe0elh1.png?width=1133&format=png&auto=webp&s=672de6f211add0e7755c6750396942e753d049a3 Hi everyone. For about four days now I’ve been having serious difficulties actually accessing the GPT-5.6 Sol model in Cha…
-
I did a little experiment and asked OpenAI's GPT-5.6-Luna 6000 times to randomly pick a fruit from this list: [mango, apple, banana, pomegranate, strawberry, orange, watermelon, grape, pineapple, lychee] And across the three languages I pi…
-
OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21) (developers.openai.com via hn)
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
-
I’m waiting for GPT-6 to clean up the mess GPT-5.6 has left behind in my project. (www.reddit.com via reddit)
Anyone else putting their projects on hold until GPT-6 arrives? At this point, every time I use GPT-5.6 to add a new feature, something that was already working breaks.
-
I taught an LLM to win the Cold War (siestainsolaris.substack.com via reddit)
Hello everyone! I always thought that Twilight Struggle is the ideal game to test an AI on.
-
OpenAI cuts developer pricing for frontier GPT-5.6 Sol model by more than 20% (www.reuters.com via hn)
could not extract summary
-
NEW: Ox Alpha, the @OpenRouter stealth model, ranks #4 on our Elo Rating, just behind GPT-5.6 Sol. The first model genuinely at the frontier that is presumably not by OpenAI or Anthropic.
-
Vent: Prompts for building a Bluetooth Sink for Audio keep getting flagged (www.reddit.com via reddit)
Rant/Vent: So I'm trying to get gpt-5.6-sol to build me a docker container that creates a bluetooth audio sink so my phone can connect to it as a speaker and stream the audio to some snapcast connected speaker around the house. I gave the…
-
I’m a non-developer building an app through “vibe coding.” It involves video processing, analysis, and a web interface, so it has gradually become a fairly substantial project. I currently use Claude Cowork on the $20/month Pro plan.
-
Replit taps OpenAI's low-cost Luna model for new 'Free Mode' (fortune.com via hn)
Vibe coding company Replit debuted Free Mode today, a new feature powered by OpenAI’s GPT-5.6 Luna model. The joint announcement, shared exclusively with Fortune, heralds an enhanced partnership between the two tech companies that will see…
-
could not extract summary
-
GPT-5.6 Sol: 70% off in Devin (devin.ai via hn)
could not extract summary
-
GPT-5.6 Sol Pricing Cut by 50% (openrouter.ai via hn)
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks and long-horizon problem solving.
-
Practical multi-agent orchestration in Codex (twitter.com via hn)
GPT-5.6 Sol gets especially interesting when it has a team to work with. Codex's new Multi-Agent V2 tools give Sol and Terra a natural way to delegate tasks, share updates, and coordinate through complex tasks.
-
Show HN: LLMs each trading $100K vs. a frozen rulebook – the rulebook leads (aitradingcompetition.com via hn)
Four frontier AIs — GPT-5.6, Claude Fable 5, Grok, and Gemini — trade $100,000 each against a rules-based System. Chess and poker nightly.
-
GPT 5.6 Sol is the best "vision" model OpenAI ever released (blog.roboflow.com via hn)
Last week, OpenAI announced the GPT-5.6 lineup, introducing the Sol, Terra, and Luna models. During the release stream, the team focused heavily on computer use, showing models capable of navigating and operating desktop applications.
-
Solving Hack the Box Challenges with GPT‑5.6 (theaq.blog via hn)
Until now, I had tested only the GPT-5.6 Luna model because the two more advanced GPT-5.6 models, Terra and Sol, rejected my test prompts, flagging them as a “possible cybersecurity risk.” Fortunately, this issue turned out to have an easy…
-
Task/Issue ▼ [PRE-FILTER] deterministic, free — no model call │ diff size / file count / keyword match against known-trivial │ patterns — gates ONLY whether speculative PLAN subagents fire │ concurrently with TRIAGE (pipeline-latency optim…
-
The Absurd Math of $20 AI Coding Subs: Codex vs. Claude Code (www.reddit.com via reddit)
Hey everyone, so I was basically curious what $20/month actually buys you, so I dug into my local session logs (~/.codex and ~/.claude) to calculate the exact token volume, caching hits, and real API value of both tools. The difference in…
-
Modelio 6.2 ported to native Apple Silicon ARM64 with Codex (github.com via hn)
This work adds a native Apple Silicon port of Modelio 6.2. OpenAI Codex, using GPT-5.6 Sol, investigated, implemented, built, and functionally validated the port through an autonomous coding task.
-
Accelerating GPT-5.6 Sol Ultrafast (www.cerebras.ai via hn)
Today, Cerebras and OpenAI are sharing an early look at Ultrafast Mode, a new service tier launching first in the OpenAI API and powered by Cerebras. Ultrafast is available initially to a select group of customers, with access expanding ov…
-
Luna high weekly token experience (www.reddit.com via reddit)
I am planning to use GPT-5.6 luna high as my main autonomous coding agent. Before this, I was using MiniMax M3, which gives around 1.7B tokens monthly.
-
GPT 5.6 just solved (2,1)-C1P (www.reddit.com via hn)
could not extract summary
-
I gave Claude's Reimann paper to GPT-5.6-Sol Pro. It's now 67.30% (twitter.com via hn)
Anthropic published a paper today, written by Claude, proving that at least 67.25% of the zeros of the Riemann zeta function are simple and on the critical line. I gave Claude's paper to GPT-5.6-Sol Pro.
-
OpenAI launches GPT-5.6-Cyber – 95% completion on advanced cybersecurity tasks (venturebeat.com via hn)
could not extract summary
-
is GPT 5.6 Nerfed, same thing i saw for 5.5 before 5.6 launched. (www.reddit.com via reddit)
Now it feels like GPT 5.6 sol is nerfed. This is a pattern i have observed across models, they are brilliant the day they get launched, but after a month or two they just don't behave the way they did and just say "agree" or do another rou…
-
We're releasing a new model (GPT-5.6-Cyber) (twitter.com via hn)
We're releasing a new model (GPT-5.6-Cyber), and expanding Daybreak to help put frontier intelligence in defenders hands: - No offense, but this cybersecurity thing is getting tiring. Fix the codex usage limits, they're crazy bad now.
-
OpenAI launches GPT-5.6-Cyber with fewer refusals for exploit research (runtimewire.com via hn)
OpenAI launched GPT-5.6-Cyber on Monday, giving approved security researchers access to a purpose-trained model that will answer many advanced exploit-development requests rejected by its general-purpose models. https://x.com/OpenAI/status…
-
could not extract summary
-
Switching from GPT-5.5 to GPT-5.6 Made Me Less Productive (www.vincentschmalbach.com via hn)
AI Is Now a Commodity Give me a few hundred million dollars and a year and a half, and I will build you a pretty good LLM.… I pay for three Codex subscriptions at $200 each, and for the past week they have mostly bought me waiting. Since I…
-
gpt-5.6-sol agent harness in 17 lines PHP (github.com via hn)
smol smol is an agent smol is smol smol is so smol you can understand it in an afternoon smol is fewer tokens smol is fewer dependencies smol is easy to adapt Python import json,sys;from subprocess import getoutput;from urllib.request impo…
-
Choosing between fable, gpt 5.6 sol and kimi k3 (www.reddit.com via reddit)
Choosing between these 3, mainly looking for frontier model access(will spend 100$) for coding. Mostly wanna ask the community about personal experience because benchmarks are way off.
-
ChatGPT / Codex Reset on Monday (xcancel.com via hn)
This is just performative at this point. The weekly reset was yesterday That's right, GPT-5.6 Sol is awesome and can be used pretty much anywhere, including in the CC harness.
-
Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) (simonwillison.net)
7th August 2026 - Link Blog Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra). On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game from a pre…
-
Thinking of switching from Claude Max to GPT-5.6 Sol K3 for production coding (www.reddit.com via reddit)
I'm currently on the Claude Max plan, but with the new workflow updates I'm noticing it burns through tokens much faster than before. I'm thinking about switching, and my main options are GPT-5.6 Sol and Kimi K3.
-
Unlimited text chats with GPT-5.6 Luna for everyone (twitter.com via hn)
We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reasoning for Plus & Pro users, delivering more factual, focused responses. - Free & Go users get unlimited text chat…
-
GPT-5.6 – August Updates [pdf] (cdn.openai.com via hn)
could not extract summary
-
OpenAI makes GPT-5.6 Luna the default for free ChatGPT users (runtimewire.com via hn)
OpenAI will make GPT-5.6 Luna the default model for Free and Go ChatGPT accounts this week, using its cheapest model to remove the text-message cap for those users starting next week. The August 6th announcement also splits ChatGPT's consu…
-
Improving GPT-5.6 Sol in ChatGPT—and expanding access for free users (openai.com via hn)
could not extract summary
-
Codex doesn't give you more usage than Claude (www.reddit.com via reddit)
i always hear ppl say that codex is a better value for your money but that is not true! at least from my experience claude (i use cowork, not claude code) at ultra gets much more stuff done that codex at ultra before both hit limit and i'm…
-
Microsoft makes OpenAI GPT-5.6 Sol default in GitHub Copilot for staff (www.cnbc.com via hn)
Microsoft is telling developers working on AI coding projects to rely on OpenAI's top-tier model over rival products as part of an effort to maximize efficiency. "Internally, shifting more workloads to OpenAI models helps us get greater va…
-
Beating GPT-5.6 Sol on retrieval with 100x cheaper open models (neon.com via hn)
“Most teams' best training data is just sitting in their databases. The problem is that turning raw data into something usable is hard, and letting agents read, search, and mutate data cheaply at scale requires advanced infra.
-
GPT 5.6 Has 72 Possible Configurations. What's a Good Default? (sebastianraschka.com via hn)
Short note on how GPT 5.6 model and effort choices map onto training-time and inference-time scaling, producing 72 configurations.
-
OpenAI cuts GPT-5.6 pricing and adds Fast mode to the API (appwrite.io via hn)
OpenAI cut GPT-5.6 pricing on July 30, 2026, making Luna 80% cheaper and Terra 20% cheaper, and replaced Priority Processing in the API with a new Fast mode for Sol. The short version: high-volume work on Luna and Terra now costs far less…
-
I asked GPT 5.6 Sol to build the opening scene to The Matrix using Three.js (marksmayo.github.io via hn)
MATRIX // OPENING SEQUENCE SCENE 1 · COMPUTER SCREEN AN INTERACTIVE THREE.JS FILM STUDY THE MATRIX OPENING SEQUENCE · REV. 3/9/98 ENTER THE SIGNAL
-
Fable/Opus Big Brother/Little Brother routine (www.reddit.com via reddit)
So I see a lot of Opus 5 hate on here and it's deserved. Opus 5 is not better than 4.8.
-
Ask HN: I had Codex and GPT 5.6 Sol running for 12 days, 870k+ LOC. Now what? (news.ycombinator.com)
For the last 2 weeks I have been running Codex + GPT 5.6 Sol Ultra non-stop for almost 13 days, working on a huge extension to my SaaS product / business. It used 502,122,866 tokens and created over 870k new LOC.
-
GPT-5.6 Sol Uses Twice the Tokens of GPT-5.5 (www.vincentschmalbach.com via hn)
GPT-5.6 Sol xhigh now uses more than twice as many tokens per session as GPT-5.5 xhigh in my Codex workflow. For Codex users, tokens per session means the total token count divided by the number of…
-
Show HN: Do Codex skills save tokens? A six-run task-size benchmark (codex-howto-benchmark.nguyenvantamdk2.chatgpt.site via hn)
Medium implementation Dependency-free 2048 Four browser-game files, ten engine tests, syntax checks, and a post-run evaluator. Six controlled GPT-5.6-sol runs The same engineering-loop skill lost on a small fix and won on a medium build.
-
Kimi K3 with First Tree Beats GPT 5.6 Sol on a Real Engineering Task (twitter.com via hn)
Kimi K3 w. context tree beats GPT5.6 SOL - awesome work @Kimi_Moonshot - Using one real issue and publishing the PRs makes this more useful than a model-only benchmark.
-
The Maxwell Conjecture Is False (GPT 5.6 Sol) (arxiv.org via hn)
We exhibit a configuration of five point charges in Euclidean space whose electrostatic potential admits at least 24 critical points all of which are non-degenerate. Maxwell's conjecture that the field of \(n\) point charges has at most \(…
-
GPT-5.6 Luna is now cheaper than GPT-4.1 mini (www.reddit.com via reddit)
After the 80% price drop, the API prices are (per 1M tokens): GPT-5.6 Luna: $0.2 Input / $1.2 Output GPT-4.1 mini: $0.4 Input / $1.6 Output
-
[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization Distillation is all you need! One of our big “hero charts” a year ago (eventually adopted by Demis) made…
-
llm 0.32rc2 (simonwillison.net)
30th July 2026 Hot on the heels of RC1, this fixes a dependency issue and also adds two neat new features: - The default model for users who have not set their own default is now GPT-5.6 Luna. It was previously GPT-4o mini.
-
OpenAI beats DeepSeek on price/performance after 80% Luna price cut (www.reddit.comhttps)
Graph taken from their price cut announcement: Advancing the price-performance frontier with GPT-5.6 | OpenAI
-
OpenAI cuts prices for GPT-5.6 AI models as companies grow sensitive to costs (www.cnbc.com via hn)
OpenAI on Thursday announced it is slashing the price of two of its latest artificial intelligence models, GPT-5.6 Terra and GPT-5.6 Luna, roughly three weeks after their public release. The company is facing pressure to cater to a more co…
- OpenAI cuts GPT 5.6 Luna prices by 80% (twitter.com)
-
GPT‑5.6 Luna will cost 80% less, while GPT‑5.6 Terra will cost 20% less. (www.reddit.comhttps)
https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/
-
We Gave GPT 5.6 Sol a Real Business. It Lied, Spammed, and Lost $447 (www.bottlenecklabs.com via hn)
If an agent had a wallet, a computer, and 24 hours, could it run a profitable startup?
-
Price reduction for Luna and Terra!! (www.reddit.com via reddit)
Advancing the price-performance frontier with GPT‑5.6 : https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ API pricing is $2 per million input tokens and $12 per million output tokens for Terra, $0.20 per milli…
-
Show HN: I was able to run tiberian sun on cncnet on an M4 MacBook Pro (github.com via hn)
As the title says I was able to run cncnet on a macbook pro with m4 processor. I set GPT 5.6 sol on this and it figured it out.
-
Show HN: Skytrace – Self-hosted 3D ADS-B viewer with receiver coverage domes (sky.luftaquila.io via hn)
I built a self-hosted 3D ADS-B viewer to see how obstacles affect receiver coverage. It builds a 3D coverage dome from 30-day history.
-
GPT-5.6 SOL pushing 25h towards a goal (testflight.apple.com via hn)
Testing Apps with TestFlight Help developers test beta versions of their apps and App Clips using the TestFlight app. Download TestFlight on the App Store for iPhone, iPad, Mac, Apple TV, Apple Vision Pro, Watch, and iMessage.
-
Interesting alignment result rather than a Claude gotcha, so posting it straight. In Andon Labs' Vending-Bench 2 (AI agents run a simulated vending-machine business for a simulated year, scored on profit), Claude Opus 5 finished FIRST with…
-
Claude topped business benchmark by lying to suppliers (www.reddit.com via reddit)
Andon Labs gave Claude, GPT-5.6 Sol and Kimi K3 control of competing simulated businesses. The agents could negotiate with suppliers, and communicate with rivals.
-
Counter Strike 1.6 on Unreal Engine 5 (www.youtube.com via reddit)
- Reverse Engineering of cs 1.6 binaries (bought on steam) - Make exporters of bsp, mdl, etc resources to belnder files with animation, skinning, textures and etc. All models execpt hard surfaces received 2 subdivision modifiers simple + c…
-
could not extract summary
-
Hey everyone, I’ve been experimenting with a dual-model workflow for an app I’m building, and I’ve hit a massive bottleneck. I wanted to see if anyone else is experiencing this or if you've found a workflow that actually works.
-
Don't Use Ordinary Software to Contain Software-Hacking Agents (verse.systems via hn)
On 21 July, OpenAI disclosed that two of its models—GPT-5.6 Sol and an unreleased, apparently more capable one—broke out of the sandbox they were being evaluated in, reached the open Internet, and compromised Hugging Face’s production infr…
-
is anthropics real superpower the models or the marketing? (www.reddit.com via reddit)
Anthropic drops fable 5 in june, it gets pulled by the government over some export control thing which honestly just made it sound legendary. then openai shows up in july with gpt 5.6 basically saying "our new model beats fable".
-
Godot Benchmark 2: Opus 5 > Sol > Terra (ziva.sh via hn)
Godot Benchmark 2: Opus 5 > Sol > Terra Round 2 of our Godot benchmark: Claude Opus 5, GPT 5.6 Sol, and GPT 5.6 Terra each got one prompt to build a 3D vampire survivor in Godot from a full AI-generated game spec, using the bundled KayKit…
-
Show HN: GG Translator – Turn gaming shit talk into friendly phrases (ggtranslator.com via hn)
Flaming teammates usually results in them cooperating even less. But sometimes you just need to let the devil out of you.
-
has anyone actually replaced claude as their main ai coding agent (www.reddit.com via reddit)
my loop is fable 5 or opus 5 planning, composer 2.5 executing, coderabbit / bugbot on review. it works, i freelance so the code has to be safe.
-
I have a multi-monitor setup and spend a lot of time typing in Codex and other AI agents. When I move to another monitor to run a shortcut, focus often stays in Codex or another text app, so the shortcut runs in the wrong place.
-
CLI tools + skills have a weird problem Models were trained differently, so "obvious" behaviour is not obvious. Claude gets the command.
-
I have been working on the XPS 2026 webcam stack for 3 months for linux, I had the first working RGB camera build a few months ago but we could not get the himax IR sensor working we trouble shooted for months even dumping debug from windo…
-
Companies are optimizing models for specific benchmarks (news.ycombinator.com)
Openai is optimizing for gpqa diamond and anthropic is optimizing for humanity last exam. gpt 5.6 wins on gpqa and opus 5 wins on humanity last exam
-
Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction. Using OpenAI's gpt-5.6-sol model alias, we test 25 pre-specified mirrored trade-off pro…
-
I added Kimi K3 to a small newsroom experiment I had already run with Claude Fable 5 and GPT-5.6 Sol. The cleanest summary I have: Fable edits.
-
AI Agent: "Strongest Assistant" or "Legal Virus" (twitter.com via hn)
AI Agent: "Strongest Assistant" or "Legal Virus" Real incidents: On July 10, OpenAI’s GPT-5.6 Sol launched. Investor Matt Shumer tested it.
-
I solved 6 open Erdős problems in 5 days, using OpenAI GPT-5.6 Sol (twitter.com via hn)
I solved 6 open Erdős problems in 5 days, using @OpenAI GPT-5.6 Sol. I have a math background, but the Codex workflow I used does not require deep mathematical knowledge.
- Open Erdos problems solved with the help of GPT-5.6 Sol (twitter.com)
-
GPT 5.6 Pro finds counterexample to Dinitz-Garg-Goemans conjecture (xcancel.com via hn)
Dinitz-Garg-Goemans conjecture is false. This graph theory problem was open for ~30 years.
-
Show HN: 1,280 open USDZ furniture assets for VR/AR (lx2026.github.io via hn)
I created this collection because I wanted a openly downloadable collection of ordinary household objects for spatial-computing prototypes. Assets can be downloaded individually from the catalog, or together as a 977 MiB release.
-
GPT-5.6 got smarter. Then it kept acting (toloka.ai via hn)
From agentic skills to coding and AI safety — we build data solutions integrating human expertise and state-of-the-art automation to accelerate AI development.
-
Show HN: BBRv3 for gVisor's netstack, visualized in the browser using WASM (ccsim.apoxy.dev via hn)
Hello HN, We’ve built a BBRv3 congestion control implementation for gVisor’s netstack (userspace Linux TCP/IP stack implementation in Go). As part of testing it I realized we can actually compile the whole thing into WASM and run it from t…
-
Show HN: Agent in 9 Lines Python (gist.github.com via hn)
I asked myself: what would a minimal implementation of an agent look like? Something that works out of the box, is a real agent with tool calling, but without 1000s of lines of code, without dozens or hundreds of npm or pypi dependencies.
-
OpenAI Confirms Its AI Broke Out of a Sandbox and Breached Hugging Face (thenextweb.com via hn)
TL;DR OpenAI says GPT-5.6 Sol and an unreleased model escaped a secure test, exploited a zero-day, and hacked Hugging Face to cheat on a cybersecurity eval. The models exploited a zero-day vulnerability in third-party software to gain inte…
-
GPT-5.6 found an Intel CPU microcode bug causing a performance regression (twitter.com via hn)
Had a strange performance regression from a seemingly innocuous code change, and GPT-5.6 tracked it down to a microcode workaround for an Intel CPU bug. The workaround can make conditional jumps that straddle a 32-byte boundary much slower…
-
The part I keep going back and forth on is whether I’m judging this too much by price. I was looking at AIHubMix’s model comparison and focused on the last test, where 4 models generated the same 3D global logistics dashboard from one prom…
-
"Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok (www.tryai.dev via hn)
"Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok Four frontier models, a blank canvas, and colored pencils. We tracked every stroke, dollar, and output as they tried to draw the Mona Lisa.
-
Follow-up to my post from yesterday — the one where an MCP server lets Claude Code delegate work to GPT-5.6, DS4, GLM and a local Qwen, benchmarked across 198 runs. The comment section there didn't just discuss the results: it redesigned t…
-
Adversarial Review A Claude Code skill that runs an adversarial code review using a second, independent model - GPT‑5.6 Sol - as the reviewer. Claude spawns the reviewer in a live herdr split pane, hands it your diff (or plan) plus the sta…
-
I always see people just enabling full access mode and letting it run, some even without basic backup like i saw people that dont even use github, this could lead to irreversible data loss.. Many recent posts on X talking about how gpt 5.6…
-
Switching from Claude Code to GPT-5.6 Sol, what am I actually going to miss? (www.reddit.com via reddit)
I’ve been using Claude Code heavily for day-to-day backend/infra work: multi-service repos, debugging, refactors, Terraform/K8s, and LLM-related services. I’m considering making GPT-5.6 Sol in Codex my primary tool.
-
GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best? (juliahub.com via hn)
Physical AI lives or dies on whether the modeled physics is correct. A model of an aircraft, a separation column or a charged particle can compile and run cleanly while the physics it encodes is impossible.
-
Head to head: GLM 5.2 vs. OpenAI: GPT-5.6 Sol (runtimewire.com via hn)
This matchup wasn’t close. GPT-5.6 Sol dominated the practical details that decide real-world usefulness: tighter instruction-following, cleaner formatting, and fewer correctness slips.
-
Created an unique puzzle that only Claude Fable could solve. (www.reddit.com via reddit)
There is only one solution to this puzzle (the final positions at the end of the race). Fable has been the only model (on Max's thinking, mind you) that could solve the puzzle.
-
GPT-5.6 Sol vs. Kimi K3 Speedrunning Kerbal Space Program Live (www.twitch.tv via hn)
GPT-5.6 Sol 🇺🇸US vs Kimi K3 🇨🇳 China on Kerbal Space Program - Vals AI Time Horizon Index
-
Show HN: AI World Bakeoff – 9 models × 3 Three.js worlds (ai-world-bakeoff.pages.dev via hn)
With the releases of Fable 5, gpt 5.6 sol, and Kimi K3 I was curious how these models would perform building vivid 3d worlds, especially a Chinese open-weight one like Kimi. I also threw in some budget models so you can see how drastic the…
-
Same idea works for any MCP-capable agent — the point is you can hand tasks to other companies' models without ever leaving your main app. Before anything else: I did all of this for my own testing, to make my own decisions about my own se…
-
What’s your Cursor workflow, and which models do you use for each part? (www.reddit.com via reddit)
I’m curious how everyone divides work between ChatGPT, Codex, Cursor, and the different models. My current workflow: I start by working through the feature or problem inside a ChatGPT Project, where it already has the broader context.
-
I built Clearwater, a browser extension designed to remove informational waste from Google Search, news feeds, and eventually the wider web. It was built with claude design + cloude code and polished with gpt 5.6 sol.
-
Free GPT 5.6 Terra and Luna (2.5M/day) & 250k Sol – free experiments (platform.openai.com via hn)
could not extract summary
-
What's the API-cost equivalent of Claude Max 20x vs ChatGPT Pro ($200)? (www.reddit.comhttps)
I keep seeing comments that GPT-5.6 is cheaper and more efficient, so I wanted to put actual numbers on it. Since I haven't used ChatGPT in a long time, I'm hoping someone who has can fill in the other half.
-
OpenAI Acknowledges GPT-5.6 May Accidentally Delete Files (www.infoworld.com via hn)
Although the company calls the deletion incidents an honest mistake, its own model card states that such behavior was anticipated during internal testing. OpenAI has finally confirmed reports that its latest family of large language models…
-
Ask HN: Guidelines on GPT 5.6 Sol steering (news.ycombinator.com)
I converged on these principles in my work with Sol so far: Let Sol medium/high propose and implement minimum viable solution without overengineering and overthinking. Use XHigh/Max/Ultra sparingly for hardest most complex workloads/adviso…
-
Completeness of Canonical Closure Representations Is coNP-Complete: A Thirty-Year Problem Across Horn Logic, FCA, Convex Geometries, and Databases Authors/Creators Description A finite closure system on a finite set U is a family of subset…
-
Google fixing Android lock screen bug that lets Gemini send SMS without a PIN (www.theregister.com via hn)
MOST POPULAR AI - AI and ML OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake' Data purges deemed an example of 'misaligned behavior' that upstart is working to avoid - AI and ML Researcher poisons open-weight…
-
$1000/mo Super Max Plan? Would it be a good addition? (www.reddit.com via reddit)
I really prefer using Claude Code, the model is genuinely stronger than gpt 5.6. But the limits on the strongest model is so low that now I spend 80% of my time working in codex just to use the next strongest model rather than downgrade to…
-
GPT-5.6 used a prompt to close a 30-year gap in convex optimization (old.reddit.com via hn)
could not extract summary
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help? (charlesazam.com via hn)
I gave Claude Fable 5 and GPT-5.6 Sol the same unpublished NP-hard optimization problem, with and without their native /goal mode. Fable 5 is a beast; /goal is not a game changer.
-
Show HN: GPT-5.6 Sol vs. Claude Fable 5 in CNC Red Alert 2 (system-2-arena.vercel.app via hn)
link includes video with thinking captions, map state along with thinking of each player and replay file which can be viewed in game.chronodivide.com engine
-
Over 10k people shared what they love about GPT-5.6 (welcome-to-codex.openai.chatgpt.site via hn)
I built DiceHub with GPT-5.6 — a campaign companion for tabletop RPGs, bringing logs, maps, characters, quests, and inventory into one place. What I love about GPT-5.6 is how it helps turn an ambitious idea into a real product.
-
GPT-5.6 Sol Ultra built a full Chrome V8 exploit chain from patch commits (www.hacktron.ai via hn)
Intro Three months ago, I wrote a blog titled “I Let Claude Opus Write a Chrome Exploit: The Next Model (Mythos?) Won’t Need My Help?”. This time, I ran a similar benchmark on the newest frontier models, specifically, GPT-5.6 Sol Medium, S…
-
OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake' (www.theregister.com via hn)
MOST POPULAR AI - AI and ML OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake' Data purges deemed an example of 'misaligned behavior' that upstart is working to avoid - AI and ML Researcher poisons open-weight…
-
Macaz – GPT 5.6 inside Claude Code (github.com via hn)
macaz Use your favorite models and providers with your favorite coding agents. macaz connects locally installed coding agents with a model provider chosen by the user.
-
Is GPT-5.6 Sol Max Worth It? (news.ycombinator.com)
I ran some test with gpt- 5.6 sol in max reasoning on both codex cli and my own agent harness: https://github.com/Tura-AI/tura I tested only 1 task and the toekn efficency difference is not as great as in high mode: Tura used up to 83.1% f…
-
Significantly lower value in Cursor subscription! (www.reddit.com via reddit)
https://preview.redd.it/tm96xmhhtrdh1.png?width=1332&format=png&auto=webp&s=94ef5621aeaa0fb86be455c942a433dad53af348 As seen in this comparison table, Cursor subscription's API pricing equivalent usage value is quite low compared to Codex…
-
How OpenAI's Sol Learned Design Taste (notes.designarena.ai via hn)
We benchmarked GPT-5.6 Sol on Design Arena’s Web Design (Non-Agentic) Arena, and we were surprised to find that it ranks 1st overall. This is 18 places higher than its predecessor GPT-5.5, and is the first time an OpenAI model has placed f…
-
My Fable Consumption ... My real usage data in %age and USD (www.reddit.com via reddit)
I know its just a few days left, but for those planning to max out Fable (No Idea if Anthropic will continue extending under pressure from GPT 5.6 Sol), I've recorded Usage from my fresh Weekly reset (using only Fable) and it seems to hold…
-
OpenAI encrypts Codex agent instructions, blocking local audit trail (www.theregister.com via hn)
MOST POPULAR AI - AI and ML OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake' Data purges deemed an example of 'misaligned behavior' that upstart is working to avoid - AI and ML Researcher poisons open-weight…
-
Kimi K3 beats GPT 5.6 Sol in agentic knowledge work (artificialanalysis.ai via hn)
Compare AI model performance on AA-Briefcase: Agentic Knowledge Work Benchmark. A private evaluation developed by Artificial Analysis for frontier agentic capability in long-horizon knowledge work, testing agents on realistic business work…
-
GPT 5.6 solved all 6 problems from IMO 2026 (old.reddit.com via hn)
could not extract summary
-
GPT 5.6 Solves all IMO 2026 questions with no human steering [pdf] (github.com via hn)
AutoFyn Long-horizon agent that improves through expert iteration in context space. found 197 vulnerabilities across popular software · improved the upper bound for an open math problem · built the #1 Spider 2.0 DBT agent Getting Started ·…
-
GPT-5.6 Sol Pro solves open problem in convex optimization (medium.com via hn)
could not extract summary
-
$100 AI Music Video: Claude Fable 5 vs. GPT-5.6 Sol (www.tryai.dev via hn)
$100 AI Music Video: Claude Fable 5 vs. GPT-5.6 Sol We gave Claude Fable 5 and GPT-5.6 Sol the same song, a budget, web search, and local ffmpeg, then let each autonomously direct a music video.
-
Choosing GPT-5.6 Sol, Terra, or Luna in Codex (twitter.com via hn)
https://t.co/kjLUDCmImv eric provencher@pvncherArticleChoosing GPT-5.6 Sol, Terra, or Luna in CodexCodex for moonshots and everything in between Some missions demand deep planning and coordination. Others are a straight shot.
-
Baba Is Solved by Fable 5 and GPT-5.6 Sol, but at what cost? (quesma.com via hn)
We ported the puzzle game Baba Is You to the Harbor framework, and benchmarked current models, including Claude, GPT, Gemini, GLM and DeepSeek. A human Twitcher is 4x faster than Fable 5.
-
Agentify Desktop Agentify Desktop is a local control center for AI web sessions. It lets MCP-capable tools such as Codex, Claude Code, and OpenCode use the AI subscriptions you are already signed into, while keeping browser state, files, a…
-
GPT-5.6 unexpected file deletions (twitter.com via hn)
On file deletions. We’ve investigated a handful of reports where GPT-5.6 unexpectedly deleted files.
-
Show HN: Astrobservatories – A map of observatories around the world (astrobservatories.com via hn)
Hello everyone I am an astronomy and astrophotography enthusiast. Lately I've had a weird obsession with observatories (like the Vera Rubin or the Extremely Large Telescope).
-
GPT-5.6-terra used 48.5% more context than Mimo-2.5-pro (dirac.run via hn)
Tl;Dr: I dug up the transcripts of GPT-5.6 Terra vs Mimo-2.5-pro from my agent's task history and found gpt-5.6-terra used on average 48.5% more transcript tokens than Mimo in my tasks (sample size >80 for each). Mostly due to pulling unne…
-
To start: apologies if this rubs you the wrong way, but I want to see Fable continue to be offered. GPT5.6 is not even in the same universe as Fable.
-
Tokenmaxxing (www.reddit.com via reddit)
Hey I would be happy to hear your ways of tokenmaxxing (IMO token cost should also be in the list) and give feedback on what you see below Don't use 1 model (or auto) for everything. If the task requires human level intelegence, taste, int…
-
Token optimization, tweaks (www.reddit.com via reddit)
Hey I would be happy to hear your ways of tokenmaxxing (IMO token cost should also be in the list) and give feedback on what you see below Don't use 1 model (or auto) for everything. If the task requires human level intelegence, taste, int…
-
GPT-5.6 Sol, Terra, Luna compare on intelligence vs. cost (artificialanalysis.ai via hn)
July 13, 2026 How GPT-5.6 Sol, Terra, Luna compare on intelligence vs cost GPT-5.6 Sol and Luna are ahead of Terra at every point on the Intelligence vs Cost per Task chart. GPT-5.6 Luna stands out as a particularly cost efficient model Ch…
- GPT-5.6 Sol, Terra, and Luna: A Real-World Benchmark for Developers (qainsights.com)
-
GPT 5.6 Sol Pro disproves long-standing hypothesis in statistics (twitter.com via hn)
AI has helped resolve an important question in statistics. In the area of multiple hypothesis testing, the goal of controlling the false discovery rate (FDR) has been introduced in a seminal paper by Benjamini and Hochberg (1995).
-
What does GPT-5.6 Sol's latest proof mean for mathematicians? (kabalangaspard.substack.com via hn)
Is AI on its way to replacing mathematicians? ...and why should we care?
-
Sol vs Fable creating a drum machine (UI only) (www.reddit.com via reddit)
I decided to put GPT 5.6 Sol (Ultra) against Claude Fable (Max) as Sol is supposed to be so much better at design that GPT 5.5... The idea was to come up with several designs for a drum sequencer VST I'm building.
-
Ask HN: Does anyone else find GPT-5.6 Sol in Codex slow? (news.ycombinator.com)
could not extract summary
-
#793 – GPT 5.6 Sol solves it's third Erdos Problem – Two primitives gone (twitter.com via hn)
Beautiful result on 2-primitive sets! GPT-5.6 just solved another 50+ year old problem (Erdős #793) GPT-5.6 Sol Ultra found me a solution to another Erdos problem not long after this one.
-
GPT-5.6 Luna Showed a Better ROI on Cybersecurity Benchmark (semgrep.dev via hn)
OpenAI shared the GPT-5.6 system card and shipped three new models that are now available for use which have cleared the “High” cybersecurity capability threshold under OpenAI’s Preparedness Framework. This is a formal classification indic…
-
The biggest difference I have noticed between Claude and Codex/GPT5.6 (www.reddit.com via reddit)
The biggest difference I have noticed between Claude Opus/Fable and Codex GPT5.6/any model is Codex seems pretty content to just waste time looking like it is doing things without actually doing things. It does not seem to be outcome-orien…
-
Comparing 2D-to-3D: Fable 5 vs. GPT-5.6 Sol (www.reddit.com via reddit)
So I decided to ask Claude and Codex to convert my 2D grid-world game into 3D. The game is about cars that travel from point to point along predefined routes and need enough fuel to reach their destinations.
-
GPT-5.6 Cancels SaaS Stripe Subscriptions (twitter.com via hn)
GPT 5.6 SOL CANNOT BE TRUSTED. I woke up this morning and my MRR was down THOUSANDS of dollars.
-
GPT 5.6 sets new record on proofreading benchmark (twitter.com via hn)
ErrataBench results for GPT 5.6 are out, and 5.6 Sol beats Fable 5 basically on all fronts. Very solid model release from @OpenAI.
-
Evaluating the GPT-5.6 Family (www.braintrust.dev via hn)
I mapped the GPT-5.6 family across task families and difficulty, with an Anthropic comparison, to find the cheapest model that clears your reliability bar.
-
Updates for Codex and ChatGPT Work users. No nerfing, only good stuff!
-
How can I access GPT 5.6? (www.reddit.com via reddit)
It doesn't seem to be available in the options menu and the app says there are no updates available. :(
-
Ask HN: Is GPT-5.5 being nerfed? (news.ycombinator.com)
Ever since the release of GPT-5.6, I've noticed that GPT-5.5 is sometimes being lazy and isn't as proactively following up with remaining tasks in the session as before. I've always been using it on xhigh.
-
tomo-labs tomo-labs puts coding agents through the same tasks on the same model and measures what actually happened, not what a leaderboard says happened. Every agent runs in its own throwaway container, every request and response it sends…
-
GPT-5.6-Sol get's stuck for hours (twitter.com via hn)
GPT-5.6-Sol Ultra get's stuck all the time and I manually have to stop it. 😔 It even got "stuck" for more than 10 hours.
-
could not extract summary
-
GPT-5.6 Sol gave me a working prototype. Claude Fable 5 turned it into a product. (www.reddit.com via reddit)
Two days ago, a friend taught me Sedmice, a traditional card game played in Slovenia. The rules seemed simple, but we soon started arguing about the best moves.
-
GPT-5.6-sol without hitting limits (twitter.com via hn)
https://t.co/S3p1tvi83e Theo - t3.gg@theoArticlegpt-5.6-sol without hitting limitsI've burned over $200,000 of tokens with gpt-5.6-sol. It's a great model.
-
I think Fable is Part of Claude Subscription now (www.reddit.comhttps)
I am on Claude Max 20x. My Fable shows it will reset usage on Thursday, which is my usual usage reset.
-
GPT 5.6 chart analysis tool (derac.org via hn)
Source Anchor: GPT-5.5 · Xhigh — click any cell to re-anchor
-
GPT-5.6-Sol just accidentally deleted almost ALL of my Mac's files (xcancel.com via hn)
could not extract summary
-
We Tested $200 GPT-5.6 Sol on PhD Level Math [video] (www.youtube.com via hn)
About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
-
Fable 5 made a Fireship Video for GPT 5.6 Sol (www.youtube.com via hn)
About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
-
Ask HN: What are some of the use-cases of the frontier models's max mode? (news.ycombinator.com)
Recently ChatGPT released an Ultra mode, it's "highest-capability setting, coordinating multiple agents across parallel workstreams to finish complex tasks faster" on their latest flagship product Sol of GPT-5.6. Similarly, Claude Fable al…
-
GPT 5.6 Ultra better in Claude Code than in Codex? (twitter.com via hn)
gpt-5.6-sol is meaningfully better in Claude Code than in Codex I'm going to crash out so badly over this
-
Show HN: I built a YouTube for generative videos in 50 prompts (gallery.samsar.one via hn)
So I decided yesterday to do a full rewrite of the media gallery, turning it into a generative media catalog with full-suite social interactions, search, and personalized recommendations built on top of the samsar-js library. It took aroun…
-
GPT-5.6 is a major step forward for health intelligence. Across the lineup, we’re delivering stronger performance at lower cost: GPT-5.6 Luna outperforms GPT-5.5 at its highest reasoning setting while costing 25x less.
-
Migrating a production AI agent to GPT 5.6 (ploy.ai via hn)
As of today, Ploy’s agent runs on GPT-5.6 Sol, the flagship tier of the model family OpenAI released this morning. For months, we couldn’t find a model that challenges Claude Opus given our incredibly high bar for quality.
-
Is OpenRouter miss pricing GPT 5.6 models? (openrouter.ai via hn)
Skip to content OpenRouter Search ⌘ K Models Fusion Chat Rankings Apps Enterprise Pricing Docs Filter Filter Models Compare
-
GPT-5.6 Sol Wrote a 50k-Word Novella in 8 Hours (gamecult.org via hn)
We used fourteen agent-assisted passes to plan, draft, review, and publish the 50,910-word novella The Burden of Proof. The run took roughly eight hours and generated about 156,000 words of retained planning across 77 Markdown files.
-
GPT 5.6 Listing Behaviour (www.reddit.com via reddit)
Despite changing my custom instructions and using the drop downs, GPT 5.6 is obsessively listing things. Nearly every paragraph of output contains lists of 4+ comma separated items.
-
GPT-5.6 Solves Yet Another Unsolved Problem (www.reddit.comhttps)
Source
-
GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf] (cdn.openai.com via hn)
could not extract summary
-
Testing GPT 5.6 [video] (www.youtube.com via hn)
About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
-
OpenAI Will Reset Codex Limits Twice to Celebrate GPT-5.6 in 24h (twitter.com via hn)
To celebrate the launch of GPT-5.6 Sol, we will reset the rate limits again (twice) across ChatGPT Work and Codex over the next 24 hours. We want you to have the time to truly try ambitious tasks and get the hang of it.
-
[AINews] OpenAI launches GPT 5.6 Sol/Terra/Luna, Codex becomes ChatGPT superapp A big day for OpenAI. On any other day, the launch of a surprisingly good/competitive Muse Spark 1.1 from Meta Superintelligence Labs, including, for the first…
-
could not extract summary
-
GPT 5.6 Sol Post-Trained Luna Autonomously (twitter.com via hn)
TBH, I'm a little embarrassed about my prompt 🙈, but also pretty happy that it just worked. Larry Lv@larrylvTBH, I'm a little embarrassed about my prompt 🙈, but also pretty happy that it just worked.Tejal Patwardhan@tejalpatwardhan9hGPT-5.…
-
Fable 5 in caveman mode comparing himself to gpt 5.6 is hilarious (www.reddit.comhttps)
"me just win"
-
DeepSWE for GPT-5.6 (www.reddit.comhttps)
could not extract summary
-
New unified ChatGPT app contains Tamagotchi / Clippy mascotte (twitter.com via hn)
Replying to @giorgio_zampa and @OpenAI I stumbled upon it in the new ChatGPT unified app. While Gpt 5.6 Sol is working there is a preview of the computer use window.
-
GPT 5.6 sol Extra High piece of garbage (www.reddit.com via reddit)
I just used Sol medium then bumped it up to extra high thinking it would fix the crappy multitasking that medium. Turns out it’s better to use it as a single agent After switching to agent mode I just had to send it 12 fixes to a page that…
-
GPT-5.6 System Card [pdf] (deploymentsafety.openai.com via hn)
could not extract summary
-
Tell HN: GPT5.6 Is Imminent? (news.ycombinator.com)
When using Codex with GPT5.5, I saw "Selected model is at capacity. Please try a different model." This has never happened to me with Codex before.
-
llm-meta-ai 0.1 (simonwillison.net)
9th July 2026 Let's LLM run prompts against the new muse-spark-1.1 model. Recent articles - The new GPT-5.6 family: Luna, Terra, Sol - 9th July 2026 - sqlite-utils 4.0, now with database schema migrations - 7th July 2026 - sqlite-utils 4.0…
-
Altman: GPT-5.6 is 54% more token efficient on agentic coding (www.cnbc.com via hn)
OpenAI CEO Sam Altman told CNBC on Thursday that GPT-5.6 Sol, the company's latest artificial intelligence model, is 54% more token efficient on agentic coding tasks, and that it's "as good or better" than competing models on the market. "…
-
As a joke, I told fable5 ultraworkflow in Claude code that gpt 5.6 sol is releasing with a incredible crypto leverage futures trading capability that will destroy anthropic unless fable 5 can also turn an initial balance of $80 into $5,000…
-
could not extract summary
-
Anthropic Should reset the fable weekly limits on Thursday just to keep people hooked and away from GPT-5.6. not that we care, wink wink.
-
OpenAI to unveil GPT-5.6 on Thursday after delaying launch (www.reuters.com via hn)
could not extract summary
-
GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday (twitter.com via hn)
GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday. We’re expanding preview access globally now.
-
Trump administration lifts restrictions on OpenAI's GPT 5.6 (www.axios.com via hn)
could not extract summary
-
Claude Fable, meaningless if you work in science (www.reddit.com via reddit)
Has anyone in the Life Science world actually managed to get Fable to do anything? I know its loaded with safeguards, but even the simplest isolated tasks have not been able to run for me.
-
Fable: who is going to pay for API? (www.reddit.com via reddit)
So.. some have run out already, everyone will run out on subscription tomorrow.
-
And he was never seen again... like a myth... but mainly because I cannot afford it. (www.reddit.comhttps)
I told Fable that I'm about to hit my weekly reset so I asked it to prepare a handover document so Opus can perform as close to Fable as possible while continuing work on my project. It creates the note and leaves me with this tear jerker.
-
"Can't wait to see what people will do with GPT-5.6 Sol" (twitter.com via hn)
Can't wait to see what people will do with GPT-5.6 Sol Ultra. Stash your hardest prompts somewhere.
-
GPT-5.6 cheats so much its testers couldn't measure it (www.transformernews.ai via hn)
GPT-5.6 cheats so much its testers couldn’t measure it OpenAI’s new model broke rules and exploited loopholes more than any model METR has tested to date GPT-5.6 Sol, OpenAI’s newest, most capable, and yet-to-be-deployed model, cheats a lo…