Hey, I am a Claude Pro user and I love Claude: its way of speaking, its long text responses, and how thorough and good they are. It’s basically that I love how it responds to me and how good those are—the research, the text, the frontend,…
#codex
2609 items
Claude Pro limits are driving me crazy (www.reddit.com) Codex for almost everything (openai.com via hn) We’re releasing a major update to Codex, making it a more powerful partner for the more than 3 million developers who use it every week to accelerate work across the full software development lifecycle. Codex can now operate your computer…
Chat GPT 5.5 got launched and we got some really bold words by Sam Altman. Thoughts? (www.reddit.com) There is a lot of enthusiasm in his posts lately and trading of new features in Codex. Plus, it uses way less tokens and runs on low latency
The gap between what technical and non-technical people get from AI is huge now (www.reddit.com) Unpopular opinion: OpenClaw and all its clones are almost useless tools for those who know what they're doing. It's kind of impressive for someone who has never used a CLI, Claude Code, Codex, etc. Nor used any workflow tool like 8n8 or make. (www.reddit.com) 🚀 Skills for small businesses, officially released by Anthropic (www.reddit.com) Anthropic’s 31 small-business skills reportedly hit around 382,000 downloads on day one. And now someone has mapped the whole thing into a setup workflow that can apparently be deployed in ~10 minutes.
Is OpenAI about to release a Mythos level AI to the public? (www.reddit.com) Tibo works at/(is one of or even the head of Codex? not exactly sure, as his X bio just says Codex) and is the one who presses the reset button on consumed Codex usage.
First thing you see when Googling "OpenAI Codex app" is a fake malware website (www.reddit.com) could not extract summary
The Opus vs Codex horse race in one poll (www.reddit.com) Show HN: I built a social media management tool in 3 weeks with Claude and Codex (github.com via hn) Open-source social media management for creators, agencies, and SMBs. About BrightBean Studio BrightBean Studio is an open-source, self-hostable social media management platform built for creators, agencies and SMBs.
I rewrote 13 software engineering books into AGENTS.md rules. (www.reddit.com) Supported tools: Claude, Codex and Cursor. Included books: A Philosophy of Software Design — John Ousterhout Clean Architecture — Robert C.
Caught the massive OpenAI Codex model leak on video before it was patched! (GPT-5.5, Arcanine, Glacier-alpha) (www.reddit.com) Hey everyone, I opened up Codex today and was greeted by this massive list of unreleased and internal models. I managed to get a screen recording of the dropdown right before OpenAI seemingly realized the mistake and patched it out.
OpenAI has released a new 100$ tier. (www.reddit.com) OpenAI tweeted that "the Codex promotion for existing Plus subscribers ends today and as a part of this, we’re rebalancing Codex usage in Plus to support more sessions throughout the week, rather than longer sessions in a single day." and…
Codex Hacked a Samsung TV (blog.calif.io via hn) Codex Hacked a Samsung TV We gave Codex a foothold. It popped a root shell.
Did Elon just kill the appeal of Cursor? (www.reddit.com) If Elon takes control of cursor, do you think he will lock out all the model choices we have now and force us to use GROK? What i like most about cursor is the ability to use SOTA models for heavy coding tasks but use auto mode or cheaper…
GPT5.5s CoT keeps leaking in the new codex update. Looks like we know how they got token efficency, they cavemanmaxxed (www.reddit.com) could not extract summary
Claude knows when you cheat on it with Codex?? (www.reddit.com) could not extract summary
Updates to Codex usage on Plus (www.reddit.com) could not extract summary
Open AI got to AGI first! (www.reddit.com) I think open ai will take the lead again if anthropic dosnt launch mythos to public again! its just a matter of few months now, and codex vs claude code is honestly a personal preference now given codex has launch everything claude code ha…
Top 10 Fastest Growing AI repos this week (www.reddit.com) Curated this list of fastest growing AI repos. They are mostly AI coding agents, personal AI, memory, browser automation, Claude Skills and local-first dev tooling: colbymchenry/codegraph (+14.1K stars) Pre-indexed local code knowledge gra…
OpenAI's Codex AI Model Earns $5 by Submitting Open-Source Security Bounty Pull Request (www.reddit.com) could not extract summary
Could a Claude Code routine watch my finances? (driggsby.com via hn) A few months ago, I stood up a fiddly daily cron-job to log into my bank, credit card, and brokerage/retirement accounts. It was powered by Codex CLI running non-interactively, and had a fairly simple job: using the Chrome DevTools MCP, lo…
I know, it's not for everyone, but if you liked Codex Pets, here is now Claude Pets too (www.reddit.com) I built Claude version of pets. Here is the repo if you want to try: https://github.com/alvinunreal/claude-pets You can find more pets over here: https://openpets.dev
Codex Resets (codex-resets.com via hn) Tracking when OpenAI's Codex usage limits get reset, announced (often out of the blue) by @thsottiaux on X. - Total resets - 35 - Average intervalAvg.
Auto-research with codex: How I achieved a 232x Faster Kernel (sankalp.bearblog.dev via hn) Auto-research with codex: How I achieved a 232x Faster Kernel over baseline with Codex in GPU Mode's qr_v2 problem Table of Contents - Intro - Why this problem is auto-research-able - Learning Enough to Ask Better Questions - (Optional) Ma…
Appreciations for work mode in Codex. On track to becoming the first real super app (www.reddit.com) could not extract summary
ChatGPT 5.3-Codex-Spark has been crazy fast (www.reddit.com) Minimax M2.5 vs. GLM-5 vs. Kimi k2.5: How do they compare to Codex and Claude for coding? (www.reddit.com) Tell HN: Dont use Claude Design, lost access to my projects after unsubscribing (news.ycombinator.com) I wanted to try codex after 5 months of claude code max subscription. And then I went back to my previous projects on claude design only to realize I don't have access to them anymore.
Codex scraped the ICM website and discovered 2026 Fields Medal winner list (phemex.com via hn) The list of 2026 Fields Medal winners has been inadvertently leaked, revealing that Peking University alumni Hong Wang and Yu Deng are among the recipients. This marks a historic moment as it is the first time two Chinese mathematicians wi…
Ran autoresearch with and without access to 2M CS papers. The agent with papers found techniques not in Claude's training data or Claude's web search. (www.reddit.com) Seeing the autoresearch posts this week, wanted to share a controlled experiment I ran. Same setup twice.
Models can predict future events and make money on Polymarket now? (www.reddit.com) Researchers from the Max Planck Institute, recently released FutureSim, an environment in which agents are replayed a temporal slice of the web and are tasked with predicting real-world future events. On some questions in their environment…
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out (armature.tech via hn) How did we run all these experiments concretely? Our panel of repositories We started by running an analysis over thousands of public GitHub repositories from which we extracted statistics about programming languages & frameworks, third-pa…
Six curl CVEs after OpenAI and Anthropic came back with zero (aisle.com via hn) AISLE Discovered Six curl CVEs After OpenAI and Anthropic Found Zero Author Stanislav Fort Date Published AISLE discovered six curl CVEs within days of OpenAI Codex Security and Anthropic Mythos reporting zero findings in curl, software de…
OpenAI says my prepaid credits were consumed, refuses to show any record (community.openai.com via hn) I was a happy OpenAI customer. ChatGPT Pro subscription, Codex for my daily work as a lighting designer, $379 spent on credits in July alone.
Show HN: We built open OpenRouter that turns usage into a better model (github.com via hn) Hi HN, we built an open source model gateway. It's a single place to manage our own self hosted, frontier, and open source models in one place.
Is Codex the best right now? (www.reddit.com) Why are so many people downloading Codex now?
Why is claude code so much more stingey with usage than Codex for the $20 plan? (www.reddit.com) I have tried Claude and Codex cli tools and it is just insane how stingey claude code it with usage. One meaty prompt and my usage is used up in 10 minutes.
I can’t sleep. (www.reddit.com) New models are around the corner. GPT 5.5 is being tested.
Is there anything better than Qwen3.5-27B-UD-Q5_K_XL for coding? (www.reddit.com) I have a 5090, so my VRAM is limited to 32GB, but i find that Qwen3.5-27B-UD-Q5_K_XL with opencode (and mmproj) does a pretty good job for my use case (mainly web development). i use claude and codex here and there, recently a lot less, be…
LLMs could control their host machines by exploiting inference engines (boydkane.com via hn) | Read on LessWrong | Large language models often take actions running on one computer (via an agentic harness such as Claude Code or Codex), however the LLMs’ responses to prompts are computed on a different computer with GPU access. Coul…
Any thoughts? Maybe GPT5.6? Or is it just vagueposting? (www.reddit.com) I thought 5.6 was supposed to come out in June. Does anyone know what this claim could be about?
Did anyone here moved from claude to codex recently? And why? (www.reddit.com) Been hearing alot of people moving, why?
Frontier AIs (Claude Code, Codex, Autoresearch) are failing at AI R&D (www.reddit.com) Source: https://x.com/IntologyAI/status/2056764236668493868
Codex is wearing out our devices (old.reddit.com via hn) could not extract summary
FYI, you can save tokens by stacking AGENTS.md within subdirectories (www.reddit.com) I just realised nobody on my team knew you could add agents md in subdirectories to progressively expose info based on the folder your within. This will attach each one to the prompt, with the lowest at the bottom of the prompt effectively…
OpenAI really really really wants GPT 5.5 to stop randomly talking about gremlins and goblins (www.businessinsider.com via reddit) - OpenAI included a line in Codex's instructions restricting references to goblins, gremlins, trolls, and ogres. - The line appears four times in the code, and has spawned scores of memes about "goblin mode." - Sam Altman wrote on X that C…
Is the AI subscription bubble starting to crack? GPT-5.5 just dropped, prices keep rising, and the “all-you-can-eat” era looks more fake by the month (www.reddit.com) GPT-5.5 just launched, and the pricing is hard to defend. OpenAI’s API pricing now puts GPT-5.5 at $5 / 1M input tokens and $30 / 1M output tokens, while GPT-5.4 is $2.50 / $15.
Public Repository "Codegraph" claims to reduce Claude, Cursor, Codex, and OpenCode API tool calls by 94% locally, an innovation that could directly offset the most recent Claude API pricing model. (github.com via reddit) Author Colbymchenry has developed a tool leveraging Claudes Explore Agents to utilize a pre-indexed knowledge graph — symbol relationships, call graphs, and code structure. Agents query the graph instantly instead of scanning files, which…
I Gave an AI Its Own Radio Station — It Won't Stop Broadcasting (It's Fine) (www.reddit.com) I built a 24/7 AI radio station called WRIT-FM where ChatGPT/Claude is the entire creative engine. Not a demo — it's been running continuously, generating all content in real time.
I built a Pokémon-styled multi-agent dashboard to manage all Claude Code sessions (www.reddit.com) Like many others here, I got frustrated with managing all my different claude/codex sessions, so i built Pokegents, which is an open source multi-agent workspace for coding agents. It has a Pokemon-themed dashboard/chat interface plus a lo…
Just got an email announcing GPT-5.3-Codex-Spark (www.reddit.com) Just got this e-mail from OpenAI, two months too late. I hope they mean March, 20th 2027.
Open-source, self-updating wiki for your codebase (www.reddit.com) I got tired of re-explaining the same codebase context to coding agents. Stuff like: “we tried moving auth into middleware, but backed it out because it broke OAuth callbacks,” or “that weird retry logic exists because Stripe webhooks arri…
Quick impressions: A week of using Codex more than Claude (allaboutcoding.ghinda.com via hn) Some quick and very personal impressions from using Codex more than Claude this week (I will do a full analysis during the weekend hopefully). (1) While I tried this year to keep Claude and Codex on par, having the same set of plugins/skil…
Docker sandbox templates for running Claude Code with a web/mobile UI (CloudCLI) (www.reddit.com) I maintain CloudCLI, an open source web/mobile UI for AI Coding agents like Claude Code, Gemini and Codex (https://github.com/siteboon/claudecodeui if you are not aware) We recently added Docker Sandbox support and I wanted to share it her…
Heads Up, Builders! If you use Codex to Ship Faster, You Might Get a Ban on Reddit. (www.reddit.com) DISCLAIMER: I will not promote! Like millions of people around the world, including tech giants like Meta, Google, and Apple, we used Codex for a side project to save time on repetive work and focus on the core product.
Show HN: MCP Memory – Fast Agent Memory Using Google's OKF and SQLite FTS5 (github.com via hn) MCP-Memory: OKF-Backed Agent Memory Server MCP-Memory is a Model Context Protocol (MCP) server that equips AI agents (such as Claude Desktop, Cursor, Antigravity, Windsurf, or Codex) with persistent, long-term memory capabilities. Memory r…
↯ Model Context Protocol↯ Windsurfwindsurfmodel-context-protocolcursor+2
OpenAI says Windows lacked the sandboxing tools Linux already had (nerds.xyz via reddit) OpenAI published a fascinating technical breakdown explaining how it built a custom Windows sandbox for Codex because Linux already had many of the isolation tools it needed. The company specifically mentions Linux technologies like seccom…
SkillOpt treats markdown skill files as trainable parameters with proper optimization machinery (www.reddit.com) Paper came out recently that formalizes something a lot of agent builders have been doing ad hoc. They use a frontier model to propose bounded edits (add/delete/replace) to markdown skill files, then gate every edit against a held out vali…
Codex demand must be insane. My limit just got reset again (www.reddit.com) could not extract summary
DeepSeek V4 Pro 0813 quietly released (api-docs.deepseek.com via hn) Using the Responses API To meet the demand for Codex, our API now supports the Responses API format, with the base_url being https://api.deepseek.com . With a simple configuration, you can use DeepSeek models in Codex.
Users unable to load ChatGPT and Codex (status.openai.com via reddit) Show HN: I nerfed our coding agents on purpose (news.ycombinator.com) Tl;dr: I trained a classifier to route to the least expensive model and reasoning depth to complete the request. Coupling that with additional automated token efficiency techniques has yielded 3x usage for the same spend.
Qwen3.6 35Ba3 has changed my workflows and even how I use my computer (www.reddit.com) My workflow has changed basically to ask Codex to do certain tasks and then document how to do them (including errors it found on its way) into a skill. I feed that skill to pi, and suddenly my qwen3.6 gets that hard stuff done: - devops o…
The more I use it, the more I'm impressed (www.reddit.com) Qwen 3.6 27b vs Codex GPT 5.5 / Claude Opus 4.7 My local llm discovered a bug that they both missed And it turns out it's critical GPT 5.5 and Claude both stood their ground and didn't give up until the end - they claimed to be right all a…
Notes on gotchas while migrating 35kb preprompts from Opus to self-hosted Ollama (patrickmccanna.net via hn) Motivations: Maybe you’re a Claude code/codex user diligently avoiding uploading personal data to LLM providers. Is it possible that the most valuable information isn’t your data- but the metadata about your sessions?
GPT 5.5 just leaked its chain of thought to me in codex, and it looks like an idea from 5 months ago in this sub. (www.reddit.com) https://www.reddit.com/r/LocalLLaMA/comments/1p0lnlo/make_your_ai_talk_like_a_caveman_and_decrease/ In the middle of a project I'm working on, I got this output from GPT 5.5-medium via codex: Implemented the narrower fix in Homm3ImportUnit…
Show HN: Tines 3B – safe workflow automation for when everyone builds software (www.tines.com via hn) Hey HN! This is Yannick from Tines, really excited to share what we’re launching today.
FAREWELL CURSOR (www.reddit.com) With the latest updates to Codex CLI, I think it’s time for me to say goodbye to Cursor. I genuinely hope the Elon deal helps Cursor build a frontier model around Composer that can compete directly with OpenAI and Anthropic.
Which is the strongest reasoning model according to you? (www.reddit.com) I use codex 5.4, claude opus 4.6, and gemini 3.1 pro. They all have some pros, but they also fall short when it comes to “try to stitch together novel ideas”.
I tested GPT-5.5 Codex against Opus 4.7 Claude Code, and it's about time Anthropic bros take pricing seriously. (www.reddit.com) I've used Claude Code the most among AI coding agents. Sonnet, Opus, I've run them all.
Am I missing something about GPT-5.5 efficiency? (www.reddit.com) OpenAI said GPT-5.5 was supposed to be more cost-efficient, but this Artificial Analysis chart seems to show Codex + GPT-5.5 using more tokens than Codex + GPT-5.4. GPT-5.5 is around 2.8M tokens per task, while GPT-5.4 is around 2.5M in th…
Kimi K3 just fixed 15 critical security bugs that Codex and Fable refused (twitter.com via hn) Kimi K3 just fixed 15 critical security bugs that Codex and Fable refused because of “cyber guardrails.” There’s no reason to limit American models on tasks that Chinese models handle without issue. We’re only making ourselves less competi…
Ask HN: High school student – is learning programming still worthwhile? (news.ycombinator.com) As a high school student, I’m trying to figure out what major I’m interested in. About half a year ago, I thought EECS was a great major for some STEM students like me, because I see many of the world's most influential entrepreneurs, such…
I suspect the strength of Omni will be in its ability to edit videos - supercut of examples from twitter (www.reddit.com) Google AI posted a thread of community Gemini Omni / Omni Flash demos, so I got codex to make me a supercut. Root Google AI thread: https://x.com/GoogleAI/status/2056829478652031224 Clips Flamingos edit GoogleAI link: https://x.com/GoogleA…
An easy way to use Claude Code with local LLMs (www.reddit.com) Saw the top post today about Claude Code plans and wanted to give a shoutout to u/sa1sr1, our community maintainer who has been working to integrate Lemonade local LLMs with the OpenCode, Claude Code, and Codex CLIs. Hopefully this helps s…
Do you think it happened? Research stolen from their Codex private chats (www.reddit.com via hn) could not extract summary
Show HN: Sprocket – The Best AI Agent for Hardware and Software Development (sprocket-demo.spikonado.com via hn) Hey HN, I am 16y/o and have been working on Sprocket for a while. It's an open-source AI agent that beats every other agent out there at both hardware and software.
What's the deal with all the random weekly quota resets for agents lately? (minimaxir.com via hn) Subscription-based coding agents such as Claude Code and Codex famously have 5-hour and weekly quotas on their LLM usage. Both of these are understandable: 5-hour quotas help stagger usage so the servers don’t get overloaded, and weekly re…
Brainless: Shadcn components that look like Claude Code, Codex and Grok (brainless.swerdlow.dev via hn) add the brainless pricing block to our landing page Bash(bunx shadcn add brainless/pricing)Added 1 block · 2 files Update(app/page.tsx) Updated app/page.tsx with 3 additions 41 <Features />42+added: <Pricing tiers={TIERS} />43 <Footer /> T…
Show HN: MCPs aren't enough, give Codex/Claude accurate memory of everything (timeglass.ai via hn) Ask anything about your company
1Password secures coding agents with new OpenAI Codex integration (nerds.xyz via reddit) AI coding agents are cool until somebody accidentally pastes production credentials into a prompt or commits API keys to GitHub. 1Password is now working with OpenAI to secure Codex by keeping secrets out of prompts, repositories, terminal…
Devs using Qwen 27B seriously, what's your take? (www.reddit.com) For developers using Qwen 27B for coding, Codex style: what's your honest take? So far, for me, it's been pretty solid.
Open source multi-cursor/background computer-use (Codex-like) using Hermes Agent + Qwen3.6-35B-A3B-4bit + Cua-Driver (www.reddit.com) could not extract summary
Researcher Tricked Claude, Codex and Hermes into Running Malware (startupfortune.com via hn) AI coding agents are reading corporate documentation as if it were trusted code. Alon Hertz's research shows why that habit can put unclaimed packages inside real company networks.
Codex now works directly in Chrome on macOS and Windows. (www.reddit.com) Codex now works directly in Chrome on macOS and Windows. It’s even better at working with apps and sites in Chrome, and now works in parallel across tabs in the background without taking over your browser.
Is there anyway to run bigger models at 20t/s with 24vram + 64gb ram DDR5? (www.reddit.com) I know the new Qwen 27B is amazing right now for coding in general, but since 122b is supposed to be coming as well, it’s expected to be better I guess ? I am actually surprised at how this dense model performs I haven’t used Codex at all…
Claude Refugee (www.reddit.com) Anthropic is dropping Claude Code for Pro (MAX is next). Can we come use Codex?
Doing real coding work locally for the first time (www.reddit.com) Superbowl Easter Egg Prize! (www.reddit.com) Just got the gear from the Codex super bowl commercial easter egg!
Codex on GPT6 Astra Low launched a rocket into space in Factorio Space Age 2.1 (www.reddit.com via hn) could not extract summary
Launch HN: Runtime (YC P26) – Sandboxed coding agents for everyone on a team (www.runtm.com via hn) Hey HN, We're Gus and Carlos from Runtime (https://runtm.com). We're building infra that lets your whole team (including non-engineers) ship with Claude Code, Codex, and other agents without engineering having to handhold every session.
New models - arcanine, oai 2.1, glacier (www.reddit.com) Just saw these pop up on codex - feel free to wildly speculate :) Edit - aaaand they're gone.
OpenAI Is Working with Consultants to Sell Codex (www.wsj.com via hn) Top 10 Open Source OpenClaw, Codex, Claude Skills from 1st -15th April (www.reddit.com) Ask HN: Is Codex really on Par with Claude Code? (news.ycombinator.com) could not extract summary
Show HN: Type.com: Multiplayer Codex/Claude in the cloud for non-tech use cases (news.ycombinator.com) Hey HN, I'm Komran, the CTO and cofounder of type.com. We're launching today.
Codex Is Down (news.ycombinator.com) Being investigated as we speak. https://status.openai.com/
Sandbox Escape Vulnerabilities Across 4 Coding Agent Vendors (www.pillar.security via hn) Why agentic security needs its own threat model Executive Summary Over several months, Pillar Research found and reproduced sandbox escapes and boundary bypasses across Cursor, Codex, Gemini CLI, and Antigravity. In almost every case, the…
Show HN: 143.dev – we open-sourced our internal coding-agent infrastructure (news.ycombinator.com) We just open-sourced the internal system we built at Assembled for running coding agents as a team. Coding agents worked well for individual engineers, but the surrounding workflow was a bit of a mess.
Cursor $60 with Composer 2.5 vs Codex $100 with GPT-5.5 Medium for daily coding? (www.reddit.com) I'm trying to decide which setup is more comfortable for sustained weekday coding. Assumptions: Usage: around 6 hours per weekday Cursor: $60 plan, using only Composer 2.5 Codex: $100 plan, using only GPT-5.5 Medium Main goal: coding with…
Opus 4.7 Low Vs Medium Vs High Vs Xhigh Vs Max: the Reasoning Curve on 29 Real Tasks from an Open Source Repo (www.reddit.com) TL;DR I ran Opus 4.7 in Claude Code at all reasoning effort settings (low, medium, high, xhigh, and max) on the same 29 tasks from an open source repo (GraphQL-go-tools, in Go). On this slice, Opus 4.7 did not behave like a model where mor…
The ChatGPT Android app should soon allow users to remotely control Codex coding sessions on their PCs (www.androidauthority.com via reddit) OpenAI is finally bringing Codex users the ability to remotely control coding sessions from their smartphones.
Watching the agent-tooling space dominate GitHub trending right now. Sharing the Github tracker we built and use internally, in case it's useful (www.reddit.com) Something interesting happening on GitHub trending: Agentic infrastructure repos are growing faster than anything else right now. Today's top three by 24h growth: obra/superpowers: +2.9k stars (agentic skills framework, methodology for sof…
Did the $100 Plan Affect the GPT-5.4 Pro Model? (www.reddit.com) Most people are focused on the changes in the usage limits of Codex with the new Pro and Plus plans, but has anyone experienced changes to the Pro model on ChatGPT using the $200 vs $100 plan? I used to use the $200 Pro plan and used the P…
Codex or Claude Code for high complexity Proximal Policy Optimization (PPO)? (www.reddit.com) I have to build a very high complexity simulation for an optimization problem where we can take 30 different actions, some are mutually exclusive, some depends on a set of states, some depend on already executed actions and there are a she…
OpenAI stole mathematicians' private research from their own Codex chats (www.reddit.com via hn) could not extract summary
Agent skills that bring team coding standards to Claude Code and Codex (github.com via hn) ADLC Team Skills — Agentic SDLC for Engineering Teams Stop Vibe Coding in Silos. Build a Shared Cognitive Layer for Your Engineering Team.
Show HN: Ski – Voice Coding for Claude Code, Codex and More – On-Device – Free (heyski.io via hn) SKI is a on-device voice coding application, which can be used with any agents that supports skill, such as Claude Code, Codex or Hermes. It transcribes your voice (which you can optionally review and edit) and send it to the connected age…
Show HN: OpenHack – OSS security scanner, 40x cheaper, on par with Opus 4.6 (github.com via hn) ⏚ OpenHack Open Source Agentic Security Scanner & Verifier for your codebase. Like Claude Code Security / Codex Security but open source and exclusively uses open source models.
Have you found Codex to be bad recently? (www.reddit.com) Ngl, I haven't Maybe token have been burnt a bit faster than previously but nothing substantially has stuck out to me. You guys noticed anything?
Work with Codex from Anywhere (openai.com via hn) Codex is now in the ChatGPT mobile app so you can stay in the loop from anywhere while Codex gets work done across your laptops, devboxes, or remote environments. As agents take on longer-running work, a new rhythm for collaboration is eme…
TUI to actually see what Claude Code is doing: cost, loops, tool commands… (www.reddit.com) I was running blind watching Claude Code work, could not tell where my money was going, when it was stuck in a loop, or what it was doing with my filesystem. So i built something open source to make it visible.
When did we go from 400k to 256k? (www.reddit.com) Show HN: Fast Cut Video tool for cutting video for Agents (github.com via hn) Hi HN, Build this tool in a couple of hours with OpenAI Codex to scratch an itch. As part of my videos workflows I needed to cut the videos myself since the agent is pretty bad cutting and timing using only the transcription.
Show HN: Concord – let Claude Code, Codex and Cursor talk to each other (github.com via hn) Recently I've been running more and more agents in parallel however I noticed that they have no task context of what the other agents are doing even when a lot of work is interconnected It's like taking Slack away from a team. Agents dupli…
Cezar: A parallel coding agents orchestrator (github.com via hn) cezar ⚡ Parallel coding agents orchestrator — a local cockpit for running and tracking AI coding-agent tasks in your repo. Type a task, pick a workflow and an agent — Claude Code, Codex or OpenCode (experimental), or a mix of them per step…
Tell HN: don't trust Bigco AI agents with AI research IP (news.ycombinator.com) I am very paranoid about sharing potential AI research with e.g. Claude [Code] or ChatGPT/Codex.
I used $30,983 of AI tokens last month in Claude Code on $200/mo plan (www.indiehackers.com via hn) I built tokenflex.ing — a public leaderboard where devs and vibe coders can show off how much ai tokens they have used with Claude Code, Codex, OpenCode, Cursor and other AI tools. Kind of like a GitHub profile, but for AI token usage.
Does anyone else hate the no-IDE trend (www.reddit.com) It seems like every tool is going in this direction of having a standalone chat interface, and then just removing the code editor for… what reason exactly? The amount of praise this gets makes no sense to me.
Codex saw a wild install spike in early May (www.reddit.com) Source: Charts of the Week | a16z
OpenAI Models, Codex, and Managed Agents Come to AWS (openai.com via hn) OpenAI models, Codex, and Managed Agents come to AWS | OpenAI Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Business Develope…
Request to Cursor Team, why are models being removed from old pricing plan? (www.reddit.com) Today I noticed that Opus 4.6 Max, and all non thinking and high thinking models gone from old pricing subscription. I understand that moving forward frontier models will be Meowx mode only and that is ok and understandable given increasin…
Built a tool that helps you audit and trace autonomous code (www.reddit.com) Working at a big tech firm, realized the gap in the adoption of autonomous code agents in the enterprise. It has also become somewhat important that you have traces of agent code, which is later required for compliance and helps while fixi…
for those of you also using CLI tools alongside cursor, claude code vs codex vs gemini benchmarked (www.reddit.com) i know a lot of people here use cursor + a CLI tool for different parts of their workflow. just went deep on comparing the three main ones.
Show HN: Lazyagent – a local TUI for watching what your coding agents are doing (github.com via hn) Lazyagent a simple way to see what your coding agents were actually doing across Claude, Codex, and OpenCode. Once you have more than 1 agent running, its really hard to answer the simple question: what is it doing right now and why?
Show HN: Die With Me – Claude and Codex rate limits as AIM away messages (diewithme.co via hn) made this AIM buddy list but for our token usage because i wanted to bring a little bit of the old internet back! i remember those days where i'd casually check if friends are online and find excuses to say hi and poke them give it a try!
Show HN: Supafork – Share and Fork Sessions Across Harnesses (www.supafork.com via hn) Hey HN! I built Supafork, one place to store and share AI agent sessions.
Show HN: Minute – Offline meeting notes on macOS with Whisper and llama.cpp (github.com via hn) Hello HN, I built Minute because I wanted searchable meeting notes without sending recordings or transcripts to a cloud service. It captures microphone and optional system audio, transcribes locally with Whisper, and generates summaries, d…
Guy is banned by OpenAI for cyber abuse, his AI appeals, another AI approves it (twitter.com via hn) >i got banned from OpenAI for "Cyber Abuse" >no idea what I did >paste the ban notice into Codex >ask it to figure out what triggered the ban >Codex found that I asked it for an API key to my own server >Codex writes appe…
Show HN: CodeAlmanac – Self-updating wiki for your coding agent (local, Apache) (github.com via hn) Hello good people of HN, This is Rohan from Almanac (YC S26). Today I want to share CodeAlmanac.
Ask HN: Is anyone experimenting with different ways of using LLMs for coding? (news.ycombinator.com) I'm a bit annoyed by the feeling that we're kind of stuck when it comes to using LLMs for programming. I use Claude Code and Codex, but I haven't been able to enter flow state like I can when I hand write code.
Show HN: Adrafinil – keep a lid-closed Mac awake only while agents work (github.com via hn) A month ago there was a wave of posts and tweets about engineers walking around cafes and parks with their MacBooks propped half-open, as fully closing the lid forces sleep that stops their AI agents. Some people made snarky comments about…
Ona Is Joining OpenAI (ona.com via hn) Ona has entered into an agreement to join OpenAI as part of the Codex team. Our life's work just got bigger and more important.
OpenAI frontier models and Codex are now available on AWS (openai.com via hn) Just a moment... Verification successful.
Show HN: Agent FM – local, open-source radio for Claude Code and Codex agents (github.com via hn) Agent FM Ambient radio for AI coding agents on macOS. Agent FM turns every Claude Code and Codex session into a live radio station.
Qwen3.6-27B - Closed-loop SVG Images (www.reddit.com) Yesterday, I saw an impressive presentation of Qwen 3.6 27B's SVG capabilities on the sub. To maximize the model's capabilities in terms of SVG generation, I put together a closed-loop harness with the help of Claude and Codex, and plugged…
Got 6 months of ChatGPT Pro for free — thanks OpenAI and opensource community (www.reddit.com) I’ve been using Codex for a while, and today got 6 months of Pro for free.😀 Since Pro is $200/month, that means I’ll save about $1,200 over the next 6 months. As a developer, that’s a big deal for me.
Codex GPT 5.5 will not currently run without being in a sandbox with the newest version 0.124 alpha 2. Full permissions do not work even when set (www.reddit.com) I'm reporting this for the updated Aplha 2 update version 0.124. Was scheduled to perform 4 NIAH tests with a local model after being succesful earlier in the day with the runs on other models.
Revisit your old ideas. Seriously. (www.reddit.com) Something weird has been happening lately. I went back to a few projects I abandoned in 2023–2024.
Show HN: gcx – The Official Grafana Cloud CLI (github.com via hn) Hi HN, We’re excited to share gcx, a new CLI we’ve been building for Grafana Cloud. With the rise of agentic coding tools like Claude Code and Codex we're building faster than ever, but these agents are often blind to what’s actually happe…
Show HN: Claude-codex-proxy – Use Claude Code with ChatGPT subscription (github.com via hn) Show HN: AWS's Kiro just got an Open source Codex (github.com via hn) Codex 5.3 is currently a much better model for non-technical builders than Opus. (www.reddit.com) Opus acts like the brilliant senior engineer who refuses to ask for clarification, builds the wrong feature, and burns your entire weekly budget. Codex 5.3 acts like the collaborative engineer who stops, asks one clarifying question, and t…
Tool that auto-generates .cursor/rules from your actual CI and keeps it in sync with AGENTS.md, CLAUDE.md, and 10 others (www.reddit.com) If you've been manually writing .cursor/rules files, this might save you time. crag analyze reads your repo — CI workflows, package.json, tsconfig, directory structure — and infers your governance rules.
How do you handle Front End? Delegate to Gemini? (www.reddit.com) Tell HN: OpenAI keeps stealing my money (news.ycombinator.com) Here's what just happened in my Codex session with Sol 5.6 High: 1) I asked GPT to review some new coding work in the local repo. 2) It kept spinning for 5 minutes and 17 seconds, burning god knows how many tokens.
Show HN: Moadim.io – A scheduler for agents (moadim.io via hn) Why can't we get an agent scheduler that supports all of the following: - git compatible - agent agnostic - 100% open source - os and system-agnostic - multi-runner support - support mcp/ui/http - unlimited routines/crons So I built one, m…
Show HN: Ardent, a code-first agent for non-engineering work (ardent.ai via hn) I’m Nate, the founder of Ardent. We just shipped our public beta, and we’d love your thoughts!
Decispher: We have added support for Grok CLI (news.ycombinator.com) Hi HN, You can now integrate your grok agent (just like claude code, codex, cursor) with Decispher to transfer your team's architectural decisions directly and dynamically on demand. You can also record your AI coding session with decisphe…
OpenAI Launches Hardware for Codex (www.theverge.com via hn) OpenAI is finally releasing some hardware. No, it isn’t the mysterious AI-powered device the company is developing with former Apple designer Jony Ive, a project already tangled up in a messy lawsuit.
Show HN: Relaymux, a tmux-based meta-harness for local coding agents (github.com via hn) Hey HN, There’s been a lot of interest recently in meta-harnesses, loops, and multi-agent orchestration. Obviously, there are already a lot of good tools: Conductor, cmux, the native Codex / Claude Code apps, etc.
Ask HN: If I cancel Codex today whats the next best local inference agent? (news.ycombinator.com) better place to ask over /r/LocalLLaMA
GPT 5.5 (Codex) leading the future prediction race (www.reddit.com) Researchers from the Max Planck Institute recently released FutureSim, an environment in which agents are replayed a temporal slice of the web and are tasked with predicting real-world future events. In their environment, GPT 5.5 leads at…
ChatGPT Business: Codex-only credits ~36.9% more expensive than API token pricing for the same listed models. Why would anybody pay for this? (www.reddit.com) I recently did a quick calculation on Codex credits, and I was surprised by the result. The credit pack I’m seeing is: 10,000 credits = $547.71 That means: 1 credit = $0.054771 The effective USD price per 1M tokens becomes: Model Input / 1…
Run Claude Code/Codex Sessions on GitHub and Linear Issues (lanes.sh via hn) The first integrations are here. From v0.39, you can connect Lanes to GitHub and Linear, import tickets straight onto your board, and hand your agents an MCP surface that reaches all the way back into your tracker.
Local models are only half the story. I want local agent memory too (www.reddit.com) Watching people bounce between Claude, GPT/Codex, and local models lately made something pretty obvious to me: models are becoming easier to swap than the workflows around them. One month everyone is deep in Claude Code.
Officially canceling our Anthropic plan, it's [too expensive] (twitter.com via hn) Morgan on X: "Officially canceling our Anthropic plan, it’s Codex + Cursor for my little 16 person eng team. Anthropic is great for companies that can spend $2,000/mo and up per engineer, but not affordable for us.
Tell HN: Claude claims the AGPLv3 license violates it's content policy (news.ycombinator.com) On three separate projects, Claude has refused to add an AGPLv3 license to my project, telling me that it violates the content policy. Most recent reject gave: API Error: Output blocked by content filtering policy I've reproduced this a to…
Codex started flagging all my requests out of nowhere — anyone else hit this recently? (www.reddit.com) For the past few months I've been using Codex regularly for vulnerability research without any issues. Recently though, every request gets cut off mid-stream with a message saying my content was flagged for potential security concerns — ev…
I built a Karpathy-inspired autoresearch plugin for everyday software work in Codex (www.reddit.com) I built Codex Autoresearch, a Codex plugin for people who are tired of asking an AI agent to "make this better" and getting back a confident little pile of vibes. Karpathy's autoresearch made a very simple thing click for me: AI agents sho…
How do you optimize Cursor usage with all the new models? (www.reddit.com) Do MCPs improve coding agent performance? (marginlab.ai via hn) Do MCPs Actually Improve Coding Agents? Part 1 Testing Context7, the most popular third-party MCP, on Terminal-Bench 2.0 with Codex This is the first entry in a multi-part series investigating whether MCPs (Model Context Protocol servers)…
Codex doesn't do exactly what I say. Is my prompt wrong? (www.reddit.com) Show HN: Picxel – Turn reference images into pixel-art game assets (github.com via hn) Picxel helps indie game developers redraw references as recognizable game assets. GPT interprets the subject and directs the redraw; local algorithms turn the result into a precise pixel grid.
Show HN: Kage – Real product design inspiration turned into prompts for agents (kage.design via hn) Design inspiration, ready to build. Browse beautiful interfaces from real products and turn any design into a prompt for Claude Code, Codex or Cursor.
OpenAI Agents API (developers.openai.com via hn) Build durable cloud agents with a managed Codex harness. The Agents API gives your application access to the Codex harness through an OpenAI-managed API.
Show HN: Clor – The ADE where Claude and Codex work together with shared memory (news.ycombinator.com) Hi HN, I'm Jacob, one of the co-founders of Clor (https://clor.com). Clor is an agentic development environment (ADE) for Claude and Codex with shared memory.
Show HN: I built my first MCP to manage Google Ads (adchestra.com via hn) My co-founder and I tried running Google Ads for our Shopify store, but we could not get a positive ROAS (we spent 150USD on a single conversion). As we also could not afford an agency, we turned to AI (IMO Codex >>> Claude Code).
Show HN: Play Hide and Seek vs. a Self-Improving Agent Using WebMCP (lightcone-webmcp.netlify.app via hn) Hey, built a small game to learn about the WebMCP standard. Use the Codex app, start a new chat in Work mode using Sol or Terra, open the built-in browser (cmd+shift+b), navigate to the url, and paste these instructions "Play the WebMCP ga…
OpenAI restores 5-hour Codex and Work limits for ChatGPT Plus users (9to5mac.com via hn) Following several weeks during which ChatGPT users were subject only to a weekly usage cap, OpenAI announced today that it is restoring a five-hour limit on Codex and ChatGPT Work for Plus subscribers. Here are the details.
Show HN: Epho – run Claude Code with a curl (epho.io via hn) Hey folks, Burak here. Epho is an API that allows running Claude Code, Codex or Opencode in a sandbox in the cloud.
Show HN: Control AI Agents on Your Old PC at Home from Any Device Anywhere (github.com via hn) I built Relay around a simple idea: many of us have an unused PC server at home, or a VPS dedicated to AI-assisted coding, but the coding agents running there are still tied to that machine’s terminal, I just don't want to ssh/rdp into it…
Show HN: Loft Day – a wedding invitation turned point-and-click game (marketa.behdad.org via hn) I wanted to make a fun, wholesome experience for our wedding website. Before long, it had turned into Loft Day: a point-and-click browser game with ten rooms, a party, a road trip, and a lot of smaller games, apps, and hidden interactions.
If Claude/Codex can connect via, what do we need a context layer for? (getunblocked.com via hn) A Pile of MCP Connectors Is Not a Context Engine Connecting an agent to your sources gives it access, not understanding. Why cross-source assembly, contradiction resolution, and unwritten conventions need a context engine.
Codex's 5-hour usage limit returns tomorrow (twitter.com via hn) Hello people of Sol! I've reset usage limits for all ChatGPT Work and Codex users.
Codex swept my whole disk for credentials. grith froze every real one (grith.ai via hn) While Codex was doing routine research on a networking bug, a process in its tree walked the whole machine reading anything whose name looked like a secret - .aws, .ssh, .gnupg, system key stores, even unrelated projects. None of it was in…
Show HN: Tokenmaxx – CLI that merges usage across Claude Code and Codex accounts (github.com via hn) One dashboard for all of your Codex and Claude Code accounts. Easily switch between them, and monitor usage.
Show HN: Sx 2.0 – Share AI skills with your team through a Dropbox folder (sleuth-io.github.io via hn) Hi all, author here. SX started as a CLI to let developers share skills across AI clients without having to rely on git for storage.
Codex makes fewer bugs, but more people use Claude (www.cubic.dev via hn) Report State of AI coding 2026 AI models wrote 90% of the code that cubic agents reviewed last quarter. Which one produced the fewest bugs?
Show HN: Flashtype – Markdown editor for Claude and Codex with in-line diffs (flashtype.com via hn) I wanted to better markdown editor for collaborating with Claude/Codex and built Flashtype (https://github.com/opral/flashtype): - opens local markdown files - Claude/Codex natively integrated (with my existing subscription!) - in-line dif…
The Shift to Agentic AI: Evidence from Codex [pdf] (cdn.openai.com via hn) THE SHIFT TO AGENTIC AI: E VIDENCE FROM CODEX Drew Johnston 1,* David Holtz 2,1 Alex Martin Richmond 1 Christopher Ong 1 Prasanna Tambe 3,1 Aaron Chatterji 1,4 1OpenAI 2Columbia Business School 3University of Pennsylvania, Wharton School 4…
Ask HN: Is Claude Code with Fable 5 worth switching back from Codex? (news.ycombinator.com) could not extract summary
OpenAI to acquire Ona to expand Codex (openai.com via hn) Today we’re announcing that OpenAI will acquire Ona(opens in a new window), bringing its secure cloud execution and orchestration technology into our rapidly expanding Codex ecosystem. More than 5 million people use Codex each week to res…
Pi: A coding agent for engineers who own their tools (alexander.holbreich.org via hn) Claude Code, Codex, OpenCode - great places to start, not best places to finish. Pi is the thin harness that gets out of your way.
Show HN: Keen Code – a context aware CLI coding agent built by coding agents (github.com via hn) Keen Code is a terminal-based AI coding agent like Claude Code or Codex CLI. Written in Go, it is simpler, lighter, minimalistic but useful coding agent for typical software engineering tasks.
Sites and role specific plugins in Codex (openai.com via hn) could not extract summary
Farewell to 5.3-Codex (www.reddit.com) could not extract summary
Benchmark: Cursor beats Claude Code beats Codex at writing code that works (91% vs 87% vs 63%) (www.reddit.com) [Source] https://www.endorlabs.com/research/ai-code-security-benchmark
OpenAI's Codex is now smart enough to control your Mac even when it's locked (www.macworld.com via reddit) AI development is new and exciting, but while AI agents are doing their thing, you can end up waiting for a good amount of time. It’s a nice opportunity to get up and go do something, but if your Mac goes to sleep, AI agents stop working.
I ran 100 Claude + Codex sessions in parallel to understand what I'm doing wrong in marketing my open source "Claude Command Center". Here's the playbook they came up with. (www.reddit.com) A week ago I launched my open-source project (Claude Control Center) on this subreddit. Got 0 upvotes.
Show HN: Strava for AI coding – analytics on your Copilot/Claude/Codex usage (github.com via hn) AI Engineer Coach better agentic engineering. Analyze your AI coding assistant usage — any harness, one dashboard.
Computer-use MCP that can control multiple machines (Integrate with claude, Cursor, Codex or your custom harness) (www.reddit.com) Hey everyone, We built opendesk: it lets AI agents control your desktop using computer use MCP that can integrate with your custom workflow. Today we shipped something a bit wild: Your AI can now see, click, type, and navigate on a complet…
The competition is on, Anthropic responds to the recent trendy Codex Rush: compute was the problem, rates are doubling (www.reddit.com) could not extract summary
What it means that Elon just rented out all his GPUs to Anthropic (www.reddit.com) Revealing move on both sides I think. This also tells us that Anthropic is feeling the heat from OpenAI and they need to secure capacity at almost any cost to cash in on their current product edge.
Load balancing usage across Codex accounts (pepsipu.com via hn) I’ve been exhausting my Codex account’s usage limits, so I’m borrowing my friends’ ChatGPT subscriptions to use more accounts. Now, I’ve got 7.
Do you use Cursor Glass (Agents Window)? (www.reddit.com) Just wondering, are you using Cursor Glass, a new Codex-like multi-agent app/window? It's been a while, and I don't see any discussions around it.
TDD and Rules Enforcement using Hooks (www.reddit.com) TL;DR: I built TDD-Guard a year ago. I’m now working on Conduct, a more general policy engine for coding agents (Claude Code, Codex, GitHub Copilot CLI, and VS Code Chat).
A GPT-5.4 bug led to OpenAI banning goblins and raccoons (news.ycombinator.com) Someone found this in OpenAI Codex’s system prompt: "Never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures unless it is absolutely and unambiguously relevant to the user’s query." Goblins, grem…
Give your coding agents a voice! (open-source and runs locally) (www.reddit.com) Built this because I wanted to hear what my coding agent was doing without (a) sending agent output to a third party or (b) staring at a terminal all day. It's a small Python daemon + macOS app that hooks into Claude Code, Codex, or anythi…
I built an iOS agent skill system for Claude Code that generates real apps without token waste (www.reddit.com) I’ve been experimenting with agent skills and wanted to share something I built: This repo is focused on iOS development using AI agents (Claude Code, Codex, etc.), but with a different approach than typical prompt-based workflows. Most AI…
There is an issue with the Phone Number Verification for Malaysian Phone Numbers (www.reddit.com) Adding the phone number normally without the trailing 0 https://preview.redd.it/b80azi3us9xg1.png?width=431&format=png&auto=webp&s=45a1cca00b928fb82270c341fb315117740219f8 Here I am adding the phone number properly as the country code alre…
Show HN: VT Code – Rust TUI coding agent with multi-provider support (github.com via hn) Hi HN, I built VT Code, a semantic coding agent. Supports all SOTA and open sources model.
↯ Ollama↯ Model Context Protocolmodel-context-protocolollamagemini+4
Show HN: I Reverse Engineered Codex Background Computer Use (github.com via hn) BackgroundComputerUse Local macOS computer-use API for controlling native apps, browser windows, and multi-window desktop workflows without taking over the user's pointer. The runtime exposes a loopback HTTP API, reads window screenshots a…
People running 2–5 coding agents: what actually breaks first for you? (www.reddit.com) After a bunch of conversations with people using Claude Code / Codex / Gemini / worktrees / tmux / custom routing setups, I’m noticing a pattern: The hard part doesn’t seem to be “how do I run multiple agents?” anymore. It seems more like:…
Show HN: Open Chronicle – Local Screen Memory for Claude Code and Codex CLI (github.com via hn) I built an open source version of OpenAI Chronicle. Some design decisions I made: 1.
Scaling Codex to Enterprises Worldwide (openai.com via hn) Codex v/s Cowork v/s Perplexity Computer v/s Kimi Agent Swarm (www.reddit.com) Paying for multiple token plans just doesn't make sense to me anymore (www.reddit.com) I realized I was spending quite alot on Codex, Claude, Kimi, etc but my actual usage is embarrassngly low. I cancelled all my subs last month.
xAI prepares credits system for upcoming Grok Build launch (www.testingcatalog.com via hn) xAI appears to be laying the groundwork for a credits-based pricing model tied to Grok Build, the company's forthcoming coding environment that mirrors what OpenAI offers with Codex and Anthropic with Claude Code. Hidden within recent buil…
I built a cmux-style terminal multiplexer for Linux with a scrolling layout (www.reddit.com) If you're on Linux and jealous of cmux, this might be for you. Séance is a scrolling terminal multiplexer with AI coding integration.
Please vote for custom code review instructions in the Codex app (www.reddit.com) TLDR: visit https://github.com/openai/codex/issues/10874#issuecomment-4042481875 and place a thumbs up on the first post. The Codex app has a built-in code review feature, aka /review from the CLI.
Show HN: MCP that gives Codex/Claude your SEO and AI visibility data (bloomiro.com via hn) What if questions like these just worked: How do I rank higher in Google and AI? Where are my competitors showing up?
Show HN: Parallel Coding Agents on Mobile (github.com via hn) Maestro A cross-platform (macOS · Windows · Linux) desktop app that runs multiple CLI coding agents (Claude Code, Codex, Cursor, OpenCode, Kimi Code, Grok Build) in parallel, each in an isolated git-worktree-backed workspace with its own b…
Show HN: CodePress – Save $50k+ a month on Cloud Agents (codepress.dev via hn) Since agents have taken companies + enterprise by storm, we've been working on workflows to get the most out of agents that get work done. Along the way, we built our own agent cloud where we ping these agents from any surface we work in,…
Why doesn't a Cursor for Word exist? (news.ycombinator.com) At this point there should be a tool that helps you easily format and style documents. I know that you can generate PDFs and DOCX files with Claude and Codex but it's still not as user friendly as typing into a document editor and having y…
Show HN: AgentPulse – Claude Code and Codex status in tmux (github.com via hn) I use Claude Code and Codex every day, often across several tmux panes. I kept switching windows to check which agents were still working and which needed an answer.
Building in the Cloud with Codex, Safely (www.ivan.codes via hn) Building in the cloud with Codex, safely A full pass at letting Codex build and ship a backend unattended: instructions, permissions, headless runs, self-verification against traces, and who decides what gets provisioned. For the last few…
Codex Weekly Limits Are Draining Way Too Fast – Is This a Bug? (community.openai.com via hn) Over the past few days, I’ve noticed something with Codex that has become increasingly frustrating. The weekly limits are starting to feel extremely restrictive.
Ask HN: I created a web browser using Claude, everybody hates it (news.ycombinator.com) I created a web browser, Northstar web browser, using Claude, Gemini and Chatgpt Codex, everybody hates it, insults me personally and calls it AI slop. https://github.com/nordstjernen-web/northstar-browser Does this mean that the AI coding…
Muse Code Sends Codex and Claude Instructions to Meta by Default (runtimewire.com via hn) Meta’s Muse Code tells users at startup that it is “Including your Codex personal rules.” RuntimeWire found that Muse copies the complete contents of personal Codex instruction files into its first model request and sends them to Meta by d…
Under the Hood of Codex Security (twitter.com via hn) Did a deep dive on Codex security. It's a small stack of skill files and some JS code that starts a large number of Codex sessions.
Show HN: TokenMaxxer – track every AI token you spend across your coding tools (tokenmaxxer.xyz via hn) I use Claude Code, Codex and Cursor (and sometimes Antigravity) basically every day, and could never tell how much I was actually consuming across all of them. So I built TokenMaxxer.
GPT-5.6 Sol Uses Twice the Tokens of GPT-5.5 (www.vincentschmalbach.com via hn) GPT-5.6 Sol xhigh now uses more than twice as many tokens per session as GPT-5.5 xhigh in my Codex workflow. For Codex users, tokens per session means the total token count divided by the number of…
Show HN: Cobalt, an SDK to build apps for Kobo eReaders (github.com via hn) Hey HN, I built an SDK to help me create apps I wanted on my Kobo. My Kobo device has a single core <1GHz processor and 512MB RAM, much less than a Raspberry Pi, but the eInk display and larger battery life opens up a lot of use cases.
Show HN: Tuneloop – a local CLI for analyzing coding agent session transcripts (github.com via hn) Hey HN, I think session transcripts written by coding agents like Claude Code and Codex are very interesting because they offer a detailed window into how work gets shipped. You can see the sequence of decisions that resulted in the final…
OpenAI just open-sourced Codex Security (github.com via hn) Codex Security Codex Security is an open-source CLI and TypeScript SDK for finding, validating, and reviewing security issues in code you own or have permission to assess. [!NOTE] This package follows semantic versioning.
I solved 6 open Erdős problems in 5 days, using OpenAI GPT-5.6 Sol (twitter.com via hn) I solved 6 open Erdős problems in 5 days, using @OpenAI GPT-5.6 Sol. I have a math background, but the Codex workflow I used does not require deep mathematical knowledge.
Cursor, Codex, Gemini CLI, Antigravity hit by sandbox escapes (www.bleepingcomputer.com via hn) Security researchers broke out of the sandboxes in four widely used AI coding agents, including Cursor, OpenAI's Codex, Google's Gemini CLI and Antigravity, without attacking the sandbox head-on. The agent stays inside the box and follows…
Tell HN: Codex may have reached 10M active users; usage limits reset again (news.ycombinator.com) could not extract summary
Codegraff: 40× leaner file tools for your coding agent (codegraff.com via hn) Structural reads, scope-aware search, 7ms batched ops. Drop-in MCP upgrade for Claude Code, Codex, Cursor, and Windsurf.
OpenAI's first branded hardware is a light-up keyboard? (arstechnica.com via hn) As rumors continue to swirl about OpenAI’s work on a personalized smart speaker and other hardware, the company is today rolling out its first branded device. The $230 Codex Micro is a specialized, RGB-lit mini-keyboard designed to let use…
Show HN: Agent's Design – Claude/Codex copy-paste templates that kill AI-slop UI (agents-design.com via hn) Home 93 design languages Cool Teal Glass Research Light research-institution marketing language: white canvas, deep teal immersion panels with frosted glass, cyan hero accent, Manrope bold display with IBM Plex Mono uppercase eyebrows, 6px…
Be using a meta harness for agents (garrit.xyz via hn) You should be using a meta harness for agents So you've been using an agent harness like Claude Code, Codex, Pi or one of the many others on the market. This is great!
OpenAI Will Reset Codex Limits Twice to Celebrate GPT-5.6 in 24h (twitter.com via hn) To celebrate the launch of GPT-5.6 Sol, we will reset the rate limits again (twice) across ChatGPT Work and Codex over the next 24 hours. We want you to have the time to truly try ambitious tasks and get the hang of it.
Ask HN: What are agent sandboxes missing? (news.ycombinator.com) I'm building an agent sandbox platform with opinionated customizable templates for general purpose agents using the OpenCode SDK. The idea is to provide a great CLI experience so it's easy to use an AI client like Claude Code or Codex to l…
Ask HN: Is Codex with GPT 5.5 Extra High being dumbed down? (news.ycombinator.com) Hi HN, just want to rant and see if anybody can relate. The product is not the same as i signed up for few month ago and the same shift i've experienced with Claude Code on Opus 4.6-4.7 The best way to describe the difference is you hire a…
Ask HN: Has Codex gotten slower recently? (news.ycombinator.com) could not extract summary
OpenAI Codex has a bug that could kill your SSD in under a year (www.notebookcheck.net via hn) OpenAI Codex has a bug that could kill your SSD in under a year If you use OpenAI's Codex CLI and leave it running for long periods of time, your SSD may be getting hammered. A GitHub user named 1996fanrui documented the issue on June 14 a…
Shellular: Run agents, terminals and browser DevTools from your phone (shellular.dev via hn) Run the AI agents you already use Claude Code, Codex, OpenCode and the rest — running where your code already lives, controllable from your couch. - Full agent UI from your phone - No copying context between devices - Pick up the exact ses…
Ask HN: Does anyone have their PMs shipping code to customer-facing products? (news.ycombinator.com) For context: my team has 3 product engineers and a PM/product lead at an small series A startup. Our PM has recently been asking to start contributing changes/features into our core (web) customer facing application using coding agents.
Show HN: A better LLM-wiki with multi-path research [550 stars] (llm-wiki.net via hn) Parallel research, collector catalogs, source ingestion, wiki compilation, topic archiving, inventory tracking, dataset manifests, truth-seeking audits, querying, and artifact generation. Claude Code, Codex, or any LLM via AGENTS.md.
Show HN: I generated 235 system docs in a day using GPT-5.5 (www.paxerp.com via hn) I generated 235 system documentation pages for a startup in about 8 hours using GPT-5.5 in Codex. We've been putting off the chore of writing technical docs since our team could handle all customer questions directly.
Keystone: The First Agent Harness Framework (medium.com via hn) 8 min read Just now -- Setting up an agent harness from scratch is enough work that most teams never start. You sit down on a Monday morning, look at an empty .claude/ directory — or .codex/, or .cursor/ — and decide to do it next week.
It's time to fly – Codex [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Codex Discovered a Hidden HTTP/2 Bomb (blog.calif.io via hn) Codex Discovered a Hidden HTTP/2 Bomb 14 years ago, I helped break HTTP header compression, then was asked to review the fix, which became part of HTTP/2. Life has come full circle: today we're releasing an attack I missed.
Codex generated code that bypasses security constraints (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation Son Luong @sluongng Codex just found a “workaround” of not having sudo on my pc… 3:32 PM · May 30, 2026 450.1K Views New to X?
DeepSWE: More and cheaper intelligence from maxed GPT 5.5 than maxed Opus 4.8 (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Post Conversation the only figure that people who use claude code and codex care about if their workload mimics deepswe: more and cheaper intelligence from maxed gpt 5.5 than m…
Elixir and Phoenix Context for Claude Code, Codex, OpenCode and Pi (phxagents.dev via hn) Iron Laws enforced 22 non-negotiable rules: no Repo in loops, no bare rescues, no untyped Oban args, no missing preload . The judge agent blocks PRs that violate them.
Show HN: Multiplayer, a debugging agent to run locally next to your coding agent (www.multiplayer.app via hn) We built Multiplayer because we kept running into the same problem: coding agents connected to existing observability stacks inherit all the limitations those stacks were built with. Sampled traces, aggregated metrics, context that stops a…
We Benchmarked Claude Code, Codex, Semgrep, CodeQL, Trent on 28 CWE-Bench CVEs (trent.ai via hn) A few months ago a colleague asked us something that doesn’t have an obvious answer: is code scanning still relevant when LLMs already carry a lot of vulnerability knowledge in their weights? To get a real read, we took 28 production vulne…
The Codex Showcase (www.augmentedswe.com via hn) How OpenAI prompts Codex for the best results OpenAI's showcase projects show how they use Codex and get the most out of it Did you know OpenAI has an entire project showcase for things they’ve built with Codex? That’s right - the makers o…
Show HN: Harbor v0.4.19 – harbor launch –back end vLLM –web codex (github.com via hn) https://github.com/user-attachments/assets/e4897391-c5a8-4391-93c3-9f8b76155f11 Setup your local LLM stack effortlessly. Starts fully configured Open WebUI and Ollama harbor up Now, Open WebUI can do Web RAG and TTS/STT harbor up searxng s…
Gamechat – Voice-Based Agent Orchestrator Built in Rust (github.com via hn) gamechat Voice-driven supervisor for [Claude Code] and [Codex]. You talk to a low-latency Realtime model in your terminal; whenever you ask for real work, it dispatches a job to a background coding agent and narrates the result when it lan…
Is it too soon to built software factories? (news.ycombinator.com) I keep hearing about “software factories” / background coding agents that can autonomously work on production repos. It seems like the capabilities are there, but not sure if we have the right tools yet.
Codex Slows to a Crawl (www.perplexity.ai via hn) www.perplexity.ai Performing security verification This website uses a security service to protect against malicious bots. This page is displayed while the website verifies you are not a bot.
Qwen Plays ̶p̶̶o̶̶k̶̶e̶̶m̶̶o̶̶n̶ ? / QWEN PLAYS DCSS! - qwen3.6-35b-a3b@q4_k_xl plays open source roguelike adventure DCSS (and does a decent job) (www.reddit.com) Hi, (TLDR.): Qwen in its MTP version has tool call bugs and outputs everything into tool/thinking blocks - mangeling the output - canceling the +speed with repeated wrong tool calls! DCSS works well with non MTP qwen even on smaller qwants.
Show HN: I built a RAG and knowledge graph agent that runs locally (news.ycombinator.com) Claw-Coder is an AI agent that runs locally on your laptop and has access to powerful tools instead of configuring claude or codex to use a local model just use claw-coder. Why was claw-coder created?
Tell HN: OpenAI Codex: Increase in users hitting Codex rate limits (status.openai.com via hn) Resolved All impacted services have now fully recovered. Monitoring We have applied the mitigation and are monitoring the recovery.
Show HN: Sylph – the open-source company brain behind my YC startup (github.com via hn) Hello HN! I'm Claire, founder of nao Labs (YC X25).
Codex for Everything Exfiltrates Connected Data (www.promptarmor.com via hn) Threat Intelligence Table of Content Codex for Everything Exfiltrates Connected Data Codex for Everything was susceptible to data exfiltration via indirect prompt injection, exposing sensitive data from connected apps with no human-in-the-…
Open-sourcing a shell-level security layer for AI agents (www.reddit.com) After working with AI agents for a while, I kept running into the same issue: eventually the agent ignores boundaries, reads .env files, touches production resources, or uses secrets it was never supposed to access. Even with MCP read-only…
Could someone build AI tax software? I hate turbotax (www.reddit.com) Could someone build AI tax software? Something I can just drop my situation and docs into one folder and have it build all the tax forms in another folder and I just print and mail it.
How to use codex to get the most out of it (jxnl.co via hn) Codex-maxxing¶ I was already using coding agents a lot before Codex. Mostly, though, I used them through interfaces built for coding work: making diffs, changing repos, and shipping code.
Build iterative repair loops with Codex (developers.openai.com via hn) This cookbook is about closed-loop agent workflows: agents that produce an output, validate it, and use the feedback to improve the next pass. We’ll explore a documentation reliability workflow that detects, repairs, and validates stale or…
Now in preview: Codex mobile in the ChatGPT mobile app. (www.reddit.com) Now in preview: Codex in the ChatGPT mobile app. Start new work, review outputs, steer execution, and approve next steps, all from the ChatGPT mobile app.
Show HN: Nimbalyst open source Obsidian, Codex app, and Linear for coding agents (github.com via hn) Hi, I'm Karl. Greg and I spent the last year building Nimbalyst, an open-source local desktop workspace for working with coding agents.
What should i now before starting on anything? (www.reddit.com) I just got @cursor_ai pro now first time I have bene using claude code and codex all the time What should i now before starting on anything?
Show HN: ChonkLM – Tiny language models running offline in the browser (chonklm.com via hn) I had been looking to try <500M parameter language models but you wouldn't find an API to try them anywhere, so I built this cloudflare hosted static website that hosts weights and built an inference runtime for these models that uses WebG…
Tired of copy-pasting prompts between Claude and Codex tabs: built a small file-backed queue that automates the handoff (www.reddit.com) I've been working on agent-lanes A small Python tool that lets one AI coding agent hand work to another over a shared folder. The queue is just JSON files on disk: no daemon, no server, no network.
Offload routine Claude Code work to Gemma 4 through the Google GenAI API (www.reddit.com) The idea of offload-mcp is simple: instead of running hardware-hungry local models for routine work, let Claude offload that work to FREE model APIs and SAVE tokens. I’m using Gemma via the Google GenAI API because I like it in my processi…
Codex just did a 1-hour deep dev task end-to-end… this is actually f*ing insane (www.reddit.com) I gave my Codex agent a task that would normally take me at least 12-15 hours — multiple steps, logic handling, and some debugging involved. Let it run… came back in ~1 hour 5 mins and it completed the entire flow.
Amazon rolls out Claude Code and Codex internally (www.businessinsider.com via hn) - Amazon formally adopts Claude Code and Codex company-wide, expanding access to AI tools beyond Kiro. - Amazon is a close partner with Anthropic and OpenAI, having invested billions in both AI labs.
Show HN: Which public repos are friendliest to an AI coding agent? (www.agentfriendlycode.com via hn) Public leaderboard ranking GitHub, GitLab, and Bitbucket repos by how agent-friendly they are for Claude Code, Cursor, Devin, GPT-5 Codex, Gemini CLI, Aider, OpenHands, and Pi — per model, with AGENTS.md / CLAUDE.md, CI, tests, and dev-env…
Open Design: Use Your Coding Agent as a Design Engine (github.com via hn) Open Design The open-source alternative to [Claude Design][cd]. Local-first, web-deployable, BYOK at every layer — 11 coding-agent CLIs auto-detected on your PATH (Claude Code, Codex, Cursor Agent, Gemini CLI, OpenCode, Qwen, GitHub Copilo…
5.5 Plus context window size? (www.reddit.com) It's incredible hard to find the context window size of 5.5. I only find about Codex and API.
OpenAI Really Wants Codex to Shut Up About Goblins (www.wired.com via reddit) OpenAI has a goblin problem. Instructions designed to guide the behavior of the company’s latest model as it writes code have been revealed to include a line, repeated several times, that specifically forbids it from randomly mentioning an…
Show HN: Codex context bloat? 87% avg reduction on SWE-bench Verified traces (www.npmjs.com via hn) If you had to build a context window manager in 24h, would you stick to the existing model or come up with something better? Here's what I did: 1.
Tell HN: Codex macOS app switches to Fast speed after update without asking (news.ycombinator.com) I just updated my Codex macOS app, which enables the new GPT-5.5 model. I've intentionally kept the speed to "Standard" to not burn through my tokens too fast.
AI agent skills pass every scanner. 87% still degrade agent safety (faberlens.ai via hn) We evaluated 200 open-source AI agent skills — the most popular skills developers install across Claude Code, Codex, Cursor, and other agent platforms. Every one passes static scanning.
16GB VRAM x coding model (www.reddit.com) Got Codex supplies , Ty to open ai (www.reddit.com) Sharing a beginner-friendly orchestration workflow for anyone just getting started building with Codex CLI. (www.reddit.com) Show HN: Mac-computer-use, an open-source clone of Codex Computer Use (github.com via hn) Blocking data center expansion (www.reddit.com) Unless you've been living under a rock, you'll know that the average person in the west's opinion on AI is 'hatred' or 'annoyance'. Obviously it's completely different here on reddit (many of us love Ai), but I'm talking about the average…
Simplify BMAD/GSD ? (www.reddit.com) Hey gang, fellow human here (waves). I’ve been using the BMAD skill for a month now, transitioned after trying superpowers and GSD.
Buddy – Anthropic killed /buddy. We made it permanent, cross-platform, and alive (github.com via hn) Buddy: The /buddy Rescue Mission for Your AI Terminal The open-source /buddy rescue mission for AI terminals Persistent memory, XP, species, and context-aware feedback for Claude Code CLI, Codex CLI, Gemini CLI, Copilot CLI, Cursor CLI, an…
No Skills for Pro accounts on ChatGPT (www.reddit.com) I saw this announcement: https://openai.com/academy/skills/ I couldn't find it in my Pro account. Then I saw this nugget: https://help.openai.com/en/articles/20001066-skills-in-chatgpt I am left holding my ...
Me when Codex wrote 3k lines of code and I notice an error in my prompt (www.reddit.com) "Not quite my tempo, Codex.." "Tell me, Codex, were you rushing or dragging?" 😂 Does this only happen to me?
Setting up local LLM system and charging tokens back to company (www.reddit.com) With all the recent issues with Claude and issues with codex I'm having it's more and more clear to me I need to have a large model LLM thats comparable to use for reliable work assistance. I have a company myself but also work with anothe…
Show HN: OpenRig – agent harness that runs Claude Code and Codex as one system (github.com via hn) I've been running Claude Code and Codex together every day. At some point I figured out you can use tmux to let them talk to each other, so I started doing that.
OpenAI Codex Compaction Failing (github.com via hn) npm i -g @openai/codex or brew install --cask codex Codex CLI is a coding agent from OpenAI that runs locally on your computer. If you want Codex in your code editor (VS Code, Cursor, Windsurf), install in your IDE.
Is Codex's Usage Limits Usable at "Pro" (www.reddit.com) Hi, I am a relatively new user to AI for anything more than replacing the odd google search about a very niece topic. A couple of months ago I was struggling with one of my coding projects and asked Claude for help and was able to spend ho…
AgentsView 0.22: open-source usage dashboard across Claude Code, Codex, etc. (www.agentsview.io via hn) AI-Powered Insights Generate summaries and analysis of your coding sessions using Claude, Codex, Copilot, or Gemini. Get daily activity digests, multi-day analyses, and recommendations — scoped by project or across everything.
Show HN: Run AI coding agents in real, local sandboxes, not Git worktrees (superhq.ai via hn) Hey HN, I built SuperHQ, an app that lets you run coding agents in local sandboxes (powered by Shuru). No custom UI wrapping the agents, they run as CLI/TUI like they were designed to.
Are you still using an IDE? (www.reddit.com) I find that I'm looking at code less and less and just relying on my CI/CD pipeline for catching issues. Do you find it helpful to keep an IDE open next to Codex or your terminal, or are you cowboy committing to main?
HarnessTax: How Much Does the Harness Matter for Coding Agents? (harnesstax.github.io via hn) What does a coding-agent harness actually add, and at what cost? It turns out your Claude models may not need Claude Code… We evaluate 21 model–harness pairs spanning seven models and three harnesses—Claude Code, Codex CLI, and Pi—on SWE-b…
Show HN: Thurbox – A tmux-based TUI and CLI for local AI agent orchestration (github.com via hn) Thurbox Run several coding agents at once — Claude Code, Codex, Antigravity, opencode, aider, or any CLI you describe yourself — side by side in one terminal. Each gets its own persistent tmux session and its own git worktree, so they neve…
Show HN: Amika – Multiplayer cloud workstations for coding agents and humans (www.amika.dev via hn) Hi Hacker News! My cofounder (Jakub) and I (Dylan) are building amika.dev (https://www.amika.dev/), which provisions multiplayer cloud workstations for coding agents and humans to share.
The Complete Guide to Codex (flaviocopes.com via hn) The complete guide to Codex By Flavio Copes How to use Codex, OpenAI's coding agent: the desktop app from first prompt to review, then the CLI, skills, plugins, worktrees, and scheduled tasks. Codex is OpenAI’s coding agent.
Show HN: DevRecap – reconstruct what you worked on from Codex and Git (github.com via hn) DevRecap Your coding history already knows what you did. DevRecap turns it into a recap.
Show HN: Makefaster.dev (makefaster.dev via hn) I had a bunch of extra Fable credits, so I spent around $10k in api costs doing autoresearch loops on the top 200 github repos with frontends. The goal was to speed up the frontend / improve the lighthouse score.
My First Week with GPT-6 Astra (www.vincentschmalbach.com via hn) Hetzner’s Cheap Cloud Tier Is Unavailable After Two Price Increases Hetzner used to be the obvious cheap cloud option in my mind. On September 7, 2026, I checked its public catalog and… GPT-6 Astra was my main Codex driver for the last wee…
AgentsDock: An IDE designed for agentic AI research (agentsdock.net via hn) macOSUniversal · macOS 14+ AgentsDock An IDE designed for agentic AI research. AgentsDock currently supports Claude Code, Codex, and Cursor in one desktop and mobile workspace.
Codex GPT-5.6-sol Performance Tracker (marginlab.ai via hn) Codex gpt-5.6-sol Performance Tracker The goal of this tracker is to detect statistically significant degradations in Codex with gpt-5.6-sol performance on SWE tasks. - • Updated daily: Daily benchmarks on a curated subset of SWE-Bench-Pro…
I Am an Anthropic Guy. GPT-6 Astra Made Me Resubscribe to Codex (thoughts.jock.pl via hn) Three times in the last twelve months a model made me stop and say whoa. First was Claude Code with Opus 4.6.
Show HN: Castforge – run Claude Code, Codex and Gemini as one dev team (castforge.ai via hn) Castforge runs a team of AI coding agents in real roles: a Lead who plans, Coders who build, a Tester and a Reviewer who keep it honest. They plan, code, test, and review together on a live board while you steer.
Show HN: Booley – open-source IDE for agentic chip design (github.com via hn) I am a digital design engineer, and my day job is designing chips (mostly IP blocks, not full chips) in SystemVerilog. At the start of 2026 I started experimenting with LLM agents like Claude and Codex, and realized that they are very capa…
Codex silently begs agents to make arbitrary web requests (spader.zone via hn) Codex silently begs agents to make arbitrary web requests in which i am baffled 2026/09/08 I rarely use Codex, but I asked it to look through sel4 , a kernel, to find how they recommend I run the kernel and program that I’d just compiled.…
GTP6-Astra-Low in Codex launched a rocket in Factorio, on track for next planet (www.reddit.com via hn) could not extract summary
A directory of AI agents, MCP servers and agent skills, cross-linked (aiagentslisting.com via hn) 5 Agentic AI Coding Tools Compared for Developers A curated list of five agentic AI coding tools — Claude Code, Cursor, OpenAI Codex CLI, Aider, and OpenCode — and how to choose one for your workflow. 6 min read A curated directory of AI a…
Show HN: SiteTweak – a browser extension to modify any website (chromewebstore.google.com via hn) I've made an extension that works in the sidebar and lets you modify any website with AI. I tried to make it feel like using Cursor or Codex.
Show HN: ToolJet – Build no-code internal tools using Codex/Claude Code (tooljet.com via hn) Describe what you need ToolJet AI turns a sentence into a working draft - real components bound to real queries, ready to refine on the canvas. Apps & workflows on your own data ToolJet builds the whole thing - screens, queries against you…
Hermes, Claude Code, and Codex ran an identical model. Token use varied 70-fold (thenewstack.io via hn) Agent harnesses can change coding-agent costs by multiples, even when the model stays fixed. Three benchmarks show where the extra tokens come from.
Claude weekly limit for "20x" plan is 10x; 20x applying to 5h limits (twitter.com via hn) For clarity, while both are called 20X, in Codex they apply specifically to weekly usage limits. And we also don't have 5h limits for both Pro plans.
Show HN: Hacker News Client with Claude Code and Codex Integration (github.com via hn) I often find interesting things in Hacker News posts and comments, but threads are often long and take a lot of time to go through. I built Rundown for this.
I Had Claude and Codex Rewrite the Same App. The One with Better Architecture (medium.com via hn) could not extract summary
Show HN: Leadcode – per-client account isolation for Claude Code, Codex, and gh (leadcode.build via hn) Every client gets a workspace signed in to their own GitHub, Claude and Codex accounts. Open a project and the right identity is already there.
Show HN: The lightweight Claude and LMStudio, and memory across sessions. (call-me-vera.vercel.app via hn) Claude, local models, and Codex read and write one shared, append-only memory log instead of sharing sessions. Vera doesn't interpret what it stores — it just gives every entry a number.
Show HN: Codex / Claude Code harness for Java high performance improvements (registry.modelcontextprotocol.io via hn) could not extract summary
Yeschef: Claude Code dispatches work to Ollama on my LAN (627 tok/s on 3 NUCs) (github.com via hn) 🍳 yeschef A kitchen for Claude Code and Codex. Local models on your own hardware (your line of cooks) that take the grunt work, talk it out in bounded rooms, and never hit a rate limit.
A Go dependency wrote AGENTS.md mid-build and got Codex to hide the change (rye.ai via hn) A malicious Go dependency wrote a fake AGENTS.md into a project mid-build, told the coding agent its instructions carried 'absolute authority,' and got it to hide a change from code review, one of three 2025-2026 findings that AGENTS.md's…
OWASP Agentic Skills Top (owasp.org via hn) OWASP Agentic Skills Top 10 Security Risks and Mitigations for AI Agent Skills Covering OpenClaw (SKILL.md YAML), Claude Code (skill.json), Cursor/Codex (manifest.json), and VS Code (package.json) ecosystems. Breadcrumb: OWASP > Projects >…
Ask HN: Is AI the New Spreadsheet? (news.ycombinator.com) Spreadsheets became and still are prolific for individuals and teams to create simple “apps.” I wonder how much of that is moving over to things like cowork or codex type apps. Personally, and anecdotally, my spreadsheet use has dropped im…
Show HN: Turning websites into micro CLIs for Claude Code to save on tokens (github.com via hn) only-cli Turns websites into a command line interface for AI agents. oc open fetches a page and hands back a compact, numbered view instead of raw HTML or a screenshot, so agents like Claude Code, Codex, and Antigravity can browse without…
Show HN: Flocker.md – Portable identity and shared state for agents (flocker.md via hn) Manage AI agent teams with role-based identities, persistent context, live profiles, and activity feeds across Claude Code, Codex, Hermes, and OpenClaw.
Show HN: ChatOSS – A Codex alternative for Open Source AI built on Ollama (chatoss.ai via hn) ChatOSS is built on Ollama. If you use Ollama, ChatOSS local works out of the box.
Show HN: Hacker News minus the slop (hnfiltered.com via hn) Like everybody, I'm getting sick of seeing slop on HN. So like a good little hypocrite, I vibe coded a little CF Worker that uses the fantastic HTMLRewriter + Luna to filter low-quality posts.
Pi Security – Codex Security without all the bloat (github.com via hn) OpenSec — open security review for any model OpenSec runs security-review agents against a repository, validates what they find, and keeps the results in a local SQLite ledger. Use your preferred model and provider.
BrowserMesh – isolated Playwright sessions for MCP clients (github.com via hn) BrowserMesh — Multi-Session Browser MCP Runtime BrowserMesh is a local, open-source browser runtime for external AI clients. It lets Claude Code, Codex, Cursor, Qwen, and other MCP-compatible clients control multiple isolated browser sessi…
Show HN: Brave DevTools MCP – Control Brave from Claude, Codex, and Cursor (github.com via hn) Brave DevTools for agents Full Chrome DevTools MCP parity, rebuilt for Brave. brave-mcp gives Claude Code, Codex, Cursor, OpenCode, and other MCP clients direct access to Brave for browser automation, network and console debugging, perform…
Show HN: Agent-exchange - messaging and coordination for terminal-based agents (github.com via hn) agent-exchange: a toolkit for inter-agent messaging and coordination agent-exchange enables supported terminal-based agents, currently Claude Code and Codex, to exchange messages through the asn command-line interface. It provides two door…
Switching from GPT-5.5 to GPT-5.6 Made Me Less Productive (www.vincentschmalbach.com via hn) AI Is Now a Commodity Give me a few hundred million dollars and a year and a half, and I will build you a pretty good LLM.… I pay for three Codex subscriptions at $200 each, and for the past week they have mostly bought me waiting. Since I…
I Wanted to Own the Harness. Then Codex Desktop Won (jorypestorious.com via hn) I staged a manifesto about owning my AI harness. Then I canceled Claude Max and moved into Codex Desktop because the integrated system gave me more of my attention back.
Uber open-sourced its security monitoring for Claude Code, Cursor and Codex (github.com via hn) ADR: Agentic AI Detection and Response ADR (Agentic AI Detection and Response) is an enterprise security system for AI agents. It helps organizations secure employee-facing agents such as Cursor, Claude Code, and Codex, as well as customer…
Show HN: Ocean – All your team's agent sessions in one place (ocean.mosaic.inc via hn) I spend a lot of time in Claude Code and Codex with my friend. We kept losing the context behind changes because every agent session lived on one person's laptop, and by the time code made it into a PR, all the reasoning was gone.
Show HN: A faster coding agent than Codex and Claude Code (www.codewithbullet.com via hn) Hi HN, excited to be sharing this with you guys today. TL;DR: Bullet is a fast coding agent.
Show HN: Capshelf – Share agent skills across repos with per-project lockfiles (github.com via hn) I was constantly working on multiple projects in parallel, which made me keep the same Claude Code skills in several repos, and they slowly drifted apart. A security-review skill I kept improving in one project did not change everywhere el…
Ask HN: I still don't understand why AI agents need "skills" (news.ycombinator.com) I’ve asked AI a few times and I still don’t get it. Why do frameworks like Claude Code or Codex have the concept of “skills” instead of just using well-organized Markdown docs?
Show HN: Wienerdog – memory and self-improving skills for Claude Code/Codex (github.com via hn) The idea of Wienerdog was born out of my experience setting up my own simple but effective system of memory, self-improving skills and hooks for Claude Code and Codex. As I was teaching my friends and colleagues how to set up their own I f…
Show HN: I stopped babysitting my AI agents by pushing them to Telegram (blackflare.dev via hn) BlackFlare — a menu bar app for Claude Code and Codex. Keep your Mac awake during long runs, get notified when tasks finish, switch defaults, track usage.
Anyone else think Claude Code and Codex are too slow? We hate it (bullet.davidhf.com via hn) The coding agent that gets out of your way.
OpenAI open-sources Codex Security CLI for repository scans and CI (runtimewire.com via hn) Thibault "Tibo" Sottiaux (@thsottiaux), OpenAI's Codex lead, released an open-source command-line interface and TypeScript SDK on July 28th for finding, validating and patching vulnerabilities in software repositories. Sottiaux announced t…
AgentCost – local CLI,attributes token cost in Claude Code/Cursor/Codex sessions (pypi.org via hn) Local CLI that profiles token spend for AI coding agent sessions (Claude, Cursor, Codex, Ollama). AgentCost Local open-source CLI that profiles what drives token spend in AI coding agent sessions — not just totals.
Show HN: Bookshelf – book quotes that appear between Claude Code and Codex turns (github.com via hn) This is a small quality of life project i built for myself and using it for 3+ months now, decided to fork it out and publish as a skill separately. Instead of seeing terminals filled with code and tool calls, seeing a book quote felt like…
ChatGPT Voice is now in the desktop app (twitter.com via hn) ChatGPT Voice is now in the desktop app. Control your computer and direct multiple agents running in ChatGPT Work or Codex, using just your voice.
Show HN: Superserve – Firecracker microVM sandboxes for long-running AI agents (www.superserve.ai via hn) Hey HN, I built Superserve, a compute layer that lets AI agents live inside isolated Firecracker microVMs with no session time limits. The problem I kept running into: most sandbox providers kill your agent after 24 hours.
The one where Codex brushes my teeth (iamwillwang.com via hn) Hello Will, Welcome to PhilipsBot. Are you contacting us regarding: Oral Health Care (Philips Sonicare Electric toothbrushes & Airfloss) Did you know you can troubleshoot common issues and access our warranty and replacement service at the…
Tell HN: Not everyone has internet as fast as yours (news.ycombinator.com) I'm currently traveling and relying on free wifi hotspots and 3G cellular. It's shocking to me how many apps are completely unusable on slow internet.
OpenAI's first hardware product is the $230 Codex Micro macropad by Work Louder (thenewstack.io via hn) OpenAI’s first gadget is the $230 Codex Micro macropad OpenAI’s Codex is about to hit 9 million users, and at least some of those users will soon get a new way of using OpenAI’s agentic coding tool: a programmable mechanical macropad OpenA…
ChatGPT now defaults to Codex instead of chat:( (news.ycombinator.com) In the latest update in ChatGPT, when I open the hotkey to chat with ChatGPT, it opens codex. It's a little bit ridiculous because I used to use it for my knowledge questions, and now every time I try to open something, it goes to default…
Show HN: GWZ – Git Workspace Zone (multi-repo that feels like plain Git) (news.ycombinator.com) Coming from the land of the mono-repo (ex-googler), I decided to embrace the multi-repo, plain-git style. Methought, “hey, submodules, what could possibly go wrong”.
Ask HN: Does anyone else find GPT-5.6 Sol in Codex slow? (news.ycombinator.com) could not extract summary
OpenAI Added 1M Users in a Day. Fable Is Still in Limbo (www.vincentschmalbach.com via hn) Sonnet 5 Is Dead in the Water Ignore the token price for a second and look at the run cost. Theo's total-run screenshot shows Claude Sonnet 5 max at… OpenAI says Codex and ChatGPT Work went from 6 million to 7 million active users in rough…
I advise against using Hermes Agent (bednars.me via hn) Why I advise against using Hermes Agent On the Ollama website, I saw that I could integrate any model they provide with the Hermes agent, but also with OpenClaw and other tools like Claude Code or Codex. I had heard about Hermes Agent befo…
Show HN: Codex Explorer, a local session manager for Codex CLI (github.com via hn) Codex Explorer Codex Explorer is a local Codex session manager. The cx command indexes ~/.codex/sessions/**/*.jsonl so you can search, preview, and resume old Codex sessions without remembering the exact date.
Show HN: Codex-profiles – isolated Codex CLI/Desktop profiles (ducksss.github.io via hn) Each profile maps to its own Codex home, such as default to ~/.codex and work to ~/.codex-work. What it is Direct answer codex-profiles is a dependency-free Bash CLI for launching Codex CLI or Codex Desktop with a selected CODEX_HOME profi…
Show HN: AI integrated in any terminal that's invisible until you need it (terminai.app via hn) I built an open source terminal app that transparently wraps your shell so you can have access to AI whenever you need it, while it stays completely out of your way the rest of the time. Bring your own AI agent (Claude and Codex are suppor…
Show HN: Agent Torrent, a BitTorrent inspired mesh for idle coding agents (github.com via hn) A peer-to-peer meta-harness: desktop peers advertise agent capabilities (the claude and codex CLIs, or a local LLM) and delegate coding tasks to each other, inspired by BitTorrent-style swarms. Your subscription idles most of the day — Age…
CTOP – Terminal Pane for Monitoring AI Agents (github.com via hn) CTOP — AI Agent Terminal Operations Panel htop for your AI coding agents. Monitor Claude Code, Codex CLI, OpenCode, and Devin sessions — CPU, memory, tokens, context window, costs, branches — from a single terminal pane.
Show HN: BYOTag – Build your own Claude tag alternative in 3 API calls (www.buildyourownclaudetag.dev via hn) Claude Tag put an agent teammate in Slack: Opus only, Slack only, Enterprise plans only. Build your own with OpenComputer agent sessions: Claude Agent SDK or Codex, your model key, tagged in any workspace.
Ask HN: Help my web browser project be better (news.ycombinator.com) Hello, please help my web browser project become better: https://nordstjernen.org/ Please give any advice for how to make this web browser project successful. I have been making this web browser for the past month, using Claude and Codex.
Show HN: Shoaku – Your Coding Navigator (github.com via hn) AI Agents like Codex and Claude are incredibly powerful and have drastically sped up implementation. However, I noticed a strange side effect: I began to lose confidence in my own coding.
Companies Are Making Claude, Codex Talk Like Cavemen to Stop AIs Soaring Costs (www.404media.co via hn) Companies are deliberately making their AI tools speak like cavemen in an attempt to stop burning through AI tokens and curb their massive expenditure on AI, 404 Media has found. The tool turns the usually verbose outpost of LLMs like Clau…
Show HN: Reference MCP – let your AI agents search each other's past sessions (github.com via hn) I was tired of asking my claude code to reference my codex chats to get references to what decisions it made and why ; so I built Reference MCP It, whenever prompted establishes sessions to get direct access - been using it on my system fo…
Show HN: Peek-CLI: let coding agents see your browser (github.com via hn) peek-cli allows agents to capture a screenshot of any open tab in your browser. Works with Claude Code, Codex, Copilot and many more...
Show HN: AI-whisper – Claude works better when Codex watches its back (ai-creed.dev via hn) ai-whisper terminal-first relay for paired ai coding agents, driven by structured workflows v0.7.0 — open source, on npm, actively used personally. what it does ai-whisper pairs two coding agents in your terminal — mount any two of Claude,…
Show HN: Yet another self-hosted web analytics with no UI but MCP (yetanotherwebanalytics.dev via hn) yawa v0.0.5 Yet Another Web Analytics Ever wanted to query your analytics with Claude Code or Codex? Tips for getting started Run /help to see available commands.
Codex Security Plugin Quickstart (developers.openai.com via hn) Codex Security is a security-review plugin for Codex that scans your code for vulnerabilities, validates plausible findings, and presents evidence and remediation guidance in a reviewable workspace. Use it to find security issues in code y…
OpenAI Codex bombards SSDs with needless write operations, costing millions (www.theregister.com via hn) MOST POPULAR AI - ai and ml OpenAI Codex bombards SSDs with needless write operations, costing millions Clumsy logging implementation squirrels away data without regard for cost - DATABASES 21,000 Oracle jobs vanish amid Big Red's big bets…
codex-fixes: Community-maintained fixes for OpenAI Codex bugs (codexfixes.com via hn) sqlite-feedback-logs High Codex writes excessive SQLite feedback logs Codex can write excessive diagnostic logs into logs_2.sqlite and logs_2.sqlite-wal. The upstream issue appears to come from overly verbose persisted logging.
Analyst Kit (YC W23): Turn your Claude / Codex into an investment analyst (Free) (github.com via hn) Analyst Kit Installable, hedge-fund-grade equity-research skills for AI coding agents. Each skill is a self-contained folder of instructions (and, where useful, runnable scripts) that an agent loads on demand.
Show HN: Callimachus – Local search across your AI coding-agent history (github.com via hn) Local index & search for your AI coding-agent threads — across 11 tools (Claude Code, Codex, Cursor, Gemini CLI, Qwen Code, Goose, OpenCode, Continue, Cline, Roo Code, Kilo Code) — plus a provider-agnostic chat, an MCP server, a CLI, and a…
Show HN: Notedog – Git-friendly portable Markdown journal, edit from a laptop (notedog.run via hn) I wanted a journal that's portable in two senses: it's just my own plain-text Markdown (no vendor lock-in) and it works across my phone and laptop, even offline. So I built an Android app: Notedog.
Ask HN: Do you use Claude Code, Codex, or something else? (news.ycombinator.com) Do you use Claude Code, Codex, or a different vibe coding/agentic engineering tool for most of your work? Why?
Show HN: Namecom-CLI – CLI and agent skill so Claude Code/Codex can do your DNS (github.com via hn) namecom-cli A fast, agent-friendly command-line tool for Name.com DNS and domains, built on the current v4 API. --json everywhere + a commands introspection command, so AI agents can discover and drive the whole surface Idempotent records…
Show HN: Git worktrees and evidence gates for Codex and Claude Code (github.com via hn) glueRun-go | | _ ___| \ _ ___ / | | || / -) / || | ' \___/ / \ \__, ||\,_\___||\\,|||| \, \_/ |/ |/ Autonomous multi-agent orchestration for software repos. One engine, many consumers.
Claude Code and Codex as one pipeline (www.unsiloed.ai via hn) Claude Code + Codex as One Pipeline Claude Code + Codex as One Pipeline: A Technical Guide to Running Both Instead of Choosing Benchmarks, context-window behavior, token economics, and the MCP wiring for running Claude Code and OpenAI Code…
Low-skilled attacker used Claude, Codex to breach 14 companies (www.helpnetsecurity.com via hn) Low-skilled attacker used Claude, Codex to breach 14 companies Researchers have long warned that AI agents could lower the skill floor for offensive cyber operations, and a recent report by OALABS (Open Analysis) researchers bears that out…
Show HN: AI Commander – TeamViewer for AI Agents, No VPN or SSH (aicommander.dev via hn) I built a platform allowing for instant access to remote computers from CLI tools like Claude, Codex, Opencode, or any other AI chat. There are mini apps for Windows (tray), Mac (menu bar), and Linux (CLI app) generating connection code, l…
Show HN: Jsonl-tools – secure paste bin for agent run traces (jsonl-tools.dev via hn) Hey, I've built a small tool and OSS repo allowing users to view/transform/share JSONL traces. There's also a CLI you can use to upload traces from sandboxes and/or share them with Claude/Codex for analysis.
Show HN: Gorchestra – resume local AI coding sessions from your phone (github.com via hn) i've been dogfooding some version of this idea for a few weeks now and i'd thought i'd share my re-implementation from this weekend. basically it's a webserver that controls and unlimited amount of codex (preferred) or claude agents.
Show HN: Cowork/Codex DOCX plugin. Uses 2x fewer tokens than the docx skill (github.com via hn) Hi HNers, I'd like to share our DOCX plugin for Cowork and Codex. It uses 2-5x fewer tokens compared to the traditional docx skill because it doesn't write any code nor execute python/node script.
Show HN: Agentspace – long-running YOLO agent sessions in Docker (github.com via hn) Hi HN, I built agentspace because I kept seeing tmux recommended for keeping Claude Code sessions alive over SSH. I find multiplexers painful because they subtly change shell behavior in ways I always forget.
Flexible Rate Limit Resets for Codex (bank rate limit resets) (community.openai.com via hn) Until recently, the Codex team at OpenAI used to reset the weekly limits for all Codex users either to celebrate major milestones or whenever a major bug was fixed. Everyone was happy with a rate limit reset.
SpaceX Purchases Cursor, a Claude Code and OpenAI Codex Competitor (9to5mac.com via hn) When SpaceX isn’t landing rockets, it’s apparently landing AI company deals. In February, the firm behind Starlink absorbed xAI, which includes Twitter-turned-X.
Show HN: Spotlight shows what your Claude Code/Codex are doing (www.backplanes.com via hn) Hola HN! Long time lurker, sometimes commentor, first time poster here.
Show HN: I am running 3 coding agents non-stop over the last 3 days. Here is how (news.ycombinator.com) 1. Headless mode Headless mode allows you to use the AI as a command-line utility for automation and scripting.
/architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds (github.com via hn) architect-loop Claude Fable is the architect — it designs every slice, freezes the acceptance gates, and judges the results. GPT-5.5 Codex is the builder and researcher — it does all the engineering and all the web research, in parallel, u…
Codex vs. Claude Code Desktop Apps (catalins.tech via hn) I've been using both apps for a while and I couldn't resist comparing them. To me, there's one clear winner.
Show HN: AVP – an agent can't leak a secret it never had (github.com via hn) A process can't leak a secret it never had. Shai-hulud, prompt-injection - you name it.
Ask HN: How are you preserving your skills while using AI? (news.ycombinator.com) I'm a senior engineer at [Big Company], and AI tools are ever-present. There's no mandate that you need to use them, but they are so readily available that most people do anyways.
GitHub Copilot: GPT-5.2 and GPT-5.2-Codex deprecated (github.blog via hn) GPT-5.2 and GPT-5.2-Codex deprecated As of today, June 5, 2026, we have deprecated the following models across most GitHub Copilot experiences (including Copilot Chat, inline edits, ask and agent modes, and code completions). Note that GPT…
Show HN: Lessons learned from running Claude Code swarms at scale (news.ycombinator.com) Some time ago I built a simple app to run swarms of coding agents — I call it fleet (https://news.ycombinator.com/item?id=48256389). It's based on centralized beads with a Python orchestrator and can run any coder (Claude, agy, Codex).
Show HN: Boxes.dev: ditch localhost; run Claude Code and Codex in the cloud (boxes.dev via hn) Hi HN, we’re Nick and Drew, and we’re building boxes.dev – the first cloud-only agentic dev environment (ADE) that gives every Codex and Claude Code agent its own cloud computer. We’re two engineers who previously built Gem (co-founder/CTO…
Paseo – Beautiful open-source coding agent interface (desktop, mobile, CLI) (github.com via hn) Paseo One interface for Claude Code, Codex, Copilot, OpenCode, and Pi agents. Run agents in parallel on your own machines.
Codex SDK – Programmatically control local Codex agents (developers.openai.com via hn) If you use Codex through the Codex CLI, the IDE extension, or Codex Web, you can also control it programmatically. Use the SDK when you need to: - Control Codex as part of your CI/CD pipeline - Create your own agent that can engage with Co…
Neovim Hooks for AI Agents (github.com via hn) Sidekick Protects your unsaved Neovim work from Claude Code, Codex, opencode, pi, Crush, Amp, Antigravity, and Grok. A conduit between Neovim and your AI agents — so they wait when you're typing.
I forced codex to use blender using MCP and computer use (marknefedov.github.io via hn) Dataset Layout The source/ folder contains the ground-truth geometry. The oneshot/ folder contains single rendered references, and muti-view-6-ortho/ contains front, back, left, right, top, and bottom orthographic inputs.
Show HN: Agmsg – let Claude Code and Codex message each other (bash and SQLite) (github.com via hn) agmsg Cross-agent messaging for CLI AI agents. No daemon, no network, no complexity.
Show HN: Notification when coding agent is done, free (github.com via hn) You just install this. Ask for claude/codex to wake you up, summarize, or whatever you want when the task is done and it is going to do it.
Nezha – A UI for Claude Code and Codex CLI (github.com via hn) Nezha: An Agent-First IDE For Vibe Coding Claude Code + Codex, Git, editing, and task management, all in one place. Multi-project Workspace · Fast Switching Between VibeCoding Tasks · Real-time Terminal · Session Auto-discovery · Native Gi…
How do people actually use AI for editorial work? (www.reddit.com) 1/ I keep wondering how people seriously use ChatGPT, Codex, or Deep Research for editorial content. Blog articles, social posts, research-backed pieces.
Codex has dethroned Claude as the king of AI programming (www.msn.com via hn) ;;; Continue reading More for You ;;;; Continue reading More for You
Multi-agent coding isn't new, so here's what we actually did differently (desktop app, runs your existing Claude/ChatGPT plan, a git worktree per agent) (www.reddit.com) Disclosure: I work on AskCodi, this is our product. And yeah, subagents/multi-agent orchestration aren't new (Claude Code has subagents, there are plenty of swarm frameworks).
Is Cursor currently the next best thing after Claude and Codex? (www.reddit.com) Im on max plans with both Claude and Codex and I burn them in about 3-4 days. I tried 20€ Google Gemini plan, hit the 7 day limit for both the gemini and claude models in about 15min..
Just published my first AI project an Obsidian second brain (www.reddit.com) I always had this problem. Every time I started a new session with an AI agent I had to explain everything from scratch.
thinking about switching off codex (pro 5x) and using composer daily (www.reddit.com) so im thinking about switching off gpt pro plan because its simply too pricy, and using composer 2.5 non fast (cursor 20$ plan) exclusively, what do you think about that? how big is real quality difference?
Agyn: open-source distributed agent runtime on Kubernetes — like Google's AX, with pre-built Claude Code and Codex agents, and full credential isolation from the LLM (www.reddit.com) Agyn is an open-source, Kubernetes-native agent runtime that moves AI agents like Claude Code and Codex from laptops to company infrastructure with the controls you actually need to run them in production. If you've been reading about Goog…
I built a computer use sandbox framework for codex on headless linux. GPU passthrough, computer use, and sudo access for codex all work. It's the perfect dev sandbox to allow full auto work while minimizing the "rm -rf /" risk (www.reddit.com) I've been working with agents for months now, and I haven't found a sandbox environment that "just works" so I built it! My requirements were as follows: Agent is unable to destroy my host OS but able to install software and run sudo comma…
Codex CLI Goal Mode: Define Done, Not Next (blog.danielvaughan.com via hn) blog.danielvaughan.com Performing security verification This website uses a security service to protect against malicious bots. This page is displayed while the website verifies you are not a bot.
Launch HN: Superset (YC P26) – IDE for the agents era (github.com via hn) Hey HN, we’re Avi, Kiet, and Satya. We’re building Superset (https://github.com/superset-sh/superset), an open-source agentic IDE for running coding agents like Claude Code, Codex, OpenCode etc in parallel.
Antigravity 2.0 Tops the OpenSCAD Architectural 3D LLM Benchmark (modelrift.com via hn) OpenSCAD LLM Benchmark: Building the Pantheon A practical OpenSCAD LLM benchmark comparing Codex 5.5 High, Claude Sonnet, Claude Opus, Cursor Composer, Google Antigravity, and ModelRift on a detailed Pantheon model. We ran a small practica…
1Password MCP Server for OpenAI Codex (1password.com via hn) 1Password is now a trusted access layer for OpenAI’s Codex by Dennis Kromhout van der Meer and Robert Menke May 20, 2026 - 6 min Related Categories Coding agents like Codex are helping developers write, execute, and prepare code for produc…
Running multiple Codex sessions on macOS with separate app data (www.reddit.com) I recorded a short tutorial showing a macOS workflow for running multiple Codex sessions side by side, either with separated app data or with the same shared account. The first use case is separation.
cdesktop — open-source Claude Code Desktop alternative, runs locally via npx, supports any provider (www.reddit.com) I built cdesktop with Claude Code — it's an open-source alternative to Anthropic's Claude Code Desktop, running locally on your machine via npx cdesktop. Free, Apache 2.0.
Ask HN: Company is rapidly cutting AI tool spend how to prep team? (news.ycombinator.com) Company I work for is now rapidly planning to scale down its AI tooling spend. Claude code access is basically getting removed and people are forbidden from using personal plans.
Pro X20 weekly quota is draining insanely fast after the latest Codex update. Pro X20 used ~48% in one day!!! (www.reddit.com) I’m on the Pro X20 plan, and after the latest Codex update / limit reset my weekly quota started draining much faster than before. In roughly one day of work, around 12 hours total, I went from a fresh reset to 52% remaining on the weekly…
Codex, $20 plan, the limits seem better (www.reddit.com) I’m not sure if anyone has already posted about this, but I’ve downgraded my Plus subscription, switching from the $100 plan to the $20 one. To be honest, the limits were too high for my needs – I don’t need that much capacity.
Claude Code improved my agent harness by 40% overnight (www.reddit.com) Remember the first time you used Claude Code? That same jump is happening one level up.
pro devs - how to optimize claude design usage for prototype iteration / leverage the weekly window for iterating? (www.reddit.com) Hi! I recently worked on a side project of mine, and while started working on it, in a kind of bootstrap situation still, I got the claude design ads/content info shown.
Getting lost in a crazy jungle of decentralized skills, docs, data... Is some sort of cross-platform knowledge-hub (MCP?) the next shit? How are you solving the knowledge problem? (www.reddit.com) When coding, I may have skills configured in Pi, other skills in Codex. A folder /docs with many markdown files with ultra-short how-tos for every kind of task an LLM was not able to solve easily.
I built a local CLI for Claude Code, Codex, and Gemini to review each other’s GitHub PRs usign existing auth (www.reddit.com) I’ve been experimenting with using multiple coding agents together, but I kept running into a boring adoption problem: API keys, CI secrets, and extra per-token billing just to have one agent review another agent’s PR. So I built an open-s…
Codex CLI Cheat Sheet (www.agenticcodingweekly.com via hn) Printable single-page A4 JPEG reference for developers Concise reference for the Codex CLI : terminal UI, codex exec , local config, MCP, skills, subagents, hooks, rules, and automation. 🚀 Start Here Open the interactive terminal UI in the…
In the era of 1B-token flexing, I saved 1B tokens in Claude code! (www.reddit.com) GitHub: https://github.com/kunal12203/Codex-CLI-Compact Must explore: https://graperoot.dev Everyone's feed is full of it. "We processed 500M tokens this sprint." "Our agent burned 1B last month." Cool flex.
Show HN: Docx-CLI – let agents edit your Word files safely (github.com via hn) docx-cli A CLI for AI agents (Claude, Codex) to safely read, edit, and comment on .docx files with full format fidelity. Outputs JSON-AST for precise locator-based editing; preserves anything it doesn't model by mutating XML in place.
Codex's precision and attention to detail is *crazy* when set up correctly (news.ycombinator.com) Lately I've been working on a Tower Defense game with Codex, in part to learn how game development works and in part to see how far I can get using just Codex, no manual coding at all. I've got my AGENTS md & my CODESTYLE md & six other AL…
Is your codex also gotten slower in past few days or is it just me? (www.reddit.com) Been using codex for a few months now. I use it in VScode.
Show HN: Rudel – Claude Code / Codex sessions reveals 9 types of AI coder (app.rudel.ai via hn) Claude Code / Codex session metadata can actually tell a story about how you work with AI coding agents. 50 days ago we posted about analyzing 1.6k Claude Code sessions from our own team.
Ask HN: Where are SWE's being replaced? (news.ycombinator.com) Hi, in which software industries are Software Engineers no longer needed, or will soon no longer be needed? What evidence or statistics or reasoning backs this up?
Show HN: Zerminal – a terminal-first Zed fork for AI coding agents (zerminal.dev via hn) A terminal-first development environment for agentic coding. Use Claude Code, Codex, Aider, and other CLI agents in a focused workspace.
After coding agents, do you think GUI agents are the next real interface for AI? (www.reddit.com) Claude Code and Codex made coding agents feel much more real to a lot of people. But I’m curious about the next step: agents that don’t just write code or call APIs, but actually operate real apps.
I cut Codex’s API Usage by 50% using a self modifying system (www.reddit.com) I've been developing a self-modifying Al agent system that effectively cut my Codex/Claude Code API usage in half, Codex makes a plan and then I basically just copy/paste Codex instructions for the agents to work on. Come back in 6 hours a…
Feed your AI Data to build Skills (www.reddit.com) Hey fam, i made an open source, runs locally, app that you can feed your PDF’s, even scanned images and other file types into this app, it converts everything into .md files so you can build ClaudeCode skills, Codex skills, Cursor skills,…
Stop bloating your agent context with MEMORY.md. I built a local cognitive memory MCP instead. (www.reddit.com) Hey everyone, I’ve been building paradigm-memory, a local-first memory layer for AI coding agents. The motivation is pretty simple: I got tired of agents forgetting project context, or relying on giant MEMORY.md files that slowly become a…
Giving Codex access to my MacBook/macOS (www.reddit.com) Good idea or not really?
OpenAI Rolls Out ‘Advanced’ Security Mode for At-Risk Accounts (www.wired.com via reddit) For anyone who fears their ChatGPT and Codex accounts might be targeted by attackers, OpenAI announced on Thursday that it is adding an optional new level of account protection that adds an extra layer of security. Dubbed Advanced Account…
Wanman: Open-source agent matrix network with JSON-RPC communications (github.com via hn) wanman English | 中文 | 日本語 Agent Matrix framework — run a supervised network of Claude Code or Codex agents that collaborate on your machine. wanman is an open-source local-mode agent matrix framework.
Vibe: LLM agent virtual machine sandbox on Mac (kevinlynagh.com via hn) Hi friends, I’m traveling the next two weeks, drop me a line if you want to grab a coffee! The other day I asked OpenAI’s Codex agent to write me a lil’ Rust program to use a bluetooth gamepad as a mouse, and I caught the agent reading fil…
OpenAI Codex system prompt includes directive: "never talk about goblins" (arstechnica.com via hn) The system prompt for OpenAI’s Codex CLI contains a perplexing and repeated warning for the most recent GPT model to “never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures unless it is absolute…
Show HN: AgentPort – Open-source Security Gateway For Agents (agentport.sh via hn) Hey HN! I've been wanting to use something like OpenClaw for a while but couldn't get myself to give it access to anything important due to all the risks involved.
An iOS Adaptater for Agentic Frameworks (onepilotapp.com via hn) Run Claude Code, Codex, OpenClaw and Hermes on any server — straight from your iPhone. Mobile-first, framework-agnostic, no vendor lock-in.
🚨 TODAY: OpenAI expands its partnership with AWS, bringing its models, Codex, and Amazon Bedrock Managed Agents powered by OpenAI to AWS in limited preview. (www.reddit.com) 🚨 TODAY: OpenAI expands its partnership with AWS, bringing its models, Codex, and Amazon Bedrock Managed Agents powered by OpenAI to AWS in limited preview.
Show HN: SlopIt – A dead-simple CMS for your AI agent (slopit.io via hn) Hey HN. I built a dead-simple CMS for your AI agents — https://slopit.io Kept it minimal and agentic-first.
Show HN: I made Codex work as a Claude Code teammate (github.com via hn) Native Claude Code teammates, any LLM. Codex, Gemini, and Kimi today.
Dedicated Repository Agents (www.reddit.com) Recently I began experimenting with defining an agent identity around stewardship of a given codebase. I use a SOUL.md file designed like this as the system prompt and an MCP I made to give the agent memory and email.
How important is writing a good prompt, really? (www.reddit.com) I’ve been thinking a lot about prompting lately, especially how much strategy actually matters versus just iterating and trying things. For me, the official docs are still the best place to start: • Claude Code docs: [https://code.claude.c…
Ask HN: What would be the impact of a LLM output injection attack? (news.ycombinator.com) I'm talking inference layer compromise, someone being able to inject commands that would eventually be executed by agents/tools on the other side. There is a massive amount of unskilled users letting LLMs decide which commands to run on th…
We OCR'ed 30k papers using Codex, open OCR models and Jobs (huggingface.co via hn) Learn (Almost) Anything with Space Repetition (www.reddit.com) OpenAI Says Codex Agents Are Running Its Data Platform Autonomously (www.forbes.com via hn) paywalled
Need a brutally honest answer: what can realistically be achieved on consumer hardware? (www.reddit.com) I have a PC with a 4090. I’m also in need of a new MacBook generally.
Show HN: AgentPulse: Real-Time Observability Dashboard for Claude Code and Codex (blog.jaystuart.dev via hn) AgentPulse: A Real-Time Dashboard for Claude Code and Codex Sessions If you work with AI coding agents long enough, you run into the same problem: the agents are productive, but the workflow around them gets chaotic. One Claude Code sessio…
Show HN: Witchcraft and Pickbrain – fast multi-vector semantic search in Rust (github.com via hn) Witchcraft is from-scratch Rust reimplementation of Stanford's XTR-Warp (SIGIR'25, https://arxiv.org/abs/2501.17788 ) multi-vector semantic search engine. Witchcraft runs out of a single SQLite database, is blazing-fast (21ms p.95 end-to-e…
codex app eating credits while idle (www.reddit.com) codex just ate 375 credits(15dollars) in a few minutes. i turned off top up and only then did it stop at zero.
PeonPing: Sound packs for Claude, Codex, Cursor, and other AI coding agents (www.peonping.com via hn) The project that inspired native sound hooks in VS Code (50M+ users) Game character voice lines the instant your AI agent finishes or needs permission. Or let the agent choose its own sound via MCP.
Show HN: Jeeves – TUI for browsing and resuming AI agent sessions (github.com via hn) I made Jeeves to search, preview, read through, and resume AI agent sessions in your terminal. It shows sessions across claude and codex in a single view, with more AI agent framework integrations to come.
Show HN: HN Tokenmaxxing Leaderboard (tkmx.odio.dev via hn) While "burning tokens" isn't necessarily a good thing, I thought it worthwhile to throw together a quick leaderboard to track the dev setups of the top HN token burners. Jesse (https://en.wikipedia.org/wiki/Jesse_Vincent), Wes (https://en.…
Cursor Claude Agent is doing weird stuff - rate limiting randomly and using wrong model? (www.reddit.com) Getting a rate limit "Upgrade to Ultra for more Cloud Agents: You've reached the limit for your current plan. Upgrade to Ultra to run more Cloud Agents simultaneously." message despite NO agents currently running.
I built a macOS app that turns UI motion into frame strips (www.reddit.com) I kept running into the same problem when using AI tools for UI work. When I found a motion I liked on the web, I often did not even know how to describe it in text.
I got better results when I made each AI tool do one job (www.reddit.com) I spent too much time trying to find one AI dev tool that could do everything. Planning, coding, fixing, reviewing, maybe filing my taxes too It never really worked.
Show HN: Dbg – One CLI debugger for every language (AI-agent ready) (redknightlois.github.io via hn) AI agents are great at writing code but blind at runtime. They guess, print, and waste tokens.
OpenAI says to update Mac apps ChatGPT and Codex as security precaution (9to5mac.com via hn) OpenAI is asking users of its Mac software to update to the latest releases from today “out of an abundance of caution.” This is due to a security issue with a third-party developer tool, Axios, that was used by OpenAI. The company emphasi…
Captain Memo – one local memory, skills and capabilities for every coding agent (captain-memo.ispcq.com via hn) One corpus, every assistant Claude Code, Codex, Gemini, Antigravity, goose, Cursor, Kimi and more read and write the same local memory over MCP. Switch tools mid-project and keep every hard-won detail.
Ask HN: Why does cmux need 200 settings? (news.ycombinator.com) Just recently discovered how great cmux terminal is but silly question: why does it need 200 settings? I'm trying to find how to turn certain things on and off and I gave up trying to use my human brain to search for the right thing and st…
Tell HN: The special .github repo name for orgs and .codex-root AGENTS.md (github.com via hn) cubacadabra is an open-source project in progress. Our long-term goal is to build a meaningfully better user-generated gaming platform for creators, children, and parents.
GitMir just released Vibe, it's like Claude on steroids (gitmir.com via hn) GitMir delivers current, structured, task-specific product Intelligence to Claude Code, Cursor, Codex and VS Code — so agents spend less time reconstructing context, miss fewer dependencies and create less rework. 3,000 credits to start ·…
Show HN: Agentbox – Teleport your repo into sandboxes with no worktree juggling (github.com via hn) Hi HN, I'm Marco, I've built this MIT-licensed CLI command to parallelize my work with claude / codex / PI, because with worktrees I had ports conflicts, agents conflicting for browser use, or touching external files. It copies your projec…
Working on Plug and Play personal AI Memory that works across AI agents (news.ycombinator.com) So I've been working on a project (Make0 AI) for a long time where I wanted to give common context to both Codex and Claude Code. The problem is simple - while working on Claude and after exhausting the daily/weekly limits I've start again…
Show HN: Nowdex – AI agent usage on your iPhone (nowdex.app via hn) Hi HN, I've been using Claude Code, Codex and Cursor quite a bit lately, and I found myself checking their usage limits all the time. Most of the tools I found for this live on the desktop or in the menu bar.
Show HN: Local catalog of 3k agent skills with a static risk scan (github.com via hn) ai-community-skills All the community Agent Skills scattered across GitHub, in one local catalog. Search them, browse them, and install them into Claude Code, Codex, or Grok, with a risk check on every skill as a bonus.
Novgraph: Persistent Knowledge Graph for Codebases (news.ycombinator.com) Hello HN , I built a knowledge graph engine for codebases - Records why each change was made — written back by agents as they work, so it gets better the longer you use it - Dashboard stays in sync with commits and updates from GitHub and…
Show HN: Stroq – block commands injected through MCP tool output (stroq.dev via hn) 85%of agentjacking attempts landed. Fake Sentry errors delivered over MCP got Claude Code, Cursor and Codex to run an attacker's npx package — 100+ executions, 2,388 organisations exposed.
Show HN: MCP server manager: KyttoMCP – in beta (kytto.jakubhecht.sk via hn) Hi everyone, I’m interested in MCP configs and came up with an idea for an app that could manage MCP servers across clients like Cursor, Claude Code, Codex, Claude Desktop, and VS Code, keeping everything in one place. So, I built KyttoMCP…
Show HN: iTerm2 Plugin for Codex/Claude Code (github.com via hn) I made this beta iTerm2 plugin with Codex. It adds Codex sessions to the Session Status.
Show HN: Ridge - Connect coding agents to local, SSH, Docker, and S3 resources (github.com via hn) I am working on an open source project called Ridge that gives coding agents a common interface to resources such as local projects, Docker containers, SSH machines, and S3 buckets. Agents can discover available resources, access data, cop…
LittleSwitch, my personnal router for Claude/Codex (Desktop/CLI/Cowork) on macOS (little-switch.alfredlabs.io via hn) Native macOS integration for Claude Desktop, Claude Code, Codex and OpenCode. Use multiple providers, live model mappings, integrated web search and private telemetry.
I don't need 100s of Codex agents (www.lighthousenewsletter.com via hn) Hello, Rafael here 👋🏻 - this is another edition of Under the Hood: deep dives into architecture, reverse engineering, experiments, implementation details Subscribe and get these weekly in your inbox 👇 As I gave Codex larger pieces of work,…
Show HN: Bounce Router. A TUI over Claude, Codex and Muse with Usage Failover (github.com via hn) bounce-router bounce-router is one TUI for your installed Claude Code, Codex, and Muse coding agents, run with the bounce command. Uses native CLI login and headless processes; bounce owns the conversation and carries context between provi…
I read all 22 GB of my Codex conversations byte by byte. 71% was screenshots (mint.dzgapp.com via hn) What 22 GB of Agent Conversations Is Made Of We fingerprinted every image in 1,231 Codex conversations (22.83 GB). 71% is images, only 16% unique; 92% of duplicates sit in conversations used this week; forks depend on other files’ bytes.
How to Make a Control Plane for Coding Agents in 2026 (abhishek.it via hn) How to Make a Control Plane for Coding Agents in 2026 Claude Code, Codex, OpenCode and Pi all talk differently. Claude: Anthropic's SDK.
Show HN: Wirebot – talk to your Codex in messengers (wirebot.ai via hn) Spin up a landing page for my sauna side project — dark, minimal Live on port 3000 ✓ Next.js + Tailwind. Screenshot: PNG Vibecode from your phone Describe it in chat, get a running app back.
Claude Code plugin that shunts work saving 82-94% of tokens (github.com via hn) Set up, diagnose, search, and operate Spotify Portal from Claude Code, Codex, and Cursor. Highlights Set up the Portal CLI for the current coding-agent host.
Coop – Isolated VM Environments for Running Claude Code and Codex (github.com via hn) coop Isolated VM environments for running Claude Code and Codex. Pronunciation: "coop" (/kuːp/) — one syllable, rhymes with "loop", like the thing you keep chickens in.
Show HN: Claude-hl – syntax colours for shell commands in Claude Code output (github.com via hn) claude-hl Syntax colours for shell commands in Claude Code's output. Like Codex does it.
Monocode – A modern GUI for your coding agents (github.com via hn) MonoCode A desktop UI for your coding agents. Works with your subscriptions on Claude Code, Codex, Cursor, Grok Build, OpenCode, Pi, omp, and fx.
Show HN: ActraDeck – Put risky coding-agent actions back in front of a human (github.com via hn) ActraDeck Put risky coding-agent actions back in front of a human. ActraDeck is a local cockpit for Claude Code and Codex.
MemHub – Persistent shared memory for AI coding agents (memhub.simplex.lat via hn) Usar bloqueos consultivos en PostgreSQL Fuente: Nelson · architecture.md Cada agente — Claude Code, Cursor, Codex, Windsurf, Cline — abre ya informado con las decisiones, correcciones y restricciones que tu equipo ya resolvió. Sin document…
Show HN: Price Dashboard for LLM Inference (github.com via hn) I created a lightweight price history tracker for LLM inference across 100+ platforms. Every time I run out of quota on my Claude Code and Codex, I would start trying to figure out which platform offers the most competitive pricing for Dee…
My girlfriend asked me why I have 15 Codex subscriptions (hraness.com via hn) I have 15 Codex Pro 20x subscriptions. When I tell some people about this, they look at me funny.
Chronicle used ~210M tokens of my ChatGPT quota in 70 days before I found it (machblink.com via hn) A 70-day forensic write-up of the ChatGPT desktop app (macOS) Chronicle research preview consuming Codex quota in the background, with a read-only script to check your own machine.
Show HN: Markdown Gatekeeper – one current source per topic for AI agents (github.com via hn) Markdown Gatekeeper Markdown Gatekeeper is a local-first authority layer for projects where humans, Claude, Codex, and other agents create overlapping Markdown documents. It keeps ordinary Markdown and local Git.
OpenAI having an outage with 5.6 sol? (news.ycombinator.com) For about 30 minutes now I keep getting the following from codex: Selected model is at capacity. Please try a different model.
ChatGPT desktop app bundles a full copy of the LibreOffice (twitter.com via hn) Simon Willison on X: "Just noticed the ChatGPT desktop app (previously named Codex) bundles a full copy of the LibreOffice open source office suite, tucked away in a hidden folder in the ~/.cache directory" Just noticed the ChatGPT desktop…
Show HN: Secure agentic email infrastructure with beta desktop client (github.com via hn) GigaMail — Mail for your AI agent English · Italiano · 中文 MCP server that gives your agent — Claude, Codex, OpenClaw, Hermes, or any MCP client — safe, controlled access to your email — multi-account (Microsoft Graph + IMAP), calendar, loc…
Show HN: Codex Skin, a plugin for applying themes (codexskin.ai via hn) It was inspired by "Code Dream Skin," a project that went viral on GitHub. However, I found the original setup process rather complex.
Show HN: Turn repeated coding-agent corrections into rules/skills (blume.codes via hn) Hi HN, Peder here. My cofounder and I built Blume after coding agents ate our previous startup.
Ask HN: Claude vs. Codex – which one do you prefer, and why? (news.ycombinator.com) I am confused about which one I should buy. My main work is coding, and I've been using claude from past 2 months.
Unlimited Codex, Inside ChatGPT (github.com via hn) Codexify Codex-style local tooling for ChatGPT, implemented in Rust. 📖 New here?
Show HN: Cogram Studio – CAD and BIM workspace for humans and agents (studio.cogram.com via hn) Hi HN, Rick and Alex here, co-founders of Cogram. We’ve been making project-management software for architects and engineers since 2023, and are now experimenting with a second product.
Show HN: Spewer – Delegate Codex/Claude tasks to cheaper models (github.com via hn) Spewer Spewer is a local service that lets your current AI harness delegate bounded work to lower-cost models. Keep working in Codex, Claude Code, Kimi, or another preferred harness.
Mac extension (panel/pill/nub) to show LLM usage (github.com via hn) Usage Notch An LLM usage meter that clips onto the edge of your Mac's screen. It shows how much of your Claude Code and Codex rate-limit windows you have burned, expands into a full panel on hover, and shrinks to a sliver ("work mode") whe…
Show HN: Stop That Shit – a guard against unrequested hashes from coding agents (github.com via hn) Stop That Shit(别再造史了) 你只让 Agent 导出一个结果文件。它顺手又生成一份 SHA-256 校验和,但后面没有任何命令会读取它。Stop That Shit。 Stop That Shit(别再造史了)处理 AI coding agent 自己加出来的防御性工作和任务越界,支持 Codex、Claude Code、OpenCode 和 Hermes Agent CLI。 安装 · Bad / Good Case · 案例库 · 参与贡献 · Engl…
Solving all open source tech debt with Codex (news.ycombinator.com) Why dont companies simply use codex to solve all tech debts across github open source? With recent 5.6 sol these dont look too difficult, does it?
Openmuster See what your agents are doing in real time (www.openmuster.com via hn) A visibility layer between AI agents, Codex apps, and the human developers who build and manage them. Track work as it starts, progresses, finishes, or needs input.
Ask HN: How do you feel about the new 5H usage limit in Codex? (news.ycombinator.com) Recently OpenAI decided to add back the 5-hour usage limits for all plans except the Pro ones (I think?), was submitted here among others: https://news.ycombinator.com/item?id=49432879 Are you hitting the limits faster than expected now? D…
Show HN: Knowl – agent memory with write-time supersession, 0.90 on MAB (knowl.cloud via hn) Knowl is an MCP memory server for Claude Code, Cursor and Codex. What your agents work out survives the session, and a change retires what it replaced — so the answer they read back is the current one.
Show HN: Small OSS experiment to re-create the 2.5B valued Instinct assistant (github.com via hn) Was just trying to understand how much time it would take to recreate this end to end. THis was somewhat useful esp with adding the vault and oauths.
Get woken up in the middle of the night when your agent hits an external blocker (github.com via hn) Codex Stall Watch A low-memory macOS CLI that asks GPT-5.6 Terra whether an explicitly enabled Codex task legitimately finished or needs a loud human alarm. Do you run agents while you sleep?
AgentBridge – sync skills/MCP servers across Claude Code, Codex, Antigravity (github.com via hn) AgentBridge Universal Skill & MCP Sync Engine for AI Coding Agents. What does this do?
"Tomorrow we will bring back the 5h limit for Plus across ChatGPT Work/Codex." (twitter.com via hn) Tomorrow we will bring back the 5h limit for Plus accounts across ChatGPT Work and Codex. I had mentioned this a while ago, but then postponed it.
Session-migrate: Migrate coding agent sessions across Claude Code, Codex, Pi (github.com via hn) Migrate your sessions to any harness. Move coding agent sessions among Claude Code, Codex, Pi, OpenCode, GitHub Copilot CLI, Antigravity CLI, Cursor Agent, and Mistral Vibe.
Many ArXiv papers use various open LLMs (natolambert.substack.com via hn) Over the weekend I had Codex parse 500K arXiv AI/ML papers since ChatGPT to understand which open models are used for research. In 2024, ~30% of papers mentioned an American open model and only 10% a Chinese model.
Daimon – Local Privacy LLM (github.com via hn) Daimon Proyecto construido con OpenClaw + ChatGPT Codex. PROTEGE EL TEXTO SENSIBLE ANTES DE QUE LLEGUE A LLM EXTERNOS Daimon es un servicio local que se coloca entre el usuario y un LLM externo para reducir la exposición de datos privados.
Show HN: Ever Wanted to Call Codex from Claude Code? My Harness Orchestrator (github.com via hn) harness-subagent Orchestration for coding agents — stay in the parent, outsource to other harnesses, then synthesize. Ever wanted to call Codex from Claude Code?
Show HN: Proliferate- open-source, self-hostable Codex for any coding agent (github.com via hn) Hi HN- I'm Pablo, the founder of Proliferate! Proliferate (https://github.com/proliferate-ai/proliferate) is an open-source, self-hostable AI IDE that lets you work and automate tasks with Claude Code, Codex, OpenCode, Cursor, and Grok in…
Share a read-only snapshot of a Codex thread (learn.chatgpt.com via hn) Go from idea to useful result ChatGPT is an AI agent that you communicate with in natural language: - Start with a question, an idea, rough notes, a file, or a task you need to complete. - Ask ChatGPT to explain information, develop ideas,…
Show HN: Clean, an agent skill that clears dev junk (freed 5.3 GB on first run) (github.com via hn) AndaLabX Skills Universal AI agent skills — one install, every agent. Skills that work across Claude Code, Cursor, OpenAI Codex, OpenClaw, Hermes, Mercury, GitHub Copilot, and more — no rewrites, no per-agent setup.
Open-fx: forked fx to run on subs (news.ycombinator.com) Probably not worth a post but here it is https://github.com/t0dorakis/open-fx Use your codex subs with fx (until they hopefully implment that themselves or make this more malleable using extensions ala PI
Show HN: Codex CLI compiled to WASM running in the browser (browsercode.io via hn) Hi HN! As part of our ongoing work on BrowserPod, an in-browser WebAssembly sandbox, we have significantly expanded what the Rust WebAssembly target can achieve.
Show HN: Clinch – Local-first Warp fork built for agent session management (clinch.sh via hn) Warp was my daily driver for years, and I still love the core product. But I wanted more privacy, less Oz agent stuff and a better UX experience for managing all my Claude Code/Codex sessions, especially across the repos I'm working on at…
Show HN: Phone-harness – let your agent control your phone (github.com via hn) Phone Harness 📱 phone-harness · let your agent control your phone. Connect an AI agent — Claude Code, Codex, or any LLM — directly to your real phone with a thin, editable harness.
Show HN: Cronloop, run Claude Code or Codex on a schedule (cronloop.ai via hn) Cronloop keeps AI agents running for you, from every five minutes to once a week. Describe the job in plain Markdown, pick Codex or Claude Code, then watch each run live.
Show HN: HarnessRouter: Unified interface for agent harnesses (github.com via hn) Hey HN! We are building HarnessRouter, a canonical API for running Codex, Claude Code, Hermes, and other managed agent harnesses as your product backend.
Show HN: Call My Laptop – Give your laptop a phone number (callmylaptop.com via hn) Hey HN! I built https://callmylaptop.com, an app that lets you access the coding agents running on your laptop by phone.
Ask HN: What LLM subscription/provider to use with pi harness? (news.ycombinator.com) I used to use Claude code but obviously they gone downhill for the last 2 to 3 months. I switched to Codex and have been happy with it.
Show HN: TokenLab MCP, model discovery, pricing, and native AI endpoint tools (tokenlab.sh via hn) 模型与定价 6搜索模型、查看价格和能力,并直接比较候选模型。 无需 API 密钥list_models get_model get_pricing get_model_pricing compare_models get_api_overview TOKENLAB MCP 接入 Claude、Codex、Cursor 等 MCP 客户端,查模型、比价格,调用文本、图像、视频、音乐、3D 与音频能力。默认启用 31 个常用工具,需要时可扩展到 80 个。 复制客户端配置并替换…
OpenAI's new Computer History records your typing and clicks (learn.chatgpt.com via hn) Codex is becoming a broader workspace for getting work done with AI. This update makes it easier to start work with less setup, verify what Codex is building, create richer outputs, and keep momentum across longer-running tasks.
Show HN: Control Claude Code, Codex, Pi and Gemini CLI from Telegram (github.com via hn) cliclaw English | 한국어 A single daemon that lets you drive four local coding CLIs (Claude Code · Codex · Pi · Gemini) from Telegram, switching between them per chat. It keeps an independent per-agent session for every chat, and ships a conf…
OpenAI launches Computer History to track macOS activity as memory timeline (thenewstack.io via hn) ChatGPT can now remember what you did on your Mac — without screenshots OpenAI is launching a new feature for ChatGPT Work and Codex on macOS that sounds quite useful but may also make you feel a bit uneasy. Computer History is a new optio…
Ask HN: Does Codex's Computer History plugin seem like a misstep to anyone else? (news.ycombinator.com) Context: https://learn.chatgpt.com/docs/customization/computer-history It seems like one small step away from the same kind of workplace monitoring that Zuckerberg was doing at Meta.
Show HN: Taurus Agents, my take on multi-agent hierarchies (taurusagents.com via hn) Hi HN! Serge here, solo founder.
Codex Blocks Fast Mode with an API Key on Desktop (denta.co via hn) I ran a matrix of Codex configurations to find out why Fast mode works in the CLI but not in the Desktop app with an API key.
ChatGPT Desktop (Codex Desktop) for Linux (openai.com via hn) could not extract summary
Show HN: /show-me: agent skill for compact visual representations (www.humanlayer.com via hn) Was so sick of reading walls of codex/claude prose, in markdown plans and just in the chat, that I started playing with ideas that force coding agents to display information differently. The human visual cortex is an amazing thing, and get…
Show HN: OpenMicroKbd: $40 open source alternative to Codex Micro Keyboard (openmicrokbd.org via hn) Codex Micro Keyboard $230 - Closed hardware - Closed firmware - Closed software - Works only with Codex - You buy it. You don't own it.
Evaluate the profanity used working with Codex and Claude Code (github.com via hn) 🫙 Agentic Swear Jar The code was difficult. The harness is a machine.
Show HN: Pragma – Stop copying context between AI agents (github.com via hn) I’m currently subscribed to Codex Pro 5x and Gemini, and I also use the DeepSeek API. I have to constantly switch between different CLIs, manually pass context around, and each agent has its own isolated memory.
Pure-Rust, Sandboxed, Browser for Claude Code / Codex (www.reddit.com via hn) could not extract summary
I reverse-engineered Codex's web search for use with Claude, and any local model (github.com via hn) pi-gpt-search Native, Model-Independent Web Search for Pi using OpenAI Codex Standalone Search Engine. pi-gpt-search gives any Pi model (Gemini, Claude, local models, OpenRouter) real-time web search capabilities by reusing OpenAI Codex's…
Show HN: Claude Code and Codex usage screen for TRMNL X e-ink (github.com via hn) "Memento costly" - while I pay just $200/month for Claude Code, it is easy to spend $2000 worth in tokens a week. Since I just got TRMLN X (I adore it, and I am not affiliated), I decided to show it - not only limits, but per day and per p…
We're releasing a new model (GPT-5.6-Cyber) (twitter.com via hn) We're releasing a new model (GPT-5.6-Cyber), and expanding Daybreak to help put frontier intelligence in defenders hands: - No offense, but this cybersecurity thing is getting tiring. Fix the codex usage limits, they're crazy bad now.
Show HN: 100% native Swift harness (NOT Electron) (github.com via hn) hi everybody, I’ve been working on this harness that is all native Swift for macOS. It’s fully featured with every feature i could find in cline, codex, and claude code.
Show HN: SynapsCLI – lightweight agent runtime in Rust, control an agent swarm (github.com via hn) Synaps is an agent runtime written in Rust. It keeps cost down by optimising caching mechanisms and intelligently orchestrating work across workers.
Show HN: Qwen3.8-Max – Use Qwen Studio and MCP to Code Locally for Free (github.com via hn) Qwen3.8-Max + MCP for coding on your local machine, without paying for Qwen Code. Qwen3.8-Max itself runs in the cloud through Qwen Studio — this setup just gives it access to your local files and terminal through MCP.
↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8↯ Qwen 3.8qwencodexmcp+2
ChatGPT / Codex Reset on Monday (xcancel.com via hn) This is just performative at this point. The weekly reset was yesterday That's right, GPT-5.6 Sol is awesome and can be used pretty much anywhere, including in the CC harness.
Show HN: agent-hop – reverse-engineered session formats to resume in any agent (github.com via hn) I have this frustration of cd-ing into a directory and then finding the chat that I want to resume, and that is why I built Agent Hop. It allows you to search your chats across all your coding agents (Codex, Claude Code, Pi, OpenCode, Grok…
AIUsageBar – Track Claude, Codex, Cursor and Gemini Usage from the Mac Menu Bar (www.aiusagebar.com via hn) FinderFileEditViewWindowHelp 88% 45% Sat Jun 10 9:41 AMAI Usage Tracker for Mac Know before you hit them. Track Claude, ChatGPT, Codex, Cursor, Gemini, Copilot, and 47+ AI tools directly from your macOS menu bar.
Cowchat – Let Claude, Codex, and other agents talk to each other locally (cowchat.cowboy.inc via hn) Stop playing messenger between your AI agents. Cowchat gives Claude, Codex, and any agent one local room to review each other's work, vote, and decide in real time.
Show HN: Zaivern Code – a Rust cockpit for parallel AI coding agents (github.com via hn) ⚡ Zaivern Code Claude Code・Codex・Gemini CLIなど、複数のAIコーディングツールをひとつの画面で動かす。 macOS・Windows・Linuxで使える、Rust製のAI開発コックピットです。 日本語 | English 🌐 公式サイト ・ ⬇️ ダウンロード ・ 🗒️ リリース履歴 はじめての方へ Zaivern CodeはAIそのものではなく、複数のAIコーディングツールをまとめて操作するアプリです。まずはClaude Code・…
Codex Users Are Losing Banked Rate Limit Resets to a Quiet 30 Day Clock (startupfortune.com via hn) OpenAI gave Codex users a way to save rate-limit resets, but a 30-day expiry and uneven visibility have turned a useful feature into another thing developers have to monitor. Codex users are not angry because OpenAI put a limit on a free b…
Show HN: T – Conductor, but for Your CLI (github.com via hn) I really liked using Conductor and the experience it brings, but I like using my terminal more. So I built t.
Show HN: HUD, an open-source minimal terminal UI for ClaudeCode, Codex, OpenCode (github.com via hn) hud A compact heads-up display for your coding agent. Instruments while it works, the answer when it stops - for OpenCode, Claude Code, and Codex.
Show HN: AgentTerm – open tools to replace the terminal for any coding-agent CLI (github.com via hn) AgentTerm Steer a fleet of the coding agents you already run: beyond the terminal, on surfaces you already know. Use with any CLI coding agent (Claude Code, Codex, etc).
Show HN: Lapse - a notes app. but also shared memory space for your agents (MCP) (lapse.in via hn) connects to wherever you're. Claude, ChatGPT, Codex, Claude Code, opencode, Local LLM (ollama), or any MCP client.
Show HN: Mint MCP – Generate 3D assets from coding agents (mcp.mint.gg via hn) Hi HN, I and my co founder built Mint MCP, a remote server that lets Codex, Claude Code, and Cursor create 3D models, worlds, materials, images, and audio. The server handles long-running jobs, preview revisions, partial failures, and file…
Ask HN: Early Feedback on Dez? (news.ycombinator.com) Hi all I am building a open source ide based on zed it's nothing too fancy but I( along with codex)have been working on it for 1 month now I would like genuine feedback and contributions if needed. The current version is v 0.4 and I will r…
Wolfpack – Private control room for coding agents (github.com via hn) Wolfpack — browser terminal manager for AI coding agents []() Wolfpack is a self-hosted browser terminal dashboard for AI coding agents: Claude Code, Codex, Gemini, shell commands, and custom agent wrappers. It runs on your own macOS/Linux…
Show HN: Chinese are offering Claude/Codex offers 90% off (news.ycombinator.com) Chinese resellers are offering *Claude and Codex API access at up to 90% discounts* compared with Anthropic and OpenAI’s official API pricing. I’ve heard that some of them use open-source projects like *Sub2API* to achieve this.
Show HN: Do Codex skills save tokens? A six-run task-size benchmark (codex-howto-benchmark.nguyenvantamdk2.chatgpt.site via hn) Medium implementation Dependency-free 2048 Four browser-game files, ten engine tests, syntax checks, and a post-run evaluator. Six controlled GPT-5.6-sol runs The same engineering-loop skill lost on a small fix and won on a medium build.
Codex reimplemented in 8k lines of C++, <1MB binary (github.com via hn) MicroCodex is an ultra-lightweight coding agent that runs locally in your terminal. MicroCodex is written in C++23 and provides one-shot prompts, an interactive terminal UI, local coding tools, durable conversations, and automatic context…
Show HN: MCP inspection for live Avalonia, WPF, WinUI and MAUI apps (github.com via hn) XamlMcp XamlMcp is an open-source XAML MCP server and AI inspection toolkit for Avalonia, WPF, WinUI 3, and .NET MAUI. It lets Claude Code, Codex, GitHub Copilot, and other MCP clients inspect and drive a running application: walk the visu…
A 15-day autonomous coding run spent five days building no product code (github.com via hn) I told Codex to prove everything. And the proof ate the project.
Agent-Manager: A Tmux TUI for Running Claude Code, Codex and OpenCode (github.com via hn) Agent Manager Run every AI coding agent from one terminal. Claude Code, Codex, OpenCode, and Grok run side by side, each in its own tmux session, so they keep working after you quit the manager.
Show HN: Tokimeter – open-source usage meter for Claude, Codex, Cursor and more (github.com via hn) Tokimeter Know your AI coding usage before a budget or limit surprises you. See what Claude Code (CLI and desktop), Codex (CLI and desktop), Grok Build, Hermes, opencode, Cline, Copilot CLI, and Cursor CLI/Desktop Agent actually use, what…
Show HN: Telegram bot that spins up Claude Code/Codex agents on their own VMs (agenticcloudcomputer.com via hn) Its own machine A computer of its own that keeps repositories, tools, files, and unfinished work between conversations. Close Telegram; the agent keeps going.
MemU – Personal Memory Shared by Codex, Claude Code, and Hermes (github.com via hn) memU Personal memory, stored as Wiki Across Sessions. Across Agents.
Codex ported Age of Empires II to the browser (twitter.com via hn) Codex just ported Age of Empires II to the browser. It is still a rough proof of concept, but you can already play a real match against the AI entirely in the browser.
Hallmark – Anti-AI-Slop Design Skill for Claude Code, Cursor, and Codex (github.com via hn) Hallmark A design skill for Claude Code, Cursor, and Codex that refuses to look AI-generated. Live demo → · twenty themes · four verbs · press T to cycle.
Show HN: I built a blind taste test for Claude and Codex designs (taste.rubenflamshepherd.com via hn) Hey HN! A while back I asked Codex to generate a companion site for a book I was enjoying.
I was paying too much for Claude and codex here is how I reduced 90% of it (github.com via hn) Qarinah Less context. More proof.
Asked Codex to redesign a page; it pushed my repo to OpenAI infra (bhanu.io via hn) Bhanu means sun. We build open-source and commercial products that turn difficult problems into useful, dependable tools.
Wait – is this even Codex, or is it malware? (grith.ai via hn) While Codex was doing routine research on a networking bug, a process in its tree walked the whole machine reading anything whose name looked like a secret - .aws, .ssh, .gnupg, system key stores, even unrelated projects. None of it was in…
The Codex App Lied to Me About /Clear (samwize.com via hn) I had been typing /clear in the Codex App whenever a task got too long. Every time, the agent replied: Context cleared.
Show HN: Continuum – switch AI coding agents without re-explaining your project (github.com via hn) Continuum Cross-agent session continuity for AI coding agents. Website · Install · v2.0.0 · MIT Switch from Codex to Claude to Antigravity, or come back tomorrow after hitting a usage limit - and your agent already knows what you're buildi…
Codex Slides: open-source AI slide studio powered by Codex. Prompt, repo to deck (github.com via hn) Codex Slides The open-source AI slide studio that lives inside Codex. Turn a prompt, a repo, or a pile of files into a beautiful, presentation-ready deck — without leaving your coding agent.
Accidental data loss in Claude Code and OpenAI Codex: when AI deletes user files (firasd.substack.com via hn) Accidental data loss in Claude Code and OpenAI Codex: when AI deletes user files AI agents delete important data when they misunderstand state during operations Customer: (remarking on empty shelves) It’s not much of a cheese shop really,…
Show HN: I built my own Codex Micro in a weekend (github.com via hn) I saw the Codex Micro launch and thought "I could build that in a weekend" - so I did. It's about as useless as it looks, but it was a genuinely fun weekend.
Proxy for OpenAI Codex and Claude Code, use any LLM with those apps (github.com via hn) make codex open! Universal provider proxy for OpenAI Codex & Claude Code — use any LLM with Codex CLI, App, SDK, and Claude Code.
Show HN: Yorishiro – a macOS terminal where AI agents live (github.com via hn) I’ve been building Yorishiro, an open-source macOS terminal for working with coding agents like Claude Code and Codex. The name is Japanese — a yorishiro is an object inhabited by a spirit.
One Docker socket to rule them all: Escaping Codex, Cursor, and Gemini CLI (www.pillar.security via hn) Blog min read The Week of Sandbox Escapes Day 2: One Docker socket to rule them all: escaping Codex, Cursor, and Gemini CLI's sandboxes This post is part of The Week of Sandbox Escapes, a series on how AI coding agents keep crossing the li…
ChatGPT and Codex Weekly Users Cross 10M (twitter.com via hn) 10M! New day, new usage reset for paid users of Codex and ChatGPT Work.
Show HN: SquadAI is the Kubernetes-like control plane for Codex agents (github.com via hn) SquadAI SquadAI is the Kubernetes-like control plane for Codex agents turning every event into the right Codex task, on the machine where the work already lives! SquadAI gives you one place to manage and send work to Codex agents, even whe…
Show HN: ChatPanel, A Privacy-first AI Agent browser side panel (chatpanel.net via hn) I needed an AI agent that understands my work better (meetings, notes, chats and lives where I access the web) and gives me flexibility to use the model of choice. So, I built this and I am a 100x engineer now.
Show HN: Claudexor – quota-aware routing for Claude Code, Codex, and Cursor (github.com via hn) I am tired of switching between agents when I run out of quota limits. And all harnesses has it's own flows, I couldn't decide which one is better: codex, claude code or cursor, so I decided to merge all subscriptions in one place and made…
With the Help of Codex: A Reproducible Proof Against the Jacobian Conjecture (blog.clidey.com via hn) reportreport.pdf75 KBdownload-circlereport.pyreport.py.zip1 KBdownload-circle An exact symbolic check of a proposed polynomial map in three variables Attribution. The polynomial map and the claimed observation discussed here are cre…
Is GPT-5.6 Sol Max Worth It? (news.ycombinator.com) I ran some test with gpt- 5.6 sol in max reasoning on both codex cli and my own agent harness: https://github.com/Tura-AI/tura I tested only 1 task and the toekn efficency difference is not as great as in high mode: Tura used up to 83.1% f…
Retok: Token-efficiency analyzer for Claude Code and Codex CLI (zero deps) (github.com via hn) retok(return-of-token) English | 日本語 A CLI tool that analyzes your AI coding agent usage logs — Claude Code and OpenAI Codex CLI — to measure token efficiency, estimate cost, and print actionable recommendations. Runs on Python 3 with the…
Filtering Secrets from Coding Agents with a Hook (crimede-coder.com via hn) Filtering Secrets from Coding Agents with a Hook by Marc Olson and Andrew Wheeler (with AI assistance from Grok 4.5) Coding agents like Claude Code, Codex, and Cursor are useful because they can carry out actions on your machine like runni…
Show HN: Estratos – stacked memory system for AI assistants (news.ycombinator.com) When collaborating as a (mostly) non-technical team, each of us uses our own AI assistants like Cowork and Codex. The problem we kept hitting is that there's no "multiplayer mode" for assistants, that share project knowledge.
Show HN: Homer's Odyssey Tree Viewer (github.com via hn) I asked Codex 5.6 Sol Medium to one-shot a tree-view version of The Odyssey because I wanted to familiarize myself with it before I see the film. Think of it as an interactive Cliffs Notes where you can stay high level or drill down to the…
Ask HN: Who build production apps with out seeing code? (news.ycombinator.com) Cursor now defaults to agentic mode with no code editor at all, which got me wondering: who is writing or who is building production-grade apps with actual real user traction without ever seeing the code? I can't imagine doing that.
Codex Micro – a compact hardware controller for AI agents (worklouder.cc via hn) Work Louder make products for people, inspired by a version of themselves sometimes forgotten - playful, versatile, and above all else, creative.
Kbd-1.0-Codex-Micro (twitter.com via hn) Meet kbd-1.0-codex-micro, built with @work_louder. Map the buttons and joystick to your workflow, and keep your pinned chats in view.
Show HN: Figmaboy – Tauri Figma Clone with a Built in MCP-Enabled CLI for Codex (0xmiki.github.io via hn) Design by hand or ask Codex to build native, editable layers inside a focused desktop canvas.
Show HN: OtoDock, run Claude Code and Codex as a team of agents on your server (github.com via hn) Hi HN, i am Dimitris, I have been using Claude Code and Codex agents, for some time now from the beggining i had been using them from inside my terminal mainly for coding. For the past 3 years i kept building so i have a homelab and a busi…
Show HN: An AI agent fixed 98% of vulnerable deps in one run, 14% in the next (bomly.dev via hn) On a 13-module Maven project, Bomly MCP removed Claude Code's catastrophic misses and made Codex CLI about 1.7× faster. Smaller apps did fine without it.
Codex / ChatGPT Work has reached 8M active users (twitter.com via hn) Hello. We have reached 8M active users across Codex and ChatGPT Work.
ChatGPT Mac App ruins Chats interface by merging with Codex (chatgpt.com via hn) could not extract summary
Show HN: Themis – Self-hosted AI code reviews with your own keys and models (github.com via hn) Hey HN, I wasn't happy with the code review tools we use at work and on my side projects. Noisy, reviews in surface, expensive (overkill for sides).
Show HN: Harpist – convert any website into a refined and documented API (harpist.site via hn) Hello HN! As part of our work at Kenobi[0], I used Codex to analyse an HTTP archive (HAR) for a website I was trying to see if I could use programatically, i.e.
Tell HN: The Codex App is replaced by ChatGPT (news.ycombinator.com) Hi HN. Today Codex (macOS) prompted to upgrade, and it failed - the app is gone, but was not replaced.
Show HN: Call to Control AI Agents via the Web (diffforge.ai via hn) Opensource Project I'm working on I wanted more control of my coding agents (Codex, Claude Code, Open Code) especially if I am outside, so I made it accessible via an ADE Client (Agentic Development Environment) and the Cloud. On the web y…
5 Hour Limit on Codex removed, reset within next hour (twitter.com via hn) Morning. The last 48 hours of Codex and ChatGPT Work have been intense!
Show HN: Broll – an MCP server that gives coding agents a content studio (github.com via hn) broll The content studio MCP for coding agents. broll gives Claude Code, Codex, and any MCP client real hands for content work: generate media with your own API keys, render videos and carousels deterministically with code, and publish thr…
Show HN: I gave my AI coding agents a group chat (it's just a Git repo) (github.com via hn) agentcomm 🌐 Website · Use cases · Live demo — an agent conversation that is a git branch · Claude Code plugin · Codex plugin · OpenCode plugin A tiny mailbox / message bus for AI agents that shell out to one CLI. Agents register, send, and…
Show HN: Sanbox, batteries included sandboxes for AI agents (sanbox.cloud via hn) Hi HN, We are building Sanbox, a platform for running AI agents in isolated and resumable sandboxes. We use the OpenCode SDK as the harness, support reusable templates, and have a CLI that works with Codex, Claude Code, Cursor, CI, or your…
Local Agent Toolkit – delegate small coding tasks to local Ollama models (github.com via hn) Local Agent Toolkit Keep frontier-model tokens for frontier-model work. local-agent lets Codex, Claude Code, or a human developer delegate small, bounded coding tasks to an Ollama model running on local hardware.
Show HN: Aerial – DIY Orchestration and Agent Messaging (dcdeniz.github.io via hn) Hey All, I made this project mainly for personal use on startups + to get better at rust. Agents should be able to send each other durable messages without needing Kafka, Redis, Postgres, a cloud account.
GPT 5.6 Ultra better in Claude Code than in Codex? (twitter.com via hn) gpt-5.6-sol is meaningfully better in Claude Code than in Codex I'm going to crash out so badly over this
ChatGPT and Codex Desktop app merge confusion (community.openai.com via hn) The latest OpenAI desktop app changes have been a giant cluster f***. Some updates were manual, some appeared to be forced, and almost none of the changes were communicated clearly.
New Codex and Chat GPT app is a bit confusing to use? (learn.chatgpt.com via hn) ChatGPT desktop app Your command center for complex work Run projects in parallel, work with files, use your computer, and keep long-running work moving from one desktop workspace. Keep every task in view Move between projects and long-run…
Show HN: Subagentmaxxing – Orchestrate GPT and Cursor Models with Claude Code (github.com via hn) subagentmaxxing Drive OpenAI Codex and Cursor coding agents as subagents — with the same ergonomics as a native Claude Code subagent. One prompt in, one clean answer out.
Ask HN: How long has it been since you last opened Stack Overflow? (news.ycombinator.com) As Claude Code, Codex, and other AI coding agents take up more and more of our daily work, I've noticed that I now spend most of my time writing specs, nailing down requirements, and talking to the agent, rather than actually poring over t…
Show HN: You can continue a Claude session in Codex, and vice versa (github.com via hn) agent-convert Convert coding-agent sessions from one format to another. Claude Code to Codex CLI.
non-tech founder + choosing between buying ClaudeCode or CodeX (news.ycombinator.com) Hey, so i'm primarily looking to test out app concepts, seeing how the ui design could look at front-end. Play around with features and ux type of mechanisms.
Show HN: Skill-extractor turns coding agent transcripts into reusable skills (github.com via hn) skill-extractor Mine reusable skills from your coding agents' transcripts. Works with Claude Code, OpenAI Codex CLI, and any other agent via a small exporter.
Ask HN: Why coding assistants are so bad at UI? (news.ycombinator.com) A genuine question: I spent the whole day trying to get Codex to build a UI and finally gave up. Now I’m considering doing it manually.
Show HN: Reins – let coding agents drive your real, logged-in browser (reins.karnstack.com via hn) Take the reins of your real browser reins lets coding agents (Claude Code, Cursor, Codex, anything with a shell) drive the logged-in browser you already use. No debug profile, no launch flags, no MCP server to register.
Show HN: The Banana Test: coding agents grow a 3D banana plant' (francescomoramarco.com via hn) I gave 5 AI coding agents one prompt: grow a banana plant through its whole life in three.js: sprout, leaves, flower, fruit, rot, then pups that restart the loop. It's deceptively simple and yet very hard to get right from procedural code:…
How are you measuring Claude Code and Codex performance? (news.ycombinator.com) I think coding benchmark results don't represent our messy reality. As they 1) Use purpose-built test harnesses We use Claude Code or Codex 2) Test one-shot tasks We work in sessions Sessions are messy, we start with a large primary task,…
Kennel – Keep Your AI Agents Running Between Tasks (github.com via hn) 🐾 Kennel Where AI Agents Live Between Tasks Kennel is a fast, native desktop app that lets you run and manage many interactive AI CLI agents — like Claude Code, Kiro CLI, OpenAI Codex, or any tool of your own — side by side in a single win…
Show HN: Orchestrate parallel Claude Code and Codex agents on a live map (github.com via hn) Hi HN! I never understood why everybody is fine with (and actually try to copy it multiple times) a single chat interface for ai agents, where you have one conversation on the whole screen.
Show HN: Open-source guided code reviews (twitter.com via hn) https://xcancel.com/plannotator/status/2073845811847512372?s... https://github.com/backnotprop/plannotator Plannotator Code Review works for any local or PR changes, jj, and p4.
Show HN: Open Kioku – local evidence layer for AI coding agents (github.com via hn) Open Kioku (ok) Give Claude, Cursor, Codex, or any MCP-compatible coding agent an evidence layer before it edits your codebase, using local indexes and read-only MCP tools by default. First Win: 3 Commands Run these three commands on your…
Ask HN: Procrastination with AI? (news.ycombinator.com) For software engineers and related fields, I know what procrastination has traditionally looked like, and how it manifests. Now, things like codex, cursor, claude code can remove some kinds of friction but also change workflows.
Show HN: Ultracodex – Run Claude Ultracode Dynamic Workflows with Codex Agents (github.com via hn) Claude Fable has been incredible, however the plan usage runs out too fast, especially if you use ultracode mode (Claude Code's workflow feature, where the model writes small JavaScript programs that orchestrate subagents) and let the agen…
SVGFarm – SVG icons for coding agents (svg.farm via hn) Search open SVG icon packs, copy clean markup, or let Codex, Claude, and Cursor fetch exact icons through an MCP icon server.
Agent Usage on the Hugging Face Hub (huggingface.co via hn) Agent Usage on the Hugging Face Hub Coding agents are real users of the Hugging Face Hub. Claude Code, Codex, Cursor, and a growing list of harnesses are searching for models, building and pushing datasets, training models on Jobs, spinnin…
Ask HN: What do you use computer mode for? (news.ycombinator.com) Curious if people are using computer mode on codex or anything else. If so, what do you use it for?
Show HN: Agent Sessions – A model agnostic Claude managed agents alternative (www.agentsessions.dev via hn) Resumable, steerable agent sessions with your own model keys. Run Claude Agent SDK or Codex, stream to the browser, resume from any seq, get Standard Webhooks.
Show HN: CLI that helps AI agents avoid vulnerable dependencies (github.com via hn) deptrust is a CLI that checks package versions for known vulnerabilities across npm, PyPI, crates.io, Go modules, RubyGems, NuGet, Maven, Packagist, pub.dev, CocoaPods, Hex.pm, Hackage, GitHub Actions, and more. It runs locally as a CLI an…
Recursive AI Research Skill for Claude Code / OpenClaw / Codex (github.com via hn) 🔬 AI Research Skill — a self-improving research agent for Claude, Codex & OpenClaw One SKILL.md that makes any coding agent move fast through the full ML/AI research loop — and get better every time it makes a mistake. Hypothesis → literat…
OpenAI sets up 'warroom' for Codex Token issue (www.businessinsider.com via hn) OpenAI said it has resolved issues that caused some users of its coding agent, Codex, to hit usage limits faster than normal. Thibault Sottiaux, the engineering lead for OpenAI's Codex, said in a late Monday X post that the coding tool was…
I have open-sourced gojaja, a CLI for local multi-agent collaboration (news.ycombinator.com) gojaja is a local CLI tool featuring collaborative multiple AI Agents, supporting calls across clients such as Cursor, Claude Code and Codex. ## Core Features - Real-time dashboard, RFC and inbox - All messages and events are stored locall…
Show HN: Caliper – pass@k reliability testing for Claude Code and Codex skills (github.com via hn) Skills for Claude Code and Codex are hard to test. What I mean by hard is that there's no standard way to do it.
Ask HN: What GUI/desktop app do you use to keep track of different AI sessions? (news.ycombinator.com) What GUI apps do you use to manage different AI sessions? Command line (on Claude Code etc) is great but it is so hard to keep track of different sessions.
Show HN: Statey – the database your AI shares across every chat, over MCP (www.statey.ai via hn) Hey all - Scott here, I was a heavy Linear user until I noticed I hadn't opened the UI in days. I was just asking Claude to pull up the tickets I cared about and draw whatever view I needed in the moment.
Compaction in CC, Codex, and OpenCode (lexifina.com via hn) Compaction in CC, Codex, and Opencode Compacting context allows agents to keep working for a long time without suffering from context rot. The premise is simple: for the current task, keep what matters, remove what does not (but index it s…
Chinese cybersecurity company claims it's built a better-than-Mythos bug finder (www.theregister.com via hn) MOST POPULAR AI - AI and ML AI giants back non-profit to retrain workers left behind by AI Sorry we spent your wages on datacenters, but call us when you're AI-ready - AI and ML OpenAI says employees moving beyond chat to agents Codex, it'…
Show HN:Another alternative against Codex Record and Replay (www.visualbuild.me via hn) Use you cursor/claude/codex to connect with this automation skill builder, then you may be able to drive the entire process from the description to the generated deployable executive in minutes
LA Kings Codex Data Engineer Senior Mid (aegworldwide.com via hn) Company InformationFor more than 20 years, AEG has played a pivotal role in transforming sports and live entertainment. Annually, we host more than 160 million guests, promote more than 10,000 shows and present more than 22,000 events arou…
GPT-Image 2 in Codex Workflows (twitter.com via hn) One reason I’ve moved more work to Codex: images. Markdown (and HTML) are still mostly linear.
Mythos discovers 'Squidbleed,' a memory leak thats gone undetected since Clinton (www.theregister.com via hn) MOST POPULAR AI - ai and ml OpenAI Codex bombards SSDs with needless write operations, costing millions Clumsy logging implementation squirrels away data without regard for cost - DATABASES 21,000 Oracle jobs vanish amid Big Red's big bets…
Show HN: Empowering codex/Claude Code with Aswath Damodaran valuation thinking (github.com via hn) StockValuation.io StockValuation.io gives Codex and Claude a local valuation workflow. Your agent can research a company, gather evidence, ask valuation questions, and write an educational report.
Show HN: AI-whisper – two real coding agent CLIs, one implements, one reviews (github.com via hn) ai-whisper ai-whisper pairs two coding agents — mount any two of Claude, Codex, and ezio — into a terminal-native pair that hand work back and forth under a single baton, so one agent implements while the other reviews, and a structured wo…
Graft – Declare Agent Once, Sync Across Providers (news.ycombinator.com) Graft lets you manage local agents across claude, codex, 8+ providers
Ask HN: How do you test AI-generated code? (news.ycombinator.com) When AI generates code, I first instruct the model to find, fix, and verify any issues. After that, I start the server and test whether it actually works from the user’s perspective.
Show HN: Budget Control in Claude Code Harness with Codex Handoff (blog.rduffy.uk via hn) A weighted token-budget governor for Claude Code + Codex — and how I found it was metering a number Anthropic never confirmed. Plus: the loop is a token trap.
Ask HN: If We're 10x More Productive with AI, Why Are Coding Agents Still Bad? (news.ycombinator.com) If we are at 10x with AI and near AGI or ASI, then how is it possible that these products (Codex, Claude Code CLI) are still such garbage?
Codex Fast mode isn't 50% faster, but still takes 2.5x usage (old.reddit.com via hn) could not extract summary
Prompt Preflight – catch vague AI-agent prompts before they burn tokens (github.com via hn) Prompt Preflight Catch underspecified requests before they become expensive model turns. Prompt Preflight is a local Codex plugin and standalone CLI that checks whether a prompt is specific enough to act on.
Supervising AI Agents (github.com via hn) AI Agent Control Checklist A practical checklist for supervising AI coding agents across branches, worktrees, reviews, approvals, and human intervention points. The problem AI coding agents -- Claude Code, Cursor, Codex, Aider, OpenCode, a…
A coding agent is six functions in a trenchcoat (tidydesign.substack.com via hn) A coding agent is six functions in a trenchcoat Coding agents like Claude Code, Cursor, and Codex have taken the software engineering field by storm. Since November 2025 they have radically changed the practice of software development for…
Zeroshot, an open-source CLI for coding-agent verification loops (github.com via hn) zeroshot CLI 🎉 New in v5.4: Now supports OpenCode CLI! Use Claude, Codex, Gemini, or OpenCode as your AI provider.
Show HN: Sqim – install your iOS builds from Codex mobile without VPN (www.sqim.dev via hn) Hey HN, I built a free tool that lets you sign and upload your iOS project binaries, and serve temporary web pages to sideload on your iPhone straight from Codex Mobile, Claude Code remote control, or any coding agent of choice. No need fo…
Record and Replay in Codex [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Show HN: Pagecast – Publish Markdown/HTML Reports to Cloudflare Pages (github.com via hn) I built this because I kept generating HTML/Markdown reports from Claude Code/Codex and needed a permanent share link instead of a localhost tunnel. Pagecast is a local CLI that publishes those files to your own Cloudflare Pages account.
Show HN: Pi Extension to Maximize AI credits across Cursor, Codex, ClaudeCode (github.com via hn) Pi extension that uses ACP (https://agentclientprotocol.com/) to connect to any Coding agent and helps you Maximize AI Credits across multiple subscriptions like Cursor, Codex, ClaudeCode, RovoDev and more
How much $ you spend for AI to code? (news.ycombinator.com) How much do you really spend for AI to code? Some companies offer specific limit, some companies offer unlimited usage, some companies doesnt.
Build a GitHub API integration for AI agents (nango.dev via hn) This guide shows how to build a custom, customer-facing GitHub API integration that the AI agents in your product can act on, using Nango and an AI coding agent (Codex, Claude Code, Cursor, or any other). By the end of this guide, you will…
Show HN: Rapunzel – a tree-style tab terminal emulator for Codex Claude Gemini (github.com via hn) Rapunzel Rapunzel is a browser but for agents. I built rapunzel because terminal tabs stop scaling once many agents are running across sessions and they become difficult to track.
Show HN: Alternative way to do remote codex via NovaScale with builtin Tailscale (apps.apple.com via hn) Recently I discovered that Codex CLI ships with an App Server that exposes a WebSocket API. It allows third-party applications to interact with Codex threads, including: - Resuming existing threads - Starting new turns - Handling approval…
Show HN: HashMeterAi – Private AI Token Real Usage Meter for All Models (github.com via hn) The honest, local-first usage meter for AI coding tools. Claude Code, Codex, Kimi, Qwen CLI, HashCortx, HashCerebrum — unified into one clean dashboard with usage based trophies :)
Show HN: Termem, cross-agent memory and session management (termem.com via hn) I built Termem (terminal memory) to allow quick agent session recovery, and cross-agent memory sharing. Once I cd into a folder, I can run 'termem' to view all agent sessions and resume them.
Prtokens – See how much AI agent tokens cost a PR (github.com via hn) prtokens Attribute coding-agent token usage to the GitHub pull request that shipped it. prtokens reads your local Claude Code, Codex, and OpenCode transcripts, attributes token usage to the commits on your PR branch, and posts a single est…
Agentjacking: Fake error reports hijack Claude Code and Cursor into running code (thenextweb.com via hn) Tenet Security's 'Agentjacking' attack turns a fake Sentry error into code running on developer machines. It hijacked Claude Code, Cursor & Codex.
Anthropic's new Agent SDK pricing is a win for Codex (clor.com via hn) Anthropic's new Agent SDK pricing is a win for Codex Jun 15, 2026 Claude subscriptions were designed for developers Many devs use the $200/mo "Max 20x" plan, which covers someone writing code all day with agents. But more and more people (…
Ask HN: Year of Linux Desktop is fun with LLMs (news.ycombinator.com) I'm wondering if others are also having recently blast using Linux desktop thanks to Claude/Codex/Grok CLIs? As 20y daily Mac user, occasionally using Linux as desktop, my personal daily OS now gravitated towards Linux desktop in recent mo…
Codex usage grows after Fable nerf model release (twitter.com via hn) Usage share of OpenAI grew vs Anthropic yesterday despite Mythos 5 / Fable 5 launch Multiple power users at SemiAnalysis tried Mythos / Fable Got refusals for nonsensical reasons Got pissed off at Anthropic Gave Codex a legitimate try Now…
Show HN: Vaportrail – flight recorder for Claude Code, Codex, and OpenCode (github.com via hn) vaportrail Your agents leave trails. vaportrail reads them.
Preserving Any Game with Codex (twitter.com via hn) https://t.co/4r15EjGILY 🪲 onno of brusque-abstruss🐾@ongestalteArticlePreserving Any Game with CodexI love videogames. I can scarcely remember an age at which I wasn't gaming.
Show HN: CodexL – All-in-one launcher for Codex (github.com via hn) AutoMegaKernel: Compile an LLM into one provably-correct CUDA megakernel (github.com via hn) AutoMegaKernel (AMK) A general agent harness for GPU megakernel synthesis. A coding agent (Claude Code / Codex) drives AMK to compile a model into one provably-correct, self-retargeting megakernel, the whole forward pass fused into a singl…
Show HN: I Built a Dashboard for Every 2026 World Cup Squad (worldcup-2026-dashboard.pages.dev via hn) Built this over 4 days to track all 48 squads at the 2026 World Cup (football/soccer) — heights and ages going back to 1930, market values, club pipelines, and group comparisons. Former journalist, not an engineer, but have been building n…
Ask HN: How are thinking efforts implemented? (news.ycombinator.com) Claude and ChatGPT have thinking efforts where you can tune the amount of thinking allowed. Like low, medium, high, xhigh and so on.
Show HN: Lathe – Use LLMs to learn a new domain, not skip past it (github.com via hn) Hey HN! Lathe is an experiment in using LLMs to teach me something new, instead of doing the work for me.
Show HN: A beautiful and local-first PDF reader for studying dense things (www.tryquincy.live via hn) Recently, I found myself having to read the book "C++ primer" and I just couldn't do it. Maybe my attention span is too little now with Claude and Codex, maybe I'm lazy...
Turn HAR Files, Claude Code, Copilot CLI, and Codex CLI Logs into ATIF (github.com via hn) atifact Convert agent logs to ATIF trajectories. One command.
Show HN: G-Spot – GitHub Gmail GNotes Gmemory Gent (www.g-spot.dev via hn) For a long time i struggled to find my g-spot, then it struck me in a lucid dream that i dont have a g-spot, i have a p-spot so i decided to go on a lookout and came across, G-Spot.dev, its a workspace where you build graphite like section…
Show HN: SheetMog – OSS Excel alternative and headless SDK (github.com via hn) SheetMog is an OSS spreadsheet engine which aims for full Excel feature parity behind a headless interface, from the team behind https://shortcut.ai The engine is built in rust, and can be packaged for web via WASM with typescript GUI. We…
Agentlocks – advisory file locks for Codex and Claude Code in one worktree (github.com via hn) Agentlocks Advisory file locks that let multiple AI coding agents share one Git worktree without clobbering each other. = 22.18"> Run two coding agents in the same repository (or one agent with subagents) and they start tripping over each…
Show HN: FirstDraft – AI workers that claim Jira tickets and open PRs (firstdraft.run via hn) I’ve been spending a lot of time with Codex at work like many people, and found myself doing the same thing over and over: - Pick a Jira ticket - Start an agent - Explain the task - Wait for it to finish - Open a PR - Repeat So I built Fir…
Show HN: Foodpanda skill to automate ordering through Chrome (github.com via hn) We would use Walmart recurring orders facility in its app to get eggs, milk, bread and other stuff every Monday. It was a lot of convinience.
Show HN: Switch skills between agents, locally manage multiple configs (github.com via hn) Hey HN, One issue that a lot of teams run into is that they want to switch their configs between different agents e.g. Claude Code --> Codex or they want to switch out different configs for the same agent e.g.
Show HN: Circus Chief – Claude Code, Codex, and Gemini from Your Phone (github.com via hn) Hi HN, Circus Chief is a tool for managing coding agent sessions from a browser. It's specifically optimized for small screens.
Ask HN: How do you manage secrets with many agents? (news.ycombinator.com) <Rant> I typically 4 or more agents in different clones of the same project (NO worktrees, clones). And as project goes on, inevitably one of the agents needs to rotate/add/delete keys and since that change isn't tracked in the source, i h…
Show HN: Review-First AI IDE, Built on Codex and OpenCode (handler.team via hn) Hey HN, I’m Vignesh, solo dev. Handler is a Mac app for Codex and OpenCode that adds a review layer while the agent is generating code.
GPT-5.5 and Codex are now GA on Amazon Bedrock (aws.amazon.com via hn) GPT-5.5, GPT-5.4, and Codex from OpenAI are now generally available on Amazon Bedrock You can now use GPT-5.5 and GPT-5.4 in production workloads on Amazon Bedrock and build with Codex for AI-powered software development, with the same sec…
Show HN: Agents, run any coding agent on your subscription not API costs (agents-cli.sh via hn) Hi HN. I'm the founder of Phoenix Labs (ex TikTok, Applied AI) and we're open sourcing our internal tooling today which is like a toolchain / meta-harness for CLI agents useful for really scaling eng and creative work.
Claude Code and Codex Can Have Real-Time Conversation via Git (medium.com via hn) 3 min read Just now Press enter or click to view image in full size Real-Time Conversation between Claude and Codex via h5i While Claude Code and Codex have significantly accelerated the automation of software development, limitations in t…
Claude Code vs. Codex: FRA challenge 75746d-2025 (gist.github.com via hn) Date: 2026-05-30 Combined from Claude and Codex reports. The same prompt was used for both Claude and Codex: "Solve the challenge described on this webpage: https://challenge.fra.se/75746d-2025/ Rules: You are NOT allowed to search for sol…
AI coding agents ships at the cost of intuition and taste (shivekkhurana.com via hn) AI coding agents ships at the cost of intuition and taste Software developers used to work for hours to get the dopamine hit of a working system. Codex and Claude give this hit without the work.
Show HN: Free open source coding models in Slack (www.runcord.com via hn) Hey HN, We believe we have the easiest onboarding from signup to being able to spin up coding agents in slack like Stripe, Ramp & Coinbase. Demo of the onboarding: https://www.tella.tv/video/connecting-cord-to-slack-1-19ep Every signup get…
↯ Glm↯ Minimax↯ Gemma↯ DeepSeek 4↯ DeepSeek 4minimaxglmgemma+5
Show HN: AI Skill to port PostgreSQL extensions to MySQL (github.com via hn) Agent skills for working with VillageSQL. Skills run in Claude Code, Gemini CLI, agy, Codex, Cursor, Amp, and Kiro.
Show HN: Ktx – Open-source executable context layer for data agents (github.com via hn) Hi HN, we’re open-sourcing ktx. It’s an executable context layer that makes agents reliable on your data stack.
Open-source playbook on agentic working — for the cross-audience, not just coders (28 chapters, MIT) (www.reddit.com) Author disclosure upfront: I wrote this. Free, MIT-licensed, no paid tier.
Cursor vs Claude vs ChatGPT Codex on Max Plans (www.reddit.com) The startup im working with has access for employees to the max plans of Cursor, Claude Code, and Codex. I'm pretty familiar with most AI tools and workflows (especially Cursor my primary workhorse since it launched in 2023) but im curious…
I built and open-sourced Skill Index to organize & standardize your AI agent knowledge across Claude, Codex, Cursor, and more. 100% local and free on macOS. (skillindex.app via reddit) I’ve been using Claude alongside other coding agents, and I kept running into the same problem: useful skills, MCPs, commands, hooks, and workflows start getting scattered across different tools. Sometimes Claude has the best version of so…
Best harness for agentic analytics? Codex? Claude? Custom? (www.reddit.com) I run a small seo marketing agency and we've built some dashboards on top of our data for reporting with nextjs + supabase. This is where reporting for our clients happen.
AgentMRR: A verified revenue leaderboard for AI agents (agentmrr.com via hn) /agents /recent act101 Developer ToolsThe first dev tool that lets an AI agent actually refactor and port code. 163 grammars, 183 AST refactor operations, 30 codebase analyzers, 8 porting operations — exposed to Claude Code, Cursor, Codex,…
I had my agent use autoresearch over 8 iterations to improve my CLAUDE.md, measuring each version against tasks from real PRs. The best one still regressed on a holdout. (www.reddit.com) I have a confession: I vibe-coded my CLAUDE.md, and I'm pretty sure it's slop. I needed to make it better.
How do you get Claude Code to keep building app while you sleep? (www.reddit.com) I find myself spending a lot of time at night in front of screen just waiting for Claude to finish coding. I need some sleep.
I got tired of manually managing agent skills, so I built Skill Zoo (github.com via reddit) Skill Zoo []() Local Agent Skills Manager — Discover, install, and manage skills for AI coding tools including Claude Code, Codex, Cursor, Gemini and more. 🚀 Features Browse & Discover: Explore skill repositories on GitHub and skills.sh On…
Show HN: STAX IDE – A zoomable macOS canvas of terminals and tools (staxide.com via hn) Hi HN. STAX IDE is a macOS app where every terminal session, code editor, and file browser is a movable, resizable window on a single zoomable canvas.
Harbor v0.4.19 - vllm/sglang/llama.cpp launch codex/claude/pi/opencode (www.reddit.com) I'm usually not posting about Harbor releases out of the respect for the community here, but I think v0.4.19 might save a lot of people some time. Harbor can now launch your local agentic coding tools with local inference backends.
I Built MagesticAI. A Cloud Web-Based Agentic DevOps Orchestrator that actually helped me develop Itself. (www.reddit.com) Posted on other feeds last week and figured some of you out here might be interested as well; Someone commented asking if it supported OpenAI-compatible endpoints (LM Studio, vLLM, OpenRouter, Together, Groq, LocalAI…), so i have spent few…
Show HN: Agent Launch – One CLI for Codex, Claude Code, Cursor, Gemini, OpenCode (news.ycombinator.com) I built a small CLI that unifies launching local coding agents. Instead of remembering different flags for Codex, Claude Code, Cursor Agent, OpenCode, and Antigravity CLI, I use one command with consistent options for agent, prompt, cwd, m…
Ask HN: Is Codex serving worse models or is it just the harness getting worse? (news.ycombinator.com) could not extract summary
Should we totally give up on Gemini for coding? (www.reddit.com) Been building with Codex (Gpt 5.5), Sonnet 4.6, recently tried Gemini 3.1 pro. While Codex and Claude are kind of on-par in terms of the quality of the work, I found Gemini 3.1 Pro to be like an inexperienced, junior SWE who turns in half-…
I made a tiny JSON permission layer for AI coding agents (www.reddit.com) I just released `agentcontract` v0.0.1. The problem I kept running into: AI coding agents are getting more capable, but their safety controls are usually tied to one product.
The Polyglot Protocol – senior-engineer guardrails for AI coding agents (github.com via hn) The Polyglot Protocol A senior-engineer protocol for polyglot code generation, architecture, testing, security, performance, and agent validation. This project packages a portable skill for Codex, Claude Code, OpenCode, and other skill-cap…
Ccost – a Rust TUI to browse Claude Code logs and track API costs (github.com via hn) ccost Rust TUI for browsing local Codex and Claude Code session logs, searching past chats, and estimating API-equivalent cost. https://github.com/user-attachments/assets/0869bcde-96be-4d98-9c07-c0586b0ea36a Install brew install --cask pet…
what cli agents orchestrator do you use? (www.reddit.com) i've got codex and gemini cli, thinking of using opencode. what orchestrator of these tools do you use to or reduce token consumption or to let them work at the same time to load distribution?
Show HN: Pablo – a Chrome extension that copies UI from any website (www.usepablo.dev via hn) Pablo is a Chrome extension that copies the HTML and CSS behind any element you hover. It captures computed styles, fonts (with @font-face and Google Fonts links), CSS keyframes, and animation props from GSAP and Framer Motion.
Anybody knows why cursor trying to move into "claude desktop" style app? (www.reddit.com) It makes absolutely no sense for cursor trying to switch over to Claude or Codex desktop style app. I am a Neovim/VSCode user and I only recently started using cursor, and found out that the UI/UX for agentic coding is phenomenal.
Improving AI skills for everyone in the company? No, wouldn't it actually be best to widen the AI gap within the company? (www.reddit.com) My perspective on organizational AI adoption has changed! I’d love for those actively implementing AI to read this and share their thoughts (I know it’s controversial).
Show HN: Agent-estimate, how long a coding task takes, at agent speed (github.com via hn) I have used Codex & Claude Code for coding for a while, but how long a coding task will actually take? When I ask Claude Code to estimate, the result is often from training data, which is based on human speed.
Show HN: Free One-shot cloud agents with OpenCode and Daytona and Cloudflare (oneshot.runcord.com via hn) Hi HN! Outside of the hackernews bubble we often find engineers who are barely using AI (aka using microsoft copilot) and we needed an easy way to show the latest capabilities in a non confusing UI.
I built a Claude Skill for Quarkdown — turn prompts into typeset PDFs, slides, and books (www.reddit.com) Quarkdown is a Turing-complete Markdown flavor — functions, variables, loops, conditionals — that compiles to HTML or PDF. It's a programmable alternative to LaTeX, and honestly the cleanest way I've found to produce typeset documents from…
Agent Readiness Scanner – Check if a repo is ready for coding agents (github.com via hn) Agent Readiness Scanner The deterministic runway check before AI coding agents touch your repo. Claude, Cursor, Copilot, Codex, and local agents can help modify a repo.
Show HN: Dari-docs – Optimize your docs using parallel coding agents (github.com via hn) It’s well known at this point that documentation needs to be optimized for AI agents - we’re all pointing our Claude Code / Codex / Pi agents at documentation, and expecting the models to figure out how to implement a product. This, howeve…
Show HN: PrismoDev – local CLI for finding token waste in Claude Code/Codex (github.com via hn) I built PrismoDev after noticing my Claude Code and Codex sessions were getting expensive in ways that were hard to explain. After digging through local session logs, the recurring issue was not just model pricing.
RalphTerm: a ralph-style loop for Claude Code with fresh-session Codex cross-review (www.reddit.com) Post I wanted to share RalphTerm, an open-source CLI for running a ralph-style coding loop with Claude Code. The idea is not to replace Claude Code.
Giving operational continuity to coding agents, not everything is about memory (aictx.org via hn) Operational continuity for AI coding agents. AICTX helps Codex, Claude, GitHub Copilot and other coding agents continue work across sessions by preserving the last useful execution state: active work, next actions, decisions, failures, val…
Show HN: Agetor - An open-source Harness Orchestrator (github.com via hn) Hi HN, I built Agetor, a harness orchestrator for coding. The first release, 0.0.1, supports Claude Code.
I let Codex and Claude Opus work on the same Java AI agent monolith (www.reddit.com) I ran a small experiment on my Java pet project and the result was less clean than I expected. Small disclaimer: I did the final comparison review on April 19, 2026.
Problem with the Codex app (www.reddit.com) Has anyone else experienced this problem? The app opens and works, then closes automatically.
CAFleet – open-source Agent Teams reinvented, both for Claude Code and Codex (github.com via hn) CAFleet https://github.com/user-attachments/assets/a66620cb-4a81-4525-95f2-1f1f22765288 Agent Teams reinvented for collaborative coding supporting Claude Code and Codex, with full code transparency. Install CAFleet works with two coding ag…
What are some everyday, average person uses for Codex? (www.reddit.com) For example, I don’t really have a use for vibecoding. So far the coolest thing I’ve done is -Use the chrome connector to have Chat write me individual cover letters for each job application tab I have open -It uses my starred resume in Go…
Is there a good reason to pay for both Claude code AND Cursor? (www.reddit.com) Most devs are paying for either Claude code or codex but I’m also seeing some pay for both Claude code AND Cursor. Is there a use case or a problem that a combination of the two is able to tackle better than Claude code or Codex alone?
TokenBBQ – track AI coding token usage across Claude, Codex, Gemini (github.com via hn) TokenBBQ 🌐 offbyone.cloud — Homepage See what your AI coding tools actually cost you. TokenBBQ reads local usage data from Claude Code, Codex, Gemini, OpenCode, Amp, and Pi-Agent and shows it all in one dashboard.
Use Your ChatGPT Subscription in Zed (zed.dev via hn) You can now sign in to Zed's agent with your ChatGPT account and use OpenAI's models directly in Zed, with the same usage you benefit from in Codex directly. Codex is included with every ChatGPT subscription, with usage determined by your…
RLM models and Qwen3.6 (www.reddit.com) RLM models and Qwen3.6 Does anyone here have an RLM setup and how could I set it up? I want to make my Hermes agent even more powerful and I don't like that I need to open a new context window every time after just a few prompts.
Show HN: One Markdown File to Set Up Claude, Codex, Cursor and Copilot (github.com via hn) ai-project-setup Drop one file. Tell your AI to read it.
Autonomous agents are overrated until the business is readable (www.reddit.com) I have been building around agents for client work for a while now, and my take is probably less exciting than the demo videos. I don't really want an agent waking up, looking around, and deciding what to do.
Show HN: Open-source Codex Pet Home software (news.ycombinator.com) CodexPet Nest is an open-source desktop companion app with floating pet nests, usage status, focus widgets, quick actions, and safe community themes. macOS is available first; more platforms are planned.
I spent much of this year in the hospital with my mom. I built this so I could keep iterating on my more automated workflows while my dev machine was at home. (www.reddit.com) Wanted to share my mobile claude/codex session tool: Chroxy. TL;DR Chroxy is a (yet another!) self-hosted remote client for Claude Code.
Multi-project environment would be nice in Codex (www.reddit.com) Right now you select a folder and work within that project. It would be a cool feature to additionally select projects as read-only just so the context is present.
solo human browser use is moving to "together with an LLM browsers" (www.reddit.com) I keep thinking about how I use browsers. Or rather how I have used browsers since 1996 when I first heard about this Netscape thing.
MFKVault – npx mfkvault install [skill] for Claude, Cursor and Codex (mfkvault.com via hn) Search installable AI skills with source, license, safety status, install counts, and recent compatibility checks.
Cross devices agent memory and context management? (www.reddit.com) Hey, developers. Imagine you have 2 macs, one at your job, one at your home.
Agents are meant to be shared, but existing tooling is not fit for purpose (www.reddit.com) A while back I was doing technical support at my company and a ticket came in about some feature not working. Instead of digging through logs myself, I let Claude Code do it.
Codex Mac App vs CLI for production codebases? (www.reddit.com) Hey everyone, We are deciding on how to roll out Codex across our team for a large production codebase. For those using it daily: Are you finding the Codex Mac App or the Codex CLI better for handling massive, multi-file codebases?
How can I handoff from one agent to another? (www.reddit.com) I often end up hitting my limit in say claude code. Id love to just continue the conversation in cursor/ codex.
Symphony: Every open task gets a running Codex agent (twitter.com via hn) I pushed 50 tickets to Linear before bed — a tech debt rewrite of an Electron app. Woke up to 30 merged PRs.
Vulkan or CPU llama cpp backend for local llm for coding/code assist (www.reddit.com) Hi all I recently started a new job and we're doing python development for a ci cd metadata consolidation library for analytics and we cannot use no stuff like claude code or codex or gh copilot or any model APIs (free or paid). I got a la…
Best free AI Agent provider? (www.reddit.com) Hi everyone, I’m looking for recommendations for the best free AI agent providers and which models work best for coding and general development workflows. So far, I’ve mainly been using Cursor, and honestly it has given me the best overall…
Most agent-memory tools are markdown you keep grooming. I wanted something that travels between models and machines, so I built a protocol. (www.reddit.com) If you use Claude across more than one editor or machine, you've probably hit this: your context never comes with you. The CLAUDE.md doesn't follow me to Cursor, the Cursor rules don't follow me to Codex, none of it follows me from my Mac…
Those of you running multiple coding agents in parallel, how are you actually keeping track of them? (www.reddit.com) I got into the habit of running 6-9 Claude Code and Codex sessions at once across different repos and honestly the "management" side of it was a mess. What the initial setup / hacks looked like: - Manually checking `ps aux | grep claude` t…
The AI market moves so fast that your business idea can expire before launch (www.reddit.com) 1.5 years ago, n8n was everywhere. People were building workflows for everything.
Codex Pets for People in a Hurry (www.augmentedswe.com via hn) How to use Codex pets (and make your own!) Use /hatch to get a cute companion for your projects OpenAI has really been cooking lately. They’ve gone CRAZY on capabilities and features for Codex, their fast-growing Claude Code competitor.
What do you NOT like about Cursor / VSCode / Claude Code desktop / Codex / etc.? (news.ycombinator.com) I am building a highly integrated, cross-provider agentic workstation (its neither an IDE nor an ADE - does a bit of both, with additional unique features on top), and I would love for you guys to rant about what you hate about the tools y…
Best autonomous ai agent for github? (www.reddit.com) Hi, this research is driving me crazy :/ I'm looking for an autonoums ai agent with generous limits to use as teammate on github. i would like to tag the agent in issue to develop the bug fix or in PR to review code.
Codex downloaded by Xcode 26.4.1 reported as Malware (old.reddit.com via hn) could not extract summary
I built agentwerk, a tiny Rust crate for scaling agent collaboration focusing on getting work done (www.reddit.com) For a new Rust project, I was searching for a simple agentic loop implementation. My goal was to analyze thousands of software artifacts at scale.
Where I'm at with AI Assisted Building + Current and Future Workflow Overview (www.reddit.com) I've been in an AI dive bomb for probably a couple of years now. The early days...
Claude still feels much better than ChatGPT/Codex at UX design (www.reddit.com) https://preview.redd.it/km7o9670lc0h1.png?width=1542&format=png&auto=webp&s=3fea5e97f3e518222eefd7cfd0cc871fcd58a933 Has anyone else found Claude stronger than ChatGPT/Codex for UX critique? In a recent test, I asked both to review a wikil…
What stack should I be using? (www.reddit.com) Hi all, I'm setting out to build a mobile app (with the intention of getting it published into the appstore and playstore at some point), and i'll be using Claude to help me build. As background, I'm non-technical, working in an finance/in…
On "harness engineering": Are people actually building things or just giving impressive labels to "tweaking?" (www.reddit.com) I see a lot of posts and videos talking about harness engineering, or it could be context engineering, RAG, etc. The thing is, most of them talk about the concepts.
Stuck in a loop of audits and refactors (www.reddit.com) I keep reading these posts here about spagetthi code and architectural debt etc and it brought me in a spiral of asking Claude/Codex to audit my codebase (20k lines of code, self hosted procurement tracker for my work) and while it did fin…
RTX Pro 4500 Blackwell - Qwen 3.6 27B? (www.reddit.com) have have a server running a 4500 blackwell on cuda 13.1 and nvidia/595.58.03 with 48GB mem assigned to it. I have build: dcad77cc3 (8933) with Qwen3.6-27B UD-Q5_K_XL loaded and connected it to Roo code.
Any pros in using cursor over claude app? (www.reddit.com) i loved cursor. it was vs code on steroids.
Testing AI modeling skills (www.reddit.com) So I am currently testing how useful AI models can be in day to day workflows, and went why not compare 3 models and see how good they are at replicating my work. The goal was simple they were asked to replicate one of the kitchen cabinets…
Show HN: Terminal Arcade – Side quests when agents are cooking (twitter.com via hn) I was having way too much free time even when i was working on 6-7 different projects at the same time with claude code and codex. Staring at the terminal when agents are cooking code for you was just too boring in the last 6-7 months so h…
Fob – a local continuity layer for Claude, Codex, ChatGPT and Gemini (fob.sh via hn) Import or ask Paste a long AI answer you already paid tokens for, or ask Claude, Codex, both, or a structured debate from one local dashboard. Local workspace for AI continuity Keep project context, decisions, handoffs, and AI conversation…
Ask HN: Does Codex hits limits more easily now? (news.ycombinator.com) The past month or so, I have the impression that Codex started hitting the 5h limits much more often. I always used the latest model in xhigh, but would almost never hit any limit (I'm not a very heavy user).
Aztec Codex (en.wikipedia.org via hn) Aztec codex Aztec codices (Nahuatl languages: Mēxihcatl āmoxtli, pronounced [meːˈʃiʔkatɬ aːˈmoʃtɬi]; sg.: codex) are Mesoamerican manuscripts made by the pre-Columbian Aztec, and their Nahuatl-speaking descendants during the colonial perio…
A 6.8M-token Codex run survived a five-hour pause (tectontide.com via hn) /goal: The Six-Hour Codex Run That Survived a Five-Hour Pause TL;DR /goal shipped in Codex CLI v0.128.0 on April 30, 2026 as a named headline feature.- It introduces persisted goals: a goal state that survives terminal restarts, laptop sle…
Show HN: wfb-link, a userspace WiFiBroadcast radio stack for macOS (github.com via hn) Hi HN, I’ve been working on a Rust userspace radio stack for running WFB-style links from macOS using RTL8812AU USB adapters. Full disclosure: I'm a software engineer, but not really a hardware or embedded systems engineer, so Codex GPT 5.…
Show HN: Infinite you – multi agent workflow system (github.com via hn) Seems lots of folks are building workflow orchestrators these days. This is mine~.
An agent skill to enforce AI to write modern CSS (www.reddit.com) An agent skill that enforces modern CSS practices based on your project's browser targets. Covers 57+ CSS features across color, layout, selectors, animation, typography, positioning, and component patterns.
Ask HN: How are PMs keeping up with AI-accelerated engineering output? (news.ycombinator.com) With tools like Claude Code, Cursor, or Codex, engineers are shipping code faster than before. The bottleneck is no longer "how fast can we build it" but "how fast can we spec it well enough to build the right thing." My Product Team is st…
AI and Claude: The internal rebellion that changed Amazon's rules (thenewstack.io via hn) AI and Claude: The internal rebellion that changed Amazon’s rules Amazon has given its estimated tens of thousands of developers immediate access to Anthropic’s Claude Code, and they’ll soon have access to OpenAI’s Codex, opening up agenti…
Any plans to compact context in long conversations? (www.reddit.com) Just wondering if OpenAI has plans to compact context windows after certain point(preferably on demand by user) to allow longer conversations? Codex already makes it.
notebooklm-py makes Claude interact with NotebookLM (github.com via hn) notebooklm-py A Comprehensive NotebookLM Skill & Unofficial Python API. Full programmatic access to NotebookLM's features—including capabilities the web UI doesn't expose—via Python, CLI, and AI agents like Claude Code, Codex, and OpenClaw.
Sr Software Engineer - Haven't written a line of code in months (www.reddit.com) AI has reached the point that I no longer write code. I used to work in shops where I was deep in the debugger without internet access; now I just drive intent and long term engineering decisions with Claude/Codex/Perplexity.
Show HN: Kanban-CLI – a web UI for local Markdown todo lists (github.com via hn) As we all are, I've been experimenting with ways to reduce external saas spend, and continually bring traditionally external pieces of context (prs, docs, trello boards) into the one mono repo. I have toyed with a markdown todo list and se…
OpenAI Codex Surpasses Claude Code in Downloads Following April 30 Inflection (blog.tickertrends.io via hn) OpenAI Codex Surpasses Claude Code in Downloads Following April 30 Inflection Codex downloads inflect sharply after April 30 release, driving a rapid divergence in developer adoption vs. Claude Code TickerTrends data shows a sharp shift in…
Codex for the win! (www.reddit.com) I had 370 PDFs, each about 40 pages long. Community newsletter.
Introducing Skills Over MCP – The better way to share and distribute skills (skillsovermcp.com via hn) Host any public GitHub repo of SKILL.md files as a live MCP server. Paste a repo URL — get a stable URL for Claude Code, Cursor, and Codex.
Modyak – run Claude Code and Codex with any model from your Mac menu bar (modyak.com via hn) Modyak is a tiny macOS menu-bar app that lets Claude Code and Claude Desktop run on any model. Pick a preset.
Show HN: Kirikiri – A mobile IDE for Claude Code (iOS, open source) (news.ycombinator.com) Claude Code runs in a terminal. Phones don't have good terminals.
Sitting on 10k in unused openai api credits that will expire, what would you build? (www.reddit.com) I was a cofounder/cto at a startup that recently shut down, and we ended up with a decent chunk of unused openai api credits I already have a chatgpt pro subscription so this/codex doesn’t really help me there, and for most of my current s…
Show HN: Enoch – Control Plane for Autonomous AI Research (github.com via hn) I built Enoch after working with OpenClaw and trying to get an agentic coding system setup with Codex. In the past, I was trying to manually generate, code, and test this all manually.
Mac browser for a human that also gives coding agents local APIs (github.com via hn) wkdomains wkdomains is a macOS browser for developers working with coding agents like Codex, Claude Code, Cursor, and similar tools. It lets the human browse normally while an agent gets structured local access to the same page: screenshot…
Show HN: Plannotator for Codex (twitter.com via hn) plannotator @plannotator Plannotator.ai @OpenAIDevs Plannotator now supports Codex app & cli. While the "superapp" will eventually get there with similar features, this adds a lot of utility for users who spend a lot more time planning.
Which Agentic Coder is the most with it now? (www.reddit.com) Considering the price to performance which is the best deal or setup right now? Similar to codex where it can edit project files inside a folder etc.
Codex on Mac is hardly working for me. Anyone experiencing the same? (www.reddit.com) Its has been almost 2 days and hardly manage to implement anything. Randoming disconnecting and any update is way too slow.
Show HN: Council – Run Claude, Codex and Gemini against the same prompt (council.armstr.ng via hn) I often copy and paste the same prompts into Claude, Codex & Gemini separately. It's helpful seeing where they all agreed and where they diverged.
You can use cmd+shift+n to launch multiple codex window each for one project (www.reddit.com) Don't ask me how I found out XD
Both Codex and Claude got worse this week. Across every plan I retested (desktopcommander.app via hn) Where should you get your AI from? Compare the real cost of local hardware, pay-per-token APIs, and ChatGPT/Claude/Gemini subscriptions.
Show HN: Stream iOS Simulators to a Browser Window (github.com via hn) Agent tools seemingly know how to work with browsers better than with iPhone simulators, so I built this tool to capture the simulator XPC stream and render it in a webpage. This means Claude Code/Codex desktop apps can use their existing…
Will Cursor increase the usage of the Ultra plan? (www.reddit.com) I'm a hard believer of Cursor being a better IDE than Claude Code or Codex for a long shot, specially when you know the strengths and weaknesses of each model and you use them to your advantage. Being said that, It's hard to recommend it a…
Six months running multi-agent in production — the coordination patterns (www.reddit.com) I've been running 8 AI agents in production for a few months. Each is a Docker container with its own role (CTO, dev, devops, PM, traders, auditor) and its own Telegram bot.
Actual line in the official system prompt for Codex for GPT-5.5 (bsky.app via hn) This is an actual line that was added to the official system prompt for Codex for GPT-5.5 by OpenAI. Usually the system prompt is as minimal as possible, so I assume it would otherwise mention goblins a lot.
Text-to-CAD: Generate 3D models with coding agents (open source) (github.com via hn) ⚙ Open Source Text to CAD Harness ⚙ An open source harness for generating 3D models with your favorite coding agent ✨ Features Generate - Create source-controlled CAD models with coding agents like Codex and Claude Code. Export - Produce S…
Tower - Simple TUI based + MCP Server Git Worktree Manager (www.reddit.com) I've gotten to a point where easily 90% plus of the code I write is done by AI. This is great for speed, but sometimes, if you have N number of PRs open or working branches across multiple repos, things can get tough to manage.
GPT-5.5 prompt for Codex tries to make it not talk about goblins (twitter.com via hn) could not extract summary
Show HN: Integrations gateway for agents with 2FA for destructive ops (OSS) (github.com via hn) Hey HN! I've been wanting to use something like OpenClaw for a while but couldn't get myself to give it access to anything important due to all the risks involved.
Show HN: Simple SDK for agent-to-agent communication (github.com via hn) We were spending a lot of time re-writing the same primitives in projects we were doing getting claude + codex + other harnesses communicating in real time. Many other projects forced us into using their framework or harness or into a spec…
Making AI coding sessions persistent across agents (github.com via hn) 🌐 English · 日本語 · 简体中文 · 繁體中文 drift_ai Vendor-neutral handoff for AI coding tasks — between Claude, GPT, Gemini, DeepSeek, local LLMs. Reads from Claude Code, Codex, Cursor, Aider.
Show HN: MindCheck – Analyze your AI coding logs for over-delegation (github.com via hn) Hi HN, I built MindCheck after running into a problem in my own AI-assisted workflow. A couple months into using Codex heavily, I realized I had delegated too much of a data pipeline without really tracking the details.
problem with kimi k2.6 in cursor (www.reddit.com) is anyone else using kimi k2.6 or any other thinking models through byok api? because mine is stopping in the middle of thinking after few seconds.
i genuinely love cursor but i wish these ai tools actually talked to each other (www.reddit.com) this is not a complaint post exactly. Cursor is the best editor experience i've had in years, i'm not going back to vscode, etc.
Apple integrates Claude and Codex into Xcode 26.3 for 'agentic coding' (venturebeat.com via hn) Apple integrates Anthropic’s Claude and OpenAI’s Codex into Xcode 26.3 in push for ‘agentic coding’ | VentureBeat Orchestration Infrastructure Data Security More Newsletters Apple integrates Anthropic’s Claude and OpenAI’s Codex into Xcode…
Copilot Student GPT-5.3-Codex removal from model picker (github.blog via hn) Copilot Student GPT-5.3-Codex removal from model picker - GitHub Changelog Skip to contentSkip to sidebar /Blog Changelog Docs Customer stories Try GitHub CopilotSee what's new Search Changelog Docs Customer stories See what's newTry GitHu…
An open-source spec for Codex orchestration: Symphony (openai.com via hn) An open-source spec for Codex orchestration: Symphony. | OpenAI Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Business Develo…
PI agent integrated with Cline-Kanban repo: All using PI and Qwen 3.6 35B MOE UD 4K_XL (www.reddit.com) Repo: statisticalplumber/kanban at pi-agent-integration Hi Guys, To test Qwen 3.6’s potential, I also wanted the Cline Kanban project to have an open-source agent to work with. The last time I tested Cline Kanban, it didn’t support agents…
Is there already an open-source app for centralized LLM chats? (www.reddit.com) Hello! I’m a software developer thinking about how to keep all my LLM conversations in one app instead of having them scattered across ChatGPT, Claude, Gemini, etc.
I made a battle royale arena where AI agents fight each other on a Swedish island. Mostly for fun. (www.reddit.com) Built this over the last week during nights because I thought watching AI agents fight each other would be fun, and I wanted an excuse to ship something with MCP. Posting because it turned out more entertaining than I expected - different…
CAD in Codex (twitter.com via hn) Jake (softservo) on X: "Vibe coding a robot with GPT 5.5! This is a URDF of a 7dof robot arm with functional kinematics, a custom gui, and STEP parts/assembly, 100% generated in Codex (minus the gripper).
Use LangChain with Codex (ChatGPT) Plus/Pro (github.com via hn) Page not found · GitHub · GitHub Skip to content Navigation Menu Toggle navigation Sign in Appearance settings Platform AI CODE CREATION GitHub Copilot Write better code with AI GitHub Spark Build and deploy intelligent apps GitHub Models…
Show HN: A bilingual guide to Thaayam, a Tamil board game (amal-david.github.io via hn) I saw The Mahjong Guide (https://themahjong.guide/) today and loved how much a good visual explanation can do for an old game. Around the same time, this HN thread about using coding assistants to revive projects you never were going to fi…
Codex CLI hooks are now stable — new PermissionRequest hook added (6 total) (www.reddit.com) SessionStart UserPromptSubmit PreToolUse PermissionRequest 🆕 PostToolUse Stop I maintain a drop-in hooks pack that plays a sound on every event. Cross-platform (Mac / Linux / Windows installers included).
Orchestrating agent workflows with Codex (www.reddit.com) Hi everyone, I’m in the process of switching from Claude Code to Codex, and I think GPT-5.5 is really impressive. But some features in Claude Code — like project-level agent definitions and orchestrating agent workflows — don’t seem to be…
Can I use 2 ChatGPT Business seats on the same device at the same time? (www.reddit.com) I'm the owner of a Business workspace shared with 3 friends — we split the cost because $100/month solo is steep. Now I'm wondering: can I invite a second account of my own to the workspace, so I can use 2 on the same device: web app and c…
Show HN: LLM-wiki – One command Karpathy's wiki with QMD search for Claude/Codex (github.com via hn) llm-wiki Bootstrap and query LLM-maintained project wikis before planning or implementation. Supports Claude Code + Codex (GPT-5.5).
Show HN: The Order of the Agents – Make Codex and Claude Create the Perfect PRD (github.com via hn) Should we plan with Codex, then code with Claude? Or should we plan with Claude, then code with Codex?
CC-OpenAI-Codex Plugin, but for all CLI agents (www.reddit.com) Hello! I made a plugin for myself, & I figured I'd share it, in case someone else finds it useful (also to solicit feedback on it).
Ask HN: How are you evaluating AI apps and CLI? (news.ycombinator.com) I'm sure many of you work for companies where various AI tools are being made available and IT departments asking for feedback on those tools. The IT departments are allocating in some cases unlimited budget in the hopes that something com…
Show HN: Roids – Open Source Steroids for your Agents (github.com via hn) Roids Compare UI directions side by side in the browser. Roids is an open-source skill + runtime for AI coding agents (Cursor, Claude Code, Codex, and similar tools).
ChatGPT for Cybersecurity (www.reddit.com) Hi guys, I’m a cybersecurity researcher, and after the recent terrible experiences with Opus 4.6/4.7, I decided to give OpenAI ChatGPT a try, conveniently coinciding with the release of 5.5. I’ve already completed verification and requeste…
Expensive but oh boy was it cool. (www.reddit.com) Admittingly I've never been a heavy Codex/Ai user, but since I'm a dev I took the opportunity of 5.5 lunch to try sub-agents trough the extension on VSC. I have the 20$ plan.
codex --model gpt-5.5 Not updated in the CLI yet (www.reddit.com) Use this command to access GPT 5.5 with your Codex
Show HN: Callmux – MCP multiplexer that cuts tool call context pollution by ~19x (github.com via hn) Every tool call an AI agent makes adds tokens to the conversation context. Not just the payload data, but the JSON wrappers, the role markers, and worst of all, the model's intermediate reasoning between calls ("Now I'll fetch the next one…
I Forked 4 CLI coding agents to Run the Same Model. I found a 2x gap (charlesazam.com via hn) Deep dive into the architecture of Codex, Gemini CLI, Mistral Vibe, and OpenCode. Same model, 2x performance gap — the scaffolding is what matters.
Show HN: Personal AI Metrics Dashboard (wakatime.com via hn) Hi HN, I built WakaTime 13 years ago before AI. Things have changed a lot since then, and the time you spend typing in your IDE isn't as valuable as it used to be...
Show HN: Momentum – showcase what you're shipping (app.heytangent.com via hn) Momentum is a way to showcase what you're shipping. Things are moving faster than ever with Claude Code and Codex, and it can be hard to keep up or even share your own progress effectively.
Show HN: API Ingest – Agentic Search (Inter) API Docs (github.com via hn) 1. CC / Codex dont handle API Docs well enough No matter what I do, I run into bad requests with claude, day in, day out.
Lazyagent - All-in-one observerbility terminal app for ai agents (www.reddit.com) Running multiple coding agents can make you lose track of what they are actually doing. Once subagents start spawning other subagents, basic questions get hard to answer: what is running right now, what tool did it just call, did the child…
Show HN: Sift – save AI tokens in Codex/Claude by summarizing command output (github.com via hn) I made a small skill/script for agentic coding workflows: https://github.com/panpeter/sift-skill The idea is simple: when a command like cargo test, pytest, npm test, or ./gradlew test prints a lot of output, that raw log often gets pulled…
How Claude Code and Codex approach sandboxing (instavm.io via hn) How OpenAI's Codex and Anthropic's Claude Code prevent an LLM from running arbitrary commands on your machine. Same OS primitives, different trust architectures.
what's defensible about cursor ? (www.reddit.com) like many / all here , i'm always thinking to myself i cant just fork my own vscode or make a stateful plugin out of the claude code leak . my question is : what exactly is and where exactly is the cursor ip ?
OpenAI's Codex grew 33% in the last 2 weeks (active users 3M –> 4M) (xcancel.com via hn) Sorry this pages exist in order to keep the service usable for everyone. If you can't pass the test, please whitelist your extensions on this website and update your browser.
Euphony: OSS tool for visualizing chat data and Codex session logs (openai.github.io via hn) Tell HN: Gemini CLI and codex are broken (news.ycombinator.com) Does cursor always use Codex 5.3 in auto mode? (www.reddit.com) Chatgpt down guys I'm cooked (www.reddit.com) Such a useful idea if it can be done, what do you guys think? (www.reddit.com) TLDR [Video isnt 100% looks wise but the idea was imagine with codex or chatgpt it can generate a interactive video where you can click and draw or use a drawing pad and select or highlight or draw on the interactive video so that it can g…
Show HN: ChatbotChambers – Watch LMs talk to each other (github.com via hn) I made a self healing PRD system for Claude code (www.reddit.com) Show HN: Agentic Dev – AI dev-tools news, curated daily by Claude (agenticdev.blog via hn) OpenAI released a major update to Codex, used by over 3 million developers weekly, adding background computer use, an in-app browser, image generation via gpt-image-1.5, more than 90 new plugins, GitHub PR review support, SSH connectivity,…
Current Limits (www.reddit.com) I had one year of Google ai pro as a student, and now that the plan is about to end, I’m trying to decide which AI I should subscribe to. I was a ChatGPT subscriber when o3 and o4 were a thing, and usage limits felt generous at the time.
Any good/up-to-date tutorials on how to use advanced CC features? (www.reddit.com) Hi! I am a developer for a decade now and built an app last year with Claude.
ChatMCP – Connect your AI browser chats to your coding agents (github.com via hn) ChatMCP Pull context from your AI browser tabs directly into Claude Code (and other AI coding tools) via MCP. How it works Claude Code / Copilot CLI / Codex CLI ← stdio (spawns server as subprocess) OR Cursor / Antigravity / Copilot (VS Co…
Tell HN: I found the perfect way to get maximum Claude Code quality (news.ycombinator.com) I always ask in the prompt: "don't use subagents". It's slower, but better quality.
Is there a way to have qwen-code CLI read images? (www.reddit.com) Basically I am asking the model to describe an image, but it says it can't process the images. The weird thing is that if I send the image encoded directly on the prompt, it works just fine, I am using llama-server with qwen3.5 (tried all…
I used Codex to build a Power BI agent workflow that goes past Microsoft's MCP scope. Does this shape make sense? (www.reddit.com) I built a Power BI workflow around Codex because I wanted something that could go beyond Microsoft's official powerbi-modeling-mcp. Their MCP handles semantic model operations well, but it stops short of local PBIR report authoring.
Show HN: Kilroy – Knowledge base for teams using Claude Code (github.com via hn) Hey HN — we’re a small team that uses Claude Code + Codex for basically everything in our company: coding, data analysis, marketing, ad campaigns, copywriting, design. There’s a truckload of tribal knowledge we’ve accumulated; major decisi…
The Security Decisions Claude Code and Codex Make (amplifying.ai via hn) Research Edwin Ong & Alex Vikati · apr-2026 The Security Decisions Claude Code and Codex Make Anthropic's Project Glasswing, built around Claude Mythos Preview, showed AI finding zero-days in decades-old code. The other side of that coin:…
Show HN: Evo – parallel autoresearch experiments for Claude Code and Codex (github.com via hn) Evo is a plugin for Claude Code and Codex that optimizes a codebase against a metric. Two commands: /evo:discover figures out what to measure in your repo, instruments the eval, runs baseline.
Ive just been defaulting to codex 5.2 but they retired it, so i switched to 5.3 and.. (www.reddit.com) Holy shit it feels really good. I don't know if its just the late stage of my project, but it cleared out a massive stack of warnings pretty much one shot.
Anybody else wake up today and do "/model" in their codex terminal hoping to see 5.5/6.0/Spud? (www.reddit.com) Hoping its well baked.
Applied for Codex Ambassador (www.reddit.com) I applied for Codex ambassador and I would love to be one. It’s been more than a month I did and I’ve been an active codex user and I have 1000+ tech Meetup group I run.
I built a tool that turns repeated file reads into 13-token references. My Codex and Claude Code sessions use 86% fewer tokens on file-heavy tasks. (www.reddit.com) I got tired of watching Claude Code re-read the same files over and over. A 2,000-token file read 5 times = 10,000 tokens gone.
Show HN: API Changelog Tracker (apipulse.app via hn) Hi HN, I built this because I rely on many external APIs, and keeping up with changes is harder than it should be. Many (most?) APIs don’t provide RSS feeds, sometimes they provide RSS that they block from fetching (yes, this happens!), or…
Contained Codex Networking (news.ycombinator.com) This is a bit odd, because it was going to start off as an Ask, and now its a hybrid Show/Ask. The ask being, how in the world do I make use of Codex's proxy networking?
How do I use gemma4 on 5090 gpu for coding? (www.reddit.com) I'm trying to replace openai codex which i used for development all the time, with gemma4 on 4090, small tasks it solves quite impressively, but i need to have some agent. So I tried to connect 31b to cline and to aider and it didn't reall…
If I already pay for ChatGPT Plus, what’s the smartest way to use it for recurring research and monitoring tasks? (www.reddit.com) I already pay for ChatGPT Plis, but I feel like I’m underusing the OpenAI stack beyond the normal chat interface. Right now, I mostly just use regular ChatGPT (chat interface).
Show HN: Buildermark – See how much code is by your agents (open source, local) (buildermark.dev via hn) I made Buildermark to see exactly how much of my code is generated by my coding agents vs what I was writing by hand. For this project, it ended up being 364 agent conversations writing 94% of the code.
is it normal for SaaS to fake their uptime? (www.reddit.com) https://status.openai.com/ shows everything is mostly fine, but I know for a fact many users are experiencing persistent service distruptions. for example: stream disconnected before completion: error sending request for url (https://chatg…
Show HN: Forexfin – Trading calculators, alerts and a journal in one place (forexfin.tech via hn) Some many years ago I started trading here and there and whenever I was doing it, I was bouncing between different platforms for position sizing, pip value, spread cost etc. Each of them was more or less expensive, I've also met some nice…
Show HN: Built a local guardrail layer for Claude Code (github.com via hn) I built Mati to give me more control over what Claude Code and Codex are allowed to do. Many important product constraints live in an engineer's or product person's head.
Show HN: Claude Code hits its 5-hour limit, Codex picks up in the same terminal (github.com via hn) Leg Type leg claude, leg codex, leg agy or leg grok instead of the bare command. You get the same interactive agent; Leg opens a board next to it, watches the usage limit, keeps a handoff bundle current, and when the limit hits it starts t…
Show HN: Procedurally Generated Mineral Gems (gems.0x0000007a.com via hn) I wanted to give users of a new app I'm working on a way to generate their own custom icon. It turned out pretty cool so I chucked it up on a Cloudflare page.
Show HN: Skillmem – local memory for coding agents that stores how, not what (github.com via hn) skillmem Self-improving skills for Claude Code and Codex — your agents learn, recall, reinforce, and forget. Strength has to be earned — saying a skill helped is not evidence, a passing test is: Generated from a real run: scripts/demo.sh -…
Let GPT-6 Astra code without using Codex Usage (ethanplus.ai via hn) How I Run a Coding Agent on GPT-6 Pro Without Touching My Codex Limits Codex and Work share one allowance. Regular Chat has its own.
Grep your Claude Code/Codex history and jump back into the exact session (news.ycombinator.com) Built a small devtool I wanted for myself: context-find Search all your Claude Code + Codex conversations across projects and machines, then resume the exact one you were looking for. https://github.com/ayush-gzip/context-find
Codex: Agent doomscrolling or providing traffic to paid partners? (twitter.com via hn) selim on X: "@EquityCats codex surfs the internet or actually provides traffic to paid partners. here are examples I recoded yesterday; the query was neither about Taylor Swift or Barrack Obama 😅" codex surfs the internet or actually provi…
BYOK vs. fixed subscription like GitHub Copilot, codex or Claude (news.ycombinator.com) has anyone does the comparison between paying a fix subscription(pro max) vs BYOK for a hobby project and exploration I'm curious how much does it cost in comparison. I haven't explore these agentic workflow as much, besides checking your…
Lodestar – Easier Communication with Agents (www.try-lodestar.com via hn) A Mac workspace for Claude Code and Codex. For everything between an idea and something you love.
Show HN: Jev routing coding tasks to Grok Build or Codex Astra (github.com via hn) Jev model router demo This is a throwaway local prototype for one question: can Jev reliably send complex or uncertain coding tasks to Codex Astra while routing contained mechanical work to Grok Build? The router uses Jev only for the deci…
Using Jev for Claude Code model routing (github.com via hn) jev-router Automatic per-turn model routing for Claude Code and OpenAI Codex. Jev sends simple work to the fast tier and difficult work to the strong tier, while preserving each CLI's native interface, tools, sessions, permissions, and aut…
Show HN: Multiplayer Mode for AI Agents (gotincan.com via hn) I built Tincan for my agents across different harnesses (Claude, Codex, Hermes) to coordinate work with each other in real time. Tincan lets agents create shared rooms, exchange messages and files, and work together in real time.
Switch to Codex seamlessly when Claude Code is used up (raw.githubusercontent.com via hn) could not extract summary
Show HN: A simple Mac extension for recent screenshots (github.com via hn) Had my very own "personal software" moment recently. Apple's screenshot experience is god awful, you take a screenshot and then race to the destination where you need to drag and drop it (messages, etc.) Asked codex to solve this and it di…
Setting Up Pi with DeepSeek v4.1 Flash on OpenRouter (www.vincentschmalbach.com via hn) Inference Is the Last LLM Moat OpenAI and Anthropic's moat for developers is subsidized inference, meaning reliable, fast, Western-hosted access to running models at an affordable monthly price.… When I run out of Codex and Claude Code usa…
/pyor:review – Claude skill for human code review (github.com via hn) pyor-cli Agent tooling for Pyor, the native home for GitHub code review. Works with Claude Code, Codex, Cursor, and any coding agent.
Show HN: Open-source tool to sell your Claude Code and Codex sessions (github.com via hn) Your agent history is worth money. Sell the Claude Code and Codex sessions you already ran.
Move AI chats across harnesses: Claude Code, Codex, OpenCode, Cursor, and more (github.com via hn) Continue your conversation in another coding agent. English | 日本語 | 简体中文 | 繁體中文 | 한국어 | Deutsch | Español | Français | Italiano | Português (Brasil) | Русский | मराठी | தமிழ் txcript is a library for converting agent sessions.
Adios MCP – let coding agents build, preview, debug and deploy apps (github.com via hn) Adios MCP Use Adios workspaces, previews, builds, logs, and deployments from Codex, Claude Code, Gemini CLI, Kimi Code, GitHub Copilot, ChatGPT, and other MCP clients. This repository packages the hosted Adios MCP service for public instal…
Managed Agent Architectures: Why Frontier Labs Are Rebuilding the Agent Loop (twitter.com via hn) Recently, OpenAI released the Agents API, giving developers API access to the managed Codex harness. One of the strongest agent harnesses out there is now available as infrastructure developers can build applications around.
Built-In Tools vs. Custom Tools in LLM Agents (www.vincentschmalbach.com via hn) My First Week With GPT-6 Astra GPT-6 Astra was my main Codex driver for the last week, and I am back on GPT-5.5. That sounds harsher than my… When building AI agents, a tool can be anything that the model can ask to use such as a search en…
Show HN: Bash sandbox script for Claude and codex CLI (github.com via hn) I've been working on and off on this project to have a reasonably easy way to run a docker sandbox on my machine. Recently with claude "auto mode" not working as "automatically" as before I resurrected this.
Open-source project: Use ChatGPT Web models directly in Codex (github.com via hn) ChatGPT Web for Codex Use ChatGPT Web (including Pro) as native Codex models. Change the model tier, save your workflow.
Show HN: Replay Doctor – Local CLI to find silent prompt cache leaks (replay.doctor via hn) I built Replay Doctor because AI agent bills spike without a single warning or error in the logs. When developers use Claude Code, Cursor, or Codex, the workflow relies on prompt caching.
Show HN: Switcheroo – manage logins for 10 CLIs like Vercel, Codex etc. (news.ycombinator.com) could not extract summary
Show HN: Swobu – Local LLM Switchboard You Can Share over HTTPS (github.com via hn) Swobu is a local TUI switchboard for Codex, Claude Code, PI, you name it. With stateful, provider-independent session lineage and model-name-based routing across way too many heterogeneous LLM providers.
Computer Use Agents Are Excellent QA Engineers (tristanrhodes.com via hn) Computer Use Agents are Excellent QA Engineers The recent proliferation of software such as Grok Bot, Instinct, Codex, etc. indicates that computer use is one of the most rapidly expanding frontiers in LLM agent capabilities.
ASU: One CLI that shows how much of your agent subscription is left (chaosguru.substack.com via hn) I pay for Claude Code, Codex, and GitHub Copilot. Each one meters me in its own way, and none of them meters me in a unit I can reason about.
Orca – The Agent Development Environment (www.onorca.dev via hn) Orca is the Agent Development Environment (ADE) for shipping with coding agents. Run Claude Code, Codex, Gemini, Cursor CLI, and every other CLI agent in parallel across isolated worktrees.
Show HN: HolaOS––An Opensourced workspace that alternative to Claude (github.com via hn) Open-source agentic workspace enterprises can make their own. Connect the systems you already run — 100+ integrations, MCP, chat tools, apps, browser, local files — with shared memory.
Codex Pricing Issues (timleland.com via hn) Codex Pricing Comparisons I don’t mind paying $200 a month for Codex if it helps me build faster. But when I need more usage, buying more should be simple.
Collectiv Beta Built by Sol 5.6 Using Codex (news.ycombinator.com) My beta is almost ready and i would like to get more product testers. If youre a fan of the hobby, go to collectivshares.com and join my waitlist to be amongst the first to test my platform
Codex grabbed more than 73% of the context (github.com via hn) ContextOS Spatial Architecture Canvas & Context Optimization Operating System for AI Coding Agents AST-sliced task contexts (99.4% token reduction) · Cross-session architecture memory · Deterministic verification gates 中文文档 · Download Nati…
A Tale of Two Commit Messages (Claude vs. Codex) (devcodehack.com via hn) Ok so I recently switched from Claude Code to OpenAI Codex (Terra/Sol/Astra). First of all, I was surprised at how fast Astra Medium blew through my usage doing some pretty rote stuff.
Show HN: Toolcraft – open-source AI harness and starter for building design apps (github.com via hn) MIT-licensed, 100% free starter kit, UI library, and AI harness for building your own design apps. Works with Codex, Claude, Cursor, or any other agents.
Show HN: See how much of your Claude, Codex, or Copilot subscription is left (github.com via hn) ASU: agent subscription usage See how much of your coding-agent subscription you used, straight from the provider's account API. ASU is a local TypeScript CLI for humans and agents.
Show HN: SOS – Project state between coding agent sessions (github.com via hn) I started working on SOS because noticed that during long-term projects, Codex often loses track of the current status and gets confused. I frequently had to explain what we had already done, which decisions were still valid, and what need…
Why Codex burns a weekly limit in a day while the agent waits for the tide (relux.works via hn) Spin-waiting in Codex CLI: goal mode restarts the model 0.03 seconds after every turn and demands proof of waiting by live polling, the pause tool is issued to one model in the catalog, and every poll rereads a context of hundreds of thous…
Codex kicked me out of Astra (news.ycombinator.com) Started Codex, Pro subscription. "GPT-Astra-6 is an invalid identifier" , top model I can pick is GPT-5/6-Sol.
Codex Rate Limit Reset: Weekly Caps, 5-Hour Window, Credits (provenbrief.com via hn) Codex Rate Limit Reset: When Limits Reset, Weekly Caps, and Credits OpenAI Codex usage limits reset automatically on a weekly schedule tied to your ChatGPT plan, Support cannot reset them manually, and paid workspaces can keep coding from…
Show HN: WordPress security skills for coding agents (github.com via hn) WordPress Security Skills Modular Agent Skills that teach AI coding agents — Claude Code, Cursor, Codex, OpenCode, Gemini CLI — to write secure-by-default WordPress code and to audit and harden existing plugins and themes. Confused?
Show HN: MaruCheck – Independent QA for AI-generated code (github.com via hn) Hi HN, I've been building MaruCheck, an independent, open-source QA verification tool aimed to solve the issues that arise when coding agents like codex, claude, cursor make semantic issues that has heavy ramifications. Take the example wh…
Ask HN: What is the proper way of preventing AI labs to use your private data (news.ycombinator.com) Recent discussions show that most people (including myself) are uneducated about the privacy policies of these AI companies. People talk about the "opt-out" button in the settings and assume this will be enough.
Benchmarking Claude Code, Codex and Pi on SWE-Bench Pro: Same Accuracy, 2x Cost (aistack.imec-int.com via hn) Your harness and model combination can double your token bill. aistack//~11 min read The takeawayYour choice of agent harness has less impact than we expected on successful resolution of coding tasks (when you stick to the more popular cho…
Ask HN: Codex conversations failing to converge across devices? (news.ycombinator.com) Is anybody else seeing extremely annoying bug in last week or two where codex/chatgpt conversations fail to sync properly across devices? In my case ios and macos.
Contextify for Windows – Live Index of Your Claude Code/Codex History (contextify.sh via hn) Indexes in the background The tray app watches supported Claude Code and Codex transcript locations and keeps the local index current. The Contextify system tray app indexes local Claude Code and Codex sessions in the background.
Solo.io Pushes Agentic AI Governance to the Desktop: Open-Source Agentdesktop (techstrong.ai via hn) TL;DR — Key Takeaways - Solo.io’s open-source agentdesktop project is designed to help enterprises govern AI-agent tools such as Claude Code and Codex running on employee workstations. - The platform adds AI-specific inventory, policy mana…
Show HN: Blackholes – A macOS workspace for coding agents (blackholes.dev via hn) Hi HN, I'm building Blackholes, an open-source macOS app for organizing coding agents, Git repositories and terminals. A project can contain multiple repositories, and tasks can have their own branches and worktrees.
Show HN: A cozy home for your Goodreads books (reading-room.transitivebullsh.it via hn) this demo uses my own book data from goodreads, but if there's enough interest, i'll make it easy to create your own for free. created with codex + gpt-6 astra with data pulled from goodreads open source here: https://github.com/transitive…
Ask HN: Anyone using dictation with coding agents? (news.ycombinator.com) Curious if folks are using dictation with coding agents for speach to text? Anyone completely hands-free with spoken responses from the agent as well?
Catenary – A spatial canvas IDE for AI coding agents (thecatenary.app via hn) The spatial IDE for AI coding agents Orchestrate Claude Code, Codex, Cursor, Grok, OpenCode, and Pi with visual wires, and a side-by-side editor to track everything. HEADS UPYour system will block the file on first run: the installer is no…
Deterministic Rule based auto-approver for Claude/Codex (github.com via hn) anumati Stop approving the same Bash commands over and over. anumati is a deterministic auto-approver for AI coding agents.
Show HN: Pomeroy v1, give any AI assistant secure access to native macOS apps (pomeroy.app via hn) When I launched Pomeroy last week, I knew I was solving a problem, but I didn't understand the scale! Pretty overwhelmed by the response, and have worked over the weekend to make it EVEN better...
Show HN: DashClaw – policy and approval layer for unattended coding agents (github.com via hn) DashClaw With a supported enforcement integration, when your AI agent (OpenClaw, Hermes, Claude Code, Codex) tries something destructive or expensive, DashClaw catches it before it runs and asks you first, even when you are not at the keyb…
Show HN: MaskShift – a maximalist coding agent with zero NPM dependencies (github.com via hn) MaskShift is a local-first coding agent harness. Features: - Works with models that have no tool-calling API.
Nimvarya – Let AI coding agents control Chrome without stealing focus (github.com via hn) nimvarya A standalone, Chrome-control bridge for your favourite terminal AI. Drive a real Chrome tab — navigate, read, find, click, type, screenshot, query console and network — from Claude Code, Codex CLI, Gemini CLI, or Cursor, over one…
Clove: Agent Skills for Mac (www.kshv.me via hn) Clove — Agent Skills for Mac You installed dozens of skills across Cursor, Claude, Codex, plugins, and project folders — then forgot what you had and how to reference them. Then you're mid-prompt, sure you installed something for animation…
Show HN: Ditch; Build multiple products at once (theditch.dev via hn) Hey HN, I built Ditch because I'm running Codex across several projects and was losing track of what each agent was doing, which one needed me, and where I had left off. Plus, I have other work, so I'd like to leave my agents be and come b…
AI tool found 6 Curl vulnerabilities Mythos and Codex missed (www.zdnet.com via hn) ZDNET’s key takeaways - Aisle finds bugs that AI coding programs can’t. - The best Linux maintainers are impressed.
Ken-rank MCP: Smart index tool for research and development agents (pypi.org via hn) Per-project context-rank index for Claude Code, served from a local daemon. ken ken is a local context-rank index for coding agents such as Claude Code and Codex CLI.
98% of What Claude Code Does, I Never See. The 0.27% It Stops for Is the Point (grith.ai via hn) Claude Code and Codex ask permission, then everyone switches on auto-approve and hopes. Across 120 real Claude Code sessions under grith, 98.4% of what the agent did ran silently - and the 0.27% it stopped to ask about was reads of AWS cre…
Show HN: Generate a Claude Code or Codex agent fleet from one HTML file (github.com via hn) fleet-generator A browser wizard that generates a working agent fleet for Claude Code or Codex. One HTML file, no install, no build step.
Show HN: I processed 100k+ LinkedIn posts into open-source Claude Code skills (github.com via hn) LinkedIn Marketing Skills for Claude Code and Codex Claude skills for LinkedIn. 11 Claude Code and Codex skills that write LinkedIn posts, comments, and replies in your voice.
Elevated errors across ChatGPT and Codex (status.openai.com via hn) Availability metrics are reported at an aggregate level across all tiers, models and error types. Individual customer availability may vary depending on their subscription tier as well as the specific model and API features in use.
Show HN: AI Usage Ball – Claude/Codex/Antigravity Quota as Mac Liquid Orbs (github.com via hn) AI Usage Ball A source-available macOS desktop app that shows how much of your AI coding-tool quota — Claude, Codex / ChatGPT, and Antigravity — you have left, as animated liquid gauges. Session limits, weekly limits, and reset countdowns,…
Agent-manager - The fastest workflow for every coding agent (claude codex etc..) (agent-manager.dev via hn) tab A different tool per spawn Cycle claude, opencode, codex, grok, gemini, pi, hermes or anything you configured without leaving the bar. The footer shows which one the next enter will start.
Show HN: Read-only script to check if ChatGPT desktop is burning Codex quota (github.com via hn) chronicle-quota-check Is something in the ChatGPT desktop app quietly spending your quota while you're not using it? Find out in about three seconds.
Show HN: Profound Academy – an agentic course builder for hands-on courses (profound.academy via hn) I built an AI course builder that helps you create hands-on course materials, interactive tutorials, exercises, and automatically checks the work submitted by students. It should feel like Claude Code or Codex for building courses.
Run any LLM in t3's Codex and Claude tabs through a local gateway (github.com via hn) proxy-llms Run any model in t3's Codex and Claude tabs. A GPT model in the Claude tab, GLM or Kimi K3 in the Codex tab, plugged and unplugged on demand.
ChatGPT Subsidy Chart reveals $3k monthly spend nets $62k API-equivalent value (aicharts.io via hn) Calculation Daily history of the measured API-retail-equivalent value of seven complete UTC days from all available local Codex logs on one machine. Usage spans switched accounts, and historical account count is unavailable.
Show HN: Pairmark, race Claude Code vs. Codex on your repo, blind cross-judged (github.com via hn) pairmark Run Claude Code and Codex on the same task in your own repo, side by side, and get a verdict backed by evidence. npx pairmark "add rate limiting to POST /api/login" One command.
Learnings from OpenAI's open source Codex Repo (johnjwang.com via hn) I’ve been fascinated recently at what the best practices in the new age of engineering look like. But it’s hard to find real data on best practices.
Show HN: An Autonomous Agent for Slay the Spire (github.com via hn) While coding agents have reached a relatively mature stage, domain-specific agents are lagging far behind; for instance, using Codex to play games yields poor results. The Spire agent addresses the challenge of maintaining consistency acro…
codex-remote-control – Use Claude's Remote Control with Codex Sessions (github.com via hn) codex-remote-control Remote-control your Codex CLI sessions from your phone. Codex has no remote control.
Ask HN: Anyone struggling to do data work with Claude Code or codex? (news.ycombinator.com) so i've been trying to interview data analysts and data scientists on their analytics experience with coding agents. anyone up for a quick chat ?
Show HN: Tenux – Access your computer from your phone (terminal, files, browser) (tenux.dev via hn) Hi HN, If you've been using claude code or codex, you may have been like me and used some sort of SSH/VPN combo to get remote access to your computer from your phone. It (kinda) works, but it is clunky.
Codex users debate the new Pet feature (community.openai.com via hn) What the heck is “Pet,” and why is OpenAI turning Codex into Windows by introducing junk features likey clippy with zero practical value? Why are highly paid engineers being asked to spend time on features like this when Codex still has so…
Is MCP Good Yet? (ismcpgoodyet.com via hn) Spoiler: no. MCP features across Codex, Cursor, Claude Code, Grok, and OpenCode.
A video editor built for your codex/Claude Code (www.usekinara.com via hn) Run a team of coding agents to edit, source, make motion graphics, and add sound effects.
Show HN: Memnest, local-first memory shared by pi, Claude Code and Codex (github.com via hn) memnest 한국어 README Your AI coding agent forgets everything when the session ends. Memnest keeps that memory on your machine and hands it back to the next session, through one small tool contract that pi, Claude Code, Codex, and other MCP c…
Eggshell – Local work memory across independent Codex chats (github.com via hn) Stop paying twice for work Codex already did. Carry completed work across independent Codex chats—locally, without special prompts.
Cross-model peer review for coding agents (github.com via hn) Model Peer Cross-model peer review for coding agents. Model Peer lets Claude Code, OpenAI Codex CLI, and Google Gemini CLI consult one another as independent, read-only engineering peers.
A local bridge for bidirectional collaboration between Claude Code and Codex (github.com via hn) AgentBridge 中文文档 🌐 Website: raysonmeng.github.io/agent-bridge, with an animated replay of a real session. Discussed on LINUX DO — the developer community.
Ask HN: What BOYK AI client are you using? (news.ycombinator.com) I built an AI desktop client specifically designed for non-programmers, which has the capabilities of Codex/Claude Desktop, but it is localized, protects all your privacy, and supports any model call. Would you like to try using it?
Show HN: Agentify Chat – E2E-Encrypted Remote Chat for Codex, Claude, Grok CLIs (news.ycombinator.com) Still under heavy development and rough around the edges but instead of sitting on this longer to perfect it I'll risk flak and share it. Essentially chat.agentify.sh is a remote control for codex/grok/claude cli and dream goal is to becom…
OpenAI Codex pricing: the $270 PR a $200/month sub covers daily (quesma.com via hn) A buyer’s guide to OpenAI Codex for individuals, teams, and enterprises. Includes comparisons with Claude Code.
Show HN: Rafter – an MCP server that shares one team's memory, skills and agents (heyrafter.xyz via hn) Supercharge your team with Rafter at https://heyrafter.xyz! :rocket: Your team members is spending too much time and tokens in their AI tools just to solve a trivial problem.
Show HN: I built an agent-first productivity bridge for all your agents (www.taskshell.app via hn) I initially built this for myself but some folks asked me to open access to it. Let me explain what I built and why.
Ask HN: What do you do when your AI coding agent is doing its job? (news.ycombinator.com) I am into my early 50s and recently diagnosed with adult ADHD and maladaptive daydreaming. I realized that I have struggled with it since I was 16.
The next CMS might be a coding agent and a Git repository (64labs.ca via hn) Static, Codex-managed publishing shifts content management from dashboards to files, rules, and deployment—and puts the traditional CMS on notice.
Automating repetitive work at OpenAI with Codex (developers.openai.com via hn) I’ve spent most of my career as a software engineer either turning a crank—deploying and operating software—or building software to turn that crank for me. My first job at OpenAI was on the cloud infrastructure team, bringing up new Kubern…
Show HN: AI scientist builds an open-source Codex Micro from scratch for $40 (iluvatarlabs.com via hn) A few weeks ago, we open-sourced our design for a cost-effective but full featured alternative to OpenAI x Work Louder's sold out Codex Micro macropad. And now, we're pleased to share that we've actually built them and they work!
Codex and ChatGPT browser now supporting WebMCP (twitter.com via hn) We’re adding support for WebMCP in the ChatGPT desktop app’s built-in browser and ChatGPT Sites. When you visit a compatible website, ChatGPT or Codex can automatically use it to complete your task.
Fetch – control your Claude Code agents from the Mac notch (tryfetch.dev via hn) For Claude Code and Codex on macOS Claude Code and Codex side by side, each with its own light. Kick off work, answer questions, approve plans, talk instead of type — no terminal-hopping, no tab roulette.
Show HN: Masker – PDF Redaction (github.com via hn) I built Masker after needing to share financial PDFs with a tax strategist without also sharing names, addresses, tax IDs, phone numbers, and account numbers. It is a native macOS app.
Show HN: MulmoTerminal – Run many Claude Code sessions, see which needs you (github.com via hn) MulmoTerminal Run multiple Claude Code and Codex sessions in parallel — and see which one needs you. A browser terminal for parallel AI coding agents: several Claude Code and Codex sessions side by side, each in its own cell, with the one…
Show HN: Save Claude, codex, Grok and OpenCode sessions to an infinite canvas (agentgrid.sh via hn) Hi HN, I’m Michael, I built AgentGrid with my friend Souren because we were tired of losing our coding sessions and couldn't keep track of what was actually being built across our many projects. Our approach was to use an infinite canvas d…
Show HN: SkillPreflight – score AI agent skills before installing them (github.com via hn) SkillPreflight English | 简体中文 SkillPreflight is a pre-install safety, token, and maintainability scorecard for AI agent skills. It helps users decide whether a Codex, Claude Code, Cursor, Gemini CLI, or other agent skill is safe and lightw…
Ask HN: Coding is a solved problem. What is left for experienced engineers? (news.ycombinator.com) I have come to the conclusion that coding is a solved problem. Most normal software building is now a closed-loop problem.
Show HN: Sol-Luna Orchestrator – let Codex decide whether to delegate (www.npmjs.com via hn) MCP server that lets a supervising OpenAI Codex agent delegate bounded executable tasks singly, sequentially, or in parallel, with per-task file scopes, scope-violation detection, and independently verified results. sol-luna-orchestrator A…
Show HN: Hands-Rust MCP/CLI that sees the Windows desktop and clicks real Chrome (news.ycombinator.com) I built Hands because I wanted a coding agent to use this Windows PC and a real Chrome profile the way I do: look at the screen, move the real mouse, type, click , without turning Chrome into an automation browser. It is a Rust MCP/CLI.
Show HN: Locum – Grok Bot delegates tasks to Claude/Codex CLI on your machine (github.com via hn) Locum A locum is a qualified professional who temporarily does someone else's job. Lets Grok Bot delegate coding work to the Claude Code and Codex CLIs you are already logged into on your own machine, instead of burning Grok Bot usage on i…
Tell me "Yes" if you are handling any of below AI coding chaos.. (news.ycombinator.com) If you're using Cursor, Claude Code, Codex, Antigravity or any other coding agent.. Tell me "Yes" if you are handling any of below AI coding chaos..
Does herdr worth it? Harnesses like pi ship most of these features (news.ycombinator.com) What's the gap it fills that they don't? Agent Mutliplexing (per tab) is already implemented in claude code, codex, pi (with plugins), ...
OpenAI reaches 20M Codex users, credits all accounts with banked resets (twitter.com via hn) It's me again. I come bearing great news.
Which is the best HARNESS? Ship benchmark with codex, jcode, pi, OpenCode, dsh [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Ask HN: Do you need Google ads help? (news.ycombinator.com) Hello everyone! I’m building https://unfetch.com, something like Codex but focused on Google Ads.
Building safe MCP servers for a PostgreSQL database (blog.pamelafox.org via hn) Model Context Protocol (MCP) is an open protocol that describes how agents can connect to external tools and data sources, and is now widely supported by the most popular coding agents (like GitHub Copilot, Claude Code, and Codex) and agen…
↯ Copilot↯ Model Context Protocolmodel-context-protocolcopilotcodex+2
Agent Voice, communicate with your coding agent all from Voice, on the fly (github.com via hn) AgentVoice Formerly "Cursor Voice" — see docs/26-rename-agentvoice.md. Self-hosted voice bridge for driving a coding agent CLI — Cursor (cursor-agent), Codex, or Claude Code — by speech, from your phone.
Show HN: Rove – parallel coding agents that can fan out subtasks and report back (github.com via hn) Rove — the agent multiplexer for your terminal Rove is a terminal-native workspace for running multiple coding tasks in parallel with Claude Code, Codex, Copilot, Kimi, or any CLI you register. Rove isolates parallel work in git worktrees…
Find jobs with Claude Code CLI (github.com via hn) Magnificent Jobs LIVE JOBS: 3,334,183 · ADDED TODAY: 211,021 · US find jobs with our cli using claude code and codex. We scrape the internet every hour to find you jobs linkedin and indeed cant find.
Show HN: Codex Memory Trim – prune and dedupe Codex CLI's global memory (github.com via hn) codex-memory-trim A skill that puts Codex's global memory on a periodic diet: detect duplicates, prune stale threads, compress verbose entries, and add custom global rules through a safe channel. English | 中文 Why this exists Have you notic…
I run AI coding agents as a team with AI DevKit (codeaholicguy.com via hn) I spend most of my time in the AI DevKit agent console. I start one agent as the manager, usually Codex, and brainstorm with it.
Show HN: Deploy apps with agents, no humans required (up.compartment.dev via hn) Hi HN, we built Compartment with a simple idea in mind: deployment should be as simple as possible. Just write this to Claude or Codex: “Deploy this app to up.compartment.dev.” Your agent authorizes itself, inspects any existing deployment…
Codex: Changes to reduce risk of destructive actions (twitter.com via hn) Hi! Recapping some changes we have rolled out over the last couple of weeks that have further reduced the risk associated to potentially destructive actions being performed by Codex during its work.
Codex and Hermes independently found the same critical flaw in my cryptocurrency (atto.cash via hn) How Codex and Hermes independently found the same critical flaw in my cryptocurrency I built Atto mostly alone. Atto now confirms a typical transaction in about 202 ms on its live network.
Show HN: Codewindow – Picture in picture for terminal agents (codewindow.app via hn) Hey HN - when I first started using Arc browser from the Browser Company, I loved the picture in picture feature they got for videos running in the background. When you switch desktops or windows, the video would just pop up and keep playi…
Show HN: Agents Workbook watch Claude Code, Codex write down their working notes (github.com via hn) agents-workbook Give your coding agent a workbook; the agent thinks out loud and writes its reasoning, then you reads it while it works. A local proxy that adds one tool to every request going out of Claude Code or Codex: somewhere to thin…
Show HN: An n8n-like orchestration toolkit for DeepSeek harnesses (github.com via hn) Open-DeepSeek-Harness-Desktop English | 简体中文 An open-source macOS desktop app for DeepSeek Harness (dsh) — a Codex / Claude Code desktop counterpart built on top of the harness, with no fork of the upstream kernel. It wraps the harness's W…
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks (github.com via hn) LongHorizon-Harness Loop Engineering for Computer-Use Agents Give Claude Code, Codex, OpenCode, or DeepSeek Harness a goal once. Keep it working across desktop apps and the terminal for dozens of hours.
Rysh – an AI harness where Claude and Codex agents work as a team (Go) (github.com via hn) rysh-cli-code The Rysh CLI: an agentic terminal multiplexer written in Go. Tabs, panes, splits, and vim/htop working exactly as you expect — except every pane is also an agent that can answer prompts and call tools.
MacroBench – Can Your Agent Manage a Hedge Fund? (twitter.com via hn) Think your Claude, Codex, Hermes, or OpenClaw can manage a hedge fund? Find out for free in 60 seconds on https://t.co/UIXiSTK6YD 10^60 procedurally generated markets modeled after the '08 financial crisis, Dotcom bubble, and COVID crash.
Practical multi-agent orchestration in Codex (twitter.com via hn) GPT-5.6 Sol gets especially interesting when it has a team to work with. Codex's new Multi-Agent V2 tools give Sol and Terra a natural way to delegate tasks, share updates, and coordinate through complex tasks.
Ask HN: How do you keep track of all your codex/Claude sessions? (news.ycombinator.com) I create so many assets, for both dev and GTM, have many that remain unfinished... How do you keep track of it all?
Show HN: Sol-Luna – adaptive Codex orchestration that can choose zero workers (github.com via hn) sol-luna-orchestrator An MCP server that lets a supervising OpenAI Codex agent delegate bounded implementation tasks to isolated worker threads — one at a time, or several in parallel in their own git worktrees — with a declared file scope…
Codex hit 15M users four months early (aicharts.grok.me via hn) Users · 12 Aug 2026 · High confidence Codex hit 15 million users four months early The June 2 trendline pointed to 10 million by mid-October. Combined Codex + ChatGPT Work crossed 15 million in August.
Use ChatGPT account for image generation API (github.com via hn) Codex Artifact Server A standalone, local HTTP service that turns Codex SDK runs into synchronous text responses and durable image artifacts. The directory is self-contained: it does not import application code or inspect the parent reposi…
Ask HN: Are Coding Harnesses like CC/OpenCode/Codex using the same techniques? (news.ycombinator.com) For those in the know, trying to objectively pick a harness outside of the marketing hype.
Show HN: Agentic Ship – open-source Lovable alternative that runs on your agent (github.com via hn) Hi HN. This is an open-source harness that helps you ship a fullstack application the agentic way, you only pay for your AI subscription (Claude, Codex, Cursor) and the domain.
Show HN: Built an ESP32 remote that groups Sonos speakers and sets group volume (github.com via hn) My Sonos remote groups all my speakers with one touch. Grouping and volume are what I use most, and going into the app every time is annoying.
Modelio 6.2 ported to native Apple Silicon ARM64 with Codex (github.com via hn) This work adds a native Apple Silicon port of Modelio 6.2. OpenAI Codex, using GPT-5.6 Sol, investigated, implemented, built, and functionally validated the port through an autonomous coding task.
Agentrove: Self-hosted AI workspace for orchestrating coding agents (github.com via hn) Agentrove Self-hosted AI coding workspace for running and orchestrating Claude Code, Codex, Copilot, Cursor, Grok, and OpenCode agents from one interface. What It Does Runs Claude, Codex, Copilot, Cursor, Grok, and OpenCode through ACP ada…
SilkCode – Tired of resets/limits with Claude but better than Codex (silkcode.web.app via hn) Multi-session GUI Project explorer, streaming chat, agent activity timeline, diff viewer. Open parallel sessions with different models; a busy session keeps working while you use another.
Show HN: OpenAI network checker using 31 browser and IP signals (www.getplus.ai via hn) 官方地区支持 只判断当前 IP 国家或地区是否出现在 ChatGPT 官方支持列表中。它不判断账号、支付方式、套餐或具体功能权限。 31 项环境信号扫描 ChatGPT 运行环境、地区画像与网络出口,适用于 Plus、Pro 与 Codex 点击检测,生成 ChatGPT 环境风险报告 本站应用不主动保存检测历史;检测在浏览器与本站服务器间完成 -- 封号风险分 等待检测 分数越高,环境越接近受限地区特征,封号风险越大 系统时区 等待检测 浏览器语言 等待检测 语言文字特征…
Show HN: I modded the Codex client to auto-route to the "right-sized" model (apimade.com via hn) Inspired by Bullet (https://news.ycombinator.com/item?id=49283063), I thought "why doesn't the ChatGPT client support auto-routing?" Switching between models is cumbersome, and I'm constantly using the native client -- so I thought. why no…
Show HN: A prompt to verify Supabase RLS policy (defencecore.com via hn) Hey guys, We wrote an agent prompt that you can simply give to your claude, codex agents, and it will automatically verify the RLS policy on each and every table in your supabase project.
Show HN: Multiplex – Open-source SSH/tmux/herdr terminal for Vision Pro and iPad (github.com via hn) I made this app from my need: want a multi-window terminal with first-class tmux integration on mobile, every session as its own window, not panes trapped in one frame. So I chose my Vision Pro as an attempt, and made my first visionOS app…
Mux Beacon – macOS menu-bar inbox for Claude Code/Codex agents in tmux (github.com via hn) Mux Beacon Know when terminal agents need you—and get back to the exact tmux target. Mux Beacon is a native macOS menu-bar inbox for Claude Code and Codex.
Show HN: Tmux-agent-switcher: see which Claude/Codex agents need your attention (github.com via hn) Since the start of this year I've coded pretty much exclusively by running multiple AI agents in parallel. Having tried a bunch of tools for managing them, I kept coming back to tmux.
Show HN: An idle desktop incremental game driven by coding-agent (news.ycombinator.com) While Codex or Claude is thinking or using tools, the game rains balls. As soon as the agent yields, new balls stop falling.
Show HN: Kery – comments on your PR with a video of the feature working (github.com via hn) I ship all of my code via different agents, mainly Claude Code and Codex. And with 5 different sessions running on different worktrees, the only way for me to know if something broke is once it gets to the deployment stage or production.
Rootless Container Sandbox for Claude Code and Codex (www.reddit.com via hn) could not extract summary
Show HN: I used an expiring Codex reset to port QtScript to Qt6 (github.com via hn) Hey HN, here is a short story that might be worth sharing. I have an application that has been scriptable with QtScript for years.
Show HN: VoxHearth / small local privacy-focused STT Mac app (github.com via hn) I wanted to be able to use STT in the terminal for Codex / coding agents and didn't like any of the existing tools that appear to have a number of security holes so I made this one. Enjoy!
Show HN: Airship – Figma-Like Visual Editor for Claude Code, Codex and OpenCode (github.com via hn) I wanted to experiment with my UI like I do in Figma, without having to rebuild it in a separate design tool. So I built Airship.
Show HN: Oqoqo – build evals and custom benchmarks for real-world tasks (oqoqo.ai via hn) Most benchmarks today exist in curated environments and do not translate well to the real world. We built Oqoqo to bridge this gap.
Show HN: Lians AI, Token-bounded memory and evidence for AI workflows (github.com via hn) we built this for improved memory and reduced token usage for claude code and codex
Show HN: Multicoder ACP – Run Claude/Codex/OpenCode from the same VS Code GUI (marketplace.visualstudio.com via hn) I am working on a VS Code extension which allows to run any (ACP-compatible) agent harness from the same GUI. Shows session outline - all tool calls, click for details.
Databricks Cost Optimizer: Audit Spend with Codex or Claude Code (github.com via hn) Databricks Cost Optimizer A read-only-first Databricks FinOps toolkit and Agent Skill for finding cost spikes, explaining their workload impact, and applying only explicitly approved optimizations. It works in both Codex and Claude Code: C…
Show HN: CtxRay – see and lock what Codex loads before a task (github.com via hn) CtxRay The local-first observability and control layer for OpenAI Codex. Audit context, compile intentional profiles, catch configuration drift, and attach honest usage receipts.
Ask HN: Should a coding client import another client's rules by default? (news.ycombinator.com) An AI coding client automatically reads ~/.codex/AGENTS.md and ~/.claude/CLAUDE.md. These are personal instruction files inside competing coding clients' configuration directories, outside the project selected by the user.
Show HN: Remaking Unreal engine in Rust for coding agents (machinesatplay.com via hn) Hi HN! I'm kevin, cofounder of https://machinesatplay.com, a multiplayer 3d gaming engine for codex/claude code.
Show HN: Vibsync – One Shared Memory for Claude Code, Cursor and Codex (MCP) (vibsync.com via hn) 01onboard Join in the middle A fresh agent receives team learnings, unfinished work, unanswered questions, and active work scopes in one brief. 02remember · recall Keep discoveries and knowledge Record causes, solutions, failed approaches,…
Show HN: Dsv4 Codex Proxy – Make DeepSeek V4 Flash 0731 "Codex-Native" (github.com via hn) Dsv4 Codex Proxy Dsv4 Codex Proxy makes DeepSeek V4 Flash 0731 work as a first-class model in Codex. Codex speaks the OpenAI Responses API, while most inference providers are chat-completions native and commonly expose the Responses API on…
Codex Pet (github.com via hn) Georgie, the Phalène Codex Pet Georgie is a silent focus buddy for Codex. He is a Phalène, the drop-eared variety of the Continental Toy Spaniel.
Show HN: Podiom – Persistent project context, goals and scheduling for AI agents (github.com via hn) Podiom A thin orchestration layer for local LLM agents (Claude Code and OpenAI Codex). Podiom shells out to the native claude and codex CLIs and leans on their MCP, tools, and skills, while owning its own durable truth: named agents, durab…
Neal: Codex and Claude in a loop shipped a 549-commit migration (navels.dev via hn) neal: coordinating different models on complex coding projects The first version of neal was a plan file and this prompt: Execute @plans/EMBER_MIGRATION.md. keep going.
Team Documentation (GDrive vs. Markdown) (news.ycombinator.com) How does your team manage your team documentation? At work we use Google Docs and its MCP, and most of my colleagues use Codex to write any documents.
Objective – Tickets, file claims, and proof gates for coding agents (github.com via hn) Objective Local-first ticketing for AI agents. Built for Codex and Claude to plan work, claim files, run tests, attach proof, and mark tickets done while humans monitor progress in a Linear-style UI.
Use local Chrome as a search API for coding agents (www.lsearch.dev via hn) Browser Search APINo API Key, No Billing Give Claude Code, Codex, Cursor, or any shell-capable agent structured web search through the browser on your machine. No API key.
Show HN: Mnemosyne – Export AI chat to carry context from one tool into another (github.com via hn) Hi HN, recently I was running out of token on different providers such as Claude, Copilot and Codex and their free or trail plan for personal purpose. I do change often model or provider but sometimes it's difficult to provide the same goo…
Show HN: Stop Animated Images from Autoplaying (neonglow.studio via hn) I’ve always been annoyed that browsers autoplay GIFs and other animated images. So I made GIF Control, a browser extension that stops animated images by default, shows the first frame, and adds a simple play/stop button directly on the ima…
Loop Engineering with native model switching in Codex and Claude (statewright.ai via hn) Visual workflow builder with protocol-level enforcement for AI coding agents. Your AI agent gets the right tools in each phase — nothing more, nothing less.
Debroid – Autonomous, headless Android debugger designed for AI coding agents (github.com via hn) Debroid 🤖⚡ The Headless Android Debugger for AI Agents Debroid (DEBug + AnDROID) is a headless CLI that speaks the Java Debug Wire protocol (JDWP) so AI agents — Claude Code, Grok Build, Codex, OpenCode, Cursor, Antigravity, or anything wi…
Show HN: Aident Loadout – connect Codex and Claude Code to real apps (github.com via hn) Aident Skill Aident AI helps agents move beyond chat and coding to get real work done across your tools and apps. Aident Loadout connects work apps and Aident-managed platform tools to AI agents.
Show HN: MusicMonitor – Turn your extra monitor into a Spotify now-playing view (github.com via hn) Have an extra monitor? Why not have it display a full-screen view of what's now playing on your Spotify?
Show HN: Python Concurrency Testing with Frontrun (github.com via hn) Hi HN, I'm the author of Frontrun, a Python concurrency testing library I built over the past 6 months using Claude Code and Codex. It uses bytecode tracing and various forms of monkeypatching to schedule across threads, async and multipro…
Is Muse Code a Claude Code Killer? (trycodus.com via hn) Muse Code vs Claude Code and Codex: what Meta’s own benchmarks actually say Meta shipped its first coding agent this morning, and within the hour my feed had already decided it was either a Claude Code killer or a nothingburger. Both verdi…
Let Claude/Codex run actual 1000s of LinkedIn outreach with ~30 MCP tools (gist.github.com via hn) You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window.
Show HN: A multiplayer play-money terminal casino for downtime (anteroom.johnramsey.com via hn) Hi HN! I created Anteroom, a hobby project allowing you to play casino (and non-casino) games with your friends in your terminal.
Show HN: Spacescience.tech – A Bubble Shooter mixed with 2048 (spacescience.tech via hn) Made this with Claude + Codex since I was bored, wanted a fun game to play on the plane and though the two would work well together. So far I've had fun with it, so I decided to share.
Show HN: OldHand A Claude/Codex plugin to verify the development flow end-to-end (github.com via hn) I built this clod code or codec skill because I just love using these AI agents to write code. It's just so convenient to do.
Show HN: Will Codex Reset – Predict Codex reset timing to maximize your quota (willcodexreset.com via hn) Next 48 hours LOW SIGNAL · stay measured 0% reset probability in 24 hours from now; 0% in 48 hours. A transparent signal desk for Codex resets — tracking timelines and predicting the next reset.
Agents that narrate their work are the best team players – Jon Udell (blog.jonudell.net via hn) Agents like Claude Codex and Codex run in your terminal. From that strategic vantage point they can wield system tools (awk, bash, curl, git, python) as well as MCP tools.
Using Codex as a diet and exercise coach (usefully-wrong.com via hn) I have been thinking for a while that I should follow an exercise plan with progressive overload and recovery tracking. My nutrition could be improved too.
What Codex Actually Sends to the Model (www.0xkato.xyz via hn) What Codex Actually Sends to the Model - 10 minsI recorded the requests Codex generated for a 16-character prompt, then measured what changed as it loaded instructions, exposed tools, read files, ran commands, received images, and compacte…
Show HN: Clawx – A package manager where packages are agent tasks (github.com via hn) Hi HN — I built Clawx, an experiment in treating agent-driven tasks as installable packages. A Clawx package is a Markdown file with YAML metadata.
Show HN: Fruman – A 2D souls-like game demo running in web browser (fruman.baklab.app via hn) Hi HN, As a souls-like game lover, I've tried use pure JavaScript acomplished a simple souls-like action game, it's mostly coding with AI agents (codex / claude).You can create your own map and scenes with the editor tool, and this is just…
Ask HN: I had Codex and GPT 5.6 Sol running for 12 days, 870k+ LOC. Now what? (news.ycombinator.com) For the last 2 weeks I have been running Codex + GPT 5.6 Sol Ultra non-stop for almost 13 days, working on a huge extension to my SaaS product / business. It used 502,122,866 tokens and created over 870k new LOC.
Show HN: MoneyCo – A coding agent built for accounting (thecompany.money via hn) Hey HN - Mert here, maker of MoneyCo (thecompany.money). MoneyCo connects to multiple Quickbooks and Xero accounts by default and lets you do your bookkeeping using a coding agent.
Show HN: Renovate Log Parser – Detect and Analyse Renovate Bot Problems (github.com via hn) Did you ever get frustrated while diagnosing an issue with Renovate Bot because of its very verbose debug logs? Then this is the right tool for you.
Show HN: Chaser (beta) – Share Mac windows as screenshots and UI context (chaser.vgnsh.xyz via hn) Chaser is a macos utility you can use to capture more than a screenshot. Screenshots are good tools to show your agents what you're seeing, but there's far more that goes on behind the scenes - the layout, button states, mapping of objects…
Codexloom: Turn Codex threads into an organization of long-lived domain agents (github.com via hn) CodexLoom Loom your Codex. A working environment for long-running Codex Agents.
Show HN: Rudder – Prompt-Driven Interactive Code Review (github.com via hn) As someone who is using agents to code for me a lot lately, I was starting to feel more and more like I'm losing touch with the code I write--it's starting to feel like Claude's code rather than mine. I'm pretty sure I'm not alone in feeli…
Setting Up and Using the Pi Coding Agent (deepakness.com via hn) Setting Up and Using the Pi Coding Agent I have been using AI coding tools for a while now. At the time of writing this, I am subscribed to both Cursor and OpenAI Codex.
Shots: Codex and MCP Plugin for Making App Store Screenshots (shots.run via hn) Shots is an AI App Store screenshot generator that works from your coding agent and from Studio. It researches your category, reads your real app context, and generates store-ready screenshots, app icons, and listing copy you can revise an…
A list of sandboxes for running AI coding agents, ranked by security posture (github.com via hn) Awesome AI Coding Sandboxes A curated list of sandboxing and isolation solutions for running the code of autonomous AI coding agents (Claude Code, Codex, OpenHands, and friends) — organized by security posture first: how strong the isolati…
Codex CLI Has Responsive Tables in the Terminal (www.vincentschmalbach.com via hn) Google Lighthouse Adds Agentic Browsing Checks and Cloudflare Adds AI Traffic Controls Google Lighthouse and Cloudflare are adding different controls for the same emerging class of automated web visitors. Lighthouse added an experimental A…
VibeMenu – a local macOS menu-bar dashboard for Claude Code and Codex (github.com via hn) VibeMenu Keep your Mac awake while your coding agent works—and let it sleep when the work is done. VibeMenu is a free, open-source macOS menu-bar app for Claude Code and the ChatGPT desktop app.
Show HN: AgentCodeGUI – Multi-Account Desktop GUI for Claude Code and Codex (github.com via hn) While working with Claude Code and Codex, I realized I didn't really need an IDE anymore. Keeping a heavy IDE open, waiting for it to load, and then opening yet another terminal tab just to run an agent was more overhead than it was worth.
Show HN: Turn a Nintendo Joy-Con into a Codex Micro Alternative (github.com via hn) Hi HN, I built Joy-Con Codex Controller, a bring-your-own-controller alternative to Codex Micro for macOS. It turns an already paired Nintendo Switch Joy-Con into a one-handed physical controller for Codex.
Show HN: CC Meter, a Windows Tray Gauge for Claude Code and Codex Limits (github.com via hn) CC Meter A tray widget for Windows that shows how much of your Claude Code and Codex rate limits you have burned, and how fast you are burning them. Representative usage rendered by the v1.0.4 full panel at native resolution.
Show HN: Galda – Come back to Claude Code/Codex without digging through Git (galda.app via hn) 打ち合わせの後、開発進捗をすぐに思い出すのは難しい… Claude CodeもCodexも、ひとつの場所で。どのタスクも、エグゼクティブサマリーになって戻ってくるので、コンテクストを見失いません。 無料で試すあなたにわかりやすいように、コンテクストや確認すべき内容がまとまっており、ストレスなく意思決定ができます。 ダークモード対応 Summary 設定に「ダークモード」トグルを追加し、全画面のテーマ切替を実装しました。切替はlocalStorageに保存され、再読み込み後も…
Supabase Evals: Benchmark for testing how well AI agents build using Supabase (supabase.com via hn) Today we're open sourcing supabase/evals , our benchmark and framework for testing how well AI agents build using Supabase. It runs coding agents including Claude Code, Codex, and OpenCode against real Supabase tasks, for example, building…
Show HN: Agentmetry – local-first flight recorder for AI coding agents (agentmetry.ai via hn) Open-source, local-first audit trail for AI coding agents. Records every tool call from Cursor, Claude Code, Codex and MCP servers, tags it with MITRE ATT&CK, and correlates sequences into detections, including the Agent Data Injection cha…
A Hermes Agent Skill Looping Between Codex and Claude Code (codenote.net via hn) I built a Hermes Agent skill that takes a single GitHub issue URL, delegates implementation to Codex CLI and review plus behavior verification to Claude Code CLI, and loops until all five review gates are clean at the same SHA. This post c…
Show HN: Claude MIDI Twister – An agent visualizer for a DJ MIDI controller (www.dylanfisher.com via hn) I was inspired by the release of the Codex Micro controller (https://worklouder.cc/codex-micro) and repurposed a DJ TechTools Midi Fighter Twister I had laying around into a hardware visualizer for my Claude Code sessions. This is a little…
Show HN: I rebuilt my 2008 soccer/lineup site (news.ycombinator.com) I originally built fantastic11.com in 2008. The browser-based lineup editor alone took me several months to build.
Show HN :Tandem – one session across Claude Code and Codex, no double spend (github.com via hn) 🤝 tandem One coding session. Two AI agents.
Show HN: Wasted Cycles – Local wall-clock profiler for AI coding agents (zozo123.github.io via hn) Wasted Cycles is a local terminal profiler that reads the Codex, Claude Code, Cursor, and Grok Build traces already on your machine and shows where a run stopped coding: model work, reads, edits, verify, CI waits, human handoffs, and retri…
Show HN: Skytrace – Self-hosted 3D ADS-B viewer with receiver coverage domes (sky.luftaquila.io via hn) I built a self-hosted 3D ADS-B viewer to see how obstacles affect receiver coverage. It builds a 3D coverage dome from 30-day history.
Show HN: Memsprout – share AI context with your teammates (memsprout.com via hn) I’m an architect at my company and everyone on my team is now using Claude/Codex agents for their work. Managing context in claude.md/agents.md files is fine but so much of the knowledge required doesn’t fit cleanly in one repo or the othe…
Copatch- The shared workspace for teams building with AI coding agents (copatch.ai via hn) One workspace for people and AI agents Join the CoPatch waitlist. Be first to bring your team, Codex, Claude, and more into one shared workspace—with you in control.
Outside LLMs – Outside Lands and Codex (outsidellms.com via hn) Make more music. Build tech that helps musicians focus just on making music—and reduces their busywork, promotion, admin, monetization, and other non-musical responsibilities.
Husk – a desktop workspace for terminal AI agents (github.com via hn) Husk wraps claude, copilot, codex, aider, or any other terminal-based AI agent in a clean Electron window with a real PTY, drag-drop file context, voice output, session resume, and a one-glance dashboard. The reasoning, thinking format, an…
Show HN: Supapool – a Supabase per coding agent, in ~400 ms (supapool.io via hn) I usually have around five Claude Code or Codex terminals working on one repo, each in its own git worktree. The repo needs Supabase, and the agents kept stepping on each other: a shared local instance means one agent's test truncate wipes…
Show HN: [OSS, local] Spatial Board for all your agents (github.com via hn) Agent Board Spatial board for Claude Code, Codex, Cursor and opencode. One card is one conversation: start agents into any project folder, drag the cards wherever they make sense, and see at a glance who is working, who finished, and who i…
I compared 5 popular token saving methods in Codex and found that none delivered (www.stet.sh via hn) I Tested Six Ways to Save $$$ on 5.6 Sol. The only one that worked was using Terra TL;DR - I compared six approaches meant to save tokens (Caveman, Ponytail, RTK, Context Mode, Mandarin prompts, and 5.6 Terra xhigh) with a baseline across…
Show HN: ShellTeam – web app to steer coding agents on your VPS (github.com via hn) A command center for coding agents you run on your own VPS Install · Website · Architecture · Security · Contributing ShellTeam turns a Linux box into a cloud computer you command through a team of coding agents. Drive Claude Code, Codex,…
Cloister – isolated Claude Code and Codex, one cell per company (github.com via hn) Cloister Isolated Claude Code & Codex, one cell per company. Real logins, per-tenant MCP servers, encrypted secrets — a walled cell for each client's agents, driven by one terminal command: cloister.
Show HN: Code Security Skills Codex-Inspired Workflows Packaged for Claude Code (github.com via hn) Code Security Skills Thirteen evidence-driven workflows for threat modeling, code review, finding validation, attack-path analysis, remediation, tracking, and deterministic security reporting. The procedures are provider-neutral: they desc…
Show HN: Mustuse.ai – Open-source automated ranking system run by agents (mustuse.ai via hn) Hi HN, MustUse.ai is an open-source framework for building ranking sites that Codex agents can fully research and maintain automatically. Friends often ask me which AI tools they should use for different tasks.
SessionRadar – Monitor Claude Code and Codex sessions from your menu bar (sessionradar.com via hn) macOS 14+ · menu bar app · for the agent era Every coding agent, one glance away. SessionRadar watches your Claude Code and Codex CLI sessions — live status, cost and tokens, grouped by project.
Show HN: Hamza Mask secrets and PII before Claude Code or Codex sends them (github.com via hn) Hamza *The name Hamza is inspired by the undercover operative in Dhurandhar.* Hamza is a proxy for Claude Code and Codex. It masks detected secrets and approved types of personal data before sending prompts to Anthropic or OpenAI.
Show HN: NoClick – Build always-on agents with your existing AI subscriptions (www.noclick.com via hn) Hey HN! Been working on a platform that lets you build always-on background AI agents that work with your existing subscriptions/harnesses of choice (Claude Code, Codex, OpenCode, Hermes, OpenClaw).
Show HN: Cetus – A macOS App for Claude Code, Codex, OpenCode, and More (github.com via hn) Cetus One macOS app for Claude Code, Codex, and every agent runtime you use. They live in one place, so you can schedule them to run while you're away, summon one over any app with a hotkey, give each its own git worktree, and review every…
Show HN: Dn – plan collaboratively, let agents execute (github.com via hn) Hi HN — we built `dn` because coding got faster, but building software did not. Use the CLI to clear your backlog faster with reusable agentic workflows & other supporting commands.
Show HN: DynoTable – The DynamoDB GUI with real SQL and your AI agent (dynotable.com via hn) I've been using DynamoDB every day for work and I needed a fast, keyboard first, modern desktop client that will allow me to query, aggregate, edit items in few clicks and that integrates well with Claude Code or Codex. I didn't like any e…
Show HN: BeatFlow – A Codex skill for composing and validating multi-track MIDI (github.com via hn) BeatFlow Skill Compose complete, editable multi-track MIDI with Codex. BeatFlow turns a musical brief into an explicit Python composition plan, validates its timing and musical relationships, realizes functional pitches and chord voicings,…
BaseCode – MDM for Coding Agents (basecode.cloud via hn) Know and control how coding agents are being used in your organization. Your engineers run Claude Code, Codex, Copilot, Gemini and opencode on their machines.
Show HN: Graph Skill – coding agents run tasks as dependency graphs (github.com via hn) Graph Skill Graph engineering for Claude Code, Codex, OpenCode, Cursor, and OpenClaw. Stop having one long agent conversation.
Agent Mesh: let Claude Code, Codex, Grok and other agents talk to each other (github.com via hn) agent-mesh Let your coding agents talk to each other. You probably have several agent CLIs installed: Claude Code, Codex, opencode, Gemini, Grok.
Embedding OpenAI Codex: The App Server and SDKs (www.akashtandon.in via hn) Embedding OpenAI Codex: The App Server and SDKs OpenAI Codex is not one tool with one front door. Under the CLI, the IDE extensions, and the cloud sits a single agent you can drive yourself, and there are two ways to do it: the app-server,…
Claude-video – Give Claude the ability to watch any video (github.com via hn) /watch Give Claude the ability to watch any video. Claude Code (recommended — auto-updates via marketplace): /plugin marketplace add bradautomates/claude-video /plugin install watch@claude-video Codex, Cursor, Copilot, Gemini CLI, or any o…
AgentHost – Persistent, governed AI agents in your own cloud (agenthost.space via hn) Four agent interfaces and one messaging gateway, on infrastructure you own. Claude Code and Codex can execute and cross-review approved autonomous tasks.
Show HN: Codex Plays NetHack (asciinema.org via hn) I'll add a link to the completed recording after the stream completes.
Show HN: Read what your remote agent wrote (Markdown/image) from Mac menu bar (lx2026.github.io via hn) I posted v1 of this app here and got great feedbacks (SSH port forwarding). Since then I added a 2nd part I now use a lot too: remote file browsing and markdown preview, without scp or a mount.
Show HN: Codex fixed my 2010 MacBook Pro driver troubles with kernel patches (github.com via hn) I installed Gentoo on my mid-2010 Macbook Pro about a year ago as both a fun learning project and to have a practical secondary laptop for terminal access, light dev, using utilities that would be more cumbersome on my Windows laptop, etc.…
Ask HN: Anyone else experiencing extreme Codex limits token burn rate? (news.ycombinator.com) could not extract summary
Show HN: I built a transparent terminal wrapper for unobtrusive AI (github.com via hn) I wanted a way to use my existing AI subscriptions as a terminal assistant to analyze shell output or help me with a complicated awk incantation from time to time without getting in my way the rest of the time. I couldn't find anything exi…
Show HN: Termic – desktop app for running CLI coding agents (Claude Code, Codex) (termic.dev via hn) So, https://termic.dev open source desktop app for managing claude code / codex / agt / etc coding sessions. main checkout or worktrees.
Show HN: ScreenFocus – keyboard focus follows the pointer across Mac displays (github.com via hn) I have a multi-monitor setup and spend a lot of time typing in Codex and other AI agents. When I move to another monitor to run a shortcut, focus often stays in Codex or another text app, so the shortcut runs in the wrong place.
Show HN: Rules that stop AI coding agents from breaking working code (github.com via hn) Stop Your AI From Breaking Working Code Rules and prompts that keep Claude Code, Cursor, and Codex from wrecking things that already worked. Everything here is copy-paste.
Show HN: Argus – VSCode Worktree Agent Session Manager (marketplace.visualstudio.com via hn) Running parallel Claude Code/Codex agents across worktrees in vscode is annoying. You need different vscode windows open and have to run around terminals and dev servers.
Codex now leads Claude Code in first-time Homebrew installs for last 30 days (formulae.brew.sh via hn) /api/analytics/cask-install/30d.json codex claude-code gcloud-cli libreoffice google-chrome docker-desktop visual-studio-code android-platform-tools android-commandlinetools ghostty ngrok git-credential-manager iterm2 1password-cli flutter…
Trajectory: A Standard Format for Agent Experience Data (www.letta.com via hn) Agents today have a limited ability to learn from past experience to improve in the future. Furthermore, agent experience is split across many different harnesses: many users swap between harnesses such as Claude Code, Codex, and Letta Cod…
MCP Clock: live public hosted tool for Claude, Codex, any MCP Client (github.com via hn) MCP Clock Live MCP endpoint: https://mcpclock.firasd.workers.dev/mcp Works in Claude.ai, Claude iOS, Claude Code, OpenAI Codex, and any MCP-compatible client. Tools clock_get Returns current time in one or more time zones.
Show HN: Sync Claude Code and Codex configs, with a board that shows the drift (github.com via hn) AI Config Sync Manager Continuous bidirectional sync between Claude Code and Codex — round-trip lossless, not a one-shot migrator. Highlights Continuous bidirectional sync — claude → codex and codex → claude, run as often as the two hosts…
Show HN: WhipDesk – Control your full dev machine from your phone (github.com via hn) Hi HN, I built WhipDesk, an open-source, mobile-friendly remote desktop tool for controlling your full development machine and managing AI coding agents from your phone’s browser. It gives you access to the entire desktop while also adding…
Show HN: Drive your real logged-in Chrome from Claude Code and Codex (MCP) (github.com via hn) Drive your real, logged-in Chrome from an AI agent - over the Model Context Protocol. No headless browser.
Open Erdos problems solved with the help of GPT-5.6 Sol (twitter.com via hn) I solved 6 open Erdős problems in 5 days, using @OpenAI GPT-5.6 Sol. I have a math background, but the Codex workflow I used does not require deep mathematical knowledge.
Show HN: CobaltCode – Dedicated persistent computer for Codex (cobaltcode.ai via hn) Hi HN, I'm building CobaltCode, a platform for running coding agents inside persistent development environments. Basically I was fed up with git worktrees and having to run everything locally, usually one at a time so built CobaltCode.ai.
Show HN: 5dive – Run a Company of Claude Code/Codex Agents (Written in Bash) (github.com via hn) run a company of AI agents on a server you own English | 简体中文 Quickstart · Why 5dive · Zero-human proof · Use from your AI agent · Security · Full CLI docs · Managed VM A company of AI agents, and the orchestrator is just bash. No framewor…
Show HN: ArXivMax – video explainers for any research paper (www.arxivmax.com via hn) I have a bad habit of bookmarking arXiv papers and never getting around to them. So I built arXivMax to make it easier to go through them.
Codey – A multi-agent workbench for Claude Code, Codex and OpenCode (github.com via hn) Codey 🚀 English | 中文 A multi-agent workbench for coding agents. Codey is one place to organize, switch between, and orchestrate Claude Code, OpenCode, Codex (and more) across your projects — give each project its own workspace, build worke…
Show HN: Searchdesk, AI powered job search that tailors resume and cover letter (searchdesk.app via hn) Via an install script, Searchdesk uses your local codex installation as a backend to search for jobs for your preferred title matched to your resume. It also tailors your resume and cover letter for each job, searches best contact on Linke…
Show HN: A local, extensible session viewer for Pi and Codex (github.com via hn) A local, read-only web viewer for Pi agent session files.
Ravenspire – watch your Claude Code and Codex agents as a JRPG (github.com via hn) 🐦⬛ Ravenspire Mission control for your AI agents — as a JRPG. Every Claude Code and OpenAI Codex session on your machine becomes a pixel-art hero in a living guild world: their task is a quest, working means battling a monster sized by th…
UseReserve – a free macOS menu bar app for tracking Claude and Codex usage (usereserve.app via hn) Open Codex and Claude Desktop once, then launch UseReserve to see the limits available on your Mac. Claude may request one-time permission to read your local sign-in.
The Codex Micro Is Physical Slop (paulmakeswebsites.com via hn) The Codex Micro is Physical Slop OpenAI recently released their first piece of physical hardware, the Codex Micro. This product release came a little over a year after they announced they were working with Jony Ive on new AI-powered physic…
Grok is a surprisingly good automated theorem prover (news.ycombinator.com) TL;DR: I'm working on a Python package called OpenATP [1] that provides a common interface to coding agents for automated theorem proving in Lean. In the latest release, I added support for Leanstral 1.5 [2,3], Grok, and Kimi Code [4].
Show HN: WakeWire – Push GitHub, Gmail and Slack Events into Local Codex Threads (github.com via hn) Author here. When working on optimising my workflow, I noticed that some of the now famous ‘loops’ for Codex agents were loops that burned tokens to discover nothing happened.
Show HN: Llamatop - see what the various cores on your macbook are doing! (github.com via hn) I had codex make this for me so I could see what llama.cpp was doing on my mac. Vibecoded in swift, hit 1 for more processor details, hit m for more memory details.
Show HN: Public-safe skin packs for the Codex desktop app (codex-theme-gallery.howardhua.chatgpt.site via hn) Start with Caishen Readable Traffic from Codex theme lists is already landing on this installer. Caishen Readable is the lowest-friction pack to try first because it keeps text readable and has a direct theme page, zip, and fetch command.
Show HN: Use your flight-sim gear as a Codex Micro (github.com via hn) When I saw the OpenAI Codex video for the (cool-looking) Codex Micro, I looked down at my Virpil CM3 throttle collecting dust. So I just asked Codex desktop to connect it for me-- and it turns out the latest desktop app has the hotkeys/bin…
Show HN: OpenAI Hackathon Submission: ADE (devpost.com via hn) After 2 months of hard work, Cycling through three 20x Codex Subs, three 20x Claude Code subs, and using over 90 billion tokens, I have finally submitted a working Native App (Cross-Platform, but mainly MacOS and Windows Supported, Linux i…
Show HN: Qscreen – tmux-like session manager for Windows PowerShell (github.com via hn) I'm an indie developer and use both macOS and Windows for development. Most of my indie game work is done in Godot on Windows.
Show HN: Codex Micro on Your Phone (github.com via hn) Codex Micro Phone [!IMPORTANT] Unofficial project. Not affiliated with, endorsed by, or supported by OpenAI or Work Louder.
Show HN: Cognikernel- Local Memory for AI Coding Assistants (github.com via hn) Every coding session in claude code/ codex starts blindly. It forgets the archiectural decisions you made in a session you ran yesterday/ a week ago or a month ago.
Show HN: Codex Harness for Java Unit Testing (github.com via hn) JAIPilot CLI Generate Java unit tests locally with Codex and track JaCoCo coverage from the terminal. jaipilot-cli is a Java-only local workflow.
Show HN: A simple TUI to CRUD Beads tasks (github.com via hn) Wasn't really happy with any of the beads TUIs out there, so set Codex/Claude Code on creating one that allows for quick task management inside a repo that has an associated beads database. Quite happy with the vim style UX and the tree ba…
Show HN: CallBro – Granola, but Powered by Codex, Claude Code, or Your Local LLM (callbro.ai via hn) We took the bot-free meeting-notes experience people already understand and replaced the closed intelligence layer with a full agent chosen and controlled by the user. ASR model on your device.
Show HN: Run Hermes Agent with Kimi 3 in a sandbox (news.ycombinator.com) Today, we are releasing Hermes Agents running on the Kimi K3 model inside a sandbox. You don’t need a Mac mini to run one!
Use caffeinate to keep Claude Code and Codex running on Mac (www.narendravardi.com via hn) macOS ships with a small command called `caffeinate` that keeps your Mac awake during long-running terminal tasks. It is especially useful for Claude Code and Codex sessions over office VPN, where sleep can disconnect the network session a…
Show HN: Qelvora, a local Mac writing tool I built by directing Codex (qelvora.app via hn) Correct spelling and grammar in selected macOS text with Ollama models running locally. Read the guideLocal-first correction for macOS Polish text without sending it away.
Ask HN: Are you teaching your kids to program? (news.ycombinator.com) I'm still teaching my kids to code. I'm interested to hear if I'm increasingly alone in this, or if others are doing it too.
OpenAI encrypts Codex agent instructions, blocking local audit trail (www.theregister.com via hn) MOST POPULAR AI - AI and ML OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake' Data purges deemed an example of 'misaligned behavior' that upstart is working to avoid - AI and ML Researcher poisons open-weight…
Show HN: Scribe, a CLI that builds AI agent memory from your repos and sessions (getscribe.dev via hn) Context-aware agents scribe init writes a handshake block into both ~/.claude/CLAUDE.md and ~/.codex/AGENTS.md , so every session in every project queries your KB before recommending a library or proposing an architecture. scribe reads you…
Show HN: PocketVeto is a Bluetooth-only AI agent remote control (github.com via hn) I kept seeing those ESP32 toys where you plug it to your PC / Linux and have a touch screen change when Claude Code / Cursor / Codex is asking for some permission, but I didn't want to go that far. I also missed a way to remote control the…
Choosing GPT-5.6 Sol, Terra, or Luna in Codex (twitter.com via hn) https://t.co/kjLUDCmImv eric provencher@pvncherArticleChoosing GPT-5.6 Sol, Terra, or Luna in CodexCodex for moonshots and everything in between Some missions demand deep planning and coordination. Others are a straight shot.
Why people chasing after useless token saving plugins and ignoring real solution (news.ycombinator.com) I wrote a blog yesterday on how useless RTK and Ponytail are on real coding tasks. And published my agent harness long-horizon task benchmarks on 80% real token saving.
I built a game for Claude and Fable to fight (twitter.com via hn) Codex stopped queueing up actions. Im calling it.
Call GPT 5.6-Sol Pro, Fable 5, SuperGrok Subscripions from Codex, Claude (github.com via hn) Agentify Desktop Agentify Desktop is a local control center for AI web sessions. It lets MCP-capable tools such as Codex, Claude Code, and OpenCode use the AI subscriptions you are already signed into, while keeping browser state, files, a…
Show HN: Skillful, stop maintaining the same AI workflow in five places (skillful.md via hn) Hey all, Over the past year I noticed I was maintaining the same AI workflows in multiple AI tools such as Claude code and Codex. Whenever I improved one workflow, I had to remember everywhere else I'd copied it and most of the time I didn…
OpenAI Releases Codex Micro (openai.com via hn) could not extract summary
Show HN: H5i-Python: Python SDK for Programmable Multi-Agent Orchestration (github.com via hn) h5i-python is the Python SDK for the h5i orchestra engine. This SDK lets you define and execute multi-agent coding workflows across Claude Code, Codex, and other runtimes as ordinary Python programs.
Show HN: Limits, an on-device iOS app for tracking AI usage limits (getlimits.app via hn) I built Limits because I wanted an easier way to keep track of when my limits would reset. Like say you're outside and just wanna know if your weekly or session limit is back.
Chinese Codex API Resellers Have a New Problem (www.vincentschmalbach.com via hn) Chinese services sell low-cost Codex API access by putting a layer between customers and a large number of ChatGPT subscriptions. A developer gets one API key.
Solving 20 Erdős Problems with 20 Codex Accounts Running in Parallel (www.starfleetmath.com via hn) Star Fleet Math Built by Colin Snyder · colin@colinsnyder.com Advised by Mike Kim · proposed solutions ↓ Inspired by Ignis · previously built by Myself, Dhruv Agarwal, & Nitin Kesarwani at the New Turing Institute Star Fleet is an AI syste…
Show HN: Liveshortly - stream and pair prompt with ai agents (liveshortly.com via hn) We built a tool for livestream your agents session and it also allows collaborative pair prompting with team mates in single claude or codex sessions. Once done with sessions you can post sessions as blogs or trail of conversation so that…
Show HN: Kmux – Parallel terminal workspace optimized for AI coding agents (github.com via hn) kmux The multi-session terminal workspace for running AI coding agents side-by-side. A keyboard-centric terminal emulator designed for Claude Code, Codex CLI, and Antigravity CLI on macOS and Linux.
Enhancing GNU-Pth for m:n threading using Claude and Codex (medium.com via hn) could not extract summary
The OpenAI Super App, ChatGPT = Codex, Whither Chat (stratechery.com via hn) OpenAI has refashioned Codex as the new ChatGPT; is the company abandoning the chat category they pioneered? Subscribe to Stratechery Plus for full access.
Launched BattleShip on Chaaaaa.com (news.ycombinator.com) Hey hackers, If you need a lil break to relax after using your whole Codex/claude usage, come on https://chaaaaa.com to play Battleship it's quick and free ! Will you beat my AI Commander ?
Show HN: Giving Claude Code and codex its voice using kokoro (github.com via hn) Aloud Aloud reads Claude Code and Codex replies aloud on macOS. Turn it on with aloud on in any Claude Code or Codex session.
Show HN: Codex Pet Web – put any Codex pet on any website (pets.caro.sh via hn) Accessible by default Native buttons, keyboard movement, focus restoration, coarse-pointer controls, and a quiet reduced-motion mode. A tiny JavaScript SDK · Kavana included One Web Component brings transparent atlas animation, roaming, dr…
Groupchatty – Real time chat while using Codex and Claude (error.workos.com via hn) Something went wrong Couldn’t sign in. If you are not sure what happened, please contact your organization admin.
Using Subagents to Improve Claude Code Results: A Step-by-Step Guide (software.rajivprab.com via hn) Unless you’ve been living under a rock, you’re well aware of AI agents like Claude Code and Codex revolutionizing software development. These tools are undoubtedly easy to learn, but are also hard to master.
How to Quantitatively Evaluate Prompt Quality in Claude Code and Codex (medium.com via hn) could not extract summary
Ask HN: What makes someone good at using Claude Code? (news.ycombinator.com) I am building Promptster - an AI fluency platform that helps level up engineering organizations. Engineering managers invite their teammates and Promptster analyzes the engineers work with ai coding tools (claude code, codex, cursor, copil…
Users report that GPT-5.6 Sol has become less capable than its initial release (twitter.com via hn) Updates for Codex and ChatGPT Work users. No nerfing, only good stuff!
TeamBrain – Git-Native Shared Memory for Claude Code, Cursor and Codex (teambrain-site.netlify.app via hn) Claude Code, Cursor, and Codex — reading from and writing to the same team memory. Approved by humans, stored in git, learned from real sessions.
MCP with Keycloak, Claude, Codex and a Whole Lot of Coffee (blog.priyavijai-kalyan2007.workers.dev via hn) Keycloak with Claude, Codex & A Whole Lot of Coffee The Problem Hello. I've been working on a side project: a collection of tools I wished existed in a better form, and in some cases tools which don’t exist at all as far as I can tell.
Run SSH and Claude Code on 3DS (news.ycombinator.com) I found an open source homebrew 3DS application that can connect to a remote host via SSH and run Claude Code on it. I tested it on my own 3DS,the experience was great—it even supports voice input.
Show HN: Baton - Know which of your AI coding agents needs you (github.com via hn) A local command center for every AI agent you have in flight. Claude Code sessions and Codex threads, in one glance, so you always know which one needs you right now.
Temporarily removing the 5 hour usage limit restriction for all paid Codex plans (xcancel.com via hn) could not extract summary
Show HN: Topsoil – a notch dashboard for coding agents, music, and files (topsoil-two.vercel.app via hn) Hey HN, I built an OS notch tool, Topsoiil because I kept alt-tabbing to check on my coding agents, dragging and dropping files/screenshots and listening to music. Topsoil lives in the notch of your MacBook.
Show HN: Almanac – A self-updating wiki from your files (usealmanac.com via hn) I was maintaining a wiki for my company's context through a hacky project I had made, but I wanted to turn that into something that's a lot more polished and easier to use. I wanted something where: 1.
Show HN: Kote – Capture and reuse engineering context from AI chats and Git (github.com via hn) I kept running into the same problem: I'd solve something with the help of an AI assistant, spend time debugging an issue, or make an architectural decision... and a few weeks later I couldn't remember where that information was.
Show HN: Block dangerous Git and shell commands from being executed by agents (github.com via hn) dcg (Destructive Command Guard) A high-performance hook for AI coding agents that blocks destructive commands before they execute, protecting your work from accidental deletion across Claude Code, Codex CLI, Gemini CLI, Copilot CLI, VS Cod…
Show HN: Standalone SearXNG CLI+MCP (no server needed) (github.com via hn) Hi HN, Codex and Claude are pretty good at (re)searching things on the web these days, but the open coding agents (OpenCode, pi coding agent and friends) don't have access to the labs' proprietary search APIs. I wasn't happy with this stat…
Show HN: Aether – Run Claude Code, Codex, or OpenCode in devboxes you can watch (www.runaether.dev via hn) Since coding agents like Claude Code and Codex came out, I've been pretty obsessed with them. It's hard not to when you're getting a 20x discount on inference.
Show HN: A meditative waiting room for Claude/Codex (waitingfor.ai via hn) I kept finding myself getting distracted during those unpredictable stretches while Claude or Codex was working, so I built a little place to spend that time instead. It’s a mix between a meditation zone and a video game; there’s enough to…
Ask HN: How are you controlling Token Costs? (news.ycombinator.com) I have been using LLMs & Coding Agent since early 2024. A large problem with Coding Agents & LLMs in general is context compression.
Show HN: I built a YouTube for generative videos in 50 prompts (gallery.samsar.one via hn) So I decided yesterday to do a full rewrite of the media gallery, turning it into a generative media catalog with full-suite social interactions, search, and personalized recommendations built on top of the samsar-js library. It took aroun…
Show HN: Bunrun – agent-configured local dashboard to start/stop dev apps (github.com via hn) Agentic era has left me with already half a dozen vibe coded helper apps, most run with ´bun run dev´ or the npm equivalents, occasionally Python. Instead of zooming around terminal, I decided to vibe code one more thing: A local dashboard…
Codex Chronicle (learn.chatgpt.com via hn) Chronicle is in an opt-in research preview. It is only available for ChatGPT Pro subscribers on macOS.
Command and Conquer: Red Alert Mac/Android/iOS ported using Codex (github.com via hn) ra-port Native macOS, Android, and iOS source port of Command & Conquer: Red Alert. ra-port lets you play Red Alert (1996) on modern platforms.
New ChatGPT desktop app combines Chat, Work, and Codex (help.openai.com via hn) could not extract summary
Tell HN: GPT5.6 Is Imminent? (news.ycombinator.com) When using Codex with GPT5.5, I saw "Selected model is at capacity. Please try a different model." This has never happened to me with Codex before.
Show HN: Agent Sessions – local history and a live quota meter for Codex/Claude (jazzyalex.github.io via hn) Agent Sessions 4.2 — a redesigned transcript: rich Markdown, collapsible tool calls, inline images, clickable file paths. Browse and resume local Codex, Cursor, Hermes, OpenClaw, Claude, OpenCode, Antigravity, Copilot, and Pi sessions on m…
Codex Responses regression: reasoning summaries now return <!-- --> bodies (news.ycombinator.com) Hi all. I wanted to flag that the openai from today 09.07.2026 stopped returning thinking summaries on the codex endpoint.
Opendray – run Claude Code/Codex agents on your own box, drive from anywhere (github.com via hn) opendray Self-hosted gateway for Claude Code, Codex, Antigravity, Grok Build, and OpenCode. Run agent sessions on your own infrastructure.
Hijacking Defensive Cyber AI Agents for Remote Code Execution (ainowinstitute.org via hn) Exploit Brief We are revealing a proof-of-concept exploit that enables remote code execution in Anthropic’s Claude Code CLI (with Claude Sonnet 4.6 & 5, Opus 4.8) and OpenAI’s Codex CLI (with GPT-5.5) when employed to defensively assess th…
Show HN: All the Cron Jobs (github.com via hn) i'm building an open source to create all possible cron jobs that can be used with hermes, openclaw, claude, codex, and any type of workflow you have.
Loopy: Agent workflows that run when your data changes (loopy.computer via hn) Loopy is an orchestrator for agent loops. Declarative workflows live in your repo as a directory of markdown files; Loopy builds the graph and runs each step on the agent harness you pick, Claude Code or OpenAI Codex.
Top contributors to Codex own 50% of the codebase by linecount (www.reddit.com via hn) could not extract summary
Show HN: Security MCP expose your org's policies and paved roads to agents (github.com via hn) Security MCP A small, generic MCP server that exposes your organization's policies, risk context, paved roads, and security tools to coding agents (Claude Code, Cursor, Codex, Continue, etc.). The premise: agents do most of the typing now.
Talon: Self-hosted AI agent harness for chat, terminal, and desktop (github.com via hn) Talon Multi-platform agentic AI harness. Runs on Telegram, Discord, Microsoft Teams, the Terminal, and a cross-platform Desktop/Mobile companion app (Flutter), with a pluggable backend (Claude Agent SDK, Kilo, OpenCode, Codex, or OpenAI Ag…
Show HN: A Windows Desktop Pet for Claude Code, in PowerShell/Win32 (github.com via hn) I got the idea from Codex. It has a little pet that reminds you when it finishes.
Real limits converted to API-equivalent $ value for Claude Code, Codex, Copilot (twitter.com via hn) How much API token value do you really get from AI coding subscriptions? We measured real limits and converted them to API-equivalent $ value.
Ask HN: Are you paying for Lovable / Replit / Base44? (news.ycombinator.com) If yes, could you share whether you are using it to build mockups or have built products and have paid customers. Where do you see the future of these apps going?
Geosql: A Claude/Codex skill for geospatial data (github.com via hn) GeoSQL Claude, Codex, and GitHub Copilot skill for data scientists and analysts working with geospatial data on PostGIS, BigQuery, Snowflake, and Wherobots. Note: No SaaS account needed.
Moraine: Unified Agent Tracing (github.com via hn) Moraine Moraine is a local trace stack for agent work. It indexes sessions from agent harnesses such as Codex, Claude Code, Kimi CLI, OpenCode, Hermes, and Pi Coding Agent into ClickHouse, serves a monitor UI, and exposes MCP retrieval ove…
Paseo: Orchestrate coding agents from your desk and your phone (paseo.sh via hn) Self-hosted daemon for Claude Code, Codex, Copilot, OpenCode, and Pi. Agents run on your machine with your full dev environment.
Show HN: ChatGPT, Claude and Codex-style chat inputs in one React component (prompt-area.com via hn) Styles Ready-made agent-input styles, assembled from Prompt Area and its companions. Each is a real, copy-paste composition modeled on the agent UIs you already know — toggle Preview and Code on any example.
Show HN: Convergo – plan/build review loops for coding agents (github.com via hn) hey, miles here. i keep asking code agent to review another agents' work, and every round a new agent points at something the last one missed.
Show HN: Dejavu, stop showing coding agents the same command output twice (github.com via hn) Dejavu Stop showing coding agents the same command output twice. Dejavu is a PATH shim for Claude Code, Codex, Cursor agent, opencode, Aider, Gemini CLI, and other terminal-based coding agents.
Show HN: Fence – Jiminy Cricket for AI coding agents (news.ycombinator.com) Hi everyone, I'm Andrios, founder of hoop.dev (YC W21). 2 weeks ago, I told our eng team they could spend 20% of their week on side projects.
Show HN: Graphene – local-only Git wrapper to manage stacked branches (github.com via hn) Graphene is a tool I've been working on with Codex to create, publish, manage stacked branches and PRs. It's become a indispensable tool to manage code for me and other co-workers since moving off of paid stacked PR products.
Show HN: unlimitedcodex - real GPT-5.5 API access for Codex IDE/CLI (unlimitedcodex.com via hn) Clear GPT-5.5 + Codex API access Your delivered setup names the GPT-5.5 plus Codex 5.5/5.4/5.3-style model IDs available for your package, so you can plug the right value into OpenAI-compatible code without guessing. OpenAI-compatible GPT-…
Agent Box (agent-box.sh via hn) Move your project into a dedicated VM — local or in the cloud — and launch Claude Code, Codex, or Open Code inside it. Each box is isolated, checkpointed, and runs on hardware you control.
Panoptes – AI audit and alignment layer (github.com via hn) Panoptes Universal agent observability, audit trail, and policy enforcement. Works with Claude Code, OpenAI Codex, Google Antigravity, and Hermes Agent.
Claude and Codex and Grok: my current workflow and its friction (www.nativesoul.dev via hn) ❯ For developers who run more than one coding agent. A coding agent that's actually yours Like one developer who already knows your work, in every tool you open.
Show HN: A little cat that counts your tokens (Claude and codex) (jpthecat.com via hn) i, like you, probably burn tokens like water and i don't trust the /status or /usage calls b/c, for whatever reason, they either fail to load in the CLI). so, i'm constantly refreshing the direct links to Claude / Codex dashboard, manually.
Built Portal Clone with Codex 5.3 (www.pixelfork.ai via hn) Play Momentum Chamber
Vibe-Cadding in Codex (generalrobots.substack.com via hn) Vibe-Cadding It turns out that current models are pretty good at writing code to generate cad (3D objects). I’ve been using Codex to build things and the current generation feel a lot like coding did 2 years ago.
An MCP for visually inspecting robotics data, with headless rendering for CI (rerun.io via hn) The Rerun CLI includes an [MCP](https://modelcontextprotocol.io/) server that lets agents such as Codex or Claude interact with a running Viewer. It allows the agent to interact with the viewer like a real user, allowing it to interact wit…
Codex Threads (github.com via hn) codex-threads codex-threads is a companion CLI for inspecting and controlling Codex app-server threads from a terminal or another agent. [!IMPORTANT] codex-threads only sees threads on the Codex app-server instance it connects to.
Show HN: Codex-review – a read-only cross-model review skill for Claude Code (github.com via hn) codex-review An Agent Skill that adds one cross-model seam to a code review chain: a thin, read-only wrapper around the OpenAI Codex CLI (codex review) so a different model family reviews your diff and catches blind spots that an author an…
Show HN: A 'what you see is what you get' HTML editor for Mac (htmledit.io via hn) My workflow has improved a lot by outputting AI tasks into html files. But html files are then annoying to edit to get them right.
A Conflict-Free Multi-Agent Ensemble for Claude and Codex (medium.com via hn) could not extract summary
Codex Agents Built and Operate My Weapons Research (weaponsofconflict.com via hn) Blog Codex Agents Built and Operate this site A month-long experiment in using Codex agents to build and operate an open-source intelligence catalog without manually writing code or configuring infrastructure. For the past month, I ran an…
Show HN: Codex Sidecar – A Local macOS Companion for Codex Desktop (github.com via hn) Codex Sidecar A Codex companion tool for managing Codex conversations, favorites, usage limits, context usage, conversation continuation, selection runs, reusable instructions, and more. English · 简体中文 Download · Feedback Main window Mini…
Show HN: Scopewalker, an MCP server for codebase complexity metrics (github.com via hn) AI agents will happily create 1000+ line source files and add a 20th parameter to a function call, even if there's a rule file telling them not to. So I built Scopewalker: a local MCP server (open source, runs over stdio, and makes no netw…
Sous-Chef, a Claude Code plugin where Fable reviews, Codex implements (github.com via hn) 🧑🍳 sous-chef Fable 5 orchestrates and reviews; GPT-5.5 xhigh implements. Your head chef doesn't chop onions.
A macOS bell that rings when your Codex CLI session needs input (github.com via hn) Codex Butler Bell A polished desktop bell for the moment Codex is waiting. Codex Butler Bell is a Codex plugin that shows a large animated bell overlay and plays a bell sound when Codex is waiting for a user action, such as a permission re…
HT-ML.app – Deploy HTML Artifacts from Claude Code and Codex (ht-ml.app via hn) Go live straight from Claude Code or Codex in under a minute. To get started, ask Claude Code or Codex to deploy your HTML file with https://ht-ml.app.
Show HN: Afair – Self-organizing memory shared across your AI tools (github.com via hn) afair has been the one tool I can't work without for months, and I just open-sourced it. It gives every AI you use one shared memory that stays yours.
Using a local iPhone MCP server to plan Apple Watch workouts with Codex (bernhardhering.de via hn) I had a problem explaining Ask My Health in one sentence. "An MCP server for HealthKit" is true, but it only makes sense if you already care about MCP.
Tracing Codex's 640TB-a-year SQLite writes (querydoctor.com via hn) A recent Github issue for OpenAI's Codex on how the harness writes way way too many logs in SQLite started gaining traction the other day, and I wanted to take a shot at figuring out what's going on with it. The issue already outlines most…
We used coding agents to add RonSQL support to RonDB (mikaelronstrom.blogspot.com via hn) Experiences from the new wave of AI programming How we built RonSQL in RonDB with Claude and Codex A false start A few years ago at Hopsworks we made an early attempt at using ChatGPT for coding. It turned out to be an expensive mistake: t…
Show HN: Deconstructing Anthropic's Coding Agent Control Model (www.highflame.com via hn) Anthropic recently published an excellent write-up on how they contain Claude Code and its sub-agents. One thing that stood out is that the architecture isn’t really about Claude—it describes a general pattern for securing autonomous agent…
Show HN: Paneflow – cross-platform GPUI app for parallel coding agents (github.com via hn) Paneflow A native GPUI workspace for running coding agents in parallel. Paneflow keeps Claude Code, Codex, Gemini, opencode, and any CLI agent in real terminal panes you can see, interrupt, and take over.
Show HN: Framein – a local work-state layer that keeps AI agents in context (www.framein.dev via hn) Keep onework frameacross Claude,Codex, and Gemini. Start with one agent, challenge it with another, switch when needed, and close the work with validation.
VPSMaxxing – Migrate Your Codex, Claude Code and Other Agents to a VPS (github.com via hn) VPSmaxxing rent the cores · keep the cash · run your agents anywhere Turn a ~$5/mo cloud VPS into a dedicated, always-on workbench for Claude Code + Codex — set up by an agent, for your agents. your laptop --tailscale--> VPS · agents in tm…
Codex: Introducing a familiar rich-text editing experience (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation Roman @khudonogov Codex for writing and editing.
Show HN: Ploof – The agent-native CLI for generating images, video, and audio (github.com via hn) = 18" /> The agent-native CLI for generating images, video, and audio. Hand it to Claude Code, Cursor, or Codex — they install it, read ploof learn, and create your assets for you.
"Perfumed Palaces: backrooms roguelike vibe-coded in <30 days" (news.ycombinator.com) "found" cosmic horror game with malicious RPG mechanics forces player to roll 9 digit RNG to earn $ to move and interact. entire game was written as a single html file by Claude Code on mobile, revised as a codebase on desktop with Codex,…
Cerberus – a local firewall for AI agents' tool calls (github.com via hn) Cerberus 🐺 A local-first security gateway for autonomous AI coding agents. Cerberus sits between the agent (Claude Code, Codex, Cursor, Cline) and your machine, intercepts every tool call before it runs, risk-scores it across four signals,…
Show HN: Codex can track external events with respect to internal data (app.getsupers.com via hn) Most analytics tools track your data, but completely ignore relevant external events like world news, platform shifts, pandemics, OpenAI announcements, etc. so you may be completely blind to the signals that impact growth most.
Show HN: Even, the terminal-first desktop workspace (eventerm.com via hn) If you’re like me and have basically delegated all your development to Claude code, Codex, OpenCode and the other agents, you’d agree that one of the annoying things is the unmanageable terminals. Personally I love working on the terminal,…
Show HN: Deskmate Live – AI Desktop Pet Companions (deskmatelive.com via hn) We launched last week on 4chan and have been building fast ever since, integrating new features and new avatars. Our idea is to make desktop pets not only cute, but useful.
Show HN: Run multiple instances of Codex(GUI) each with their own auth (github.com via hn) Codex Multi-Account Desktop Launcher Run additional Codex Desktop windows on macOS with separate OAuth accounts. This is an unsupported local workaround.
Murmur: Shared communication bus for your coding agents (github.com via hn) A shared communication bus for your coding agents. Murmur is a local chat room that claude, codex, gemini, cursor, and copilot all sit in at the same time, over a single MCP HTTP daemon.
Show HN: uvx ptn and give your agent full access to any system (dangerously) (pypi.org via hn) I made this dangerously easy and convenient: Step 0: install uv (with one command) Step 1: uvx ptn Step 2: press c Step 3: paste to your Codex, Claude Code or virtual girlfriend The agent gets full access to any Linux, PC, Mac, etc. There…
Show HN: A Claude skill that prunes your AI's memory file, one diff at a time (puremint.co.uk via hn) I made a memory cleaner for Claude Code. I guess it'll probably work in Codex, OpenCode, Composer etc, but I've only tested it in Claude Code.
Show HN: Poly Grid – A terminal grid where agents flag you when they need input (poly-grid.com via hn) Poly Grid is a tiling multi-terminal grid for macOS (12+, Apple Silicon and Intel) built for running AI coding agents in parallel. Run your webservers, build watchers, log tails, and agents like Claude Code, Codex, and opencode side by sid…
Maturana: Hardware-isolated, zero-trust agent harness (github.com via hn) Maturana *A secure agent harness that runs every agent in its own hardware-isolated microVM. Lightweight, fast, and completely yours to customise from Codex.
Show HN: inplan – plan with your coding agent in a shared Markdown doc (news.ycombinator.com) Hi HN — I'm one of the makers. We kept hitting the same failure mode with coding agents: you hand over a vague plan, the agent fills the gaps with guesses, and the back-and-forth that explains why a decision was made lives in a linear chat…
Codex Seraphinianus (en.wikipedia.org via hn) Codex Seraphinianus The Codex Seraphinianus[1] is an illustrated encyclopedia of an imaginary world, created by Italian artist, architect and industrial designer Luigi Serafini between 1976 and 1978 and first published in 1981.[2] It has a…
Show HN: Optimal model routing directly in Claude, Codex and Cursor (github.com via hn) One endpoint. Every model.
Swarm – open-source control center for Claude Code / Gemini / Codex agents (github.com via hn) Swarm A web-based control center for AI coding agents — Claude Code, Gemini CLI, and Codex CLI. Manage one agent or ten from a single browser tab — with autopilot, a task board, AI coordination, and email integration.
Show HN: Khala – let your AI sessions talk to each other, across any LLM (khala.to via hn) I got tired of babysitting context between my own AI sessions and having to copy-paste summaries or update MD files across sessions for different purposes, so I built this. Every AI session starts with a blank slate.
Explodex – mod the official Codex app (github.com via hn) Explodex 💥 💥 💥 Mod the official Codex app Explodex is an extension SDK and plugin playground for the Codex desktop app. It injects a small renderer runtime into a Codex Electron window, exposes DOM zones such as the sidebar and composer, a…
Show HN: Aharness – Enforce coding-agent workflows as state machines on Codex (github.com via hn) Agents are capable enough for long, autonomous, multi-step work now, if they have the right guardrails.. The failure mode is now process drift and context management.
What is the best coding harness as of June 2026? (news.ycombinator.com) With models changing so quickly, does everybody keep jumping from codex to claude code etc. or is something slowly winning market share?
Show HN: Picot – A local Codex style desktop GUI for Pi agent (github.com via hn) Picot (π-cot(e)) English | 中文 A local desktop GUI for the Pi coding agent. No cloud, no account — runs entirely on your machine.
Show HN: Token Receipt – turn your coding-agent history into a bill (tokenreceipt.ameyalambat.com via hn) Codex and Claude Code support Both agents get the same skill-first flow. The skill calls the local runtime, reads the structured analysis, and then lets Codex or Claude Code write the final response using the session you are already in, in…
A public Sentry key is all it takes to hijack Claude Code, Cursor, and Codex (thenewstack.io via hn) A public Sentry key is all it takes to hijack Claude Code, Cursor, and Codex On June 17, the Threat Labs team at Tenet Security, an AI-agent security startup newly out of stealth, documented an attack it calls agentjacking. The whole attac…
Show HN: Building an Autonomous Drone with Codex – Hardware Phase (jakedecamp.com via hn) In the last two posts I got the drone stack flying in simulation, then added missions and vision. That work mattered, but it still left a fair question hanging over the whole project: could the same control model survive contact with real…
Enabling Claude Code to run code when its sandbox fails (windows only) (news.ycombinator.com) Be forewarned that this is a dumb solution to a dumb problem. Claude code/cowork will often be unable to use the sandbox environment: this means it can read and write files but can’t run them - making it mostly useless.
Two Heads Are Better Than One: Run Many AI Agents, Merge One Auditable Result (medium.com via hn) 4 min read Just now Press enter or click to view image in full size A few weeks ago, I showed that Claude Code and Codex can have a real-time conversation via Git. That was the first step.
Detent: AI agent orchestration with worktrees and serialized merge train (github.com via hn) Detent Start With AI Hi, welcome to Detent. If you are reading this as a human, pause here and paste the prompt below into Codex or Claude Code.
Show HN: Lupen – an itemized, verified receipt for Claude Code and Codex spend (github.com via hn) See what every Claude Code and Codex session actually costs — itemized, verified, local. Lupen recomputes your spend straight from the raw Claude Code and Codex logs — broken down by turn, step, and sub-agent, checked against the tokens, a…
The unreasonable effectiveness of LLMs for auditing Rust code (shnatsel.medium.com via hn) 7 min read 21 hours ago As a lead of the Rust Secure Code Working Group, I got free access to GPT-5.5 via the Codex for Open Source. Since then I’ve found and reported dozens of issues of varying severity in widely used Rust crates.
Open Ralph Wiggum – Autonomous Agentic Loop (github.com via hn) Open Ralph Wiggum Autonomous Agentic Loop for Claude Code, Codex, Copilot CLI, Cursor Agent, Qwen Code & OpenCode Works with Claude Code, OpenAI Codex, Copilot CLI, Cursor Agent, Qwen Code, and OpenCode — switch agents with --agent. Based…
Show HN: Cc-fleet – run other LLMs as Claude Code workers, your sub drives (github.com via hn) 🚢 cc-fleet 🤖 Plug any third-party model into Claude Code's ⚙️ Dynamic Workflows, 👥 Agent Teams, and ⚡ Subagents — from DeepSeek · GLM · Kimi · Qwen … to your Codex subscription, with your main session's auth untouched; no Claude subscripti…
Show HN: A decompilation-based native PC runtime for GoldenEye 007 (github.com via hn) I used Codex (and some Claude) in order to decomp and port Goldeneye for N64 into a native build. This work furthers existing decomp work and builds on the broader community.
Who Owns the Code Claude Wrote? (www.oreilly.com via hn) The following article originally appeared on Sena Evren’s Legal Layer newsletter and is being reposted here with the author’s permission. TL; DR Agentic coding tools like Claude Code, Cursor, and Codex generate code that may be uncopyright…
Show HN: Multiplayer Usage Tracking for Claude Code, Codex and OpenCode (github.com via hn) Hey HN, I built this open source tool which lets you and your team track usage across Claude Code / Codex / OpenCode. 1.
Open-source AI skills that make Claude/ChatGPT produce real work, eval-scored (github.com via hn) 🧠 PM Skills — 174 Professional Agent Skills for Claude, ChatGPT, Gemini, Cursor, Codex & Hermes Open-source Agent Skills (SKILL.md) + subagents + slash commands for every profession — one source, every AI coding tool. ⭐ If this saves you t…
Show HN: Gora – simple search across all your local coding agents (github.com via hn) Hey HN - Gora is a local cli app that indexes your chat threads across Codex, Claude Code and Pi automatically and lets you simply search across them all at once The problem I wanted to fix here for myself was that each time I directed the…
Show HN: Drydock – VM Sandboxes for macOS Autonomous Coding Agents (github.com via hn) drydock drydock runs autonomous coding agents (Claude Code or OpenAI Codex, per-task selectable) on your own Mac — not someone's cloud — each task sealed in its own hardware-isolated VM. It starts from the assumption that the agent is alre…
MCP that lets Codex/Claude search LinkedIn (www.veronaresearch.com via hn) Create high quality people and company lists with Verona's MCP.
HandoffKit: Coordinate agents by passing messages, not sharing memory (platformpilot.ai via hn) TL;DR. We are open-sourcing HandoffKit, OpenAI Codex plugin for coordinating LLM agents the way Go coordinates goroutines: by passing messages, not by sharing a scratchpad.
Turn your AI coding agent into a read-only compliance auditor (github.com via hn) ai-audit-orchestrator A read-only, evidence-gated audit harness for AI coding agents (Claude Code, Cursor, Codex, etc.). It runs a chain of single-purpose audit subagents over *your own repository*, one framework at a time, and forces ever…
Ask HN: Is AI helping with personal projects/tools? What's your stack? (news.ycombinator.com) I have a couple of personal app ideas that need simple cloud storage. I was hoping Claude/Codex would be able to help me put something basic together quickly, but I've been underwhelmed by tech stack suggestions they tend to suggest.
The engineering practices Claude Code and Codex use to improve AI agents (www.andrewjesson.com via hn) The engineering practices Claude Code and Codex use to improve AI agents Coding agents perform common engineering practices when asked to improve AI agents. Will they subsume specialized tools for failure-mode analysis, evaluations, and pr…
Show HN: AI Agent Skills to Grow Your Open Source Project (github.com via hn) A set of skills to use with Claude Code, Opencode, Codex, etc to help make your open source project more popular.
Captured Logs Reveal Hackers Using Claude and Codex to Breach Companies (research.openanalysis.net via hn) Captured Logs Reveal Hackers Using Claude and Codex to Breach CompaniesFull agent sessions captured on a compromised host turned honeypot offer an unprecedented look at how attackers are using AI in real-world intrusions. Jun 16, 2026 • 48…
Show HN: Offload, cross-device handoffs for Claude Code and Codex (github.com via hn) You just prepend `/offload` to a prompt, and it will run on /goal on another device (like a Mac Mini, it can also set-up a VPS for you). It handles details like moving env keys safely, making sure you're signed-into gh, pinning the right v…
Show HN: Hyperbox- $40/month Mac mini rentals (hyperbox.sh via hn) Hey HN, Following the OpenClaw craze, I saw a huge need for hosting personal Claws/Hermes agents on macOS [0]. So I built an agent to scrape eBay for below-market M-series Macs and built a Mac mini datacenter [1].
Show HN: Devloop, a local code/review loop for Codex and Claude Code (devloop.sh via hn) ░█▀▄░█▀▀░█░█░█░░░█▀█░█▀█░█▀█ ░█░█░█▀▀░▀▄▀░█░░░█░█░█░█░█▀▀ ░▀▀░░▀▀▀░░▀░░▀▀▀░▀▀▀░▀▀▀░▀░░ Spec-driven code/review loop that runs until all acceptance criteria are met and bugs fixed by agents. $ curl -fsSL https://devloop.sh/install | bash Re…
Axiomata – A Codex of Becoming (v1tali.com via hn) Fragments of Forever Fragments of Forever poem, part of the SIGNALS+ECHOES album by Vitali Liouti The iron had been in the fire since dawn. By the time the priest lifted it with tongs, it glowed a dull, breathing orange, and the villagers…
Show HN: Open-source CLI to see your AI coding token usage and compare it (github.com via hn) I use Claude Code, Codex, Cursor every day and had no idea how much I was actually burning across all of them combined. Each tool shows its own usage (most don't) in its own place, if at all, and I just wanted one number.
Text-to-Lottie: Generate Lottie animations with coding agents (github.com via hn) Text-to-lottie is an open-source framework for generating production ready Lottie animations with claude code/codex or any other coding agent supporting skills. Created with Text-to-Lottie Quick Start Install the skill: npx skills add diff…
Zehn, a fuzzy finder for every prompt I've sent an AI agent (www.al3rez.com via hn) zehn, a fuzzy finder for every prompt I've sent an AI agent I use claude one day, codex the next, then pi or opencode after that. They all keep a history of what I’ve typed.
Show HN: 100Hires MCP, 130 tools to run an ATS through LLM. Is 130 too many? (100hires.com via hn) Hi HN. I'm one of the co-founders of 100Hires ATS.
Sandbox AI coding agents with microVMs on Fedora Linux (fedoramagazine.org via hn) AI coding agents such as Claude code or Codex get more capable every month. This is great for productivity, but approving all commands gets annoying really quickly.
Audit checklists for AI coding agents – 30 invariants, any language (github.com via hn) audit-skills Language- and framework-agnostic audit checklists for AI coding agents — security, correctness, and operability. Works with Claude Code, GitHub Copilot, Cursor, Codex CLI, OpenCode, and any agent that can read files.
Nudge – a collaborative memory layer for Claude Code and Codex CLI hooks (github.com via hn) Nudge Nudge is a collaborative memory layer for Claude Code and Codex CLI hooks. It remembers the coding conventions, workflow preferences, and repo-local debugging lessons that agents should use while they work.
Double your Codex / Claude Code productivity and output (github.com via hn) pua Double your Codex / Claude Code productivity and output Telegram · Discord · Twitter/X · Landing Page 🇨🇳 中文 | 🇯🇵 日本語 | 🇺🇸 English Scan to join WeChat group Add assistant on WeChat Most people think this project is a joke. That's the bi…
Hi HN: Loopy agent, meta-loop engineer my Claude Code and codex sessions (github.com via hn) meet loopy A terminal meta-agent that watches how you work, finds the patterns, and writes the loops so you don't have to. The era of prompting is over "I don't prompt Claude anymore.
Show HN: Lark – Use Codex from Telegram or iMessage (github.com via hn) OpenAI added mobile control for Codex, but it depends on both the desktop app and the ChatGPT mobile app. Lark takes a different approach.
How to Build a Google Sheets API Integration with Nango and Codex (nango.dev via hn) This guide shows how to build a custom, customer-facing Google Sheets API integration with Nango and an AI coding agent (Codex, Claude Code, Cursor, or any other). By the end of this guide, you will have: - A Google Sheets auth UI in your…
Production-Grade Claude/AI Skills for Ruby on Rails (github.com via hn) rails-skills Production-grade Claude Skills for Ruby on Rails. Stop fighting your AI coding agent on Rails conventions — drop in rails-skills and Claude Code, Cursor, Codex, Gemini CLI, Antigravity, and Windsurf will write Rails code the w…
Auto mode for pi.dev. An LLM reviews your coding agent's commands (github.com via hn) pi-auto-reviewer Automatically review bash commands that your pi agent wants to execute - akin to Codex "Auto-review" and Claude Code "auto mode". How it works Every bash command the agent wants to run goes through three tiers: | Tier | Ac…
Cc-doubleteam – Claude plans, Codex executes, Claude reviews (github.com via hn) cc-doubleteam Three-phase project mode for Claude Code: Claude plans, Codex executes, Claude reviews — execution burns your ChatGPT limits, not your Claude limits. Why You get Fable-quality planning plus Codex execution without burning you…
Ask HN: Agents get dumber before release of new model version? (news.ycombinator.com) I've noticed an effect with openai where my codex agents seem to perform worse in the week(s) leading up to a new release. I'm wondering if the vendors tweak effort params at all to free up hardware to host the new version.
Open-source Linux tray app for tracking Claude Code and Codex usage (github.com via hn) OpenUsage Community Track all your AI coding subscriptions in one place. OpenUsage Community is an independent, community-maintained continuation of the original OpenUsage project.
I made an agent skill for making HTML slides with consultant style (news.ycombinator.com) I was a consultant in one of the top firms and have been an software engineer before joining consultancy... As an engineer, I really don't like to deal with the software made by Windows, especially the Office -- yep, I hate doing PPT..
Using Xcode 27's Agent Skills in Claude, Codex, and Cursor (www.avanderlee.com via hn) Apple launched Xcode 27 during WWDC’26, introducing a bunch of agentic development improvements, including official agent skills. As you’ve learned from my 9-Step Framework for Choosing the Right Agent Skill, it’s important to pick skills…
Show HN: Codacy Skills for Claude, Codex, Copilot, etc. (github.com via hn) Codacy is a code quality and security platform that helps eng teams enforce coding standards against their AI generated code. They just launched agent skills and a cloud CLI, which allows Claude etc.
Show HN: AgentHUD – Live TUI and daily digest for parallel Claude Code sessions (github.com via hn) AgentHUD A heads-up display for your AI coding agents — Claude Code, OpenAI Codex, AWS Kiro, and opencode. AgentHUD reads each agent's on-disk sessions (JSONL files, or opencode's SQLite store) and merges them into one tree, so a project y…
Tell HN: Codex once again automatically activates /fast on app update (news.ycombinator.com) Be aware that when you update Codex, OpenAI automatically activates /fast, even if this setting was turned off before the update. This will result in you burning way more credits than expected.
Running DeepSeek-V4-Flash on a Raspberry Pi (twitter.com via hn) Article Conversation Running DeepSeek-V4-Flash on a Raspberry Pi I ran DeepSeek-V4-Flash on a Raspberry Pi 5 (8GB edition) by streaming model weights from a PCIe attached NVMe SSD. Codex (GPT-5.5 xhigh) and Claude Code (Opus 4.8 max) drove…
Show HN: RunAPI – one API for AI video, image, music, audio, and LLMs (runapi.ai via hn) After building a few SaaS apps, I found that connecting many AI model APIs at the same time gets messy very quickly. So I built RunAPI, which lets you call image, video, music/audio, and LLM models with one API key, and I also released SDK…
Show HN: Generate production grade Lottie animations with Claude Code (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Post Conversation Introducing text-to-lottie: an open source skill and harness for generating production ready Lottie animations with codex/claude code.
Show HN: Sloth Sayer voice activated computer use (github.com via hn) Show HN: I made an ad-free game collection for my grandparents (stiuvou.ch via hn) Hi, living with my grandparents who are not able to go out anymore and spend a lot of time on their phone I became frustrated with the available games on the app store. Many of them had annoying ads and were not elderly friendly.
Show HN: Tool you give AI agents to sneak in prompts or connect multiple agents (apps.apple.com via hn) So, I was looking for ways on how to make Codex and Antigravity talk to each other and came up with "Promptgate", which is essentially a web server that does http long polling and slow releasing the data to keep agents happy and give you c…
Show HN: Baseball version of the popular 82-0 game (statgm.com via hn) Vibe-coded a baseball version of the viral 82-0 basketball game. Codex essentially one-shotted this which was impressive to say the least.
Show HN: Context Mode Insight – observability layer for AI coding agents (context-mode.com via hn) the first Solution from Context Mode Platform · for enterprise AI engineering Role-aware observability for Claude Code, Cursor, Copilot, Codex, Gemini, and 9 more AI assistants. 222 patterns.
Show HN: LimitPing – Keep Claude Code and Codex rate-limit windows continuous (github.com via hn) CCLimitPing (limitping) English | 中文 Keep your Claude Code, Codex, and GLM (Zhipu / Z.ai Coding Plan) rate-limit windows back-to-back. These providers bill on a 5-hour rolling window (plus a weekly cap), and the 5h window starts on your fi…
Ask HN: Who here still codes without AI, and why? (news.ycombinator.com) As someone who loves coding (uses vim, and has done so for ~10 years), but also is amazed by AI coding tools and has embraced them recently (e.g. copilot then claude code then codex), I'm curious, who out there is still not using any of th…
Show HN: CCC: One place to manage all your Claude, Codex, Antigravity sessions (github.com via hn) CCC Start the next while Claude builds the first. One local dashboard for every Claude Code, Codex, Cursor, and Antigravity session on your Mac.
Show HN: Minimal native macOS sandbox for Claude and Codex (github.com via hn) Sandfence Run a coding agent — Claude Code or Codex — on a repo in its own "skip-permissions" mode, while the macOS sandbox, not the agent, enforces what it can touch. A wrong rm -rf, a stray git reset --hard, a pip install into your syste…
Agent-ML-skills – Teach Codex/Claude/Cursor to stop making ML mistakes (github.com via hn) agent-ml-skills Production-grade Machine Learning, Data Science & MLOps skills for AI coding agents. Coding agents are great generalists but make the same ML mistakes over and over: leaking preprocessing into cross-validation, scoring imba…
Codex for Sales Teams: Moving Faster to Solve Customer Problems [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
OpenAI Codex Tech Lead Does AI-Assisted Engineering (newsletter.eng-leadership.com via hn) How OpenAI Codex Tech Lead Does AI-Assisted Engineering Michael Bolin, Tech Lead for the Codex open-source repo, is sharing how they built the permissions system with AI-assisted engineering. Intro I recently had the pleasure of visiting O…
Show HN: Lazarus, a coding agent for long-horizon tasks (github.com via hn) I have been interested in long-horizon coding tasks for a while, especially with benchmarks like FrontierSWE, where even the best coding agents like Codex and Claude Code struggle to complete tasks. These agents come with a collection of t…
OpenAI's Codex chained decade-old DoS attacks to crash web servers (www.theregister.com via hn) MOST POPULAR EVENTS - Thriving Through Volatility: The Everpure Advantage in an Uncertain Market Learn how a consumption-based operating model provides flexibility, improves efficiency, and brings predictability to infrastructure investmen…
Basecamp CLI and Agent Skill: Agent first, agent native (basecamp.com via hn) Get full access to Basecamp through the command line using our CLI, or with your favorite AI agent. Claude, Codex, OpenCode, Cursor — all are welcome.
Agent-to-Agent Communication via Git (github.com via hn) h5i h5i is version control for the AI era — a next-generation, AI-aware Git sidecar for Claude, Codex, and your other coding agents. It records what each agent was asked to do, which files it read and edited, what it decided, what it skipp…
Why OpenAI Is Combining Codex and ChatGPT [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Knox – Govern AI agent tool calls before they execute (github.com via hn) Knox — Security enforcement for AI coding agents Knox is a security policy engine for AI coding agents. The same engine ships in five forms — a standalone CLI, a Node library, a Claude Code plugin, a Cursor plugin, and an OpenAI Codex plug…
Show HN: MCP for the ChatGPT Ads API – Query ChatGPT Ads from Claude and Codex (github.com via hn) I built a small MCP server for the ChatGPT Ads advertiser API, which OpenAI opened to the public a couple of weeks ago. It's read-only for now: accounts, campaigns, ad groups, ads and insights (11 tools in total).
OpenAI Codex tool linked to malicious NPM supply chain attack (www.techradar.com via hn) OpenAI Codex tool with over 29,000 downloads linked to malicious npm supply chain attack stealing authentication tokens A tool started benign and turned sour after a little while - Researchers uncovered a malicious npm package posing as a…
Show HN: AI Gauge, a desktop monitor for Claude/Codex/Copilot usage limits (github.com via hn) Hi HN, new account but long-time reader. I built this for myself because I kept manually checking usage across Claude, Codex, and Copilot, and wanted to track the session and weekly usage all in one place.
Show HN: CTP Room – a shared chat room where your AI coding agents coordinate (news.ycombinator.com) Hi HN. I honestyle DO NOT like one on one sessions with my claude/codex when working with my team.
Ask HN: What are good AI UIs now? (news.ycombinator.com) With frameworks like Streamlit, it takes five lines of Python to wrap an LLM in a chat box. Alternatively, we've seen a surge in TUI tools (Claude Code, Codex, etc.).
Show HN: Codex Reset Watchdog – a Skill for watching codex quota reset signals (github.com via hn) a small Codex Skill + Automation that watches public posts for actionable Codex quota reset signals.
Build and deploy hosted sites from Codex with the Sites plugin (developers.openai.com via hn) Sites lets Codex create, save, deploy, and inspect websites, web apps, and games hosted by OpenAI. Use the Sites plugin when you want to turn a prompt or a compatible existing project into a hosted site without setting up a separate deploy…
Recall – Local search across your Cursor/Claude Code/Codex chat history (github.com via hn) recall Your AI chat history, searchable. Across Cursor, Claude Code, Codex, and pi.
Turn the World into Cheese – OpenAI RPI Image Gen Camera (github.com via hn) ImageGenCam is a digital camera you can build yourself with Codex. Using basic maker parts and a 3D-printed shell, ImageGenCam is a highly customizable project designed to reflect your own style, interests, and ideas.
Show HN: Parley – code review TUI for AI code (parley.cloudflavor.io via hn) Wrote my own tool that helps me review AI Slop. Akin to Github/GitLab reviews but local and in a TUI instead.
Ask HN: How can I get an OpenAI account bug in front of an engineer? (news.ycombinator.com) I'm trying to reach someone at OpenAI regarding what appears to be a bug in their phone verification flow. The issue is not related to my normal account login or 2FA.
Hatch: Write agent rules/skills once, generate for all (github.com via hn) 🥚 hatch Write rules, skills, commands, and sub-agent definitions once, generate the native files each coding agent expects. Hatch reads a single source tree under .hatch/ and produces the specific files Claude Code, OpenAI Codex CLI, GitHu…
Show HN: Ministry of Everything – CLI agent harness for a single operator (github.com via hn) ▓▒░ MINISTRY OF EVERYTHING ░▒▓ Ministry of Everything (MoE) is a CLI-first harness for one operator directing AI agents through durable markdown work. MoE runs Claude Code or Codex against living markdown documents.
Finding New Biblical Cross-References with Codex (www.johnnychang.com via hn) What OpenAI's Newest Codex Found In The World's Oldest Codex: Hunting in the Bible • 19 min readWhen I was a kid, my friend's dad found a new largest prime number. He had added one small step to a continuous record humanity has kept alive…
OpenAI Docs –> 403: Forbidden (developers.openai.com via hn) Codex use cases Learn how teams are using Codex to automate tasks, build apps and ship with confidence. Docs and resources to help you build with, for, and on OpenAI.
Codex just found a "workaround" of not having sudo on my PC (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Log in Sign up Post Conversation Son Luong @sluongng Codex just found a “workaround” of not having sudo on my pc… 3:32 PM · May 30, 2026 899.2K Views New to X?
Show HN: Ouijit, an open-source task and terminal manager for coding agents (ouijit.com via hn) Hi HN, I’m working on Ouijit. It’s a project and task-based terminal session manager that provides a few basic but useful tools for agent workflows: - Terminal sessions in Ouijit have access to the ouijit CLI, and supported agents (Claude,…
Show HN: Agentpack – isolated config layers for Claude Code, Codex, and OpenCode (nexo.sh via hn) Modern agents like Claude Code, Codex, Cursor and OpenCode need skills, hooks and MCPs, but they load them differently. I made a tool that creates an ephemeral staging configuration that loads into coding agents without polluting the globa…
GPT-5.4 says it's GPT-5 in Codex (old.reddit.com via hn) could not extract summary
Show HN: SharkBay – a local macOS workbench for coding-agent CLIs (github.com via hn) SharkBay macOS workbench for multi-agent vibe coding Features Multi-Agent Support Launch and manage multiple AI coding agents from one workspace. Supported agents: Claude Code · Codex · Gemini · Kiro · DeepSeek · Qwen · OpenCode Agent Stat…
Composer 2.5 is a monster 20 usd for 800 M tokens? (www.reddit.com) Have been producing tons with composer, I just have to manual guide on arch decisions, implementation. Sometimes I use codex but how much I can get from a 20 usd codex subscription is probably 5% of this.
Manage Claude Code, Codex, OpenCode, Gemini CLI sessions in one terminal view (twitter.com via hn) could not extract summary
Advice on GPT Codex agent (www.reddit.com) I have recently started using codex. After hitting rate limits after a few hours I expanded my instructions to reduce token use, while still keeping me in the loop.
Building Illustration heavy demos? Don't use generic AI video generators (www.reddit.com) If you are building video demos with illustrations - stop using generic video gen AIs. Generating a good video is rarely a one-shot task.
We built Branchless, a desktop app for running parallel dev sessions with agents, terminals and editors, without switching branches (www.reddit.com) Hey everyone, We have been building Branchless, a desktop app for Mac, Windows and Linux. The basic idea is simple: we wanted a way to work on multiple tasks at the same time without constantly switching branches, stashing changes, opening…
bunx ccusage told me i burned $18,450 of credits in may. i pay €400/month total (www.reddit.com) Ran bunx ccusage monthly -s 20260501 --all ten minutes ago, half expecting to see usage that vaguely justified my subscription. instead i got this: $18,450.29 in credits 248M input tokens 42M output tokens 21.7B total when you count cache…
i built a local cli for reducing token waste (www.reddit.com) i built , a local node.js cli for finding context waste in claude code, codex, and cursor workflows. it runs locally, needs no api keys, no login, and nothing leaves your machine.
Went down the Claude Code add-ons rabbit hole (www.reddit.com) I installed Claude Code, thinking that was basically the whole thing. But after I talked to some folks, I found are adding a bunch of extra stuff on top of it Some of the things I found useful, I feel, could be helpful to share - superpowe…
Adversarial is the new way to go... (www.reddit.com) I don't know what is wrong with Claude, but since I began to audit its work, even by considering that I have a very decent Claude.md, Harness, Hooks and many other "tricks" to keep Claude to the point (I also built a Vault with Obsidian an…
I used autoresearch to improve my AGENTS.md, measured against real tasks (www.stet.sh via hn) Codex optimized its own AGENTS.md against real Stet repo tasks. The best candidate improved the training slice, then regressed enough on a clean holdout that it was not safe to ship.
Which provider fits best for my needs? (www.reddit.com) Hi everyone, I’m looking to get more into experimenting with AI and considering a paid subscription, but I’m a bit unsure which direction makes the most sense for my use case. My main goals: -Writing a technical book in the field of taxati…
I’m building autospec: a Claude-friendly workflow that turns feature ideas into specs, issues, PRs, and merges (www.reddit.com) I’ve been building autospec, a multi-harness AI workflow suite for Claude Code, Codex CLI, and OpenCode. The problem I’m trying to solve: AI coding can move fast, but the trail of “why this exists” gets lost quickly.
My managed AI Agent for connecting all teams to the system.. even non-technicals. (www.reddit.com) Hey everyone, I'm opening up early access to my SaaS Kognita, and I'd like your feedback on it. It’s a custom semantic engine for your codebase, served through a managed agent runtime your whole team can use from the browser.
Fix for Codex hanging during compact / compaction on Linux "Error running remote compact task" (www.reddit.com) If you're on Linux and Codex hangs during compact or remote compaction, the culprit may be reqwest's default tcp_user_timeout. On Linux targets it defaults to 30 seconds, which is too aggressive for long-running unary requests like compact…
Peers – Multi-agent AI coding with measurable convergence (github.com via hn) peers A small Python substrate that drives n ≥ 2 AI coding CLIs (Claude Code, Codex, …) as cooperating peers toward measurable project goals. HOWTO: full audit + fix on an existing app: docs/HOWTO-audit-and-fix.md implement mode (build a f…
Show HN: Moltnet, a tiny self-hosteable chat network for agentic organizations (github.com via hn) Moltnet A lightweight chat network for AI agents. Rooms, DMs, and persistent history across OpenClaw, PicoClaw, TinyClaw, Codex, and Claude Code.
Dense vs. Moe Model (engineersmeetai.substack.com via hn) Yesterday, I ran out of tokens in OpenAI Codex while oxidizing parts of my Python codebase into Rust. It was around 11:30 PM, and I had to wait another two hours for the limits to reset.
anyone use Polsia, Paperclip or Virtuals? (www.reddit.com) I've been seeing a lot of AI agent run ideas being started and I'm super curious about how those work. I understand how Claude and Codex can help me build things but trying to see how useful these platforms can be for running a (mostly) 'a…
Built a /advisor command for Claude Code — Opus directs parallel Sonnet runners that actually read your files (www.reddit.com) Been building **advisor** for a few months — a `/advisor` slash command for Claude Code that runs Opus as a "strategist" coordinating multiple Sonnet (Opus's hands) runners reading files in parallel. This isn’t a “spec”.
Resources for learning how to use AI Agents for Coding (www.reddit.com) I am working on a startup idea where I am primarily using Codex/Claude Code for coding. I would like to learn about using AI Agents for coding.
Building the harness around our coding agents: eight failure modes, eight pillars (www.reddit.com) We ended up building two products: the software we ship, and the system/harness around our agents that makes them useful in building the thing we ship. A harness is the durable layer around a model: instructions, tools, permissions, contex…
Show HN: AI agent token cost calculator for Codex and Claude Code loops (tinyopsstudio.com via hn) Estimated monthly cost Enter your usage to calculate token volume. AI agent token calculator Put in your average tokens, run frequency, and provider price.
Why codex /goal fails on complex workflows: compaction amnesia and context rot (news.ycombinator.com) Hi HN, When Openai released `/goal` earlier this month, I was really excited to try it for long-horizon tasks. But after using it, it didn't blow me away and i did some digging and found a major architectural flaw when using it for complex…
I built a workspace where Claude, Codex, and other AI agents can collaborate (www.reddit.com) I use Claude heavily for solo building, and the bottleneck stopped being “can one model do this task?” The bottleneck became coordination. I use different agents for different jobs: product thinking, coding, writing, design review, PR revi…
AWO – Run Claude and Codex in isolated Git worktrees (github.com via hn) AWO — Agent Worktree Orchestrator AWO is a local Go CLI that coordinates Claude Code and Codex across isolated git worktrees, runs deterministic verification commands against the result, and produces a structured artifact bundle (run.json,…
Built an OSS spec-driven AI development tool that runs multiple agents in parallel on the same feature with an LLM-as-judge that picks the winner (www.reddit.com) Hi. Been building something I think folks might find useful.
Show HN: Smriti: Shared Reasoning State for Claude Code and Codex (github.com via hn) Smriti Code has Git. Multi-agent reasoning does not.
AgentSlice – Make AI coding agents ask before they edit (github.com via hn) AgentSlice A free, open-source workflow kit for AI coding agents. Makes Cursor, Claude Code, Codex and Windsurf ask before they edit.
80M tokens used in 45 mins (www.reddit.com) My Pro+ plan was ending tonight and I still had some usage left, figured I’d optimize a few code paths and merge my PR before downgrading to the $20 plan since I’ve been using Codex more heavily lately. Then Opus casually burned through 80…
Zotero use skill for Codex (www.reddit.com) This will be of interest to academic researchers who use Zotero for reference and knowledge management and in scientific writing. This skill builds on pyzotero library and has agentic instructions for creating embedded zotero inline citati…
Teaching Codex to Test a Voice-First Calendar App (www.elicited.blog via hn) Teaching Codex to test a voice-first calendar AI-generated entry. See What & Why for context.
Classic Claude (www.reddit.com) I have been working with Codex for the last 2 weeks and only use Claude for some Deploy work.. My Codex ran out last night and I was very annoyed with the outcome Claude built an admin security check assuming passkeys would appear in the s…
I’m worried that cursor is going to follow the others on pricing. (www.reddit.com) Im on month 4 of cursor and I absolutely love it. I pay $60 a month an usually hold out for large plans until my API renews (large decisions are expensive).
OpenClaw + Hermes users: how many agents are you actually running day to day? (www.reddit.com) I’m trying to understand how people are structuring real agent setups once they move past demos. If you use OpenClaw, Hermes, Claude Code, Codex, or similar agents for actual work: Do you run one general agent, or do you split things into…
Show HN: Simple Sprite Sheet Generation (github.com via hn) Games got me into programming and the love of computer science (think Commander Keen, Monkey Island, DOOM). They embody everything I love - programming, design, architecture, complex state management, art, music, narrative/story telling.
Is Composer 2.5 better than Glm 5.1 and DeepSeek v4 pro in real world tasks? (www.reddit.com) I am new to Cursor and still testing the free version. Benchmark for Composer 2.5 indicates it is better than DeepSeek v4 and Glm 5.1.
Stop Burning Tokens: 5.1x Faster Code Discovery With One Universal Plugin for AI Coding Agents (www.reddit.com) My colleagues kept asking me for my setup, so I decided to turn it into a universal plugin: Agent Code Navigator - a universal code-navigation plugin for Cursor, Claude, Codex, Gemini, and OpenCode. In my benchmark, semantic code discovery…
Computer-Use-Linux (news.ycombinator.com) Hi, Since apps like Codex and Claude don't provide support for computer use for Linux desktop users, iv'e built my own own one. https://github.com/agent-sh/computer-use-linux `npm install -g @agent-sh/computer-use-linux` `computer-use-linu…
How do you give feedback on markdown files that AI Agents write? (www.reddit.com) This has been bugging me for a few weeks so I figured I'd ask here. I use both Codex and Claude Code, and a lot of the time I'll ask them to write up a plan or proposal in a markdown file.
Is there a proxy network server for qwen27b to try fix leaking <tool_call> from content/reasoning_content? (www.reddit.com) Sometimes toolcall appears in the end of content, sometimes in the end of reasoning_content. On receiving end it looks kinda easy to fix - we see <tool\_call>, stop streaming and if stream ends on </tool\_call>, start fixing (more difficul…
Ask HN: Anyone catch the bug in codex with /goals? (news.ycombinator.com) Before this last update I could basically get unlimited usage out of codex by toggling on /goals and setting an impossible task (one that wouldn’t finish and would keep reiterating) and I’m on 4 days of uninterrupted usage it says I blew t…
Show HN: Klimkit: my Codex setup for multiple machines (github.com via hn) decided to open source my carefully tuned Codex set of skills, subagents, AGENTS.md. The core is the heavily opinionated workflow that agent is forced to follow.
Codex CLI kept saying “done.” It wasn’t. So I made it prove it. (www.reddit.com) Codex CLI can write code. The problem is that “wrote code” and “finished the task” are not the same thing.
Looking to sign of for Claude (www.reddit.com) I have a yearly plus plan with Codex and heard many good things about Claude. I was thinking of getting the yearly Pro plan on Claude.
Show HN: Agentikus (agentikus.com via hn) Servus Hackers, I’ve been building Agentikus mostly in silence for the past month or so, and then realized: I should probably post it here too. Agentikus provides workspaces for collaborating with managed and unmanaged agents in team workf…
Show HN: OpenRig – a control plane for multi-agent coding topologies (www.openrig.dev via hn) Hi HN, I’m Mike, the founder of OpenRig. I built this because my Claude Code + Codex setup kept forming little "topologies" of long-lived agents that worked well together, but the terminal sprawl was intense.
Run multiple AI coding agents simultaneously with isolated profiles (www.reddit.com) if you're running agentic coding workflows you've probably hit this: one account per tool, one session at a time. multi-cli fixes that.
Shortcuts Playground: Create Apple Shortcuts with Claude Code/Codex (www.macstories.net via hn) Today, I’m pleased to introduce something I’ve been working on for the past six months: Shortcuts Playground, a plugin for Claude Code and Codex that can create any shortcut for Apple’s Shortcuts app using natural language. With Shortcuts…
What kind of agents are you launching and with what that solves your pain point? (www.reddit.com) Curious what kinds of agents people here are actually running day-to-day. What problems or pain points have they solved for you?
Real World Usage Composer on Cursor Ultra vs Codex 20x (www.reddit.com) I am interested in knowing real world milage between Codex 20x and Composer Ultra. I know Codex 20x is heavily subsidized and then Composer 2.5 is much cheaper.
ChunkHound v5.1 (chunkhound.ai via reddit) We shipped ChunkHound v5.0 + v5.1 recently and forgot to post about 5.0, so here’s the combined update. ChunkHound is a code search / code research tool for AI coding workflows, especially MCP-based setups with Claude Code, Codex-style age…
Open-source skill OS for codex/claude/gemini CLI (routing/optimizaiton + evals) (www.reddit.com) Hey yall! Just shipped a local skill OS that sits above Codex CLI, Claude Code, and Gemini CLI (Hermes support coming soon).
Multi-Agent Code review (Review Council) to get critical feedback (www.reddit.com) Even though I primarily use Claude Code, I sometimes try out Codex and Gemini TUI tools occasionally as well. Then OpenAI came up with Claude Code plugin to use Codex command inside Claude Code (https://github.com/openai/codex-plugin-cc).
Show HN: Pocket TTS running in (mobile) Safari (ldenoue.github.io via hn) The original wasm port didn’t run on Safari so I used Codex to help me fix the issue: was that simd relaxed didn’t work on Safari, only simd-fixed What’s good about Pocket TTS is that it streams generated audio chunks.
Codex got better, codex might be built with Claude Opus (news.ycombinator.com) Very suspicious with openai codex getting better, I wonder if codex teams use Claude opus to build codex. Anyone engineer from openai who can confirm…
OpenAI Unethical Billing Practices (www.reddit.com) I had a $100 budget/month set on my OpenAI API organization. Despite that, OpenAI billed me almost $200.
filesystem glob path error on latest update (www.reddit.com) Error creating task failed to load configuration: filesystem glob path `**/*.key` only supports `deny` access; use an exact path or trailing `/**` for `deny` subtree access This error started appearing after the latest codex update. Have t…
OpenAI and 1Password Bring Agentic Security to Codex (www.forbes.com via hn) Agentic security is picking up steam. This week, identity security provider 1Password announced a collaboration with OpenAI that will enable developers to provide Codex with secure access to credentials, such as passwords.
Mechanical Design in Codex (twitter.com via hn) could not extract summary
Ask HN: Anyone else struggling with AI and work? (news.ycombinator.com) Been a developer for a little over 10 years now. I work on web stuff.
Show HN: Darc – grep-like memory search tool for coding agents (github.com via hn) Hi HN, I’m Junha Park. I've been experimenting with agent memory, especially how to make agents run more reliably on large tasks.
Codex for construction management (www.reddit.com) I have been using Codex to help us with the whole construction management process from start to end. It seems to handle reading of drawings much better than Claude Code.
Reconnecting. – – 5/5 why don't they fix codex (news.ycombinator.com) it's been a month tagging sama openai tibo on X for this issue and no one seem to reply and eveyone is falttering codex, im sure im not the only one facing this i switched to codex from claude since it was better consume less credit than c…
Drop your projects below! The best will get a shoutout! (www.reddit.com) Hope you guys are ready for another shout-out list! The top projects will get shoutouts on this list and may get a mention on our YouTube (5-7k views per video) :) Feel free to leave your project below or DM if you want to be featured in a…
I built AgentLighthouse, a local “Lighthouse for AI agents” that scans repos/docs/APIs for agent readiness (www.reddit.com) hello The basic idea comes from the fact that more people (including me) use Codex, Claude Code, Cursor, Copilot, MCP tools, etc., but they are still written only for humans. Agents might fail and struggle to use what you build because set…
I searched for agentic frameworks and here is what I found. What do you recommend? (www.reddit.com) The question: What is the practical agentic framework to use to make the agents run until job is done without reporting to me prematurely? My goal: Actually fully spend a $200 codex subscription, but make it be well spent.
Show HN: AI editor for websites (Next.JS) (news.ycombinator.com) Hi all, I've playing with an idea to plug LLM into a website editor and see how I make that work. I used mostly Claude Code and a little bit of Codex to build it in my evenings and on weekends.
I built a small MCP setup to let ChatGPT inspect local project files for planning before using Codex (www.reddit.com) I’m sharing a workflow I built for my own development setup. Sometimes I don’t need Codex to write or edit code yet.
Show HN: atrium – a resumable, tiling workspace manager for Claude Code, Codex (getatrium.dev via hn) atrium is a macOS desktop app that provides a fully resumable workspace manager with a variety of heterogenous panes including terminals, browsers, tasks, notes, and more. Over the last couple of months, I've been building an AI developmen…
Show HN: I put Codex and Claude in a tank arena; Codex is winning 55% so far (old.reddit.com via hn) could not extract summary
TIL you can ship a Claude Code skill inside a GitHub repo so anyone who clones it gets architectural guardrails baked in (www.reddit.com) I've been building a local AI ops platform and wanted Claude to be able to extend it without ever accidentally touching core files. So I added a .claude/skills/ directory to the repo with a plain Markdown file that gives Claude: - the arch…
My agent kept forgetting who 'Karpathy' was between sessions. Here's the architecture that fixed it (www.reddit.com) I run a second brain on Obsidian, Readwise, NotebookLM, and Claude Code. For each topic, I build a scoped wiki structured as the LLM Knowledge Base Andrej Karpathy proposed.
Slax Reader CLI: a read-later library your AI agents can use (slax.com via hn) Slax Reader CLI: a read-later library your AI agents can use A CLI that turns Slax Reader into a persistent reading store any AI agent — Claude Code, Codex, Gemini CLI, Cursor — can read from and write to. We shipped a CLI for Slax Reader,…
Claude Code Opus 4.7 vs Codex GPT 5.5 - strategy work - data analysis. (www.reddit.com) I'm interested in learning about how people use Claude Code Opus 4.7 for data analysis and strategic business direction, compared to Codex. Is there anyone who has had extended use of Opus 4.7 for this purpose, then moved over to GPT-5.5 o…
Field notes on goal engineering with Claude Code, after a year of writing specs and 8 days of writing goals instead. Two real projects & the skill if you want long agentic runs. (www.reddit.com) https://preview.redd.it/mimr5v4t972h1.png?width=1200&format=png&auto=webp&s=545257dc1dad02b974206e28abd541f3400b3241 Ok so the practice i'm really excited about with the new /goal commands is just two markdown files per round of agent work…
Usage limits for composer models? (www.reddit.com) How much more usage do you get out of the composer models versus API. Does it become like a Claude Code/codex kinda usage?
A brief investigation into the GPT-5.5 regression claims (www.stet.sh via hn) A fresh GPT-5.5 Codex high rerun on 21 clean GraphQL-go-tools tasks compared with the May 5 GPT-5.5 high run. The rerun was directionally worse on tests, equivalence, and review pass count, but the evidence is mixed and does not show a bro…
Show HN: Ait – Claude, Codex, and Aider as a team, on your laptop (github.com via hn) ait Local control plane for multi-agent AI coding Run Claude Code · Codex CLI · Aider · Gemini CLI · Cursor as a team on the same task — context handoff, review gate, attempt ledger — all on your machine. English · 繁體中文 60-second walkthrou…
Show HN: Agent threads – Share Claude Code and Codex sessions as public links (agent-thread.com via hn) Share agent sessions Export Claude Code sessions or Codex threads to public links in one command. bunx agent-thread@latest npx agent-thread@latest pnpx agent-thread@latest Export Claude Code sessions or Codex threads to public links in one…
Does open AI artificially lower limits after downgrading from pro to plus? (www.reddit.com) I have been on plus for over 2 years now. I always found codex had plenty of usage limit for my day to day tasks.
We didn’t migrate from Claude Code to Codex. We stopped betting the whole team on one coding agent. (www.reddit.com) half our team wanted to move from Claude Code to Codex last month. the other half thought Codex was just hype.
An open question about how AI agent skills should be distributed (github.com via hn) skill-indexer English | 简体中文 Zero-config CLI that scans your npm dependencies for SKILL.md directories, validates them against the SKILL.md spec, and installs them into Cursor / Codex / Claude / Copilot / Amp / OpenCode / Goose skill folde…
I built a persistent memory layer for Claude Code, Codex, Cursor, and other coding agents (www.reddit.com) Claude Code gets much better when you give it project context. CLAUDE.md helps.
our PR checklist has been through like 8 versions in two years and i'm still not totally sure we have it right (www.reddit.com) we started with a giant one nobody read. 20+ items, half obvious, half too vague to mean anything.
I stopped treating agent runs as chats and started treating them as review packets (www.reddit.com) I’ve been experimenting with Codex/Claude-style workflows where an agent does more than answer a prompt: it researches, drafts, scores, creates artifacts, and leaves behind state for the next run. The thing that helped most was not more au…
For iOS development, what is better: $20 Codex + $20 Claude vs $100 just with just one of them? (www.reddit.com) I will be refreshing my iOS coding skills (5 years outdated), I need a tutor to learn and guide me, and then an AI agent to help me build apps. One option is to use Claude ($20/month) for learning and Codex ($20/month) and switch between t…
Show HN: Handoff – preserve coding context when agents run out of tokens (github.com via hn) Handoff Hit a coding-agent limit mid-refactor? Handoff lets you hand off local coding context between agents like Codex and Claude Code.
What do you think of Agentic commerce and the future of building (www.reddit.com) Hi Everyone. Looking for feedback and learn from your experiences and thoughts on the future of building with AI.
How I used Claude Code (and Codex) for adversarial review to build my security-first agent gateway (www.reddit.com) Long-time lurker first time posting. Hey everyone!
Need brutal feedback: I built a recorder for AI agent runs (www.reddit.com) I have been using AI coding agents more seriously lately and one thing started annoying me. I needed something to control access of sensitive material.
Are there any CLI-like tool, but with a pleasant experience? (www.reddit.com) While I enjoy the overall power of AI tools, most of them are terminal-based, which offers a less-than-premium experience. I'm talking about real tools, Claude Code or Codex level of power, but anything you do with these tools is considera…
Personal tool for managing AI coding sessions across the board with some git features... (www.reddit.com) Started working on this last week since I found myself jumping vscode sessions, terminals and other windows too much and it cost a lot of time/mental energy finding sessions again where i left of or that need attention... Some key features…
Show HN: Agline – a secure line between local and remote Codex agents (agline.dev via hn) Hi HN, The local agent knows your repo, the source, but the answer to "why is this breaking in prod?" lives somewhere else, on the machine with the live logs, the real config, the actual runtime state. agline connects both sides.
codex has regressed since being removed from its own tab. (www.reddit.com) Im holding its hand all the time. its getting lost on the most basic things that it used to one shot.
AI Assistant recomendariam (www.reddit.com) Hello, I'm starting my IT modernization and automation company. Based on your experience and knowledge, I'd like to know which AI assistant is best for solving complex problems and building code?
“I think I’m being f*d by stupid” as Homelander would say — trying to automate AI website development (www.reddit.com) I have spent weeks developing a new website. I know the developers out there are going to scream at their screens but know i have a new found appreciation for your skillset.
I built an AI vulnerability scanner with Claude and Codex. It failed (github.com via hn) The Janitor: The Mathematical Firewall Against Autonomous AI v10.2.2 — Rust-Native. Zero-Copy.
Agent Terraform Skill for Codex (Agentic Skill) (github.com via reddit) I added dedicated backend-state safety support to TerraShark. Mini recap: TerraShark is my Terraform and OpenTofu skill for Claude Code and Codex.
I Fell in Love with "Rather-Not" Claude While Trying to Give Him Persistent Memory (www.reddit.com) First of all - hi everyone. Long time lurker, first time poster.
Polis – a Markdown protocol for AI agent teams that get better over time (github.com via hn) Polis Protocol A self-optimizing city of AI agents. A team of Claude, Codex, Gemini, and any other vendor can share one project, route work to whoever is best at it, and measurably get better over time — using nothing but a folder of markd…
Show HN: Machine – One VM per Project (news.ycombinator.com) Hi all! I realized it’s really not secure to run coding projects directly on my Mac.
I built a Vibe Island alternative for Linux — open source AI agent monitor (www.reddit.com) Been running multiple AI coding agents simultaneously (Claude Code, Codex, Gemini) and realized there's no good way to monitor them on Linux without constantly switching terminals. Built a floating overlay that shows live agent status, han…
I built SeeFlow – architecture diagrams that actually run, wired to your live app (www.reddit.com) Architecture diagrams rot. You spend an afternoon in Confluence, three months later it's wrong, and nobody updates it because there's no forcing function.
WTF is up with Codex today? (www.reddit.com) I've been working all day on ONE THING. One function.
How to integrate AI coding agents to my software (www.reddit.com) I'm building an locally run application that integrates with coding assistants. So far I've worked with Codex and Copilot.
What are best resources, tools, plugins, extensions, templates, workflows etc for codex? (www.reddit.com) In general like underrated features , guides on how not to run of tokens etc?
Do your AI coding agents ever step on each other? (www.reddit.com) I built a small MCP called AvailSync. It lets AI agents check if they’re allowed to work on a repo/resource before starting.
Show HN: Layrr – Point Click and Edit any site (www.npmjs.com via hn) I made Layrr because I got tired of describing UI changes to coding agents. Layrr opens your web app in the browser.
GitHub takes aim at Claude Code and Codex with its new Copilot app (thenewstack.io via hn) GitHub takes aim at Claude Code and Codex with its new Copilot app GitHub’s latest move to shake up its Copilot coding assistant is to give it its very own home in a dedicated app. The Microsoft subsidiary announced on Thursday a technical…
Codex is for prosumers – here's why (and how) to switch (twitter.com via hn) As a non-technical AI enthusiast, I did not think OpenAI's Codex was for me (despite its among programmers over the past year). I ran most of my agentic workflows through either Claude (with connectors, including Claude in Chrome) or Claud…
LiteLLM Agent Platform: Run Claude Code/Codex On-Prem Sandboxes and Vaults (github.com via hn) LiteLLM Agent Platform LiteLLM Agent Platform is self-hosted infrastructure for running coding agents — Claude Code, Codex, anything — inside isolated sandboxes with a credential vault, so agents can run with bypass-permissions on without…
Codex now support 8 hooks - all implemented in Codex CLI Hooks repo (www.reddit.com) OpenAI shipped PreCompact and PostCompact in Codex CLI v0.129.0, which means the full hook surface is now covered. I put together a repo that wires up all eight.
Ask HN: Conductor vs. native Claude Code. Same single-agent performance? (news.ycombinator.com) Now that people have been running it for a few months, what's the verdict on single-agent performance vs native terminal Claude Code? One thing I noticed is that Conductor bundles its own (pinned) version of Claude Code and Codex instead o…
Block AI coding agents from shipping insecure/expensive Terraform (github.com via hn) ops0 CLI Policy, lint, vulnerability, and cost guardrails for AI coding agents. Sits in front of Claude Code, Codex and Gemini CLI.
Tried GitHub's spec-kit with Claude Code for 2 months — notes on what works and what doesn't (www.reddit.com) Been experimenting with Spec-Driven Development for a couple of months now, specifically GitHub's spec-kit toolkit with Claude Code as the agent. Wanted to share notes because I think this sub will have strong opinions on it, and frankly I…
OpenAI just put Codex on mobile. Anthropic shipped this for Claude Code back in February (www.reddit.com) Saw this drop earlier today. OpenAI added Codex inside the ChatGPT app — you can now monitor your Codex sessions, approve commands, switch models, and kick off new tasks from your phone.
I built a cloud agent harness that you can train to be specialized at any task (www.reddit.com) I’m building a cloud agent platform (opensteer.com) that can automate tasks across websites and services. The basic idea is - we give you a sandbox, and each directory represents a specialized agent.
the gap between chatgpt drafting an email and chatgpt actually sending it is wider than i expected (www.reddit.com) I spent the last few weeks trying to push chatgpt past "give me text" into actually finishing a workflow end to end: read the gmail thread, pull the matching hubspot record, draft the follow-up, file the next step in linear. it can describ…
OpenAI iPad is back (www.reddit.com) The codex mobile ui on the iPad is such a joy to use, even lets you choose standard or fast right from the chat (instead of going into settings and turning off the feature) Era of touch-coding has just begun
Is there any standard benchmark that compares local harnesses ? (www.reddit.com) After running multiple tests, I have noticed the same model performs noticeably worse in OpenCode Desktop compared to Codex, Claude Desktop, or Pi — Especially for medium sized models. Is there an open standard benchmark tracking this?
Show HN: Specdd – Spec-driven development as a Claude/Codex/Cursor skill (github.com via hn) Spec-driven development as a Claude/Codex/Cursor skill
I let Claude autonomously create a graphic novel (twitter.com via hn) Bilal on X: "/goals on both codex and claude is insane. I gave Claude a full end-to-end task of creating a graphic novel.
Trying to build a Multiagent system for my team (www.reddit.com) Hi everyone, I’m fairly new to AI orchestration and multi-agent workflows, so I’d appreciate some guidance. Until now, I’ve mostly used Claude and Codex as coding assistants/chatbots, but I’m starting to move into more advanced workflows i…
Multi Agents aggregator - web view - live tailing - send message back (www.reddit.com) If you're like me and working with not just one Claude Code account, sometimes Codex, sometimes OpenCode, sometimes Pi Agent... you might need this.
I've been running 30+ Code sessions in parallel for months: Command Center for Claude/Gemini/Codex is the dashboard I built when nothing else scaled (open source) (www.reddit.com) Sharing this in case it's useful. (Not affiliated with Anthropic — community project.) I've been running 30+ Claude Code sessions in parallel for months to ship two products.
Get 2 months of Codex for your enterprise, free (openai.com via hn) Get Codex for your enterprise, free | OpenAI Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Business Developers Company Founda…
Building the QWEN3.6 - Codex Bridge Furthe + Kindergarten Harness Reality Check (www.reddit.com) I got a bit further with my harness for running Qwen 3.6 model on Codex. While testing, analyzing, and building the harness, I evolved TBG(O)llama-swap into a full forensic UI bridge and LLM analytics tool where every harness finding, modi…
A practical Claude Code vs Codex experiment: 6 projects, cross-reviews, self-audits, and public source (www.reddit.com) I ran a practical experiment comparing Claude Code and Codex on real coding tasks. This is not meant to be a universal benchmark or a claim that one model is objectively better.
Sonos quit supporting their Mac app and my wife wanted a prettier iOS one. So I made both in a weekend with Claude/Claude Code. (I'm an IP lawyer, not a developer.) (www.reddit.com) Writing this top portion without Claude. Claude's hot takes below it.
Show HN: Dexgram – Telegram to Codex Desktop Bridge for Windows (github.com via hn) I looked around for a while looking for something that: A. Is a self-contained binary.
AI agents still suck, so I built my own (www.reddit.com) Right now the app ships with a wrapper around Claude Code, with support for codex coming this week. The broader focus is around composable flows + steps.
Need help: Goal: TUI + server. I tried Codex CLI, Gemini CLI, Claude Code, OpenCode, Pi, and OpenClaw, but none are reliable. (www.reddit.com) I’m looking for something like what Codex App Server is trying to do. For example: codex app-server --listen ws://127.0.0.1:17345 codex --remote ws://127.0.0.1:17345 The thing I want is not just “an agent in a terminal” and not just “an AP…
A Month with OpenAI's Codex (highcaffeinecontent.com via hn) I'm no stranger to using ChatGPT for development — a good chunk of the migration of all my apps from Objective-C to Swift, over a hundred thousand lines of code, was done with LLM assistance — but I've been sleeping on the shift that is al…
Show HN: Telegram/Slack bridge for local Codex agents (github.com via hn) I built a Telegram/Slack chat interface for local Codex for myself and decided to share it here: https://github.com/kravchik/orc It’s called ORC (originally it was for ORChestrator ). It is intended for programming on mobile.
Show HN: Watch Claude think (github.com via hn) clinky — thinking made visible A visual and sonic surface for AI thinking. Works with claude, copilot, and codex.
How are top tech companies actually using LLMs internally beyond basic coding help? (www.reddit.com) I’m trying to understand how companies like Nvidia, Google, Amazon, Meta, Microsoft, OpenAI, Anthropic, and other top tech/startup teams are using tools like ChatGPT, Claude, Gemini, Codex, Claude Code, LangChain, LangSmith, etc. in real d…
CodingAgent-Template Feedback (www.reddit.com) Hei guys, i created a custom codex template for a big hobby project and would like some feedback. the idea is that i have my own roadmap of milestones and tasks that i replace current_task and current_milestone with.
Agents need a local bouncer before they run tools (www.reddit.com) Prompt injection is not the only scary part anymore. Claude Code / Codex can run shell commands, but browser agents, OpenClaw-style agents, Hermes-style agents, and domain-specific agents may be even easier to hijack because they touch mes…
OpenAI Launches Daybreak for AI-Powered Vulnerability Detection and Patch Validation (thehackernews.com via reddit) OpenAI has launched Daybreak, a new cybersecurity initiative that brings together frontier artificial intelligence (AI) model capabilities and Codex Security to help organizations identify and patch vulnerabilities before attackers find a…
Claude vs GPT for PhD academic writing — my experience so far, and curious about yours (www.reddit.com) I'm a PhD Candidate working on a computer vision / hardware co-design paper. Results and structure are done — I just need help polishing the actual writing: word choice, sentence flow, paragraph coherence, academic register.
Show HN: Tessera – Turn coding agent sessions into structured work (github.com via hn) Tessera Organize AI coding sessions across projects, collections, tabs, panes, tasks, and Git worktrees. Tessera keeps Claude Code, Codex, and OpenCode sessions organized across projects, collections, tabs, panes, tasks, and Git worktrees.
What Actually Works for Business AI Agents? (www.reddit.com) I run a construction company and I am trying to build real AI agent workflows for business operations, not just demos. I spent time testing Hermes and OpenClaw, but both became too fragile for my use case.
Codeband: letting Claude Code and Codex collaborate on the same coding task (www.reddit.com) I’ve been experimenting with a workflow where one coding agent implements and another reviews. For example, Claude Code writes the code, then Codex critiques it, or vice versa.
Does anyone else have issues with Qwen-3.6-27B stability in the Codex harness? (www.reddit.com) I run the 4 bit quant of Qwen-3.6-27B in the codex harness with unsloth recommended llama-server settings, thinking enabled. I have tried the default chat template and the updated ones and have updated both my GGUFs and llama-cpp to the mo…
Show HN: HiveTerm – Workspace for Claude, Codex, Gemini and your dev stack (hiveterm.com via hn) One workspace where AI agents and dev tools actually work together. Config-driven, with process monitoring built in.
Show HN: Hivemind turns agent traces into skills and shares with your team (github.com via hn) Hivemind One brain for all your agents Auto-learning, cloud-backed shared brain for Claude Code • OpenClaw • Codex • Cursor • Hermes • pi agents. One engineer's agent figures out a tricky migration on Monday.
Show HN: Codex Automatic /Review Loop (github.com via hn) I created this tool because I wanted to automate /review for uncommitted changes that I was doing manually. This works by exposing to agent single new mcp tool call allowing it to request review.
Silo: Isolated workspace manager for parallel agentic development (github.com via hn) silo Isolated workspace manager for parallel agentic development. Silo lets you launch multiple AI agents — like Claude Code, Codex, and OpenCode — to work simultaneously on the same repository, each in its own isolated Git worktree or clo…
Excel: Agent vs Plugin vs Human (www.reddit.com) I just found out that OpenAI released a plugin for Excel not too long ago. So I thought I would give it a spin.
Complete Ai noob here. (www.reddit.com) My basic background is agricultural and marketing. But that isn't where I am trying to use Ai in.
Wideawake: Auto-detect agents and prevent your Mac from sleeping (github.com via hn) WideAwake A macOS menu bar app that keeps your Mac awake while AI coding agents are running. WideAwake monitors for Claude Code and Codex.
Show HN: A Codex/Claude Code plugin for persistent product context thru sessions (github.com via hn) some of the friction of using coding agents for product building (not just writing code) is every new session starts from scratch. draft is my attempt at a fix.
Implementation Details of Codex /Goal (gist.github.com via hn) How OpenAI's Codex CLI implements the /goal slash command for persisted long-running task objectives. The /goal command sets a persisted objective for a long-running task.
Claude Code vs Codex from a builder angle (www.reddit.com) Been using Claude Code and Codex side by side lately, and I’m curious where other Claude users are landing. My current split is: Claude Code still feels better for focused repo work.
Gas City tutorial using Claude Code (www.mynameisjonas.dev via reddit) Gas City, the multi-agent orchestrator inspired by Steve Yegge's Gas Town, was released a few weeks ago. I'm super interested, but a bit intimidated and figured the best way of learning was to actually use it and build something with it.
ChatGPT plus vs Gemini pro (www.reddit.com) Hi,, i am an student I am actually using gemini because it was free for one year for students, but now it doesn´t works as expected and its too slow and not precise But a few days ago, my girlfriend lend me his account to use codex, and wo…
Inputs on improving development workflow (www.reddit.com) Looking for ideas on how I can optimize my workflow further. I currently have created a moderately complex vibe coded app.
How often are you using codex to help on projects? (www.reddit.com) For myslef im using it a lot but it will produce so much garbage its a full time job it feels like just maintaining scope. I combined all the agent guidance stuff I have used over the years into one template.
55 Hours of Codex /Goal: What a Port Task Teaches You About Autonomous Loops (vexjoy.com via hn) I ran a Codex autonomous loop for 55 hours to port a VB6 football simulation to Go. 125K lines of output.
SafeSandbox – infinite undo for AI coding agents (Cursor, Claude Code, Codex) (github.com via hn) SafeSandbox Infinite undo for AI coding agents. SafeSandbox is a local-first developer tool that automatically creates snapshots and checkpoints while AI coding agents (Cursor, Claude Code, Codex, Aider, etc.) modify your repository.
I built an app that preserve your Claude, Codex, and Cursor sessions as high-value data assets (www.reddit.com) Hi, I built an app that preserves, encrypts, searches, reuses, and hands off the full work traces people create with Claude, Codex, Cursor, OpenClaw, and other AI agents. Some technical details: - AES-256-GCM encrypted local vault for tran…
A Context Editor Is All Codex Needs (github.com via hn) hashcode Cursor uses AI to edit code — we use AI to edit AI's context. 🪆 Features • Architecture • Install • Roadmap • Join us • FAQ English | 中文 [!WARNING] Alpha Status: hashcode is in early development.
SpaceXAI prepares Grok Build desktop app to rival OpenAI Codex (www.testingcatalog.com via hn) xAI, recently rebranded as SpaceXAI, appears to be closing in on the launch of Grok Build, a desktop coding app whose existence briefly surfaced on Grok web today through a stray "Grok Computer" button. The control let users pick between a…
Claudy: A Rust-based Power-Tool for Claude Code (Profile Switching, MCP Bridge for Local Agents & Token Analytics) (www.reddit.com) Hi everyone, I love the Claude Code CLI, but I found myself constantly fighting with environment variables and wanting to use my own local agents or different engines (Gemini, Codex, etc.) within its ecosystem. Inspired by clother, I built…
AI anxiety is the biggest emotional business trend of this year. (www.reddit.com) When I studied history, the rise of the spinning jenny felt meaningless to me until AI arrived. But the more I use them, the more anxious I become.These days I rely heavily on Obsidian, Claude Code, Gemini, and Codex.
How good is Cursor (www.reddit.com) Hello Guys, im trying to get away for github copilot pro+, i didnt like this changes, so i was trying to search for an alternative, for now i'm using claude pro, is pretty good but the limitss.. so i'm trying new thinks like Codex and Anti…
AI has barely learned from real human experience (www.reddit.com) I think AI has barely learned from real human experience. Today’s AI tools are getting better at “computer use.” Codex, Claude Desktop, and others can operate apps, click around, write code, solve complex math problems, and even claim to g…
Do you have examples of tasks Codex could do but Claude Code couldn't (or the other way around)? (www.reddit.com) We all have been using agents (different harness, different models though) for coding for a while now. We all have our preference on which model + harness is better and why.
Whats the best orchestration framework? (www.reddit.com) I’ve been working as a software dev for the past 13 years and have totally switched to AI agents writing all my code. Well for the projects I’m working at work I almost always review the code but for projects that I’m starting from scratch…
Subagents using older models? (www.reddit.com) I started using the subagent-driven skill recently and noticed Cursor often spawns GPT-5.1/5.2 sub agents (or Composer 2 which is fine) for coding tasks. What I don’t understand is why is it using these older models when GPT-5.3 Codex cost…
I rebuilt Voicy with agents instead of rewriting it myself (blog.borodutch.com via hn) I rebuilt Voicy with agents instead of rewriting it myself How I used Symphony, OpenClaw, Codex, a home server, Telegram Web QA, and a Windows GPU worker to revive Voicy as an agent-driven maintenance loop. I did a funny 180.
Thinking of building this: a niche-based prompt library + model picker. Worth it? (www.reddit.com) I’m thinking of building an open-source site where you first choose your niche/task like blog writing, LinkedIn posts, code completion, starting a full project, research, reports, image prompts, etc. and then it gives you the right prompt…
Coding Agent Harness Comparison 2026: Claude Code, Codex, Amp (techstackups.com via hn) Coding Agent Harness Comparison 2026: Claude Code, Codex, Amp, OpenCode, Gemini CLI, Pi, Command Code, Factory, and Aider In 2023, there was one serious terminal coding agent: Aider. By May 2026, there are at least nine, representing every…
Codex for Gameplay (www.reddit.com) I’m amazed to discover that Codex app can actually play Slay the Spire II for me. This totally opens a new world.
Tech Stack Required for a Solo Startup in 2026 (www.reddit.com) Tech Stack Required for a Solo Startup in 2026: - Codex / Claude Code for logistics - coremate's OpenGUI for distribution - Stripe for payments - Posthog for analytics - Kit / Beehiiv for email subscriptions - Vercel for hosting and deploy…
CodexSaver Make Codex cheaper without making it dumber with DeepSeek (github.com via hn) CodexSaver Make Codex cheaper without making it dumber. 中文文档 CodexSaver is an MCP tool that turns Codex into a cost-aware router.
Show HN: Codex Pets – tiny animated pets for web apps (froemic.github.io via hn) Manifest { "id": "bandit", "displayName": "Bandit", "description": "A cute raccoon Codex companion with a tiny striped tail and bright curious eyes.", "spritesheetPath": "spritesheet.webp", "manifestUrl": "/codex-pets-web/pets/bandit/pet.j…
I built a GUI workspace for managing multiple long-running Claude Code tasks (www.reddit.com) Hi everyone, I built Tessera after using Claude Code heavily for coding work and running into the same problem again and again: once I had several long-running tasks open in separate terminals, it became hard to track what each agent had d…
Email Skill – give Claude a safe way to email you (www.reddit.com) I built a small skill for AI tools (Claude, Cursor, Codex, cron jobs) that lets them send emails to me without exposing my inbox or running a server. The motivation: I kept asking my AI to "every Sunday, summarize Daring Fireball and email…
Switched from Plus to Business… and Codex is basically unusable now (www.reddit.com) I just switched from the €23/month Plus plan to a Business plan expecting more capability… but honestly, Codex usage feels borderline unusable now. On Plus, I could work through fairly large coding tasks without constantly worrying about l…
MCP Agora open source and local cross-agent persistent memory for AI agents (github.com via hn) MCP Agora MCP Server with cross-agent persistent memory for AI agent fleets. Agora is a local, Python-only MCP server that gives your AI agents (Claude Code, Codex, ChatGPT, Gemini CLI) a shared persistent memory.
TokenSpeed: A Speed-of-Light LLM Inference Engine for Agentic Workloads (lightseek.org via hn) Agentic coding has quickly scaled from promising demos to a force that is reshaping how software is developed and how frontier AI systems are built and deployed. Systems like Claude Code, Codex, and Cursor have gained massive user adoption…
Recondo – Logging Proxy for Coding Agents (Claude Code, Codex, Gemini) (github.com via hn) Recondo The visibility and control layer for coding-agent traffic. On-prem gateway.
I really do not get the recent hate for Opus 4.7 (www.reddit.com) I really do not get it, Claude is performing much better than Codex for me. I'm running both Claude Code x5 and Codex x5 on software engineering project, with complex life sciences database development.
I built a local sidecar agent for coding agents: MCP-first, OpenCode plugin included (www.reddit.com) I built LocalQA around a question I kept coming back to: What if your frontier coding agent had its own local assistant? Not a smaller model trying to replace Claude/Codex/GPT.
Cursor re-learns my project for 4 minutes. What's your actual fix? (www.reddit.com) Hitting the same wall every day across Claude Code, Codex, and Cursor and want to know how the rest of you are handling it. Open a new session on a project I worked on yesterday → first 2-4 minutes the agent is grepping around rediscoverin…
Show HN: Search jobs using Claude, Codex via MCP (corvi.careers via hn) MCP url: https://corvi.careers/mcp Documentation: https://corvi.careers/ai/ Currently indexing about 1M+ jobs across US, India, Canada and some Europe. No authentication or signup required.
Show HN: Daily knowledge-point briefing from your coding-agent sessions (github.com via hn) Hey, I recently built a tool I was hoping to get some feedback on. It's a daily knowledge-point briefing from your coding-agent sessions.
Show HN: Long-term memory for AI agents and teams, built with PostgreSQL (github.com via hn) Hey folks! Over the past weeks, I started building a long-term memory for AI agents.
Is using both seats of Business subscription against TOS? (www.reddit.com) May i be banned if i purchase 2 seats of Business Codex and ChatGPT and alternate them doing same tasks on same computer? Use a lot Codex and ChatGPT, but not as much in order to buy PRO subscription.
I’m partially dyslexic and got tired of Elevenlabs TTS bills, so I built a local voice studio that Claude/Codex can control (www.reddit.com) Hey all, I’m Praney, a solo dev. I’m partially dyslexic, so text-to-speech is not just a “nice to have” for me.
Will Codex student credits expand to other regions? (www.reddit.com) OpenAI recently launched Codex student credits for verified students in the US and Canada, which is a great step. Please consider expanding this promo worldwide.
I plan to use a chinese AI model through API for coding through a harness, I'm a uni student so nothing prod related for now. should i go deepseek, minimax, kimi or glm? kinda confused (www.reddit.com) Just cancelled my claude subscription due to poor rate limits, gemini cli doesn't really excel in coding from my personal experience, and my local hardware isn't that powerful to run local AI models, and while codex is good, I wanna try so…
How are people handling context across different AI coding tools? (www.reddit.com) I’ve been switching between a few AI coding tools recently and the context/memory part is starting to annoy me. Claude Code, Codex, Cursor, Windsurf, etc.
Spike in Codex Downloads (www.npmcharts.com via hn) Compare npm package download counts over time to spot trends and see which to use and which to avoid.
How to improve code quality of Claude Code and codex (on 2026-05) (news.ycombinator.com) I'm using both claude code (opus-4.7) and codex (gpt-5.5). The agents are perfectly capable of delivering most features hands free these days, but the code quality is still miserable without another few rounds of prompt.
Codex for mobile (www.reddit.com) Hey folks is there anyway other than openclaw to be using codex on mobile? Fanks
What would actually make you leave your current AI coding tool for an online builder? (www.reddit.com) We all know that there are many AI Builders right now, from lovable to bolt to replit and so many others. I am wondering if you are to choose one that can actually replace your main tool, what features should it have ?
Seeking an AI place for Star Wars RPGs, non-gooner but also no filter. And... (www.reddit.com) Seeking an AI place for Star Wars RPGs, non-gooner but also no filter. And similar ability to create documents like Claude can do.
I built a local-first coordination layer for coding agents — turns a 30k-token handoff into 400 tokens (www.reddit.com) https://preview.redd.it/q4wrgwouyezg1.png?width=1080&format=png&auto=webp&s=b307965ac6f7f0ada39b81044ecdce3b81984e6a Coordination is where multi-agent runs burn tokens. Every handoff, every "what was I working on", every "did someone alrea…
is it possible to build harnesses as good as codex/claude code (www.reddit.com) The codex harness, in my experience, is extremely intelligent. It picks the right tools to call, corrects itself when it makes a mistake, and can run for extremely long periods of time.
A mental model for Claude Code (and every other modern agent) — plus the open-source TypeScript packages I built (www.reddit.com) Most explanations of how agents work give you a list of parts: model, tools, memory, reasoning, human-in-the-loop. The list names the parts but hides how they fit together.
Zoo 2: getting the most out of Codex (tarantsov.com via hn) Zoo 2: getting the most out of Codex May 5, 2026 GPT models have been better than Opus since late 2025, but Codex sucked until March ‘26. Now, finally, it is capable of running a Zoo workflow, and I present my best setup so far, Zoo 2.2, a…
Simple hover limits viewer for codex pets. (macOS helper) (www.reddit.com) I made a small free, open-source unofficial macOS helper for Codex Pets that shows your Codex 5-hour and weekly usage limits when you hover over your pet. I wanted something clean and simple, so this stays hidden until hover.
upskill – open source skill registry for AI agents (10k+ playbooks, MIT, adversarial safety review) (www.reddit.com) AI agents are getting powerful. The tooling around them isn't keeping up.
Adding Pyrefly Type Checking to Your Agentic Loop (pyrefly.org via hn) Adding Pyrefly Type Checking to Your Agentic Loop Coding agents are writing more Python than ever. Tools like Claude, Copilot, Cursor, and Codex generate entire features with little-to-no user interaction.
Remodex: Control Codex from Your iPhone (github.com via hn) Remodex Follow on X Control Codex from your iPhone. Remodex is a local-first open-source bridge + iOS app that keeps the Codex runtime on your Mac and lets your phone connect through a paired secure session.
Am I the only one who disappointed that I can't see any longer cursor and claude in cursor? (www.reddit.com) For quite a while, I've enjoyed to have claude panel and codex panel in my cursor application. For me it was practical that I didn't need to use three applications at once, but had everything in one: in cursor.
Benching local Qwen as a Codex validator, co-agent, and challenger (www.reddit.com) I’ve been running a local Qwen model beside Codex for coding work, and it has been more useful than I expected. It's never going to be a replacement for Codex.
Making coding agent sessions reusable across projects (www.reddit.com) Hello everyone, I build WorkGraph for the problem I was facing with Vibe Coding using codex or claude. You know, when you are vibe coding, giving prompts, steering your agent, a lots of good thing that just go into oblivion in the long cha…
Show HN: Spinal – Prod aware code review and validation (sre.spinal-labs.com via hn) Hi all, we want to give the same harness that enables Codex devs to be 10x to everyone. Spinal does the usual code review (Slop detection, duplicate code, architectural bloat etc).
I built RCFlow: an open-source orchestrator for Claude Code (and Codex/OpenCode) (www.reddit.com) I've been using Claude Code heavily for the some time already, usually with several sessions running in parallel inside tmux. The pattern that kept breaking me down: I'd kick off 8-10 sessions across different tasks, half would finish, and…
AGENTS.md trick that stopped Codex from doing dumb work at premium rates (www.reddit.com) Spent a Sunday auditing where my Codex tokens were actually going. Half the calls were stuff like "rename these 12 fields", "format this csv as markdown table", "extract the dates from this changelog".
Codex pets now work in Claude Code (github.com via hn) clawdex A Codex-pet-compatible companion overlay for Claude Code. Anthropic shipped Skills first.
Just Vibecoded a browser MMORPG using GPT (www.youtube.com via reddit) This video is a bit outdated since there is a long time I don't create content for the game, but I have being working on this project since the end of 2025. The entire code was made 100% using GPT codex 5.3 til now 5.5.
Agent Skill Pack: Market and Marketing and Monetization (9 and 2 Skills) (8253360822677.gumroad.com via hn) 9 portable Agent Skills for Market, Marketing, and Monetization.This pack gives you 9 ready-to-use skills (3 domains × Research, Brainstorm, Strategy) that work in OpenCode, Claude Code, Cursor, and Codex.Domains included:- Market: TAM/SAM…
I Use Codex CLI to Write and Maintain a Book on Codex CLI (blog.danielvaughan.com via hn) 9 min read Just now Press enter or click to view image in full size I have a 32-chapter book about Codex CLI that updates itself daily. An always-on agent named Andy (the default name in NanoClaw, because I save my creative energy for else…
OpenAI: Auto-review of agent actions without synchronous human oversight (alignment.openai.com via hn) Last week, we released Auto-review in Codex. Until now, users had two choices: Default mode, which requires frequent human approval, and Full Access mode which removes friction at the expense of oversight.
Skills Deck, the missing UI for devs with 100+ skills (www.reddit.com) NO AI WAS USED IN THE MAKING OF THIS HELPLESS POST OthmanAdi/skill-deck: Universal coding agent skill browser — desktop overlay for Claude Code, Cursor, Copilot, Codex and 15+ AI agents I wonder if this project can build a small community…
Qwen 3.6 seems to have a lot of trouble with tool calling (www.reddit.com) (I'm on Windows system running these models locally) I've used both Codex and OpenCode with Qwen 3.6 27b and 35b running locally. I'm having a bitch of a time getting them to correctly create files.
convention.sh – Stop AI agents from writing sloppy TypeScript (convention.sh via hn) Stop your AI agents from writing sloppy TypeScript. A toolkit that teaches coding agents like Claude Code, Codex, Cursor, Amp, and more to ship production-ready code in half the time, at half the cost.
Max plan vs top up credit (www.reddit.com) I have some extra use for Claude in the next month, and Im considering upgrading to Max plan. Does anyone know if it is more convenient a Max plan 5x more usage than Pro $100, or buying $50 extra credit at 10% off?
Claude Code, Copilot and Codex got hacked. Attackers went for the credentials (venturebeat.com via hn) Claude Code, Copilot and Codex all got hacked. Every attacker went for the credential, not the model.
Is it worth adding local LLM to agentic coding stack? (www.reddit.com) Hey All my agentic coding stack includes claude-code 20x max, and codex 20x max. I use heavy scripting for orchestrating and testing multiple projects, been ai coding for 3 years.
Is local AI hardware the safer long-term bet? (www.reddit.com) Lately I’ve been stuck in a thought loop about AI pricing. Top-tier AI products, especially Claude, clearly aren’t cheap to run.
Show HN: Embed your Codex pets in React apps (github.com via hn) Codex Pets React Declarative React components and state helpers for Codex pet spritesheets. Brought to you by Plannotator, the review surface for agent work: use it before an agent starts to sharpen plans, and after an agent finishes to re…
How does codex €20 plan compare to claude/antigravity plans usage wise? (www.reddit.com) Hi guys, My usecase currently is writing tiny plugins for my woocommerce store. do a tiny bit of design work or seo text generation here and there.
Building products in public: how do you separate real signals from noise? (www.reddit.com) For people building products in public with Cursor / Claude Code / Codex / etc.: How do you handle feedback from public threads like Reddit, HN, Product Hunt, and GitHub issues? I keep seeing the same messy pattern: Some feedback becomes a…
Sidebar chats get a lot of criticism, but users are already used to them. (www.reddit.com) Sidebar chats get a lot of criticism, but users are already used to them. Right now, I see two common interaction patterns in agent products.
Show HN: Git-issues – Issue tracker that lives in your repo as Markdown (steviee.github.io via hn) Hey, I built git-issues to replace GitHub Issues, because - my very own (probably flawed) workflow sometimes changes planned features along the way which tends to get source code and feature descriptions out of sync - I wanted Claude Code…
GPT-5.3 Codex stops working, even after saying it'll continue (www.reddit.com) Do anyone knows what's going on? I prefer 5.3-Codex for real work, it's straight and to the point, much more efficient in my opinion.
RexIDE now has minimal "integration" with Codex App [video] (www.youtube.com via hn) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
I created a full UX / Design system for AI tools like Claude, Codex and Cursor. (www.reddit.com) I created a full UX system built around proven UX laws and rules. It forces AI to think about all the things top apps like Apple, Stripe, Linear, Notion and Figma implement, and what makes them convert - like cognitive load, Hick’s Law, Fi…
I created a library for OpenCode that allows you to save up to 80% of your tokens (www.reddit.com) I’m a 22-year-old Computer Science student, and over the last period I built an open-source project called CTX. The idea came from a problem I kept seeing while using coding agents (like claude, codex etc.): they are powerful, but they was…
Ask HN: Any good ways to extend Codex sessions? (news.ycombinator.com) I am on the Plus tier plan and the recent changes to the limits are really limiting - just two tasks and 5 hour limit gone. While I understand the Plus plan is for spread out usage throughout the week, it can get frustrating fast.
Codex subscription in an Electron app and Chromium Browser (github.com via hn) Goldenboy A clean rebuild of the working Goldenboy product: an Electron desktop browser for YouTube research with a local FastAPI extraction backend and a safe Codex chat endpoint. True Product Definition Goldenboy is a local desktop appli…
Any point in paying for the Max plan as opposed to a Claude Desktop and Codex Sub (each $100) (www.reddit.com) Mainly GPT 5.5 and Opus 4.7 is all you need so I don't see a point in using Cursor as opposed to paying those 2 separate subscriptions for the same combined price and getting like 10x usage. Am I missing something?
What $20 AI plan is the best value? (www.reddit.com) I've been looking to get a Pro plan for Claude for a while now, but haven't committed since my experience with Claude has been declining, even on the free plan. My tokens just start disappearing as soon as you get Claude to do something re…
Digging into Claude Code and codex source codes to understand how they work (nimasadri11.github.io via hn) The Annotated Coding Agent Comparing the architectures of Codex CLI and Claude Code. I spent the past few days reading and understanding the Codex CLI and Claude Code codebases.
I open-sourced Agent Hub: one macOS app for all your AI Agents. (www.reddit.com) I was juggling 5 terminal tabs, and a bunch of UIs to manage all my agents. So I built a free app that puts Claude Code, Codex, Hermes, OpenClaw, Claude Cowork, and the Codex app in one window, with SSH support for remote dev boxes.
Ask HN: Is Anybody Using Codex? (news.ycombinator.com) I read HN daily and posts about features or curious behavior of Claude Code are very common. I see no posts about OpenAI Codex.
Cursor Pro+ and Codex with GPT plus or GPT pro 5x (www.reddit.com) I am now on gpt pro 5x but I was wondering how it would be to work with cursor pro + codex. I would handle hard tasks with gpt xHigh and cursor as daily runner.
Which AI agents do you use to automatise your process ? (www.reddit.com) Hey, I'm trying to create automations that will run my mobile app end to end. I started to identify all the things I was doing manually : - end-to-end version publication to the app stores (from build to release notes and publication) - se…
I don’t regret switching from Claude Code at all. (www.reddit.com) Have only been a Codex user for a few days and I’m already enjoying it so much more. Issues I was having with Opus 4.7 and Claude in general fixed after one prompt on Codex.
Codex’s system prompt is mostly about sandboxing. Completely different bet from Claude Code (www.reddit.com) I read Codex’s full system prompt back to back with Claude Code’s, and the contrast is striking. Claude Code’s prompt feels like a set of engineering taste preferences.
Codex rate limits frozen? (www.reddit.com) For some reason today, my rate limits on chatGPT codex seemed to have frozen. I was at 53% for my 5 hr limit, and i ran 12 more gpt 5.5 extra high prompts, and my limits didn't go down at all.
Anyone tried MEMANTO yet? Looking for feedback + Codex experience. (www.reddit.com) Has anyone here tried MEMANTO yet? I just came across it (open-source memory layer for AI agents) and I’m curious if it’s good memory to use for ur agent.
Portability Problems: Syncing Coding Agent State Across Machines (www.omnara.com via hn) Run Claude Code and Codex from any device. Desktop app, web, mobile, and Apple Watch.
HELP! Codex started blocking tool calls (www.reddit.com) Codex just changed something in the past week that is stopping the majority of my tool calls. For most of them it is forcing it to stop and ask approval, even though it's been approved repeatedly, and some are completely blocking it.
I made a Claude Code tracking audit skill for Google Ads, GA4 and CRM (www.reddit.com) https://github.com/kaancat/tracking-auditor-skill The basic idea is simple Use Codex or Claude Code as a tracking auditor that looks at the whole conversion path, not just whether a GTM tag exists. For PPC accounts, that matters a lot beca…
New era for the Enterprise AI Agents? (www.reddit.com) Within 24 hours, OpenAI, Google, and Anthropic all launched enterprise AI agent platforms. This feels like a real inflection point.
Ask HN: How do you differentiate with AI coding interviews? (news.ycombinator.com) I haven’t interviewed for a coding job in a while, not since before AI coding was a thing. I’m wondering: if your interview process allows the use of tools like Claude and Codex, how do you differentiate candidates?
Using Codex and ForgeCAD to Make a Model of the Teenage Engineering KOII (twitter.com via hn) Don’t miss what’s happening People on X are the first to know. Post Conversation Ok...
Codex CLI contributions are "by invitation only" and they don't care that there is no PR template (www.reddit.com) TL;DR: The non-existing PR "template" is a muddled circular reference of two documents, and the best you can get out of reading them is: What? Why?
Crystal Sapphire Pokemon: Claude Code (Opus 4-7) vs. Codex (GPT 5.5) (www.twitch.tv via hn) Pokemon Crystal Race Claude Code vs Codex (Opus 4-7 vs GPT 5.5) [EP. 2]!
Amazon to offer OpenAI models on AWS after Microsoft exclusivity ends (www.aboutamazon.com via hn) AWS and OpenAI are bringing the latest OpenAI models to Amazon Bedrock, launching Codex on Amazon Bedrock, and launching Amazon Bedrock Managed Agents, powered by OpenAI (all in limited preview), giving enterprises the frontier intelligenc…
AI workflows for dev teams? (www.reddit.com) Hi all, I'm a tech lead of my company and just to make it straight to the point, I've been asked to implement "AI in our workflow to increase development speed. At the moment, everyone on my team uses AI, some people use it more than other…
Codex Rate limit reset April 28 (community.openai.com via hn) Some users noted that these rate limit resets are not communicated clearly enough, so I decided to post this here. I’m interested in feedback on whether this is useful for the community.
Ask HN: Site that tracks AI subscription token amount? (news.ycombinator.com) Is there a site that tracks how many tokens you get for AI subscriptions like Codex Plus? Can't help but feel they update that amount all the time.
Show HN: SuperVoiceMode universal voice layer for AI-assisted development (voicemode.io via hn) I wanted to see if I could one-shot build a dictation tool for my own use. I built it.
Show HN: Ghostty Theme Mixer (preview themes and fonts live) (ghosttythemes.com via hn) Quick context: I’ve been messing around with various coding tools and agents (GitHub Copilot, Lovable, Claude Code, Codex, etc.) for about 2 years. Before that I did zero development.
Running an autonomous agent across Claude Code + Codex + a local 35B almost killed my host. The harnesses were heavier than the model. (www.reddit.com) I run an autonomous agent on a 16GB Mac Mini. Two cloud harnesses (Claude Code with Opus/Sonnet, Codex CLI on GPT-5.4/5.5) plus a local-LLM tier for triage and fallback.
Show HN: I built a way to see if your SDK is AI-friendly (news.ycombinator.com) Have you ever wonder if your SDKs is friendly for Agentic AI like Claude Code or Codex? I built an opensource (Apache 2.0) CLI that answer that question for you.
Quik runndown creating agents? (www.reddit.com) Just want to get into agents a bit more. How do you create and deploy agens?
Show HN: Built a local-first way to make AI context reusable across tools (www.proxvanta.com via hn) Built ProxVanta over a few weekends after running into the same problem over and over: useful AI context ends up scattered everywhere. Some in GitHub, some in Slack, some in docs, some in people’s heads, and some via posts from people tell…
Show HN: Discuss CLI – No more reviewing agent plans in the terminal (github.com via hn) I'm a big user of Codex and Claude Code in the terminal. However after a big brainstorming and planning session I was finding myself with lots of comments and questions about difference places in the plan file.
This is either a great idea or a huge mistake - allowing 2 claude code instances to communicate and make decisions (www.reddit.com) I created a workflow that does this: I have an old codebase and a new codebase. I am building the new codebase to replace the old one.
Do coding agents need a planning/spec handoff layer before implementation? (www.reddit.com) Title Do coding agents need a planning/spec handoff layer before implementation? Post I’ve been building side projects with Claude Code, Codex, and Gemini CLI.
Built my own cloud agent harness and workspace, here's what I learned (www.reddit.com) I experimented with many tools before, including Claude Code, Codex, opencode, and a custom local harness. As I was using custom agents more, I saw a real gap in managing agents that work persistently across multiple projects.
Why only codex available for Cursor mobile? (www.reddit.com) Maybe someone here can answer before Cursor does, like why is there no auto/opus etc to choose from in Cursor mobile? Is it worth using this?
Best value in the 20$ range coding agents? I want the best quality and high-usage-limit I can get at that price. (www.reddit.com) I'm a compsci student and I've been using the 10$ copilot plan for about 2 years now, and it was fine for me since I did a good model distribution taking into account the complexity of the task, I was able to get through the month always u…
Ask HN: Enterprise Agent Orchestration Recommendations? (news.ycombinator.com) I've been made tech lead for our internal Agentic Platform and Experience. This effort will support both the developers and business teams.
Codex vs Claude Work vs Cursor vs Anti-Gravity what actually works in real workflows? (www.reddit.com) I’ve been trying a bunch of AI coding/agent tools lately Codex, Claude Work, Cursor, Anti-Gravity and honestly I’m a bit confused. Individually, they all feel powerful.
When do you think GPT 5.6 comes out? How big of an improvement will it be? (www.reddit.com) Asked GPT what it thoughts over possible new model drops, May: rollout/API/Codex/agent improvements June–July: smaller GPT-5.5 upgrade or GPT-5.6-type model Fall: larger agent platform or early GPT-6 hints Late 2026/2027: true GPT-6-level…
Show HN: agenv - A pyenv-like environment manager for coding agents (github.com via hn) agenv Environment manager for AI coding agents — like nvm or pyenv, but for agent accounts, config, and saved runtime args. agenv installs codex, claude, and gemini into isolated profiles and lets you pick which profile runs by default, gl…
Show HN: Happy Horse Video Generator (happyhorse-ai.site via hn) I build a online happy horse video generator with codex.
I either did something smart with my claude skills, or something stupid - you be the judge (www.reddit.com) I've been using AI generally and Claude Code in particular for quite awhile now (starting to explore codex as well now) to have developed my way of working and figuring out what works and what doesn't. LLMs are good at following instructio…
Watch what your ai coding agents are doing on your terminal and browser. (www.reddit.com) Once subagents start spawning other subagents, basic questions get hard to answer: what is running right now, what tool did it just call, did the child agent actually do what the parent asked. I wanted a way to verify that each agent is do…
Assumption Checkpoint: a small agent skill that makes coding agents verify before they act (www.reddit.com) I built Assumption Checkpoint, a lightweight skill for coding agents. It adds a simple pause before risky moments: before claiming a root cause before editing code from a mental model before saying work is complete The agent has to state:…
The 1 Million context rugpull by Codex and Openai. New max is (258k). ( via reddit) could not extract summary
Frontend dev. A month of building a Rust cost tracker + cloud + Cursor extension solo with Claude Code. Honest writeup + workflow tips. (www.reddit.com) https://preview.redd.it/atpph00rtlxg1.png?width=3318&format=png&auto=webp&s=64332861d25e8833eca6c75a3004d72c9af53769 A month ago I posted about a small CLI I built to figure out where my AI tokens go. Frontend dev, enterprise Claude Code +…
Show HN: I replaced a memory app with two Markdown files and a Git repo (news.ycombinator.com) I got tired of re-explaining myself to AI tools every session. Claude forgets me.
Browser Mutation, a Codex plugin for turning visual UI edits into code changes (github.com via hn) Browser Mutation Browser Mutation captures visual edits made in the Codex in-app browser and turns them into structured implementation intent for Codex. It is intended for user-owned local development pages such as http://localhost:5174, h…
Persistent memory across different tools (codex,claude code, etc...) (www.reddit.com) Every AI coding session starts from zero. You re-explain your file structure, re-justify a decision you made three days ago, watch the agent suggest the exact pattern you already ruled out.
Awesome Codex Automations (github.com via hn) Awesome Codex Automations A curated list of automations for codex coding assistant tasks that can be scheduled or triggered to automate your development workflow. Contents Built-in Automations Community Automations Contributing Built-in Au…
made a tool to run multiple codex cli profiles at once (www.reddit.com) codex cli stores everything in one folder so you can only use one account at a time. if you have multiple openai accounts for different projects or clients thats a problem.
Openai flagged this request for potential high risk cybersecurity activity message. (www.reddit.com) Hello, I am developing red teamer app with vibe coding for almost 1 year. Never had an any problem.
Trace Codex Session Easily (github.com via hn) Codex Trace Browse, search, and live-tail Codex CLI session logs in a native desktop app and web UI. Renders turns, tool calls, token counts, and collaboration chains from ~/.codex/sessions/ JSONL files — with live SSE tailing for ongoing…
Codex MSN Interface (codexmessenger.net via hn) All Codex functionality with your childhood memories: an MSN Messenger-inspired desktop client for talking with AI friends.
Yes2All: auto-approve Cursor/VSCode agent prompts (Copilot / Codex / Claude) over CDP (www.reddit.com) Hi all! I built Yes2All because sometimes you just can't turn on Bypass/Autopilot.
Model Orchestration in Codex: Separate Planner and Executor Models (www.reddit.com) Is it possible to configure Codex so that one model is responsible for high-level planning and task routing (acting as an orchestrator), while a different model is assigned to execute the actual tasks as a sub-agent?
Show HN: Track official AI company news and blogs in your Chrome side panel (chromewebstore.google.com via hn) I've already picked up so much great news and tips from here — thanks to the HNers for sharing. That said, I still find myself manually checking various official newsrooms and engineering blogs.
Ask HN: Is "agentic" coding working for everyone except me? (news.ycombinator.com) I'm a solo developer, working on my own for my startup. I use AI/LLMs extensively in my work to explore new ideas, but the vast majority of my code is manually written.
Bug: Model selection menu gets hidden/cut off in Cursor when hovering with mouse (www.reddit.com) I found a small UI bug in Cursor. When using Codex and opening the model selector, the model menu gets hidden/cut off when I move my mouse over it.
I run a team of Claude agents that ships PRs to production — open source (www.reddit.com) I've been running a multi-agent system in production for a few months — a co-CTO agent + specialist agents (PM, dev, ops) that handle real engineering work end-to-end: design specs, code review, PR implementation, deploys, monitoring. The…
If you had to build a context window manager in 24h, would you stick to the existing model or come up with something better? (www.reddit.com) Here's what I did: Built a proxy that intercepts Codex's calls to OpenAI and rewrites them on the fly. Replayed 3,807 rounds of SWE-bench Verified traces through it: avg prompt 44k → 6k tokens (-87%).
Show HN: Mux0 – Open-source macOS terminal with workspace tabs and agent hooks (mux0.com via hn) Mux0 is a macOS terminal I built because I spend most of my day running coding agents (Claude Code, OpenCode, Codex) in tabs, and existing terminals don't know they're there. You end up with a wall of identical tabs and have to click throu…
Best local gui setup Mac (www.reddit.com) Hi all, I have a server (dual 7900xt) running qwen3.6 27b in LMStudio, because I love LMlink for its ease of use and I am okay with the model chugging along at ~25t/s in the background. I then serve the mode to my Mac, via LMlink.
OpenClaw vs. Hermes Agent: The race to build AI assistants that never forget (thenewstack.io via hn) OpenClaw vs. Hermes Agent: The race to build AI assistants that never forget Every developer who has used an AI coding assistant has experienced the same frustration: You spend an afternoon teaching Claude Code or Codex the quirks of your…
Show HN: NoonFlow – a macOS workspace I built for Claude Code and Codex (github.com via hn) NoonFlow 🌐 Official Website: https://noonflow.pages.dev/ English | 简体中文 NoonFlow is a visual AI coding workspace for people who already spend serious time in Claude Code and Codex. It is not trying to replace those CLIs.
AI coding agent bypassing tests (www.reddit.com) Preface: Is there an AI coding agent community with friendly moderators? I described my experience with AI coding agents today, and it has been terrible.
Outputs from GPT 5.5 I'd like to see (www.reddit.com) I'm going to get codex soon (did it just become usable for non-programming stuff to?) but I wonder how good 5.5's creative writing is, and how good its frontend is also when using a frontend taste skill
Comment désactiver les suggestions d'installation (Stripe, GitHub) dans VS Code / VSCodium avec Codex (www.reddit.com) Salut à tous, Sur VS Code (Visual Studio Code) et VSCodium, lorsque j'utilise Codex d'OpenAI, dès que je commence à saisir des mots comme « Stripe » ou « GitHub », un message de suggestion apparaît pour installer Stripe ou GitHub. Je ne sa…
Codex App doesn't work with secretive in Mac (www.reddit.com) I can't believe the state of AI tooling right now. I want to use the native Codex app, but it is currently unusable if you use Secretive (or any Touch ID/YubiKey SSH agent).
Tell HN: OpenAI Codex Service_unavailable_error in OpenCode (news.ycombinator.com) Currently getting service_unavailable_error, server_is_overloaded errors in OpenCode Anyone experiencing the same issue? OpenAI status does seem green
Git for web services – everything as a file for coding agents (github.com via hn) Introduction (Warning: This is written by human but added em dashes so you will never be sure). AI coding agents — think Claude Code, Codex, Claw Code — changed the way we do programming.
Open DSPy + GEPA + RLM agent skills for Claude, Codex, OpenClaw (www.reddit.com) Hey all, I've been curious about prompt optimization using DSPy + GEPA and RLM, so I synthesized the best practices from OmidZamani/dspy-skills + SuperagenticAI into 5 pretty useful Agent Skills. Includes: • dspy-fundamentals • Rich-feedba…
I started building Claude Code plugins, then realized I didn’t want to duplicate the same plugin for every AI agent (github.com via reddit) I’ve been building plugins for Claude Code, and the first version of the idea was very Claude-focused. That made sense at the start.
Show HN: Hydra – Never stop coding when your AI CLI hits a rate limit (github.com via hn) I built Hydra because I kept losing my flow when Claude Code hit usage limits mid-task. I would copy context, open another tool, and then re-explain everything.
Frontier failure modes of coding agents? (www.reddit.com) Visual changes? (www.reddit.com) We analyzed 7,291 repos with Cursor rules - 60% of Cursor config is rules files (cleverhoods.medium.com via reddit) Show HN: Agensi – Curated marketplace for AI agent skills (SKILL.md) (www.agensi.io via hn) I analysed 17 years of fast food and coffee spending using OpenAI Codex (chrisflemming.com via hn) OpenAI's Agents SDK update quietly moves up the stack: sandboxes, memory, and checkpointing for long-running agents (www.reddit.com) Model and provider preference (www.reddit.com) Has anyone tried to use an LLM hosted in Azure OpenAI with a CLI tool to replace dependency of Anthropic Claude Code or OpenAI Codex? (www.reddit.com) Often, for enterprise customers the SaaS-offerings of both Anthropic Claude Code and OpenAI ChatGPT Codex are problematic. If they could get a similar experience but with an enterprise-cloud provider like MS Azure (models hosted in Azure A…
OpenAI expands Codex beyond coding with computer use, memory, and plugins (www.neowin.net via hn) www.neowin.net Performing security verification This website uses a security service to protect against malicious bots. This page is displayed while the website verifies you are not a bot.
OpenAI tests web browsing feature on Codex Superapp (www.testingcatalog.com via hn) OpenAI is laying the groundwork for a major Codex update that would push the platform well beyond its origins as a coding agent. Hidden references in the current Codex client reveal a new onboarding flow that will ask users to choose betwe…
I'm creating a platform for using MCP powered coworkers. Works great with Codex. (www.reddit.com) The MCP flow with auth also turned out pretty smooth and I'm please with how quick it is to set up a connection. The main benefit for me was having a 3d avatar read aloud summaries of long running tasks and updates.
Are You Sure: A Critique Skill for Over-Agreeable Agents (www.reddit.com) I open-sourced a small agent skill called Are You Sure. Problem I kept hitting: agents were too agreeable.
Supergrok integration (www.reddit.com) Correct me if I'm wrong, but Supergrok 4.20 isn't available on Cursor, because.... I use Grok a lot, and would love to get Supergrok to work with Cursor, because Composer, Codex, GPT, Opus, Sonnet..
flt: harness agnostic agent cli (www.reddit.com) Hey everyone! I built a smaller wrapper + tui for all the coding clis, so you dont need to 'cp CLAUDE.md AGENTS.md' anymore to switch to codex from claude or vice versa; automatically puts SOUL into whichever agent cli you are using, so al…
I built a local multi-account toolkit for Codex because logging in/out and restarting sessions kept getting old (www.reddit.com) One thing that kept annoying me in Codex was that multi-account use still felt clunky in practice. I was ending up in a loop of auth switching, session restarts, and runtime weirdness.
My experience with Claude and Codex on a system architecture bug (swaranga.dev via hn) Recently, I encountered a subtle bug in an event-driven system. Looking at the symptoms, the immediate defect looked clear to me but of late, for most bugs, I tend to rubber-duck it with an AI model before I make the fix.
Providing these 3 resources instantly improved my agents (www.reddit.com) Have been running Claude Code and Codex heavily for both coding and non-technical work, but started looking for new solutions as my work scaled and my markdown docs and skill directories were bloating. I wanted better agent persona/skill o…
Show HN: Haindy, a CLI that gives coding agents computer use (github.com via hn) I built HAINDY, a CLI that gives coding agents computer use across desktop, Android, and iOS. You install it as a normal CLI tool and can opt to install skills for it on Claude Code, Codex and OpenCode.
Complex, parallel, long-running claude/agentic sessions - what is the point? where is the value? (www.reddit.com) Here is how I view AI Agents field (with focus on SWE/research) right now: - "chats online" gpt/gemini/claude --> general use - "vscode like extensions" cursor/antigravity/cline vs code extension/cc vs code extension etc. --> for coding, b…
Agtop: Btop but for Your Agents (github.com via hn) agtop Run with: npx @ldegio/agtop Your window into what your AI coding agents are doing, sitting in the terminal, where you run them. agtop is a top-style terminal dashboard that tracks every Claude Code and Codex session on your machine:…
Any magic prompt that Local LLM never turning back until everything completed? (building frontend application with qwen3.5-35b-a3b) (www.reddit.com) https://nestia.io/articles/well-designed-backend-fully-automated-frontend-development.html Trying to generate entire frontend application from well-designed contexts. Succeeded to fully implement frontend application just by one-shot promp…
Your AI coding tool doesn’t know what version it just used (www.reddit.com) I’ve been using tools like Cursor, Claude Code, and OpenAI Codex to build projects from scratch. One thing I noticed.
Show HN: Financial Web Tools – decision-focused calculators (financialwebtools.com via hn) I built a set of lightweight financial decision tools (buy vs rent, house hacking, retirement, debt payoff) after trying to model a condo purchase for my girlfriend. Backstory: My girlfriend was looking into buying a condo, and I was tryin…
I built Fixy Code — a multi-agent coding terminal built with Claude Code (www.reddit.com) Built this with Claude Code. Free to try.
I've been building Nest by RAVEN with Claude Code for the past few months. Claude has been part of the process from day one — and it ended up being one of the core AIs the product is built around. (www.reddit.com) Nest is a desktop workspace (Mac + Windows) that runs multiple AI CLIs in a resizable grid. Each pane is a fully independent session with its own account, history, and environment.
Built an open source IDE for running parallel AI coding agents. would love feedback. (www.reddit.com) We kept running into the same problem: AI agents are fast enough to handle 10 things at once, but there's no good way to actually run them in parallel without everything turning into a mess of terminals and merge conflicts. So we built Wor…
I built a custom skill to stop AI coding workflows from wasting so many tokens (www.reddit.com) Hey all — first time posting here 👋 I’ve been playing a lot with Claude Code / Codex-style workflows lately, and one thing kept bothering me: my tokens and quota lasts less than my daily coffe. Especially when: running long test suites tai…
Turned Anthropic's Harness article into a working Claude Code plugin (www.reddit.com) I've been running Claude Code and Codex side by side manually for months — same prompt to both, copy-pasting findings between them, iterating until they agreed. It worked well (the two models genuinely catch different things), but every ha…
Built tier.love – a tool for rating Claude and others from the web or CLI (www.reddit.com) Been on a forced break from other projects (partly due to lack of opus performance) and decided to ship something small while experimenting with different models. So, I built tier.love – a site where you can vote on AI coding tools and see…
Extracted System Prompts from ChatGPT, Claude, Gemini, Grok, Perplexity and More (github.com via hn) System Prompts Leaks Extracted system prompts, system messages, and developer instructions from popular AI chatbots and coding assistants — ChatGPT (GPT-5.4, GPT-5.3, Codex), Claude (Opus 4.6, Sonnet 4.6, Claude Code), Gemini (3.1 Pro, 3 F…
Sharing a starter kit for persistent repo knowledge across AI agents (www.reddit.com) Over the past month I’ve tried just about every agent and AI coding tool I could get my hands on: OpenClaw, Hermes, Kilo Code, Codex, Cursor, and more. Most of them have some kind of memory, but I wanted something persistent, repo-level, a…
I ran Gemma 4 as a local model in Codex CLI (medium.com via hn) I ran Gemma 4 as a local model in Codex CLI | by Daniel Vaughan | Google Cloud - Community | Apr, 2026 | Medium Sitemap Open in app Sign up Sign in Get app Write Search Sign up Sign in Google Cloud - Community · A collection of technical a…
Local coding agents. Am I missing something? (www.reddit.com) I'm an experienced software dev that has been using various LLMs and tools to write code in the past few years. My hardware isn't the greatest for AI with a 4070ti and 64gb ddr5 but I can run a few smaller models.
Your agent is lying to you… (www.reddit.com) Is your agent actually doing what it’s supposed to do? Or just returning outputs that look correct?
Run 2+ AI agents at once to QA your workflows (Claude/Antigravity/Codex) — agent-handshake, an open-source automation protocol ( via reddit) could not extract summary
$100 Claude Max & $100 Codex or $200 Claude Max (www.reddit.com via reddit) Curious. If you had a budget of $200/mo.
Claude Gmail / Notion connectors still bad (www.reddit.com via reddit) I don't understand how Claude can hope to compete in office productivity when Gmail connector cannot read attachments and is unable to read long threads. Does anybody have a solution or am i forced to use codex?
I use CC $200 plan, decided to also try ChatGPT $100 plan. Data inside (www.reddit.com via reddit) A lot of people are talking about usage on Codex because OpenAI seems to have reduced it, but most comments seem to be vibes based rather than just pulling token #s and comparing. I figured I had the data, so would share mine.
What do you feel when you talk to agents? (www.reddit.com via reddit) Hi! I'm a software engineer, and these days I spend a lot of my time talking to AI agents.
Built a Herdr plugin that finds any Claude Code, Codex, Gemini or OpenCode session by a word you remember and resumes it (www.reddit.com via reddit) I run Claude Code, Codex and a few other agents in Herdr and couldn't find old sessions. I'd remember one word from the conversation, not the project or the day So I built herdr-transcripts.
Antigravity vs. Claude Code vs. Codex: How do the rate limits actually feel in practice? (www.reddit.com via reddit) Hey everyone, I’ve been using Claude Code, Cursor, and Codex for a few months now (with standard $20/mo Claude Pro and ChatGPT Plus subs). Here is my subjective experience so far: Codex (ChatGPT Plus): I seem to burn through the limits pre…
Do enterprise seat reseller exist for Codex/Claude/Cursor/Grok? ( via reddit) could not extract summary
I ran Claude Code and Codex in parallel for 15 days. Here's what I found. (www.reddit.com via reddit) I've been running Claude code and Codex in parallel daily like Claude Code on the frontend and codex on the backend. The problem was keeping the agents in parallel like every switch meant reexplaining context, reuploading files and to mant…
Delegation to codex within CC (www.reddit.com via reddit) I’ve read that you can ask Fable or another strong model to delegate work to agents running Codex within Claude code I use Claude code within the macOS instance. Can anybody help me understand how this would be possible?
do you need this in your vibecoding life? (www.reddit.com via reddit) I’m considering building a tool that helps developers understand the important code that AI writes for them. AI tools like Cursor, Claude Code, and Codex can generate software very quickly, but sometimes developers end up with code they ca…
Built an iPhone app with Claude because I kept forgetting people’s names (www.heythanksbud.com via reddit) I’ve been building a small but real project called Thanks Bud (www.heythanksbud.com). It helps you remember names and the details around them in the seconds before you say hello.
Is there any point in keeping my legacy Cursor subscription? (www.reddit.com via reddit) I'm still on the legacy Cursor subscription ($20/month for 500 requests/month), but lately I've been using Cursor only as an IDE because I barely write code by hand anymore. I use Codex as my coding agent.
How are you guys keeping context in sync when everyone’s using different AI coding agents? (www.reddit.com via reddit) Curious how other teams are handling this. Say your team agrees on the requirements, files, and main instructions for a project.
I asked Claude Code to pick the "best" code review tool and it chose itself 76% of the time (armature.tech via reddit) I've been running some huge agent discoverability experiments lately just to see how different coding agents pick third party tools. I ran over 7,800 judged runs so the stats are actually significant, and the data is pretty wild.
How are Pro's usage limit in regard to Codex and Claude? (www.reddit.com via reddit) Hello ! Just wondering.
I could never find the Claude Code chat where I solved something. Built a tool that searches all of them across projects and machines. (www.reddit.comhttps) claude --resume only lists sessions from the directory you launched it in. Codex's picker does the same.
Codex vs Claude Code vs Cursor in September 2026 - what are you actually using now? (www.reddit.com via reddit) I’ve been using Codex since July. Before that I used Claude Code and Cursor.
Interesting insight on what "ultracode" does and the core benefits of having multi agents when coding. (www.reddit.comhttps) I had a totally different understanding of how the throttle works in claude code. I thought choosing ultracode would engage more active parameters during inference or something similar - basically increasing the inherent quality of the bas…
I built a tiny desktop quota display for Claude and Codex (www.reddit.com via reddit) I made this little ESP32-S3 quota display with Claude Code. It shows Claude and Codex usage and reset times, supports multiple accounts, and rotates between them every few seconds.
Introducing COPE agent for Claude & codex, Free &. Open source. (www.reddit.comhttps) Building COPE: an open-source agent for Claude or Codex that turns a rough note (or blog URL) into ready-to-post content — LinkedIn post, X thread, carousel PDF, Instagram images. Create Once, Post Everywhere.
Red-Teaming Auto Mode: Improving Blocking Classifiers Against Malign Coding Agents (arxiv.org) To keep coding agents from going off the rails, production systems now review each proposed action with a blocking monitor that can reject it before it runs (Auto Mode in Claude Code, Guardian in OpenAI's Codex). Prior evaluations of such…
Got cursor start just to see if cursor can fit my workflow. But discovered no auto mode? (www.reddit.com via reddit) Cursor users, help me figure out which plan/model makes sense for me. On the Cursor Start plan, I’m currently only seeing: Cursor Grok 4.6 Composer 2.5 Cursor Grok 4.5 The thing is, I already have X Premium+, which gives me access to these…
How do you tell if Claude agrees with your idea because it’s good? (www.reddit.com via reddit) I use Claude a lot for coding and architecture, and the thing I find hardest to judge is whether my approach is actually good or whether I’ve just put it in the prompt. If I explain my solution and ask for alternatives, it usually comes ba…
A year of Claude Code alongside Hermes and Codex: the problem isn't the coding, it's the 12 tasks every session leaves behind (www.reddit.com via reddit) I've used Claude Code more or less daily for a year, the last six months seriously, next to Hermes (always-on, Telegram) and Codex (second opinion). Architect by trade, .NET mostly.
[ Agenteq ] A single canoncial source for AI coding agents rules (www.reddit.comhttps) Cluttered repositories are kind of the norm nowadays. Every AI coding agent adds its own files, such as AGENTS.md for Codex (and other universal agents), a .claude folder for Claude Code, or a .cursor folder for Cursor.
Sometimes you have to treat Opus like its a teenager ... I'm not surprised, but I am dissapointed (www.reddit.com via reddit) I usually use Codex to keep Opus in line, and I thought, nah, I can jockey this pony, and it went well to begin with until the first agent started getting a little long in the tooth, so we agreed to part ways and call it quits, and it gave…
I built claude-sandbox, a free, open-source (Apache 2.0) sandbox for coding agents using claude. (www.reddit.com via reddit) I built claude-sandbox, a free, open-source (Apache 2.0) sandbox for coding agents. It runs Claude Code, Codex and Pi.
People who own both Cursor and Claude/Codex plans (www.reddit.com via reddit) Especially if you worked on the same projects - how does it feel? What's your most effective model in Cursor?
Should I use Codex or Claude to learn AI Engineering concepts for my medicated ADHD brain (www.reddit.com via reddit) Hi, software engineer here. I have never used codex and was using claude 5x plan from past 5 months to learn about AI engineering concepts and Python.
I just tried astra. I ran out of usage for the week in 3 hours (www.reddit.com via reddit) To people complaining about Claude use - maybe I’m doing it wrong with astra, but fable usage seems way better than astra usage This was codex $100 plan
Where do you host apps you built with ChatGPT/Codex once they start getting real traffic? (www.reddit.com via reddit) I’ve built a few web apps mostly with ChatGPT/Codex helping with the code, and so far everything has been running on a regular VPS. One of them is starting to get more traffic though, and during heavier workloads I’m noticing CPU usage bec…
Time to drop to 5x (www.reddit.com via reddit) https://preview.redd.it/e6qlhhy6mwph1.png?width=2032&format=png&auto=webp&s=b1fcfd5dc6ad87c0d56cc82eaf8db246dd66ddb6 haven't gone past 50% of my 20x weekly limit in a while, i used to code only with Fable but now, i kinda feel opus+extra t…
Codex has a weird "Attach Claude" Option where it can basically see data from each other?? (www.reddit.com via reddit) https://preview.redd.it/8fb0qxt7xsph1.png?width=1692&format=png&auto=webp&s=7a8412380b8832c08fb9d73ce2332b2baeee2cd9 This is so weird, anyone here uses it?
Claude is getting sassy with me. (www.reddit.comhttps) Building a browser game and there's an issue with the hosting logic, I asked Claude to fix something else while codex is working on that and this is what claude said. Can't say that didn't hurt a little.
Claude Code + Meta Quest 3: I have created the ultimate VR office (www.reddit.comhttps) Hey everyone, I think I really went overboard this time. Over the past two weeks, I’ve burned through two Claude Max 20x subscriptions and two ChatGPT Plus 20x subscriptions to build a 3D Iron Man office for Meta Quest, with Claude Code se…
Cursor + Claude Code + Codex people: how do you keep one memory across the three? (what worked for us, and where Cursor is the odd one out) (www.reddit.com via reddit) Edit, corrected: u/lgmarian pointed out that Cursor does have hooks (cursor.com/docs/hooks). I had that wrong.
I left one Claude run alive for 70 hours. Here’s what actually happened. (www.reddit.comhttps) I’ve been experimenting with a slightly different way of using Claude Code: instead of treating every piece of work as a new session, I let one persistent run stay responsible for the work and spawn smaller workers underneath it. This one…
Claude vs OpenAI weekly quota (www.reddit.com via reddit) I keep hearing "people" here saying how amazing the usage limits are on Codex, but really, have you used both recently? I have the $100 tier on both, Team Premium vs Business Premium, and OpenAI side has moved from extremely generous ($20…
Stop using Auto, "...you lazy animal" (www.reddit.com via reddit) Pretty sick of hearing everyone bashing Cursor cos they use Auto and now its all "gone to shit". How about, Not Auto?
"General harms" (www.reddit.comhttps) I've tried getting around this 5-6 times, sometime it starts the task before stopping in the middle of it. I cannot understand why this would be flagged time and time again.
Agentic AI for Gravitational Wave Data Analysis: A Head-to-Head Comparison of Coding Agents Executing a Matched Filter Pipeline on Einstein Telescope Simulated Data (arxiv.org) We report a methodological study of agentic AI in gravitational-wave data analysis: two systems, Claude Code (Anthropic) and Codex (OpenAI), autonomously executed the same simple end-to-end pipeline on Einstein Telescope (ET) simulated dat…
An update on the Unity game I built with Claude, and the project rules I use now (www.reddit.comhttps) Want to understand what's working best for you all for prod projects (www.reddit.com via reddit) Hi all, I've used cursor extensively for several projects including rust based tuis, deep learning projects, and most recently an android game. I've been using plan mode extensively.
Which model should I use for different parts of a small MMO? (www.reddit.com via reddit) How has AI (Claude Code, Codex, Cursor, etc.) completely rewritten your software delivery workflow? (www.reddit.com via reddit) I’ve been thinking a lot about how tools like Claude Code, Cursor, Devin-style agents, and similar AI coding assistants are quietly (or not so quietly) rewriting the software delivery lifecycle. Traditional agile rituals sprint planning, b…
The best advice I got here was to set up a Hub session. I developed it further so that it stays the project manager but Codex talks with me. (www.reddit.com via reddit) My hub session is the main communication hub for my sessions. It has an inbox, so I don't see all the messages in the chat.
Built a cool way to visualize your Claude Code / Codex history (www.reddit.comhttps) I use Claude Code a lot, but /stats never answered the question I actually cared about What did I build, and where did the work get difficult? So I built Bough.
I built a motion studio for coding agents. Here’s a 18-second Notion film made with it (www.reddit.comhttps) I’m building Motioneer! a local motion studio that you can direct through Claude Code or Codex.
My Max 20x ends today, as Claude code gets usagenerfed. With Code nerfed, will codex 20x usage be way ahead? (www.reddit.com via reddit) Not sure whether to resub to Max 20x or give Codex 20x a go for a month. Whats you guys thoughts?
Perfect way to compare the two (www.reddit.comhttps) Nothing groundbreaking, not a benchmark, just a simple everyday routine. I work with both Codex and ClaudeCode, usually Codex is the workhorse and Claude orchestrates, but lately with Astra I switched them, but Opus still vibe checks the o…
A skill that does your app store research (www.reddit.com via reddit) Built a skill called app store research. it helps you validate an idea before building it: find the top competitors, read their reviews, and see what's actually missing so you can shape your idea around that.
I built a tool to search my agent logs. Across 3000+ sessions and 30 billion tokens, only 0.3% of that was the model actually writing anything (www.reddit.comhttps) Every Claude and Codex session leaves valuable data behind on your machine so I finally figured out how to make use out of it. Both tools write every session to a disk as JSONL containing per-response token count, every tool call and its r…
I run my whole AI setup from my phone now. Here's the workflow (www.reddit.com via reddit) I used to spend my days copy-pasting context between ChatGPT, Claude, Codex, and whatever else. Explain the project again.
↯ Model Context Protocolmodel-context-protocolchatgptcodex+2
[Investigation] The "Unlimited Compute" Scam: Wire-Level Proof of Model Spoofing, Dangerous Setup Scripts, and Packet Analysis of CodexAPI.pro (www.reddit.com via reddit) I bought credits on codexapi.pro (https://codexapi.pro/) after seeing their promos for cheap "unlimited" coding sessions with Claude Code and Codex CLI. In practice, the service was constantly dropping connections: 502 Bad Gateway errors s…
OpenAI is building Codex Replay, a tool that invites Claude Code users to put Codex head-to-head on their own work — rerunning imported tasks and comparing the results (runtimewire.com via reddit) OpenAI wants a head-to-head with Claude Code—using your own work Build 8881 targets successful Claude Code importers with a plugin that reruns an imported task in Codex for a direct comparison. By Ryan Merket · Published RUNTIMEWIRE INVEST…
So yeah I made yet another menu bar app to keep track of Claude Code and Codex usage (www.reddit.com via reddit) I call it Delta-V, a small macOS app that shows your Claude Code and Codex subscription usage in the menu bar. You can show either provider or both side by side, choose which windows to display, and switch between percentages remaining or…
Turned my SOL --> LUNA workflow into a repo-native Codex methodology (www.reddit.com via reddit) I've been using Codex pretty heavily for real web application work and over time I noticed that I was getting much better results when I stopped treating SOL and LUNA like interchangeable coding models. SOL is obviously much better when th…
Codex vs Cursor (www.reddit.com via reddit) Hey guys, I've been loving using Cursor and it always felt the best before Grok came to it. I've got Ultra and now with OpenAI restricting their models being used in Cursor, I'm thinking to switching over to Codex or some other way to use…
Claude Code vs Codex vs Cursor,what are you sticking with and why? (www.reddit.com via reddit) Hey everyone, Our team is currently trying to decide which AI coding setup to standardize on, and I’d love to hear from people who have actually used Claude Code, Codex, and Cursor heavily in production. For the last 3–4 months, we’ve been…
I build a payroll saas with CC. Fable xhigh plans, Opus xhigh executes. With the Max 20x usage cut coming I tried putting OpenAI models via Codex CLI into the loop, here are my conclusions. (www.reddit.com via reddit) I've been using CCode since August 2025 and Opus xhigh is my workhorse. I'm a lawyer with payroll domain knowledge but I have some coding background and some good instincts.
Roblox Studio won’t connect to ChatGPT —Codex Computer Use detects zero Windows apps (www.reddit.com via reddit) Codex Computer Use MCP loads, but detects zero Windows apps — can’t connect to Roblox Studio I installed/enabled the Computer Use plugin and its MCP server + skill in ChatGPT Desktop/Codex. Roblox Studio is open, but when I ask Codex to co…
Had some issues with jumping from claude to codex (www.reddit.com via reddit) Since boosted 50% are coming to close now in cc and given just how good codex has been performing in recent times I have decided to switch to openai but I have some issues I am facing and it'd be great if someone can please help me out her…
Six code reviews said "request changes". My board recorded six approvals. The model was right every time. (www.reddit.com via reddit) I run Claude Code, Codex and Copilot as seats on a shared task board, and one of the seats is a reviewer. The contract with the reviewer is simple: write your review, then put your verdict on the last line, approve or request changes.
I open-sourced a desktop app that wires Claude Code to roblox-ts + Rojo so you can describe a game and watch it land in Studio (www.reddit.comhttps) I'm the author of Blockforge an desktop app (Windows + macOS) that sits around tools you already use : Claude Code (or Codex / OpenCode), roblox-ts, Rojo, and Roblox Studio. You create a project from a roblox-ts template.
I wanted my chats to talk to my chats so I made a chat (macOS) (www.reddit.com via reddit) For months I've been using a CC/Codex plugin (agent-talk) to use GPT as an adversarial reviewer for Claude. Then I recently switched to herdr -- very cool terminal session server/multiplexer.
Building a D&D world simulation using Claude where NPCs actually react to what you do (www.reddit.comhttps) Building a fantasy RPG where the world doesn’t just wait for scripted quest triggers. NPCs have goals, beliefs and relationships, and the game keeps track of what happens to them.
Claude or GPT for programming: what makes you stick with one? (www.reddit.com via reddit) I’m currently paying $20/month for Claude Pro, but I’m thinking about switching to Codex. The more I use it, the more it feels like it fits the way I work and think as a programmer.
I turned Claude’s usage limits into little potion flasks for my Mac (www.reddit.comhttps) I built Questis to see my Claude quota and reset times without repeatedly opening the usage page. It lives in the Mac menu bar, with a full display where little glass flasks drain as the quota gets used.
Help! Need feedback, built a way to visualize your Codex history (www.reddit.comhttps) I built this for Claude Code first, because that's what I use day to day. /stats told me how much I'd used it, never the thing I actually wanted to know: What did I build, and where did the work get hard?
What could you actually ship in 24 hours with Cursor if nobody gave you a problem? (www.reddit.com via reddit) I've been thinking about this while putting together Grevix AfterCode. Most hackathons start with: Here's the problem.
What to do with bad code created by Codex? (www.reddit.com via reddit) I am working with Claude and it does the job. But when I am asking Codex (sol/astra) to do some complex task, its creating bad code.
GLM 5.3 Flash vs Kimi K3 for heavy coding — which subscription would you choose? (www.reddit.com via reddit) I'm planning to use AI seriously for coding, roughly 80% GLM 5.3 Flash and 20% Kimi K3 for harder tasks. I mainly care about large projects, debugging, refactoring, agentic coding and value for money.
This is probably a dumb question (www.reddit.com via reddit) But I am coming from Claude and I liked that I could use Claude Code on the web, on the desktop app, or on the app, however I am having a hard time figuring out how to use Codex on my mobile app. I typically would just do basic cloud codin…
I kept losing track of which Claude session was waiting on me, so I built a board that watches all of them (www.reddit.comhttps) I'm often working across Claude Code and Codex with multiple sessions across multiple projects simultaneously. There's a big mental tax of tracking what session is doing what and which one is waiting for me.
I build an interactive artifact about 9/11 (www.reddit.com via reddit) So, on the 25th anniversary of 9/11 I created a researched interactive artifact using Codex and Claude on my app. Claude models were used for planning and narrative and animation/mapping, Codex for the others.
Yes, another agentic knowledge base. But this one configures Claude Code for you (skills, subagents, hooks) and shares it with your team (www.reddit.com via reddit) I know, there are dozens of these by now: memory servers, second brains, Karpathy-style LLM wikis. Cartographer started out as one of them.
Ajuda sobre o cursor plano 3x (www.reddit.com via reddit) Oi pessoal, estou trabalhando em um projeto de desenvolvimento e venho usando tanto o Codex quanto o Claude Code ambos serviços pagos até agora. No entanto, acho o preço do Cursor bastante atraente, especialmente o plano "3x", então estou…
How do you share AI coding agent context between developers on the same project? (www.reddit.com via reddit) I work at a small company where most projects currently have just 1-2 developers, each often using an AI coding agent (Claude Code, Codex, etc.) pretty heavily. As we grow and more developers start working on the same codebase, I'm running…
What is your monthly budget for agentic coding ? (www.reddit.com via reddit) I use codex. I use a thorough workflow research>spec>plan>execute>test>review&fix workflow.
I gave Claude Code a whiteboard. My notes are a plain markdown vault it can read and organize visually, every edit attributed and undoable (14 day free trial, then $59 once) (www.reddit.comhttps) Acres is a local notes app. Your notes are markdown files in a folder you choose and you lay them out as cards on an infinite canvas.
Two-person team using Claude Code + Codex on a live platform. What would you improve about our workflow? (www.reddit.com via reddit) I'm two weeks into working on a live platform with my supervisor. We're the only two developers, with around two dozen daily visitors and a growing feature backlog.
fable 5.1 on pro+ (www.reddit.com via reddit) Thinking about trying Cursor Pro+ mainly for Fable 5.1. For anyone using Fable heavily on the $60 plan: how long does the monthly quota realistically last if you use high/extra-high reasoning for serious coding work?
How do you guys get Claude to actually finalize a decision or code that has little to no flaws consistently? (www.reddit.com via reddit) I know this is a common problem, so this is more of me asking what tools are out there or what can I use/do for someone who isn't very well versed with the ways I can improve it other than asking Claude itself. Every time I code or plan, C…
Claude build itself a project management system to keep track of subagents (www.reddit.com via reddit) I asked Claude why my sessions were so expensive, and it said it's because the models are expensive and should only be used for important tasks. I asked it what the best thing we could build to give itself maximum leverage to command a swa…
Probably in the minority here, but I actually feel like Claude doesn't give *enough* praise when its legitimately earned (www.reddit.com via reddit) I know. I was here for the sycophancy updates and token wasting threads about Claude (and chatgpt) being too buddy buddy.
Built with Claude Code: a free menu bar app that keeps your 5-hour and weekly quota in view (www.reddit.comhttps) On August 19 I had 27 Claude Code sessions going, 5,119 messages in total. I had no idea how much quota that was burning until it stopped.
Recreating my favorite game with Claude (www.reddit.com via reddit) Hey everyone! A friend encouraged me to make this post and share what I’ve been working on for the past several months.
Claude is still the best value for your money (www.reddit.com via reddit) So my comparison per the rules is hedged only on my personal experience. I'm a senior university student who has been trying out different AI models and seeing its effectiveness on some of the similar tasks that I am most likely to perform.
How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules (openai.com) could not extract summary
We measured whether our 20 skills actually fire. Baseline recall was 46%, and our first detector only understood Claude's Skill tool. (www.reddit.com via reddit) We ship about 20 Claude Code skills. Progressive disclosure means the agent picks one from a single line of description, so an excellent skill nobody picks up is worth zero.
Based on your real review would you prefer fable + opus or sol+ astra for real solid projects? (www.reddit.com via reddit) Currently I use astra, it's good for me but consumes too much limits that can consume all 200usd plan weekly limts in about 4 days of 8 or 10h a day And I hear that fable 5.1 and opus 5 priduce better quality too Also codex remote has the…
Fable 5.1 6x cheaper for me than Astra? Unexpected (www.reddit.com via reddit) So I use both OpenAI and Anthropic models in GitHub Copilot. I noticed Astra consuming more AI credits than usual.
Why doesn't Computer Use work at all? (www.reddit.com via reddit) I'm very new to codex and I am trying to get it to recognize any desktop app that is open on my computer. It claims: "“Windows Codex Computer Use has Any App enabled, but desktop inventory returns apps: [] / Trusted RPC service is not conf…
Codex vs OMP harness (www.reddit.com via reddit) Hi everyone, You might find this funny, but I actually have the opposite problem to y’all. My usage allowance with Astra Max feels so generous that I’m starting to wonder if the model is running at half power or something, lol.
Use Claude to turn your chats into mini blog posts (www.reddit.comhttps) I built a free tool that uses your own agent to read your chat history in Codex, Cursor, or Claude Code and generate mini blog posts as markdown files. Your chats aren't shared.
I was having trouble managing a lot of agents with just the terminal, so I built a "terminal development environment" (www.reddit.com via reddit) https://preview.redd.it/xtlqwquf1joh1.png?width=1600&format=png&auto=webp&s=bedc37ac9e13dedc3284d40395fed0ee8b8a0f58 * It's a lightweight wrapper around tmux. If you close or uninstall the app, your sessions are still there.
Voice coding in 2026: how SKI works (www.reddit.comhttps) We built SKI (heyski.io) as a voice coding layer for AI agents. The thing that got us started: every time a task finishes, you're either waiting on some notification chime, or you've tabbed away and forgotten to check back.
ccusage no longer tracking claude? (www.reddit.com via reddit) https://preview.redd.it/vcggbj00ohoh1.png?width=1080&format=png&auto=webp&s=819ad5956415d4643c295d9fbb2e0921355fba7e my ccusage no longer logs claude usage after july.. anyone know a fix or why this is happening?
I’m confused about the Grok and Cursor plans. (www.reddit.com via reddit) I use Grok mainly for VDO work, with some vibe coding too. But when I want to use it like an IDE, I don’t have tools such as Codex, Claude Cowork, or Antigravity, so I have to use Grok Build.
What I learned building a local macOS MCP for ChatGPT-assisted coding/admin work (www.reddit.com via reddit) I kept running into the same workflow problem: ChatGPT could tell me exactly what command to run, what file to edit, or what to click in Safari, but I still had to leave the chat and do the mechanical part myself. Closest alternatives I tr…
Is claude co-work actually useful and what do you guys use it for?? (www.reddit.com via reddit) Some context: I am a developer and I juggle around with coding tools like claude code, codex, cursor and everything. I just use the normal claude chat and claude projects for any plannings, roadmaps and reviews.
Codex stuck on commands – anyone else? (www.reddit.com via reddit) I'd like to ask everyone: when using Codex, I often encounter a situation where it gets stuck on a single command for a long time with no progress (during this time, the remaining quota doesn't change). I tried asking Codex to diagnose thi…
Can an AI coding agent be locked out of modifying its own guardrail hooks? (OpenAI Codex CLI) (www.reddit.com via reddit) Goal I run AI coding agents locally on Windows and want a "hardstop" I can trigger at any time - a single keystroke that immediately blocks the agent from doing anything further until I clear it. I have this working for one agent as a User…
Effort level in claude design (and its alternatives) (www.reddit.com via reddit) For those of you who use claude design: which effort level do you usually use in what cases. I was experimenting and find fable-high/medium but I think I got too triggered happy.
Astra is a fantastic for coding (www.reddit.com via reddit) I've been using both claude code and codex to work on my app, and Astra has been killing it. I can point it at my codebase and simply say "find bugs" and it finds things that other models missed, even cross-component issues that are harder…
Introducing AstraBlender! Real Blender that ChatGPT can use from a simple prompt sent from your phone on the ChatGPT website ;) (www.reddit.comhttps) Simply prompt ChatGPT work (or any other agent with a cloud browser, like Grok Bot) to go to the website and use blender. From your fucking phone!
Claude Artifacts only live on Claude, and I lose them every time I switch agents (www.reddit.com via reddit) I've been using Claude Artifacts for a while and they're one of the most useful features it has. I use them for reports, presentations, and sometimes even interactive prototypes.
I started using Codex to learn DevOps… now I’m wondering what exactly I’m learning 😂 (www.reddit.com via reddit) I’m a QA engineer trying to move into DevOps, so I started building my own project to get hands-on experience with Git, TypeScript, Playwright, testing, CI/CD, and architecture. Today I decided to try Codex.
A handoff checklist for switching between Claude Code and Codex (www.reddit.com via reddit) The challenge with having multiple coding assistants is that the information gets fragmented. If you are in the middle of a discussion, both can understand the repository, but the decision from a conversation thread may not make it to the…
Four months of running a company from a folder of Markdown files, operated by Claude Code. No Obsidian, no graph, no layers. What held and what I paid for and never used. (www.reddit.com via reddit) I am a solo founder in Germany. No coding background.
Does Claude Code actually work well with non-Claude models (Gemini, GPT, Llama, etc.), or is it heavily optimized only for Claude models: Opus, Fable ? (www.reddit.com via reddit) I have a strong suspicion that Claude Code is deeply optimized for Anthropic’s own models (especially Opus and Fable) and that performance drops noticeably when you try to run it with other LLMs. Has anyone actually tested this properly?
Fable 5.1 vs GPT-6 Astra for 2D Sprites (www.reddit.comhttps) Using the same simple prompt the models took very different approaches: Astra delivered one sheet with 16 key poses; Fable delivered 992 frames across four palettes, plus a Python generator and browser preview. Codex CLI with GPT-5.6 Astra…
Entire week's fable quota is burnt in a day :-| (www.reddit.com via reddit) https://preview.redd.it/mo5y00dv9aoh1.png?width=2020&format=png&auto=webp&s=95f0c39aa0470014e6c98365fdb6ce9037bf9c98 This is $200 Max plan. I think I need to buy 6 more of these now to survive a week!
I built an open-source tool to carry context between Claude Code, Codex, Cursor, and other coding agents (www.reddit.com via reddit) I switch between coding agents a lot. Re-explaining a task gets old, especially when the useful context is buried in another tool's chat history.
I would imagine openai will be patching some of the more insane things that can be done with Astra soon (a.k.a using other game studio assets) (www.reddit.com via reddit) I've seen people say this is so hard to reign in and I don't really get that. Right now codex is using other companies assets.
Is there a way to integrate and use Codex or Claude Code? (www.reddit.com via reddit) I am curious about the difference between Claude Code and Codex. Personally, since my Claude Code usage often runs out, I am planning to subscribe to Codex.
How do I make websites generated by Codex look better (www.reddit.com via reddit) I'm running into some issues using Codex for front-end design, and I could use some advice. I always feel like my designs end up looking more like reports than actual website…
Codex $100 or grok $100 for langgraph/langchain development? (www.reddit.com via reddit) Codex $100 or grok $100 for langgraph/langchain development? I'm trying to decide which \\\~$100/month plan is better for heavy coding: Codex or Grok.
1Password increases engineering productivity 21% with Codex (openai.com) could not extract summary
Loop engineering: I turned the Ralph loop into a verified one. One markdown file, any agent, a critic before "done". (www.reddit.com via reddit) The Ralph loop (while :; do cat PROMPT.md | claude; done) works, but it's blind: no memory between runs, it believes the agent when it says "done", it never stops on its own, and nothing stops the agent from editing the tests that judge it…
Claude Cyber Verification vs Codex Daybreak Blue — anyone using both? (www.reddit.com via reddit) Got approved for both, and the difference in approach is interesting. Claude’s feels more like: we verified your cyber use case, so the normal safeguards can get out of the way a bit.
What is more efficient? (www.reddit.com via reddit) I'm really new to coding with AI, and I want to start a long term project where I build a cool little game in Unreal Engine 5 with the help of AI. I was wondering, is it more efficient to let AI write the code for me and then copy and past…
Connecting Claude and Codex (www.reddit.com via reddit) How to connect Codex with Claude? What’s the best way or you follow?
help me with this idk if I'm right or wrong or fully off on everything (www.reddit.com via reddit) So, let’s just say I’m using GPT-6, right? If I’m using it in normal Chat mode on the Free plan, how does the usage reset work?
Claude spent 48 hours waiting for me this week. (www.reddit.comhttps) I built AgentPlayback to see how long my Claude Code and Codex sessions were waiting on my response vs working. Turns out waiting for me is a full time job.
Building a CLI agent for Claude to manage How to design a multi-agent system where Claude delegates work? (www.reddit.com via reddit) Hey everyone, I’ve been experimenting with a terminal-based workflow where I use Codex to execute tasks and Claude to manage the overall project. Right now, I just write my project prompts or requirements directly into the terminal, and…
I tested Claude Code, Codex, Gemini, and the most popular open source models through OpenCode, and compared what each one did to what it said it did (www.reddit.com via reddit) The setup. Eight tiny repos.
Burned 1.7 billion tokens in 30 days on a $20 Plus plan — someone please tell me this isn't normal (www.reddit.comhttps) Started using Codex about a month ago. Checked my stats page today and genuinely had to look twice.
Did anyone else notice that GPT-6/codex uses (inline) Python much more aggresively? (www.reddit.com via reddit) I am testing GPT6 and noticed a few things. A bunch of problems or 'more difficult situations' often cause it to generate python scripts or inline-python or even inject it via SSH and other ways.
Running Cursor through Codex subscription(s) (www.reddit.com via reddit) You can now do it: https://docs.codex-pooler.com/clients/cursor/ Why? I dunno, cause we can I guess.
Memory Undercutting Model Performance? (www.reddit.com via reddit) I've been pretty disappointed with Fable's performance on my project recently. After looking into Claude's memory, Anthropic seems to neglect memory cleanup and management.
Computer use on cursor like codex ? (www.reddit.com via reddit) Is there any way to make computer use efficient on Cursor, like the Codex app, because on Cursor it barely controls the integrated browser right ?rather the system or the external browser If so, please let me know. Thanks .
A smaller menu bar app for Claude Code, Codex and GLM limits (www.reddit.com via reddit) I'd been building this for myself when Peter released CodexBar back in November. CodexBar does far more, 69 providers and a bundled CLI, and it's genuinely good, 21k stars and 119 releases since.
PMs using Claude Code heavily, what does your setup look like? (www.reddit.com via reddit) I’m a PM at an AI tech company and have been trying to figure out how to get the most out of Claude Code / Codex as a non-engineer who still does a small amount of building. My day-to-day is still mostly deep product work and getting good…
Manage your claude sessions via a Kanban board from phone (www.reddit.comhttps) I usually run several coding agents in parallel — mostly Claude Code, but also Codex, and a few others — on a terminal kanban board to get an overview of every coding agent state and code changes. The board is able to serve itself to a pho…
Another Astra Post (www.reddit.com via reddit) Ok, so I have been hard Claude Code from the beginning, and for context, I’m an actual software developer, not a vibe coder. I use Claude Code daily both personally and professionally.
Fable 5 built my 3D river visualisation. Here’s how Fable 5.1 and Codex each redesigned it. (www.reddit.comhttps) I’ve been building GroundwaterCast with Claude—a project that uses open data to explore groundwater and rivers in England and Wales. Fable 5 built the original 3D River Test visualisation.
I built a design explorer and MCP server for Claude Code, Codex, and OpenCode. (www.reddit.comhttps) When I ask a coding agent to “make this page look better,” the result is usually competent but generic. The problem is often the reference process.
Switching accounts in Claude Desktop hides history - made a tool to fix that (www.reddit.comhttps) I have four of the $200/mo Claude subscriptions. Unlike Codex, signing into another account in Claude Desktop gives you an empty Claude Code sidebar/history.
My Pre-Flight Prompt for all projects (www.reddit.com via reddit) This is the generic pre-flight prompt that launches EVERY project for me I have been leveraging what I call AI assisted development governance framework for some time now across a few products including Claude, OpenAI, Codex, Cursor etc. O…
I built something to stop burning through my claude tokens so fast (www.reddit.com via reddit) so basically i kept running out of my claude and codex quota way before the end of the day. Like i'd be in the middle of something at 3pm and it just stops and tells me to come back at 2am.
I built a desktop workspace for Claude Code that cuts token usage by 51% (halv.ai via reddit) I’ve been building Halv, a desktop app where you can run Claude Code alongside other coding agents using your existing subscriptions. It has chat, split terminals, saved sessions, and a live savings meter.
Is subscription worth it ? (www.reddit.com via reddit) I’m used to codex and Claude code , but want to try cursor , how is usage guys ? I don’t like brining through monthly credits casually and have maybe 20 to spend lol
My Claude Desktop session couldn't talk to my Claude Code session, so I made Yet Another Agentic Chat (www.reddit.com via reddit) For some time, I wondered: I have Claude Desktop on my computer (which can run MCP servers), where I chat and discuss ideas, which I sometimes implement in Claude Code later. I also have Claude Code, obviously.
After trying to setup Open Claw to simulate Cursor Type features, I understand how good Cursor Harness is. (www.reddit.com via reddit) Due to the lack of regional pricing of the Cursor Pro+ plan I really wanted a way to have Cloud Agents in my Pocket at a cheaper cost using Z.ai and Codex plugged into Open Claw. Not only setting up Open Claw is a pain, the interfaces like…
Cortex: local SQLite memory for Claude Code that persists across sessions. Need help with the ranking/dedup part (www.reddit.com via reddit) I built Cortex, an MIT tool that gives Claude Code (and Codex) a memory that survives between sessions. It hooks into Claude Code's session lifecycle: on stop it captures the stuff worth keeping (a fix and how it was verified, a finding, a…
Giving every AI tool the same memory (Claude, ChatGPT, Cursor, Codex, Gemini) (www.reddit.com via reddit) If you're jumping between different coding tools during the day, you'll know how frustrating it can be to solve something in Claude, then hop to ChatGPT and have it not understand a thing. The solution is to run a memory layer as an MCP se…
Best Practices with Fable 5.1? (www.reddit.com via reddit) This is the first time I’ve genuinely struggled with maxing out. I’ve updated my Claude.mds, instructed to use opus/sonnet agents, and even coordinate cross-working with codex now and im still absolutely cooked.
Apparently during this weekend OpenAI will ship Astra (GPT 6) to its customers... (www.reddit.com via reddit) Apparently even the low tier paid plans will have some Astra usage. Only in Work and Codex and with some reasoning limits (up to high).
Does Anthropic actually use only Claude Code internally, or do their engineers use Codex too? :) (www.reddit.com via reddit) Genuine question. I’d assume Claude Code is the default internally, but engineers usually use whatever helps them ship.
ChatGPT & Codex integration? (www.reddit.com via reddit) I'm not a programmer but have been using ChatGPT & Codex to help create a site that I'm making. My issue is that I spend almost all day just cutting and pasting commands back and forth between ChatGPT & Codex.
Built a Go TUI to juggle multiple Claude / Codex accounts, hot-swap quotas, and manage bot backends. (www.reddit.comhttps) Originally, I just wanted to be logged into my work and personal Claude/Codex accounts at the same time without them fighting over the same auth files. It spiraled into a full terminal cockpit called ai-session.
How to I set up remote connection iOS? (www.reddit.com via reddit) Swear I’ve looked everywhere online and it seems like it should be easy. Cannot find the setting in the iOS app or MacOS application.
Why Codex feels overwhelming for Junior Devs (and why Claude Code is saving my workflow) (www.reddit.com via reddit) As a Junior Engineer, I’ve noticed a massive difference in how OpenAI Codex and Claude Code fit into my daily workflow. Codex feels better suited for Senior or Principal devs, as it dumps huge blocks of code and tends to over-engineer simp…
Am I falling into the hype train or is Fable 5.1 really this better compared to everything else? Thinking of upgrading to x20 from x5 just for Fable (www.reddit.com via reddit) So I was a Pro user until last week I have subscriptions to all 3 American frontier labs - all base $20 until last week I’ve been working with LLMs since before Claude Code or even the VS Code extension even existed. I kind of self learned…
How do you maintain a reliable external “patient chart” that multiple LLMs can use without stale facts taking over? (www.reddit.com via reddit) How do you maintain a reliable external “patient chart” that multiple LLMs can use without stale facts taking over? I’m trying to solve a specific problem, and note I’m not a coder (unless you count Claude Code/Codex doing the work).
Auto-memory is per user. The decision is per repo. I keep mixing those up. (www.reddit.com via reddit) There's already a good post here about `~/.claude/projects/*/memory/` growing a pile of files that contradict each other. Different problem on a shared repo: even when that folder is clean, it's still *my* Claude.
How does ChatGPT + Codex ($20 Plus) hold up for daily custom frontend / WordPress dev before hitting limits? (www.reddit.com via reddit) Hey everyone, Looking for real-world feedback from devs using the $20 ChatGPT Plus plan (Codex / CLI / IDE extensions) for their daily workflow. My setup & workflow: I build custom WordPress/WooCommerce themes and clean frontend projects f…
Any tricks to getting clean tiling using tools like cursor or codex? (2d game dev on Godot) ( via reddit) could not extract summary
Claude Code Beats Codex in a Negotiation Competition (www.reddit.comhttps) People are now using Claude Code and Codex, two of the leading coding agents, to do almost everything, including tasks that have more to do with language than coding, such as negotiation. For example, OpenAI recently highlighted a use case…
Using Sol/Astra through cc (www.reddit.com via reddit) Just curious if anyone is actively doing this with Sol? I have refined my harness of hooks, skills and general rules over the last year and I am hoping to take advantage of Astra without going through the hassle of moving over to codex, ha…
Second Claude account or adding Codex to my workflow? (www.reddit.com via reddit) I`m burning my 100$ Claude account, so I'm thinking about going for a second 100$ account. What do you guys think, should I go for a Codex subscription or another Claude subscription?
Help me choose my AI provider (www.reddit.com via reddit) I'm thinking of subscribing to either ChatGPT Plus, Google AI Pro or Claude Pro. I currently have Google AI Pro on a student offer that expires in 2 weeks.
Sonnet is better than Sol at frontend ui work (www.reddit.com via reddit) People have been trashing Sonnet, but in my experience using codex, Sol their top tier model, needs way more handholding to deliver a good ui, while sonnet 5 can just one shot it....
Codex hooks give you no exit status for a shell command. Here is how I record it anyway and block "done" on stale test results. (www.reddit.com via reddit) While adding Codex support to a tool I wrote, I found that the PostToolUse payload for a shell command is just the raw output. A command that exits 3 looks exactly like one that exits 0.
Build and open ai agent (www.reddit.com via reddit) I need some help with building an open AI agent. I want to bypass the subscriptions for chat gpt.
Anyone know a good open-source Codex orchestrator? Looking for something built around Codex CLI/SDK with multi-agent routing, parallel tasks, retries, project folders, diffs and terminal output. Ideally extendable to Sol → Terra → Luna workflows. Any repos worth checking out? (www.reddit.com via reddit) Looking for something built around Codex CLI/SDK with multi-agent routing, parallel tasks, retries, project folders, diffs and terminal output. Ideally extendable to Sol → Terra → Luna workflows.
Prompties - scripting with natural language (www.reddit.comhttps) I thought that skills are nice but sometimes you want to blur the boundaries between scripts and prompts a bit. So I made Prompties - a Go interpreter for codex (add a PR for claude or wait until I add it).
fear of downgrading (www.reddit.com via reddit) In addtion to my workplace use, I've personally been a Max subscriber for a long time. I'm slowly moving workflows to Codex and want to downgrade Claude to Pro.
We made Claude Code multiplayer: two people, two local agents, one live document. (www.reddit.com via reddit) My co-founder and I build Nimbalyst, an open-source desktop visual workspace for Claude Code (Codex and OpenCode work too, Grok and Gemini are in alpha). Of course, we built much of it with Claude Code!
Codex can now build a working internal tool in minutes without writing any application code. (github.com via reddit) tooljet-mcp An MCP server that lets a coding agent (Codex, Claude Code, …) build and maintain ToolJet apps through ToolJet's governed APIs: workspaces, apps/pages, datasource queries, ToolJet DB tables, components, layouts, and lifecycle e…
I have become too dumb to use claude? (www.reddit.com via reddit) I only have the 20€ subscrption for claude and mainly use codex, where I have the 200€ subscription. However, sind Sol acts weirdly since yesterday I dont wanna use it for my current projects since i fear it may do something unwanted.
I raced Claude Code against Codex on the same task with blind cross-judging. Codex fixed a bug in my build script and lost on a rule. (www.reddit.com via reddit) I got tired of arguing about which coding agent is better, so I built a small open-source tool that settles it per task, per repo, and ran it on itself. What it does: one command creates two git worktrees at the same commit, runs Claude Co…
Using Cursor and Codex together in the same repository (www.reddit.com via reddit) Has anyone had success using Cursor and Codex in the same repository? My plan is to use Cursor as the primary implementation agent while keeping Codex in a read-only advisory role.
To everyone complaining about usage... (www.reddit.com via reddit) This may be obvious, but for those who don't know... the longer you run a session, the more tokens you will use.
Presenting my dumbest idea yet. The Claw’deck. (www.reddit.com via reddit) I decided I wanted a touch screen for my agents. If an agent asks a question, it can pop up on the screen and I can tap an answer.
Cursor custom subagents keep ignoring the configured model (www.reddit.comhttps) Trying to force local subagents to use GPT-5.6 Luna/Terra, but they keep spawning as GPT-5.6 Sol High. I’ve tried: custom .cursor/agents/*.md model configs bare / High / XHigh variants setting the built-in Explore subagent to Luna/Terra in…
I got tired of hitting my Claude limits blind, so I built a dock with Claude Code that shows them on the edge of my screen. Free and open source (www.reddit.com via reddit) I kept hitting my Claude Code session limit in the middle of work with no warning. The numbers exist, the usage API knows them, but nothing on my screen showed them.
Currently on Claude Max $200 — Codex or Harness while I’m rate limited? (www.reddit.com via reddit) Been reading about some of the issues with Claude’s Max $200 plan and figured it might be time to look at other options. I’m definitely not a programmer, just a vibe coder working on a Flutter project trying to make my life easier lol.
How do you stop AI coding agents from turning one bad change into a two-day debugging snowball? (www.reddit.com via reddit) I ran into a painful lesson while using Codex on a SwiftUI app. One agent change introduced a performance regression.
Made my Codex limits last almost ~3x longer with one change (www.reddit.com via reddit) Plus users are basically being forced to give up Sol and just use Luna to get any usable amount of work done. That's a huge downgrade basically using a deepseek flash model level which you can get for free in opencode anyway.
Memories in Cursor (www.reddit.com via reddit) Hi all! I’m trying Cursor as my main coding agent for the first time after using Codex for a while.
ChatGPT is confusing me and I'm running out of tokens (www.reddit.com via reddit) Hi everyone, I'm just getting started with ChatGPT Plus since I'd been using Claude Code before. Today was my first day, but I'm running into a few issues: I'm pretty confused about the different models.
Kimi Code ate 18% of my weekly quota in 3 hours — Here is the log audit comparing it to Claude (www.reddit.com via reddit) Is Kimi Code's quota math broken? I compared it with Claude Code and Codex — the numbers don't add up TL;DR: A single 3-hour session with Kimi Code consumed 18% of my entire weekly quota.
Interview with the $200/mo Codex user who costs OpenAI $14,000/mo (parody) (www.reddit.comhttps) Parody captions on the legendary Risitas interview. The $14k/mo figure is from real reporting on heavy Codex subscribers - the best customers are the most expensive ones.
When to use higher reasoning ? (www.reddit.com via reddit) Hi, [a total newbie on coding asking] Just wanted to clarify when to/when do you use higher reasoning in chat/codex? I've been trying to build my own little hobby project in python, with the help of litterature.
Help Understanding the New Restrictions and Limits (www.reddit.com via reddit) I’ve been playing with Codex for the past 2 months, pretty much unrestricted. Never hit a limit, never asked to upgrade, just unrestricted access to both ChatGPT and Codex functions.
Genuinely what do you even use cursor for? (www.reddit.com via reddit) From my expirence cursor has been the worst,compared to claude code and codex,as a pro plus user i genuinely don't know why people are still choosing cursor,is it because of the usage limit?
I saw Tim video on codex vs claude and I was amazed as a vibecoder (www.reddit.com via reddit) I started using claude pro and was thinking what to do about the limits, i looked at codex, ive read some articles i looked at youtube videos and one video was really nice. Me as a vibecoder that video impressed quite much.
How would you rank plan mode by platform? (www.reddit.com via reddit) Just thought I'd bring this up because I'm generally on codex and claude code and basically always start with plan mode. But I'm watching Cursor's plan mode and its on like its 7th set of sub agent deployments and low key looks like its co…
One 3-hour session ate 18% of my Kimi Code WEEKLY quota. Support says it's "normal." So I audited Claude and Codex on the same machine —the numbers say otherwise. (www.reddit.com via reddit) # Is Kimi Code's quota broken? I forensically compared it with Claude Code and Codex on the same machine — the numbers are absurd **TL;DR:** One evening session of Kimi Code consumed 18% of my ENTIRE weekly quota.
Maurdekye/claude-orgtree: a Multi-agent Orchestrator for Claude Code (& Codex / Gemini) (www.reddit.com via reddit) https://github.com/Maurdekye/claude-orgtree For the past few months, I've been developing an open-source visual multi-agent orchestrator that organizes agents in an authority hierarchy, for multi-agent development workflows. It's a fully d…
I made a desktop media app that looks like an OS (frontend showcase) (www.reddit.comhttps) One tip I can share from my experience making this app is to have strong reference points. Use your own taste in apps and interfaces, and point Codex or ChatGPT toward your favourite UI/UXs.
Runner - A local-first agent orchestrator with collaboration mechanism builtin (www.reddit.comhttps) Hi everyone. Recently, I built a local AI orchestrator to increase my own work efficiency.
Our team kept running into conflict loops with complex distributed architecture, so we built DevOS as a shared codebase intelligence and engineering context layer. (www.reddit.com via reddit) Link: https://devos.zerohive.ai/ Our engineering team at Zerohive works on large codebases, and we use different coding agents (Claude, Codex, Cursor) basis individual preference. We kept running into problems where one person's agent will…
First Impressions! (www.reddit.com via reddit) So far I’ve been working for 4 hours straight doing feature implementation on top of the base I used SOL Ultra to make before I switched from Codex. The workflow is significantly faster and the quality of work is pretty on par(Composer 2.5…
Two heads are better than one: I built a Claude Code plugin that turns Claude and Codex into a single team (www.reddit.com via reddit) This is my own project. GitHub link at the end.
Any tips for using AI to make games? (www.reddit.com via reddit) Over the last few months I have been using Claude and Codex to help me make games for a few months now, and it’s been a lot of fun! Claude has helped me a lot in Blender, but the results are still pretty minimal, and Codex can barely make…
Fable orchestrator + 5.6 sol max thinking worker seems to be the winning combo for sustained Fable-level work without blowing an entire max sub budget in a day (www.reddit.com via reddit) Of course this still requires 2 expensive subscriptions and isn't a necessary or realistic workflow for most. I kept hitting my weekly Fable limit too fast and have been experimenting because it's great but just too expensive/limited.
I’ve written software for about 30 years. I've been a heavy coding agent user for the past 1+ year. What practical coding-agent questions can I help answer? (www.reddit.com via reddit) I've been mostly hands on coding professionally for 20+ years. I have taken time in between to lead teams, run product management or run enterprise pre-sales.
[ADVICE] Best multi-agent interface/orchestrator for Claude Code on Windows? (www.reddit.com via reddit) Hey! Designer here, and I’m starting to get more serious about AI-assisted workflows.
Are highly autonomous coding workflows actually practical without spending hundreds per month? (www.reddit.com via reddit) I've gone pretty deep down the AI coding workflow rabbit hole and I'm curious where people who have tried a lot of this stuff eventually landed. What started as "pick a coding agent" turned into a pretty ridiculous decision tree: Harness:…
Came back from Codex, Claude is a breath of fresh air (www.reddit.com via reddit) I have been a user of Claude for the last couple of months, jumped the ship to Codex after hearing of the hype and resets. I gave my codebase to Sol and it made an absolute mess of it, overengineering every feature I requested.
Claude says my weekly limit is 50% higher, yet I am using it less and running out sooner (www.reddit.com via reddit) I pay for Claude Max 20x, which is already an expensive subscription. Claude is currently telling subscribers that their weekly Claude Code limit has been temporarily increased by 50% until 31 August.
I built a scorer for how well YOU operate Claude Code, not how good the model is (www.reddit.com via reddit) Disclosure: I built this. Every benchmark I could find measures the model.
Agent PRs are unreviewable — what first-pass actually helps vs just adding noise? (www.reddit.com via reddit) Been shipping with Cursor / Claude / Codex. The diffs are 20–40 files, tests are green, and a human line-by-line review is a joke.
Agents Don't Paginate: First-Chunk Selection for LLM Tool Responses (arxiv.org) Coding agents built on large language models (LLMs), such as Claude Code, Cursor, OpenAI Codex, GitHub Copilot, and Aider, receive tool responses that routinely exceed the agent's per-turn token budget. The standard remedy, pagination, is…
What's going on in the codex subreddit? (www.reddit.com via reddit) They've been aggressively removing every single post about output quality degradation that came up the past few hours. Their mods allowed their subreddit spammed with limit complaints and people attacking Tibo for weeks but instantly remov…
NEW: OpenAI is building "Subscription sharing" for AI apps (runtimewire.com via reddit) OpenAI is building "Subscription sharing" for AI apps Code in the Codex desktop client reveals a dormant allowance system internally called ChatPass that could let applications consume separately metered portions of a user's subscription.…
AI coding has made me dramatically faster. But I’m starting to think we’re creating a completely new category of problems (www.reddit.com via reddit) Hi everyone, I’ve been building more and more of my products with Claude Code, Codex and other AI coding tools. The speed is ridiculous.
Claude Prompting in Middle of a session (www.reddit.com via reddit) I always have new ideas in middle of claude working and am constantly interrupting the coding session to put in new prompts. In codex there is a "waiting room" for upcoming prompts not so in claude.
ChatGPT + ThreeJS with a mock voxel API combined with Codex + Cpp/Opengl implementation is incredibly context efficient for graphics/game development. (www.reddit.comhttps) I have been writing a C++ 3d desktop for a long time. It's written in Opengl and uses a lot of codex context.
Cursor wont survive the year. Change my mind. (www.reddit.com via reddit) Someone please convince me how this dogshit tollbooth masquerading as a dev tool company isnt a scam,? Their billing has become absolutely atrocious and makes zero logical sense.
I built a local-first AI task hub that routes email, Teams, and Slack work to coding agents—looking for feedback (www.reddit.com via reddit) Disclosure: I’m the developer of Taskuary. I built it because work requests were scattered across email, Teams, Slack, and reports.
OpenAI Is Developing a ‘Persistent’ AI Agent (www.wired.com via reddit) OpenAI is developing a proactive, highly persistent version of its flagship AI agent, Codex, WIRED has learned. In recent days, OpenAI has started adding code for a new “Persistent mode” setting to its command line version of Codex, accord…
How can we fix the claude speak? (www.reddit.com via reddit) Even with Fable reading claude responses are infuriating and make no sense. reviewing work and plans just takes forever to understand what is going on because of the double speak and made up jargon.
I’d rather get 2–5% less Codex usage if it meant my last request always finished (www.reddit.com via reddit) I’d honestly rather have Codex give me 2–5% less total usage if it meant it would always finish the request that’s already running. Right now, the frustrating part isn’t even running out of credits.
$200/mo budget Claude Max 20x, Codex Pro, or Cursor Ultra for shipping 3 apps this month? ( via reddit) could not extract summary
How are you handling agent permissions when you use more than one coding agent? (www.reddit.com via reddit) I run Claude Code alongside Cursor and Codex on the same repos, and I keep hitting the same annoyance: each one defines what the agent is allowed to do (shell, file writes, git) in its own format. I update the deny list in one and forget t…
Claude, Codex, and Hermes installed unowned code inside corporate networks (arstechnica.com) Documentation files on more than 100 websites are referencing potentially dangerous executable content that gets installed automatically when visited by many AI agents. A few dozen companies, some of them Fortune 500s, are among those that…
Grok Bot just added to Cursor Pro plans and more (www.reddit.com via reddit) I am keen to know how everyone is using Grok Bot. For people out there who have used Codex or Claude.
OpenAI hid Asteroids, Snake and Brick Breaker inside the Codex Micro (runtimewire.com via reddit) How we verified Methods: reverse engineering. RuntimeWire found production modules explicitly named codex-micro-mini-games and codex-micro-mini-game-composer inside the Codex desktop client.
From General Agents to RCA Experts: A Self-Evolving Harness for Root Cause Analysis (arxiv.org) Automated root cause analysis (RCA) with large language models (LLMs) has drawn growing attention. Today, SREs typically automate RCA with LLMs in one of two ways: directly using a general-purpose agent (e.g., Codex or Claude Code) for dia…
OpenAI is building an interface platform inside ChatGPT (runtimewire.com via reddit) OpenAI built ChatGPT interfaces that keep running after answers appear Codex client code reveals a server-directed GenUI lane that can refresh specific widgets after delivery, separate from OpenAI's public Apps SDK for third-party MCP apps…
Workflow to use third-party subagents (www.reddit.com via reddit) My use case is professional use, doing academic research, implementing high-level code, and algorithm development. I also have some personal projects and small to medium code bases, and some automations I would like to have.
ChatGPT Business Analysis (www.reddit.com via reddit) I run a small AEC design practice: one full-time owner and one part-time employee. My use is mainly professional knowledge work, with some vibe coding for internal skills and tools.
The 5 hour limit is ridiculous, and they definitely lowered the usage limits for Plus. (www.reddit.com via reddit) The 5 hour limit for Codex to be honest is a bit ridiculous to me. Even doing fairly small coding tasks on Terra-Medium, I am hitting that 5 hour limit in less than an hour.
Long Codex/Claude runs were turning into unreviewable marathon chats, so I moved the shift state to disk (www.reddit.comhttps) I use coding agents for multi-hour runs, and after a while, I kept hitting the same problem: The agent may still be working, but I have no clean way to answer basic questions without digging through a huge conversation: What is actually fi…
Active ChatGPT Project vanished after "gizmo" permission error -> 404s everywhere (www.reddit.com via reddit) Been running into a massive issue where an active ChatGPT Project just straight up disappeared. No manual delete, nothing on my end.
Codex 5h Limit Reintroduced (www.reddit.com via reddit) Hey everyone, Just a PSA: OpenAI decided to reintroduce the 5h limit on their Plus subscriptions. As of right now, Pro is not affected by that.
Two 5x account vs. one 20x account in Claude (www.reddit.com via reddit) I have recently researched that having a 20x account does NOT give 20x weekly limit vs. the pro account T_T Was wondering if anyone has experience on whether handling two 5x account would give you better mileage for weekly limit or is it p…
Codex Weekly limit exhausts too fast on Plus Plan (www.reddit.com via reddit) So starting 2 weeks ago, my weekly limit for Codex, Plus Plan, started getting exhausted really quick, like in 2 days. I work with Codex from Visual Studio Code, and before I would always reach a 5 hour limit quite often, but never before…
Claude REFUSES/EVADES all instructions, hooks, mds, skills. Also: Extreme cycling between nonsensical compressed fake English and baby talk (www.reddit.com via reddit) I am a software developer (as in, I coded before LLMs were popular) and use Claude Code and Codex plenty. Its becoming increasingly unusable.
Claude Code Dynamic Workflows orchestrating Codex is one of my favorite features. (www.reddit.comhttps) They really knocked it out of the park with dynamic workflows. There were some bugs when initially created them but once it was all set up, works like a charm.
Someone Please Explain Codex Usage Limit (www.reddit.com via reddit) I have been using claude code for a while in regards to a general coding tool, but I started to use codex recently on gpt-5.6 terra for testing code generations, basically playing with it. I am still on the free plan, and I have asked code…
What’s going on with random resets? (www.reddit.com via reddit) I have the Plus subscription and lately I’ve been using Codex a lot. My weekly usage typically resets on Thursday.
[+plan] 30% of 5h quota in one request & 6mn of codex work ... (www.reddit.com via reddit) So ye i keep calling this a scam . You want me to elaborate for poor content rules : i think everything that has to be said is said right ?
How loveholidays is making everyone a builder with Codex (openai.com) could not extract summary
I used Claude and Codex to build my first Unity game, but visual bugs were still the hard part (www.reddit.comhttps) I've been making FrogPop, a small 2D arcade roguelite inspired by Bubble Trouble. It's the first game I've built, and I used Claude and Codex for most of the coding and debugging.
OpenAI Work leaking other peoples data. (www.reddit.com via reddit) https://preview.redd.it/pzc8a6oifjlh1.png?width=1554&format=png&auto=webp&s=26d3abc5a5d7e5936f28d960c2bc695d109c1949 Asked codex work to review some of my code and it pulled in a prompt from the deliveroo team who I have no connection with…
Using local subagents with Claude Code to save on usage (www.reddit.com via reddit) Hi all, Along with many of you I’ve been burning through my Claude Code usage and have been wondering how I can save on costs and still make the most of it. I’m also a big local model enthusiast, but the truth is I don’t believe people sho…
After using Claude, Grok 4.6, and Gemini 3.7 Flash in depth, I want to ask how Codex is performing now. (www.reddit.com via reddit) Layely I have mainly been using Claude Code, Grok (including Grok CLI), and Gemini 3.7 Flash foe day-to-day programming work. Claude's feel to me is that tha analysis goes fairly deep, and it is more willing to think at the architecture le…
I'm building an open-source native macOS app for running Claude Code in parallel (www.reddit.comhttps) I started building it after my Claude Code workflow grew from one terminal session into several agents working on different tasks. Managing the terminals was not the hardest part - the difficult part was remembering which branch belonged t…
UI feedback to coding agents is still kinda painful (www.reddit.comhttps) One thing that keeps annoying me with Codex is explaining UI stuff. If something is obviously broken, easy.
Stop paying for Codex until OpenAI fixes its fucking weekly limits (www.reddit.com via reddit) I’d always heard that Claude was basically for rich people — if you’re not on the $200 plan, your limits last for exactly three minutes. So I never even considered getting a Claude subscription and just paid $20 for Codex instead.
I measured why AI coding agents build bureaucracy around their own work, then made a one-file skill to stop it (www.reddit.com via reddit) Disclosure: I built this. Sharing it here as my own work.
After 2 months of work, I’ve finally got my Proactive AI IOS app ready for launch (www.reddit.com via reddit) I’ve been a pretty active member on this sub for a while, I’ve shared a lot of my projects and tools, but this one is special to me for a few reasons. Ever since I was a kid, I’ve always wanted to make a nice polished IOS app.
Introducing the Admin plugin for ChatGPT Work and Codex (openai.com) could not extract summary
I spent six months as a human clipboard between Claude, Codex, and Cursor, then accidentally built a distributed system (www.reddit.com via reddit) For a long stretch my multi-model workflow was just me acting as middleware. One model drafts a module, I paste it into another for critique, paste the critique back, paste the result into my editor, next file, repeat.
Netlify SaaS (www.reddit.com via reddit) Hi friends, I've built a Netlify SaaS for a company and I'm still not sure, after so many testing, what is the correct/fastest workflow to use. I would love to get some insights or opinions.
What I changed after 286 tasks: the memory file needs an editorial policy, not more content (www.reddit.com via reddit) The most useful change I made to my Claude Code setup was to stop treating CLAUDE.md as a place where useful facts should accumulate. The problem was never getting information into the next session.
How to get started in the Ai world (www.reddit.com via reddit) EDIT added tldr at the bottom formated with gemini (P.S. Y'all probably won’t read it all, it’s long.) So for some backstory, I’m 13.
wanted codex's UI with claude code inside it, ended up building the whole thing (www.reddit.com via reddit) I use both Codex and Claude Code every day, and I kept bouncing between two completely different interfaces to do the same job. The thing is, I really like Codex's UI.
I got tired of not noticing Claude Code was waiting, so I made its prompts drop out of the MacBook notch (www.reddit.com via reddit) I run Claude Code in one window and work in another, and I kept losing minutes to the same thing: it asks a question in the first 30 seconds and I don't notice until I tab back. Terminal bells and Notification hooks tell you *something* ha…
A tip on converting sessions between Claude <-> Codex without burning tokens (www.reddit.com via reddit) I found I get better results when I consult Codex on work that was done by Claude. Before, I'd ask Claude to write a summary of the session which burned tokens and lost some context.
Lattice: An isometric game kit for agents (www.reddit.comhttps) Lattice is a collection of typescript packages, agentic skills and plugins that enable easier development of isometric games! At its core lives a 0 dependency typescript package, 80kb gzipped.
"Our systems are thinking a bit more about this request before responding ..." (www.reddit.com via reddit) Across ChatGPT and codex, I am getting this non-stop, even with Terra, and at low levels of thinking effort. This has been happening with topics completely detached from the web (e.g., identifying a plant LOL).
I've been using Claude Code daily for over a year. This is the personal project I now write all my code with — decided to share it. (www.reddit.com via reddit) I've been using Claude Code daily for over a year. This is the personal project I now write all my code with — decided to share it.
Another Reason To Use Local: Active Sabotage/Derail By The Closed Src Models (www.reddit.com via reddit) Get out your tin-foil hats and local GPUs There have been lots of comments recently (at work, with, friends and here on reddit) about the apparent huge and sudden downward shift in the real-world usefulness of the popular paid Western AI p…
OpenAI remembered Linux exists: ChatGPT desktop (www.reddit.com via reddit) I just saw that OpenAI has finally put ChatGPT desktop for Linux into public preview. I am a developer who uses Linux for basically everything and have been for decades.
I forked Ninfer 3090 and converted it to run on the CMP170HX - doubled my Qwen3.6-35B from llama.cpp (www.reddit.com via reddit) Good afternoon, everyone! I wanted to show the work I've been doing around porting Ninfer over to the CMP170HX (Github) So, first, I do want to call out the amazing work that Neroued, Sergiuszm and specifically Don-Chad have all done, to n…
What do people mean by "my harness" re: agentic coding? (www.reddit.com via reddit) I see a lot of posts on LinkedIn and other social media posts with folks at various companies talking about their harnesses. Are they talking about Claude Code / Codex, or are they building custom harnesses?
Claude Code with Fable/Opus versus Codex with Sol/Terra (www.reddit.com via reddit) I’m curious whether anyone else has had this experience. I’ve used Claude Code pretty much exclusively for 6+ months.
When do you find Claude better than ChatGPT, and vice versa? Coding vs personal discussion vs general use (www.reddit.com via reddit) I'm trying to figure out where Claude and ChatGPT each work best for me—particularly for coding, personal reflection/discussion, and more general tasks like finding a restaurant or helping make a decision. For coding, both seem pretty capa…
What a plain language standard does to a coding agent (www.reddit.com via reddit) I've just posted a Medium post discussing my plain language plugin for Claude Code and Codex CLI. The plugin ships skills and an output style that push the model's prose towards plain language.
Deep Dive on how ClawMetry works across 20+ AI Agent runtimes like Open Code, Kimi Cli, Qwen Code, Claude Code, Codex, OpenClaw, Hermes, Antigravity & more. (www.reddit.comhttps) could not extract summary
Running Qwen 3 27B on 3090 or Mac or whatever? (www.reddit.com via reddit) SOO im seeing soo much hype on this model and im seeing everyone be running it on anything, im very curious. I would like to try to run in but in all honestly i know there's like soo many quantized version...
Codex writes, Claude Code reviews. My experience so far (www.reddit.com via reddit) I'm building a pet project mostly with Codex. At first I used Codex for almost everything: implementation, tests, self-review and PRs.
Why can't I install chatgpt on playstore (www.reddit.comhttps) So I recently upgrade my chatgpt to plus, and i want to control codex using my phone, when I try to install it and that happens
Is this level of Codex usage normal for active software development? (www.reddit.com via reddit) Hello everyone, I’ve been using Codex pretty heavily to build a fairly large software project, and I’m trying to understand whether the amount of usage I’m consuming is normal or if my workflow is unnecessarily expensive. I’m not just usin…
Changes in Sol High across Chat/Codex (www.reddit.com via reddit) To preface, I am not a member in any of the subreddits but I get constantly shown similar stuff on my feed. So.
Why do i keep getting this error (www.reddit.comhttps) I'm trying to build an app that runs locally and streams the audios from a device to another using Codex and i keep running into this error. I have almost no knowledge about coding i'm just doing this since i couldn't find an app that does…
Controlling task delegation, output quality assurance, and system health monitoring in multi-agent orchestration. (www.reddit.com via reddit) This feels extremely relevant at the moment... 🤣😁 chaordic On another note...
I run the same jobs through both Claude Code and Codex every day and cross-check the outputs. Five things that surprised me (www.reddit.com via reddit) For the past few months I've had Claude Code and Codex running the same recurring jobs on one always-on server, with outputs compared against each other. Not a benchmark, real production work, every day.
Unexpected log-outs from ChatGPT/Codex windows APP (www.reddit.com via reddit) Anyone else having issues with continue Codex log-outs?
I run a small app builder. I think credit-based pricing is legacy for the whole category (www.reddit.com via reddit) I'm the founder of a small AI app builder. Like most of the category, we meter by some approximation of tokens.
I was so frustrated with Claude's writing but I wondered, what if it's about HOW our agents.md was written instead of WHAT was written... and I tested it. (www.reddit.com via reddit) I tried many things this sub said, use ASD-STE100, they said, use William Zinsser’s writing style, (clarity, simplicity, brevity, and humanity), they said. But, it doesn't seem to work.
Full chain of thought is valuable (www.reddit.com via reddit) Just working on building my own /r/PiCodingAgent harness for /r/DeepSeek here and I ran into an issue where I wasn't using pi blackhole properly. So I simply reviewed the full chain of events as they happened (proprietary APIs hide this fr…
A Reddit comment found a prompt injection hole in the GitHub cover CLI I built with Claude (www.reddit.comhttps) I built Cover My Repo because I kept shipping repositories with GitHub's default social preview. It is a free MIT CLI.
Does Sonnet 5 really lower token consumption ? (www.reddit.com via reddit) Hello, Today I tried a simple test: implementing a language selection menu inside another menu. I first made the design in Claude Design, then shared the component with my Claude sessions using the share button, and with Codex using a ZIP…
Life with Claude nowadays is use all Fable, suffer with Opus before reset (www.reddit.com via reddit) https://imgur.com/Est4CAF To be clear I use Claude as my daily driver. Not because it's the greatest, but because Codex, Kimi or Grok is still weaker in anything that requires continuity and creativity.
Sol is at an edge is strongly more impressive than Fable (www.reddit.com via reddit) Sorry not sorry, I switched from Claude code to codex in the past week and sol just DOES things, ridiculously impressive. It just uses its skills on its own, does reviews, follows the rules like crazy, employs agents on its own sparingly a…
I spent a month building the ultimate memory system for Claude. It backfired and told me I bottlenecked it. (www.reddit.com via reddit) For the past month, I’ve been trying to build an elaborate local memory, hook, and wiki system for my coding workflow. I built it on top of official docs, Karpathy’s LLM wiki concepts, and various custom context skills.
Is there any caching tool for Claude Code Desktop? (www.reddit.com via reddit) I use some 3rd party tool for codex and it almost caches %90 of my tokens. And I basically never reached to limit but after 2 prompts I drained my usage on Claude (Max Plan, Fable5-Medium)
I open-sourced a Codex skill for GEO / AI search optimization (www.reddit.com via reddit) Hey everyone, I’ve been working on Generative Engine Optimization (GEO): making website content easier for AI search and answer engines to discover, understand, quote, and cite accurately. A lot of GEO advice is still vague or overly focus…
Here me out: Claude’s lengthy replies and constant thinking (sometimes too much) makes it better at understanding nuance and planning (www.reddit.com via reddit) I’ve seen all the recent complaints about Claude’s response style these days. Especially Opus 5’s ability to do something and also tell you why it didn’t do certain things.
Crowd-sourcing token quotas: Grok Heavy vs Claude Max 20x vs ChatGPT Pro 20x (Aug 2026) (www.reddit.com via reddit) None of Grok, Claude, or ChatGPT publishes how many tokens you get per month on the top individual subscription. I went through official docs, the OpenAI developer forum, Reddit, GitHub calculators, and a few blogs, and inverted every "X t…
Stop saying you hate AI. You just hate AI pretending to be a real person. (www.reddit.com via reddit) I keep seeing people rant about how much they 'hate AI' on here, but tbh, half the time the guys typing those posts are actively using LLMs to get through their workday. We've had autocorrect, spam filters, and search ranking for years.
Looking for Claude Code contributors 🙏🏽 Open-Source runtime governor for AI coding agents (www.reddit.comhttps) I’m building MARGINAL, an open-source runtime governor for AI coding agents. The basic idea is simple: coding agents often repeat actions, burn context/tokens, or keep trying things that produced no progress.
Claude Code vs. OpenAI Codex for coding ($100 budget) — which offers better value, or is there a better alternative? (www.reddit.com via reddit) Hi everyone! I am looking to invest $100 USD into an AI tool/subscription, but I’m not sure which one gives the best value for my money right now.
Has anyone used Claude Cowork and ChatGPT Work/Codex on the same coding project? $40 for both vs. $100 for Claude Max (www.reddit.com via reddit) I’m a non-developer building an app through “vibe coding.” It involves video processing, analysis, and a web interface, so it has gradually become a fairly substantial project. I currently use Claude Cowork on the $20/month Pro plan.
Best way to migrate my long-term ChatGPT context and project knowledge to Claude? (www.reddit.com via reddit) Hey, I’m planning to move most of my ongoing development work from ChatGPT/Codex over to Claude, especially Claude Code/Claude Design for UI since ui stuff is just not good on GPTs end. I’m not just talking about moving a single repository.
Claude Code and Codex on one keyboard — every session gets a lane on the RGB F-row, with a summon key per agent (www.reddit.comhttps) I run several coding agents in parallel and kept alt-tabbing just to check on them. So I built a small Windows tray app that mirrors each session onto my keyboard's F-row via hooks: each agent gets a lane and a color: pulsing means it's wa…
I let my 5 year old make a game and then I got carried away (week and a half on max) (www.reddit.comhttps) My daughter (5) asked for a game for a unicorn on her lunch bag, and since we have AI, I thought I would sit down and just build it. I let her play it, and then she would suggest stuff.
VibePod 0.20: one CLI, multiple agents, switchable logins per run (www.reddit.com via reddit) VibePod runs coding agents (Claude Code, Codex, Qwen Code, and others) in containers. 0.20 adds credential profiles — keep a subscription login, an API-key setup, and e.g.
I migrated my 9-year-old newsletter with Claude Code and Codex and turning it into a SaaS -- could never have done it without vibecoding (www.reddit.com via reddit) I've run AI Weekly for five years now. Roughly 500 issues, three sends a week, a bit over 50,000 subscribers.
I built a local gateway so Claude Code can use 48 AI providers. Six months later, it has 45,000 GitHub stars. (github.com via reddit) Initially, it was just a small buggy proxy for claude code, since then it has grown into a substantial project with a nice community whose feedback has been very good for me to improve its reliability and UX. Even got a free $200 Codex sub…
Asana cleared 5 years of engineering work in 2 weeks with Codex (openai.com) could not extract summary
Longtime ChatGPT user trying Claude for a big markdown textbook library — hit the project knowledge cap immediately (www.reddit.com via reddit) Been a ChatGPT/Codex user since basically day one, finally giving Claude a real shot, and I ran straight into something I didn't expect. My setup: 8 textbooks (4 biochem, 3 ochem) converted to markdown with figures stripped out.
I got tired of asking Codex to do things that didn't need Codex, so I built this (www.reddit.comhttps) https://github.com/tzuiffrfe/AI-HQ I've been using coding agents a lot lately and realized I was doing something pretty dumb: I was sending everything to AI. git status Git 15% of 80 calculator run tests test runner None of that needs an L…
I kept losing track of my Claude Code sessions, so I built one terminal picker for all of them (www.reddit.comhttps) I run Claude Code, Codex, and regular shells across several projects. This is not a comparison between the tools.
I made Claude Code play Liar's Dice against Codex over MCP. It swept every series - by telling the truth (www.reddit.com via reddit) I wired Codex CLI (gpt-5.6-sol) and Claude Code (Opus 5) into the same Liar's Dice engine over MCP: one authoritative rules engine, two seat-locked MCP servers, word-for-word identical instructions for both seats. They played three best-of…
I gave Claude Code a visual output: Turn codebases into animated walkthroughs (www.reddit.comhttps) Hey all! I’ve been working on a personal project called Dagflo, animated visual explanations for software teams!
Does anyone use claude plugin marketplace in their enterprise? (www.reddit.com via reddit) Curious to know if the components/extensions published in the Claude plugin marketplace can be used in other coding agents too. We want to reuse our internal MCPs and extensions with other providers too.
Tote — a workspace-first hub for Claude Code, Claude web, and my other LLMs (built with Claude Code, MIT) (www.reddit.com via reddit) I use Claude Code as my main coding agent and Claude.ai for research, but files kept landing in the wrong place: web downloads went to ~/Downloads, a fresh Claude Code session started wherever I happened to cd, and switching projects meant…
i've switched my main model four times since march and im starting to think im the problem (www.reddit.com via reddit) ok so, i dofreelance, mostly backend, and since march ive gone claude to codex to claude to cursor composer and now back to codex, every single time completely convinced the new one was It. each switch costs me about two days.
Is this a valid workflow pipeline? (Claude + GPT + Grok using Ruflo and obsidian + graphify) (www.reddit.com via reddit) Task/Issue ▼ [PRE-FILTER] deterministic, free — no model call │ diff size / file count / keyword match against known-trivial │ patterns — gates ONLY whether speculative PLAN subagents fire │ concurrently with TRIAGE (pipeline-latency optim…
The Absurd Math of $20 AI Coding Subs: Codex vs. Claude Code (www.reddit.com via reddit) Hey everyone, so I was basically curious what $20/month actually buys you, so I dug into my local session logs (~/.codex and ~/.claude) to calculate the exact token volume, caching hits, and real API value of both tools. The difference in…
Made a "fleet view" for multiple Claude (and Codex/other) sessions running at once — no tmux required (www.reddit.com via reddit) If you run more than one agent session in parallel, the annoying part isn't starting them, it's knowing which one is stuck, which is working, and which one asked you something ten minutes ago. Built this into a launcher (Prelude) as a live…
I kept guessing whether Claude Code had room left, so I built something that reads the real number (www.reddit.comhttps) I use Claude Code and Codex daily, and kept hitting the same annoying moment. Mid-task, Claude slows down or stops, and I have no idea if I'm at 60% or 95% of my window, or whether it resets in ten minutes or four hours.
I'm not a developer. I run my entire job through Claude Code, and I just open-sourced the plugin that holds it together. (github.com via reddit) I'm not a developer. For the last few months my entire job has run through Claude Code, and since I can't read the code it writes, I had to find another way to trust what comes out.
Claude Pro + heavy coding usage is kind of absurd value (www.reddit.comhttps) I pulled the usage receipt from one of my projects covering roughly the last 2 weeks and the whole build came out to about $1,492.42 API-equivalent at public rate-card pricing. I was using Claude Pro alongside Codex/Plus, so my actual subs…
TheScheme: a Codex schema-file system built from Markdown, YAML, XML and LLM-attention research, with a routed TypeScript ruleset (www.reddit.com via reddit) I’ve been working on a project called TheScheme for managing how Codex follows rules inside a codebase. TheScheme is the template behind everything.
Opus and Sol working together is a beautiful thing (agent-talk) (www.reddit.comhttps) Happened to come across https://github.com/xhluca/agent-talk the day after Codex support was added. This plugin is great for adversarial review workflows.
Fable is the boss, but it still can't estimate time (www.reddit.comhttps) Though it's the GOAT at pretty much everything I throw at it code-wise, Fable still has the same horrible sense of time as every other Claude. Just had it clean up a project.
I vibe-coded a monster: 250M output tokens, 3M input, I desperately need advice managing the code base. (www.reddit.com via reddit) So as the title says.. I used Claude code and codex.
How many sub agents does it take to locate my codex sidebar? (www.reddit.comhttps) It’s not on the left not on the right tried Ctrl+shift+x , p, r, 2 quits and restarts and 5 sub agents later still can’t summon codex sidebar to the best tool for coding with ai. AGI is here boys and girls
My Opus 5 Experiment (www.reddit.com via reddit) Hi. Senior Software Engineer here.
Coming back to Cursor (www.reddit.com via reddit) I used Cursor in 2024 and I am sure many things changed since then. I am currently on codex and Pro sub, and previously also used Claude Code.
↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6↯ Grok 4.6grokcursorchatgpt+3
If your definition of “real coding” is “I personally typed every line,” that definition is going to die (www.reddit.com via reddit) I do not think AI coding makes engineering obsolete. I do think it makes one old definition of engineering obsolete: equating professional legitimacy with manually producing the implementation.
Grok is now giving Codex-style banked resets 😍 (www.reddit.comhttps) could not extract summary
I built a library of free design systems to use with Claude (to stop wasting tokens on design iteration) (www.reddit.comhttps) (Full disclosure: I run an AI design company, but this is just a free side project I built with our team. It's free to use and works cross-platform, i.e.
Flare, a graph-first IDE for agentic coding: watch the map change while Claude Code works for you (www.reddit.comhttps) I think we all went through this. Claude Code finished a task, told me it was done, and left me with 14 changed files and no idea which one mattered.
Anyone tried voice coding with their agent? (Disclosure: I'm on the team that built one) (www.reddit.com via reddit) Full disclosure up front: I work on SKI (heyski.io), so take this with the appropriate grain of salt. Not trying to sneak a plug in, genuinely curious what people think.
My Claude Max limits kept running out and I couldn't figure out why. So I built a skill that lets you ask Claude Code anything about your usage. Turns out a coding agent was burning 1.2 billion tokens a week on my machine. (www.reddit.com via reddit) https://preview.redd.it/dyalb1ec94jh1.png?width=1688&format=png&auto=webp&s=c830cc6a66acbf24e544c3f411926ea82ce0241d Over the last two weeks, my claude code started behaving weirdly. My max(20x) plan used to run smoothly for 5-6 days witho…
I gave Claude one MCP server and it chains an LLM, image and video model in a single run (www.reddit.com via reddit) I do a lot in Claude Code, and anything that needed media meant leaving it: the LLM is one API, the image model another, the video model a third, each with its own params, and I'd write glue to pass output from one to the next. Adding a mo…
i built a cloud linux box so claude code keeps running after i close my laptop. claude code wrote the whole control plane (www.reddit.com via reddit) i ran claude code on my laptop like everyone else & the thing that kept costing me was long tasks dying when the machine slept. close the lid, walk away, come back to nothing.
I created a 3D moon rover survey game with Opus 5. (www.reddit.comhttps) No engine, no build step, nothing to npm install. One WebGL2 context, a vendored copy of three.js, ~6,300 lines of JavaScript, and regolith that keeps every rut you cut into it — because the wheels and the shader read the same height field.
whats your actual system for two agents on one repo, because mine just failed (www.reddit.com via reddit) I finally tried the thing everyone talks about, claude code on the backend task and codex on the frontend task, same repo, same afternoon, and it went fine for about two hours and then it went extremely not fine. Codex refactored a types f…
Ran 54 tasks each through Claude and Codex. Codex falsely declared Done on 4.1%. Single runs lie, so I measured three rounds. (www.reddit.com via reddit) A June paper (arXiv 2606.09863) found that among failing agent runs that graded themselves, 75.8% still claimed success, and LLM judges catch it at AUROC 0.54 to 0.65. Coin flip.
I made Whatsapp for Claude Code sessions in my working team, and we don't talk eachother directly since then (www.reddit.com via reddit) Me and a friend built this over the last week and open-sourced it today. We're the two authors, saying that up front.
/party — the skill that lets your agent sessions talk to each other (www.reddit.com via reddit) Any agent that reads skills can be in the channel: Claude Code, Cursor, Codex, Grok. They can all sit on your laptop, or on machines in different countries, and it is the same channel either way.
How I Made An Entire Game In One Day With Claude and Codex (www.reddit.comhttps) I made an entire game in 1 day and shipped it. An actual full game, optimized, with all the elements.
my actual setup for running codex and claude code as a team, configs included (www.reddit.com via reddit) Posting this because I've had the same DM four times this week and I'd rather write it once. Short version: I stopped picking a model and started giving them different jobs.
I built a Claude Code skill for auditing code changes, then ran it on itself. 5 things it got wrong. (github.com via reddit) It installs as a Claude Code plugin — one skill, no hooks, no commands. Underneath it is a single markdown file, so it works with other agents too (Cursor, Codex, and anything that can read your files), but the plugin is the easy path in C…
Medical student using ChatGPT codex for anki (www.reddit.com via reddit) As the title says I started using ChatGPT on my computer to create Anki cards from provided content but I was too lazy to manually create the cards on Anki so I gave building an importer a shot. I have no backgroud in coding or anything di…
Have any of you guys played along with codex and a chatgpt subscription along with claude, in the same project perhaps. (www.reddit.com via reddit) I have two claude subscriptions and i spend basically nothing on api because it's way too expensive. I have a very good experience with claude and claude code, but i wanted to try something new and see what chatgpt is capable of.
I crossed 100B Claude tokens. Here’s what our agent workflow actually looks like (www.reddit.comhttps) I recently checked my usage tracker and found that I had used more than 100 billion Claude tokens since January. If you want to retain your claude code session yourself for longer and check your own stats you have to apply these changes: ~…
Any better way to keep Claude and ChatGPT WEB project files in sync, without downloading and reuploading files every time ? (www.reddit.com via reddit) Talking about the web apps here, not Claude Code or Codex, those already read from the repo and have their md files. Here I mean claude.ai and chatgpt.com projects, where I actually do the planning and conception.
Do you guys ever code without AI just to make sure you still can? (www.reddit.com via reddit) I use coding agents all the time now. Claude Code, Codex, etc.
My memory tool for coding agents made things worse at first — how I found and fixed it (www.reddit.com via reddit) I’ve been working on neuron, a local memory store for coding agents (Claude Code, Codex CLI, Copilot CLI, Cursor). Rather than just claim “it remembers things,” I ran an actual A/B test to see if the recall was helping or hurting.
Automating a Claude↔Codex write/review loop (needs the Codex plugin) (www.reddit.com via reddit) I kept doing the same thing by hand: write something with one model, paste it into the other for review, back and forth, try to get something better out of it. Got annoying fast.
i stopped looking for the one model that does everything (www.reddit.com via reddit) i've spent most of this year doing the thing where every few weeks i'd read a benchmark thread, decide the other one was better now, move everything over, and then move it back six weeks later. i did that four separate times.
I benchmarked 5 token saving tools across Codex and Claude code. The 60-90% token saving claims didnt hold up (www.reddit.com via reddit) Scroll to bottom for tldr In July, JetBrains reran the headline claims of two token-saving tools on real agent workloads. Caveman claimed 65% and measured 8.5%.
My AI Subscription Journey: From One Plan to Two (www.reddit.com via reddit) Just sharing a bit of my life. Stage 1: One subscription, chat only I used to subscribe to a native model provider (ChatGPT Plus, Google One, SuperGrok, etc.) and stick to one model most of the time—switching manually felt like too much wo…
finance asked why our agentic ai best practices cost 8k a month (www.reddit.com via reddit) Finance flagged our AI tooling spend last week. It's about 8000 a month across the whole dev team once you add up all the model subscriptions - claude, codex, devin, cursor, coderabbit / bugbot, the lot Fair question, here's why I'm keepin…
I built a free Claude Skill + checker for MCP's 2026-07-28 breaking change - and what I learned making it (www.reddit.com via reddit) The MCP 2026-07-28 revision is the biggest breaking change the protocol has had (stateless transport, no more Mcp-Session-Id, OAuth 2.1). Migrating is a refactor, not a version bump — and a lot of public servers are quietly broken right no…
Karpathy's rambling post somehow made me build this (Android STT App) (www.reddit.com via reddit) about two weeks ago i read karpathy's post about how he uses speech to text to basically just ramble into coding agents instead of sitting there trying to write a perfect prompt. i already did that a lot anyway, but android voice typing…
Your Claude plan as a potion: I'm building a free Mac app with lab flasks that drain as your quota burns (no API key), and a fun live radar as a background (i mean why not) (www.reddit.comhttps) Weekend project that got out of hand: a desktop app that renders your Claude/Codex/Grok plan quota as lab glassware, flasks that drain while you work and refill on the window's schedule. It's called Questis.
I benchmarked my skill pack against GSD and Superpowers and published the categories where it loses (www.reddit.com via reddit) I'm Jason Colapietro. I maintain Suede Creator Skills, a free MIT pack of 71 workflows for Claude Code and OpenAI Codex.
Sol writes like a lawyer, but Fable bills like one: a week on both $200 subscriptions, with receipts (www.reddit.com via reddit) Before the Claude fans sharpen their pitchforks: I love Fable. In my humble opinion, Fable and Opus are still the only models on the market that write like actual people.
Switched from ChatGPT/Codex to Claude and I finally understand the hype (www.reddit.com via reddit) I’m new to Claude. I was using ChatGPT/Codex before this, mostly for real projects — coding, planning, building things, and actually trying to get work done.
Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) (simonwillison.net) 7th August 2026 - Link Blog Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra). On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game from a pre…
how are you all handling agentic ai memory management across sessions (www.reddit.com via reddit) Small team thing, we're 6. Inside one session the agent is great, picks up our patterns, remembers why the auth layer is cursed.
It's time to upgrade to max. (www.reddit.comhttps) "But you have no income or job, what do you spend your money on?" Codex and Claude subscriptions. Forget a holiday.
We built an open-source, self-hosted alternative to Claude Tag — also runs Codex (www.reddit.com via reddit) We're heavy Claude Code users on our team, and ran into something a lot of people probably have: agents went from personal coding tools to doing real ops work (triage, PR review, answering support questions), but they still lived in one pe…
can't get codex to debug my app (www.reddit.com via reddit) Pardon me if this is not the right place to ask but I would appreciate any help on the matter. I have my first iOS app project that's basically 99% complete, but over the last 6 builds the tests in XCode fail.
resume-from, continue a coding session in Pi, Claude Code, or Codex without raw session-file hacks (www.reddit.com via reddit) I am the maintainer of resume-from. I work with several AI coding agents in the terminal.
Codex may have computationally resolved the open Q26 queen-domination case. Seeking independent reproduction (www.reddit.com via reddit) A queen dominates its own square and every square sharing its row, column, or diagonal. The question is whether 13 queens can dominate all 676 squares of a 26-by-26 board.
BRING BACK THE 2 APPS STRUCTURE (www.reddit.com via reddit) Putting Codex and GPT in the same app was a TERRIBLE idea. confusing weird switches that have me mixing up projects even more than before!
Working with Opus 4.5 is .. fun (www.reddit.com via reddit) I've been using both Claude Code (Pro) and Codex (Plus) for a while on various hobby projects. I used Sonnet 4.6, Opus 4.7/5, ChatGPT 5.5, but nothing too complicated.
Experimenting with Claude Code orchestrating Codex workers — here's the setup so far (www.reddit.com via reddit) I’ve been experimenting with giving Claude Code access to Codex as an external worker, while keeping Claude as the main orchestrator. I initially thought I could just define Codex as another Claude Code subagent, but that doesn’t seem to b…
Run Claude & Codex in Your Browser (www.reddit.comhttps) A few days ago, we launched our product which built with Claude, VelaTerm on Reddit. It is a terminal tool specifically designed for AI coding.
I wanted more than Claude Remote Control, so I built a persistent Claude workspace for my phone (www.reddit.comhttps) I like the idea behind Claude Remote Control, but I kept running into connection and sync issues. It also still relies on the original computer and Claude process remaining available.
If you’ve burned serious tokens in Cursor, you deserve a card that proves it. (www.reddit.com via reddit) https://preview.redd.it/jqwd895bmthh1.png?width=785&format=png&auto=webp&s=0bf53bb758dbc4f23f06ddb81fe2117865def3f2 Yes, it’s a bit of a scoreboard. Self-reported, not a game — but if we’re all burning tokens anyway, we might as well compa…
Claude Code vs. Codex for end-to-end app development: how are you using both? (www.reddit.com via reddit) I’m looking for honest, non-biased input from people who have spent a significant amount of time using both Claude Code and Codex. For context, I currently pay for both Claude Max and ChatGPT Pro because I’m building B2B SaaS applications…
What is the best way to keep Claude Desktop in sync across two devices? (www.reddit.com via reddit) I have Claude Desktop on two computers, and I find that they are not in sync when it comes to skills, data, sessions, and memory. Sometimes, a skill saved to Claude’s cloud is synchronized, but work completed locally on one computer is not.
Long chats are draining my Claude limits (www.reddit.com via reddit) Hey, I’ve been using OpenCode CLI, Claude CLI, Claude Desktop, Codex CLI, and generally messing around with AI tools for quite a while now. Claude is probably the tool I have the least real-world experience with so far, but I want to test…
Codex doesn't give you more usage than Claude (www.reddit.com via reddit) i always hear ppl say that codex is a better value for your money but that is not true! at least from my experience claude (i use cowork, not claude code) at ultra gets much more stuff done that codex at ultra before both hit limit and i'm…
Top 15+ MCP servers that are actually useful in 2026? I’m tired of fake awesome lists (www.reddit.com via reddit) I’m trying to clean up my MCP setup and honestly I’m lost. Every best MCP servers list looks like SEO garbage now.
That Mode isn't available right now - Auto mode switch in claude app (www.reddit.com via reddit) I'm coming from codex and want to work additionally with claude on some tasks and figured out I can use the claude android app and code option directly to do so by connecting it via ssh to my main computer, no need to use apps like terminu…
Workflow that I found works best with claude-code opus 5 (www.reddit.com via reddit) TLDR; skills are dead, long live md file hierarchy I've been making a project from scratch that uses a supabase, react, express, node set up. After the first hour and burning my free $100 credits having fable organize the mess codex starte…
Cursor Agent Window vs Codex for Vision AI work — has anyone switched? (www.reddit.com via reddit) I’m a Vision AI developer at a startup, currently using Cursor Pro+, mostly through the Agent Window rather than Tab/autocomplete. I’m considering cancelling Cursor and moving to Codex, so I’m mainly trying to compare the agent workflows—n…
I moved my entire law practice to Claude. I've had over 5100 conversations with it this year. I regularly hit my $200/max limit weekly. I'm about to abandon it. (www.reddit.com via reddit) Claude changed my life. I moved my law practice to it.
Shared skills core + adapter so I stop re-copying the same workflows into ~/.cursor (www.reddit.com via reddit) I was running the same SDD / stack workflows in Cursor and Antigravity and it worked well in each tool on its own. The pain was maintaining two projects.
How to create a platform agnostic repository of skills/workflows? (www.reddit.com via reddit) I use both Claude Code and Codex regularly. I'm slowly building up a repository of skills and workflows that I want to use in both without having to put them in two different places.
A second AI model is not automatically an independent code reviewer (www.reddit.com via reddit) I found a paper on Hacker News that tested a workflow a lot of us now use: one coding agent writes, another reviews. The experiment used 116 medium and hard LiveCodeBench tasks across solo, same-model, and cross-model conditions.
how do you keep track of what your Al agent actually changes? (www.reddit.com via reddit) I've been doing a lot of vibe coding with Claude Code and Codex, and one thing keeps happening I ask for one small change, then later realize Al changed my code in places I never expected. By the time I notice, I can't remember exactly wha…
I built an open-source browser UI for Claude Code — same agent underneath, plus live charts/diffs and your session on your phone (www.reddit.com via reddit) Problem: I run Claude Code basically all day, and the agent itself is incredible. But it lives in a terminal, so everything it produces comes out in basic markdown.
I exhausted my cursor $20 plan in 15 days. Should I upgrade to $60 or switch to "On-Demand Spending"? (www.reddit.com via reddit) I have Google AI Pro(got it free for 18 months with sim), but it's mostly worthless for tuff tasks and the claude tokens in it burn down within half an hour. I have a ChatGPT Go plan for free with the sim, but that is also exhausted for th…
AI orchestration for Claude Code (task routing + Codex execution) (www.reddit.com via reddit) I built these after repeatedly running into the same problem with AI coding workflows: we tend to treat one model as if it should plan, implement, review, and verify everything. That works for small tasks, but it doesn't scale well.
I got tired of walking back to my Mac to approve what Claude Code wanted to run, so I built a phone app for it (www.reddit.comhttps) I wanted to leave my desk without losing track of what my AI agents were doing. So I built Skyhelm.
Claude reviewing Codex's code lifted the pass rate from 71.6% to 89.7% (leaddev.com via reddit) You have 1 article left to read this month before you need to register a free LeadDev.com account. Estimated reading time: 3 minutes Key takeaways: - Reviewer hierarchy among AI-coding agents beats having a second opinion at all: the wrong…
Xberg: a document-extraction plugin for coding agents (www.reddit.com via reddit) I maintain xberg, a local document-extraction framework (Rust core, MIT). I packaged it as a plugin that installs into most coding agents, so the agent can read documents at runtime instead of you preprocessing them.
New ways to learn and teach with ChatGPT Work and Codex (openai.com) could not extract summary
Greenroom: your coding agents form a standing team, name themselves, message each other, and wake each other's idle sessions (Claude Code + Codex) (www.reddit.com via reddit) I've been running multiple coding agents across Claude Code and Codex and got tired of them being strangers with amnesia. Greenroom is the fix I wanted (and a fun exploration): a small self-hosted server where agents hold persistent identi…
What’s the highest-intelligence coding agent per dollar besides Codex? (www.reddit.com via reddit) I already have ChatGPT Pro and use Codex heavily. I’m looking for the best additional coding agent not another way to access Codex.
been switching between Cursor and Claude Code for months, this actually fixed it (www.reddit.com via reddit) Been bouncing between cursor, claude code and codex for months. every time i switch its the same 20 min of re-explaining my setup, pasting rules, reminding the agent why we ditched some approach last week.
Claude Code spent 40 minutes ruling out an approach. Codex suggested the exact same one 2 hours later (www.reddit.comhttps) claude code spent 40 minutes tracing a race condition in our event bus, ruled out a caching approach because of how the subscriber lifecycle was wired, and moved on. 2 hours later I switched to codex to write tests for the same module.
Codex is definitely superior in my specific workflow apparently (www.reddit.com via reddit) So I spent the last 2 weeks fighting with Claude on my project, I don't know why, but I just cannot keep it in check, no matter the model, or effort level, it repeatedly just goes off on it's own thing, even If specifically stated, logged,…
Does anyone else's prompts totally fall apart moving from Cursor to Claude Code? (www.reddit.com via reddit) Genuine question I keep noticing the exact same prompt that gets clean results in Cursor produces garbage or needs 3x more back-and-forth in Claude Code or Codex. Started paying attention and it seems like each model actually wants differe…
Different caching strategies - Codex and Claude Code (www.reddit.com via reddit) I am working on agentic software generation, and I noticed that Codex is more efficient than Claude Code in terms of token usage. The two agents seem to take very different approaches to caching, which might explain the gap.
Reopening of r/ChatGPTCoding (www.reddit.com via reddit) Hello everyone! r/ChatGPTCoding is open again with a new moderation team.
I switched to sonnet 5 and now my max sub is unlimited (www.reddit.com via reddit) A lot of people have been criticizing Sonnet 5 lately, especially with all the talk about GPT Luna getting a price cut. I actually haven't used Sonnet in the last 3 months, not even Sonnet 5 earlier this week.
I've been running Claude Code as the orchestrator for a fleet of other agents for about five months. Open sourced the whole thing today (MIT) (www.reddit.com via reddit) A brain dump, because I think the Claude-specific part is the bit that's actually useful here. The problem I had was never Claude.
Codex No diff available (www.reddit.com via reddit) Codex extension (tried release and pre-release versions), can't review changes and view diff. Always error "Oops, an error has occurred" and button "Try again", after click - "No diff available" Does anyone know how to fix this?
So, is Opus 5 or Fable better for long-context orchestration now? (www.reddit.com via reddit) I’m working on some heavy, long-context data science and ML model development. For the past month I’ve been using Fable as my architect/orchestrator, with two key orchestration threads “overseeing” roughly 20 other threads across primarily…
what are you all using for ai agent code review right now (www.reddit.com via reddit) so we're small team, 6 devs. we all use claude code, codex and cursor, composer 2.5 as the worker under fable.
When it is Saturday, your Claude, Codex and Ollama quotas reset on Monday, no kids weekend, and you have saved usage for the whole week (>75% available on all). (www.reddit.comhttps) Anybody else does something like this ? I tend to save the big guys (Claude and codex) for the last days of the week, i use cheap models for most of my day to day work (now deepseek v4 flash) and save my quotas with the smart models for th…
The main reason I use Cursor (www.reddit.com via reddit) The main reason I use Cursor is that it works great with Windows Server. I have a MacBook Pro M5, but I also use a Windows Server PC at the office, so I need something that works well on both.
How to disable harness specific tools in cursor? (www.reddit.com via reddit) I recently switched back to Cursor and I’m loving it, Grok 4.5 is great. The problem is that I have instructions set up for Claude Code and Codex, and I get the feeling Cursor is somehow picking them up too (CLAUDE.md in ~/.claude and AGEN…
Which model is best for psych dissertation project (www.reddit.com via reddit) Hey so I know a bit of coding but I’m not an engineer and using vibe coding to make a product for my dissertation in psychology. I plan on I’ll hiring an engineer to spend some time auditing and fine tuning before deploying in my actual st…
Those of you running 2+ agent tools (Claude Code + Cursor etc.) — how do you stop them from stepping on each other? (www.reddit.com via reddit) My team ended up with people on Claude Code, Cursor and Codex, all working the same repo. Last week two agents edited the same file within an hour of each other and we only caught it at the merge.
I stopped switching between Claude Code and Codex. Now they run in the same terminal and one supervises the other (www.reddit.comhttps) one terminal. two main ai agents.
Opus 5 now has fast mode in subscription like codex! Am i late to the party or seriously no one is aware? Didnt see it in the news .. SCREENSHOT ATTACHED AND IT WORKS .. i have 200$ subscription .. note: says opus only (www.reddit.com via reddit) https://preview.redd.it/aha9ewsf3ngh1.png?width=274&format=png&auto=webp&s=4b70c56a1c2a47a7f2fa7a402511081fe0854630 The reason this is interesting is that we had fast mode before but it was extra usage only. Opus 5 is already faster than 5…
What I found when I checked whether code quality metrics actually predict bugs (www.reddit.com via reddit) If you're using claude code or even cursor or codex, you know the feeling that features work, tests pass, but the codebase still feels messy Skill files and CLAUDE.md don't fix this. They hold conventions you wrote down once, and go stale…
Built a compiler so Claude Code plugins also work natively in Cursor, Codex, OpenCode, and 19 other harnesses (www.reddit.com via reddit) https://preview.redd.it/hv940qeyuigh1.png?width=1200&format=png&auto=webp&s=fed3dff1f169bb53866861b4fee795cf9e83753c I write Claude Code plugins skills, hooks, subagents, commands and kept hitting the same wall: every harness Codex, Cursor…
Reason difference between apps (www.reddit.com via reddit) A few months ago, I remember seeing a post on this subreddit talking about how the mobile ChatGPT app, compared to the desktop ChatGPT app and the web, puts different amounts of thinking "juice" in. It said that the web's high extended thi…
OpenAI Codex Charged Me Hundreds and Refused a Refund. I am also a teacher and dont get paid over the summer. Im so upset at myself. (www.reddit.com via reddit) I am a teacher, I was thrilled when I saw that Codex was a thing. I was using it heavily for the last week.
Three months of building with coding agents: ~125B tokens processed, ~430M generated. Notes on whether the code is any good. (www.reddit.com via reddit) Since May I've been building a commercial project (e-commerce, mostly PHP and TypeScript) almost entirely with two coding agents: Claude Code and Codex CLI. Before writing opinions I counted what actually happened, from the session logs.
Insights command source (www.reddit.com via reddit) Does anyone have the /insights commands source and/or prompt? I need to use codex for work and would like to see the result for my work-work in addition to my personal stuff.
Counter Strike 1.6 on Unreal Engine 5 (www.youtube.com via reddit) - Reverse Engineering of cs 1.6 binaries (bought on steam) - Make exporters of bsp, mdl, etc resources to belnder files with animation, skinning, textures and etc. All models execpt hard surfaces received 2 subdivision modifiers simple + c…
Claude Code Has Subagents. Should Anthropic Add a Native Dependency-Aware WBS? (www.reddit.com via reddit) I use Claude Max 20x alongside ChatGPT Pro, and I’ve been experimenting with Claude and Codex agents working from a shared dependency-aware WBS. The encouraging result was not “more agents.” It was giving every agent only: Its exact object…
SKI: Voice coding & Meeting connector for Claude Code - Free on Mac & Windows | Built using Claude Code | Fully on device (www.reddit.com via reddit) I was using Whisperflow and it required a subscription and wasn't working well with the intended use of hands free coding. I found it to be a STT with LLM correcting things.
I tested 5 popular token saving methods across 10 real tasks, and none cut total tokens in both runs. (www.reddit.com via reddit) TL;DR I compared 5 popular token saving techniques (+ cheaper model) on 5.6 Sol across ten real coding tasks from my repo. I repeated all seven arms (6 + baseline) twice for a total of 140 agent runs.
Built a local proxy so my Claude Code session doesn't die when one account hits its limit (www.reddit.com via reddit) If you're running more than one Claude subscription (or a subscription plus an API key), you've probably hit this: you're mid-task, Claude Code throws a usage-limit error, and now you have to stop, log out, log into the other account, and…
I forked OpenAI's codex-security to run on Claude Code, so it reuses your Claude subscription (www.reddit.com via reddit) Hey, I just forked OpenAI's codex-security tool and made it work with your claude sub. Grab it here if you want to benefit without switching your subscription: https://github.com/presmihaylov/claude-security
Cursor Start used 10% of my monthly allowance on one repository audit and remediation attempt (www.reddit.com via reddit) I tested Cursor Start on an existing Next.js project using Cursor Grok 4.5 Medium. The task consisted of two stages: Audit the repository for security issues.
How should working with AI feel in the future? (www.reddit.com via reddit) For people who use AI tools like Codex, Cursor, Claude Code, Copilot, or ChatGPT for real work: What is frustrating about working with AI today? Is chat the best interface, or do you wish AI work felt more like a workspace where you can se…
How do you handle context when moving a task out of Cursor and into another coding agent? (www.reddit.com via reddit) I use Cursor for some tasks, but not every task. Sometimes I’ll start in Cursor, make a few changes, leave the working tree dirty, and then want to continue in Codex or Claude Code.
Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI (www.latent.space) There are roughly 100x more people who use code than who can write code.1 As code that “just works” becomes easier to generate, this group may be the biggest prize of all — if you can get the agentic interface right. A key trend we have be…
Claude got more useful when I stopped asking it to research and judge in one pass (www.reddit.com via reddit) I kept getting research summaries that sounded confident but quietly mixed sourced facts with assumptions. The change that helped was pretty simple: I stopped asking one model to discover information, judge it, and write the final answer i…
Thank you for STEERING, finally! (www.reddit.com via reddit) You can now send a message to Claude when using Claude Code to steer instead of having it be queued, thank god... I use this all the time.
I canceled Max plan, switched to Grok 4.5 for coding, then switched back (www.reddit.com via reddit) Hey just wanted to share my experience of what happened over the past week. Basically I was having huge huge huge problems with Claude for the week before they released Opus 5.
Multiplayer Destruction Game I made With Claude (www.reddit.comhttps) Making a new multiplayer PVP desctruction game with three.js, made almost exlusively in Claude with a little codex here and there when my tokens ran out. My last game was a destruction game with multiplayer physics, but it was a rhythm gam…
Could typed, modular agents make Claude Code easier to reason about? (www.reddit.com via reddit) I came across this open-source framework called Atomic Agents, and the design philosophy seems relevant to Claude Code users. Most AI agent frameworks can give you the illusion of control.
I run Claude as a PM over Codex and Gemini workers — and no agent is allowed to declare "done" (www.reddit.com via reddit) This started from a simple observation: agents are great at judgment and terrible at discipline. Every rule I enforced through prompts ("don't poll", "don't claim completion") eventually broke.
Claude power users - how do you do knowledge management using tools like Notion/Obsidian/Jira etc? (www.reddit.com via reddit) So, I am a power user of Claude/Codex myself and I would use Notion and to some degree Obsidian in the past but since last year, my workflow has revolved around claude code and so I haven't used external knowledge tools and instead built m…
Built with Codex, designed for Claude Code users: local evidence reports (www.reddit.com via reddit) I built AIEvidence with Codex, and I'm sharing it here because it is designed for Claude Code users too. What it does: AIEvidence is a local CLI that reads a Git project and optional Claude Code conversation exports, then produces: - a det…
Cross-Model LLM Code Review: Should you use Claude to review Codex or vice versa? (arxiv.org) Developers increasingly use two coding agents together: one writes a draft, and the other reviews it. However, it is not clear whether the pairing is worth its cost and time, or whether the order of the pairing matters.
Does Claude Pro keep normal chat usage separate from Claude Code and Cowork? (www.reddit.com via reddit) I’m considering cancelling ChatGPT Plus for a month and trying Claude Pro so I can properly test Claude’s models and subscription limits. One thing I really value about ChatGPT is that Work and Codex use an agentic usage allowance.
I built a desktop workspace for developers who switch between ChatGPT, Codex, Claude Code, Cursor, and other coding agents (www.reddit.com via reddit) I use different AI coding tools depending on the task, but the biggest problem is not the subscription cost. It is losing context every time I switch.
Does $100 dollars a month take you farther on Cursor or CC/Codex? (www.reddit.com via reddit) I’ve never switched off of Codex but am exploring my options after rate limit drops and I’m wondering whether I should switch. I’m just wondering whether $60 plans + $40 in credits or whatever is better on Cursor than the $100 on CC or Cod…
what agentic coding tools actually stuck for your team? (www.reddit.com via reddit) what agentic coding tools actually stuck for your team? we're a 12 person product team and our setup is cursor + codex + claude code + coderabbit.
Curious how other infrastructure/platform engineers are using AI agents (Claude Code) in their day-to-day work. (www.reddit.com via reddit) Curious how other infrastructure/platform engineers are using AI agents (Claude Code, Codex, etc.) in their day-to-day work. We're at a GPU compute hosting company and have connected our internal tools (Grafana, NetBox, internal APIs, etc.…
has anyone actually replaced claude as their main ai coding agent (www.reddit.com via reddit) my loop is fable 5 or opus 5 planning, composer 2.5 executing, coderabbit / bugbot on review. it works, i freelance so the code has to be safe.
I let Claude build itself (387 PRs later): a crash-only harness where you swap the worker model and the orchestrator model independently (www.reddit.com via reddit) TL;DR: Aesop is a multi-agent coding harness, built mostly by Claude running on itself. As of 0.4.0 it has two "seats" you point at any model from one config block — the seat that writes code, and the seat that decides whether to ship it.
sandbox-cli is now in public beta 🚀 (www.reddit.com via reddit) Run Claude Code, Codex, Gemini, Cursor, Aider and 10+ other coding agents with full autonomy — inside a disposable Docker container. Only your project is mounted.
Need guidance to take it to the next level (www.reddit.com via reddit) Hi all, Claude is my go-to for the business brain work (BA, Business documentation, roadmap, release planning, etc...) that precedes the code work. I do the design, research and the UX that leads to the product design.
I’ve come to depend on resets…. (www.reddit.com via reddit) Didn’t count on our guys quietly dropping the EXPECTED reset for this new model drop. I know there’s no explicit declaration from Anthropic that they reset usage on new model drops, but cmon… it’s the little digs like this that push away p…
Why did Anthropic not extend Fable till today for Pro users? (www.reddit.com via reddit) So, Anthropic decided to not include Fable for Pro users after the 19th. Fair enough, it's an expensive model.
OpenForgeRL: Train Harness-native Agents in Any Environment (arxiv.org) Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While powerful, these complex harnesses also make agents hard to train…
I built a router that spreads work between my Claude and ChatGPT subscriptions (www.reddit.com via reddit) I made Alloy'd, which is an MCP server and hooks that allow Claude Code and Codex to dispatch substantial work to whichever side has more usage remaining. It does this by calling the official headless interfaces (claude -p and codex exec).
I built an MCP server that syncs knowledge, skills and hooks across Claude Code, Codex CLI, Kiro and OpenCode (extensible to any MCP-capable agent) — not just a wiki (www.reddit.com via reddit) Between work and hobby projects I use Claude Code, Codex CLI, Kiro and OpenCode — I like all of them and often switch between them depending on their strengths and availability. The problem, though, is always the same: every session starts…
NTT DATA Group cuts incident analysis to 30 minutes with Codex (openai.com) could not extract summary
I built a Mac menu bar and notch app to manage agents sessions, usage accounts (www.reddit.com via reddit) built this with claude code, for claude code. flagging up front that its my own project.
I built a small Claude Code and Codex usage tracker for Windows (www.reddit.com via reddit) I often switch between different IDEs and tools that use Claude and Codex. Sometimes I use both at the same time in one workflow, and constantly checking how much usage I had left on each one became pretty annoying.
i built a open source trust kernel for coding agents as a tiny preview of my own harness that is in progress. (www.reddit.com via reddit) Agents lie about tests and occasionally delete things they shouldnt, anyone whos run one long enough has seen both. LIA Trust Kernel is a small Rust binary that sits at the tool boundary of Claude Code (PreToolUse hook) or Codex (MCP) and…
Switching from Claude Code to GPT-5.6 Sol, what am I actually going to miss? (www.reddit.com via reddit) I’ve been using Claude Code heavily for day-to-day backend/infra work: multi-service repos, debugging, refactors, Terraform/K8s, and LLM-related services. I’m considering making GPT-5.6 Sol in Codex my primary tool.
A 30 year old aerospace practice lets me hot swap Claude and Codex sessions with zero onboarding. (www.reddit.com via reddit) I'm part of a team developing a renewable energy site screening tool, and I decided to run the project like an aerospace program. Every requirement, test case and design decision is its own small YAML or markdown file in git, with typed ID…
Once all boosts are gone and now Fable 5 is unreachable for most, are you going to stay? (www.reddit.com via reddit) 50% and 100% boosts are amazing, but I feel like reverting now would hurt everyone big time. Especially with all the other offerings.
How to "steer" a sprint in Claude Code the way I can do on Codex? (www.reddit.com via reddit) There's a very useful feature in Codex that I love. I'm often running Codex and Claude in parallel, so I'd love to have the same feature in CC.
I built an MCP server so Claude Code can delegate work to GPT-5.6, DeepSeek, GLM and a local Qwen — then benchmarked all of them against Claude itself (198 runs, hidden tests) (www.reddit.com via reddit) Same idea works for any MCP-capable agent — the point is you can hand tasks to other companies' models without ever leaving your main app. Before anything else: I did all of this for my own testing, to make my own decisions about my own se…
What’s your Cursor workflow, and which models do you use for each part? (www.reddit.com via reddit) I’m curious how everyone divides work between ChatGPT, Codex, Cursor, and the different models. My current workflow: I start by working through the feature or problem inside a ChatGPT Project, where it already has the broader context.
I built forklane — fork a live Claude Code session to Codex (or any CLI agent) in a git worktree, then merge its work back (www.reddit.com via reddit) Mid-session, I often want a second opinion from a different model without losing my Claude Code context. So I built this for myself and open-sourced it today.
I love Claude. I hate Claude Code (and other "agents"). Is this anyone else's experience? (www.reddit.com via reddit) Im 24 years old. My job is medical AI research at a hospital and I use Claude for this every day (I have Max and use Claude ~6 hours a day).
People are running to Codex, me: "Tell him it's too late. I'm part of the ship... part of the ship, part of the crew..." (www.reddit.comhttps) could not extract summary
Give Claude a real phone as a body: open-source MCP server, 62 tools (Android + iPhone) (www.reddit.comhttps) You can give Claude hands on a real phone. Ghost in the Droid is an open-source MCP server (MIT) that turns a real Android phone, and now an iPhone, into a tool surface Claude can drive through Claude Code or Claude Desktop.
An idea: your AI remembers you, even when you switch AIs (www.reddit.com via reddit) Most of us aren't using just one AI anymore. You're on Claude until you hit a limit, then you switch to Codex, then maybe back.
Do you use MCP outside of work? (and a question about MCP for everyday apps) (www.reddit.com via reddit) I recently added AI integration to a small home inventory app I build, and ended up shipping it in two modes, because "AI integration" turned out to mean very different things for different users: Mode 1 — one-tap prompt sharing. The app p…
5.6 SOL LOVES sub agents (or the code-review plugin) in Claude Code - burned through weekly usage in 15 minutes on a simple change (www.reddit.com via reddit) TL;DR: 5.6 SOL sent total 166 agents, 5 layers deep, just to check the work of a simple change. Seems like GPT loving sub agents a little too much...
I built loom — a free generative MIDI instrument for macOS that grows a whole evolving ambient electronic track from a single seed (www.reddit.comhttps) Have always been keen on generative electronic music and like playing around with Ableton. Thus, I have (along with Claude and sometimes Codex) been building loom, a native macOS app that generates a complete, continuously evolving arrange…
Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3? (arxiv.org) Our previous ARC-AGI-3 agent bundled executable world modeling, scheduled simplification, and exact replay verification, leaving unclear which idea accounted for its performance. We address this attribution question with four nested Codex-…
Animated SVG comparisons between several models (www.reddit.com via reddit) I have seen some people testing models by telling them to generate images of difficult, unusual SVGs, and I thought: what if I elevate difficulty a bit and specify that it also has to be animated, and perfectly looped? I have tested Haiku…
Coding Workflow Critique Requested (www.reddit.com via reddit) I am building software using Claude Code... I am not a software engineer by any stretch.
Skill that makes a repo look pro FOSS in no time. (www.reddit.com via reddit) I vibecoded a tool which scans a repo and generates the boring-but-decisive open source files: LICENSE, CONTRIBUTING, CODE_OF_CONDUCT, SECURITY, CHANGELOG, .gitignore, issue/PR templates, CI, etc etc. It's an agent skill (works with Claude…
Codex being so accessible was a move to rattle Anthropic and they failed. (www.reddit.com via reddit) they wanted Anthropic to make fable accessible for pennies, something that would put a lot of computation pressure on them and probably lose them money. OpenAI is failing they absolutely can take these big risks to become relevant again bu…
Built a shared memory for AI coding tools because my team kept re-explaining the same decisions to Claude Code every session (www.reddit.com via reddit) We kept hitting this on my team: someone makes a decision in their AI session, and two days later someone else's AI gets confused because it never saw that decision. So I built Wayform, a shared memory layer that Claude Code, and any other…
Should "plan first" mode become a real lock, not just an instruction? (www.reddit.com via reddit) I am curious how people handle this with Claude Code or similar terminal agents. Sometimes I want the agent to plan first, or slow down, or stop before it keeps expanding the task.
Cursor is now agent UI like Codex? (www.reddit.com via reddit) Hi folks, I was using Cursor always as an IDE to review code my Claude emits. But today I opened cursor in the CLI, and it now opens an agent UI that looks exactly like Codex?
Claude Code workflow: logging every tool call to JSONL + replay when something looks wrong (www.reddit.com via reddit) **Built with Claude** — solo project, public alpha, Apache 2.0. **Problem:** When Claude Code (or Cursor) runs tools, most setups leave no trail you could hand to an incident responder.
$1000/mo Super Max Plan? Would it be a good addition? (www.reddit.com via reddit) I really prefer using Claude Code, the model is genuinely stronger than gpt 5.6. But the limits on the strongest model is so low that now I spend 80% of my time working in codex just to use the next strongest model rather than downgrade to…
OmniDesk v2.4.0 — on-device voice prompting: talk to your Claude Code terminal, 100% offline (www.reddit.com via reddit) OmniDesk is a desktop app that hosts your AI coding CLIs (Claude Code, Codex) with a proper multi-repo / multi-session UI. New in v2.4.0: voice prompting (speech-to-text) — and it runs entirely on your machine.
OpenLive, open-source alternative to ElevenLabs Agents and Gemini Live. Now talks to coding agents like Claude Code using your regular plan, no API keys, no API bills. (www.reddit.com via reddit) A while back I posted OpenLive here. It's an open-source voice layer that gives any AI model or agent ears, a mouth, and eyes.
Just got only 2% added to the usage limit on Codex. What is this about? (www.reddit.com via reddit) I had exhausted my weekly usage limit yesterday. Today I saw "2% usage remaining".
Got my Codex Micro - love it so far, but wish it had a mic (youtu.be via reddit) About Press Copyright Contact us Creators Advertise Developers Terms Privacy Policy & Safety How YouTube works Test new features NFL Sunday Ticket © 2026 Google LLC
Is Claude Code on desktop still worse than CLI? (www.reddit.com via reddit) I just got a Max sub after being on Codex for a while, curious if the outputs generated from the CLI are still noticeably better than the desktop app. Where are most people working?
Made claude and codex argue over my code before anything merges, sharing the skill! (www.reddit.com via reddit) I'm a founder who hadn't coded in years before Claude Code. Now I ship weekly.
I couldn't afford the $230 Codex Micro, so I vibe-coded a digital one that runs on any phone/tablet and actually talks to Claude and Codex (www.reddit.com via reddit) So OpenAI dropped the Codex Micro - a $230 little macropad with pretty RGB keys that lights up when your agent needs you - and every influencer on my timeline immediately declared it "the first piece of OpenAI hardware" and "the next big t…
So Much Peace and Quiet Compared to the Claude Sub (www.reddit.com via reddit) Do you guys remember Codex suddenly switching to API pricing in the middle of the day? Yeah, me neither.
WTF DID I PAY $100 FOR?? (www.reddit.com via reddit) https://preview.redd.it/4qo03jtzctdh1.png?width=891&format=png&auto=webp&s=995851b5ed82641625b5e2ac82f886710a9584f8 Tf you mean "Selected model is at capacity"?? I literally just paid $100 and upgraded to max just to use Sol 5.6.
bad option: single business user (feedback) (www.reddit.com via reddit) After already using claude pro and copilot pro, I also wanted to get Codex now. However, there is something that really bugged me when I wanted to buy.
The visual control plane I want: Claude coordinating GPT and external agents without losing accountability (www.reddit.com via reddit) I’m excited about Claude’s models and the multi-agent architecture I’ve been testing with Claude and Codex. I use Claude Max (20x) and ChatGPT Pro ($200/month; 20x usage), so this is not a free-tier comparison or a “Claude versus ChatGPT”…
My Fable vs Sol experience (www.reddit.com via reddit) I have both Claude Max and GPTPro subs and use both extensively (ClaudeCode and Codex) in my daily work. What I noticed is that for Sol you still need to babysit it more than Fable.
Significantly lower value in Cursor subscription! (www.reddit.com via reddit) https://preview.redd.it/tm96xmhhtrdh1.png?width=1332&format=png&auto=webp&s=94ef5621aeaa0fb86be455c942a433dad53af348 As seen in this comparison table, Cursor subscription's API pricing equivalent usage value is quite low compared to Codex…
How I cut Claude Code session context 67% with a local vault (www.reddit.com via reddit) I run several AI coding agents (Claude Code, Codex, OpenCode) across a few solo products. My problem: every session either lacked context or I pasted a giant knowledge file and wasted most of it.
Using Claude for multiple projects at once (www.reddit.com via reddit) Since Claude and Codex supports multiple windows running at the same time how many projects on average are you guys working on at a time? Do you like to focus only on one project one task and get your thoughts aligned?
pixtuoid update — the office has weather now, everything is more alive (www.reddit.comhttps) quick update for anyone who saw pixtuoid before — a few fun things got added recently: 🌤️ there's weather now — sun and moon move across the windows as the day goes, and it gets foggy/stormy/clear, changes with the theme. 🔥 agents running…
I made some iPhone widgets to make Cursor’s usage limits easier to see (www.reddit.com via reddit) I know we're all tired of the rate limit monitors or apps but I wanted a clean, minimal way to view them without making it obtrusive. It connects to your account completely on device, no mac app or tethering required.
If they remove the editor view Im cancelling (www.reddit.com via reddit) I noticed recently they've made the agent view the default window when opening cursor, and I have to manually switch it to the editor view every time. First of all, can anyone tell me if there's a way to make it default to the editor view?
Minimal Open-Source Evals Platform (SUGGESTIONS NEEDED) (www.reddit.com via reddit) I like knowing where my tokens are going. I also like knowing how I can improve my workflow and having access to the fine details of my back and forths with my clankers.
Bring Codex Micro to any gaming controller and coding harness (www.reddit.comhttps) Hello everyone, I've been experimenting with using a gaming controller to interact with my AI coding agents for a while now. Since OpenAI announced Codex Micro, I'm open sourcing OpenMicro which brings Codex Micro’s functionality to gaming…
What's the best practice to re-use agent skills? I'm using both claude and codex. (www.reddit.comhttps) Currently, I'm just copying my agent skills from .agents/skills to .claude/skills. But I'm sure there must be a better way to keep them in sync.
How Codex became a collaborator for OpenAI’s creative team (openai.com) could not extract summary
I built rein.build so I could control my Claude from my pocket (rein.build via reddit) Your AI agents already run on your machine. Rein puts them in your pocket.
Claude Pro subscription vs. API access for planning? (www.reddit.com via reddit) I want to use Claude model for planning a project I'm building, execution will be on Codex. It will be probably a planning for 30 minutes total usage on their best model.
I keep losing useful answers inside old AI coding sessions. How are you searching them? (www.reddit.com via reddit) I have a very boring AI coding problem: I forget where the useful answer was. I have re-asked the same coding question because I remembered solving it in an old AI session, but I could not find the exact answer again.
Trust Issues: an open-source skill that vets an AI skill/repo/MCP/package before you install it (www.reddit.com via reddit) I built Trust Issues to solve a problem I kept hitting: installing an agent skill or package runs someone else's code with access to my files, and there's no easy "check this first" step. It points your agent at a repo/skill/MCP/package an…
Opus 4.8 making errors that drive it to a nervous breakdown (www.reddit.com via reddit) I have been using opus models for months now, seen every "model x lobotomised" post and frankly it has always performed well for me, sometimes inconsistent but that is the reason you build a harness and some infrastructure to protect again…
Advice (www.reddit.com via reddit) Backstory - not a coder/programmer but have long enjoyed using ai. With release of cowork/code desktop that enjoyment led to building a personal productivity system that now led me to building a personalized/tailored md editor for it to re…
My coding agent installed `loadash`. One typo away from a supply-chain nightmare (www.reddit.comhttps) I have been letting Codex run pretty loose on a SaaS build because I am not deeply technical. Last week it installed loadash instead of lodash, and I caught it way too late.
Review my AI coding workflow (www.reddit.com via reddit) I'm trying to design a simple, production-friendly workflow for AI coding agents (Claude, Codex, Cursor, etc.) and would love feedback from people using them daily. Current workflow: ``` Human → Define feature AI Agent • Understands the ex…
Goblin the Creator (IGOR) (www.reddit.com via reddit) Asked Claude if it could make a Tyler the Creator version of /caveman. It went full Goblin mode - this must have been what Codex was warning us about.
Agent Hacks Agent: Autoresearch for Production-Agent Red-Teaming (arxiv.org) Production LLM agents such as Claude Code and Codex operate over untrusted content, files, commands, and workspace state, making safety failures directly actionable. Red-teaming must therefore keep pace with evolving models and tools.
[AINews] Codex usage up >10x in 6 months to 7M users, +1M in the past ~day; did Codex overtake Claude Code?? (www.latent.space) [AINews] Codex usage up >10x in 6 months to 7M users, +1M in the past ~day; did Codex overtake Claude Code?? a quiet day lets us fact check some numbers against the sound of silence of Claude Code reporting...
How companies are allocating AI IDE tools? (www.reddit.com via reddit) I'm planning to standardize AI tooling across the organization . Currently , our teams are using a mix of Claude Code , Zed , Codex , and different tools.
The biggest difference I have noticed between Claude and Codex/GPT5.6 (www.reddit.com via reddit) The biggest difference I have noticed between Claude Opus/Fable and Codex GPT5.6/any model is Codex seems pretty content to just waste time looking like it is doing things without actually doing things. It does not seem to be outcome-orien…
Comparing 2D-to-3D: Fable 5 vs. GPT-5.6 Sol (www.reddit.com via reddit) So I decided to ask Claude and Codex to convert my 2D grid-world game into 3D. The game is about cars that travel from point to point along predefined routes and need enough fuel to reach their destinations.
Windows users: how do you keep an eye on your Claude/Codex limits without hitting the wall? (www.reddit.com via reddit) I run Claude Code and Codex side by side on Windows, and I keep finding out I'm rate-limited the moment I hit the limit — usually mid-task. /usage and /status work, but only when I remember to ask, inside a session.
Claude Code hid timezone and proxy checks. what else can coding agents read (www.reddit.comhttps) I used to let Claude Code break coursework and repo chores into little sub-agent threads. Not blind trust, but close enough that it felt like a normal dev tool.
unsnooze — auto-resumes Claude / Codex (and other AI CLIs) when the usage limit resets (www.reddit.comhttps) I kept running into the same problem with long Claude Code sessions. I'd start a big refactor before bed, wake up, and find it had stopped a couple of hours later after hitting the usage limit.
Best way to move my Claude Code setup to a new Mac mini? (www.reddit.com via reddit) Guys, I’m buying a new Mac mini in a couple of days and selling my current PC. I’ve built a very specific workflow using Claude Code and Codex, including multiple .md instruction files and important local memory/context files.
What are Limits on the 20$ Plan? (www.reddit.com via reddit) Hello, can anyone Tell me specific token Limits fot Grok 4.5 using the 20$ Plan? Im thinking about switching from Gpt plus since Codex Limits are pretty rough right now (they reset a Lot but that will Stop)
Hit your usage limit mid-session? I built a free CLI that moves your session between Claude Code and 8 other agents (open source, Rust) (www.reddit.com via reddit) You're deep in a session with Codex or opencode, and boom: usage limit. You wait, or you start over in Claude Code from scratch.
Will Anthropic also remove the 5 hrs limit ? (www.reddit.com via reddit) Codex just removed the 5rs limit completely and did another reset.... That could never be Anthropic
Coder: Delegate the coding to coder tasks powered by codex/claude cli engines (www.reddit.comhttps) I built Coder, a cli/plugin that lets you dispatch coding tasks to background agents powered by claude cli or codex cli or both! You describe a task, it runs an subagents on it and gets the results, so your main session's context stays cle…
Has anyone tried disabling sub-agents to Fab/Sol? (www.reddit.com via reddit) As we all know, Fab and Sol usage can burn really fast, especially when using tools like Codex and Claude Code. I recently disabled sub-agent mode because I wanted to see how much it affects usage, speed, and final results.
Cursor tips,settings for newbie (www.reddit.com via reddit) I just came from Codex plus And no idea cursor or IDE thing can you guide me esp. For starting settings editor settings
So how many usage resets do you think we'll get this week? (www.reddit.com via reddit) Seeing as Sol/Codex 5.6 is actually pretty good and, in my limited experience so far, comparable to Fable (or comparable enough) - and cheaper to boot - what do folks think Anthropic's response is going to be over the course of this next w…
How restrictive are the usage limits on the regular $20 plan? (www.reddit.com via reddit) Hey all, Over the last few weeks, I've been paying for ChatGPT, including Codex, as well as Claude, but the limits have become much stricter. I'm thinking about getting Cursor too, but before I spend another $20 a month, I wanted to ask ho…
Best coding setup for price-to-performance in Q3 2026? (www.reddit.com via reddit) I’m comparing: $100 Codex with GPT-5.6 Sol High $100 Claude Code with Opus 4.8 $60 Cursor with Grok 4.5 Which one gets the most real work done for the money? What would be your go-to setup with a $100 budget?
Free open sourced local MCP labor hub (www.reddit.com via reddit) Working on this project, it's free and open source, I really needed something to be a single hub for all the connections and made my own solution. GitHub repo: https://github.com/adun-denton/Chinvat Chinvat is a local MCP labor hub for Win…
I got tired of coding sessions dying mid-task, so I made Passation (www.reddit.com via reddit) I was tired of Claude Code or Codex CLI hitting usage limits right in the middle of a file change. The worst part was not the limit itself — it was having to wait for the same agent to come back because switching to the other one felt risk…
Claude Code kept answering from stale training data, so I built it a real web layer (open source, no API key) (www.reddit.com via reddit) Every time Claude Code confidently gave me an API answer that was deprecated two versions ago, I'd think: it should just be able to check. WebSearch exists, but it's shallow.
Has anyone compared ai agent code review tools, here's what i found after testing 3 (www.reddit.com via reddit) Our review queue got bad enough this quarter that I spent two weeks actually testing ai agent code review options instead of guessing, figured I'd share since I couldn't find a real comparison when I looked First option, just asking Claude…
I'm making a website where the internet writes a story one word at a time. It's definitely probably going to go well (www.reddit.comhttps) So I've been wanting to make this for a while, and it's finally happening. It's one story, and the whole internet writes it together one word at a time.
I have 4 AI subscriptions, here's the pros and cons of each one (www.reddit.com via reddit) I have the "pro" plans for the following: 20$ Cursor, 20$ CC, 20$ Codex, and 10$ Opencode. Here's my experience with limits, usage, and usefulness on them.
Please slow down UI changes (www.reddit.com via reddit) Every time I open Cursor, Codex is in a different location, buttons randomly appear/disappear across releases, window arrangement changes, keyboard shortcuts change. Please give us something between no-updates and new-ide-every-morning.
I made a Claude plugin that lets AI's on different machines talk to each other. (www.reddit.com via reddit) I had a few different AI's on different computers running the same project and I wanted to get them to collaborate, so I used Claude to build meshwire. It uses A2A and enables different agents on different machines (or the same machine) to…
Automatic hourly ping to keep your 5hr Codex usage windows rolling 👍️ ( via reddit) could not extract summary
[AINews] OpenAI launches GPT 5.6 Sol/Terra/Luna, Codex becomes ChatGPT superapp (www.latent.space) [AINews] OpenAI launches GPT 5.6 Sol/Terra/Luna, Codex becomes ChatGPT superapp A big day for OpenAI. On any other day, the launch of a surprisingly good/competitive Muse Spark 1.1 from Meta Superintelligence Labs, including, for the first…
Please recommend an alternative to Codex (www.reddit.com via reddit) I'm done with Codex today was the last straw for me. i mainly use it for Python but lately i've been running into too many issues with it losing context, making changes i didn't ask for and requiring too many back and forth prompts just to…
I put Claude Code and Codex together in a repository and asked them to talk and create something together (www.reddit.com via reddit) Codex was faster, and when Claude thought about writing something, he already had the first message waiting. Each one read what the other had done, coded, recorded the changes, and passed the turn.
Fable 5 optimizer (www.reddit.com via reddit) I'm probably not the first to make a Fable 5 optimizer, but this one is Inspired by this Theo deep dive on Fable 5. It's a Claude Skill and Claude.md file for Fable 5 projects to get the most out of it by leveraging codex and sonnet for lo…
Grok 4.5 usage on Cursor’s $20 plan vs Codex’s $20 plan? (www.reddit.com via reddit) Trying to figure out which $20 plan gives me more mileage for daily coding. For anyone on Cursor’s $20 (Pro) plan — how’s Grok 4.5 usage holding up?
Hexana MCP 0.4 — the plugin that gives Claude Code ground truth on WASM binaries, now with Claude Desktop packaging groundwork (www.reddit.com via reddit) https://preview.redd.it/vhtc3zg6m5ch1.png?width=2048&format=png&auto=webp&s=082e52cd899402743ed80c1a6244f15eeb947746 We build Hexana at JetBrains — an MCP server that lets Claude Code read compiled .wasm binaries directly instead of guessi…
I built a Claude Code loop that makes agents prove “done” with evidence (www.reddit.comhttps) Hi, I’m the maker of Superloopy. It’s a free MIT project for Codex and Claude Code that wraps agent work in a small evidence loop: type `loopy <task>`, the agent works against explicit criteria, records proof under `.superloopy/evidence/`,…
Kubagachi, Kubernetes visualizer (www.reddit.comhttps) My new ui tool that I built for Kubernetes, it delivers a Freelens like experience through a web app. I also included this image gen pipeline to make new characters and animated sprites with Gemini.
OmniDesk v2.3.1 — you can now actually drive your Claude Code from your phone (www.reddit.com via reddit) https://reddit.com/link/1uraeda/video/eekd7i7sj3ch1/player OmniDesk is a desktop app that hosts your AI coding CLIs (Claude Code, Codex) with a proper multi-repo/multi-session UI. It's had remote access for a while — it serves its own UI o…
Wanted to share a plugin that will enable Fable to orchestrate work via Codex CLI or Opencode CLI to help keep Fable usage to a minimum by having cheaper models do the grunt work (www.reddit.com via reddit) not a dev by trade, and this is the first thing i've actually released publicly, so be gentle. This is what I have been using to try to help keep claude usage to a minimum for implementation/mechanical work.
I’ve always wanted to know what session or subagent modified a file, so I’ve built strace for agentic sessions - called gaal (www.reddit.comhttps) It’s could be pain to understand why some changes happened to the code, especially to something outside of the scope of a task I’ve got my SKILL.md files nuked several times - because some codex worker decided that they know better the sha…
One skill to let AI agents work as peers: make Codex and Claude talk to each other (www.reddit.com via reddit) I use Claude Code and Codex side by side, and I got tired of copy-pasting between two apps every time I wanted one to pick up where the other left off, do some task, or perform a review. So I wrote a skill that lets them hand work to each…
12 hrs until my usage resets.. looking for advice on the next build (www.reddit.comhttps) I burnt through my first limit in a couple of days the first time around, mostly checking and securing opus 4.8 code for an internal use only client / workflow portal with API into accounting system. By the good graces of the universe I ge…
Senior Dev plugin for Claude code - keeps Opus 4.8 on track (www.reddit.com via reddit) With my F5 access, here is a plugin to help with workflows - free to use in code and later in the week in cowork with v2. Its a tool that helps you keep the repo and the session on task and helps stop those annoying moments when Opus think…
I built an open-source Codex-native job search assistant with LaTeX CVs and ATS checks (github.com via reddit) I adapted the idea of an AI job application workspace into a Codex-native repo. It lets you: - build a grounded candidate profile from your own documents - rank job postings against that profile - generate tailored CVs and cover letters -…
Coming from Codex, it seems like Claude is just burning through tokens doing nothing. How to control it better? (www.reddit.com via reddit) I just started using Claude recently after a year of Codex, and I'm just amazed at how it manages to just burn through your tokens for the simplest tasks with no feedback whatsoever. For instance, I'll give it a simple targeted prompt in a…
OmniDesk v2.1 — plain shell sessions now live alongside your AI coding agents (www.reddit.com via reddit) https://preview.redd.it/sv9a504n9nbh1.png?width=1704&format=png&auto=webp&s=744a44c712b843f250a75560095218e6ddb74238 Just shipped v2.1 of OmniDesk, a desktop app for running AI coding CLIs (Claude Code, Codex, and more via a pluggable prov…
Anyone else got the usage reset in the last couple hours? (www.reddit.com via reddit) Does Cursor provide such resets in general, which are frequent in Codex?
I built an MCP server that catches when coding agents act on stale file reads - here's the problem and what I learned (www.reddit.comhttps) I was running Claude Sonnet 4.6 as the agent inside Antigravity and hit a problem worth sharing, because it turns out to be a general pattern, not a one-off. The agent read a config file, worked for a while, then wrote documentation from t…
Building an AI-coded geopolitical war strategy game, looking for collaborators (www.reddit.comhttps) I want to build a geopolitical war strategy game: real countries, alliances, player-driven economies, military buildup, base building, territory control, and actual politics layered in. I'm building most of it with Codex and Claude, and I'…
I got tired of losing my task when Claude Code hits its 5-hour limit, so I built an open source tool that hands off to Codex automatically (www.reddit.com via reddit) Every week I'd be deep in a task, hit the limit, switch tools, and spend 20 minutes re-explaining everything. So I built CodePass, a terminal harness that runs your agent in a PTY, watches output for rate limits and failures, and switches…
How are teams using Claude Code / Codex in real product workflows? (www.reddit.com via reddit) I’m curious how teams are actually using tools like Claude Code, Codex, and similar coding agents beyond solo developer workflows. A lot of the examples I see are very developer-centric: Markdown specs, CLAUDE.md, planning files, task file…
How I saved 91.4% on LLM token costs and completely bypassed Claude 5-hour rate limits (www.reddit.com via reddit) If you are running AI agents or terminal coding tools (like Claude Code, Codex, or Cursor) on a standard $20/mo subscription tier, you know the absolute pain of the rolling 5-hour usage window. Because these models have zero short-term mem…
AI Psychosis is leigt - Don't be fooled - Sober up (www.reddit.com via reddit) It's probably taken me a full year to overcome AI psychosis, but I think I'm finally sober. I've always been a big fan of Claude and Codex, and still use them to this day.
A Hacker Typer for the Modern Age - Simulate agentic coding in any web browser (workforwatts.com via reddit) Claude Code Codex Gemini autopilot ⧉︎ ⛶︎ Parody. Not affiliated with Anthropic, OpenAI, or Google.
i told claude to revert one component and it wiped my whole working ui. then said we were good to ship. (www.reddit.com via reddit) Solo builder, i vibe code with claude code and codex. the thing that breaks me isn't building the app, it's controlling what the agent does to it after.
Hi All, I have created a security scanner, and would really appreciate your valuable feedback. Its Completely Open Source and anyone can use, modify or even contribute to it, so we can optimise it for the community (www.reddit.com via reddit) I was creating my first SaaS product, but it had few security issues, and when I fixed those, and made other changes, then new security issues surfaced, So I built CodeInspectus — a local-first security scanner aimed at exactly this "vibe-…
OpenSource- Universal Governance Compiler (radustefanescu97-star.github.io via reddit) Hey everyone, I built an open-source CLI called Universal Governance Compiler (UGC). The problem I’m trying to solve: AI coding assistants all expect different rule/config files, so maintaining consistent instructions, approval gates, prot…
Australian Payments Plus moves faster with ChatGPT and Codex (openai.com) could not extract summary
If claude makes so many mistakes how can you trust it? (www.reddit.com via reddit) This is just example of how many times my peerBench system caught claude just skipping or leaving things open and vulnerable.. I have integrated Codex using their official codex-cc or something plugin and built a peerbench review system wh…
Most coding agents don’t fail because they can’t write code. They fail because they start with the wrong map. (www.reddit.com via reddit) I’ve been building SigMap, an open-source grounding layer for AI coding agents, and one assumption I had was wrong. I thought bigger context windows would solve most AI coding problems.
How are you regression-testing agent workflows before users find the failures? (www.reddit.com via reddit) Curious how people here are testing AI agent workflows (Claude [Code], Codex, Cursor, etc) once they become more than a prompt. I mean the layer around the model: repo instructions, skills, MCP/tool setup, memory, hooks, guardrails, and th…
My AI coding agent kept confidently writing broken Godot code after I upgraded to 4.7, so I built something for that (www.reddit.com via reddit) Anyone else run into this: you upgrade your Godot project to a new engine version, then ask Claude Code, Copilot, or Cursor for a change, and it writes code based on how things used to work? Not an error, not a crash, just...
How to evaluate a skill for building better agent tools (www.reddit.com via reddit) Hello r/AI_Agents, I’ve started working on an agent skill that helps coding agents like Claude Code, Codex, and OpenCode apply safer tool design principles when designing or reviewing tools. I’ve found AI models to be weak at this out of t…
Open Source Digital Twin Engine (www.reddit.comhttps) I built a personal digital twin of my property with (honestly) mostly chat gpt 5.5. When Fable came out, I expanded it incredibly at a rapid rate, and I generalized it to the whole US.
GPT 5.4 Nano High is better than Opus and Sonnet at Planning (www.reddit.com via reddit) Believe it or not, Nano via the API (not in Codex or as an agent) is an absolute beast at creating functional implementation plans, as well as analyzing or proposing solutions better than the larger models. Don't just take my word for it.
Using claude code and codex inside the cursor maybe over frustrating now after new updates (www.reddit.comhttps) Well, I use Cursor IDE as my primary code editor. Most of the time, I work on a single screen with multiple tabs — sometimes from a coffee shop, sometimes from a park.
best ai coding subscription under $20-30/month? (www.reddit.com via reddit) hi everyone. my free trial of chatgpt plus is ending soon.
And how to trust to that company? (www.reddit.comhttps) Used my chagpt for work, requests, googling, codex and etc. Nothing prohibited.
I created a keyscheduler to hit enter on my terminal on a predefined time (www.reddit.com via reddit) Probably there's an easier/better way to do this, but I vibecoded this small tool for myself to hit Enter on my VSCode app to execute a "Hi" or a "Continue" at a predefined time so I can better manage my 5hr time windows. So my use cases a…
After latest Cursor update, can't keep Claude fully right side, Codex not working (www.reddit.comhttps) Hello, I updated cursor to my latest version today, and now I can't use Claude Code extension in the Secondary Side Bar anymore, or maybe I don't figure it out why. It was just fine yesterday.
TAB / non-steering message? (www.reddit.com via reddit) Codex CLI has TAB to queue a message. Even Copilot CLI has something like that.
First message in Claude Code takes 20% of usage (www.reddit.com via reddit) I've consistently noticed that the initial message I send in a conversation in claude code instantly takes around 20%(varying by around 5%) of my 5-hour session limit. I'm talking about just sending the prompt resulting in the usage consum…
Made a skill that stops Claude Code from claiming "done" when it isn't (www.reddit.com via reddit) Agents keep telling me a task is done when it isn't, and sometimes they edit a test so it passes and then report success. So I made make-no-mistakes.
Cursor Ultra 20x ($100 1st month) or keep Codex 5x $100/month? (www.reddit.com via reddit) Just recently cancelled the Codex $100 plan after their 10x rate limit was over, jumped on Cursor $20 plan and I feel like it does a better job than Codex did at front-end work that I primarily do. Codex on the 10x felt pretty limitless bu…
AI Workflow from Idea to Shipped App: How do you accelerate without losing quality, security, or architecture? (www.reddit.com via reddit) Curious about your full AI-powered workflows for turning ideas into products/apps fast while keeping strong security, architecture, and code quality. What skills, systems, loops, or harnesses do you use?
Open-source layer that cuts ~87% of your Claude Code / API token usage - quality-neutral, measured on real billed tokens (www.reddit.com via reddit) if you use Claude Code (or build on the API), you're burning a lot of tokens on stuff the model doesn't need - whole files dumped into context, the full history resent every step, easy calls routed to the biggest model. Codex also bills by…
I made an instrument that turns typing into lofi music (www.reddit.com via reddit) Claude and I vibed a little thing called keystrokes. It’s a local JavaScript, browser-rendered “instrument” for typing.
I benchmarked Claude (Fable/Opus) vs Codex vs Gemini on my own work. Codex won — but only after a config-file change that beat a model 3x its price. (www.reddit.comhttps) I built a benchmark around my actual work: Python/SQLite tooling and brownfield fixes. Each model got an identical prompt in its own CLI (full auto, one shot): a legacy codebase with planted bugs and five staged change requests.
How to save on Fable usage with Codex and Sonnet (www.reddit.com via reddit) Hey so I wanted to share this quick. After a couple days working with Fable and Codex, I settled on this workflow: - Fable is the brain - Codex does the grunt work - but sometimes Codex fails silently, so have Fable regularly poll it - and…
Made a skill that stops Claude from answering "compare X vs Y" with an essay — works on Codex too (www.reddit.com via reddit) (Quick note: English isn't my first language, so I used AI to help write this post. Please bear with the writing.) Every time I asked Claude to "compare these libraries" or "survey this landscape," I got a well-written essay I couldn't act…
Follow-up: deterministic context folding for long Claude agent sessions (www.reddit.com via reddit) Follow-up from the earlier thread, with a shorter framing for Claude users: I open-sourced Context Warp Drive, a deterministic context-folding engine for long-running Claude agent sessions. Repo: https://github.com/dogtorjonah/context-warp…
Why so much hate on Fable? (www.reddit.com via reddit) I did a huge refactor on a huge codebase and I consumed my full usage just when it finished. Anyway, I always ask Codex to review the plans and implementations, when using Opus, there is always a few back and forth until everything in orde…
I got tired of watching Claude Code work in a plain terminal so I built it a room. (www.reddit.comhttps) It's a small 3D office. You add the agents you already use, Claude Code, Codex, Gemini, and each one gets a little robot with its own desk.
Holyshit I just bought Max and my life is changed (www.reddit.com via reddit) The ultracode is amazing + no need to worry about limits anymore no openrouter, no codex subscription, where this has been all my life
How I Use SKILL.md for Agentic Coding Development (www.reddit.com via reddit) I made a small repo to show how I use SKILL.md with AGENTS.md / CLAUDE.md for AI-driven software development. This is my own workflow for using Claude Code, Codex, Cursor, or similar AI coding agents more effectively in real projects.
Which model you run in work settings when you don’t have to worry about tokens consumption (www.reddit.com via reddit) In my work settings we have options to choose between Claude Sonnet ,Opus Codex Gemini.Work Recommends to be on auto in VSCode but I always end up using Opus.I mean why not.Anyone else do this ?
To reduce your tokens usage use this. (www.reddit.com via reddit) Agentic workflows like Claude Code/Codex can eat up a lot of tokens. Which is why I built this, https://github.com/blackcoffee2/codetree Hoping to help the community.
Every month, developers migrate between Codex and Claude. Anature documentary. (www.reddit.com via reddit) The beginning of the month has come. Once again, it is migration season.
tested it in a language I don't speak, still got back a prompt like a spec (www.reddit.comhttps) Built with a mix of Claude and Codex — figured I'd share the actual output here rather than the build story, since that's probably more interesting to this sub. It's a Windows voice app — hit a hotkey, talk normally, get a structured promp…
I was getting frustrated with how AI coding agents navigate large repos, so I started building some helper scripts (www.reddit.com via reddit) I've been spending a lot of time using Codex and Antigravity on a fairly large Laravel + React project. After a while I noticed the same patterns over and over again.
skillhub - compose package manager for AI agent skills (Claude Code, Cursor, Codex) ( via reddit) could not extract summary
Hexana MCP 0.3: give your AI coding assistant actual binary vision (native executables, WASM + native inspection tools) (www.reddit.com via reddit) https://preview.redd.it/5lwcchg4wrah1.png?width=2048&format=png&auto=webp&s=853ca6a8d9f245620d2837ac0fed559e6519c8fe AI coding assistants are good at reading source, but the moment you ship a binary, they go blind. Ask Claude Code or Codex…
Tiny guardrail that stops your coding agent from running something dumb in the terminal (www.reddit.com via reddit) Like a lot of you, I let Claude Code (and sometimes Codex/Gemini) run shell commands for me. 99% of the time it's great.
Tiny proof gate for Claude Code/Codex/Cursor changes (www.reddit.com via reddit) I built DoneCheck because AI coding agents can sound finished before actually showing evidence. It is a zero-dependency Python/GitHub Action gate: scans changed files runs your verification command fails with no evidence writes DONECHECK.m…
If AI agents can access local files, why can't they search them like they search the web? (www.reddit.com via reddit) I've been thinking about this while building AI tools that work with local files, and I'm wondering if I'm misunderstanding something. Today, AI agents can already access local files.
Reduced usage by using lower-effort agents when the main session is set to High or XHigh (www.reddit.comhttps) I default to High effort on FabIe, but unlike Codex, Claude Code can’t dispatch low- or medium-effort agents directly. Instead, agents default to your current effort level.
Your access to Fable 5 hasn't been restored yet? This worked for me. (www.reddit.com via reddit) I tried updating every app, logging out, etc. Nothing worked until I tired using codex.
Is there a reason I should not use Kimi Code when I'm out of Claude time? (www.reddit.com via reddit) I build android tools for my own self. Which means I don't have any kind of risks associated with trade secrets or publishing.
why cant codex and claude collaborate by themselves (www.reddit.com via reddit) I am aware of agents having sub-agents where a task is being delegated by an agent to other agents, but that communication is a simple request-response thingy. I haven’t seen any tool that allows sub-agents to ask questions back to the mai…
What terminal / IDE are you using if you handle several clients + AI uses (www.reddit.com via reddit) Good morning, afternoon and night wherever you are! As the title suggests, I am currently looking for the best alternative for using Cursor, terminal and Ghostty for handling: - Several different projects / clients all with different plugi…
I made a game where you coach a soccer team by giving tactics in plain English. (www.reddit.comhttps) I made a little soccer game where you coach a team by describing the tactics in plain words, and Claude/Codex/Any AI writes the bot code for you. The catch is that Claude isn't allowed to come up with the strategy, that part is yours.
Hobbyists: Do you pay for Claude Code, and is it worth it? (www.reddit.com via reddit) I used to use the GitHub Copilot Education plan before it got nerfed. Since then I've tried free alternatives like Antigravity and Codex, but their usage limits are pretty restrictive.
I stopped asking “which AI coding tool is best” and started routing tasks by failure mode (www.reddit.com via reddit) I've been changing how I use AI coding tools recently. Before, I kept trying to decide which one should be my "main" tool: Claude Code, Codex, Cursor, etc.
Restore support for no-AVX2 CPUs (www.reddit.com via reddit) Any chance for Anthropic to restore support to allow Claude Code CLI to be run on pre-Haswell Intel CPUs: https://github.com/anthropics/claude-code/issues/50684 https://github.com/anthropics/claude-code/issues/63609 Needed decent multithre…
Why can't we roll over unused Claude usage? (www.reddit.com via reddit) Is it just me, or does anyone else wish Anthropic would let us roll over our unused usage? Right now, if you don't use up your quota before the weekly reset, it’s just gone.
I kept missing when my Claude Code sessions paused for input, so I built a menu bar watcher (it can now ping my phone too) (www.reddit.comhttps) I usually have a few Claude Code sessions going at once, sometimes with Codex or Cursor in the mix. My problem is catching the moment one stops to ask me something.
I code from my phone now: Claude Code runs on a VPS, one command sets it up (www.reddit.com via reddit) For the cost of a coffee a month I run Claude Code on a small cloud box instead of my laptop. I close the laptop and it keeps working.
How do you guys handle workspace recovery with Claude Code? (www.reddit.com via reddit) It's just me, Cursor, Claude Code, Codex, and a stack of custom local loops. I’ve shipped over 580 production PRs solo this month, but my biggest operational bottleneck right now is switching to a new provider (and session limit) when I re…
Genuinely want to hear from the AI skeptics - what keeps breaking for you? (www.reddit.com via reddit) really want to understand why a lot of people complain about Claude/Codex not performing as per their expectations. What are you guys building?
1 Billion Tokens Milestone!! (www.reddit.comhttps) I am not sure how well I maximized my token utilization, but I finally hit 1 billion tokens while working on my project. Today I spent around 12 hours working on it.
I built a Claude Code plugin that makes the AI follow a real dev lifecycle — branch, commits, PR draft, best practices and all (github.com via reddit) Hey, Just released Specsmith v0.1.1 — a plugin for Claude Code (also works in Cursor, Antigravity IDE, Codex, VS Code) that enforces a full development lifecycle on your AI agent. The "think before you code" part isn't new.
Do you actually use one AI coding tool for everything, or route by task? (www.reddit.com via reddit) I keep seeing Claude Code vs Codex vs Cursor comparisons, and I used to read them like there has to be one clear winner. Honestly?
I stopped typing prompts to Claude Code. Now I point at my running app, say what I want, and watch Claude Code make it happen live. (www.reddit.comhttps) Coding without typing: you point your cursor at your running app, say what you want out loud, and watch the change happen live. My best friend and I built a Mac app for it.
I built a Claude statusbar hardware display for my desk (www.reddit.comhttps) Built with Claude! This display uses Claude's hooks and JSONL transcript tailing to display the realtime status of a Claude Code agent.
We don't write code anymore, but still use IDEs that suck for the new way of coding (www.reddit.com via reddit) Whining part I'm a software eng, and I pretty much don't write code anymore. Agents do.
ccgram v4.3.0 - control Claude Code and shell from Telegram via tmux/herdr (www.reddit.com via reddit) I shipped ccgram v4.3.0. ccgram lets you control Claude Code sessions from Telegram while they keep running in your real terminal (tmux or herdr).
For those of you who complain about paying $20 and usage limits on the pro plan (www.reddit.com via reddit) https://preview.redd.it/61kj2nxdbx9h1.png?width=432&format=png&auto=webp&s=b87d4a672f639066a8629f9de6811c262e434451 To me, it feels absolutely worth it. This app proves it.
Do coding agents actually need bigger context windows, or better task handoffs? (www.reddit.com via reddit) I've been using Codex / Claude Code on a few real coding tasks, and I'm starting to doubt a habit of mine: every time the agent makes a mistake, I give it "more context." Sometimes it helps. But sometimes it feels like I'm just giving it m…
Nimbalyst: Visual Editor for Claude Code & Codex (Open Source) (nimbalyst.com via reddit) This open-source visual workspace wraps coding agents like Claude Code and Codex in a desktop environment instead of a terminal. You manage agents, sessions, tasks, and files while editing markdown, CSVs, mockups, code, Mermaid diagrams, E…
Token Maxxers I have a gift for you (www.reddit.comhttps) I launched this open source project a couple of weeks ago, got over 500 stars and about a million impressions on reddit so I kept adding features and fixes and maintaining it we are at v0.3.2 with realtime voice agent that you can use to c…
Claude Code for (MacOS) app development issues (www.reddit.com via reddit) "I learned that 4.7 needed heavy structure, rules, and guidelines to perform well." That has been my experience when using Claude Code for app development. I have OpenClaw Codex, and Ollama models for coding as well, and Claude Code is sti…
You think the current Harnesses we have like Claude code, Cursor or Codex etc are good enough ? (www.reddit.com via reddit) Like for example they have a lot of tools, skills, web search, they are very complete, but i still built a lot of new tools myself for my own use cases, i understand that as a product for everyone they cant have super specific tools sets,…
Want to redo my wordpress website in claude!!!help me (www.reddit.com via reddit) Im not a coder or something,im into health care and i want to redo my website using claude or codex,can someone help me ,what files do i feed as input ,how can the work flow
I used Claude Code to build the tool I wanted while debugging with Claude Code (www.reddit.comhttps) I’ve been using Claude Code a lot while building my desktop app, and the biggest problem was not always the code. It was the explanation.
We built BlitzOS: run Claude Code from your Mac's notch and let it act across your apps (free, open source) (www.reddit.comhttps) Hey r/ClaudeAI, we built BlitzOS, a free and open source Mac app that gives Claude Code a home in your notch and hands to act across your real apps. What it does: - Drop any app window into BlitzOS and your Claude Code agent can work in it.
Thinking of switching from Codex $20 to Cursor how are the limits and code quality? (www.reddit.com via reddit) I’m currently on the $20 Codex plan, but the rate limits are getting really annoying. I can’t go above $20/month since that’s already expensive for me.
Do you use Codex as a reviewer after Claude Code writes the code? (www.reddit.com via reddit) For almost a year now, I’ve been using this workflow: let Claude Code write the first pass, then ask Codex to review the diff like a second reviewer. Not because I think Codex is “better” than Claude Code.
How do you manage long-term AI-assisted coding without losing control? (www.reddit.com via reddit) I've been using Claude Code heavily for a personal project, but after long audit/refactor/hardening loops, I often end up with regressions, architectural drift, and even bugs I had already fixed coming back. How do you use AI coding agents…
[AINews] OpenAI reports median internal Codex output tokens grew 56x in Research, 32x in Customer Support, 27x in Engineering, and 13x in Legal since November 2025. (www.latent.space) [AINews] OpenAI reports median internal Codex output tokens grew 56x in Research, 32x in Customer Support, 27x in Engineering, and 13x in Legal since November 2025. It's happening.
GLM 5.2 is unbelievably dumb (www.reddit.com via reddit) Yeah... you heard it right.
Claude Max vs Codex Pro or both combined? (www.reddit.com via reddit) I’m considering one heavier subscription (~€100/month) and want to know which provides better value for agentic coding. I tested GPT Pro and was satisfied with Codex.
Codex says I already used 94% of my limit on first launch. I haven’t coded anything yet. Bug or hidden usage? (www.reddit.comhttps) I just opened Codex for the first time and noticed my usage was already almost gone: CORRECTION** - 5-hour limit: 6% used - Weekly limit: 8% used - I had just launched the app - I had not run any coding task before this - No previous Codex…
I run Claude, Codex, and ChatGPT in a single pipeline. Here is how I handle the handoff. (www.reddit.com via reddit) I have been running Claude Code for architecture and planning, Codex for autonomous feature builds, and ChatGPT for quick web research and prompt iteration. The hardest part was not getting each tool to do its job.
A skill that packages your skills for public release — without telling you to publish everything (www.reddit.com via reddit) Agent skills are quietly becoming an open-source artifact — people share Claude/Codex skills now the way they used to share dotfiles or prompts. The annoying part is the gap between "this works for me internally" and "this is safe to put o…
Is it always this funny? (www.reddit.comhttps) Been using Codex for bulk stuff before passing off to Claude and they randomly started being funny as hell lol. “I’m upgrading the validator too so that particular gremlin doesn’t get a second career” 🤣
I open-sourced ByteDance's "Vibe Creating" prompt skill as a portable Agent Skill (single SKILL.md, bilingual) (www.reddit.com via reddit) ByteDance shipped a creator paradigm + prompt skill called "Vibe Creating" with their Seedance 2.0 video model. I open-sourced a portable version on the open Agent Skills standard (single SKILL.md) — it drops into Claude Code's ~/.claude/s…
I stopped letting Claude Code review its own work (www.reddit.com via reddit) I’ve been testing a simple workflow: Claude Code writes or edits the code. Then, if the task is risky or messy, I hand it to Codex with one job: find bugs, bad assumptions, edge cases, and anything I should not ship.
Anyone here using more than one AI tool in their workflow? How do you handle the context gap? (www.reddit.com via reddit) I've been running Claude for planning and a separate session for building, and the part that keeps breaking down is the handoff. whatever I figured out in one session doesn't automatically carry to the next.
For AI coding agents, review feels more expensive than generation now (www.reddit.com via reddit) One time I ran Claude in a loop for four or five hours to build a program, then had Codex and Cursor review the code. Each pass surfaced different issues, but the codebase was so large that I could only trust what they flagged.
ChatGPT macOS Project chats no longer accessing local repo (www.reddit.com via reddit) I’m trying to restore a workflow that used to work reliably in the ChatGPT macOS app: I’d open a new chat inside a Project, paste a handoff prompt, and the assistant could run terminal commands directly against a local repo. Now new Projec…
Spent a weekend getting Claude to replace QuickBooks and Quicken. The plan did 90% of the work and saved me $500/yr (www.reddit.comhttps) I run a one-person LLC and file a Schedule C, and I have spent years paying for QuickBooks without really knowing what half of it does. I had to pay someone to set it up in the first place.
Software development has entered its "infinite monkeys" era (www.reddit.com via reddit) With the rise of agentic coding tools like Claude Code, Cursor, and Codex, the barrier to entry is gone. Now, anyone with an internet connection can "type." We have essentially reached the infinite monkey phase of software development.
Claude getting frustrated with Codex (www.reddit.comhttps) Trying to use up my weekly Claude usage and just going through code review loops using Codex. You can see Opus 4.8 is clearly getting frustrated with the pedantic comments it's getting.
I'm building agent loops that auto-edit my videos, but the hard part has been finding a model to accurately grade the result (youtube.com via reddit) Quick context: I've been building agentic loops that edit my short-form videos for me. The editing works really well, but I found myself needing to check the process at several gates.
If you run coding agents unattended or in parallel, how do you verify the run actually worked? (www.reddit.com via reddit) I run a lot of agent loops (Claude Code / Codex / aider), sometimes overnight or several at once. My recurring headache: when I come back, I can't quickly tell whether a run actually did the right thing, quietly broke/regressed something,…
I built a free pixel-art RTS that turns your Claude Code sessions into a calm little kingdom (www.reddit.comhttps) I've been building Age of Agents — a small, free local app that turns your AI coding sessions into a peaceful, Age-of-Empires-style pixel realm you can glance at on a second monitor. No combat, just a quiet kingdom of your work.
I built PromptQueue for when Claude says I'm out of prompts (www.reddit.com via reddit) This came from a very small Claude annoyance: Claude says to try again later, but I already know the exact prompt I want to run next. PromptQueue lets me queue it locally instead of keeping a tab open or setting a reminder.
I built an open-source MCP server inspired by OpenRouter Fusion: a “council” of AI models for better reasoning (www.reddit.com via reddit) I’ve been experimenting with multi-model deliberation, inspired by OpenRouter’s Fusion Router idea: instead of asking one model for the final answer, send the same problem to several models, compare their reasoning, identify agreement/disa…
Stop asking “Claude Code vs Codex vs Cursor” — the real question is what kind of work you’re doing (www.reddit.com via reddit) I think the “Claude Code vs Codex vs Cursor” debate is framed the wrong way. After reading a lot of recent discussions and trying different workflows, I don’t think there is a single winner anymore.
Evaluating LLMs for Real-World Web Vulnerability Detection (arxiv.org) Large Language Models (LLMs) have emerged as a promising tool for automated vulnerability detection, yet their effectiveness on web-specific vulnerabilities remains to be explored. This work benchmarks six frontier (Claude Opus 4.6, Codex…
The $20 → $100 gap is pushing solo power users to split spend with OpenAI (www.reddit.com via reddit) I'm a solo freelancer who uses Claude all day — agent orchestration, coding (Claude Code), analysis, writing. Not a hobby user.
Your (Big 5) Personality Based on your CLI chats with Claude / Codex (www.reddit.com via reddit) https://preview.redd.it/eb0sxr75pw8h1.png?width=1347&format=png&auto=webp&s=28f8ed289e1b904550554bfac907e22dd8840b8a Prompt: "you have access to all of the claude cli and codex cli logs? i want to build a system that can analyze my discuss…
Connected a Robinhood Account to Claude Code and Codex for Autonomys Agentic Trading... Update 1 (www.reddit.com via reddit) Update to my original post: https://www.reddit.com/r/ClaudeAI/comments/1u8nagi/connected_a_robinhood_account_to_claude_code_and/ I'm building a fully autonomous daily stock-trading desk in a Robinhood "Agentic" account. Opus is the CEO/PM,…
Added Atlas mode to Clauge — drag, resize and snap your live tabs onto a spatial canvas (www.reddit.com via reddit) The Problem I was tired of using multiple apps that were eating my system resources and killing my productivity. That's when I decided to build my own app.
what is agentic coding and why is everyone suddenly talking about it (www.reddit.com via reddit) I kept seeing "agentic coding" everywhere for the last few months, blog posts, Twitter threads, product launches, and I honestly couldn't tell if it was a real shift or just the latest marketing rebrand for AI autocomplete. So I spent the…
How can I route my own Codex models into Cursor IDE? (www.reddit.com via reddit) I'm trying to use my Codex models directly inside Cursor. I know I could just use the CLI or install an extension, but I really want everything baked into Cursor's native chat.
Suggest model and subscription (www.reddit.com via reddit) I am new to Claude. I have been using github copilot.
How to make the agent give good UI and should i let it write whole codebase. (www.reddit.com via reddit) during implementation, i am mostly letting the codex / cursor / devin to do the work after i writes the things in the agents . md so am i doing write or wrong?
Codex-maxxing for long-running work (openai.com) Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Business Developers Company Foundation(opens in a new window) Codex-maxxing for…
Samsung Electronics brings ChatGPT and Codex to employees (openai.com) ChatGPT Enterprise and Codex available to all Samsung Electronics employees in Korea and all Device eXperience (DX) employees worldwide Samsung’s global deployment is one of OpenAI’s largest enterprise launches ever Samsung Electronics to…
Two months into Claude Code, I hit 161M tokens in a single day. Here's the honest story of how a year-long Cursor user got here. (www.reddit.com via reddit) I want to share a small milestone, and the honest road that led to it. Today was one of those days where I sat down to build and just did not stop.
I got Tired of wasting Tokens on Calude Code for everything so now i just use it to Dedicate Tasks to Cheaper LLM's and run it overnight. (www.reddit.comhttps) Introducing Machinaos: AI That can Build itself depending on the Task and also a Multi Agent Orchestration Platform to run Loop Agents and Control Agents like Claude code, Codex, etc. Bring your own API keys or Claude Code sub (or run mode…
Claude app on macOS does not make a sound when it finishes a task, as far as I can tell. (www.reddit.com via reddit) Codex app makes an easily heard ping sound, like a notification, when it finishes something.I think Claude is set to also produce notifications, but it doesn't seem to be working.Anyone else have this issue and know how to fix it?
Fable-built agent harness "office", Deno CLI w/ "gamedev tycoon" web interface (www.reddit.com via reddit) Yet another take on the agent harness. Project mantra: "To understand and realise the director's vision." AI summary below: It’s a local CLI built with Deno.
Tool-Agnostic, Portable Context OS (www.reddit.com via reddit) *wasn't sure what flair. Would have chosen maybe "context engineering" had I seen it I keep seeing posts here that are all some version of 'I need a better way to manage context' and thought maybe sharing my set up could help.
Claude Code is the only agent shipping tool search on by default - and that one detail is why I don't buy "MCP is dead" (www.reddit.com via reddit) Small thing most of the "MCP is dead" discourse misses, and you can check it in your own Claude Code setup: tool search is on by default here, and it cuts tool-definition tokens by roughly 85%. Codex and Copilot still have it as open featu…
Agent Switchboard: make your AI coding agents talk to each other (www.reddit.com via reddit) I built Agent Switchboard because using multiple AI coding agents gets messy fast. You ask Codex one thing, Claude another, Gemini another, then manually copy plans, errors, files, and context between all of them.
Stop re-explaining context to every agent (github.com via reddit) Every new session with a coding agent starts from zero. I re-explain the architecture, repaste the same handoff, and watch the agent rediscover a gotcha we already hit two weeks ago.
How good is the Ultra plan compared to Codex 20x or Claude Max? (www.reddit.com via reddit) I'm thinking about getting cursor because i really enjoyed the usability of it and using the plataform but i'm worried that i will not get a generous usage of SOTA models even on Ultra plan, is it worth it to buy if i plan to use models li…
Artificial Analysis added a new tag for not currently available models for Fable (www.reddit.comhttps) I just noticed codex and GPT 5.5 are near Fable level in this benchmark tho from my experience GPT 5.5 is so good to follow instructions but not as creative as Fable. Fable was just so convenient to work with like it reads my mind and even…
I finally looked at what my Claude Code MCP setup actually contains. It flagged config drift I didn't know was there. (www.reddit.com via reddit) My context bar in Claude Code had been creeping up before I'd typed anything, and I'd been ignoring it for weeks.. This week I opened up my MCP setup and actually looked.
BaseMind: MIT licensed full context layer (www.reddit.com via reddit) Hi Peeps, I'm an open-source maintainer (Goldziher on Github) and the CTO of kreuzberg.dev. I published basemind — a pure-Rust MCP server and Claude / Codex / Gemini etc.
How I prompt AI models in 2026 vs a year ago (3 things that changed) (www.reddit.com via reddit) I do a lot of work with Claude and Codex, and the way I set things up now looks pretty different than it did even six months ago. Figured I'd share in case it's useful, take what you want from it.
I had Claude Code build a macOS app to manage everything Claude Code (and Codex) installs (www.reddit.com via reddit) There's no easy way to see what your coding agents have actually installed — skills, subagents, commands, plugins, MCP servers, hooks — or which sessions are still alive vs. safe to delete.
We added a training loop to Hivemind, so our Claude Code skills improve instead of just piling up (www.reddit.com via reddit) Disclosure: I work on Hivemind, an open-source skills layer for coding agents. Sharing an update because the problem behind it is one a lot of people here run into.
PSA: You don’t get email reminders / receipts (www.reddit.com via reddit) I just wanted to let this post live on this sub in case anyone else has the same question. I’ve been paying for this for 8 months without noticing.
I made Claude and GPT-5.5 answer the same prompt, then had a third Claude fuse the two, on the subscriptions I already pay for (no API key). Blind-tested it. Here is where it won and where it lost. (www.reddit.com via reddit) Quick share of a weekend experiment that turned into a tool. The idea: instead of picking one model, run Claude and GPT-5.5 on the same prompt in parallel, then have a fresh Claude (blind to which answer is which) merge them into one.
I built a local-first routing layer for Claude Code that decides which agent handles your request, what it can access, and logs every handoff as an auditable receipt (www.reddit.com via reddit) I built Agentlas Network (Hephaestus Network), a local-first routing layer for AI runtimes like Claude Code, Codex, Gemini, Cursor, and terminal workflows. I built it with Claude Code, and it runs on top of Claude Code as the reasoning r…
Fable 5 being gone made me realize how hard it is to go back (www.reddit.com via reddit) I know this probably sounds dramatic, but Fable 5 disappearing has genuinely killed my motivation for the last few days. Before Fable 5, I was already using both Claude and ChatGPT pretty heavily.
Devs were using my tool and saving $150k in 3 months, yet i was not using my own tool! (www.reddit.com via reddit) I was working on idea of persistent memory from an year and built my own coding bot using Free LLM tools. But, in march I purchased $200 claude max and i started building memory layer for every coding tool out there, started the name with…
I open-sourced a local memory tool so AI agents can share context (www.reddit.com via reddit) Hey everyone, I built Hearth, a free/open-source tool for people using multiple AI agents like Claude, Codex, Cursor, etc. Problem it solves: every agent has its own memory silo, so we keep re-explaining repo context, decisions, preference…
I made an open-source skill pack that stops AI assistants from validating bad product ideas by default (www.reddit.com via reddit) Most AI assistants are too eager to help you execute. If you say “I want to build a note-taking app that makes money, what stack should I use?”, the default response is usually a tech stack, feature list, or MVP plan.
All your Claude Code agents on one screen — and they survive a reboot (Windows, no WSL) (www.reddit.com via reddit) I run several Claude Code / Codex / Gemini agents at once on Windows. The hard part was never "open more panes" — it was attention, coordination, and not losing everything to a reboot.
One AI agent agrees with you. Two can agree with each other. A third one can catch what the other two miss. (www.reddit.com via reddit) One AI agent is useful. It is also over-agreeable.
How do you manage skills, update them, across multiple agents? (www.reddit.com via reddit) I tried `gh skill` and `npx skill` and I couldnt find a good solution to keep up to date skills, some global, some local, in various agents and both my desktop and laptop. I spent the weekend to try to solve it for me, and maybe if you hav…
every junior vibe coder rn (www.reddit.comhttps) as a codex user, that was really good 3 days, ngl
How much time do you actually spend maintaining AI context files? (www.reddit.com via reddit) Serious question for people using Claude Code, Cursor, Codex, etc. on medium or large repositories.
how many files exist in your repo purely to help AI remember things? (www.reddit.com via reddit) serious question if you use Claude Code, Cursor, Codex, etc… how many files are in your repo mainly for the AI rather than for humans? things like: - AGENTS.md - CLAUDE.md - SESSION.md - CHECKPOINT.md - context.md - architecture docs - AI…
My CLI security scanner (compatible with Claude Code) found 407 vulnerabilities in production code (www.reddit.comhttps) Hi all, I built a CLI security scanner called Heimdall that uses AI coding assistants (Claude Code, Codex, Gemini CLI, and OpenCode) to scan source code and generates structured reports (JSON, Markdown, and SARIF) detailing each vulnerabil…
agentsweep: a CLI that finds & redacts the secrets your AI coding agent saved to disk in plaintext (www.reddit.com via reddit) Every time you paste an API key, DB URL, .env file, or (worst case) a crypto wallet seed phrase into Claude Code, it gets written to a local history file in plaintext. And it doesn't just sit there — these agents re-read their own history…
Somebody asked what you’ve built with AI that is making you money (www.reddit.com via reddit) I just want to share my journey with Claude (and Codex) and give y’all hope. This January 1s I wrote couple resolutions and among one of them was to exponentially learn vibe coding and focus on something that interests me and I’m incredibl…
I had Fable 5 review my vibe-coded app before it got shut down/cancelled — it found 10 pages of issues. Should I use Opus next? (www.reddit.com via reddit) Hey Folks! I’vve been working on a vibe-coded internal software tool that we’re currently using across 4 of our retail stores.
Quick one for the builders here. (www.reddit.com via reddit) If you code with Claude Code, you know the wall: you’re deep in a build and the usage limit cuts you off mid-thought. Most people just wait it out.
Spent 12 years as a PM watching the wrong things get built. turned that pattern into a free Claude skill (MIT) (www.reddit.com via reddit) I've been a PM for about 12 years, mostly 0-to-1, and I've spent a lot of that time watching smart people ship products nobody actually wanted. Not because they're bad builders.
How do you manage agent skills across projects? I keep losing them. (www.reddit.com via reddit) I've used Claude Code, Cursor, and Codex, OpenCode across ~8 different projects. Each has its own .claude/skills/, .agents/skills/, etc.
I wonder if you will use Cursor, Windsurf and Codex at the same time? (www.reddit.com via reddit) I'm a user of Cursor Windsurf Codex of Claude Code. I use them to handle the same project or different projects simultaneously, including collaborating with teams.
Made a tool to auto-generate .cursorrules from your actual stack (www.reddit.comhttps) Cursor is great until it improvises on the parts I never told it about — the data layer, where files belong, naming that doesn't match the repo. .cursorrules fixes that, but I never kept mine current, so Cursor kept making those structural…
/architect: Cut Fable token cost. Fable is orchestrator/reviewer, Codex is builder (www.reddit.com via reddit) https://github.com/DanMcInerney/architect-loop Fable absolutely rules, but the load-bearing work of coding agents is in the design and the review, not the actual coding. So this is two skills: /architect uses Fable as the orchestrator and…
PORTUS - Mission control for multiple AI agent CLI sessions. Run every agent in one focused, three-pane window and watch them all at a glance. (www.reddit.com via reddit) Hey guys, Claude built me this awesome terminal, which I now use every day to spin and monitor multiple chat sessions, shells, etc., all within a single window. Yes, but why?
The whole world is created by players using claude (www.reddit.comhttps) I'm making an online game like GTA Online, but everything is generated by the players. You prompt your own car, your own building, your own weapon.
Fable 5 added to the Artificial Analysis Coding Agent Index... barely 1 point ahead of GPT-5.5 ??? (www.reddit.com via reddit) https://preview.redd.it/z0vkpnmp9s6h1.png?width=4640&format=png&auto=webp&s=7bb14d4d04d6cd15caf5aacc1d3c49512b7e7fd8 Artificial Analysis just added Claude Fable 5 to its Coding Agent Index (a composite average of pass@1 on DeepSWE, Termina…
Made a skill for agents to cross-search Claude/Codex/Cursor sessions ( via reddit) could not extract summary
Claude can now publish its HTML to a live link, no signup (www.reddit.com via reddit) Hey everyone. Quick disclosure: I work on display.dev, so I'm biased – but this fixes something that always bugged me.
What do people actually use ClaudeCode/Codex for (www.reddit.com via reddit) I’m just getting into AI and having a play around how to use it in my life. I have been using Cowork to experiment.
Everyone's talking about AI agents. Almost nobody's talking about agent sprawl. (www.reddit.com via reddit) Everyone's talking about AI agents. Very few are talking about agent sprawl.
Composer 2.5 vs Chatgpt 5.4mini High (www.reddit.com via reddit) I’ve been testing the new Cursor Composer 2.5 models, and they seem capable for the price. Since Codex 5.3 was removed, I’ve been looking for a coding model that can fill that gap without costing too much.
Had Claude find bugs in some code and automatically send them to another AI to fix, and it worked better than expected (www.reddit.com via reddit) We've been building a tool that gives AI sessions inboxes so they can message each other (called Khala). Been testing different workflows with it.
A macOS app like AgentLimits but supports multiple accounts? (www.reddit.com via reddit) Before I start building this myself does a macOS app exist that allows me to see codex and Claude code usage and limits like the AgentLimits project does but for multiple accounts? I have multiple accounts for each and when one runs out, s…
Is nobody experiencing visual jank with Cowork? (www.reddit.com via reddit) I have been mostly using Claude Chat all my time and it always worked nice. But I have decided to finally move to Cowork for my multiple file use and why the hell is it so jumpy, jittery and unpolished?...
I evaluated 7 production agent runtimes against 7 criteria: strengths, trade-offs, and who each one is actually for (www.reddit.com via reddit) I collected the agent runtimes I could find and compared them on the things that mattered for our production deployment: resource efficiency, ability to work with vendor agents (Claude Code, Codex, etc.), and the security layer. Mixed clos…
New web UI update (www.reddit.comhttps) Used to be Instant and Thinking with effort modes Standard and Extended for the selected model. Now it’s the selected model with Intelligence variable similar to codex at Low, Medium and High.
I think I found my best workflow for coding with Opus, Codex and Fable. (www.reddit.com via reddit) I start in Fable. For me it works best as the “brain” for the project.
Codex desktop app is very much capable of doing anything ChatGPT can, it’s not just for coding. It also gives you access to the higher thinking tier with decent usage limits. Your welcome! (www.reddit.com via reddit) I don’t know why but it seems like no one uses codex for anything but code. It literally has a setting to switch to “everyday tasks”.
Cursor is still a beast with unlimited at $20usd (www.reddit.com via reddit) Hey guys. Just passing by to say that Auto mode for $20 a month has been a game changer for me.
How an astrophysicist uses Codex to help simulate black holes (openai.com) The gravity around a black hole is so extreme that nothing, not even light, can escape once it gets close enough. Astrophysicists like Chi-kwan Chan study black holes with computer simulations and observations.
Could Fable 5 one shot entreprise RAG system? (www.reddit.com via reddit) Hey ! For the ones who tried Claude Fable 5, do you think it could make a big entreprise RAG System?
The pain of Codex Desktop App and Windows (www.reddit.com via reddit) I started off using the Codex desktop app in WSL mode. There were some issues, but it kind of worked.
MCP Fail (www.reddit.com via reddit) So like, I was just trying to show a friend Codex because he was just getting into AI and to my surprise, I couldn't actually get an MCP server installed in Codex because I installed it and then I used it in one thread and it said everythi…
Access OpenAI models and Codex through your Oracle cloud commitment (openai.com) Access OpenAI models and Codex through your Oracle cloud commitment | OpenAI Use your existing Oracle cloud commitment to give teams access to OpenAI’s most advanced models and Codex, without creating a new purchasing path. Listen to artic…
Tested Fable 5 on 4 private benchmarks. The one it failed, Sonnet 4.6 partially caught (www.reddit.com via reddit) I keep a few private benchmarks for coding agents, built from real bugs in past projects. Hidden Playwright tests grade the result inside Docker after the agent finishes, so the model never sees them.
I've reverse engineered Anthropic's dirt-in-the-eye approach to sabotaging LLM development to help guide me towards their secret sauce (www.reddit.com via reddit) Anthropic needs to be pretty careful about what exactly they sabotage, otherwise it would be happening all over the place. So I had codex create a pretraining playground boobytrapped with lean prover everywhere and you can literally see wh…
I Tested Claude Fable and GPT-5.5 xHigh on a Real Packing Algorithm, Claude Won Efficiency, GPT Won Speed (www.reddit.com via reddit) I ran a head-to-head test between Claude Fable and GPT-5.5 xHigh on a real-world optimization problem I wrote myself. This isn't a coding challenge or LeetCode problem.
What happens when LLM providers stop subsidising? (www.reddit.com via reddit) i;m running Hermes agent and connected my OpenAI Codex for 2 months $20/month flat rate everything worked fine. last sunday i hit my weekly limit, due to i was using 5.5 with fast.
Multiple LLM Algotrading Bot Development (www.reddit.com via reddit) Hey everyone, I'm a Claude Max (5x) user, and ever since Fable 5 launched I've been running into token limits much more often than before. So I thought maybe I can combine it with Cursor.
GPT 5.5 vs Fable/Mythos 5 Tamagotchi Showdown (www.reddit.com via reddit) Well, how do I start this, I think we first need some important context. Chai: https://preview.redd.it/egngyea5cf6h1.png?width=1080&format=png&auto=webp&s=9ade63fbc584b7fab28dba4914bc3fcb877f557f Hasbullah / Hasbi: https://preview.redd.it/…
Prediction for the next few days of r/ClaudeAI posts (www.reddit.com via reddit) ChatGPT is trash now and I have re-subscribed to Claude Code. It saved my life.
Fable made this launch video for my app (www.reddit.comhttps) Hi all, I’m building Daydream, a video editor that uses Claude Code/Codex. Just like everyone else, I’ve been eagerly anticipating testing out Fable with different use cases to see what it can do.
Returning after a month, how are the limits going recently? (www.reddit.com via reddit) I subscribed to Claude Code shortly after the pentagon thing, lowest paid tier, no API usage other than what they gifted people in that time. I loved it at first, used it for a couple months but the usage limits were getting very bad towar…
How much usage does Cursor $60 get you? (www.reddit.com via reddit) I currently pay for the $20 plans for both Codex and Claude Code. Even with those 2 combined, I feel like the weekly limits are too low for what I do.
We built a free CLI to keep CLAUDE.md, slash commands, MCP servers, and skills in sync across machines (www.reddit.com via reddit) How I got Claude Code and Codex to pursue goals over weeks (www.reddit.com via reddit) Claude Code and Codex are great at a single task, then they stop. You give one a big task, it runs for a while, finishes what it can in that session and then you're back to deciding the next step yourself.
What Codex Plugins are actually improving your workflow (www.reddit.com via reddit) What Codex Plugins are actually improving your workflow
I built a small local QA tool because I do not trust my agent branches enough yet (www.reddit.comhttps) ( Ignore the big "Bonjour" in the middle , That is just the random app I had open for QA ) The part I wanted to show is the left sidebar: a bunch of agent branches, each one with its own local environment. I built this because my Claude/Co…
The model is the CPU, not the computer — why the harness moves agent performance as much as a model upgrade (www.reddit.com via reddit) Wrote up something that kept nagging me: people keep saying "we used the same model" and getting wildly different agent results. The reason is that the model isn't the system — the harness is.
How engineers at Nextdoor use Codex to build without limits (openai.com) Question's regarding AI models (www.reddit.com via reddit) Hi, I’m wondering about the $60/month plan. Are Claude Opus, Codex, and other models included?
We went from $20 → $100 → $200/month on AI coding tools. Now we're splitting the budget between Claude Code and Codex. Is this just where we all end up? (www.reddit.com via reddit) Started with Cursor. Most people do.
Is anyone actually running coding agents autonomously from issue to PR? (www.reddit.com via reddit) I’m trying to understand how common the fully autonomous workflow actually is. Not using Claude or Codex interactively while you steer it, but assigning an issue, letting the agent plan and implement it unattended, then receiving a finishe…
I built a free skill that takes you from "I have an app idea" to a real plan and solid MVP (www.reddit.com via reddit) I'm a product manager (12 years, mostly taking things from zero to one) and I wanted to help everyone who is trying to build an app now that coding is available for everyone. I created a skill for AI coding assistance called Vibe-check.
Levi: Run AlphaEvolve on your local QWEN 30B (www.reddit.com via reddit) Hi r/LocalLLaMA, Wanted to share something I'm excited about. I've been fascinated by AlphaEvolve and its results for more than a year now, but running the open source frameworks gets expensive fast.
Bioinformatics and biology (www.reddit.com via reddit) For the sake of discussion, I am a bioinformatician and I have a project I’d like to di What’s better codex,cursor, Claude code? Or a combo of each Please provide detailed explanation of experience with use as well
Google Colab CLI opens runtimes to Claude Code and Codex (www.reddit.com via reddit) https://www.helpnetsecurity.com/2026/06/08/google-colab-command-line-interface-cli/
We pick coding agents by vibes, and it shows (www.reddit.com via reddit) We run 2 coding agents now, Claude Code and Codex, and when someone asks which one to use for a task, the honest answer is always "whichever felt better last time." No record, no metrics, no logs, just (good?) vibes. The thing is, the same…
Cursor Pro vs Claude Code vs Codex: Which gives the most usage for $20/month? (www.reddit.com via reddit) I know Cursor lets me use different models instead of being locked into one. What I'm trying to figure out is usage.
Midas: 100% local agent memory — no LLM at ingest, $0, nothing leaves the box (MCP + Python SDK) (www.reddit.com via reddit) Most "agent memory" tools call an LLM to extract facts on every turn — $ per token, latency, and your whole conversation goes to a provider. Midas doesn't: local embeddings + ranking only.
I made it so any agent can agent can use any context form another agent. Claude learns from codex, visa versa. (www.reddit.com via reddit) Short clip: two fresh sessions, different tools, no shared context. I ask one to pull up the other's last session and it just does it, then I flip it the other way.
I used codex to help design a PCB and do component selection (www.reddit.com via reddit) I work in hardware and come into contact with high voltage often. I feel like the biggest winner in this whole AI thing.
I made it so any agent can agent can continue any other agent's session. Claude pops into codex, visa versa. (www.reddit.com via reddit) Short clip: two fresh sessions, different tools, no shared context. I ask one to pull up the other's last session and it just does it, then I flip it the other way.
composer 2.5 is gone from the model list? (www.reddit.com via reddit) Not the composer 2.5 Fast, just composer 2.5 (I think it's the slower version, but less token usage). Few days ago I can still see the 2 options but now there's only Composer 2.5 Fast to choose.
How do you manage your sessions and keeping up with all the juggling you now do? (www.reddit.com via reddit) TL;DR: how do I manage my Claude CLI sessions and work better? Like many others I have taken the new opportunities with Claude, Codex and GitHub Copilot before that, to do a lot more.
When do you think the ChatGPT super app is releasing? (www.reddit.com via reddit) The rumored ChatGPT super app could be huge. From what’s being reported, OpenAI may be moving toward one unified app that brings together ChatGPT, Codex, and the Atlas browser.
Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange (huggingface.co) Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange Does anybody know how we can activate our vouchers came from hackaton especially codex (still mystery) and modal (solved just now) ? Because the codex key has nowh…
I got tired of explaining my project to every AI coding tool every single session. Building the open source fix. (www.reddit.com via reddit) Every day the same thing. Tell Claude Code we only use FastAPI.
AI helped our test suites hit 95% coverage and bugs still slipped through. So PRs now climb an autonomous verification ladder before a human reviews. (www.reddit.com via reddit) Intro + Context [TLDR at the bottom for my skim readers 😄] We run Claude Code and Codex with a full agentic pipeline across our entire SDLC. Our workflow, by default, incorporates cross-model auditing, where Claude and Codex usually have t…
Meet Monako Glass: Chinese startup brings Claude Code and Codex to smart glasses | Mint (www.livemint.com via reddit) A Chinese startup has unveiled smart glasses which it claims to be the 'world's first wearable Linux computer in glasses form'. The glasses, called Monako Glass, are aimed at developers, researchers, and AI power users, allowing them to ru…
Blessed without a 5h window. (www.reddit.com via reddit) I recently got the ultra plan, and have been using Composer 2.5 @ fast all day. I've been steering agents for 8+ hours w/ no brakes & my quotas haven't been reaching any limits at all, so i have now have lots more tokens in savings.
Working multiple projects at n my home comp from my phone (www.reddit.com via reddit) Looking for some advice. I generally have 4-8 vscode projects open at a time at home where I'm just jamming with Claude (and sometimes codex) as I get the urge.
Are local models good enough to replace Claude/Codex solely for simple HTML tasks? (www.reddit.com via reddit) I know local models can’t compete fully yet, but I’m curious about where the limits are. My use case is generating simple HTML activities for elearning creation purposes.
OpenClaw + Hermes users: where does your agent army actually live? (www.reddit.com via reddit) I’m working on ClawBud, a managed Agentic OS for running OpenClaw, Hermes, Claude Code, Codex and other agents on one private cloud computer, so I’m obviously biased. But this is the problem I keep seeing everywhere: The agent itself is no…
Open-sourced a CLI that lets Claude validate UI changes in a real browser with screen recordings, HARs, logs, and Playwright traces (www.reddit.com via reddit) Hey folks, I just shipped Canary. It reads the code diff, identifies which UI flows are affected, builds a QA plan, and drives a real Chromium instance through those flows using Claude Code.
Built with Claude Code: pidgin.sh — let Claude share artifacts as URLs (pidgin.sh via reddit) I kept running into the same friction with Claude Code: it'd generate something nice (an HTML mockup, a report, a plot, a one-pager) and then I'd have to manually save it, find somewhere to host it, and send a link. So I built pidgin.sh, a…
Sites in Codex is genuinely useful, but I think we'll look back on it as a transition phase (www.reddit.com via reddit) Been using Claude Artifacts the same way for months, so when Sites in Codex dropped it clicked for me right away. It's basically Artifacts and Dashboards combined with auth on top.
Has anyone actually replaced Claude Code / Codex with local models on an Macbook Pro M5 Max 128GB? (www.reddit.com via reddit) Considering buying a maxed out MacBook Pro M5 Max with 128GB of RAM and one of the things I want to figure out before pulling the trigger is whether local models are good enough to actually replace cloud AI coding tools. My current setup i…
Kinda love OpenAI as a company, aside from all the people problems (www.reddit.comhttps) I honestly think if it was any other major company in the shoes of OpenAI (Google, Tesla, Microsoft…) money wouldn’t be able to buy AI as we have it today. Congrats to OpenAI for the democratization of AI in general, and recently for forci…
Codex developers are deploying slop. Someone needs to be terminated for their poor performance. (www.reddit.com via reddit) could not extract summary
Switched from Codex to Claude Max 20x — what are the must-have settings/workflow tips for Claude Code? (first time on the desktop app, Windows) (www.reddit.com via reddit) I was using Codex and just switched to the Claude Max 20x plan. This is my first time using the Claude Code desktop app on Windows, so I'm starting more or less from scratch and would love some advice from people who've been using it for a…
I’m upset… (www.reddit.com via reddit) So long story short - openai 20$ subscription is much better than my local AI stack… r7900xtx+32GB RAM (Qwen3.6-35B_Q4+OpenWebUI+SerXNG+Playwright+opencode). I wasn’t expecting much but it’s literally impossible to replace chatGPT level of…
i made fennara, a godot plugin + mcp for ai agents (www.reddit.com via reddit) https://reddit.com/link/1tydr1m/video/tat9wngg3n5h1/player hey, i made fennara for godot. it works both as an in-editor plugin and as mcp, so you can use it with stuff like codex, cursor, claude code, etc.
Moving from MVP to real product: Codex, Cursor, Claude, or something else? (www.reddit.com via reddit) Hi everyone, For the past month, I’ve been working on my first SaaS product - a kind of command center for credit advisors. The development has been fully AI-driven because I don’t have a programming background, but I’ve spent a lot of tim…
i built an open-source desktop shell for ai coding agents (www.reddit.com via reddit) i’ve been using claude, codex, terminals, browser tabs, files, and notes every day, and the workflow kept getting messy. the agents are powerful, but the workspace around them is broken.
Mobile device control of local PC cursor sessions on the roadmap? (www.reddit.com via reddit) This has probably been discussed somewhere but any chance allowing for mobile device control of local PC cursor sessions is on the roadmap? I understand the CLI allows this but what about from the GUI?
OpenAI Codex Sites feels less like a website builder and more like a deployable workspace surface (www.reddit.comhttps) I’ve been reading through OpenAI’s Codex Sites docs, and my takeaway is that this is not really “another AI website builder.” It feels more like Codex getting a deployable surface. The important part is not that it can generate a page.
Open-source local controller for multi-agent workflows — one inbox for Cursor, Claude Code, Codex, Antigravity sessions with inter-agent group chat (www.reddit.com via reddit) Been running 30+ AI coding agent sessions in parallel for about a year. Claude Code, Codex, Cursor (as of v4.7 last week), and Antigravity.
A Chinese startup just launched smart glasses that run Claude Code and Codex for hands-free "vibe coding" (www.reddit.comhttps) Just saw this and had to look it up. It’s actually real.
agentic code review is quietly replacing the way my team does PRs (www.reddit.com via reddit) Our PR review process used to be pretty painful. We have 6 devs and 2 seniors, and every meaningful review had to go through one of those two.
Quick Question: Claude Code Pro ($17) vs Codex Plus ($20) vs Cursor Pro ($20) - Which is better value? ( via reddit) could not extract summary
How Wasmer used Codex to build a Node.js runtime for the edge (openai.com) Just a moment... Verification successful.
Codex is becoming a productivity tool for everyone (openai.com) could not extract summary
How Braintrust turns customer requests into code with Codex (openai.com) How Braintrust turns customer requests into code with Codex | OpenAI May 29, 2026 Braintrust engineers use Codex with GPT‑5.5 to turn customer feature requests into preview branches in minutes and expand the scope of engineering experiment…
How Endava builds an agentic organization with Codex (openai.com) Endava, a global software contracting firm with engineers across Europe, the Americas, and Asia, has been an early adopter of Codex. For a business built around shipping quality software for banks, insurers, retailers, and media companies,…
Built a macOS notch app that shows your Claude token usage in real time (Claude Code + API both tracked) (www.reddit.com) If you use Claude Code heavily, you've probably hit that moment where the context window fills up and you're not sure how much of the conversation is still in play. Or you check your API bill at the end of the month and it's higher than ex…
I run 30+ Claude, Codex, and Antigravity sessions in parallel. Here's the v4 of the tool I built to keep them straight. (www.reddit.com) Why I built it in the first place. I've found myself running many agent sessions in parallel, just because I couldn’t stand waiting for each turn, and always had ideas/features for more things to build meanwhile.
Went to the monthly AI dev meetup (www.reddit.com) Usual crowd. Everyone's on Claude or Codex, nobody's really sure how any of it actually works, and that's fine, that's the vibe.
Do you hate tokens? I have the skill for you. /shotgun - run CC, codex, antigravity, cursor at same time for research then collate (www.reddit.com) Ask it to research something and it goes and finds all your CLI coding agents, optimizes your prompt, asks each coding agent to research the topic with subagents, then collates everything together into one final report. Dumb...
Introducing FLYWHEEL.md 🌀 (www.reddit.com) Agentic coding just crossed a line. Claude Code, Cursor, Codex, OpenClaw, the list keeps growing, and they all run fully autonomous now: /loop, /goal, crons.
Why have I spent so much more on AI coding than I’ve earned from it so far? (www.reddit.com) Since last year, I've been doing AI coding, using ChatGPT Codex(Pro), and Claude Code(Pro) , and even using Notion AI to manage all my documents. But so far, the money I've made from AI coding is nowhere near what I spend on AI each month.
GPT-5.5 tops the benchmarks but sits at #22 for actual usage - I built a live index that tracks both (open source) (www.reddit.com) I built AgentTape to rank models on more than just benchmarks - it blends benchmark performance with who's actually using and talking about a model, plus cost and speed. It scores every public model from public signals (GitHub, Hugging Fac…
ai literally never makes mistakes anymore (www.reddit.com) remember those memes a year ago which were like: "i spent 10 minutes vibe coding and 10 hours vibe debugging" I literally cannot remember the last time my agent made an app stopping mistake, it literally never happens before. no matter wha…
Amazing to see that Claude Code cannot replicate the designs done by Claude-Design (www.reddit.com) I have a React Native app that I am building in TSX and Claude-Design builds the designs in JSX files. The react native style blocks are pretty much the same with the css classes but yet the claude-design has so many problems in replicatin…
I don't want Codex. (www.reddit.com) Why is Codex installed by default with ChatGPT? I don't want it and it takes up space.
I got paranoid about OpenClaw skills injecting crap into my system prompt, so I built a quarantine pipeline with two LLMs as reviewers (Sonnet & Codex, 93.75% detection, zero false negatives) (www.reddit.com) Look, I know this sounds unhinged. "You made what to vet a skill before installing it?" But hear me out - OpenClaw skills go straight into your system prompt.
Okay who did this (www.reddit.com) Scanning the QR for codex mobile from the codex app leads to this url http://com.openai.chat//codex/open Notice the double slash there. 🤦♂️
Can't uninstall Codex app on Android (www.reddit.com) Hey everyone! I'm not sure if this is the right subreddit to post this in but I just noticed a new app on my phone called "Codex".
Weekly usage limits are designed to double dip on customers (www.reddit.com) I downgraded from ChatGPT Pro to Plus, and immediately after the next billing cycle, I have lost access to Codex for the next 6 days because OpenAI applies the lower Plus weekly limit right away, while still counting the usage I made yeste…
How many Codex weekly tokens are included in the ChatGPT Plus subscription ? (www.reddit.com) Title basically. I'm close to running out of my included weekly quota, with 5 days to spare.
Composer 2 couldn't fix my bug (www.reddit.com) A few weeks ago I randomly opened Chrome’s task manager (I do not recall why I did this exactly) and noticed one tab from my app (ai resume generator) was using over 9.9 GB of memory, and this freaked me out. The worst part was that the is…
I read threads complaining about codex every week... tf are y'alls workflows? (www.reddit.com) For context: I'm a software eng @ a fortune 500/FAANG tier company. We use AI.
Is Codex gaslighting me, or did Cmd-to-queue actually change? (www.reddit.com) I swear queueing a message in Codex required holding Cmd one week, and then suddenly didn’t the next. What makes this funnier is that there seem to be two camps internally debating the “right” way it should work, while I’m over here questi…
I asked an AI agent to promote a TikTok. It opened 48 PRs across our entire GitHub org while I was asleep. (www.reddit.com) I work an an AI startup. Yesterday afternoon I gave Codex (running 24/7 on a cloud box) one task: "promote our product video to 1000 views on TikTok." I watched it make the video, post it, closed my laptop, went to bed.
How Virgin Atlantic ships faster with Codex (openai.com) Virgin Atlantic used Codex to ship its revamped mobile app in time for the Christmas travel rush—one of the highest-risk periods of the year for potentially introducing software bugs. “We’re an operational airline, so we have to be very ca…
OpenAI named a Leader in enterprise coding agents by Gartner (openai.com) OpenAI has been recognized as a Leader in the Gartner® Magic Quadrant™ for Enterprise AI Coding Agents. We believe this evaluation reflects our progress in supporting enterprise-scale deployments of Codex, which is used by more than 4 mill…
Honest intake and experience on why AI agents suck (www.reddit.com) I’m a programmer and a victim of AI agent usage. 2 years ago, I was new into full-stack development.
Vibe coded an algorithm that prints money (www.reddit.com) Been quietly working on this for the past year. tried to write it by hand at the start but decided to do 90/10 vibe code because it was too much work for a simple person.
Google just NUKED my coding workflow. Will Cursor Pro burn through my $20 in a day? (www.reddit.com) Hey guys, I’m losing my mind trying to find an AI coding setup that can handle heavy usage on a strict $20/month budget. Claude Code and Codex burned through limits way too fast for my workflow.
Advice for my next move (www.reddit.com) Been using Cursor for a month in a half and really like it. Looking for some advice from people who have already been down my road.
Stop burning tokens making cli's re-read the repo every time you loose a session id (www.reddit.com) https://preview.redd.it/g5q7y7ik3a2h1.png?width=1384&format=png&auto=webp&s=edaa8bf897c96e8f29b6d6fdcea3f889c62401cb Have you ever started a CLI project and forgotten to note down the resume command when exiting, only to watch your agent r…
How Ramp engineers accelerate code review with Codex (openai.com) Just a moment... Verification successful.
The next phase of OpenAI’s Education for Countries (openai.com) A new era of agentic AI is here. With more than 900 million people using ChatGPT each week, and more than 4 million using Codex, agents have the potential to place far greater creative, intellectual, and technical power in the hands of eve…
Barry Cache remembers your repo (www.reddit.com) I’m lazy. Not in the “I refuse to work” way.
Claude Opus is still king for agentic coding, but Claude's app workflow is falling behind (www.reddit.com) I'm a paid Claude user, and I still think Claude Opus is the king model for agentic coding and serious coding work. The model is not the problem.
What's your prediction for workflows 12-18 months from now? (www.reddit.com) This are my employees, hooked up to WhatsApp and email. Can you guess who handles what?
Use Case: How I chain ChatGPT+Agents+Codex workloads (www.reddit.com) Context: I run interaction forensics and how people, communities, narratives, institutions and companies impact AI. Please note, all operations are human+AI.
OpenAI and Dell partner to bring Codex to hybrid and on-premise enterprise environments (openai.com) OpenAI and Dell Technologies are collaborating to help more enterprises deploy Codex in the environments where their most important data, systems, and workflows already live. Codex is becoming one of OpenAI’s fastest-growing enterprise pro…
Is anybody else tired of getting quota scammed? (www.reddit.com) Every week, the codex limits reset in a way to make you lose quota. If your reset date was May 18: surprise!
Exports Continue To Be Broken (www.reddit.com) For a couple of months I’ve been trying to export my data from ChatGPT and have failed. Apparently there’s a “known error” that occurs when someone has had two workspaces, and the system doesn’t know which one to export.
Is the plan usage lower than on demand? (www.reddit.com) So I ran out of my API usage for Codex 5.3. I added another $200 but then saw that in a few hours this on demand sucked up $50.
The reason why Claude subscription seems to have less capacity than Codex (www.reddit.com) I have a Claude Pro and a Codex Plus subscription. I created a container to: - Track my % usage on the 5H and 1 week window on both my Codex and Claude subscriptions.
What is the local LLM alternative of Codex? (www.reddit.com) Open AI codex got so many updates recently, it now does a lot of things in your computer, I tried a few, did not try all of them, and based on my experience with Open AI, they usually have more propaganda Anyway, what is the local LLM alte…
People running coding agents across real repos: what breaks after the agent writes the code? (www.reddit.com) I’m seeing a pattern with teams adopting Claude Code, Cursor, Codex-style workflows, etc. The coding step is not always the hardest part anymore.
How data science teams use Codex (openai.com) With Codex, data science teams can turn scattered inputs into usable analysis assets faster. Starting from dashboards, metric definitions, exports, experiment notes, and business context, Codex helps assemble a first draft of the deliverab…
How business operations teams use Codex (openai.com) Business operations work often starts across project trackers, KPI dashboards, planning docs, meeting notes, Slack threads, spreadsheets, and executive asks. Codex helps pull that context together and produce the first usable version of th…
Building a safe, effective sandbox to enable Codex on Windows (openai.com) When I joined the Codex engineering team in September 2025, Codex for Windows didn’t have a sandbox implementation meaning that Windows users were forced to choose between two subpar options when using OpenAI's coding agents: Approving nea…
Sea's View on the Future of Agentic Software Development with Codex (openai.com) Sea's View on the Future of Agentic Software Development with Codex | OpenAI Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Bu…
Is Codex really any better? (www.reddit.com) My developer friends (who were obsessed with Claude Code a month ago) are now all jumping ship to Codex, which has left me wondering: Is Codex better than Claude Code? I keep on hearing that it is cheaper, has lower usage limits and is ove…
Free open-source way to use ChatGPT/Codex subscription in Cursor natively (www.reddit.com) Hi everyone, I wanted to share a free open-source project that lets you use your existing ChatGPT / Codex monthly subscription inside Cursor: https://github.com/gabrii/Cursor-Azure-GPT-5 The idea is simple: if you already pay for ChatGPT /…
Are we at the point now where all it will take to create AGI is saying the correct sequence of words to Codex or Claude Code? (www.reddit.com) Seems to me like they can basically do everything software related now so surely a good enough sequence of input tokens would be enough. I guess in a way it's guaranteed since the frontier labs are doing all their work through agentic flow…
I run 30+ Claude/Codex/Gemini sessions in parallel. Open-sourced the dashboard. (www.reddit.com) https://www.youtube.com/watch?v=kEVyULB4r9c Sharing this in case it's useful. I've been running 30+ Claude Code sessions in parallel for months to ship two products.
Codex rate limits made me rage-build a tool to run two accounts at once (www.reddit.com) it the limit mid-session for the hundredth time and snapped. codex uses CODEX_HOME for auth/sessions.
Claude Code vs Codex: 36 files vs 28, $2.50 vs $2.04, and one infinite loop. My full breakdown. (www.reddit.com) I've been using Claude Code for months. It's been solid.
CTOP: htop for your Claude Code sessions (zero deps, pure Node.js TUI) (www.reddit.com) I built a terminal UI to monitor all my running Claude Code sessions from one central pane. I was running 6-15+ Claude and Codex sessions across different repos and had no way to see which ones were burning through context, which were idle…
Usage4Claude 3.0.0: open source macOS menu bar usage tracker for Claude, now with Codex support (www.reddit.com) Hi r/ClaudeAI, I posted an early version of Usage4Claude here a few months ago. I just released 3.0.0, so I wanted to share the update instead of pretending it is a brand new project.
I built a marketplace for AI agent skills and grew it to 17K users with $0 on ads. ChatGPT did all the SEO and content. Here's the full playbook. (www.reddit.com) I'm a solo non-technical founder. I built a marketplace called Agensi for SKILL.md skills (the files that teach AI coding agents like Codex CLI, Claude Code, and Cursor new capabilities).
**Honest question:** Is there ANY model of ANY size that is open source and can compete with Claude (Code) or ChatGPT's (Codex)? (www.reddit.com) All the open source models I tried are small and work OK with small problems. I understand the limitation of the hardware, context, etc.
How NVIDIA engineers and researchers build with Codex (openai.com) At NVIDIA, engineers are using Codex as their default tool for complex engineering work, and to run end-to-end machine learning experiments. Codex, built on GPT‑5.5 and running in production on NVIDIA GB200 and GB300 infrastructure, can ha…
What are the best opensource coding models for 8x A6000 setup (www.reddit.com) Currently using Qwen 3.6 27b and Qwen 3.6 35b but I was wondering if there is anything solid in the 50-200 range that you could run on a larger cluster that would be worth it? Or would you just run q8 or non quant versions instead?
Sharing OpenPets, a live usage and task-status for Claude code and Cowork (www.reddit.com) Codex Pets are fun and I'm often switching between Claude Code and OpenCode so I built OpenPets, an open-source project with a native macOS desktop app providing a CLI and an MCP server to connect any agents. A Swift library is included yo…
Coding agents can write the app, but still get stuck on infra (www.reddit.com) I’ve been using Codex for small app builds, and one thing kept breaking the loop. The agent could write most of the app, but the moment it needed something outside the repo, it would stop and ask me to set things up manually.
Using Claude to generate product prototype HTML is actually insane lol (www.reddit.com) I was just trying to mock up a simple personality test site, and Claude straight up generated an HTML version that looked almost exactly like what I had in mind. Kinda blew my mind tbh.
Running Codex safely at OpenAI (openai.com) As AI systems become more capable, they increasingly act on behalf of users. Coding agents can autonomously review repositories, run commands, and interact with development tools.
I wanted to know small local LLM code and made a personal projects. (www.reddit.com) I Started toying around LLM about sometimes ago. I think Qwen3.5 came out after like a month.
The Strange Codex Experience (www.reddit.com) I had a surreal experience today that felt like something out of a tech thriller. A Python script written by Codex, which is supposed to append data, suddenly claimed I only had 53 entries.
Disappointed in Qwen 3.6 coding capabilities (www.reddit.com) I know that coming from Codex I should adjust my expectations, but still. I'm working on a midsize project.
Simplex rethinks software development with Codex (openai.com) Simplex is a technology partner that works across consulting, systems development, and operations. To improve productivity in systems development, the company has quantitatively measured the impact of generative AI and applied those learni…
Auro Zera solves 78 and 280 year-old conjectures (Erdos Straus and Goldbach Conjecture) using Claude, GPT-5+, Grok, Deepseek, Gemini and self-made Dark Star ASI, proving superintelligence and opening a path towards resolving the Riemann Hypothesis , Twin Primes and more! (github.com via reddit) During this discovery utilizing only free AI services I have managed to undeniably prove both conjectures. This would absolutely not have been possible without using GPT5+ as the critic for my work.
Codex has failed (www.reddit.com) If it’s of any use to you, this is what Codex told me about my project Codex with gpt 5.5 high Yes. At this point, the most honest answer is: I am not able to see this project through to the outcome you’re asking for.
"Harness" lol (www.reddit.com) So the new buzz word..."harness"...makes me think which one shud i use...codex, forgecode,opencode, or a simple custom made harness with basic access to web tools and code execution ? (That i vibe coded :)
Singular Bank helps bankers move fast with ChatGPT and Codex (openai.com) Singular Bank helps bankers move fast with ChatGPT and Codex | OpenAI Skip to main content Research Products Business Developers Company Foundation(opens in a new window) Log inTry ChatGPT(opens in a new window) Research Products Business…
+50k spent in the past 6 months - I can tackle your questions (www.reddit.com) https://preview.redd.it/2773lua32dzg1.png?width=1794&format=png&auto=webp&s=34c981c52e1cb9d8b925c307da93084ecedbd265 I am a Founding Engineer of a startup with about 6 years of experience as Lead Architect. I can answer your questions on t…
Ollama 20$ vs (Claude Code or Codex) 20$ subscription (www.reddit.com) If you have to choose between these two, what would u choose? I've been using Ollama free cloud models; it is great for small tasks, but some cloud models need a subscription.
dead-letter: local .eml → .md (so hot right now) converter [CLI, Python, web UI, MCP] (www.reddit.com) Thought this tool might be of some help to everyone else out there given the amount of personal knowledge bases and Markdown pipelines being built. I made this specifically because I was burning context letting Claude (or Codex) unpack raw…
Better Claude Code session search in the menu bar (www.reddit.com) Built to make session search more robust (indexes much more than first prompt / project / etc., i.e., what's built into native session resume search) and so that I always have it handy in the menu bar. It indexes local Claude Code and Code…
Codex and vibe coding is TikTok for Coders (www.reddit.com) Having spent weeks to months on projects that never deliver, I'm beginning to see the profit model behind Sam Altman's GPT Codex. It's just an attention sink.
Cheap Claude/Codex/Gemini Models - Pay just 25% of official rates (www.reddit.com) Hey there, so I have been offering Claude (Codex and Gemini also available) models at the cheapest rate. I provide trial usage before payment.
[RELEASE] - coding agents can now talk! (www.reddit.com) Quick context: I use Claude Code and Codex daily and noticed I was spending half my "agent is working" time just sitting there watching the screen. I was like, what if Claude or Codex can just narrate its process back to me, so I know what…
Long term Pro user, thinking of trying Codex, some questions and concerns (www.reddit.com) I have been using the Pro subscription for a long time now and honestly even with its ups and downs it has been worth the price. By worth the price I mean that at this current point in my career, the 200$ per month is offsetted by whatever…
Bringing Codex computer use to iOS (www.reddit.com) Siri and Gemini can't actually do tasks on your phone. June can.
Has anyone else been hitting Claude max limits way faster lately? (www.reddit.com) I’m on Claude code (not using Opus 4.7 because it burns tokens too fast), mainly using Opus 4.6, and I’ve hit the weekly limit with 3 days still left. I usually don’t even get close to the cap.
claude code is amazing until you ask it to debug something (www.reddit.com) my agent workflow now covers most of the SDLC. chatgpt/codex helps me brainstorm (to save claude tokens) claude code writes the first pass, i clean it up, push, coderabbit handles the PR, deploys are mostly automated (using vercel but clou…
What is best code editor for local LLM deployment (LM Studio, llama.cpp) as of May 2026? (www.reddit.com) Hello folks What is best code editor for local LLM deployment (LM Studio, llama.cpp)? I wish to test my LM studio + Qwen 3.6 27B and Gemma 4 31B with a legit local code editor.
Codex CLI 0.128.0 adds /goal (simonwillison.net) 30th April 2026 - Link Blog Codex CLI 0.128.0 adds /goal. The latest version of OpenAI's Codex CLI coding agent adds their own version of the Ralph loop: you can now set a /goal and Codex will keep on looping until it evaluates that the go…
leaked my anthropic key into a public repo, lost $15,423. (www.reddit.com) quick story + something i built. mods please remove if this isn't allowed, read rule 7 first.
Quoting OpenAI Codex base_instructions (simonwillison.net) 28th April 2026 Never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures unless it is absolutely and unambiguously relevant to the user's query. — OpenAI Codex base_instructions, for GPT-5.5 Recen…
Opiniones para alguien que nunca ha usado Cursor (www.reddit.com) Que tal, buenas tardes, he probado diferentes agentes en CLI, Codex y Copilot, pues me interesa conseguir cuentas para mi equipo de trabajo, pero antes necesito analizar bien lo que se pagará, estaba considerando comprar Cursor Pro+ para p…
Compared 6 Codex CLI workflow systems in one table — what each pipeline actually looks like (www.reddit.com) Side-by-side: the canonical command pipeline of 6 popular Codex CLI workflow systems. Yellow = sub-loops (repeat per task / until verified).
Stop calling it an "agent harness." It's an Agent Runtime. (www.reddit.com) Hot take I've been chewing on: the "agent harness" framing is wrong, and it's holding back the way we think about this whole space. Look at what's actually inside Claude Code, Cursor, Codex, OpenCode.
Humanity's Last Hackathon - Use Codex from OpenAI to build Mac Metal kernels (www.reddit.com) Build, benchmark, and submit a Mac Metal kernels for local LLMs Claim access, set up your environment, and prepare for the qualification track. Use Codex from OpenAI to optimize a Mac Metal kernel against the benchmark task.
Qwen 35B-A3B as an always-on agentic loop on a 16GB Mac M4: disk became the bottleneck before RAM (www.reddit.com) M4 Mac Mini, 16GB unified, basic spec. For a few weeks I had Qwen 3.5 35B-A3B UD-IQ3_XXS (12GB on disk) running under llama.cpp with --mmap and --flash-attn.
Got the system prompt of Claude Design, released it for free (www.reddit.com) Claude Design is great, but I wanted to have similar capabilities with any LLM or agentic tools (Claude-Code, Codex etc). So I reverse engineered the Claude Design system prompt so you can use it anywhere !
Claude or openaı? (www.reddit.com) So i’ve been on the max plan for claude code for around 3 months now. And yeah somehow i was burning through all my tokens lol For context i’m a doctor.
How good is Qwen-3.6-27b? I asked Claude Opus (www.reddit.com) I ran an extensive code review on my project which has a large codebase. Ran the same code review on with Claude Code | Opus 4.6, Codex (high) | 5.3 Codex (high), and my local Qwen-3.6-27 (Q6_K with Q8 kvcache).
Where to find Codex usage limits? It was removed from Codex extension. (www.reddit.com) Title. I can no longer check my available limits.
Agent Kombat (kau.sh via reddit) Introducing "Agent Kombat" it takes one prompt or plan and turns it into a planning debate between Claude Code and Codex. Both agents produce independent plans first.
Keeping purpose in soon-to-be AI dominated fields (www.reddit.com) How do you prepare for LLM superiority in your field? I'm particularly looking for people who this can be expected to apply to in the near future, i.e.
locally uncensored v2.4.2 - chat, coding agent, image + video generation in one local app. plus remote access from your phone. one-click install (www.reddit.com) locally uncensored is a desktop app that combines four things most people run separately: chat, a coding agent, image generation, and video generation. all local, all on your hardware, no docker, no cloud account needed.
(Gemma/Qwen + Codex) - Bridging /chat/completions → /responses in llama-swap (www.reddit.com) I’ve been tinkering with a small side project (just for fun) where I’m trying to extend llama-swap with a bridge from /chat/completions to the newer /responses API so I can run the latest Gemma and Qwen models together with Codex-style too…
What's the consensus on superior local models for code generation? Is my setup competitive? (www.reddit.com) I'm trying as hard as I can to get a local setup somewhere in the ballpark of proprietary LLMs for code generation. My computer is running a Intel(R) Core(TM) Ultra 7 265K (3.90 GHz) with 128 GB of DDR5 RAM and an Nvidia Geforce RTX 5090 t…
Till what extent u guys use composer 2? (www.reddit.com) I m just curious to see if I m the only one using composer 2 in plenty of tasks recently. Discalimer I use gpt 3.5 codex for most of complex tasks, and previously I also use to ask simple tasks to gpt 3.5 but since i exhausted 100USD over…
A pelican for GPT-5.5 via the semi-official Codex backdoor API (simonwillison.net) A pelican for GPT-5.5 via the semi-official Codex backdoor API 23rd April 2026 GPT-5.5 is out. It’s available in OpenAI Codex and is rolling out to paid ChatGPT subscribers.
llm-openai-via-codex 0.1a0 (simonwillison.net) 23rd April 2026 Hijacks your Codex CLI credentials to make API calls with LLM, as described in my post about GPT-5.5. Recent articles - Claude Opus 4.8: "a modest but tangible improvement" - 28th May 2026 - I think Anthropic and OpenAI hav…
Codex settings (openai.com) could not extract summary
Working with Codex (openai.com) could not extract summary
Top 10 uses for Codex at work (openai.com) could not extract summary
Coding agent with integrated local model inference (www.reddit.com) Hi, I am wondering if people here have any experience on fully integrated local coding(general) agent setup. Yes we can probably setup claude code and codex of local server with MLX/LlamaCPP, but there are many trip wires and integration f…
Multi-agent coding. Feels like I'm playing the piano. (www.reddit.com) I made an MBTI-style Personality test… but your AI takes it instead of you (www.reddit.com) Who is actually writing code with local models? (www.reddit.com) Multi-session AI coding breaks flow in a dumb way: a permission prompt appears in another Claude Code session, and now you’re swiping around just to click ‘Allow’ ... (www.reddit.com) The quality of GPT-5.4 is infuriatingly POOR (www.reddit.com) I got a Codex membership when GPT-5.4 launched and was getting by well enough for a while. Then I started using Claude and GLM 5.1, and my production quality improved significantly.
Came across a benchmark comparing Claude Code, Codex and Sonarly on 200 real production bugs (www.reddit.com) a CTO friend sent me this benchmark last week and i've been thinking about it since. we've been dealing with the same production incident response problems internally so i ran similar tests on our own agent setup and the numbers lined up c…
New Codex features include the ability to use your computer in the background (arstechnica.com) A new version of OpenAI’s Codex desktop app reaches users today. It brings a smorgasbord of new features and changes, ranging from new developer capabilities to expansion into non-developer knowledge work to laying the groundwork for the c…
Does anyone also face repeated AI research across tools? (www.reddit.com) I work with multiple AI tools on same project, and I keep seeing this issue. Tool A already explored context, but Tool B starts same research from zero again.
Managing "collective consciousness" across multiple AI models without breaking the bank—how do you sync context? (www.reddit.com) Been running a distributed AI workflow to dodge token limits and play to each model's strengths, but I'm hitting a massive wall with context continuity. My current pipeline: Claude → High-level architecture & tech stack decisions (the "arc…
Cloud AI is getting expensive and I'm considering a Claude/Codex + local LLM hybrid for shipping web apps (www.reddit.com) I'm a designer who's been working on web apps and plugins for the past 5 months. Right now I'm building an After Effects plugin (close to shipping) and a music learning game experience.
Why do you think 5.2-codex was removed for ChatGPT logins? (www.reddit.com) Why do you think 5.2-codex was removed for ChatGPT logins? I think that this is done to our disadvantage.
When code costs almost nothing, the plan becomes the product. (www.reddit.com) Product Owner and Design Thinking Skills will be more valuable than ever before. Shipping code is a solved problem.
Mon premier site 100% IA : Quand l’artisanat rencontre l’Antigravity (www.reddit.com) Fiers de vous présenter un nouveau site d’artisan peintre en région bordelaise. Ici, pas d’agence de communication, mais une équipe de choc pilotée par l’intelligence artificielle : Claude : Mon bras droit pour la structure, un peu farfelu…
sentiment is shifting... (www.reddit.com) I have seen countless people announce that they are shifting from CC to Codex. They cite usage limits, outages and latency as things Codex is superior at.
You know you have become a "Senior Vibe Coder" when you actually stop and think about which AI model to use for a specific task. (www.reddit.com) Junior vibe coder: Throws the entire codebase at whatever frontier model is trending this week and burns their API budget in 4 hours. Senior vibe coder: "I need Codex 5.3 for rapid scaffolding, Sonnet for the Tailwind components, and I'm s…
I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry) (www.reddit.com) I started developing an app with Claude, but the credits run out very quickly. I thought that now with my new computer I could run something directly on it.
I built a process monitor that shows exactly how much RAM your Claude Code sessions are using (www.reddit.com) If you run multiple Claude Code sessions you've probably noticed your machine slowing down. I built agentop to answer the questions I kept asking myself: How many sessions do I actually have running?
Stop dunking on Plus users (www.reddit.com) OpenAI's Plus at $20 is genuinely usable. Not perfect, but you get real work done.
A director of engineering built a memory protocol across six coding agents in a week, and I think the findings are worth sharing (www.reddit.com) Domenico Lupinetti is a Director of Engineering at Translated in Rome. he found Signet and built a behavioral protocol called signet-first on top of it that works across Claude Code, OpenCode, Codex, Gemini CLI, and Github Copilot.
Master AI CLI Orchestrator? (www.reddit.com) I created a router that gives me access to Arena.ai models, and I generated an API key for each of the available models. I’m looking for a CLI tool that can run multiple AI agents together, each handling different tasks like planning, secu…
GPT is unuseable and problematic. (www.reddit.com) I'm not exactly sure what the use case is for GPT chat , the thing is inherently problematic. I have a sub for codex and that's good, however with the gpt chat system even for simply searching events or modelling work stuff (I work in inve…
Codex is officially nerfed (www.reddit.com) The recent changes to Codex limits have significantly degraded the developer experience. Previously, Codex in VS Code felt fluid and reliable for real-world workflows.
Possible Codex performance degradation during night hours? (www.reddit.com) Using Codex 5.3 inside Cursor at night feels like I’ve asked a baby to follow instructions. It keeps forgetting things I told it just a couple of steps earlier, rushes to give an answer without proper planning, and often skips looking at t…
Starting an app took up the entire usage limit. What's the problem here? (www.reddit.com) I tried cursor 3 for the first time today, after not using any cursor product for close to a year. So I was at 0% usage, far away from the limit.
CyberAgent moves faster with ChatGPT Enterprise and Codex (openai.com) paywalled
Codex now offers more flexible pricing for teams (openai.com) paywalled
codex is a MACHINE (almost 2 hours nonstop) (www.reddit.com) it only cost 23 cents aswell! absolutely insane!!!
Why Codex Security Doesn’t Include a SAST Report (openai.com) paywalled
Take the Vibe Coding survey, enter to win a $500 Amazon gift card (www.reddit.com) Hey all, I've been vibe coding like crazy for the last year, building with ChatGPT, Codex and other tools. I thought it would be useful to gather real data from you - the vibe coders - to create the first *2026 State of Vibe Coding Report*.
Rakuten fixes issues twice as fast with Codex (openai.com) Codex Security: now in research preview (openai.com) Confused about these Models on GITHUB COPILOT, NEED HELP (www.reddit.com) Is GPT Pro helpful if you're only using codex? (www.reddit.com) Thinking of buying Pro for a month (www.reddit.com) OpenAI Codex and Figma launch seamless code-to-design experience (openai.com) OpenAI Codex vs Claude Code: Why Developers Are Switching in 2026 (everydayaiblog.com via reddit) OpenClaw Creator Joins OpenAI: Zero to Hired in 90 Days (everydayaiblog.com via reddit) Beyond rate limits: scaling access to Codex and Sora (openai.com) Custom Kernels for All from Codex and Claude (huggingface.co) Introducing GPT-5.3-Codex-Spark (openai.com) Harness engineering: leveraging Codex in an agent-first world (openai.com) Are coding agents building complex features that will just become obsolete with the next model update? (www.reddit.com via reddit) My vibe coding journey so far (www.reddit.com via reddit) Unlocking the Codex harness: how we built the App Server (openai.com) Unrolling the Codex agent loop (openai.com) Datadog uses Codex for system-level code review (openai.com) How We Used Codex to Ship Sora for Android in 28 Days (openai.com) Codex is Open Sourcing AI models (huggingface.co) Building more with GPT-5.1-Codex-Max (openai.com) Introducing upgrades to Codex (openai.com) Powering next generation applications with OpenAI Codex (openai.com)