Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses like Claude Code, but a common complaint is to make even small edits we need to edit the code either manually or via the harness. To avoid this loop, I ended up creating Bento, a single HTML file with everything you need in a slide tool including animations and shared editing. There's no install or cloud login, everything works offline. The default deck is around 56
personal-ai-feed
AI-ranked personal intelligence dashboard. debug
More
First off I love 5.6 so this is not about the models but instead about the tools. I've noticed that the interaction between Terminal CLi and Gpt work (which replaced codex app?) is very janky since they emerged Codex and Chatgpt apps. On Codex work I can only select Terra Light or Sol Models. In terminal I get every Luna, Terra, Sol combo. But when I run in Terminal the permission despite what I put in terminal are always defaulting to the App permissions (which isn't always respecting my chatgp
I’ve seen a lot of people say they used a banked reset , only for Tibo to trigger a wider usage reset shortly afterwards. They still got some use from the banked reset, but it felt wasted because everyone’s usage was reset soon after. Now that Codex and ChatGPT Work have reached 10 million users , I had GPT look through Tibo’s recent X posts and replies to answer two questions: Are resets likely to continue after 10M? When could the next one happen? What Tibo has actually said When Codex reached
Today it is extremely stupid, and making too many mistakes, burning tokens at like never seen before. I'm really frustrated. Anyone else? Edit: I'm using gpt 5.6 sol medium reasoning. Some projects got a rule of 'no sub-agents' but it still launches here and there, that sneaky b*st*rd.
I was using Sol 5.6 High on a task with an active goal. Codex had reached a point where it needed to pause, so I explicitly told it to wait. Instead of simply remaining idle, it repeatedly produced variations of: “Still paused. The goal remains active and unchanged.” After that waiting period, Codex then reported: “Context automatically compacted” Approximately 15 minutes later, I returned and gave it instructions to resume the same task. Almost immediately after resuming, it compacted the conte
Every other time I open Codex there's an update. If I could stay off the pre-release update cycle I wouldn't have to deal with this nearly as much. I'm also not sure the point of calling it "alpha" if we don't have the choice to stay on official releases. I wouldn't care as much if I didn't have to restart every time, but on that note- if it could actually restart itself after an update like it says it will that would be great.
https://xcancel.com/mkratsios47/status/2079933645888880708
Heights based on benchmark scores. Was built with the help of these frontier models. Live at dat.city What dataset would like to see added next? I'm open to wild ideas.
Codex recently released its new keyboard, and honestly, I love the design. But spending more than $200 on it felt a bit hard to justify, so I tried recreating the experience with an app instead. It works surprisingly well as a software-only alternative. Still, if you’re buying it for the physical controls, the tactile feel, and the fun of having a dedicated device on your desk, the real keyboard is probably the better experience.
I ran a goal that took about 17h to finish, the goal was to create exact replica of my python desktop app but in react so that we could faster work on UI etc. it used 16 million tokens, and left me at 4% weekly so i guess the weekly limit for pro 5x (on 5.6 sol only) is 16-18mtok kinda little to be honest, dont you think?
I've been on the Claude Pro plan for a while, but my subscription expires tomorrow. One of the main reasons I'm considering switching is that I relied quite a bit on Claude Code, and after the recent changes (including the removal of the Fable feature from the Pro plan), I'm wondering if it's a good time to move to ChatGPT Codex Pro instead. For those of you who have used both recently: Is Codex Pro worth switching to today? What do you like or dislike compared to Claude Code? Have you regretted
If you ever wondered why your "dont make assumptions" isn't working: https://github.com/openai/codex/blob/main/codex-rs/collaboration-mode-templates/templates/default.md "In Default mode, strongly prefer making reasonable assumptions and executing the user's request rather than stopping to ask questions."