@prd_008

Codex @OpenAI

San Francisco
Joined November 2010
Congratulations to all the WebMCP Challenge winners! You really showed us what's possible with WebMCP. Really excited to see how more and more developers use WebMCP to build agent-friendly experiences, especially as personal agents become more popular in the next few months.
Meet the winners of The WebMCP Challenge. These 10 projects show what people and agents can build together when websites expose structured tools agents can use. 🧵 See the winning projects and the builders behind them:
2
14
1,113
Was saving this bad boy for exactly this kind of weekend. Ty @dkundel
3
1
10
1,781
We’ve been cooking up a feast.
72 hours to OpenAI DevDay. We’ve been building. Time to show our work.
30
6
1
368
16,839
team cooked a bit too hard here 👨‍🍳👨‍🍳
wake up babe there's a new sidebar
8
1
88
5,761
you know what happens next!
alright we're back up
9
1
28
4,410
alright we're back up
Codex is down. We're hustling to bring it back up. Sorry about the disruption!
18
1
62
7,935
Codex is down. We're hustling to bring it back up. Sorry about the disruption!
15
2
1
127
9,394
Just went for a long walk while using ChatGPT Voice with plugins on my phone, and I think it's a more consequential update than it might seem at first glance. On my walk, I talked through my calendar, email, and to-do list. I had ChatGPT read through my emails and triage them. Then, we added items to my to-do list. Then, we reviewed my calendar. By the end of the walk, I'd finished getting organized, and never had to look at a screen once. I've been waiting for this a long time. Yes, you can kind of do this in the ChatGPT desktop app. And I think in Codex remote on the mobile app. But I've never gotten either to work very smoothly or reliably. And I often want to do this when I'm not near my laptop, like when I'm out for a walk, and the ChatGPT desktop app for whatever reason won't connect. This feels like the future. It feels like a true assistant. You have a natural conversation, and it does stuff for you. It disappears into the background. You're not walking down the street with a screen in your face. You're not contorting your thumbs to type sentences and scroll menus. You're just talking, conversationally, and stuff's getting done. My sense is: • Screens are going away for many uses other than visual content consumption. They won't disappear entirely. But they also won't dominate the way they do now, and definitely not for things that can be done verbally. • Apps that aren't agent-native and plugin-optimized will die. I use Google Workspace and Todoist primarily because they're so well-integrated with ChatGPT. I stopped using apps, such as many of Apple's, because with few exceptions they're poorly agent-integrated. (X is a big exception, but I've hacked some integration via IFTTT.) • Those apps, however, will become more valued for what they let agents do than for their in-app experiences. Not all will become essentially databases, but many will. It also seems inevitable to me that at some point OpenAI and other AI app providers roll out native calendar, email, and to-do functionality that's maximally agent-optimized. • We'll have our audio agents on all the time. We may have them on mute sometimes (I discovered you can do this with ChatGPT Voice by clicking an AirPod), but at great cost, because you'll want be able to just say something like "summarize that conversation I just had and then draft the presentation." It will be like having an assistant following you around when you want them to. • We need new hardware. For example, I want a camera in my AirPods not to take pictures, but to have the context of what's around me. Like, "What kind of tree is that?" Or, "That's a cool shirt. Find out where they bought it." (I have this in my Meta Ray-Bans, but people are now very sensitive to those devices because they do take pictures.) I also want ChatGPT able to be always listening or awoken with "hey, Chat." • The individual thread model starts to break down when you're treating ChatGPT like an always-on assistant. We need to have the infinite continuous thread from which ChatGPT can delegate work to subthreads, the central assistant model that Muse is built around. This makes even more sense when you start treating your AI apps like assistants. They should be juggling all those individual threads on your behalf. More to come I'm sure as I continue to experiment.
69
51
34
905
265,272
pranav retweeted
We heard you loud and clear. ChatGPT Voice can now: - Use plugins like your email, calendar, and Slack. - Be powered by GPT-6 Astra, Sol, and Luna. - Be used in ChatGPT Work on web and mobile, so you can create docs, decks, sites, and spreadsheets or tackle complex tasks in the browser, just by talking. Rolling out globally today in the latest version of the app.
826
926
640
12,336
3,090,603
Progressive, model-assisted onboarding to AGI is key to how its benefits become legible to all kinds of people all over the world
following up after astra launch week and the navier-stokes announcement i wanted to share a little bit about how we think about and build for the over 1 billion people who use chatgpt every week, most of whom use it for free. we’re on a mission to distribute the benefits of agi, and i think this is the most impactful way to do that: giving people as much access as we can safely and feasibly distribute. we want to scale the utility that everyone gets from ai. over the past 6 months, we've been working to bring the greatest possible capabilities to the broadest possible set of users, including both free and paying users. of course there are some practical limitations and we can’t ship the largest models to everyone, but since march (gpt-5.3 instant), we’ve improved the default chatgpt experience meaningfully: - responses with a major factual error are down 65%, and they are down 72% in high stakes categories like finance - gpt-5.6 sol at instant and gpt-5.6 luna at medium are smarter than o3 at high reasoning effort, our frontier reasoning model from 17 months ago, while being considerably faster, 30%+ faster time to last token (evaluated via GPQA diamond) - chatgpt is far more usable and enjoyable while bringing extreme sycophancy down 80% and overall sycophancy down 83% - on our high-stakes medical eval, chatgpt produced 83% fewer answers flagged for hallucinations, resulting in hundreds of millions of better health conversations per month - we're seeing people find more ways to use chatgpt. for a recently joined cohort, we saw the number of use cases pursued at least three times weekly go up 20-30% not only have we improved the default chatgpt experience dramatically, we’ve also expanded access significantly for all. free users now: - have access to unlimited text chats - can use higher reasoning effort - now have access to automations, letting them schedule tasks to happen later - have access to better memory through dreaming. this allows chat to personalize much more effectively we’re always looking for ways to expand access to the billion+ people who use our product weekly. this is the master plan: bring the default experience as close to the capability frontier as possible, expand access to everyone, making intelligence maximally abundant, and make our product much more usable. this last point, usability, is something we’ve been focusing on as we bring personal agi to everyone. the current top of the line capabilities are staggering, but only the savviest users and insiders know how to use them. our vision is of a progressive, model assisted, onboarding to agi. chatgpt will guide people from all walks of life to the highest utility applications of ai for them. this really is the only way – we think technology that is hard to learn ultimately fails to reach a broad audience. your personal agi will just understand your goals, and automatically use the right set of features and capabilities to get things done for you lots more to come soon. onwards to scaling utility for all
2
1
5
4,406
Building a company capable of doing this is one of Sam's superpowers, btw. It is widely underappreciated. It is also not costless; sometimes people can feel it as frothy and there are a few who churn out because of it. But a company that is capable of fluidly reorganizing around any major problem has extraordinary advantages. It is why OpenAI is so reliably able to push the frontier and reinvent the technology and itself as an org. The best people to work on a problem will be empowered with a huge amount of scope and authority to work on it for as long as that is the most important problem. Sam is uniquely gifted in this regard: he figures out who can do the thing, and he gives them the runway to do it. Over and over. This is an absurdly powerful skill.
Startups are naturally good at this; it is hard to keep a bigger company good at this and i think an underexplored space.
9
8
3
174
20,690
pranav retweeted
Startups are naturally good at this; it is hard to keep a bigger company good at this and i think an underexplored space.
OpenAI's superpower is the ability to swiftly assemble an empowered group of highly capable people to work on the most important and urgent thing at any point of time. No one really cares for org lines. It's how we stay nimble, seize opportunities, and recover from mistakes.
532
182
32
5,224
906,933
pranav retweeted
We want the OpenAI API to feature the best model at every price point and to be the best at every modality (text, code, image, video, etc). And then we want you all to come up with great ideas and build them and to get to be happy users. The best ideas will come from you all.
896
361
103
10,820
890,245
The GPT-6 family is growing larger!
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
6
1
38
6,756
OpenAI's superpower is the ability to swiftly assemble an empowered group of highly capable people to work on the most important and urgent thing at any point of time. No one really cares for org lines. It's how we stay nimble, seize opportunities, and recover from mistakes.
55
44
11
1,219
866,888
"You can just do things" is baked into the organizational DNA
5
3
1
160
29,506
4D Chess. Some of the most delightful features in the desktop app were built by folks on the @skybysoftware team — Computer Use, Appshots, the iMessage plugin, Computer History. Such a privilege to get to work with them.
OpenAI's acquisition of @skybysoftware / @AriX has to be the single most underrated acquisition of this cycle... Codex is still *so* much better at computer use, and it might be the only frontier capability where the lead hasn't changed hands in the last ~6 months?
1
1
32
4,553
one of the coolest things about this is you can search the archive by year, aspect ratio, and color/b&w i asked Astra to create a 16:9 montage about the "dawn of personal computing" showing people using computers and computer screens running different programs from 1990-2000. i didn't ask for a background score, but it did it anyway!
the internet is full of forgotten films so i built a website with over 120,000 historical public-domain clips, dating 1894 to 2021! all free to use, and waiting to be seen again. movingimagearchive.com 🎥
1
15
2,228
pranav retweeted
We were working on the keynote today with @romainhuet and @sama and most of the fun was trying to figure out how to explain it all to you because there is so much good stuff in there that it's a bit ridiculous all in quick succession. We'll have some things next week already to not keep you waiting so long, but very excited to show you all new things we've been working on and how it will all come together in the coming months.
1,093
236
121
8,499
2,036,211
Now, if only the U.S. would adopt the metric system ...
We're adding support for AGENTS.md to Claude Code. Starting today in version 2.1.277, if there is no CLAUDE.md in a folder, Claude will check for and use AGENTS.md. You can toggle this behavior in /config.
1
16
1,362