Ep 816: ChatGPT Work and GPT-5.6 Sol: What’s New, 5 Overlooked Features and 1 Hot Take

Resources:

Join the discussion on LinkedIn: Got something to say? Let us know on LinkedIn and network with other AI leaders


Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup

Connect with Jordan Wilson: LinkedIn Profile

Start Here Series in our Inner Circle Community: Join for free access


Unlocking New Productivity: A Close Look at OpenAI’s Latest Model and Super App for Work

OpenAI’s recent release brings a suite of concrete improvements for organizations seeking tangible gains in operational efficiency. This release, which introduces a new family of models (Sol, Terra, and Luna) as well as the fully integrated ChatGPT Work app, offers immediate and specific upgrades to both the underlying AI and its direct application in the modern workplace. The real headline is not a single new feature, but the convergence of advanced capabilities, platform integration, and cost efficiencies designed for professional environments.

Here’s how the details break down for decision makers assessing where and how to allocate resources for AI-powered productivity.

Model Performance Benchmarks: AI for Coding, Knowledge Work, and Cost Savings

The conversation focused on the introduction of the Sol, Terra, and Luna models, with Sol positioned as the flagship. A key theme that emerged was performance-per-dollar, with Sol achieving higher benchmarks in coding, knowledge work, and cybersecurity compared to both prior OpenAI models and leading competitors.

Sol outperforms on cutting-edge, agentic benchmarks like the “Agents’ Last Exam”—comprising questions unseen in public training data—demonstrating reliability for complex problem solving and hands-off, long-running tasks. In head-to-head technical assessments, Sol edges out models such as Anthropic’s Fable Five, not only in raw performance but also in the cost of execution, coming in at nearly one-third the price per task (23:01). Terra, the mid-tier offering, matches or exceeds previous top models on coding while being five times more cost effective (39:20).

Organizations looking to upscale their automation and AI-driven workflows gain distinct advantages: precise execution, reduced overhead, and a smoother upgrade path without significant retraining or migration costs.

ChatGPT Work: Platform Unification and Actual Workplace Impact

The discussion explored the practicalities of ChatGPT Work—positioned by OpenAI as a super app. This “Work” app is not an overhaul, but a strategic re-bundling and interface refresh of tools previously segmented across ChatGPT, Codex, and Atlas (10:16). The real value emerges in surfacing previously hard-to-find settings and unifying tasks across devices.

Work can gather context, plan tasks, and execute across files, desktop apps, and browser environments, providing output in the user’s preferred document or analysis formats (08:27). It also integrates over 1,400 plugins for streamlined workflow connectivity. The outcome: employees and teams can consistently access—and actually locate—the most useful capabilities in one place, reducing onboarding and wasted time chasing disparate features (12:00).

Notably, the desktop app’s integration (formerly Codex) offers one unique business advantage—it can read, write, and act on local files and systems, enabling direct automation that pure cloud apps cannot offer (24:36). For businesses reliant on localized data operations, this extends AI’s reach and saves hours otherwise lost to repetitive file management.

Sites and Dashboards: Real-Time Collaboration and Single-Source Truth

Several points were raised, including the expansion of “Sites”—the built-in tool for creating and sharing interactive dashboards, trackers, and reports. No longer restricted to team plans, Sites now provides paid subscribers with a single source of truth, version tracking, and live updates to shared documents and internal web apps (27:05). This reduces the chaos and confusion of multi-version decks, databases, and collaborative assets, allowing information to be current across all stakeholder views without manual consolidation (28:04).

Mobile and Remote: Managing Workflows from Anywhere

Specific updates were highlighted for those working away from the desktop. The ChatGPT mobile app now fully supports plugin management and scheduled tasks, unifying oversight for recurring processes. The app’s remote function allows for seamless control of desktop work and scheduled automations directly from the mobile interface (29:14). The renaming and simplification of the mobile remote feature lowers the technical barrier for typical users, positioning it as an accessible and practical feature for distributed teams.

Browser Integration: True Cross-Platform Agentic Workflows

The discussion explored changes to the now-embedded browser environment, with the formal sunsetting of the Atlas browser and its replacement by improved, multi-tab browser access built into the app (30:02). This integration is fundamental for workflows that require interaction with password-protected systems and multiple third-party sites. Now, AI agents can use stored credentials and cookies, opening up new areas of automation previously gated by login barriers (31:11).

For organizations with recurring, manual digital tasks—think file exports, uploads, and multi-stage approval flows—this enables consistent, unattended execution, boosting staff productivity by multiples (32:45). Socketed into the rest of the platform, this browser integration is supported by live monitoring features via a Chrome extension, allowing staff to track AI-driven work in real time without constant context switching (34:17).

Competitive Implications: Market Positioning and Pricing Pressure

One concept discussed was the impact of these releases on the broader market, especially Anthropic. The medium-tier Terra model outpaces Fable Five in critical coding metrics while operating at a small fraction of the cost (39:20). This undercuts core value propositions in the enterprise and API segments, where price-per-token differentials are multiplied at scale. The competitive pressure is already visible, with rivals resetting limits in response (36:28), signaling imminent subscriber churn and a looming need for model and pricing adjustments.

Key Takeaways: Specific Business Advantages

  • Lower AI operational costs via superior performance-per-dollar, especially in coding and complex, hands-off tasks.

  • Streamlined workflows through unified tools; formerly disparate features consolidated under ChatGPT Work for both web and desktop.

  • Edge in site collaboration and document versioning, reducing time lost to fragmented edits and manual sync.

  • Direct, agentic desktop automation unlocks higher efficiency in environments with legacy or localized file structures.

  • Mobile-enabled management offers true anytime, anywhere oversight for ongoing or scheduled AI-driven work.

  • Enhanced browser automation with secure credential management further expands potential for hands-off execution of business processes.

  • Market shift indicates price pressure and performance benchmarks as primary differentiators, affecting procurement considerations in real terms.

OpenAI’s new releases provide not just incremental model improvements, but a specifically practical toolkit for business operations, enabling actionable steps toward higher productivity, faster turnaround, and real cost savings across a range of day-to-day professional applications.


Topics Covered in This Episode:

  1. GPT-5.6 Soul Model Launch Overview
  2. ChatGPT Work Super App Introduction
  3. Codex Platform Rebranding Explained
  4. Sol, Terra, Luna Model Tier Comparison
  5. Unified Plugins and Workflow Integration
  6. ChatGPT Sites Expanded Access Features
  7. ChatGPT Work Mobile App Remote Updates
  8. Atlas Browser Integration in Super App
  9. Advanced Agentic Browser and Automation
  10. Performance Benchmarks: GPT-5.6 vs. Fable 5
  11. Pricing Structure and Cost Efficiency
  12. Anthropic Competitive Landscape & Model Impact




Episode Transcript 




Jordan Wilson [00:00:18]:
Most AI tools at work don't know anything about you. Every time, you're stuck reexplaining your projects, your team, what actually matters. Slack just fixed that. The all new Slack bot is a personal AI agent built into Slack, and it starts with your context, your messages, files, and channels, translation, and AI that already knows how you work. Check it out at slack.com. Yeah. Yeah. Yeah.

Jordan Wilson [00:00:48]:
OpenAI's new GPT five six soul is amazing and the best and most practical AI model in the world right now. But they actually just pulled off one of the smartest naming tricks I've seen in a while. See, they took the popular agentic coding platform codex, soften the edges, got rid of that scary word code, and just repackaged it as chat g b t work, the super app. But everyone is rightfully paying attention to the new model itself, g p t 5.6 Soul, OpenAI's flagship release built for longer, cheaper, and more reliable execution of all type of work outputs. This sounds like one launch of two related products, but it's really a model story, a platform story, and a competitive story all wrapped up in one, and there's some details that I think people are overlooking. So that's exactly what we're gonna be tackling today on everyday AI. So here is the big picture. OpenAI just released their new 5.6 models.

Jordan Wilson [00:02:00]:
Yeah. GPD 5.6 models. There's multiple and a product they're calling chat GPT work or the super app. So the models themselves, well, by benchmarks and my limited use so far, look and feel like real upgrades built for long running, hands off professional work. So chat GPT work, because you're gonna be hearing me say work and the words ChatGPT work a lot. It's well, it's just kind of repackaging of some of the tools you didn't know you had inside of ChatGPT. And in case you did miss it, it's been a very busy, like, thirty six hours for OpenAI because not only did they come out with probably now the best model in the world, the best harness and making it more accessible for more people, but they also brought what I think could be more impactful in the long run, yesterday no. Two days ago.

Jordan Wilson [00:02:56]:
My gosh. Time's flying by. The new GPT live model, which brings that, duplex two way, voice interaction. But the big thing is GPT five six. So on today's show, you're gonna learn why Chad GPT work on the web is mostly just pretty packaging and where the real upgrade actually lives. I'm gonna tell you five overlooked features that people aren't really focusing on, and they probably should, including the full merge super app in the quiet death of Atlas. But don't worry. It's a good thing.

Jordan Wilson [00:03:33]:
And I'm gonna give you my hot take. OpenAI's positioning play and why I think Anthropic should be very nervous of a model that's not g p d five six soul. Yeah. A different opening eye model is gonna cause anthropic headaches. Let's get into it. Welcome to everyday AI. If you're new here, my name is Jordan Wilson, and, well, do this every day. It's for you.

Jordan Wilson [00:03:56]:
It's your daily livestream podcast and free daily newsletter, unedited, unscripted, helping business leaders like you and me keep up with the AI updates. I go through sleep deprivation, so you don't have to. You get to benefit from, just the good stuff to grow your company and your career. So it starts here, but make sure you go to your everydayai.com. Sign up for the free daily newsletter. We're gonna be recapping all of my random ramblings into the too long didn't listen version, so make sure you go do that and check out everything else that's happening in the world of AI in our newsletter. Alright. So let's get to it.

Jordan Wilson [00:04:31]:
We're talking about chat g b t work, the new g b t five six soul. I'm gonna tell you what's new, five overlooked features, and one hot take. So g p t five six soul. If you have missed the boat or maybe you've been on a long vacation and haven't listened to the podcast in two weeks, OpenAI already announced this. This wasn't new. We knew it was coming. They actually announced it about two weeks ago, but they had to work with the federal government now that we have kind of the permission slip, you know, way of doing AI here in The US because of anthropic. But here it is.

Jordan Wilson [00:05:08]:
Brand new model. But the difference is, well, it is actually maybe a little more confusing in the long run. Because before, we just had things like g p t five five. Right? And then you had different levels of reasoning and all that. So now we still have the different levels of reasoning. We still have the number, the five six, but now we also have Sol, which is the big model. We have Terra, which is the media model, and then we have Luna. Right? So sun, earth, moon, which is great.

Jordan Wilson [00:05:40]:
And I I think that that's gonna carry on. So let's talk a little bit about the new model, and this is from OpenAI's release. So they said we're launching the g p d five six family of models for general availability following our limited preview, our new flagship Soul alongside Terra, a balanced model for everyday work, and Luna, our most cost efficient model. The g p d five six six Sol sets a new standard for both intelligence and efficiency, achieving state of the art results across coding, knowledge work, cybersecurity, and science, while outperforming previous and competing frontier models with fewer tokens and at a lower estimated cost. The result is stronger performance per dollar, more successful work for the same spend, or comparable results at a lower total cost. We also introduce a new way to accelerate the most demanding work. Ultra is our highest capability setting, coordinating multiple agents across parallel work streams to finish complex tasks faster. Stronger computer use and design judgment make g p t five six Sol our most polished collaborator yet, helping it inspect, refine, and deliver ready to use results.

Jordan Wilson [00:06:46]:
Alright. So we're gonna talk a little bit more, about GPT five six Soul. I'm gonna do a little bit comparison, not too much on this show, but obviously, Fable five. Right? We have to talk about those two models together when we go over the benchmarks here in a couple of minutes. But I wanna quickly tell you a little bit about the other big release. So we knew, GPT five six Soul was coming out yesterday on Thursday. What we didn't necessarily know is, well, we were essentially getting the super app. Right? Internally, they've kind of saying it's it's the merge, but this is the super app.

Jordan Wilson [00:07:23]:
So this is kind of the combination of chat GPT, codecs, which has been on fire the last, five months, and the Atlas browser. So the Atlas browser is no more. Now codecs is chat g p t and chat g p t is codex. It's actually both an amazing play, just call simply calling it chat g p t work, but also kind of confusing with how it rolled out because if you go download codex or if you go update codex, you know, now it says, you know, chat g p t, but you can change it. You know, you can say, hey. I want it to be the codex icon and still call it codex, but it's the same thing. Right? But, kind of how OpenAI, says, put their posts out about check work. They said powered by g b t five six, check g b t work brings together context from your team's tools to turn scattered notes, drafts, and ideas into finished work, and it keeps projects moving while you stay in control.

Jordan Wilson [00:08:27]:
They say, ChatGPT work gathers context, plans the approach, and takes action across your tools, files, and desktop apps to create polished spreadsheets, docs, and slides. CHEDGPT can turn context from your tools and files into polished documents, presentations, and analysis to better, that better follow your templates and preferred formats. And then they say sites. Yeah. Updates on sites, and I'm gonna share about that more in a minute. Sites let you turn ideas, plans, and data into interactive websites and web apps that are easy to share with your team. Create dashboards, project trackers, launch calendars, prototypes, reports, and more, and keep them up to date as information changes. So then talking about the unified plug ins, they say with more than 1,400 plug ins available, JetGPD can pull context from the tools and workflows you already use to help move projects forward.

Jordan Wilson [00:09:19]:
They say create one time recurring tasks, monitor updates, and check progress from your phone while you're away from your desk. And I'm gonna be talking about those chat g b t remote updates because I think they're really good. Then the new, side by side, in chat gbt, which I like. It's the new built in browser experience in the chat gbt desktop app. Makes it easier to work across your tools, files, and accounts with support for multiple tabs and richer agentic workflows. And I'm gonna share at the end some of my, tips on the best way to do that because that's, for me, the biggest unlock. A very small piece of this is the biggest unlock. And then review the approach before work begins in plan mode.

Jordan Wilson [00:10:02]:
Chat g v t gathers context, ask questions, and creates a step by step plan. You can suggest changes or approve the play, the plan to start the work. Alright. So let me translate all of that marketing speak. This is codex. That's it. There's really nothing new. So I don't know if if most people remember this, but, OpenAI added a feature about two weeks ago or or sorry, about two months ago.

Jordan Wilson [00:10:30]:
Essentially, it allowed you to choose how you wanted codex to display. So you could either have a more technical display or something they just called for everyday work. So you could just choose how much technical detail codex showed when it was working. So So did you wanna see all the tokens spinning and all the Python's calling and all the, you know, tools spinning off the shelf? Right? If so, you would choose the coding approach, which is technically now codex. If you didn't, if you just wanted a more minimalistic and you didn't really care about what was going on under the hood, well, that's now just called chat GPT work. I spent so much time, and I'm like, this can't be true both on the web because they did the same thing on the web. On the web, now you have chat g p t, and there's a toggle for chat and a toggle for work. And, you know, on the web, at least, chat g p t work looks a little different.

Jordan Wilson [00:11:29]:
There's technically no real new functionality aside from being powered by the new model. The plugins are unified, but there's no new anything. Right? So you look at it and you're like, oh my gosh. This is gonna be amazing. What's new here? All this is, which it's smart. Don't get me wrong. Right? I just was like, okay. I just, you know, read the blog post six times, watched every single video that OpenAI put out, you know, played with every feature, and I'm like, oh, it's literally the same.

Jordan Wilson [00:11:58]:
But I don't think that's a bad thing. I think it's smart what OpenAI did here. Because if I'm being honest, even power users of Chad GPT, I could go in. I feel show them things, and then people would be like, wait. What? Like, where'd you find that setting? Right? There's literally be some things that were some of the best parts of Chatt GPT that you couldn't even find. Right? You had to go into, like, a a setting, like, four or five rows in, and it's like, oh my gosh. You can do that in Chatt GPT. It's like, yeah.

Jordan Wilson [00:12:26]:
You can't. You know, like, even, like, scheduled tasks, they used to, they took that away for a couple of months, but you could still access it via a URL or by clicking, like, four or five links deep inside of a setting panel. So you're gonna probably be hearing a lot of chatter about the new chat g b t work. I do think it's smart. So on the web, it's just a more, you know, it does look like more codex on the web, but you don't really have the codex features per se. But on the desktop, it's just codex. ChatGPT work is codex, and you can choose to call it ChatGPT or call it codex, but you will have a, kind of a switcher at the top, and you can switch between, ChatGPT work and ChatGPT codex, but nothing really changes. But like I said, I think it's smart.

Jordan Wilson [00:13:21]:
And here's the play. It's still baffled me how many smart people I've talked to over the last five or six months about codecs. Right? Because on one hand, yes, It is an extremely powerful agentic platform that you can automate your life. You can code. You can build things. Right? And that's intimidating for people. Right? But and I've been saying this on the show. You you know, I've been, you know, captain codex since February, and it absolutely crushes every competitor, and it's not even close.

Jordan Wilson [00:13:56]:
It's way better than clawed desktop, clawed code, clawed co work. It's better than cursor. I think cursor is probably in second place. It's better than anti gravity Gemini. It is the best desktop AI app, and it is not even close. But the problem is is the word codex. Right? People said, oh, I'm not a coder. Right? Yeah.

Jordan Wilson [00:14:14]:
I I've had so many recent conversations. People who are even somewhat technical, but they're not coders. You And they're like, oh, I can't use codex. And I'm like, you absolutely have to. So from that aspect, it's a very smart play. They literally just slap the word work on, didn't really change anything, just shows fewer details. That's it. Alright? So don't get confused when you're like, what's this new CHAD TBT work? It's really nothing new.

Jordan Wilson [00:14:39]:
Alright. So, yeah, here on my screen, I have the CHAD TBT work on the web just looks kind of like codex light, but it doesn't really do anything new or different that you couldn't previously do on the web inside of Cheggitt. So, yeah, the web version just bundles agent work, deep research projects, connected apps. So it just brought some of the more, technically advanced features, the ones that I use all the time, and it kinda gave it its own, dedicated panel in the work. You can still do all those things in the normal chat panel. It just, you know, makes the chat panel maybe a little bit easier, and less intimidating with, you know, fewer interfaces. And then the work panel just makes it easier, and then it suggests different things. Alright.

Jordan Wilson [00:15:25]:
And then well, there's one new thing. There's a nice little toggle, right, inside the chat chat g p t, desktop app. Yeah. There's nothing really new there. But Atlas is built in. Sorry. I take that back. Atlas is now built into the browser, and it is really good.

Jordan Wilson [00:15:43]:
It is improved. And I'm gonna give you here in a couple of minutes, when I go over my five overlooked features. One of them is going to, hopefully, I think, change the way a lot of people work. Alright. So more on g p t five six, you know, gotta get in the benchmarks, the comparisons, all that, but I need a quick water break, and we have to quickly hear, word from our partners. Here's what most AI tools still can't do, work outside their own little box. The all new Slack bot just changed that. It's your AI teammate inside Slack, and now it can read, write, and act across the other apps your team already uses.

Jordan Wilson [00:16:22]:
No more tool switching. No more reexplaining yourself every time you open a new tab. One ops team at engine says the summary feature alone saves them fifteen to twenty minutes of use. See what Slack bot can do at slack.com. Alright. So a little bit more on the new models. So g p d five six, here's the basics. There's three models, Sol, Terra, Luna.

Jordan Wilson [00:16:48]:
Sol is the flagship. It is aimed at the heavy users, coding, knowledge work, cybersecurity, and science. If you are someone that's using AI for three, five, six, twelve hours a day, you're probably gonna wanna use Sol because you will be able to take advantage of those full capabilities. We always talk about the capability gap. Right? If if you're doing lighter work, maybe you can use Terra. Right? That is the more everyday work at a lower cost. And then LUNA is the fastest cheapest tier. So, yeah, that's our our new three naming tier.

Jordan Wilson [00:17:24]:
And then on top of that, you have, Sol Ultra as well. So let's look at that. So with Sol so we're gonna get into the, the benchmarks here. It did extremely well. So even on a newer benchmark, which I really like, if you've heard of humanity's last exam, it's essentially all these questions that are not in training data. So there's a similar one called agents last exam, and it's been one of the more popular and trending benchmarks of 2026, and it just crushed the competition. So, you you know, and even if you're sitting there and you're like, I don't need agents. I'm not a coder.

Jordan Wilson [00:18:07]:
Yes. You do. And, yes, you will. You just don't know it yet. And and and I'm being dead serious. Right? The future of work. If you are a knowledge worker sitting in front of a computer, you are an agent orchestrator. It's it's how I've been doing my work since February, since codex came out.

Jordan Wilson [00:18:28]:
You know, constantly providing a lot of feedback, on the front ends, you know, keeping an eye on your agents, seeing what they return you. That's your work. You know, you polish it a little bit, but you're able to get four, five, six, ten times, done that of the type of work that you would normally do and, you know, unlocking a whole new skill set that you never thought you could do. So but the company's core pitch here is just performance per dollar, and we're gonna get into that in some of the charts. And then there is technically a tier ish on top of Sol, and that's Sol Ultra. And that's just when they run, essentially four, kind of sub agents in parallel, all of the highest setting, and you kinda get the best of the best. So Ultra does kind of coordinate those four agents in parallel by default to finish those demanding tasks faster. We've gone over pricing before, but Sol is $5 input and $30 per output 4,000,000 tokens.

Jordan Wilson [00:19:26]:
So that's just if you're paying on the API side. Right? So if you're in, if you have a chat GPT subscription, FYI, that now means you have a chat GPT work or a codec subscription. It's the same thing. It shares those limits. So, obviously, you're not, you know, having to worry about that, but the pricing is extremely important. And then, obviously, the cheaper per task, can still drain, plans quickly. So, you know, if you are using, the ultra, Sol Ultra, that's that's gonna eat eat things away. Alright.

Jordan Wilson [00:20:00]:
Let's talk a little bit about some benchmarks. We're not gonna go too deep into it, but it's important to talk about. So I think probably the best, single source for benchmarks, is artificial analysis. So it is, conglomerate. It just brings in, you know, different benchmarks all at once. So it averages all these different benchmarks. So it's one of the best sources to look at. So, GPT five six Sol max, pretty much on par, with Claude Fable five.

Jordan Wilson [00:20:33]:
It's a one point difference, which is pretty minuscule. Right? I'd say you can start power users can start noticing at about a three point difference. Average users maybe, like, a six point difference, and they're not gonna tell the difference. But, you know, if if you are a heavy user, you're not gonna be able to tell the difference between a 59 and a 60 on artificial analysis intelligence index. That means they are essentially one a and one a a b. It's not one a and one b. It's one a and one a a b. But that's just on that, the artificial analysis intelligence benchmark.

Jordan Wilson [00:21:10]:
Here's where it gets interesting. On, they have essentially, they have three main benchmarks. I usually just talk about the, intelligence index cause that brings a little bit of everything. But in this case, when we're talking about work knowledge work becoming more and more agentic, I think that's actually when the other two artificial, analysis, benchmarks become more important. And that is their artificial analysis coding index, which GPT five six soul, beats Fable five. Alright. So it beats it on coding. And on the agentic index, five five also beats it.

Jordan Wilson [00:21:46]:
So, essentially, I would argue by most standards that would mean that, GPD five six soul is better. I mean, we'll see when all the benchmarks come out in the LM arena. So what do humans prefer? Right. And, you know, OpenAI has said they're working on front end design. I know so much, which I never understood. Right? Because so much of, like, vibes that are out there for models. It's always like, okay. I created this video game, and, like, look at how good it looks.

Jordan Wilson [00:22:17]:
It's like, okay. Well, maybe, like, what? Point 001% of the world uses models for that. Right? I use models for knowledge work. I use models for personalizing research, carrying my context over from different apps, you you know, building, you know, building decks, building little sites that I can share. Right? You you know, helping me grow my business. Right? I'm I'm not creating three g three j s worlds. You know, I I do a a little bit of design things maybe for some of the apps that I do build. But, you know, I think it's important to look at these three benchmarks in tandem.

Jordan Wilson [00:22:54]:
But the one that is the most telling, and this is we're gonna get to this in a little bit more, is the price. So for artificial analysis, this is essentially the cost per intelligence. Right? Because these models can all get the tasks done, but, well, how much does it cost? And right now, GPT five six, I will argue, is a slightly better model if you look at all these, you know, the three different artificial analysis indexes that are the most important. You know, they're one point behind on one, winning the other two, but they're doing it at, yeah, a dollar per task, a dollar and 4¢ versus, Fable five, $2.75. Yeah. Infrapping. Alright. So, we went over Chad GPT work.

Jordan Wilson [00:23:43]:
We went over codex. Now I'm gonna give you the five overlooked features and my one hot take. Alright. Overlook feature number one, Even though OpenAI didn't say it out loud, this is the super app. Right? I think there there are a lot of super app, you know, rah rah back in, like, March, April, May. And at the time, OpenAI said, well, codecs is the super app. Right? They said, you know, they said we're building it in disguise, but this is it. But it's bringing together chat, work, codex, browser, files, and actions.

Jordan Wilson [00:24:23]:
And if you know absolutely nothing about, like, well, codex, right, or why does it matter, you know, why would why do I need a desktop, you you know, something running on my desktop versus something, you know, that I can just do in the cloud. Right? If chat g p t has this new chat g p t work on the web, why would I ever use codecs? Well, the main reason is it can access your local system. So it can use your entire computer. It can use your browser. Right? It can access, read, and write all of your local files. So those are, you know, three big advantages. And, you know, there's a little bit more that you can do with scheduling and things like that if you have a a machine like I do that just stays on my poor Mac Studio, I don't think it's taken a nap since I got it. Right.

Jordan Wilson [00:25:10]:
So I have a lot of things that are scheduled, you know, loops, goals, heartbeats, all these things. There's a little bit more, technical, capabilities that you can do with the desktop app, but it is now all one. There's some syncing issues, FYI, and I'll probably reach out to OpenAI about it. I I don't know if they're fully, aware. Because as an example, some things change names. Right? So if you were using codex before, there was a thing called chats. Now there's no codex chats. Codex chats are now tasks, so they could be called codex tasks or they could be called chat g p t work tasks anyways.

Jordan Wilson [00:25:50]:
But you can there's a new chats panel, and that brings in your chats from chat g p t. But right now, at least, only chats from the chat tab inside chat g b t are syncing and not the chats from the work tab inside chat g b t. Yeah. Confusing. Right? So there are some things that, you know, the the the super app in bringing all this together in one cohesive ball. I think that there's a couple of things that the team is probably gonna tighten up. But, I mean, first iteration, it is really, really good. Alright.

Jordan Wilson [00:26:23]:
The other thing, small footnote sites. Alright. Chachapiti sites rolling out to all paid plans. Alright. I think there's some might be some, geography restrictions at first, but, rolling out to all paid plans, and this is important. Also, there's public publishing, live URLs, and it's on Chad GPT on the web, not just codex. So we are starting to see, you know, a little bit more feature parity. You know, certain like, a good example is, Chad GPT sites used to only be on codex.

Jordan Wilson [00:26:58]:
So, well, now that's also, on Chad GPT on the web as well. So here's why this is important. Because before, sites was, number one, only available on codex. So now it's available on the web as well. But it's also, available for all paid plans. Before, it was just for team plans. If if if I'm being honest, like, it it was one of the reasons that I still was holding on, to my team plan, even though I'm mainly the only person that uses it was for ChatGPT sites. I really like it.

Jordan Wilson [00:27:30]:
But on my $200 month pro plan, it's like I wanted to be able to use sites. Now I can. So if you have a paid plan, ChatGibidi sites is great. It's like a a lovable light. Right? So all you know, I I I did a full episode on it a couple of weeks ago. Think of it like this. You know, think of how many versions of decks that you have. Right? There's 10 different versions on 10 different people's machines, v one, v three, v four final.

Jordan Wilson [00:27:55]:
Oh, did we have edits on this one? Right? You probably waste hundreds of hours a year just dealing with, who has that file. Right? Chatchifyd sites can just get rid of that because there's just one source of truth. And anytime one person updates it, it's updated everywhere. So you can build full on, you know, databases, dashboards, you you know, little apps that your team can use or just decks and websites to keep track of anything type of work. So it's nice that it's rolling out to all plans and available on the web. Next, Chad GPT Mobile got a huge update. Now you can manage plug ins from Chad GPT app. You can see and manage all of your scheduled tasks.

Jordan Wilson [00:28:40]:
So, that's great. But also, they changed the name. So you used to open up the chat g p t app, and it would say codex, and then you could essentially control your, any physical machine that was running codex. So now it's just called remote, which I kind of like. Right? So they're not slapping the name codex everywhere and confusing people. And if you see remote, you're like, cool, even if you're a nontechnical person. And now all of a sudden, you're like, wait. I can use the chat gbt app, which, you know, OpenAI has, like, a billion weekly active users, and everyone knows the chat g b t app.

Jordan Wilson [00:29:14]:
And then they see this thing called remote, and they're like, oh, I can, you know, control my desktop if I have the chat g b t work app. Wow. Great. So, again, not really new functionality there. It's just cleaned up, and it does work a little bit better. But, actually, no. New there is new functionality because you have more control over plug ins and schedule tasks on the actual code app. Sorry, on the actual remote section of the chat g p t app.

Jordan Wilson [00:29:41]:
Alright. Number four, yeah. Atlas is technically gone. So OpenAI's browser experiments, it is no more, but it's for the best. Because now this is gonna sound small. Alright? But two big updates to, the browser, formerly known as Atlas that is now just baked into the super app. Right? So Codecs had always had a browser, but it was a singular browser. And it didn't have all the normal browser capabilities.

Jordan Wilson [00:30:12]:
Like, you couldn't open multiple tabs, and it didn't keep your saved passwords. I actually spent, like, 1,800,000,000 tokens building what I called I forgot what I called it. I think I just called it, like, codex vault. But I essentially built an app that would save all of my passwords so I could run, you know, certain things agentically that required you to log in because, you know, the codex browser, you know, couldn't, log in to things without saved passwords. So that is available now. And you can import from any of your Chrome profiles because the browser is based on Chromium like Atlas was. So, it supports cookies to keep you logged in, and it'll support password. So I was playing around with it a little bit, because it was a little finicky to actually get computer used to click log in because some websites, depending on how they're set up, you know, the actual browser control can log you in.

Jordan Wilson [00:31:11]:
So it'll pull up your saved, you know, websites, your saved credentials. Right? Just like in Chrome, if you, you know, sign in to something and you've saved your password in Chrome, it'll be there, and you just have to click log in. So I had to toy around with it a little bit. I might just, if I can refine it, I just might save it as a skill and, you know, put it in our newsletter or on our website because it is a little finicky now. I'm sure they'll get it fixed. But this is the big unlock here y'all, and here's why. Yes. You you know, there's there's connectors and plugins and MCPs and all these things.

Jordan Wilson [00:31:44]:
Right? But I think for knowledge work, so much of what we do is going to all of these different websites that don't have MCP, and they don't have direct connectors. Right? So in my example, you know, the our podcast is hosted on Buzzsprout, which is an older host. They don't have any AI anything. I spend so much time, you know, going from this program, you know, downloading this, uploading this file here. I clean the audio, then I bring it in, you know, get the time stamps, and it's like five different programs. Well, now I can just fully automate that completely because codex can open multiple tabs, at once, and it can save the passwords because, you know, it does save them. It had previously saved them for a certain amount of time, but every once in a while, it just kicks you out like any normal, you know, SAS program would. So now with the cookies and the saved passwords, huge in terms of the new knowledge work that you can do on a schedule.

Jordan Wilson [00:32:45]:
Think of all those monotonous tasks yet to go log in, click this thing 12 times. You're like, why am I doing this? Right? Why you know, even if you're a high paid executive, you're like, I I, you know, I just every single day, I have forty five minutes of things that, you know, a monkey could do. Right? I just go here, click this, wait for it to load. Three minutes. Oops. Got distracted. Click this. Export.

Jordan Wilson [00:33:07]:
Download. Right? You don't have to do any of that anymore. It's amazing. Alright. And then last but not least, computer use is much faster. It is faster. It is better, and it has this real cool new thing. It's this picture in picture.

Jordan Wilson [00:33:20]:
So when computer use is happening, you can see, you know, if you've ever seen picture in picture on your TV, it's just right there in codex, or if you decide to call your codex chat g p t work. So you can see what's happening, right there in codex. Oh, another thing that I didn't mention, which I should have mentioned on number four with Atlas, related to this and related to this super app. And I don't even think OpenAI announced this, and I didn't see anyone else talk about it. But in the Chrome extension, because you can also from codecs not just control your computer, not just control the built in browser, but you can also control a Chrome browser, which is a little different. Alright. But they actually have a Chrome extension, and there's an update in the Chrome extension where you can have a side panel in Chrome, and you can see what is being worked on on your desktop, which is really cool. So, you know, if I have a a window open, which is actually unlocks a lot.

Jordan Wilson [00:34:23]:
Because normally, on on my setup, right, I have two screens. One screen is just full, you know, full codex, you know, and then I'll have, you know, probably a full clawed desktop, and then I have to swap between them because I'm like, oh, okay. And I know, you know, codecs has the little pets, and they say, alright. Your thing's ready. But I'm still, like, having to always keep an eye on that. So having a sidebar in Chrome, and I can see what codecs is working on and track it live. That's been, like, a small little thing, and I'm like, wait. Why didn't they talk about this? This is huge, especially if you're like me and you're giving, you you know, agents long running tasks.

Jordan Wilson [00:35:00]:
You know, at any point, I might have, you know, five to 20 different threads running. And it's like, I don't wanna always have to check-in on them, and sometimes the little codex pet for me is a little distracting. And I'm always in a Chrome browser, so I'm like, oh, okay. I can just have that going and work with it right there. Actually, amazing. Alright. So that is the five, overlooked features. And let me end with this hot take.

Jordan Wilson [00:35:28]:
For those of you that stuck around, Anthropic's in trouble. They are in trouble. If they do not immediately pivot, it's not going to look good. All right. And here's the thing. Yeah. Anthropic is scared and they should be right. And this is great because it's going to drive competition.

Jordan Wilson [00:35:53]:
Anthropic did something that I don't think they've ever done right after this announcement. And this, and this is how I know INTROPIT is in huge, huge trouble. They reset weekly limits, which I don't think they've ever done except one time when they made a big mistake. There's this big miscommunication. I think they might have reset five hour limits, not weekly limits. That's something if you use codex, like, we've been spoiled. Right? Tivo over there, at OpenAI always resetting your weekly limits. Infrappy did this.

Jordan Wilson [00:36:28]:
They never do it, and they did it right when OpenAI announced something. And that to me is like, they know they're scared, and they know I'm I'm I'm not gonna make projection. I'll make a projection. I I'd say it's very, very safe to say that yesterday, they lost probably 8 figures in revenue, I would guess. And I will continue to say that until they make a change, they're gonna be losing 8 figures of revenue every single day, maybe more. It's they're gonna be losing a lot of money if they do not pivot immediately. So I even heard some people say, and I'm not gonna say this. Right? People are like, oh, Fable five is dead.

Jordan Wilson [00:37:06]:
It's not dead. It's still, you know, a top tier model, and there's still gonna be some people out there paying for it. But I would absolutely not pay a dime for Fable five. And, in Propic, they said, oh, we're gonna, you know, extend Fable five in, subscriptions through July 12. Well, if they don't extend it past that, they're gonna lose subscribers. But the majority, of Anthropic's business comes from two main things. It comes from coding via the API. Right? And up until May, they had a sizable lead.

Jordan Wilson [00:37:40]:
Right? Fable five, had a sizable lead over five five from GPT. And, you know, Google, we'll see what happens. But here's the thing. I don't think Anthropic is worried about GPT five six Soul. They are worried about GPT five six Terra. Okay. Here's why. So.

Jordan Wilson [00:38:04]:
Anthropic makes money from coding. Yes. They have co work. Yes. That's popular. They make their money from coding. They sell tokens at an extreme premium. All right.

Jordan Wilson [00:38:17]:
Looking at the artificial analysis coding index, and that takes into account a lot of the popular coding benchmarks. You know, Sol, GPT five six Sol is in first place, but guess who's in second place? GPT five six Terra, the medium model. It is better than the big bad fable. The one that we had, you know, have a government anthropic fight, the one that, you know, anthropic had been, you know, saying this thing's too dangerous. They've been they were hyping it up for for three three months, and now we have this, you know, permission slip AI, in The US because of anthropic over hyping, Fable and Mythos. Okay. The medium model is eating fable fives lunch. And why are they eating it? Not just because it's incrementally better at coding.

Jordan Wilson [00:39:12]:
It is more than five times cheaper according to artificial analysis. Five times cheaper to do the same tasks versus Fable five. Yeah. So what I said on the on the Twitter machine, I said, Anthropic has, like, a week or two until enterprises start catching up, because that's how long it takes for even the decision makers that are spending multiple millions of dollars, sometimes a month on these products. And, yes, there's vendor lock in, but any company that's up for, you know, they signed a one year, deal with Anthropic, there's always companies rolling. They're gonna churn. There is zero, zero reason for any company right now to be paying for anthropic. Zero.

Jordan Wilson [00:40:03]:
Because of what OpenAI just did. They undercut on price. They over delivered on performance. And if Fable if Anthropic doesn't change Fable five pricing, it is DOA. It's done. So they've gotta either quickly release Fable 5.1, so they can, you know, knock down the pricing on Fable five, or I do think that they will be bleeding 5 figures potentially daily. Right? So not just from people going over on contract renewals, but just on the API side. I mean, it's it's gonna be interesting to look at these third party sites, like OpenRouter that track, you know, third party token usage.

Jordan Wilson [00:40:44]:
Because I I would assume you're gonna see the OpenAI share go up exponentially, and you're gonna see people that were using at least for Fable, go down. Opus go down because why would you use Opus now when Terra is cheaper? So a big portion of Anthropic's revenue just got cut off at the knees. And if nothing else, this is just fun for us. Right? We get to watch, but that means we're gonna get, number one, a better model from Anthropic sooner rather than later. They have to, and they're gonna have to cut their prices and or, extend and keep Fable five in subscriptions. There's really no other option. Right? Because they can't just keep bleeding money because they're gonna be bleeding money. Right.

Jordan Wilson [00:41:31]:
So that's it. Now you know, chat g b t work. You know, g p d five six soul, what's new, the five overlooked features, and my one hot take. I hope this was helpful. If so, please go to your everydayai.com. Sign up for the free daily newsletter. Thanks. We'll see you, well, not tomorrow, but Monday and every day for more everyday AI.

Jordan Wilson [00:41:50]:
Thanks y'all. Here's the thing about AI at work. You can't use what you don't trust. The all new Slack bot runs inside Slack security boundary only sees what you've already allowed it to see and never trains on your data. Personal AI agent with the trust to actually put it to work. Slackbot from Slack. Learn more at slack.com.

Gain Extra Insights With Our Newsletter

Sign up for our newsletter to get more in-depth content on AI