Resources:
Join the discussion on LinkedIn: Got something to say? Let us know on LinkedIn and network with other AI leaders
Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup
Connect with Jordan Wilson: LinkedIn Profile
Start Here Series in our Inner Circle Community: Join for free access
AI Upgrades: Immediate Business Impact from Desktop Integrations and Open Model Advances
This week in artificial intelligence brought a cascade of tangible updates across the desktop workspace, open-source model ecosystem, and productivity app integrations. Several crucial features now available to business leaders—and soon to be available—provide direct operational advantages for companies navigating today's fast-changing AI environment.
Desktop AI Browser Integration Increases Workflow Efficiency
Anthropic’s recent enhancement to its desktop application introduced an embedded browser, allowing users to interact directly with external websites from inside the Claude desktop environment. This built-in browser facilitates code tasks and documentation reviews without toggling between multiple applications. Key layers of security are implemented through classifiers that review every AI-initiated click and keystroke on third-party websites; additionally, users must explicitly grant website-specific permissions upon first access, offering granular control and a mitigation against unauthorized actions 03:05.
Operationally, this update brings parity between Claude and leading desktop “super apps” in the coding and knowledge worker space, especially as desktop solutions become hubs for daily work. Unlike previous versions, teams can now seamlessly reference bug trackers, dashboards, and API documentation within a single unified workspace. Companies can anticipate increased speed-to-solution, more streamlined onboarding for technical talent, and improved security visibility—particularly vital for development and security teams integrating AI-driven assistance into core processes.
AI-Powered Universal Search Unifies Knowledge Storage
OpenAI’s addition of universal search to its desktop, mobile, and web platforms transforms historic chat log and project analysis. All types of conversations, project threads, images, and uploaded documents are now accessible from a single sidebar search box 06:00, with advanced filtering and immediate access to related projects or files.
This feature mitigates information silos and lost institutional knowledge by making months-old or otherwise buried assets retrievable in seconds. Organizations previously forced to export and manually parse archives for critical content now benefit from easy repurposing of existing work, reducing redundant effort and costly re-analysis. For organizations with significant investments in AI chat-based workflows, this upgrade drives measurable reductions in lost productivity and knowledge attrition across departments, especially for teams managing high chatter and long project histories.
Superhuman and Personalized Email Automation Offers Tailored Communication
The new auto-draft enhancement for Superhuman’s suite leverages the latest models from Anthropic and OpenAI to autonomously draft replies in the user’s tone before the inbox is opened 09:39. These drafts cover both outstanding responses and follow-ups on unanswered emails, and now substantially improve on previous offerings due to more advanced underlying models. Drafts are synced with Gmail and Outlook for business continuity.
This feature addresses one of the most persistent time sinks in knowledge-driven organizations: email follow-up. Business users with high-volume correspondence flows will observe tangible bandwidth expansion, higher responsiveness, and a consistency in brand or personal voice—without the need for plug-in-heavy custom automation. Enterprises can consider this embedded intelligence as a way to preserve executive cycles for decision making rather than manual message crafting.
Voice-Driven AI Assistance Enhances Digital Music Workflows
Spotify’s conversational feature (currently in beta in the US, Ireland, and Sweden for paid users) allows back-and-forth requests by voice or text, including queries into personal listening history, dynamic playlist suggestions, and multi-turn interactions such as “play more upbeat tracks from this artist” 11:54. Unlike general-purpose chat integrations, this system leverages proprietary listening data for personalized user experiences.
For organizations managing brand experiences through digital channels or seeking to personalize offerings in consumer-tech verticals, this showcases a model for leveraging internal user data within conversational interfaces. The beta rollout’s limited initial geography also signals the regulatory and operational considerations that accompany deploying highly contextual, data-leveraged AI outside the US.
Personalized Video Avatars Enter Mainstream Business Communication
Google’s introduction of personal avatars within Gemini and the Google Vids environment allows users to record their likeness and voice, generating branded video communications on-demand 14:52 with a simple prompt referencing their avatar handle. Available for paid users based in the US and aged 18+, this feature does not yet extend to Europe due to biometric privacy restrictions.
Functional uses include internal training, executive messaging, and social content creation where leadership or subject-matter expert availability is limited. No longer must creative teams depend solely on real-time camera access; instead, content can be rapidly generated and iterated, freeing up leadership time while maintaining authenticity and presence. This marks a shift in B2B and B2C media practices, allowing for faster company-wide messaging that maintains a human touch, provided privacy protocols are carefully observed.
Open AI Model Innovation Reshapes Enterprise AI Adoption
Perhaps the most consequential shift is the emergence of Kimmy K3 from Moonshot, a Chinese AI startup. This new large language model (LLM) boasts 2.8 trillion parameters and a native vision mode, ranking it as the largest open model currently built; it supports a 1 million token context window and outperformed many proprietary models on critical benchmarks like web research, automation, and document reading 24:15. The promise of open weights by July 27 highlights a rapidly closing gap between open and closed model performance tiers.
Businesses focused on autonomous agent deployment, long-context research, and those seeking to avoid vendor lock-in will find strategic value in this development—especially as pricing pressure intensifies and frontier models commonly unavailable for self-hosting now enter open status. The ability to self-host an LLM at this scale, with competitive performance to flagship commercial models, increases sovereignty and lowers operational risk across AI-centric enterprises.
Takeaways for Organizations Navigating the New AI Landscape
Each update described above is actionable now—or imminently—for organizations running critical work on AI platforms. The distinguishing characteristic of this week’s news is the convergence of open-source innovation, advanced desktop integration, data-centric personalization, and compliance-aware deployment. Enterprises can expect immediate leverage: faster workflows, information accessibility, tailored communication, and enhanced control over AI-based solutions at both tactical and strategic levels.
The prompt pace and breadth of these changes require leadership teams to maintain close scrutiny of AI product announcements and to adapt integration roadmaps accordingly—capitalizing on cross-platform improvements and open model advancements that directly impact the bottom line.
Topics Covered in This Episode:
- Anthropic Claude Desktop App Browser Upgrade
- OpenAI ChatGPT Work Desktop App Improvements
- ChatGPT Universal Search Feature Launch
- Superhuman Email Auto-Draft with GPT-4
- Spotify AI Voice/Text Conversation Feature
- Gemini Omni Personal Avatar Video Creation
- Google Vids Integration with Personal Avatars
- Moonshot Kimmy K3 Open Source Model Release
- Kimmy K3 vs Fable 5 and GPT-5.6 Benchmarks
Episode Transcript
Jordan Wilson [00:00:16]:
It seemed like this week, we were gonna have a quieter week in terms of big AI releases that you can actually use today. But then late Thursday, we got a new, soon to be open source model that completely reset the proprietary versus open framework. So, yeah, as always some surprises, some new updates and some quality of life improvements. So if you spend most of your week inside of AI tools, the problem is under the hood and sometimes only on a random tweet. These companies release some big updates that change how all of these products work. And unless you spend hours, you're gonna miss those, but that's what we do on our Friday features show. So on today's show, you're gonna learn what open source model is. Again, resetting the open versus proprietary landscape.
Jordan Wilson [00:01:13]:
You're gonna learn how to claw on desktop is getting a lot more useful, and you're gonna see why OpenAI made another big change to its new and very popular chat g b t work desktop app. Alright. Let's get into it. Welcome to Everyday AI. My name is Jordan Wilson, and we do this every day. This is your daily livestream podcast and free daily newsletter, helping business leaders like you and me keep up with the nonstop avalanche of AI updates. I tell you what matters, what doesn't. You use that information to be the smartest person in AI at your company.
Jordan Wilson [00:01:45]:
Everyone's like, wow. How'd you do this? And you're like, well, I went to everydayyoureverydayai.com and sign up for the free daily newsletter. And, well, you listen on the podcast and, maybe on the live stream. So thanks for doing that. And if you are looking for that AI news, make sure to check it out in today's newsletter. Alright. Let's get straight into it and talk about our first new update, which I'm like, bless up. It's about time because I don't know about you, but with all these desktop apps, I've been trying to use them just like, never leaving.
Jordan Wilson [00:02:19]:
So when I go into, as an example, codec slash ched g b t work, I try to not leave. Right? Really change how I do my work. And one of the big downsides with clawed desktop is you really couldn't do that until this week. That's because Anthropic did finally add their new desktop browser. So, yeah, bringing a browser inside of clawed desktop. So here's what's new. Anthropic added a built in browser to the cloud code desktop app, which lets cloud open external websites and actually click, type, and interact inside of them while it works on your code. So Claude can now open whatever you need inside of the Claude app, whether that's your API docs, a bug tracker, dashboard, whatever it may be.
Jordan Wilson [00:03:05]:
So it's not just reading the pages inside of there. It can act on them. Infropic says safety is layered in as classifiers review every click and keystroke Claude makes on external sites. And the first time Claude touches a new site, you get a permissions card to allow or deny it. Every site needs its own approval. So that's one of the things for me. I'm like, that's a downside. But overall, this is a big step forward for anthropic, putting it more in parity, mainly with codecs slash chat g p t work and also cursor, which I've always said were the top two in terms of super app.
Jordan Wilson [00:03:41]:
So now, Infropic Claw Desktop can finally enter the conversation, and we'll see, presumably, we'll have Windows, with their new, super app, Microsoft Copilot entering we think it should be probably by next month. So, who has access? Anyone running the Claude code or, you know, I just call it the Claude desktop app because it's not just Claude code, but the clawed desktop app on Mac and Windows, and you do need a paid plan to take advantage of these features. So an important choice though that they made, the profile uses a completely clean profile. So none of your saved logins, none of your history, etcetera. And that's one of the big updates that actually OpenAI rolled out to their chat g p t work, last week as they unveiled the nontechnical version of, codex called the chat g b t work, well, it can import all of your Chrome history, everything. So a much more robust browser on the codex slash chat g b t work side. However, at least even for me personally, the way that I've changed my work, this was one of the big things that I was just like, I wasn't really touching cloud desktop anymore. At at least not very much.
Jordan Wilson [00:04:58]:
Right? I I still use it for, you know, adversarial, kind of working with codex. Right? I do some more advanced things that I'll I'll use cloud desktop more. Like, I, you know, running g p t five six soul inside of cloud desktop. Conversely, I'm running Fable inside of codex slash chat g p t work, but I really wasn't using cloud desktop as much as I was maybe, like, three or four months ago, mainly because of all the updates that codecs has. But at least for me, this one brings it a little bit more back into the conversation. So who's gonna find this valuable? I think Deb's obviously already working inside cloud code, security teams, right, or if you're a general knowledge worker and you've been doing some things inside of cloud code and you're like, wow. It would be nice to bring, you know, some additional work that I do on the web into the fold. I think that's gonna be a big one for you.
Jordan Wilson [00:05:49]:
Alright. Let's go to our next update. So this one is more of a quality of life one that I'm like, yes. I need this, and I think most people will enjoy it. So OpenAI launched universal search in chat g p t, which lets you search in one place for all of your past chats, projects, images, and documents all from one search box in the sidebar. So you can filter by content type, and clicking a request that opens the chat project or file directly. So, yeah, this just launched a couple of days ago, kind of an under the hood. But, again, this is one that should hopefully bring a big quality of life for the, you know, hundreds of millions of weekly active chat GPT users.
Jordan Wilson [00:06:36]:
The good thing, this rolls out to everyone, whether you have a free pan, a free plan, paid plan, etcetera, and it did already roll out, on the web in iOS and Android. So that's the other good thing. This also works inside of chat GPT. So if you're anything like me, you spend a lot of time regardless of what platform you're using, trying to find old chats, especially on the web. The good thing is, like, ChatGPT work and codecs are great at this because one thread can read every single other thread and direct you. But at least on the web, this has always been something I've struggled with because sometimes chats, you know, these searches will only search normal chats, or they might not search through your project. So it is good to have a single universal chat. Just makes chat GBT much more accessible.
Jordan Wilson [00:07:22]:
So why is it useful? Well, I mean, your chat GBT history just stops becoming a graveyard. Right? That analysis that you ran in March that you put a ton of time in, and maybe you use some of it, but there was a lot of it that you were like, hey. We could benefit from this. Where is it? You know, the image that you did last week, the the contract you uploaded in May. Right? It's all findable in seconds versus having to scroll forever. Right? I would have loved this, like, a year or so ago. There were actually a few instances, both in chat g b t and Claude where I had a very important, couple very important chats that I worked on that I literally couldn't find because sometimes that, you know, searching will literally only search the thread name. So in multiple instances, I've had to go through, export all of my chats, import them in, use a thinking mode just to find, you know, a certain chat that was at least for me very valuable or very useful that I needed to repurpose or reuse.
Jordan Wilson [00:08:20]:
So, this is good. And like I said, this is rolling out to anyone. And who's gonna find this useful? I mean, if you're a power user of chat GBT, this is an instant just quality of life upgrade. Alright. Our next one, not everyone uses superhuman, but, I was actually surprised. So to do you know, to help me plan for this show, I also see, you know, hey. In our newsletter, what are those stories that people are clicking on the most? And I was actually surprised. This is one of our more clicked on stories this past week.
Jordan Wilson [00:08:52]:
So, apparently, a lot of you all use superhuman for your mail. So, well, they launched a helpful new feature if you are a Superhuman mail user. And if you don't know what Superhuman is, well, it's kind of confusing now because it used to just be the, you know, the, I guess, minimal and fast email program, but now superhuman itself, the name is, you know okay. Let me just explain this a little better. So so Grammarly, you know, popular, writing tool that a lot of people use, technically acquired Superhuman, but they applied the Superhuman, name to all of Grammarly's other tools. So it is a little confusing now, but, here's what's new. If you do use Superhuman with their new auto draft so Superhuman launched a new version of auto drafts, which writes replies to your emails in your voice before you even open your inbox. So it does draft automatically two things, responses to messages waiting on you and follow ups for emails other people haven't answered yet.
Jordan Wilson [00:10:02]:
So you open your inbox, and the replies are sitting there for your review and to hit send. So the actual new thing here is the engine running it. So it now runs on Frontier models from Anthropic and OpenAI, replacing an old GPT 3.5 version. Yeah. Apparently, they were still using GPT 3.5, which is maybe why the results weren't that good. So that's why now, the drafts can actually sound like you. So right now, this is for paid plans only on business and enterprise. So you can set it up on the desktop, and it syncs to mobiles, and it also drafts the drafts sync with Gmail and Outlook.
Jordan Wilson [00:10:42]:
So who's gonna find it useful? Well, obviously, if you're a paid superhuman user struggling to keep up with emails, this is a good quality of life upgrade. And, you know, at first, I was almost like, should I even include this? Because this is one of those things that you can very easily do inside of, Claude. You can very easily do inside of chat GPT, codecs chat GPT work. But I'm like, yeah. They're I mean, if you're using these tools, keep using them. Right? But for me, it it is technically very easy to build this inside of, like, CHAD GPT work or codecs, and it can send them all there. So part of me was like, all of this already exists and probably platforms you're already using, but if you all already a power superhuman user, this is for you. Speaking of app specific, updates, this one I'll probably use because I'm a heavy Spotify user.
Jordan Wilson [00:11:37]:
And if you are too, this is one that you might enjoy. So Spotify launched talk to Spotify, which lets you have an actual conversation with the app by voice or text to control your music, discover new stuff, and also ask about your own listening history. So the cool thing here is, well, it's back and forth, not just one shot commands. So you can ask for an artist, then keep steering, and you can say, just hit his recent stuff or make it more upbeat, and then you can save and queue and follow right on from the conversation. So who has access right now? It is in beta. Alright? But it's rolling out to people in The US, Ireland, and Sweden on iOS and Android. So, yeah, you do have to be a paid, user in one of those countries like I am. I haven't actually checked this one yet because this one just came out, but I'll be using this one.
Jordan Wilson [00:12:30]:
So, why is it useful? Well, the differentiator is your data. So Spotify knows your playlist, your favorite artists, and your repeat listens. So it answers things like, when did I first listen to this song? And, you know, no general chatbot, at least right now, can do that. Right? Although, Spotify does have some integrations with, like, Chat GPT as an example, it doesn't have that level of granular data. So this one for me and the reason why I think it's very helpful. If you have a paid Spotify account and multiple people in your family use it, maybe if you have kids, you have, you know, people with wildly different taste in music, but everyone's using the same Spotify account. Right? Like, if you have, you know, an Alexa or, you know you know, home speaker, like, that's what I do, and everyone's listening to all different kinds of music on my Spotify account. So, you know, Spotify used to have, you know, and they still do.
Jordan Wilson [00:13:21]:
It's it's these weekly playlist that update based on your history. So I used to love these playlists. They would update every Friday. I would literally used to open it every Friday and be like, oh my gosh. I can't wait for my you know, the two playlists are called, I think, like, discover weekly and release radar based on, like, your, listening history. And, you know, so recently, mine is just all off. So this is one I'm gonna go in there and try to kind of, you know, discover new music based on the genres that I actually listen to. So this one, obviously, maybe a little bit more on the personal side, but I think it's a pretty big update, especially if you are a power Spotify user like I am.
Jordan Wilson [00:14:01]:
Alright. Our next one. In this one, if you are missing Sora from OpenAI, right, remember, about five five or so months ago, you know, OpenAI kind of announced that they were killing off all their side quests and really focusing, you know, on their core products, which in theory has been great, and the company is growing just exponentially. But one of the downsides is there was these fantastic products that are now just shelfware, and one of those was Sora. And one of the most popular features of Sora was being, able to essentially import your own self, your own avatar. Right? You would have to verify your face, and then you could create videos with yourself in them. Well, Gemini Omni has finally rolled that out. So this is technically a new feature in two different parts.
Jordan Wilson [00:14:52]:
So first, Google launched personal avatars inside of Gemini, which lets you record your face and voice, then generate AI videos of yourself on demand. So you just drop your avatar into prompts by typing the plus at sign, and then your username. So as an example, you would type, you know, create a video of at me, singing with an orchestra, and then your likeness becomes a reusable asset inside of Gemini. So right now, you do have to have a paid account and be in The US and 18 or older. So, yeah, all the, you know, EU, UK, not available right now. So, which is all you you know, people are always like, why doesn't, you know, the EU get this? Well, the European exclusion is almost always about, you know, regulatory caution around biometric likeness. Right? That's why so many features, especially ones like this that are hyper personalized, for your own likeness or others, you know, don't usually roll out sometimes at all to the EU. So and and so it's not only inside of Gemini, but it's also, inside of Google Vids.
Jordan Wilson [00:16:03]:
So here is what they said about using this new feature inside of Google Vids. They said today, we're rolling out two new updates to Google Vids to make it easier than ever to create, edit, and personalize your videos. Gemini Omni and personal avatars. Gemini Omni takes the hard work out of the editing process, so generating and refining high quality clips is as easy as writing a simple prompt. And with personal avatars, you have an entirely new way to star in videos without setting up a camera. So, yeah, if you missed Google's big Gemini, Gemini Omni announcement back in May, essentially, it is the new family of VO models, but just much more powerful. Where their VO, you know, so v o three, v o 3.1, were technically just video models. Right? Gemini Omni is much more than that.
Jordan Wilson [00:16:54]:
It produces video, but it is like a world model. And I think it's much better at editing scenes, than any other platform out there. So this is a pretty big one. I think, you you know, who has you know, why is it useful? Because this is like the talking, you know, so many talking head videos that just need, like, background b roll. Right? So if you're someone in your company that's whether it's for internal or external purposes, Right? I wouldn't use these to actually do, like, a full avatar, like, talking head video. That's not the thing. But if you or someone else well, like me. Right? I'm constantly talking head video right now.
Jordan Wilson [00:17:36]:
I am a talking head video. You know, I don't necessarily not part of my brand or the everyday AI brand necessarily, but when I'm yapping about all these things, I could very well instead put an avatar of me working on these things. So you don't just have to, you know, look at my, face made for radio and just be like, alright, guy. You know, put something else on the screen. So but I think there's a ton of great use cases for this. So if you are in l and d, you know, training, marketing, content creation, show social media, anything like this. And if you have a CEO as an example, that's hard to nail down and you are trying to get more video content out there or something like this. Right? Again, if you go through the proper channels, do all the approvals, data security, privacy, all that good stuff.
Jordan Wilson [00:18:24]:
But now this is such a a a weapon to have in your creative arsenal. Right? I would have loved to have this, like, eighteen years ago at one of my first jobs where I was creating a lot of video, and usually it was just kind of talking head of the CEO. And I would try to get some of this more, like, b roll stuff, and it was just sometimes hard. So I think if you are a creative, if you are trying to have a a stronger, you know, presence, on social media for your company, but it's a little hard, This is great. So, yeah, content creators, marketers, educators, you you know, is is gonna be huge. But I think the simple framing of this is it's kind of like a personal version of HeyGen inside of Gemini. Right? So avatar video just went from being kind of a specialist tool, to now going mainstream. Alright.
Jordan Wilson [00:19:15]:
Speaking of mainstream, let's tackle our next update because this one is very mainstream. So, the new chat g p t work desktop app has a big update that I think a lot of chat g p t power users are going to enjoy. Alright. So, I'm just gonna go ahead and read, what Thibault, the head of product at OpenAI, posted. It's a little easier just to read it, and this is brand new. I've been playing around with it a little bit today, but, yeah, it hasn't even been out a full day yet. So, Thibault said, evening, we've got we've gotten lots of great feedback on the new chat g p t desktop app. So the work app, which we didn't get totally right on the first try.
Jordan Wilson [00:20:00]:
And as a result, we made some changes. Number one, chat gbt conversation history and projects are now visible in the sidebar. Also, your chat and work history now sync across web, mobile, and desktop. Local task will stay on your computer. Then you can now easily switch between chat and work modes inside chat g b t on desktop, which is now also consistent with how it shows up on web and mobile. Nothing is changing for users on codex mode. Tivo says it's still the OG and best at what it does. So what does this mean? Well, to actually say what it means and, you know, Tivo did kind of say it there.
Jordan Wilson [00:20:43]:
He said, yeah. We made some we didn't totally get it right on the first try, and that kind of explains the update. So long story short, if you missed this, I just did, an episode on this on Wednesday where we went over, chat g b t work, what it is, all that good stuff. But ChatGPT work is essentially the nontechnical version of codecs. Right? OpenAI's, autonomous desktop agent. But there is a ChatGPT work on the web, and there is a ChatGPT work on the desktop. So the problem was it the chat g p t work didn't exactly sync up very well with what you were doing on chat g p t on the web. It was kind of this separate pop up that came up in the right hand corner.
Jordan Wilson [00:21:33]:
So your chats were kind of there, but it just wasn't really intuitive to use because you would have all of your, essentially, your tasks in your projects that lived or started with chat g p t work slash codex on the left sidebar. And then your chat g p t history was kind of this orphan page on the right hand side that wasn't really attached to anything. And then the big thing that people were like is, hey saying, like, hey. I love chat g p t, but I run everything inside of projects, and projects did not sync. So, essentially, OpenAI changed and technically fixed all of that because now not only do your projects sync to the desktop version of chat g p t work, right, but they're also in the left hand sidebar. And the other good thing is from an aesthetic standpoint, now it does look the exact same as it does on the web, which I think is gonna help people, and it looks the exact same as it did on mobile. So the first variation of this, it looked very similar, on the web and on the mobile app. But then when you open the CHAD GPT work app, it looked and functioned completely different.
Jordan Wilson [00:22:38]:
So this is a great update, for people who are trying out, ChattGPT work or if you're like me. Right? I've been using codecs since day one. And the good thing is, well, now you have that ChattGPT work experience, which is just a very similar version of codex, and then you have your familiarity with all of your, normal chat g p t chats and your, projects. So if you're looking at this, if if you scroll down, it's all gonna be under your recents. So whether you, start a new task inside of chatgpt work that's not attached to a project, it will go to your recents, and that's also where all of your new chats inside of chatgpt will live. So if you do start a new chat inside of chatgpt.com, you don't attach it to a project, and then you go to chatgpt work on the desktop, it will be there, which is great in the same place under recent. Alright. Our last piece of AI news, and this is technically the biggest one.
Jordan Wilson [00:23:37]:
So not just a new heavyweight on the soon to be open source, and I'll explain that. But this one is definitely resetting the open versus closed model paradigm. Yeah. We all thought it was the, GLM 5.2 that was gonna do this. Not anymore. Get ready because you're gonna be hearing a lot, especially if you are an AI large language model dork like I am. A lot of the conversation, I'm guessing, for the next month or two is gonna be set around Kimi k three, and this has huge implications. So let me first explain what it is, what's new.
Jordan Wilson [00:24:15]:
So, Chinese AI startup, moonshot, released Kimmy k three, a 2,800,000,000,000 parameter model that is now the largest open model ever built with a native vision mode and 1,000,000 token context window. So here's the thing. It came in at third place overall on the artificial analysis intelligence index, not far behind quad fable five and GPT 5.6. So, that alone is mind boggling. Right? Because we've been seeing this race just go back and forth, back and forth. And we thought that when anthropic released mythos and fable five, that this was essentially a new category, a new tier that no one else was ever going to touch. So not only, you know, about a month later did GPT 5.6 enter that category and on many of the most important benchmarks past Fable five. But now we have Kimmy k three, a soon to be open model that has not only entered the conversation, but it is actually surpassed.
Jordan Wilson [00:25:27]:
Yes. An open model has surpassed. Well, both Fable and GBD 5.6 on many important benchmarks. So, the best open model in the world now sits at three behind the two closed flagships. And the important thing that is thrusting this all back into the conversation is, well, Entropic is supposed to be pulling Fable five from from subscriptions this Sunday. So not only do they have this continued pressure, from GBT 5.6, But now there is an open model that you is gonna be better. Even if you're paying $200 a month like I am, you're not gonna get access to Fable five. There are obviously rumors that maybe today or maybe, early next week, we might be seeing Opus five, and maybe that will even reset and be much closer to Fable level than currently Opus 4.8.
Jordan Wilson [00:26:24]:
Alright. Anyways, let's talk a little bit more about Kimmy k three who has access all that stuff. One important thing to note though, right now, it's not technically an open source model, but it will be soon. Moonshot did say that they will be releasing the weights, so it will be an open weights model. They just haven't released the weights yet, but it just came out literally hours ago. And everything else that moonshot has released in the past has been opened. So this is live today in the Kimi app on kimi.com. That's k I m I, and also in the Kimi work desktop app as well as through the API.
Jordan Wilson [00:27:01]:
So they did say that they will release the full model weights by July 27. So that's within, like, a week and a half. So at that time, anyone can download and self host this. Obviously, it is a large 2,800,000,000,000 parameter model. So, if if you think you're gonna host this on normal consumer hardware, no. Unless you've spent, like, 20 or $30,000 on your setup. So this is more for enterprise companies. Well, if you have the, the GPUs, you can do this.
Jordan Wilson [00:27:33]:
So who's gonna find this valuable? Actually, no. Let's first talk about why it's useful. Well, for the benchmarks alone, what we're seeing is its specialty is in long autonomous work. Right. So they kind of shared two different case studies where k three designed and verified a working chip in a single forty eight hour unsupervised run, and it reproduced in astrophysics sorry, astrophysics research result in about two hours that they say would have normally taken a researcher one to two weeks. And then the benchmarks obviously shows that it beats many models tested, including Fable five on web research, automation, spreadsheets, document reading, and more. And the crazy thing to me is on arena, which goes head to head, blind taste test. One of the things that anthropic has really kind of dominated this space is front end design.
Jordan Wilson [00:28:27]:
Right? So, on the design arena, GPT five six, surpassed Anthropic's models, including Fable five. And then on the LM arena, which is kind of like user preference, this model is now number one, Kimmy k three. So it's interesting because that has always been one of, you know, in Propix kind of niches that they've owned, like front end design. And now on the two most important front end design benchmarks, they are no longer number one in either. So I'm personally excited to see what anthropic is gonna cook up. Maybe we'll see this with the open five drop, but I would assume that their next model, maybe they've let their foot off the gas in terms of front end design because they've been so far ahead of everyone for so long. So I do and would assume that Opus five is probably gonna surpass Fable five at least in those areas because that's something I'm sure that anthropic is gonna be feeling a lot of pressure on aside from, you know, pulling Fable five. We'll see if they extend it again.
Jordan Wilson [00:29:30]:
But on the subscription package, between GBT 5.6 and now Kimmy k three, yeah, they're gonna have to maybe justify why people are gonna be subscribing, but a lot of people are pointing to that just means we're getting an Opus five here soon. So who's gonna find this new Kimmy k three valuable? Enterprises and developers who want frontier adjacent agents without that frontier pricing or vendor lock in. So, I mean, the crazy thing is, I mean, this is literally now an open soon to be open model from a now from a Chinese, you know, startup that now benches ahead of Opus 4.8, and it is in the same tier, although technically lower than Fable five and GPT five six Soul. Alright. So an interesting one here, resetting the open versus closed race. And you know what it means for all of us y'all? We're gonna continue to get more and more models, better models, and hopefully at cheaper prices now that we have an open model, presumably pushing the cost down. So this one here with Kimmy k three, definitely more of an enterprise play. Like I said, this is not really for inner, for consumer hardware, although the prices are also much cheaper via the API.
Jordan Wilson [00:30:45]:
Alright. So that's a wrap. A lot new that we went over today. I hope this one was helpful. Like I said, we do this every single Friday, our Friday feature show where we bring you the latest and the greatest of what's new that you can actually use today. I was actually kinda bummed because we also had from Google, Gemini notebook that came out, but it wasn't available to all paid users. So that's the thing. We do this if you have a base paid account.
Jordan Wilson [00:31:09]:
These are the things that you can use today to grow your company and career. Alright. If this was helpful, do me a favor. Please subscribe on the podcast and go to your everydayai.com. Sign up for the free daily newsletter. Thanks for tuning in. We'll see you back Monday and every day after that for more everyday AI. Thanks, y'all.
