Ep 753: Anthropic Goes Full OpenClaw, Meta Muse Spark Drops, Google Gets Notebooks and More . 7 New AI Features You can’t afford To Miss

Resources:

Join the discussion on LinkedIn: Got something to say? Let us know on LinkedIn and network with other AI leaders


Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup

Connect with Jordan Wilson: LinkedIn Profile

Start Here Series in our Inner Circle Community: Join for free access


7 Essential AI Features for Businesses in June 2024: Anthropic, Meta, Zapier, Google, Microsoft, and More

AI technology is progressing rapidly, and keeping pace means understanding exactly which features deserve attention—especially as they are released. Recent upgrades across Anthropic, Meta, Zapier, Google, Microsoft, and OpenClaw are not just incremental, but have introduced specific capabilities that business owners and decision-makers can act on now. This detailed overview translates those advancements from technical jargon into actionable business insights, highlighting practical opportunities with clear value.

Zapier SDK: Direct AI Agent Integration for Workflow Automation

The Zapier SDK now offers coding agents—such as CursorCloud, Claude Code, and OpenAI's Codex—programmatic access to Zapier’s ecosystem: 9,000 apps, 30,000 actions, and raw API connections to 3,000 apps. In practical terms, this eliminates manual OAuth flows and token management, enabling streamlined authentication across platforms.

Business Value:
Automation experts and marketing teams can centralize workflows spanning Gmail, CRM, Slack, WordPress, and email newsletters, using natural language to link these operations. Developers building AI agents gain efficient access to Zapier actions, significantly reducing friction and accelerating setup, even for non-technical business leaders. Existing Zapier connections work immediately upon authentication, so companies can instantly leverage past integrations.

OpenClaw Updates: Built-In Creative AI for Video, Music, and Persistent Memory

OpenClaw has introduced built-in video and music generation, memory upgrades, and ‘dreaming’ to consolidate agent knowledge over idle time. Video/music tasks utilize APIs such as Runway and Google’s Lyria, enabling video/music generation and editing (trim, mix, splice) directly through natural language commands within the agent interface.

Business Value:
Content creators and marketing teams can quickly produce or edit media assets for websites, presentations, and training materials. Persistent memory, now structured like a Wiki, supports Obsidian export and solves long-standing recall challenges for workflow continuity. For developers and power users acting as creative assistants, these features allow for rapid prototyping or campaign iteration without external tools.

Google Gemini Notebook Integration: Organized AI Research and Project Sync

Gemini's new notebook feature, integrated with Notebook LM, allows conversations, files, and custom instructions to be organized under project-specific folders. Notebooks sync between Gemini and Notebook LM, creating a persistent space for context, files (PDFs, documents), and ongoing instructions.

Business Value:
Research, education, and knowledge work are streamlined because all chats and source materials can be centralized and grounded via Notebook LM—reducing hallucination rates to near zero. The seamless movement between Gemini and Notebook LM ensures project-related data is always organized, searchable, and actionable, making this valuable for teams managing long-term research or development workflows.

Meta Muse Spark Model: Free Multimodal AI for Text, Image, and Tool Use

Muse Spark, Meta’s new model from its Superintelligence Lab, presents a natively multimodal architecture with text, image, and tool use, featuring parallel sub-agents for reduced latency. The model is free to use via meta.ai and Meta's apps; API access is currently limited to select partners.

Business Value:
Digital product teams using Facebook, Instagram, or WhatsApp can access advanced multimodal AI without cost, enabling internal assistants, coding support, and creative ideation. While benchmarks show Muse Spark as impressive for a first-generation release, it trails Gemini, Opus, and ChatGPT’s 5-series models, meaning enterprise use will depend on matching specific technical requirements and integrations.

Microsoft MAI Models: Fast Transcription and Voice Generation at Competitive Pricing

Microsoft launched MAI Transcribe One, Voice One, and Image Two, all accessible in Microsoft Foundry and MAI Playground. Transcribe One beats Whisper, Gemini Flash, and 11 Labs Scribe with a 3.8% word error rate across 25 languages; Voice One renders 60 seconds of natural audio in under one second via a single GPU.

Business Value:
Teams needing high-volume, multilingual transcription or real-time voice agents (customer support, sales, product demos) can achieve more precise, faster, and cost-effective results. MAI Image Two ranks top three for image generation in Arena AI leaderboards, providing marketing and branding teams with quality visuals from a Microsoft-centric environment.

Google Vids: Free AI Video Generation for Marketing and Training

Google Vids now integrates VO 3.1 for free AI-powered video creation, offering up to 10 eight-second clips per month for any Google account. Paid features include AI avatars, custom music via Lyria 3, and expanded clip limits on the Ultra plan.

Business Value:
Small and mid-sized businesses can leverage Google Vids to rapidly modernize websites, corporate training, or marketing campaigns with visually rich content. Even without a paid subscription, companies gain access to capabilities previously behind enterprise paywalls.

Anthropic Managed Agents: Simplified AI Agent Deployment for Enterprise

Anthropic’s Claude Managed Agents platform, available in public beta, abstracts hosting, scaling, monitoring, and error recovery for AI agents. Agents are defined in natural language or YAML, and session orchestration manages tool integrations, context management, and credential handling—all via a guided setup with authentication support.

Business Value:
Enterprise engineering teams can deploy Claude-powered agents without months of infrastructure setup. Permission systems, state management, and sandboxing are handled by Anthropic, enabling 10x faster time to production compared to manual agent development. Project managers and non-developers gain simplified agent deployment while maintaining enterprise-grade oversight.

Conclusion: Key AI Capabilities Ready for Business Deployment

The most recent AI feature releases across Zapier, OpenClaw, Google, Meta, Microsoft, and Anthropic offer specific, immediate business utility. Automation, creative production, research organization, transcription, and agent deployment have all become easier, faster, and more integrated—often requiring less technical overhead and offering broader access. These advancements are not abstract; they translate to real-world benefits that operational leaders can activate now to drive efficiency, clarity, and innovation in their organizations.


Topics Covered in This Episode:

  1. Zapier SDK Launch: 9,000 Apps Integration
  2. Anthropic Managed Agents: Automated Scaling
  3. OpenClaw Built-in Video & Music Generation
  4. OpenClaw Persistent Memory Wiki Features
  5. Google Gemini New Notebook LM Sync
  6. Google Vids Free AI Video Generation
  7. Meta Muse Spark Model Multimodal Release
  8. Microsoft MAI Transcribe One Benchmark Results




Episode Transcript 



Jordan Wilson [00:00:16]:
Anthropic continues to add new features that look and function kind of like OpenClaw, the same platform they pretty much shut out last week. Google just dropped a notebook feature that's kind of like notebook l m, but kind of not at all, but it also adds new features to both platforms. We got new models from Meta, Microsoft. We got AI video updates and a ton more. But the most meaningful updated AI feature that you may have missed this week comes from a company we haven't talked about in a while, Zapier. And I think that one will be big. Alright. Let's get into it.

Jordan Wilson [00:00:57]:
If you're new here, welcome to Everyday AI. But on today's show, if you stick with me for the next twenty ish minutes, here's what you're going to learn on our new Friday features segment. You're gonna learn how Anthropic is going full on OpenClaw after banning its models from being used by OpenClaw. If Meta's new Muse Spark model is worth the billions they've invested in, You'll learn what Google's new notebook feature in Gemini unlocks. And last but not least, you'll see why I think Zapier's new SDK might be your favorite tool even if you're not technical. Alright. Welcome to everyday AI. My name is Jordan Wilson.

Jordan Wilson [00:01:39]:
If you're new here, we do this every day. This is your unedited, unscripted guide to keeping up and getting ahead with AI. So if you haven't already, please make sure to go to youreverydayai.com. Sign up for the free daily newsletter. We're gonna be recapping today's show and all the other AI news that you need to know to stay ahead. Alright. So there's so much that's happening in the world of AI. Y'all I have a daily AI podcast from Monday to Friday, at least.

Jordan Wilson [00:02:10]:
And I realized, after like 700 some episodes that still 80% of all the actual useful stuff that we were all using, I wasn't covering. So that's what this new Friday features is all about. So on Mondays, we do the AI news that matters. On Wednesdays, we go hands on, you know, with a new model or release. Last week, we did the, or this week. Wow. 10 flies. This week, we did Google, Gemma four, the local model, so make sure to go check that out.

Jordan Wilson [00:02:43]:
And then on Fridays, well, we're gonna fill in the gap in between because I realized that in between the Monday AI news and the Wednesday deep dive, most of the things that we talk about aren't getting covered. And then on Tuesday and Thursday, we do different shows, rotate, sometimes interviews. Alright. But let's get into it. There's a ton to cover this week. So, livestream audience bringing up my, my screen here. Podcast audience, you can always see the video version at youreverydayai.com. Go click on episodes, not only today's, but literally, you know, 750 plus, videos, podcasts.

Jordan Wilson [00:03:19]:
You can read it. It's all in the newsletter for free. It's a free generative AI university. Alright. First, let's start with the one that I think is I'm the one that was this is honestly probably the one I was most excited for, even bigger than the anthropic agents one that we're gonna get to in a little bit. But first, let's talk about the Zapier SDK. So here's the way that Zapier explain it. They say, let your agent connect to anything authenticated, govern access to the full Zapier catalog in code on behalf of your users.

Jordan Wilson [00:03:52]:
No OAuth flows. No token management. Zap Zapier handles the keys. So here's what that means. Zapier opened its SDK to everyone, including giving coding agents like CursorCloud in, Cloud Code in codecs programmatic access to Zapier's full ecosystem of 9,000 apps, 30,000 actions, and raw API access to 3,000 apps. So here's the good thing. If that if you heard that and you're like, well, what the heck does that even mean? You know, OAuths and raw API, you're like, wait. What? This isn't a show for nontechnical business leaders.

Jordan Wilson [00:04:34]:
Yes. It is. So even if you don't know what most of that means, well, you can just use natural language to set it all up. So you will need to use a coding agent to obviously tap into this SDK. So that's just a software development kit, which used to be kind of technical, but in the age of AI and large language models and natural language processing, it's not actually very hard at all. But what this means is, let's say, as an example, if you use Claude Code or if you use, OpenAI's codex. Right? Up until well, before this Zapier, kind of announcement, what you could connect to was extremely limited. Right? Obviously, Anthropics' cloud code has a growing list of MCPs.

Jordan Wilson [00:05:20]:
With codex, they do have some plugins, not a ton. But now essentially, you have Zapier, so you have everything. Right? So your agent or different agents that you use can now access your everything from your Gmail to your, CRM, to Slack, to your project management tool, like whatever you're working on. So, I haven't started to set this one up yet because it literally just came out like, like, forty eight hours ago. But this one, I think, is pretty big. So who has access right now? Actually, no. I will point out. I watched the video, with Wade Foster, their CEO, and he called it the most powerful thing that they've launched in years, which I think was pretty telling.

Jordan Wilson [00:06:03]:
Anyways, so this is an open beta right now, and it's free during early access with no billing charges. So, that's pretty cool. And then enterprise and teams plans, are off by default, and they do require manual opt in, or contacting Zapier. So if you have a personal plan, this will be a little bit easier, to enact right away. So here's why it's useful. It just eliminates the single biggest friction point for AI agents, and that's managing all the authentication, right across, you know, thousands of apps. You're probably not gonna be using this across thousands of apps, but you'll be using it maybe across a dozen or dozens of apps. Right? It's difficult to, manage all of that, but now Zapier just kind of does that.

Jordan Wilson [00:06:46]:
And that's where, Zapier's MCP gives agents a curated menu of prebuilt actions, and then the SDK lets agents write loops, handle edge cases, and chain complex logic across apps. And your existing the cool thing is your existing Zapier connections work immediately, so you don't have to re auth anything. So, once you do authenticate, the Zapier SDK in your coding agent of choice, anything that you've previously connected, it works right away. Right? So I'm like, even as I'm saying this, I'm like, oh, freak. I should have prioritized this over other things because, you know, with Zapier, you know, I have our WordPress website set in there. I have our Beehive email newsletter. I have our our circle community. Right? It's all there.

Jordan Wilson [00:07:32]:
So now I can instantly, start working with those things inside of Claude Code and inside of codex are the two that I use the most. So here's who I think is gonna find it valuable. Well, developers building AI agents, that's gonna be extremely valuable, you know, automation experts. So if you are someone or, you know, anyone in marketing that has used, Zapier for years, you're gonna find this extremely helpful. But I also think solopreneurs, entrepreneurs who maybe don't have the time, to usually do all of these things, well, now it made it a lot easier, with this from Zapier. Alright. Let's go to our next one. Doing some a little different.

Jordan Wilson [00:08:12]:
Open claw updates. Right? And, hey, let me know. Livestream audience, podcasts, you know, if you're listening on Spotify. If I should make the Friday feature, right, if we're gonna do this permanently, should I just always like, we've been doing essentially seven things. I think that's easy. Right? It's not too many. It's enough. But should one of them be OpenClaw every week? If so, right, because OpenClaw ships updates, like, every day.

Jordan Wilson [00:08:39]:
So let me know. Just say OpenClaw in the comments if, you think. I have a number. Right? I actually do this. I always have a number, and I say, between LinkedIn and Spotify, if we get enough comments on this, we'll do it. If not, we won't. So, if you want to always include, the weekly open claw update, let me know. But I will do it this week regardless because I think it's pretty big.

Jordan Wilson [00:09:01]:
Because now in open claw, we got to built in video generation, built in music generation. We got some new memory features and an experimental kind of dreaming mode. Alright. So well, if you don't know Open Claw at all, I don't know. You must literally be a lobster sleeping under, the sand somewhere. So Open Claw is the world's technically, the most popular piece of open source software ever. It is an autonomous agent that acts on your behalf. Right? You can talk with it, communicate with it via various channels, but let's get into the actual new update.

Jordan Wilson [00:09:39]:
So, we got three major releases. So one is built in video and music generation. So these are obviously all things that you pay for. Right? You have to, connect, you know, your API keys unless you're using an an open model. So if you do have a, you know, super fat like Mac Studio like I do, you'll be able to do some of these things for free or your agent will be able to do them for free. But otherwise, you will be using API, so you'll be paying for your usage. Keep that in mind. So the new things, we got the, built in video and music generation just from text.

Jordan Wilson [00:10:15]:
Right? Video music and editing tools, and then the persistent memory, kind of like a Wiki knowledge system. So the video generation, the agent can create videos and music tracks directly using configured providers like Runway, Google's Lyria, etcetera, without even leaving the chat. Then music and video editing, agents can trim, mix, splice, and refine existing video. That's really cool. And music files through natural language commands. So, that's pretty sweet. Right? Turning Open Claw into a creative editing assistant. You know, this this is the first versions.

Jordan Wilson [00:10:51]:
Right? I'm not gonna try that one, you know, until it's gone through some updates. You know, some of these things also keep this in mind. The Open Cloud team is amazing. Right? There's a reason why it's literally the, most popular piece of open source software software ever. But for some of these things, when they first are released, they're usually rough around the edges. That's how it goes. Right? And then the open source community, you know, finds fixes, you you know, they ship them out pretty fast. But, you know, if if if you need something like this for production for, like, a work project, don't think that you're gonna get anything, you know, usable or great, but it is important to know what they're working on and what they're shipping.

Jordan Wilson [00:11:27]:
Alright. So that is the, the music and the, video editing, and then there's the memory wiki. So, this kind of replaces the old fuzzy recall system with a more structured, kind of persistent knowledge base that works a little bit more, like Wikipedia. And it does support Obsidian compatible export. So if you do use the Obsidian app, which I know a lot of people do, that will be helpful as well. And then last but not least, there's a new dreaming, experimental mode. So this is opt in, and this is essentially now where the agent processes conversations during idle time, through light deep in REM style, phases to consolidate short term memories into permanent knowledge. So pretty pretty cool.

Jordan Wilson [00:12:12]:
Then it kind of has a diary timeline user interface. So, obviously, who has access to these? Well, everyone. Right? So as long as you're, updated your Open Call to the latest version, you will have all of these, features. And, the cool thing is obviously, they, you know, added g b t five four support. Recently, Gemini clawed only via the API because obviously, anthropic, cut off access, to using your, Claude pay paid plan, via open claw, which we talked about on the show last week. So, I mean, here's why it's useful and who will find it valuable. Well, I don't think some of these things are gonna be super useful now, but they will be very useful soon. So I think specifically, if you work in, anything creative, right, the fact that you now have, video and music generation is pretty big, you know, runway support, Suno, you know, comfy UI, whatever you use, and then the same thing with, Google's, Lyria.

Jordan Wilson [00:13:16]:
So if you work in any creative industry, I think it's gonna be good to just start, you know, toying around with this. I'm sure it's going to get much much better, But then the memory in the, the memory the memory Wiki solves a lot of, you know, the common problems that people have been having with OpenClaw. Even if you have your, you know, your soul MB file set up, your heartbeat set up everything correctly, still the memory piece is pretty tricky. So the new memory wiki, is the next attempt to solve that. So who's gonna find this valuable? Like I said, content creators, great. If you just want, you know, quick and dirty, you you know, videos, developers building AI agent workflows, obviously, and then power users and tinkerers, who are kind of using OpenClaw as a primary assistant. Right? And you wanted to remember more, the new memory features are their attempt to do just that. All right.

Jordan Wilson [00:14:10]:
Next, speaking of remembering things, we have a new notebook integration in Gemini. So it's technically a notebook LM integration, but not the same as the one that they added at the 2025. Right. So let me explain, like, what that means. So what they added at the 2025 is you could add your notebook l m, kind of notebooks as a source inside of Gemini. So this is a little different. The new notebooks feature is more of kind of like a folder, but not the folders that you would normally use inside of like chat GPT or, Claude. These are more like a notebook LM folder, but that you can use in Gemini and notebook l m, and then it syncs.

Jordan Wilson [00:15:03]:
Alright. Let me kind of read the, official word here from Google. So it says notebooks in Gemini help you organize your chats and projects in the Gemini app with our air with our AI powered research partner, notebook l m, for easier learning and working. Alright. So here's, kind of the the gist of what it does. So it says with notebooks, you can keep conversations about a topic organized in one place. Just click new notebook on the side panel and of the Gemini app to get started. You can move past chats into notebooks.

Jordan Wilson [00:15:34]:
So that's that is good. You can organize your past chats into notebooks, and then give Gemini custom instructions and add relevant files like documents and PDFs to give Gemini more context. So in that regard, it is like standard, you know, chat GPT projects and clogged projects, but here's where it's not at all. Because all of those notebooks technically sync to notebook l m, which is really cool. I probably, which is crazy to think, outside of Google Gemini's deeper research, because Gemini is everywhere now. Right? I use AI studio a ton and I use notebook LM a ridiculous amount. It might be one of my most used tools, aside from, like, g p t five four pro. And I am using, like, most people clogged more and more, you know, with all these new features that come out.

Jordan Wilson [00:16:29]:
But I actually am not using the I'm not using the normal Gemini app to chat as much as I used to be. Right? I'm using a lot more for deep research. I think Google's deep research is still goated. It's so good. But now with this, I'm like, okay. Now I'm gonna have to start doing more of my chats, inside, Gemini. Right? Because this feature to sync all of those, put all of those in a notebook, and have that be synced to notebook LM. Right? You have all of these sources that you bring in a notebook LM.

Jordan Wilson [00:17:00]:
If you're a power user or maybe if you don't know, notebook LM is a very unique product from Google. It is powered by Gemini, but it's grounded. So that means the hallucination rate is essentially zero. Right? Where all other large language models, you you know, you're constantly having to be vigilant against, hallucinations. Right? So this is so cool now that your chats inside Gemini can become your source material for notebook LM just by adding it to a notebook. So that's really cool. So like I said, you do have that persistent project space where you can organize chats, upload documents, and PDFs, and give Gemini custom instructions for those ongoing products, and the notebook sync automatically, which is really cool. So if you add a file, you know, in notebook l m, that file will be added in Gemini, which is so so cool.

Jordan Wilson [00:17:56]:
I I I'm pretty sure, I I messaged one of the, notebook l m leads of something like this, like, a a year ago. They did add one of my features that I requested, finally, getting on the on the mobile app, you know, plus and minus, you know, plus and minus ten seconds. I asked for 15. They did 10, but, you know, can't complain. So who has access? So right now, this is rolling out to paid subscribers first. Alright? So if you're on a free plan, you know, maybe this will roll out eventually. I'm guessing it probably well, I could be wrong. So I'm guessing this is going to be a more of a compute intensive feature, especially if it's popular.

Jordan Wilson [00:18:39]:
So it is rolling out to, paid users on the web this week, and then mobile access and European expansion in the following weeks. And free users, said they'll get, access after all the paid rollout completes. So no clue when that is because sometimes it does take a little bit longer, for all of the paid customers in Europe to get access. So sometimes those, you know, free users getting access gets delayed a little bit. So here's why is this useful, and I think who will find it valuable. First of all, literally anyone is gonna find this valuable. Right? I don't care if you're a researcher, student, general knowledge worker. If you're not using notebook LM literally every single day, I don't understand why not.

Jordan Wilson [00:19:23]:
Unless you're not able to because of your job, but I mean, use it for your personal life. If you have a personal Gmail account, even the free version of notebook LM is can't miss. It's so good. Right. But here's why I think that this new updates useful. It solves that scattered conversations problem, right? Obviously, if you've used, projects in chat GBT or if you've used projects in Claude, especially if you've used, like, project memory inside of chat GBT, you'll understand how useful it is, to essentially that way you're not wasting all of that back and forth. Right? So I think sometimes when we're working with large language models, you know, ultimately, we just use whatever that last piece is. Right? It's the last thing that you copy and paste.

Jordan Wilson [00:20:08]:
Right? Or maybe you're going through a lot of iterations of something just to get a couple of facts. But what about the other 70%? The other 80%? The other 90% could all be really good stuff. So now this lets you automatically use that inside of notebook l m. That's probably the way I'll be taking advantage of it most is, you know, moving from Gemini to notebook l m, not vice versa necessarily because I am a power notebook l m user, but regardless, this is pretty freaking sweet. Alright. We have a couple more. Alright. Next, we have Muse Spark meta meta after billions of dollars, quite literally.

Jordan Wilson [00:20:48]:
Right? They just had a a single $14,000,000,000 acquisition, of, Scale AI essentially in acquihire. And then they reportedly were spending hundreds of millions on individual researchers. And then over the past year, they've done nothing with Llama. A year of silence in 2025 and 2026, it's like a decade. I think people forgot Meta existed, but they came back with what I will say to their credit is a fairly impressive model in Muse Spark. So here is Muse Spark. This is their new model, but the first model from the meta the new, meta superintelligence lab or MSL led by former Scale AI CEO, Alexander Wang. And it is a natively multimodal reasoning model with text image and tool use.

Jordan Wilson [00:21:37]:
But unlike the previous llama models, this is a complete rebuild. So this is not an open source model that you can download and fork locally. Right now though, it is free. So free, but not open source. Alright. So it uses a multi agent architecture where parallel reasoning, sub agents tackle part of a problem simultaneously reducing latency on those complex tasks. So who has access right now? Like I said, it's free to use on meta.ai, and in the meta AI app across Facebook, Instagram, and WhatsApp. There is also a thinking mode that they call contemplative.

Jordan Wilson [00:22:15]:
And then, there is a private API preview right now, but it's not public. It's only open to select partners. So I mean, why it's useful? I mean, you get it. Right? I don't know. If you're if you're a heavy, meta Facebook, WhatsApp user and you like AI, that's cool. You're gonna have a much better, AI assistant to chat with. Will this be, a model that's good enough, for businesses to, replace their stack? I don't know. Probably probably not.

Jordan Wilson [00:22:47]:
Right? It's actually really good. Right? The benchmarks for technically being a first model, like, this might sound crazy because yes, it is technically their first model, but it's not right. They had Lava before, but this is a complete rebuild. So if you give it the benefit of the doubt of being a complete rebuild, it's actually one of the best first models ever. But is it a top three model right now? No. It's not because it's still behind Gemini three one pro, Opus four six, and, Chad GPT's GPT five four. However, I have been using it. It's really good.

Jordan Wilson [00:23:23]:
It's good at coding. Its ability to write is pretty good. It's obviously a huge step ahead, from Llama. So, like I said, if if I think you're gonna find value out of this if you are already using, Meta's products on a daily basis, or if for whatever reason, the certain benchmarks really just kind of hit what you're, you know, what you're, trying to get out of it. Like I said, we actually and I will share it. I'll reshare it in today's newsletter. I put together a chart comparing, Muse Spark, with its biggest competitors. And then also just for fun, I put the, the Mythos, model from Anthropic that no one's gonna get access to for the most part.

Jordan Wilson [00:24:10]:
So, yeah, if you wanna see how it ranks on the benchmarks, I'll have that in the newsletter. Alright. Next, Meta wasn't the only company with new models today or this week. Microsoft also did. Yeah. This one, I kinda came out right at the end of last week. I'd already, you know, planned the show out. So this one is, like, eight days old.

Jordan Wilson [00:24:32]:
Alright? But stuff happens fast. Right? So, this is from Microsoft. I'm just gonna read their, little blurb here. So they said introducing m a one, or sorry, m a I Transcribe one alongside m a I Voice one and m I m a I Image two, what they call world class quality at lightning speeds, now available at the most competitive prices. And it is available now in the Microsoft Foundry and the MAI Playground. So I will say if you want to, you know, compare everything on the image side, right, it's I think it's a top three family of models. It's not a top three AI model, image model. But, you know, for their first technical model, it's fairly impressive on the image side.

Jordan Wilson [00:25:24]:
And then on the, the transcribe or the voice side, it actually got some pretty good, benchmarks here. So it was the, the lowest for mean word error rate. Right? So when transcribing, it actually is the top model for that. So if you are building, something with voice or transcription, it might be a model worth looking at. But some more details. So this is the first big series of releases from Microsoft's new in house AI team now led by Mustafa Suleiman, and they released these three proprietary foundational models. So MAI Transcribe One claims the lowest word error rate 3.8% across 25 languages, beating the likes of Whisper, Gemini Flash, and 11 lives scribe. Then you have m a I voice one.

Jordan Wilson [00:26:19]:
I don't know why the m a I is just hard to say. We had, m a I voice one that can generate sixty seconds of natural audio in under one second on a single GPU. So, that's text to, speech model or text to audio model very, very fast. And then the, m a I image two debuted as top three on the arena AI leaderboard. So, these are all available in right now. Right? So, you can access them on the Microsoft Foundry or the MAI playground. So if your company for whatever reason, right, if they're not a big Microsoft shop, maybe you're a Google Gmail shop, or you just don't have access to the Microsoft Foundry, which is kind of there, you know, formerly the, Azure AI Foundry. Well, you can just go sign up for the MAI playground.

Jordan Wilson [00:27:10]:
Right? That's the way I've been kind of playing around these models in the same way that, you know, OpenAI, Claude, etcetera, you know, they have, back ends essentially for developers. That that sounds daunting. Trust me. It's not. You just sign in with your credentials. You usually have to collect, connect a card, credit card, because then at that point, you are charged even if you're just playing around with it in the sandbox, you are charged. So, well, here's why it's useful. I think if you have already built or your company has already built transcription or voice products, it's worth looking at it.

Jordan Wilson [00:27:50]:
Just because at least I do think on the transcribe and voice side, it's fairly impressive. Or maybe, for whatever reasons, maybe your company has wanted to use AI images, but you are locked into Microsoft. In this case, this is their first image model that I think is definitely worth worth using. So who's gonna find this valuable? I think enterprise developers who are building voice agents, if you need accurate, fast, and affordable transcription as, at scale, that's one great use cases, or, you know, marketing teams who need high quality, image generation, and you only have access to Microsoft. Right? I obviously wouldn't choose this, at least not now, over the nano banana, over the GBT 1.5 image. We talked, earlier this week that, you know, there's rumors OpenAI is gonna be coming out with their next image model. So it's definitely not the best, but I have seen, some of the, examples so far, and it looks fairly impressive. Alright.

Jordan Wilson [00:28:51]:
We have two more quick ones for you here. So our next one, Google vids update. Stick with me here. Why does this matter? Well, you probably know Google VO. Right? And if you're on a paid plan, then you have access to Google's extremely impressive AI video generator in VO 3.1. Now it's free. Well, a version of it. And if you use Google bids, so here's, what's new.

Jordan Wilson [00:29:23]:
So Google bids now includes free AI video generation powered by v o three one for anyone with a Google account. Alright? You do only get 10 clips per month. But, y'all, if if if you don't remember, like, when v o three came out or even v o two. Right? But when v o three came out, it's like, everyone's like, oh my gosh. I'll pay a million dollars, you know, for this. Because at the time, it was so good, and it was, you know, so far ahead of, you know, the original Sora and so far ahead of all the other, you know, Chinese models. Now that gap has been closed. But the fact that now we're saying that you can get VO 3.1 for free, it's pretty impressive.

Jordan Wilson [00:30:06]:
So let's talk a little bit more about how this works. Well, one of the big features of Google Vids, I did do a dedicated show on that, a couple of months ago, but it has AI avatars that can be directed via text prompts, placing custom scenes, and dress to match your brand and interact with products and props, which is really, really cool. It also has a new screen recorder, and, sorry, a new Chrome screen recorder and extension to direct YouTube publishing as well. So here's who has access. Like I said, free, free. V o 3.1, free. How many times can I say that? So, up to 10 clips a month, even if you have a non paid account. Also, you have custom music in Lyria three and AI avatars, but that is for paid subscribers only.

Jordan Wilson [00:30:58]:
And then Ultra subscribers gets a thousand, video generations per month. That's obviously on the, 200 plus dollar plan. So here's why it's useful, and I think who's gonna find it valuable. Well, it's useful because it creates really good, v o 3.1. It generates eight second video clips from text prompts or photos at no cost. Right? So, this is one of those. I think if you've been, looking at some some ugly visuals on your company's website for many years, which be honest, this is probably most of us, you y'all liven it up. Or if you have this old, you know, corporate training video and maybe you can't replace all of it, you could at least splice it up and breathe some life into it, with VO 3.1.

Jordan Wilson [00:31:48]:
Alright. And then our last feature, which may be the one that's most talked about, although at least for me personally, I think I'm most excited about the agents SDK, but here we go. We got, managed agents from Anthropic. All right. So more or less, if it sounds like every week that Anthropic is releasing one or two things that sound very much open claw esque. Well, that's because this is very OpenClaw esque. Alright. So this is technically a little bit more technical than using clawed.ai on the front end because you do have to use, Anthropix Council.

Jordan Wilson [00:32:33]:
Right? So that's their kind of back end, but it's super easy. I did mess around with this a little bit. The good thing is is you can do it all in the natural language. You can do it all just by chatting with Claude, and it will walk you through how to set it up, you know, authenticating things, etcetera. But, here's essentially what's new. So, they launched the anthropic launched the Claude Managed Agents right now in public beta, and it's a platform that handles hosting, scaling, monitoring, and failure recovery for AI agents so teams can focus on agent logic instead of infrastructure. So you just define agents in natural language or you can use YAML, then you can set guardrails, and then Anthropic runs them. Alright.

Jordan Wilson [00:33:23]:
So this isn't like you have to set up your own, infrastructure or you don't even need to necessarily run these on, a website or, you know, set up a local instance on your computer or anything like that. It is all running on anthropic's instance. So this has built in orchestration that handles tool calling decisions, context management, and error recovery, with session tracing. Right? So, the easiest way to explain this, think Opus 4.6 is agentic by default. Right? And I think that we've all seen, I've said this many times. Right? Anthropic is crushing it in 2026. They are winning the year. And one of the things that's most impressive is as an example, you know, adding skills, you know, inside of Claude, adding all of these different MCPs inside of Claude.

Jordan Wilson [00:34:18]:
Right? Interactive ones as well. These more interactive apps. Now imagine kind of similarly to Zapier being able to control that harness a little bit more. Right? Because if you didn't have, you know, as an example, a connector, you know, for whatever program that you want to use or if it wasn't available as an MCP, yes, you could go set up a custom MCP server, but this is just a different way to set it up, by doing it just in natural language. And then the thing that I love is that their kind of agent builder walks you through it, and it'll do the authentication and everything for you. Right? It'll walk you through it. It'll give you options. Right? It'll say, hey.

Jordan Wilson [00:35:02]:
You can do it via, you know, OAuth where you can just click the screen, or you can do it via API key as an example. So it's gonna walk you through it. It's gonna give you these different options. So who has access right now? It is a public beta. So anyone. Alright? For API users and subscribers. So like I said, you do have to be on the back end. The cool thing is, well, you technically don't need to even be a paid Claude user.

Jordan Wilson [00:35:28]:
You do though obviously have to connect a credit card, go into the back end. I think the minimum, you know, put in, like, $5. Right? You can, like, prepay $5 to go, like, mess around with this. Right? So here's why it's useful. Because you can abstract months of infrastructure work, right? The sandboxing, the state management, credential handling, permission systems, all these things that typically delay ancient projects and just let Anthropic manage that piece. Right? So Anthropic claims 10 times faster time to production compared to building the agent infrastructure yourself. Alright. So here's who's going to find it valuable.

Jordan Wilson [00:36:10]:
I think enterprise engineering teams who wanna deploy Claude powered agents without building infrastructure from scratch, you know, project managers and non developers who can define agents in natural language, or companies that are already using Claude's API, who wants managed scaling and monitoring for their agent workloads. So we will see how, you know, how popular that this is from Anthropic. On the surface, it seems like it will be fairly popular, although it is a little bit more technical. But then at the same point, I thought that OpenAI's, their version of this with their agent builder, it seems like that never really took off. And it's very similar to what, Claude just released here in managed agents. So, I could be wrong. We'll see. Maybe most of the momentum, goes into things that you use on the front end, which is something I've been saying for many years.

Jordan Wilson [00:37:03]:
Right? I've always said, you know, when you're talking about buying your building, right, I'm like, well, you should technically be buying and using these systems on the front end because the front ends are becoming more and more powerful. And I do think, you know, even as we see OpenAI eventually go to the super app, I think then it'll become apparent to people why using these things in the front end is probably gonna become more prominent, because, well, number one is it's gonna share and keep your context. Right? So, yes, I think things like the, managed agents from Anthropic are really good. I don't know if it's gonna take off as an example the way that Claude Co worker has, the way that Claude code has. Because when you do it there, right, you have it all kind of under the front end, quote, unquote, front end system. Whereas sometimes moving on the back end, yes. Obviously, for developers, that's not who I'm speaking to. You all, can see the the promise of this.

Jordan Wilson [00:37:57]:
But for everyone else, I don't know if this is gonna take off. Although I do encourage you to go play around with it just like I did. I think you're gonna find plenty of use cases. Alright. So that is a wrap. There are the seven features that I don't think you can afford to miss. So whether you are a Zapier power user, or if you want a managed agent inside of Anthropic, or if you've really, been wanting some more content creator, you know, capabilities inside of OpenClaw. This week brought a whole lot in a whole lot from Google and Microsoft as well.

Jordan Wilson [00:38:32]:
So I hope this was helpful. If so, let me know about it. If you could, leave me a rating on the podcast. I'd really appreciate it. Make sure to follow and subscribe to the show, then go to youreverydayai.com. Sign up for the free daily newsletter. Thank you for tuning in. Hope to see you back tomorrow and everyday for more everyday AI.

Jordan Wilson [00:38:50]:
Thanks, y'all.

Gain Extra Insights With Our Newsletter

Sign up for our newsletter to get more in-depth content on AI