EP 473: Claude 3.7 drops, OpenAI releases GPT-4.5 and more AI News that Matters

AI Model Innovations: Breaking New Ground

In a groundbreaking development, Anthropic released the Claude 3.7 SONNET, the first publicly available hybrid AI model. This innovative model marries traditional transformer capabilities with advanced reasoning, allowing users to toggle between rapid responses and deeper logical thinking. Particularly impressive is its performance in coding and front-end web development, signaling a valuable tool for software engineering sectors.

Meanwhile, OpenAI launched GPT-4.5, its latest and largest language model. Notably, GPT-4.5 emphasizes human-like interaction, boasting enhancements in writing skills and world knowledge. Although initially accessible only to research preview users, broader accessibility is anticipated in the upcoming weeks, promising a significant upgrade in AI-human communications.

Big Tech's Strategic Shifts: A Competitive AI Race

Major shifts were also observed as tech giants realign their strategies to maintain competitive edges in AI. Meta, for instance, is reportedly developing a standalone AI app, moving beyond integration within social media platforms to directly compete with OpenAI and Google. This move reflects a broader strategic initiative to reach non-social-media users and expand its AI services.

Conversely, the Google cofounder advocates for intensified efforts to develop Artificial General Intelligence (AGI), pushing for enhanced productivity and in-office presence. This highlights an industry trend where tech giants are increasing pace and operational change to emerge as leaders in the AI domain.


Enhancing AI Accessibility: Microsoft and Eleven Labs Lead

Microsoft showcased its commitment to democratizing AI technology by offering users free unlimited access to its Copilot's advanced voice interactions and enhanced reasoning capabilities. This move positions Copilot as an essential tool for businesses and individual users interested in leveraging AI for productivity and innovation.

Similarly, Eleven Labs expanded into the speech-to-text market with the release of Scribe, their first standalone model supporting over 99 languages. This development promises high accuracy and could fundamentally change how businesses manage and transcribe verbal data.


AI in Action: From Consumer Tools to B2B Solutions

While Amazon refutes claims of Anthropics AI powering its new Alexa Plus features, it remains evident that the evolution of voice assistants continues to be a focal point across tech companies. Integrating advanced AI capabilities into everyday tools provides businesses with new opportunities to enhance consumer interaction and streamline operations.

In the realm of business applications, Sesame's Maya chatbot is gaining traction for its human-like conversational flow, although opinions vary regarding its practical utility in professional settings. Despite this, innovations like Maya suggest potential new pathways for integrating AI into customer service and client interactions.

Conclusion

The dynamic landscape of AI continues to offer transformative opportunities for businesses willing to adopt and adapt. Leaders and decision-makers must remain informed and proactive, leveraging these technological advancements to drive strategic growth and maintain competitive advantage. Whether through adopting pioneering models like Claude 3.7 SONNET, GPT-4.5, or exploring new applications in speech recognition, the future of business is undeniably intertwined with AI innovation.


  • Topics Covered in This Episode

  1. Anthropic releases Claude 3.7 Sonnet, the first hybrid AI model
  2. Google cofounder Sergey Brin pushes for 60-hour workweeks to win AGI race
  3. Meta reportedly developing a standalone AI app to compete with OpenAI & Google
  4. Microsoft Copilot now offers free unlimited access to voice and Think Deeper
  5. Eleven Labs launches Scribe, a speech-to-text model supporting 99+ languages
  6. Amazon denies reports that Anthropic’s AI powers new Alexa Plus features
  7. Sesame AI’s voice chatbot, Maya, gains attention for its human-like conversations
  8. Apple announces $500B US investment, including AI server factory in Texas
  9. Bloomberg reports Apple’s advanced Siri won’t be ready until 2027
  10. OpenAI releases GPT-4.5

Podcast Transcript



SJordan Wilson [00:00:16]:
Anthropic Claude dropped 3.7. SONNET, OpenAI responded days later with their much awaited model, GBT 4.5. And we may be waiting many more years before we actually see a full AI from Apple. We might even see AGI before we see full Apple intelligence. Yeah. It was one of those kinds of weeks in AI. It's like, wait. This all happened in one week.

Jordan Wilson [00:00:46]:
Yes. It did. And if you missed any of it, and that's just the tip of the iceberg, don't worry. We're gonna be going over those stories and a whole lot more today on everyday AI. What's going on y'all? My name is Jordan Wilson, and I'm the host, and this thing, it's for you. This is your daily livestream podcast and free daily newsletter, helping us all not just keep up with AI, but what it all actually means. Right? Deciphering all this PR and all these news releases from all the biggest companies so we can actually use that information to grow our companies and our careers. If that sounds like what you're trying to do, maybe this is the first time you're listening, welcome.

Jordan Wilson [00:01:25]:
This is your new home. Your other new home is youreverydayai.com. There, you can sign up for our free daily newsletter. So each day, we recap, exclusive insights that we bring you only on this very podcast as well as we keep you up to date with everything else that you need to know in the world of AI to be the smartest person in AI at your company or your department. Right? So if that sounds like what you're trying to do, make sure you go sign up for that free daily newsletter at youreverydayai.com. Alright. Quick reminder. We're gonna be oh my gosh.

Jordan Wilson [00:01:59]:
It's like two weeks away. We're gonna be broadcasting live with NVIDIA at their GTC conference, starting, March. What day are we actually gonna start there? Probably March 17, the Monday. So, kind of for that week, at least the first couple of days early in the week, we're gonna be partnering, with NVIDIA to be bringing you a lot of exclusive, insights, some great expert interviews, maybe breaking a little bit of news as well. So, really excited for, this year's GTC conference. We were lucky enough last year to partner with NVIDIA as well. So, hey, let me know. Hit me up if you're gonna be, at the GTC conference in San Jose.

Jordan Wilson [00:02:41]:
Would love to say what's up. Alright. With that, let's get into what's happening in the world of AI for the week of March 3. Let's get after it, y'all. Alright. So Anthropic. Yeah. That seems like it was a week it it it was on Monday.

Jordan Wilson [00:03:02]:
Right? Right after this show. Seems like it was a month ago that Anthropic released 3.7 SONNET, but they did. So, Anthropic has introduced Claude three point seven SONNET, the first publicly available hybrid AI model, which combines that traditional transformer capabilities with advanced reasoning. So it does kind of merge these two AI paradigms. It combines the traditional transformer model with reasoning capabilities, allowing it to switch between rapid responses and deeper logical thinking. So extended thinking mode is a key feature, and right now the extended thinking is only for paid users, and they can enable that mode where the model spends more time solving complex problems, also showing a summarized chain of thought for transparency. So free users can use the Claude three point seven SONNET model, but at least as of right now, they can't toggle on that extended thinking. So Claude three point seven SONNET is very impressive when it comes to, coding, software engineering, anything like that.

Jordan Wilson [00:04:10]:
So, if that's something that you're in charge of at your company, you're probably gonna wanna check out three seven SONNET. So it scored a very impressive 70.3 on the sweet bench coding benchmark, outperforming competitors like OpenAI's o one and o three many. It is particularly strong, Claude three point seven SONNET that is encoding and front end web developments, making it a great choice for software engineers. Anthropic also released Claude Code, which again, live stream on us. Let me know if you want if we should dive into that. It's slightly more technical, But essentially, like, anthropic released, I'm not gonna say it's a cursor competitor or a competitor to to Bolt and Lovable and Windsurf and all these kind of AI IDE's, but it kind of is. Right? Even though cursor, has said that, you know, hey. Anthropic Claude is our default model.

Jordan Wilson [00:05:04]:
It does look like with Claude code, Anthropic is trying to get into this, AI, straight up development space, which I think is a smart move. So they did launch quad code, a command line tool that allows developers to interact with and update entire code bases directly from their computer's terminal. So it integrates also with GitHub and supports debugging, signaling, and tropics push into the AI powered coding assistance space. So despite these advancements, Claude stood firm on their API cost, which at the time seemed kind of silly until we got OpenAI's API pricing for their latest model. So just wait on that one. So it is still the same price, $3 per million input tokens and $15 per, million output tokens. But, I mean, here's what I think with with with this latest model. Right? So, a lot of people are like, oh, looks like Anthropics, you know, trying to, you know, get get the best of OpenAI right with, kind of a logic based model with a a model that reasons.

Jordan Wilson [00:06:15]:
I'm being honest. I I I was not very impressed, with anthropic clawed three point seven's, ability to reason. But again, I am a power I am a heavy user of OpenAI's o three Mini. I I'd say that's my most used model. I use o one pro as well. So, you know, looking at Anthropic's first, foray into the, kind of, you know, reasoning models, not super impressed by that. Also, a lot of people are reportedly rolling back to three fives on it for certain tasks, especially when it comes to coding because it seems like sometimes, Claude uses that reasoning when it maybe shouldn't, and it takes things a little further and, you know, changes a bunch of things that you maybe didn't even want changed. So, I am using, three seven SONNET every single day for certain use cases.

Jordan Wilson [00:07:05]:
I think it's great. It's the first right. You have to tip like, I keep saying this. You have to tip your cap to anthropic, because they are the first company with a hybrid model. And I do think that that is, going to be the big, kind of the the the future of large language models is gonna be kind of combining this, quote, unquote, old school, transformer approach with the new, reasoning models. So, yeah, I'm curious, you know, because we're gonna talk about GPD 4.5 at the end. You know, I'm curious for our livestream audience and, you know, hey. Let me know if you're listening on the podcast as well.

Jordan Wilson [00:07:41]:
What are your thoughts on these newest releases? Right. So, Sandra here joining us from YouTube says I haven't been able to tell the difference yet. I've I personally I have been able to tell the difference with, SONNET three seven. You know, when I'm testing it just for, you know, niche coding tasks, which is something I don't necessarily do a lot unless I'm testing models, it's great for that. Everything else that I might use Claude SONNET for on an on an everyday basis, I I'd say it's about the same. I think the, the artifacts feature has actually improved, for that reason. Right? If you're trying to visualize data or something like that, I think it is better. But for non coding, non data visualization tasks, I don't know if, you know, we we necessarily saw a huge leap, with 3.7 on it.

Jordan Wilson [00:08:27]:
Again, that's in my testing. I'm using it maybe, I don't know, forty five minutes to an hour every single day since it came out, you know, for the show last week. I probably tested it for a good four or five hours. So, you know, I'm not using it, you know, five hours a day since it came out or anything like that. But I'm I'm curious, what everyone's thoughts are. Are you still running in circles trying to figure out how to actually grow your business with AI? Maybe your company has been tinkering with large language models for a year or more, but can't really get traction to find ROI on Jenna AI. Hey. This is Jordan Wilson, host of this very podcast.

Jordan Wilson [00:09:08]:
Companies like Adobe, Microsoft, and NVIDIA have partnered with us because they trust our expertise in educating the masses around generative AI to get ahead. And some of the most innovative companies in the country hire us to help with their AI strategy and to train hundreds of their employees on how to use Gen AI. So whether you're looking for chat g p t training for thousands or just need help building your front end AI strategy, you can partner with us too, just like some of the biggest companies in the world do. Go to your everydayai.com/partner to get in contact with our team, or you can just click on the partner section of our website. We'll help you stop running in those AI circles and help get your team ahead and build a straight path to ROI on GenAI. Alright. Our next piece of AI news. Google is pushing hard for AGI, and apparently, they just need to work more.

Jordan Wilson [00:10:05]:
Alright. So Google cofounder Sergey Brin has reportedly called for an increased productivity and more in office presidents, presence for Google as the company intensifies its push to develop artificial general intelligence or AGI. So Brin believes AGI is within reach according to reports if employees just work a little bit harder. So in an internal memo seen by the New York Times, Brin stated that Google has all the ingredients to win the AGI race, but needs to, quote, unquote, turbocharge its efforts. He suggested that employees work at least sixty hours a week, calling that the sweet spot of productivity. Also, return to office policies are being emphasized as Brynn recommended employees come into the office every weekday, exceeding Google's current three day a week on-site policy. He argued that remote work and reduced hours are demoralizing to others. So Google's AI teams are already working long hours.

Jordan Wilson [00:11:10]:
So according to CNBC, some staff working on Google Gemini's AI projects clocked up to a hundred and twenty hour weeks to address critical flaws in their image recognition tool. Imagine putting in a hundred and twenty hours a week. Right? That's nuts. That's like, I don't know. I'm not great at math, but that's more than fifteen hours a day. Imagine, right, fifteen can someone do a live calculator? Right? Let's let's do fifteen fifteen times seven. I should be able to do that in my in my brain right now, but I can't. Okay.

Jordan Wilson [00:11:47]:
So that's a hundred and five hours a week. So that's even more. Hundred and twenty hours. That's like seventeen hours, a week. No. Thanks. Or seventeen hours a day, working to debug this. Right? But imagine working fifteen plus hour days, and then they say, ah, you know, we'll achieve AGI if you just work a little harder.

Jordan Wilson [00:12:07]:
Not probably what most people wanna hear. So, high pressure work culture, in AI development is raising concerns. So while the push for AGI could lead to groundbreaking advancements, the demanding workload such as, as an example, the routine twelve hour days reported by employees at XAI developing Grok highlights the toll on the industry. So I don't know how I feel about this. Right? Number one, I thought AI was supposed to allow employees to work less. Right? And and and focus on, you know, higher quality, work and to, you you know, like so I don't know. Part of this to me just isn't adding up, especially since Google, I think, should have been light years ahead of their closest competitors, in the AI race considering they essentially developed the GPT technology, but it was other companies that really ran away with it. And I would say that Google did not even catch up in the AI race probably until late twenty twenty four.

Jordan Wilson [00:13:16]:
So now it seems like they really want to win the AGI race and are really just, pushing employees to work more, work smarter, be more productive in the process. So, I don't know. Seems seems kind of ironic that, you know, we're supposed to all be benefiting, from AI in large language models and, you know, we're supposed to be you know, we're focusing on higher level creative and strategic tasks, and it's like, nope. Just work double. Just work triple. Right. A little wild. Yeah.

Jordan Wilson [00:13:47]:
Suresh from, from YouTube just said Google finally woke up after missing out for a couple of years. Yeah. It's like, oh, we we weren't really, like, in the in the, the race for, you know, 2022 and 2023. So now we just gotta work, double or work triple, to catch up. Yeah. So much for that four hour, work week we are all promised the utopia of AI, at least not yet. Alright. Other big tech trying to play catch up.

Jordan Wilson [00:14:16]:
Meta is reportedly developing a standalone AI app to compete with OpenAI and Google. So a new report suggests that Meta is launching a dedicated app for its Meta AI assistant, potentially signaling a major shift in its AI strategy. So according to reports, Meta is working on a standalone app for its AI assistant, Meta AI. The app would mark a departure from Meta's current approach of just integrating the AI services into social platforms like Facebook, Instagram, and WhatsApp. So the standalone app could help Meta reach users who avoid social media, like me, or use messaging services from competitors addressing a gap in its current strategy. This move could potentially bring in millions of new users who were previously out of reach. So Meta AI, their online service, which launched in 2023, offers features like question solving, image generation, and answer suggestions. So while it has seen improvements, it still lacks advanced functionalities, offered by competitors like OpenAI's ChatGPT and Google's Gemini.

Jordan Wilson [00:15:25]:
So Meta CEO Mark Zuckerberg has ambitious plans for AI stating earlier this year that Meta AI could become the leading personalized AI assistant reaching over 1,000,000,000 people. And the standalone app does align with that vision. Also, monetizing. Right? It's obviously a key part of, Meta's reported new strategy because they could roll out paid plans and premium features that they might not as easily be able to roll out, you know, in something like Facebook or, Instagram or WhatsApp where a lot of their users may be using the Meta AI technology. Fred says, I just won't use meta. Right. I use it I use it online. It's actually a great online, resource.

Jordan Wilson [00:16:13]:
Right? We did a head to head on this show, probably like three or four months ago just how, accurate certain large language models are when using the Internet. So I think we ran down what we did OpenAI. We did Google. We did Meta, and we did Copilot. And And I was actually surprised by how well meta performed. Right? Essentially, the biggest thing, the biggest takeaway from that one was how well it just surveyed the Internet and could, return back, an accurate answer using its llama model connected to the Internet. So, I was I was actually pretty pretty impressive and, yeah, their their augmented reality is what they're definitely focused on. But, yeah, I mean, they're definitely wanting to marry kind of those two different technologies.

Jordan Wilson [00:16:58]:
Right? The wearable technology and the, you know, just the stand alone, LLM that's outside of their social media network. So alright. Speaking of big tech, yeah. It's just been a big tech kinda week, but pretty impressed here with Microsoft. So Microsoft Copilot is now rolling out some major updates including free unlimited access to voice and their think deeper capabilities. So that is the o one. Yes. The OpenAI o one model, you can now use it for free and unlimited access.

Jordan Wilson [00:17:37]:
So if you just go to, Copilot.Microsoft.com, you have to have an account, but you can use essentially the, their Microsoft Copilot's voice mode, which is not quite as good as OpenAI's advanced voice mode, although it uses the same technology, but you can use the o one model. OpenAI's, you know, what was, you know, a couple of months ago, their most powerful model. You can use it for free in unlimited. So huge move here from Microsoft that I think kind of got slept on. So the voice feature allows users to interact with Copilot hands free. Use cases include practicing a new language, preparing for a job interview with mock q and a, or receiving step by step cooking advice. I've actually used, Copilot for that exact reason. I don't think it still helped me, and that wasn't Copilot's fault.

Jordan Wilson [00:18:31]:
It's just I just can't cook. You know, maybe maybe I do need that, figure o two robot, that just silently does work in your kitchen for you. But, think deeper. I think it's it's it's really good. So, we we did a review of it when it first came out, I don't know, five months ago. So probably you're gonna have to revisit that, but, I think it's great. So anything that really requires some advanced reasoning, some logic. Right? Maybe you're using, you know, chat GPT or Claude or Gemini or something else, and you don't have a paid plan.

Jordan Wilson [00:19:04]:
And you're like, wow. I'd love to be able to use a reasoning model. So right now you can use very limited, o three minutei for free on chat g p t. But if you wanna get your hands on, o one, which is a very, very, very good capable model, now you can do it. So Copilot Pro users, yeah, which is myself. I'm on Copilot Pro. I paid $20 a month for that. And I was like, wait.

Jordan Wilson [00:19:29]:
What are we getting? You know? What are we getting now for $20? But, you know, I do appreciate that Microsoft at least emailed Copilot Pro users, and they're like, yo. We're making this free version really, really good. If you wanna cancel, here's the link. More more big tech companies should do that. Right? Like, if they make something free and all of a sudden the paid plan isn't that good, they should be like, hey. It's fine if you cancel. Here you go. However, here's what the Copilot Pro, kind of, account still has.

Jordan Wilson [00:19:57]:
So, you know, think of this differently. Right? I'm still a Mac user. Yes. I I have a Windows Copilot Plus PC. I still gotta use I still gotta set it up. I've been so busy. I'm actually super excited to do that, but, you know, maybe your company uses Microsoft three sixty five Copilot. For the most part, then this is not going to impact you.

Jordan Wilson [00:20:16]:
Right? Unless you're, you know, using it, you know, not logged in outside of your company's, you know, biz chat maybe. Right? So in that case, sure, you can you can use this that way. But this is for, I think, is going to appeal appeal to a lot of people who are Mac users, and who maybe don't have that Microsoft three sixty five Copilot integration. But Copilot Pro users, so though that's, those that still pay $20 a month, can still continue to enjoy and use Copilot, across the different Microsoft three sixty five apps like Word, Excel, PowerPoint, etcetera. So, yeah, even though, you know, as an example, I'm using a Mac right now. I still have Microsoft Word on my computer and I can use, Microsoft Copilot via Copilot Pro, in Microsoft Word, in, Excel, in PowerPoint, etcetera. So, you know, pro users aren't like SOL. Right? They're not just out of luck.

Jordan Wilson [00:21:10]:
They still have, some Copilot capabilities that normal, free users do not have. But y'all, if you if you do not use Copilot yet, I would just go right away and try out, the unlimited voice and the think deeper, features. They're actually super impressive. Yeah. I I I like what Graham here from, from LinkedIn is saying. Copilot is the gateway to AI for most people yet now. Yeah. So, a couple of weeks ago, I would probably say, no.

Jordan Wilson [00:21:41]:
Probably chat g p t free. But now I'll say those are like one a and one b. I see the free version of chat g p t, is probably a little bit better than this free version of Copilot. But now they're at least, hand in hand, especially if your company, is a Microsoft organization. I think it's I think it's pretty big here. So, alright. Next piece of AI news. One of the biggest names in AI on the on the text to speech side is changing up their biz model a little bit.

Jordan Wilson [00:22:15]:
So Eleven Labs, they're an AI startup company that was recently valued at $3,300,000,000, has introduced its first standalone speech to text model called Scribe, following a $180,000,000 funding round. So, yeah, you maybe heard of, Elephin Labs as the, text to speech, but now they're flipping it, and going speech to text. So, yeah, even if you listen, to this show on the podcast, there's a there's a little intro. Right? It's it's this, AI intro to but that's eleven Labs. Right? That was like the very first version of 11 Labs, by the way, which was I thought, pretty impressive for, you know, two and a half years ago. So the Scribe model supports over 99 languages and boasts exceptional accuracy for more than 25 of them including English, French, German, Hindi, Japanese, and Spanish. So the company, claims an impressive 97% accuracy rate for English with a word error rate below 5% for its top performing languages. So according to ElevenLabs, Scribe outperformed competitors like Google Gemini two point o Flash and OpenAI's Whisper Large v three, that is OpenAI's, kind of, speech to text model in benchmark text such as fluors and common voice, setting it apart in speech detection market.

Jordan Wilson [00:23:46]:
So, one cool feature that I liked, hopefully I pronounced this right, but it includes advanced features like smart speaker diarization, which is essentially just automatically identifying the speaker that is speaking, which like for someone that does a podcast and usually I have guests on, that's huge. That's why I can't use, like, personally, I can't use something like Whisper or Google Gemini two point o. Right? Because I normally have a guest. Right? So I actually use a a tool called Cast Magic that has that built in. So I'm I'm gonna definitely be checking out, this new offering from eleven Labs with which can auto identify speakers. I think that's huge. It also has word level time stamps for precise subtitles and auto tagging of sound events like laughter. That's kinda cool, kinda creepy as well, but kind of useful.

Jordan Wilson [00:24:40]:
So right now, Scribe only works on, it only works with prerecorded audio, but will soon offer low latency real time, options. So that's pretty cool. When that comes out, I'm gonna have to hit up 11 and, you know, always have that going live when I have a guest on the show. That would be super helpful for me. You know, I'm curious. Does anyone out here like, do you all use tools like Eleven Labs? Like I said, I think it's been great, for text to speech. It was one of the leaders, in that, field. I think it still is for for when people just needed voice overs, you know, audio books.

Jordan Wilson [00:25:19]:
You know, I think people overused it maybe early on and didn't put enough care into it, but it's actually a a very, very good platform. So George here is saying Eleven Labs Jibberlink, allows AI to talk to AI. Yeah. In a faster than human language. I saw that. That was a pretty cool demo. Essentially, you know, two AI agents were talking to each other. They identified that they were both AI agents.

Jordan Wilson [00:25:44]:
I believe this was an open source project, and then they just use their own, Jibberlink, kind of to talk to each other. Sounded like two fax machines, you know, talking to each other. It was it was pretty cool. Yeah. Samuel here says I pay the $5 a month to 11 labs just to be able to listen to my docs. Huge point, Samuel. I do that as well. That's one of the things I actually use 11 labs for the most.

Jordan Wilson [00:26:07]:
I pay for it. Sometimes I have a big block of text, and I don't wanna go into, as an example, OpenAI's back end in their playground because you can only do, a certain amount of text at once. So yeah. I often will do that if, you know, many times I'll just throw it into notebook l m and get more of a summary or more of a conversation around it. But if I actually need to read something, you you know, point by point and I'm super busy, a lot of times I'll just grab that, you know, couple thousand words, throw it into 11 labs, you know, crank the output up to two x because, you know, that's why I speak at so quickly because I listen to this stuff so quickly, I guess. But, yeah, a great great use case, I would say, for 11 Labs. Look at this. Here we are, you know, just rounding out the big tech lineup for this week.

Jordan Wilson [00:26:55]:
So Amazon is denying reports that Anthropics AI is powering their new Alexa Plus features. So, yeah, if you pay attention, to this show, we cover this. So, Amazon finally, announced their smarter Alexa, right, powered by large language models. And earlier reports was it was powered by Infropic Cloud's AI model, but apparently not because Amazon well, at least not entirely. So Amazon has publicly refuted claims that its recently announced Alexa Plus capabilities are powered by Anthropic's clawed AI models, and this is sparking a lot of discussion now online. So Amazon insists that its in house models, which is called Nova, powers the majority of Alexa plus conversations. So in response to a CNBC report claiming Infropix quad model was handling most customer interactions, Amazon stated that Nova has managed over 70% of conversations, including complex requests in the past month. So, yeah, this is just starting to roll out over the next, like, week or two, to paid Amazon users.

Jordan Wilson [00:28:12]:
So maybe that was just in testing. They also didn't really say, like, hey. What's the other 30%? So, I'm assuming maybe the other 30% is clawed. We'll have to see if this report just came out. But, anthropic obviously is a key investor in, sorry. Amazon is a key investor in anthropic. And Amazon, the company maintains that its own proprietary AI, Nova, is responsible for some of these advanced features. So the upgraded Alexa Plus boasts generative AI capabilities and dubbed Alexa Plus, the new iteration is designed to be more conversational and capable handling tasks like grocery shopping, booking services, sending texts, and browsing websites.

Jordan Wilson [00:28:56]:
Yeah. It was funny. In the, in the Alexa Plus kind of demo, I'm like, why why are all these demos just like you buying more things from Amazon? Right? Why can't you just show me an instance where, Alexa just isn't dumb? Right? It's like so much of the demo is just like, oh, you're buying more stuff from Amazon. Like, I I don't know. Like, literally, I'll ask for, like, I don't know, the weather or hours at a store, and then Alexa, old school dumb Alexa is still just being like, do you want me to add this to your cart? And I'm like, I asked about the weather. So I don't know. I'm I'm I'm not super, pumped, about this, but, you know, it's gotta be better than what we currently have. And this was supposed to come out many, many months ago, but the launch has faced delays and early challenges as Alexa Plus was delayed due to issues with hallucinations and incorrect answers, during testing.

Jordan Wilson [00:29:49]:
However, Amazon CEO Andy Jassy highlighted the transformative transformative impact of Jet AI on making such advancements possible. So yeah. Michelle says Alexa and Siri are both useless. Don't worry, Michelle. It looks like Siri might be useless until 2027. So more on that here in a couple of minutes. Alright. Speaking of voice assistance, a new one is taking over the web.

Jordan Wilson [00:30:22]:
So, yeah, kind of all weekend and even early, early today when I was sleuthing online to bring you all the latest news. Sesame, the AI chatbot, and their chatbot's name is Maya, is really grabbing a lot of headlines for its ability to mimic human conversation with uncanny realism. So Sesame's Maya aims to cross the uncanny valley of conversational AI. So the company showcased Maya in a demo emphasizing its ability to replicate human speech and interaction, making it feel more like talking to a real person than a chatbot. So, yeah, a lot of people that, like, I actually follow and respect online, were losing their noodles over, Sesame and Maya over the weekend. So, Maya impressed a lot of users with its conversational flow and realism. So during, test conversations, which you can go right now, you don't even have to have an account. It it shows it does how am I gonna say this? If you're not a heavy conversational user, you might really be impressed, with this new Sesame voice assistant.

Jordan Wilson [00:31:39]:
Alright? It is more neural. It responds with very low latency. The voice does sound, more realistic and more human. For me, it drove me crazy. I absolutely hated it. I probably won't be using it. Sorry, Sesame. You're not gonna be sponsoring the Everyday AI Show anytime soon.

Jordan Wilson [00:32:02]:
Is it great? Yes. Does it have a very high ceiling? Sure. Right? I don't know. For me, one thing I noticed in testing this new Sesame AI voice model, you know, I'm and I'm curious, livestream audience. Did any of you guys use this over the weekend or, you know, early today? It it doesn't seem to do a good job of actually answering your question. So I think a lot of people are fooled by this low latency claim that a lot of companies, you know, AI voice companies are putting out there because usually the initial response to a question that you ask it is this kind of a delay tactic. Right? Or it just says, I mean, kinda like a human would. Right? Like, they kinda laugh about your question or go like, oh, that's a good one.

Jordan Wilson [00:32:49]:
Right? So is it actually low latency? I mean, yeah, like, yes and no. Right? I think they achieve that, like, immediate human to, you know, AI conversational rate, like, where it can respond to you almost immediately because it just responds back with some useless, needless, unrelated. Right? It's just this little quip that buys itself some time to then answer your question. Also, at least for me, kind of a default of this Sesame I found extremely frustrating, because at least for me, when I talk to an AI voice assistant, I don't want fluff. I don't. And that probably puts me in the minority of people. Right? Maybe people want, you you know, this, I don't know, unrelated quips and stories. No.

Jordan Wilson [00:33:37]:
I want facts. I want stats. I want fast. I don't want any I don't want any gibberish. Right? So if you're like me and maybe would prefer talking to, you know, robots sometimes versus humans. Right? Like, yo, I just want facts, stats, and I want it quick. Right? So at least for me, Sesame wasn't really, appealing, probably not something I'll be using a lot, and it did struggle, it seemed like with just fact recall. Right? Something simple.

Jordan Wilson [00:34:06]:
One thing I always do is just be like, tell me about the everyday AI podcast. Right? And it didn't. Right? And and it should be in the training data. Right? Because we've had, hundreds of episodes dating back to 2022. Right? Is that right? And we've been doing it for this long. No. Twenty twenty three. So, you you know, it did struggle with just fact recall and some other things that I tried, but it is free.

Jordan Wilson [00:34:26]:
Go try it out for yourself. Let me know, let me know what you all thought. So Sam Samuel says Maya's EQ is next level. That's true. So if you are more on the emotional, you you know, if you are looking to get some, EQ benefits out of an AI model versus just IQ, I've talked about this. I prefer IQ. The EQ is super nice. You know, it it is pretty good.

Jordan Wilson [00:34:52]:
George says it feels it it is very touchy feely, but the voice, is really good. Yeah. I'd agree I would agree with those, with those observations. Yeah. Nisiani. Nisiani knows. That's a former journalist in me. Yeah.

Jordan Wilson [00:35:04]:
Anytime someone puts something out, I'm like, I don't know about this. Let me go test it, and I'll tell you guys how it is. But, I still think it's pretty impressive. You can go go check it out. Right? Alright. Our last couple pieces of AI news. So Apple has announced a $500,000,000,000, investment in The US, including a server factory in Texas. So according to reports, Apple has announced a massive $500,000,000,000 investment in The US over the next four years, which will include a new AI server factory in Texas and reportedly the creation of 20,000 research and development jobs nationwide according to Reuters.

Jordan Wilson [00:35:47]:
So the investment will span a variety of areas, including purchases from US suppliers, manufacturing expansions, and content production for Apple TV. So Apple will, reportedly partner with Foxconn to develop a 250,000 square foot facility in Houston that will assemble servers for its AI powered services. These servers are currently made outside of The US, marking a shift toward domestic production. The company also plans to double its advanced manufacturing fund from 5,000,000,000 to 10,000,000,000 with a significant portion allocated to producing advanced silicon at the Taiwan semiconductor manufacturing, facility in Arizona. So most of Apple's products are assembled overseas, but many components such as chips from Broadcom and Skyworks Solutions are made in The US. So as part of this, investment, Apple will launch a manufacturing academy in Michigan to offer free courses in project management, and manufacturing process optimization to small and midsize companies. Hey. More news on AI voice assistance that apparently aren't going to be super smart anytime soon.

Jordan Wilson [00:37:07]:
Well, at least not series. Right? So we just heard that Alexa is getting smarter, and that's gonna be rolling out, to paid Amazon, subscribers here in the coming weeks. But you might have to wait until 2027 to get that fully smart Siri from Apple and Apple Intelligence. Yes. I did not get that wrong. New reports are suggesting, from, reports from Bloomberg, always on the spot that Apple's long awaited overhaul of Siri described as a modernized conversational version is now reportedly delayed until 2027. So this up yeah. We might have, someone on the live stream this morning, I think Michael, said we might actually have AGI, before we actually have, a smart Siri.

Jordan Wilson [00:38:03]:
So, the upgraded Siri is expected to debut with iOS 20, combining a generative AI approach with the assistant's classic features, for a more advanced seamless experience. I think by classic features, it just means an assistant that's not super useful. So while Apple is planning to release a limited LLM powered version of Siri with iOS fifth, 18.5, so that would be pretty soon. It will reportedly run as a separate model and fall short of the significant improvements users are anticipating and Apple is marketing. So Bloomberg reports noted that the real upgrade to Siri will begin to take shape in nineteen point four, but won't reach reach full maturity until iOS '20. Wow. The enhanced Siri is expected to feature contextual understanding and improved autonomy, potentially rivaling advanced AI assistance currently dominating the market, but that's AI assistance of today. Right? I don't I I fully don't understand how Apple and their Apple intelligence has so severely fumbled the bag.

Jordan Wilson [00:39:16]:
That's I don't know. What's the euphemism for, like, fumbling the bag but 10 times worse? Apple had all the money, all the resources. They know where this technology is headed. They're partnered with OpenAI and, you know, chat g b t. There's a chat g b t Siri integration, so they have to be getting, a good amount of data, from this partnership yet 2027. Right? Alright. I get it. So one part of me is like, alright.

Jordan Wilson [00:39:43]:
It's better to, underpromise and overdeliver. Right? Then a lot of companies are like, oh, we're releasing, the world's best model tomorrow, and then it takes, like, three years. So I get it. I just do not I cannot fathom how Apple is so so behind at least when it comes to bringing all of these things to market. Yes. Apple, I think, is one of the leaders in putting privacy and security first. Right? But I don't know, like, at what cost. Right? If there were actual smartphones on other companies that were just as intuitive as Apple, right, Samsung has great phones.

Jordan Wilson [00:40:23]:
Right? There's a lot of great phones outside of Apple, but I don't know. Maybe it's just because the Apple interface is so stupid easy. When I pick up, you know, a Samsung phone or whatever, right, if I'm at, like, Best Buy and I'm just scrolling through, I'm like, I don't even know how to use this thing. Right? I don't know. Maybe Apple does that on purpose. Maybe they make all of their iPhone users. They just make it so easy that it seems impossible to pick up and use, a non Apple device. Maybe maybe that's what we're, you know, looking at here.

Jordan Wilson [00:40:53]:
Alright. Our last piece of AI news. OpenAI has launched its newest model, GPT 4.5. So this is the company's latest and largest AI language model, offering improved writing skills, better world knowledge, and a more refined conversational experience. So, GPT 4.5 is Apple's largest model yet describing it as the most knowledgeable model to date. It is now available as a research preview for ChatGPT Pro users. Alright. So with a broader access rolling out in the coming weeks.

Jordan Wilson [00:41:31]:
I do assume it'll probably be by about mid March, that ChatGPT Plus users, will have access to this new model, GPT 4.5. I do know that other, you know, third party providers, such as Perplexity, Po, and others. So if you have a paid subscription to something like a Perplexity, Poe,U.com, etcetera, right, you can probably start using this 4.5 model in very limited capacity right now, and you don't have to wait for OpenAI to roll it out to other tiers. But, they did say that it will be out to most pro tier or or sorry, most paid tiers in the coming weeks. But right now, if you are using, you know, chatgpt.com, right, so if you're using chatgpt on the front end, you are only gonna have access to 4.5 if you are on the $200 a month pro plan. So, speaking of EQ earlier, that's where this model shines. And, yeah, I said, I don't really need it. But, you know, in my use in my use so far of JPD 4.5, I do see, the benefits of having a model that just feels more natural, more intuitive, and just more human.

Jordan Wilson [00:42:44]:
Right? Because that's the thing. OpenAI straight up said this is not a frontier model, right, which was kind of surprising. And they did say, hey. The like, the big thing, I broke it down in two words. Right? I would have liked Apple or or sorry, OpenAI to break it down this way. So you can go listen. I cover this in, episode, four seventy two, which I believe was on Friday. I said what they're trying to do is make it more reliable and more relatable.

Jordan Wilson [00:43:12]:
So here's what that means. On the reliability side, OpenAI shared some benchmarks and metrics that showed the hallucination rate is going down, and essentially its ability, its knowledge rate is going up. So, it is much more reliable than past models like GPT four o or even their reasoning models, you know, the the o three, you know, o one, o one pro, etcetera. So it is more, reliable, which is huge. Right? That's one of the main reasons that I think a lot of companies and individuals don't even jump in, to these models to begin with because they don't they feel they can't trust them. So it's not hallucination free. Right? But it scored much higher on some of OpenAI's benchmarks in terms of just getting things right, and hallucinations are drastically down. So that's number one.

Jordan Wilson [00:44:01]:
And then number two, it's more relatable. Right? Sometimes when you're speaking to chat and GPT, right, either verbally, or just typing. Right? Let's just say typing because right now the voice mode is still powered, by GPT four point o, not by the new 4.5, but it does feel more human. And so here's the thing. If you're if you're the type of person, that likes to use, ChattGPT as as a friend, as a as a life coach, you know, as a therapist, something like that, this is a no brainer. Right? Especially when it rolls out to the $20 a month chat GBT plus plan. You're gonna love it. For everyone else, but where I've actually have started to see some value in having this, you know, equally smart EQ, large language model is as a business strategist.

Jordan Wilson [00:44:48]:
Right? That's something I use large language models for a lot. And, you know, I've I've noticed that OpenAI's, latest model in GPT 4.5 does a much better job sometimes of picking up on nuances of what I'm trying to say, but maybe might not be saying. Right? Sometimes I might just be giving, chat g b t just a ton of data and like asking it for suggestions. Right? Asking it for strategies. One thing you should always be using for I don't care what model using. You should always be using, models to second guess yourself. Right? To fight back, on a decision that you're making because if you do that, I think your decision either you are gonna have to defend it and make it even better, or you're gonna be thinking about things that you maybe weren't thinking about before. In that case, GPT 4.5 runs laps around everyone else.

Jordan Wilson [00:45:37]:
So in certain use cases, I think it's fantastic. Traditional benchmarks, this thing was a meh. Right? Literally, there's I mean, yes. The benchmarks, across the board improved from their GPT four o models, but this thing did not bench off the charts, which I think a lot of people were expecting. Right? But here's the other thing with this model. This is a foundation for future models. Right? So in the same way that anthropic is going to this hybrid approach. Right? They're essentially combining, transformer models with reasoning models.

Jordan Wilson [00:46:13]:
OpenAI has said that is their future as well. So when we get quote unquote, when we get GPT five, it's going to be a hybrid model like Claude three point seven SONNET is right now. And I think that's the thing that people are overlooking. This is not opening eyes said this. They're like, this is not a Frontier model. This is not supposed to be benchmarking off the charts. It is a new fresh model that understands humans, which I think is huge because I think the future, reasoning models, even though they may not get a name. Right? Essentially, OpenAI said, yeah.

Jordan Wilson [00:46:49]:
In the future, they're just all gonna be one model, but the reasoning models, right, are going to be exponentially better in future versions of OpenAI's offerings because of this stronger and much more capable, GPT 4.5 model. That's how these models are built. Right? The the o series models were built on four o. Right? So now when you think when you have a much more human 4.5 model, imagine what that means for the future of these reasoning models or a hybrid model. So I think it's going to be, extremely impressive. So like I said, following the launch for pro users, GPG 4.5 will be expanding according to OpenAI and kind of their, release announcement. It will be going to plus in team users in the coming weeks, so that could be as soon as this week. I'm guessing it might be next week.

Jordan Wilson [00:47:47]:
Essentially, OpenAI said, yo, we're out of GPUs. We can't serve this thing, which I thought was, interesting. Right? And also, the API pricing on this is wild. It's wild. Right? You know, we were kind of, you know, complaining or, you know, rolling our eyes that Claude three point seven SONNET, didn't reduce their prices. Right? But the price for GPT 4.5 via the API is astronomically high. Right? $75 per million input and then a hundred and 50 per million output. So that's wild.

Jordan Wilson [00:48:30]:
So compared to GPT four o alright. So, we're going GPT four five, 70 five dollars per million input. GPT four o, $2.50. Right? That's wild. And then on the output side, GPT4.5, a hundred and $50. Alright. GPT4O10. So 15 x more expensive on the outputs.

Jordan Wilson [00:48:56]:
On the input, I think that's what? Like, 25 x or 30 x more, expensive. Yeah. 30 x more expensive. So the API prices are out of this world. So I'm guessing, maybe OpenAI may reduce that, once they're able, they said that they were trying to secure more, GPUs. I'm sure the costs are gonna go down eventually, but, you know, maybe they're saying, hey. Right now, there's people that are really gonna, there there's companies and customers out there, I'm sure, that are going to find value in that relatability and reliability, of that new model. So, wow.

Jordan Wilson [00:49:34]:
It is mind bogglingly expensive. Alright. Let's quickly very quickly re clap, those top stories for the week. So first, Anthropic has released Claude three seven SONNET, the world's first AI hybrid model. Google cofounder Sergei Breen, is pushing for harder work for Google to win the race to AGI, reportedly asking, employees to work up to sixty hour weeks, maybe more. Meta is reportedly developing a standalone AI app to compete with OpenAI and Google. Oh, funny by the way, Sam Altman responded to that on Twitter and said maybe we'll make a social media app. Microsoft Copilot crushing it, you know, offering free users, free unlimited voice and advanced think deeper features that is, using OpenAI's o one model.

Jordan Wilson [00:50:27]:
Eleven Labs has launched Scribe, a stand alone speech to text model that supports more than, 99 languages. Then Amazon is denying reports that, Anthropics AI is powering its new Alexa plus features and saying that it's their own internal models. Next new story, the Internet is going berserk over the new Sesame, AI, chatbot that offers, voice capabilities. Me, it's okay, but go check it out for yourself. It's free to go ahead and try. Apple has announced a $500,000,000,000 US investment plan, including a AI server factory in Texas that will reportedly, create 20,000 jobs there. A Bloomberg report shows that we might not get Apple's truly modern Siri until 2027. And then last but not least, OpenAI has launched GPD 4.5, a model that really emphasizes, relatability and reliability.

Jordan Wilson [00:51:32]:
Woo. That was a lot of AI news, y'all. I hope this is helpful. If so, please share this. Right? I know some of you everyday AI, it's like your secret. It's your cheat code. It's it's how you're the smartest person in AI at your company. Please share the love.

Jordan Wilson [00:51:49]:
Right? People are always like, hey, Jordan. How can I help? Click that repost button. That helps. Right? If you're listening on the podcast, please, follow the podcast. Leave us a rating. Tell someone about it. Right? You can send individual episodes. Please share this as much as you can.

Jordan Wilson [00:52:04]:
I know AI can be tricky. It's hard to keep up with. It can be scary. I spend and our team spends countless hours trying to keep you up to date so you can grow your company and grow your career confidently. Speaking of that, if you haven't already, make sure you go to youreverydayai.com. Go listen to our twenty twenty five AI predictions and road map series. That's episodes four forty three to four forty seven. They are, I'm telling you, bangers.

Jordan Wilson [00:52:31]:
Bangers. Alright. So thank you for tuning in. Please go subscribe, to our newsletter at youreverydayai.com. Thanks. We'll see you back tomorrow and everyday for more everyday AI. Thanks, y'all.

Midroll [00:52:46]:
And that's a wrap for today's edition of Everyday AI. Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rating. It helps keep us going. For a little more AI magic, visit youreverydayai.com and sign up to our daily newsletter so you don't get left behind. Go break some barriers, and we'll see you next time.

Gain Extra Insights With Our Newsletter

Sign up for our newsletter to get more in-depth content on AI