Episode Categories:
Resources:
Join the discussion: Got something to say? Let us know here
Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup
Connect with Jordan Wilson: LinkedIn Profile
Try Our Free AI Prompting Course: Register for our free Prime, Prompt, and Polish AI Course!
Revolutionizing AI in Business: Key Insights from the Latest Developments
In the rapidly evolving world of AI, business leaders can gain significant advantages by staying abreast of the latest technological advancements. We delve into some ground-breaking updates from the AI domain, including NVIDIA's hardware breakthroughs, the enhancement of AI models' capabilities, and integration features that promise to transform industries. Here's what you need to know to leverage AI for your business success.
NVIDIA's Game-Changing AI Supercomputers
NVIDIA has unveiled a line of personal AI supercomputers, the DGX Spark, and DGX Station, designed to bring immense AI capabilities directly to end users. These systems empower developers, researchers, and data scientists with advanced desktop solutions, enabling them to run neural networks and AI applications locally. The DGX Spark, akin to a compact powerhouse, offers up to 1,000 operations per second, making it perfect for prototyping and refining AI models. Meanwhile, the DGX Station, with its staggering 784 gigabytes of coherent memory, allows for the handling of more complex tasks. Enterprises seeking to maintain data privacy and efficiency will find these systems transformative as they support the latest open-source models, offering businesses an edge in internal AI deployment.
The MCP Integration: Bridging the AI Communication Gap
The recent adoption of the Model Context Protocol (MCP) by Zapier and Microsoft signifies a new era of interconnected AI services. MCP enables seamless integration of AI models with external apps and tools, allowing businesses to automate tasks and gain insights from thousands of applications without the need for complex coding. For decision-makers, this translates to increased operational efficiency and a competitive advantage. By simplifying how AI assistants communicate across platforms, organizations can ensure more fluid data integration and enhance productivity.
Public Sector Innovation: AI in Government Operations
Pennsylvania is pioneering the use of AI in government services through a pilot program incorporating ChatGPT to streamline operations. This initiative, which reportedly saves workers up to eight hours per week, illustrates the potential of AI to enhance government efficiency by simplifying job descriptions and data management. These early results underscore AI’s ability to minimize administrative burden, freeing human resources to focus on tasks requiring a more nuanced approach. For private enterprises, this successful model of public sector AI adoption provides a blueprint for integrating similar techniques to boost efficiency.
Innovative AI Tools: Adobe and Microsoft Collaboration
The integration of Adobe's AI tools into Microsoft 365 presents a new horizon for creative and marketing teams. This collaboration enables users to perform complex tasks like generating high-quality visuals and data analysis directly within familiar Microsoft applications. Moreover, agents like the Adobe Marketing Agent can help refine audience targeting, paving the path for more personalized marketing strategies. This seamless integration could mean reduced complexities and lead times in content creation, allowing businesses to maintain a competitive edge in dynamic markets.
Enhanced Communication with New OpenAI Models
OpenAI’s release of new speech models offers businesses enhanced transcription and text-to-speech capabilities. These models, building on the renowned Whisper technology, provide superior performance that can revolutionize customer service, internal communications, and user interactions across languages. By incorporating these advanced models, companies can ensure highly efficient communication platforms that foster better customer experiences and faster internal workflows.
Insights and Opportunities
These advancements in AI technology signal profound shifts that business leaders must navigate to stay ahead. From NVIDIA’s supercomputers for on-premises AI applications to Adobe and Microsoft’s productive junctions and the adoption of MCP for broader integration, these innovations provide transformative possibilities across industries. Embracing these technologies could lead to significant operational efficiencies and elevated competitiveness.
Stay connected with everyday AI developments to strategically position your business for growth and advancement in the AI-driven future.
Topics Covered in This Episode:
- NVIDIA's AI Desktop Supercomputers: DGX Spark and DGX Station
- NVIDIA's AI Data Center and GPU Roadmap
- Zapier and Microsoft Support for MCP
- Pennsylvania's AI Pilot Program with ChatGPT
- Adobe and Microsoft Collaboration on AI Tools
- OpenAI's Sora Unlimited Access for Paid Users
- Apple's AI Integration with Wearables and Cameras
- Federal Government's Generative AI Chatbot by GSA
- Google's New Features in Gemini: Canvas and Audio Overviews
- OpenAI's New Voice Models: GPT 4 o Transcribe
- Claude's Real-Time Web Search Capability
Podcast Transcript
Jordan Wilson [00:00:17]:
NVIDIA changed the future of AI computing at their GTC conference. I was there. I'll tell you what that means. Claude finally gets the Internet, but is it too little, too late? The federal government is rolling out their own AI chatbot. I'll tell you what that means. AI services are going all in on MCP. ChatGPT got new voice models even though it had the leading voice models, and Gemini added two pretty small new AI features inside Gemini that I think are actually going to be really big. There was a ton of AI news this week as just about every week.
Jordan Wilson [00:01:00]:
But I don't want you spending hours every single day, like, going down rabbit holes and being like, what what is all this? What does it mean for my company, for my career? Don't do that. I do it for you. Alright? So welcome to Everyday AI. What's going on y'all? My name is Jordan Wilson, and I'm the host of Everyday AI. This is your daily livestream podcast and free daily newsletter, helping us all not just keep up with AI, but how we can use it to get ahead to grow our companies and our careers. So if that sounds like what you're trying to do, you are in the right place. Almost every single Monday, we bring you the AI news that matters. Yeah.
Jordan Wilson [00:01:39]:
I do this every single day. So on Monday, I'm like, y'all, here's here's what you need to pay attention to. Here's what doesn't matter. Right? And then hopefully give you some good advice that you can take that back to work and be the smartest person in your, company or in your department when it comes to AI. So, I'm excited to go over some of these huge stories. Love to see our livestream audience in the house. If you have any questions, get them in. I'll try to get some at the end if you do have any.
Jordan Wilson [00:02:07]:
But thanks for joining, you know, Samuel and and Sandra from YouTube. Brian joining us from Minnesota. Parimi, thanks for joining from India. Sandra, Renee, Marie, Fred holding it down in Chicago just like me. Alright. Love to see it. Alright, guys. So, as a reminder, if you haven't already, please go to youreverydayai.com.
Jordan Wilson [00:02:28]:
I don't know if you knew this, but, you know, the podcast and the livestream, it's one thing. That's where you learn what's going on. If you want to leverage this, you do that in our newsletter. So make sure you go to our website and sign up for that as well as you can go and listen to, like, 400 and, I don't know, 85 now, episodes, of Everyday AI, all for free. You can go watch it. It. You can go listen to it. You can go read about it, all on our website sorted by category.
Jordan Wilson [00:02:54]:
So no matter, where you're at in your AI journey, our website is going to be your best friend, your BFF. Alright. And as a reminder, I am gonna be talking a little bit about, NVIDIA, to start off the show actually, but make sure you check out in our newsletter. Even though the GTC conference, has wrapped up, you can still access everything online for free, for a limited time. So I'm gonna have that link in today's newsletter, so make sure you go check that out. So, yeah. Thanks again, to NVIDIA for, partnering with the Everyday AI Show. We're actually gonna have a couple more, fantastic, interviews, this week.
Jordan Wilson [00:03:32]:
Yeah. Had so many, that we couldn't even do them all last week. So, with a telecom, leader, a health leader, a a Dell leader, you know, so many good and new shows, coming to you that, were recorded from GTC. I'm still putting them together, actually. And we might have a kind of a March Madness style AI startup tournament, which should be pretty cool. So, livestream on these. Let me know. Do you do you wanna see something like that? I I I talked to, eight different, AI startups, and I was thinking of, you know, I just recorded their little five minute pitches.
Jordan Wilson [00:04:07]:
I think it's all kind of tools and, you know, services that many of you all could use. So, let me know. Yes or no? Should should I bring a a tournament style, kind of AI startup pitch competition, and have you all vote for the winners? So let me know. Alright. Enough chitchat. Let's get to the AI news that matters for the week of March 24. A lot to go over, today y'all. Let's get to it.
Jordan Wilson [00:04:35]:
So NVIDIA, has announced a groundbreaking new suite of AI desktop supercomputers, the DGX Spark and the DGX Station. So these were announced at their NVIDIA GTC conference, and they are designed to empower developers, researchers, and data scientists with local AI capabilities. So, during his keynotes, NVIDIA CEO Jensen Wong, introduced two new personal AI supercomputers, during the keynotes. So the DGX Spark and the DGX Station, and both are powered by the Grace Blackwell platform. So these systems are tailored for running neural networks and AI applications locally. So the DGX Spark, this one was actually just kind of an upgrade and a refresh because previously this was the Digits system. So now Digits is DGX Spark and that features the GB 10 Grace Blackwell super chip delivering up to 1,000 or sorry. Yeah.
Jordan Wilson [00:05:41]:
Let me get this right. 1,000 operations per second. So, yeah, a thousand tops. So that's for AI tasks making it a compact yet powerful tool for prototyping and refining AI models and running local AI models. And then you also have the big boy, the bigger version of this which is brand new and was just announced. That is called the DGX station. And that is a more advanced system. So, you you know, essentially the DGX spark, it's like, you know, it's like a big hockey puck.
Jordan Wilson [00:06:12]:
Right? You know, if you've known these, you know, if you know the Mac mini or something like that, it is that size. The DGX station, is more something it is kind of a plug and play and, modular system, but, you know, you can get the full desktop version of this thing as well. And this thing, ready y'all, has 784 gigabytes of coherent memory. That that that's wild. Right? Like, I thought, like, a couple of years ago, you know, it's like when I got a laptop with, like, 16 gigabytes of memory, I'm like, oh my gosh, I'm in the future. No. This thing has 784 gigabytes of memory. So, yeah, like, in terms of what you can run locally, right, because that's that's what this is all about.
Jordan Wilson [00:06:59]:
This is all about giving, you know, power users and everyday people, I I I think as well, the ability to run, you know, state of the art open source models on your computer. Right? Because proprietary models like, you know, ChatGPT and and Gemini and Quadanthropic, you can't download those and run those. Right? But there's some fantastic open source models, like Meta's, Llama, MetaLama. Google actually has an open source model that I think is very impressive in their Gemma three, Mistral's models, even, NVIDIA's own, NemoChon based on Llama. There's so many very capable models now. Right? So the, as well as, you know, DeepSeq for those of you DeepSeq fans, I'm not a big fan. Right? But the the gap between, you know, proprietary cloud models and, open source models, it's it's down to, like, almost nothing. Right? So, the fact that you can now, run so the DGX Spark, is having a price tag of $3,000.
Jordan Wilson [00:08:05]:
I don't believe we have pricing yet, on the DGX station, although, you know, that could drop in the next couple of hours, and, you know, I will double check that when we put out today's newsletter. But, I mean, y'all, this allows anyone in your organization. Right? I think one of the biggest reasons some companies, you you know, haven't got it, you know, haven't gone all in on the, you you know, large language model, the AI bandwagon, is because they don't understand, number one, data privacy and security. Right? Again, I'm not gonna go off on a a a rant, but it's like, hey. If you if you use cloud storage, right, that's the exact it's it's more or less the exact same thing as if you're, you know, uploading documents to a large language model. Anyways, this allows anyone to run AI models locally, very powerful ones at, you know, speeds you can actually use. Right? So on my not so powerful computers, you know, I'll I'll download and and run, you you know, llamas, but they're, you know, smaller versions and it's kinda slow, but still, y'all, like, me, like, being able to even run a a a a llama mile on an airplane when there's no Wi Fi. Right? Like, it's it's something extremely powerful to have.
Jordan Wilson [00:09:14]:
So I think this is, pretty big as well as, you know, major PC manufacturers including Asus, Dell, HP, Lenovo, Supermicro will produce and sell these systems as well. So it's not just right, oh, you just have to buy it from NVIDIA. All the major players are gonna be coming out with DGX, you know, Spark and Station equipped, you know, PCs. I actually have a great interview coming up, with a leader from Dell talking a little bit more about what this announcement and what these capabilities actually mean specifically for enterprise users. Right? That's y'all, like, I can't get over the amount of power and how small these things are. Right? Even three or four years ago, I mean, you couldn't fit all of this compute in in a room. Right? Like, you literally couldn't. And now it is, you know, in theory, the DGX Spark, you you know, fit in your the palm of your hand.
Jordan Wilson [00:10:11]:
Right? Jensen Huang literally had it in his hand. And the, the the station, the DGX station, a little bigger, but, you know, you can still carry it around. And that thing has 784 gigabytes of RAM. So there's there's very few, local open source models that you cannot run on that thing. I'm I'm pretty I'm pretty impressed. Yeah. Mohammed just said, wow. Yeah.
Jordan Wilson [00:10:35]:
Yeah. I agree with that one. You know, Marie says some cutting, edge technology. Alright. Our next piece of AI news also from the NVIDIA conference. So NVIDIA did reveal their plans for the AI data center and just their GPU roadmap through 2028. So, yeah, if if you don't follow too closely, essentially, NVIDIA powers the AI industry. Right? Their GPUs, you know, all the companies essentially use them to train, their models.
Jordan Wilson [00:11:07]:
So, you know, if you love using any generative AI, there's a good chance that NVIDIA's GPUs were, used at some point or actively being used by these companies, to run and train their models. So also, NVIDIA announced updates to the Rubin platform launching later this year, that will significantly boost performance delivering 3.6 axaflops of FP four compute. Yeah. So that's probably above, you know, my head and many others. Right. But, yeah, it's stinking powerful, as well as, NVIDIA talked about their upcoming Blackwell b 300. So Rubin will interdict will introduce next generation HBM four memory technology, which just means faster interconnects and greater bandwidth to support increasingly complex machine learning models. Also, there is the entirely new Vera CPU that will accompany the Rubin GPUs, replacing Nvidia's Grace CPUs.
Jordan Wilson [00:12:08]:
Alright. So faster CPUs to accompany, GPUs. Also, Rubin Ultra is slated for 2027 and will take performance to very new levels, with a new rack configuration featuring up to 576 GPUs. Yeah. So, unless you're running the IT department, this might not necessarily pertain to you personally. Although this is, like I said, all of the AI systems that we're using will be benefiting, from NVIDIA's, new and refreshed GPU roadmap. So NVIDIA also, talked about the growing demand for AI factories capable of handling vast amounts of data and compute intensive tasks. So, looking beyond, Rubin, NVIDIA also teased its next GPU architecture named after physicist Richard Feynman suggesting even greater advancements in compute density and efficiency for 2028.
Jordan Wilson [00:13:07]:
So, yeah, we should we should see, like I said, some, some Rubin updates, in later this year, then, some VERA, Rubin updates in, 2026, Robin Ultra 2027, and then the Feynman GPU series coming in 2028. Alright. Our next piece of AI news, two pretty big companies are going all in on MCP support. Alright. Don't worry. I'm gonna break that down and tell you what it means. But, Zapier and Microsoft have both announced support for MCP, which is the model context protocol integration. So MCP was developed by Anthropic, and it's an open source protocol designed to facilitate the integration of AI models with external data sources and tools.
Jordan Wilson [00:14:01]:
So MCP is a protocol that enables your AI assistants to securely connect with thousands of apps and perform actions such as sending messages, scheduling events, and updating records with any other complex coding tools. So think of it, it is sort of like an API. Right? It's technically, I believe, a layer on top of an API, but this just allows all of these different AI tools, to talk to each other. So, you you know, keep an eye out for what other big companies, start offering MCP support. So on Zapier's side, so Zapier MCP connects to over 8,000 apps without complex integrations. So Zapier's MCP enables AI assistants to perform real world tasks like sending messages, managing data, and scheduling events across 8,000 apps in 30,000 actions. Alright. And then Microsoft introduced modelcom context protocol or MCP support in Copilot Studio.
Jordan Wilson [00:15:08]:
Pretty pretty interesting. So building, your own low code or no code AI agents. So Microsoft has launched MCP in Copilot Studio, enabling seamless integration of AI apps and agents with just a few clicks. So MCP simplifies connecting to knowledge servers and APIs allowing real time data access while maintaining enterprise security like virtual network integration and data loss preventions. So on the Copilot Studio side, users can access prebuilt MCP enabled connectors in the marketplace, dynamically add tools to agents that you build, and reduce maintenance effort through automatic updates. Are you still running in circles trying to figure out how to actually grow your business with AI? Maybe your company has been tinkering with large language models for a year or more, but can't really get traction to find ROI on GenAI. Hey. This is Jordan Wilson, host of this very podcast.
Jordan Wilson [00:16:12]:
Companies like Adobe, Microsoft, and NVIDIA have partnered with us because they trust our expertise in educating the masses around generative AI to get ahead. And some of the most innovative companies in the country hire us to help with their AI strategy and to train hundreds of their employees on how to use GenAI. So whether you're looking for chat g p t training for thousands or just need help building your front end AI strategy, you can partner with us too, just like some of the biggest companies in the world do. Go to your everydayai.com/partner to get in contact with our team, or you can just click on the partner section of our website. We'll help you stop running in those AI circles and help get your team ahead and build a straight path to ROI on GenAI. So, let me know. Should we be doing a a dedicated MCP show, y'all? It's something it is a newer protocol. To be honest, it's something I'm still even learning about myself, and experimenting with, but, I think and I hope maybe that's what everyday AI is good for.
Jordan Wilson [00:17:19]:
Right? There's no, you you know, there's no experts in MCP technology. It is brand new, but I do think that it's going to be a very important protocol, moving forward in the same way that most businesses, most enterprise businesses nowadays can't run, without APIs or maybe webhooks, right, that just allow all your different apps to talk to each other. Right? So now you have essentially this AI version of APIs. I'm simplifying it there, that allow your different AI tools and software both to talk to each other and to talk to your data. So I do think this is pretty big. It is a little bit technical, but I just think, you know, kind of how, you you know, we've we've had this, these different periods of generative AI. Right? So we have the, you know, the large language models. Right? The AI chatbots, and then we had, you know, rag, retrieval augmented generation, and now we have, you know, agentic AI.
Jordan Wilson [00:18:17]:
Right? I I I think one of those next big steps forward is probably gonna be this MCP protocol. Alright. Pedro says definitely. Joe says I'm down with MCP. Yeah. You know me. Alright. Nineties rap reference for you.
Jordan Wilson [00:18:34]:
Alright. Our next piece of AI news, Pennsylvania's AI pilot program is saving workers eight hours a week according to their governor. So Pennsylvania governor Josh Shapiro revealed some promising results from the state's groundbreaking pilot program integrating CHAT GPT into government services. So, Pennsylvania's CHAT GPT pilot program saved state employees an average of eight hours per week according to early results shared by Shapiro. So launched via an executive order in January of twenty twenty four, the program initially provided 50 licenses for ChatGPT Enterprise and has since expanded to a 75 employees across 14 agencies. So despite nearly half of participants having never used ChatGPT before in this first wave of the state study, 85% reported positive experiences using the tool, underscoring its accessibility and effectiveness. So employees participating in the program reported saving approximately ninety five minutes a day, so just over an hour and a half, which leads you to that eight hour a week figure, allowing them to focus on more complex tasks and direct interactions with Pennsylvania citizens. So specific successes included simplifying job descriptions, which reduced hiring and onboarding times from ninety days to sixty days and consolidating 93 IT policies into 34 for the state streamlining their operations.
Jordan Wilson [00:20:11]:
So roles such as state attorneys and construction project manage managers benefitted from AI assistance showcasing its versatility across different sectors. So governor Shapiro emphasized that AI served as a, quote, unquote, job enhancer rather than a replacer, reiterating the importance of keeping humans involved to ensure nuanced decision making. So the program's first phase will conclude on May 31 with plans to expand access to more employees in the second phase. My take on this, only eight hours per week? I don't know. I don't I don't understand that. I don't understand. Like, anyone that's not saving at least two to three hours a day. Right? If you're not saving two to three hours a day minimum, by using AI operating systems.
Jordan Wilson [00:21:07]:
Right? I think there's a difference, and I think that ChatGPT is really right now the only AI operating system that runs, you know, quote, unquote, in the cloud as an AI chatbot. I obviously think, Microsoft three sixty five Copilot is its own beast. But aside from that, I think Google Gemini will get there. I think, anthropic Claude may get there, but I think right now, Chatt GPT is the only what I would call AI business operating system, and I think they are in a league of their own. But the fact that, you know, these, Pennsylvania state employees are only saving, about ninety minutes a day leads me to believe they need to be trained. Right? And so many organizations. Right? Companies reach out to us, when they want to train their employees. Right? So whether it's, you know, 50 employees, 500, they reach out to us and, you you you know, I'm I'm usually pretty shocked.
Jordan Wilson [00:21:53]:
Well well, first, it's good that companies reach out because this is what we do every day. Right? We live inside of ChatGPT. That is, you know, my personal, in our team's kind of home base when it comes to, you know, AI, you know, large language models or your, you know, AI business operating system. But I can't see a way that most most employees aren't saving at least two to three hours a day. If so, that means your employees don't know what they're doing. Right? I mean, yes. It depends on what their work actually is. It depends on, you know, data access, data security.
Jordan Wilson [00:22:29]:
Right? But at the enterprise system, I I I will say if you use any cloud storage, it's the exact same thing. So, yeah, I'm personally shocked, at how little it was. You know, eight hours per week, means, hey. Those Pennsylvania state employees need some, need a little bit of training like I think most companies do. Alright. Adobe, which I don't I don't get why all these big companies have their, you know, conferences on the same day. Right? Adobe had their conference right in the middle of NVIDIA's GTC. So I think they had some pretty exciting, announcements that maybe got slept on.
Jordan Wilson [00:23:07]:
But Adobe and Microsoft announced a major collaboration to integrate AI tools, from Adobe directly into Microsoft three sixty five apps. So Adobe unveiled their Adobe Marketing Agent and Adobe Express Agent enabling marketers to create content, analyze data, and collaborate without leaving Microsoft apps like Teams, PowerPoint, and Word. So the Adobe Marketing Agent aims to simplify tasks like audience targeting and campaign tracking, while the Express Agent allows users to generate high quality visuals directly within Microsoft apps, eliminating the need to switch platforms. So, like I said, the Adobe Express agent lets users generate images for presentations, social posts, and documents through a conversational interface, and the Adobe marketing agent helps refine audience targeting, pulling insights from Adobe Analytics tools, and creates reports directly within Microsoft apps. So, yeah, if you are a Microsoft organization and you are a heavy, Adobe user, this is going to be huge news for you and just for any marketers. Also, integration with Adobe Workfront streamlines project management and boost collaboration across teams. So it should be interesting, to see what other, big companies start partnering with Microsoft, kind of for these, you know, kind of brand or company specific AI agents. So, yeah, I think in total, and we shared this in our newsletter, in the middle of the week last, last week as well.
Jordan Wilson [00:24:52]:
Adobe also rolled out 10 different kind of premade, prebuilt AI agents as well. So I think pretty pretty exciting news. If you're a marketer, if you, use Adobe, you know, time savings right there. Right? And then being being able, to run all of these, you know, presentations, and and and generate reports, just from within, you know, Microsoft Teams, Microsoft PowerPoint, Microsoft Word, not having to, you know, jump around. I mean, that's a huge boon, for productivity as well. Alright. Our next piece of AI news, nothing new necessarily, but an update that I think is worth mentioning. So over the weekend, OpenAI has made Sora, its AI video tool unlimited for paid users on the $20 a month plan.
Jordan Wilson [00:25:51]:
So this is pretty big because originally, OpenAI's Sora, their AI video generating tool was only rolled out to ChatGPT Pro users on that $200 a month plan. And then plus users got very limited access. So now, OpenAI just quietly took away all limits. So if you are a paid, ChatGPT user on any tier, you should now have unlimited access to Sora. So this is, pretty pretty exciting, I'd say. Previously, they did have that credit system, for both, plus and pro users technically. Just the the pro, limits were much much higher and the, you know, $20 a month chat g p t plus limits were pretty low. Well, why did they do this? Well, I think it's because of competitors.
Jordan Wilson [00:26:43]:
Right? So, right now, I think when OpenAI initially teased Sora, which was now almost like a year ago, I don't think that there was another AI video tool that was in the same conversation. However, it took OpenAI a super long time, to actually release Sora. So here's the trick. Google Google, Veo. I always forget if it's Veo, Veo, even though I talk to Google people about it. Sometimes I just forget things. Veo two, Google's version, Google's AI video generator is much better. Google's version is in a class on its own.
Jordan Wilson [00:27:23]:
However, you do right now have to access it through third party platforms. So, even Google doesn't have it available, kind of as for a front end user. If you just wanna go in, to and and use Google's, Bayo tool, you can't. You have to use either a third party, obviously, a paid service, or use their API which is pretty expensive. But there's been some other great AI video, kind of, companies especially out of China. Cling AI is one, here in The US, you know, Runway. So I like, essentially, I think all of these other AI video generators have gotten so much better in the last year. And OpenAI, even though when they first, you know, kind of teased Sora, right, it broke the Internet because we hadn't seen anything like that.
Jordan Wilson [00:28:08]:
But it took them, you know, what, like, eight months to actually just start releasing it. I do think Sora, and OpenAI has just better you, like user interface, user experience than a lot of the other tools. And there's a lot of these cool, like, remixing features and the ability to build, a storyline, very simply. Right? Piecing together a lot of these different, you you know, AI video generations. So I do think even though Sora is not the best model out there, I do think now between, these more, unlimited usage and, some of these unique features, I I do think now OpenAI is kind of in that one b slot next to Google, Veo, which is definitely in the one a slot. And then I think you have everyone else, kind of, you know, your your Cling, your, Runway, your Pika, your Luma, all of those I think are are right underneath Sora. Although, you know, it's I think I think I think at that point after Google Vale, it's kind of up for interpretation. It's up for your personal taste after that and what you really value in an AI video generator.
Jordan Wilson [00:29:14]:
But y'all, I'm I'm telling you, the technology on the video base is so fast. So, yeah, I actually, believe talk about that with a Dell leader in a conversation that I'm gonna be debuting here pretty soon. We were just talking about it is unbelievable, how far this space has come. So if you haven't really looked at it in, you know, three to six months, maybe you don't think that AI video, is is something that your company, will use, you're wrong. You're wrong. Right? Let me tell you this. The days of, hiring, super expensive, you know, videography companies, Those days are unfortunately dwindling down. I'm not saying that's going to be something that, doesn't exist anymore.
Jordan Wilson [00:30:02]:
You you know, obviously, you're still gonna have your, you know, your high end video production and and creative agencies. But I think more and more small and medium sized companies are gonna be using these AI video tools, right? Because you can also start with an image, an AI image. Right? And now we have these capabilities through the AI image tools where you can upload an image, right? I could have uploaded an image of me, you know, interviewing someone at the NVIDIA GTC conference and, you know, you can use that as, a beginning point. Right? So I could just create videos based on real images. Right? If you have a a a stock photo that you've been using, right, on on your blog post that looks like it's from 1998, you can finally update it and bring some life to it. So, I do think if we if we were having this conversation a year ago, I'd be like, yo, AI video is not for everyone. AI video is for every business. If you haven't already started using it, you need to get in there.
Jordan Wilson [00:30:59]:
You need to understand it because consumers demand video. They want video. And if your company is not already using video, right, you probably already know, you you know, maybe you just don't have the talent on your team. Maybe you don't have the budget. Well, these videos or these AI video tools, are are are really leveling the playing field. Yeah. Good question here, from Joe saying when is OpenAI going to focus its attention on DALL E four? So yeah. Actually, Grok, just added the ability to edit images, and and updated, their AI image generating to be available via an API.
Jordan Wilson [00:31:43]:
I think Google in their Gemini two point o, I might have to do a, like a dedicated episode on this. They kind of are killing Photoshop. Right? You can upload in in inside Google Gemini's, two point o. Let me know, livestream audience or podcast audience. I always put my, you you know, my LinkedIn information, my email, even though I'm a little bit behind. So sorry if you've reached out to me in the last couple of weeks. But the Google Gemini two point o, what you can do with images, it is wild to me. As someone that's used Photoshop for more than, twenty years, what you can do with simple text commands, inside, Google Gemini, two point o, it is mind boggling.
Jordan Wilson [00:32:29]:
Right? You can just upload, you know, an image of yourself and anything that you could think that you could do in Photoshop, you can essentially do in there. Right? Changing what you're wearing. Right? Maybe you're doing a fashion shoot, a product shoot, you you know, maybe there's just something annoying in the background and you don't wanna have to learn Photoshop or run something super, you know, processor, compute heavy, on your computer. You can literally just go into Google Gemini now two point o in AI Studio and just with text prompts. So, I did and the reason I bring that up, Joe, is because, Axe slash Grok, is starting to roll out a similar feature that we saw in Google Gemini two point o as well as we just saw rumors, that chat GPT may be opening this up as well. So that is not confirmed yet. But, you know, there there have been some rumblings on the Internet over the last, you know, twenty four, forty eight hours that ChatGPT is offering, is starting to test out image editing, which would lead me to believe, that they will be rolling out, an improved image generation as well. We know that Sora, OpenAI's video tool, that we just started this AI news piece with, can generate photos.
Jordan Wilson [00:33:40]:
So I I I don't know if the next, you know, version of DALL E might just be called Sora photos or if, as an example, DALL E four may just be powered by Sora. We'll see. But I do expect, some image generation updates, from chat g p t soon, especially y'all. Like, I I I I talked about this on the show, I think, last week. I mean, the ability for Google Gemini two point o where you can, in one shot, you can create, as an example, a blog post. Right? My example was I did a blog post, you you know, the top five tourist attractions in Chicago. I had it write a blog post and then it did, image, for each of those five all in one shot and in line. Right? That's wild.
Jordan Wilson [00:34:25]:
It is wild. The the the the increased capabilities of of these multimodal, kind of AI chatbots. So, yeah, hoping, hoping we'll see that soon from OpenAI. Alright. Let's keep going. A couple more AI stories. This one, not a huge fan of. So Apple is reportedly working on integrating advanced AI capabilities into its wearable devices, including cameras for future Apple Watch and AirPods models, and that's according to Bloomberg's Mark Gurman, who is kind of the, the leader in getting all scoops on all things Apple.
Jordan Wilson [00:35:03]:
So, according to reports, Apple is developing multiple versions of future Apple Watch models equipped with cameras to enhance AI functionality, allowing the device to, quote, unquote, see the outside world. So this aligns with Apple's focus on expanding its visual, visual intelligence technology. So right now, visual intelligence, which currently relies on third party AI models like ChatGBC, is being repositioned to use proprietary Apple AI systems. So this shift could reduce dependency on external AI providers and strengthen Apple's control over its AI ecosystem. So cameras on standard Apple Watch models may be embedded within the display potentially using under display technology or a camera cutout. This would give users discrete access to visual AI features directly from their wrist. So the Apple Watch Ultra, which has more design space or more room to play with, is expected to feature a camera embedded near the digital crown and a side button. This placement would make it easier for those Apple Watch Ultra users to scan objects or interact with the environment using their wrist.
Jordan Wilson [00:36:20]:
Apple aims to bring similar camera equipped AI functionality to future AirPods as well, further integrating visual intelligence across its product lineup. So the release of these AI powered wearables is not expected until at least 2027. And I guess if if if we're, you know, following, the queues here, you know, if reports right now are saying something Apple intelligence related is coming out in 2027, you might as well just tack, like, another three years on that. Right? A lot of things that Apple promised last year in its WWDC, keynote address talking about Apple intelligence have not even begun to roll out yet. Even though Apple was running marketing commercials on a huge scale promoting features that literally are not available, which now they're facing some class action lawsuits on. So, again, I'm not gonna turn this into the, I'm gonna accidentally poo poo, on Apple intelligence and go off on a side rant, but, you you know, I'm not personally a fan of cameras in watches. I don't know why. I still think we need some semblance of of privacy and and trust, when it comes to AI and just when it comes to cameras.
Jordan Wilson [00:37:38]:
Right? As an example, right, I was on the NVIDIA GTC, show floor for all of I think I had, like, six minutes, to spare literally. I was jam packed with interviews. So, you know, I had my, Meta Ray Ban glasses on, which is great. So, but I think when people are wearing those, they can kind of see and understand that you're probably recording something. Right? There's a little status indicator. So I don't know. Like, having this on the watch, cameras on the air on on the AirPods, I don't know. I'm not I'm not a fan of it even though I I I use and I enjoy, Apple technology.
Jordan Wilson [00:38:16]:
Not a huge fan of of throwing cameras and trying to bring AI to my watch. That's just me. Like, first, Apple, get it figured out on the phone. Right? You're, like, thirty years, behind. Yeah. What yeah. Fred said new Apple Watch camera, watch out locker rooms. Yeah.
Jordan Wilson [00:38:33]:
So many, like, just, like, intrusive, like, just, like, red, like, red alarms blaring. It's just like, I don't know. Like, does anyone actually want this? This just seems like a bad idea. So I don't know if Apple is just trying to put out, you know, a bunch of new, you know, AI enabled, you you know, products. I mean, the biggest thing is everyone wants to collect more visual data, to make their AI models better. Because essentially right? I was having a conversation with this at the NVIDIA GTC conference. You you know, obviously, large language models aren't at the point where they've hit the ceiling, in in in terms of data available that they can learn and be trained on. But essentially, I like to say it like this.
Jordan Wilson [00:39:14]:
You you know, today's current models, and by today, I mean, you know, years past models, have essentially been trained on knowledge worker data or, you know, data that you would presumably, you know, read on a screen, text, images, etcetera. Right? The next big frontier for data is real world data, world model data. Right? Data with how us humans interact with the real world. And that is what is ultimately, I think, going to, you know, be that next big leap, in terms of AI. Right? I I talked about that with, the, agility, labs CTO at the GTC conference was a fantastic, conversation if you didn't listen to that one. But, you know, the the the the next big piece is is, you know, companies are gonna try to get more and more data from the real world. So they want all of us essentially out there with as many cameras as possible, you know, training their bottles. Right? So when we talk about humanoids, when we talk about, you know, AGI, ASI, right, like all of that doesn't happen without a ton more data from the real world.
Jordan Wilson [00:40:20]:
Right? So, these, new versions or these newer models can understand, how we interact with the world around us. Alright. Next piece of AI news. The General Services Administration, the federal government, so that's the GSA, has unveiled a new generative AI chatbot designed to improve efficiency and automate repetitive tasks. So the chatbot, which is now available to GSA federal staff, utilizes large language models from companies like Anthropic and Meta to assist with basic tasks including writing. So according to Wired reports earlier this month, the Department of Government Efficiency or, Doge. Right? Is it Doge or is it Doge? I think it's Doge. Deployed a similar chatbot called GSAI.
Jordan Wilson [00:41:10]:
That's not confusing. Deployed that to, 1,500 workers. So the tools released coincides with the closure of GSA's eighteen f digital services team and downsizing of the technology transformation services, raising questions about the future of federal tech innovation teams. So GSA officials clarified that the chatbot is not for replacing jobs, although that's kinda what it's doing. They did say it's not also intended for official agency decisions, and it operates separately from GSA knowledge bases. So safety controls are in place to prevent sharing sensitive information, and prompts are logged but not classified as federal records. So the agency aims to measure the tool's success through adoption rates rather than workforce reductions, signaling a focus on cultural integration rather than immediate cost savings. I don't know.
Jordan Wilson [00:42:19]:
It just seems like to me, right, right now the stance of the federal government and Doge is to get rid of as many federal workers as possible and integrate AI tools in their place. So although they're not really coming out and saying, yeah, we're using AI to replace jobs. That is literally kind of what the US government is doing, with these huge, cuts, across the board, across federal agencies. So you're seeing, you know, thousands, and thousands of humans being laid off, and you're seeing more and more AI tools, being used in the US federal government. Alright. A couple couple more stories. I'm personally excited about this one. So Google has launched a few new features in Gemini.
Jordan Wilson [00:43:11]:
They may be small. You might not have seen them, but I'm personally excited to use them. So Google has unveiled two new AI powered features for paid Gemini users, Canvas and audio overviews. So Canvas, like we've seen from some other companies, you know, ChatGPT has their Canvas tool. Claude has their version of it, which is called Artifacts. So, Canvas inside Google Gemini enables real time collaboration with Gemini for document creation and editing. So users can upload, I mean, whatever documents you might want. So class notes, research, ideas, and have Gemini draft speeches, edit content, or suggest improvements.
Jordan Wilson [00:43:56]:
The feature allows users to adjust tone, length, and other aspects of the text directly. So, yeah, if you use Chat GPT Canvas, I'd say that is, the closest, you know, other AI tool out there, right now to Google Gemini's Canvas. So it also supports coding projects offering an interactive learning experience. So Google highlighted how Canvas can help users learn by creating simple coding projects like a tic tac toe game, complete with explanations and previews to guide learners through the process. Alright. So that's feature number one. Feature number two is audio overviews. So if you use NotebookLM, you are definitely familiar with Google's audio overviews, tool.
Jordan Wilson [00:44:40]:
So, I'm not saying I had anything to do with this, but I do DM the, and talk with the Google team a lot. And I told them about, I don't know, six months ago, I'm like, you guys really need to be rolling out the audio overviews inside of Gemini. I I got kind of a a positive affirmation, so I'm sure it was already on the road map. But I even told them, like, hey. Six months ago, you know, you should be doing this. This this, you know, the audio overviews is amazing. Right? It should be rolled out, across Google's suite of products. So cool to see this.
Jordan Wilson [00:45:13]:
So, audio overviews can transform documents into podcast style audio clips for easy listening. This feature lets users upload PDF files, slides, or research reports and generates conversational audio summaries, making complex information more digestible. So the audio overviews tool builds on Google's previous AI experiments. So like I said, this was initially introduced through NotebookLM, And this is, now available for paid Gemini users in both the mobile and the web platform. So, here's how you use it. So in the new so like I said, you need to be on a paid plan first and foremost. For Canvas, there should be a new Canvas button, where you would normally put in your, text prompt. And for audio overviews, I I'm guessing there's going to be a way to do this a little bit better visually.
Jordan Wilson [00:46:08]:
But the easiest way to do that right now is you can just upload a bunch of documents and just say, you know, please create an audio overview for, you know, these documents. So, you know, I have a little screenshot here, for our livestream audience showing you how to trigger. So it would be cool, and I'm sure Google will, get a visual way, to kind of trigger the audio overviews because I'm sure most people aren't gonna notice. So literally, you can just, you you know, upload one long PDF or a couple text files and, you know, say create an audio overview of that and then you're gonna get that, you know, very cool, you know, kind of podcast style. But you it doesn't have the interactive feature, that you have in notebook l m. So obviously notebook l m, I think is still one of my most used AI tools and there's still benefits to using it, versus the new features inside, Google Gemini. Alright. Two last stories.
Jordan Wilson [00:47:01]:
So first, OpenAI has unveiled three new text to speech models. So those models are called GPT four o Transcribe, GPT four o Mini Transcribe, and GPT four o Mini TTS, designed to improve transcription and text to speech capabilities. So OpenAI, their new models are available immediately through APIs for third party developers and on a new website that they launched, which I think is pretty cool, called openai.fm. So that's a demo site for individual users to test and customize voice inputs. So this essentially you're like, okay. What's this mean? Well, essentially now you can integrate, this new technology into any of your apps. So if you are a developer, software engineer, or a big company, right, this model, it's extremely, extremely impressive. OpenAI was already the leader in this space with their Whisper technology.
Jordan Wilson [00:48:01]:
I love this. Right? When a company is like, yeah. Even though we're the leader in the space and we could sit by and, you know, maybe not even update Whisper for another couple of years, nope. They completely, changed the text to speech game. So, it's almost so good it's scary. But, the model allows users to change accents, pitch, tone, and emotional qualities of AI voices via text prompts. So these models are built on the GPT four o base model, which was launched in May of last year, but have been post trained for superior transcription and speech performance across 100 plus languages. So, yeah, you can do both text to speech, but you can also transcribe audio or speech to text.
Jordan Wilson [00:48:53]:
Right? So the new model boasts an impressively low word error rate of 2.4% in English, outperforming OpenAI's older whisper model, which was already a leader in that category. So OpenAI has introduced streaming speech to text for real time real time transcription. I'm gonna have to build, myself a little app using that to help me ask better questions for my guests on the Everyday AI Show, and that makes conversations feel a little bit more natural. So pricing for the model start at $6 per 1,000,000 audio input tokens and scales down for smaller models. That is the more, powerful model. So this is now essentially OpenAI is competing with Eleven Labs and Hume, AI. And like I said, OpenAI.fm is hosting a very rare, for OpenAI, a very rare public competition to find the most creative uses of its demo sites. So, yeah, we'll be sharing the links and more information in our newsletter.
Jordan Wilson [00:49:57]:
Alright. Last but not least, Claude. Welcome to the twenty twenties, Claude. Claude finally has the Internet. WTF. Alright. So Claude, the AI tool developed by Infropic, now offers web search capabilities, like, five years after everyone else. Not actually, but, you know, multiple years.
Jordan Wilson [00:50:16]:
So the new feature is currently available in The US for paid users only and expands Claude's ability to deliver actionable insights across various industries. So Claude's web search feature enables users to access the latest events, trends, and information by integrating real time data, from the Internet into its responses. So when using web search, Claude provides direct citations, which is important, for sources, allowing users to fact check information easily and ensuring transparency in responses. So web search, like I said, right now, it is a feature preview, so you have to enable it at the bottom of your screen. And there are plans to access, to, there are plans, to roll this feature out to other parts of the world as well as free users in the future. So according to reports, Anthropic is using the Brave browser, for this web access integration as revealed by updates to its sub processor list and identical citations found in both tools. So Claude dropped the ball on this one. The drop it team dropped the ball.
Jordan Wilson [00:51:31]:
Right? Having real up to date information is essential. Right? This is one of the three reasons why, about a year ago, I put out a podcast. I said, don't use Claude. In like, enterprise companies should not be using it, as a front end chatbot. It's different if you're using it on the back end via the API. Right? Because then there's a lot more that you can do to ensure accuracy. You can ensure, right, set up your rag pipelines, all of those things. I still up until this this web search, kinda dangerous.
Jordan Wilson [00:52:04]:
Right? Kinda dangerous if if if you were rolling out, quad access. I mean, unless you were using it for coding, which I think it's it's by far, the best AI model for coding software development, not even close. Right? But for the most part, for everything else, having an offline large language model in 2025 was straight up asinine. Right? I think ultimately, it like, Entropic, it was leading billions of revenue on the table. Right? And, you you know, you you like, you all might think I'm crazy by saying that, but, you know, when I did that show a year ago, I got messages from multiple Fortune 500 c suite people, and they were like, you are spot on. We will not touch Claude on the front end because it does not have Internet access. Because in those instances, you are relying on very old data. Right? You always have this knowledge cut off.
Jordan Wilson [00:52:56]:
Right? Oh, August 2024, you know, October 2024. But that is the best case scenario. So many large language models are trained on these huge datasets, and the data in there might actually be multiple years old. So so many people just blindly trust what an AI model spits out, not knowing it might be pulling in data that's two, three, four, five years old. So, I guess you gotta tip your hat, I don't know, to Anthropic for finally bringing this out. Right? I made a joke. I'm like, this is like if you're a phone manufacturer and you just release text messaging for the first time, it's like, how did you still exist without having this? Wild that, you know, Infopnik just rolled this out now. But with that being said, I probably will use, Claude for some more tasks that I just haven't, touched with Claude for that very reason because it can actually be a little dangerous, especially if you don't know what you're doing, to blindly use Claude for a lot of things, prior to it having web access.
Jordan Wilson [00:54:03]:
So that's at least good, but kind of nutty that we got there in the first place. Alright. Let me very quickly recap the top AI news stories, the AI news that matters, for this week. So first, we talked about NVIDIA unveiling their new AI focused desktop systems GTX, or sorry, DGX, that were available that were unveiled at the GTC conference. Also at the GTC conference, NVIDIA revealed, its GPU roadmap going all the way to 2028. Zapier and Microsoft, now officially, support the MCP, the model context protocol. Pennsylvania just released some initial findings from their AI pilot program where they're rolling out chat GPT access to state workers. Adobe partnered with Microsoft and, is rolling out some Adobe marketing AI agents inside the Microsoft three sixty five Copilot platform.
Jordan Wilson [00:55:06]:
OpenAI quietly made Sora unlimited for paid chat GPT users. According to reports, Apple is planning to bring AI powered cameras into its wearables like watches and AirPods. The federal agency, GSA, has launched a generative AI tool amid concerns over worker layoffs and surveillance. Google has announced, I think, two really good features in their Canvas and audio overviews for paid front end Gemini chatbot users. OpenAI has unveiled three new voice AI models and text to speech models even though their whisper model was a leader in the space. And last but not least, Claude has finally unveiled real time web search capabilities. Alright, y'all. I hope this was helpful.
Jordan Wilson [00:55:59]:
If so, please let me know. Repost this. Alright? It doesn't help. Right? People are always like, oh, Jordan. I've learned so much. What can I do? Share this. Click that repost button. Right? I think so many people are scared of AI and so many people are getting lost, and really not taking advantage.
Jordan Wilson [00:56:15]:
So if you don't stay up to date, you risk falling behind. And y'all trust me because I do this. This is my job. It it it is nearly impossible for you to have a real job and also stay up to date with everything in the AI world and how it's gonna impact how you do your work, how it's gonna impact your company, how it's gonna impact your future career. So we do that for you. So please help me out by telling someone about this. Email your department. Right? If you're giving a presentation on AI, throw our podcast up there as a free resource, you know, on our website.
Jordan Wilson [00:56:50]:
So make sure you go sign up, for our free daily newsletter to recap today's show. But on our website, you can now listen to more than 480 episodes. The podcast, the videos, we put up text recaps as well. It is a free generative AI university. So I hope this was helpful. Please make sure you join us tomorrow and the rest of this week. A lot of, exciting, announcements and shows and interviews that I did at GTC as as well as I think we're gonna do that tournament style thing. I saw some of our livestream audience be like, heck yeah.
Jordan Wilson [00:57:21]:
Bring that. So we're gonna do that. So thank you for tuning in. Hope to see you back tomorrow and every day for more everyday AI. Thanks, y'all.
Midroll [00:57:31]:
And that's a wrap for today's edition of Everyday AI. Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rating. It helps keep us going. For a little more AI magic, visit your everyday AI Com and sign up to our daily newsletter so you don't get left behind. Go break some barriers, and we'll see you next time.
