Resources:
Join the discussion: Ask Jordan questions on AI
Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup
Connect with Jordan Wilson: LinkedIn Profile
Try Our Free AI Prompting Course: Register for our free Prime, Prompt, and Polish AI Course!
Meta's MovieGen: Revolutionizing Video Content Creation
Meta, formerly known as Facebook, introduces an AI-powered video tool, MovieGen, that could play a significant role in video content creation. It can generate 16-second videos from just a single text input, allowing for precise edits like object replacements. While a public release timeline isn't available yet, this tool may revolutionize the way small and medium businesses create high-quality marketing content.
OpenAI: Investing in the Future of AI
OpenAI is showcasing remarkable growth in the AI industry, securing a further $4billion in credit, taking its total liquidity to $10billion. After securing a record-breaking round of funding, OpenAI's valuation currently stands at $157 billion. The funds will be invested in vital areas like research, infrastructural expansion, and talent retention.
Microsoft's Wave 2 Copilot: Facilitating Real-Time Teamwork
In the bid to promote real-time teamwork, Microsoft announced collaborative features in their Wave 2 Copilot. The introduction of in-line editing features by major tech giants like OpenAI, Google, and Microsoft may significantly impact business processes, calling attention to the benefits of using advanced models like GPT-4.
AI News Recap
OpenAI and Microsoft have both unveiled new tools and features in recent times, pushing boundaries and laying the groundwork for an AI-heavy future. These updates promise to revolutionize a host of industries by delivering AI that's faster, smarter, and more interactive.
OpenAI's Canvas: Enhancing Coding Capabilities
OpenAI's recent introduction of Canvas, a new coding and content tool, has caused quite a ripple in the AI industry. Making coding and content creation simpler and faster, Canvas democratizes access to cutting-edge technology. The tool allows for easy code conversion across programming languages and provides opportunities for code review, debugging, and enhancements.
Google's AI Developments: Making Strides in AI Reasoning
Google has been developing its own AI reasoning models, giving major competition to OpenAI. Gemini series models and enhancements to Google Lens and AI search point towards an AI future with enhanced reasoning capabilities, more immersive shopping experiences, and improved search functions.
Market Positions: OpenAI, Google, and Beyond
As OpenAI and Google continue to develop and unveil powerful tools and features, they're setting the path towards a future where AI will be a powerful, collaborative tool. They're not only enhancing their own product offerings but also constantly redefining what's possible within the realm of AI.
Anticipated Rollouts from Microsoft
Microsoft is set to rollout new features for users this fall, including an enhanced AI-powered photographic memory feature and an improved voice chat. These advancements are expected to offer a more comprehensive AI experience and make the daily operations within businesses more efficient and interactive.
NVIDIA's Open-Source AI Model
Recently, NVIDIA released an open-source AI model, named NVLM 1.0. With 72 billion parameters, it presents a significant performance boost in vision and language tasks, potentially bolstering the abilities of researchers and developers. The move to release the model’s weights and training code provides a vast field of opportunities, though it also raises questions about security and potential misuse.
Conclusion
In the AI industry, developments taking place today have the potential to significantly shift the corporate and technological landscape. As these advancements continue, it's critical to stay informed to maximize the impact of AI on business operations, customer engagement, and revenue. Whether it's Meta's MovieGen, Microsoft's Wave 2 Copilot, or OpenAI's Canvas, these AI tools and platforms have the potential to revolutionize how we operate in a business environment.
Topics Covered in This Episode
1. Meta's MovieGenAI Video Tool
2. OpenAI Funding and Growth
3. Windows 11 Updates
4. New OpenAI Announcements
5. Google AI Developments
6. NVIDIA's AI Model
Podcast Transcript
Jordan Wilson [00:00:17]:
Is Meta's new movie gen better than OpenAI's, Sora? What's this one feature inside of chat gbt's new canvas mode that everyone's missing? And what's going on with all of these new Microsoft Copilot updates? Speaking of questions, is Google changing the way that search works at least when it comes to AI? Yeah. A lot of questions this week if you've been following AI news and updates from big tech companies. But don't worry. We've got answers. What's going on y'all? My name is Jordan Wilson, and I'm the host of everyday at AI. Thank you for tuning in. Everyday AI, it's for you. It's for me.
Jordan Wilson [00:01:01]:
It's for all of us. It is your daily livestream podcast and free daily newsletter, helping us all learn and leverage generative AI to grow our companies and to grow our careers. So if that sounds like you, if that sounds like something that you're trying to do, and who the heck isn't trying to keep up with AI to grow their company and career? That's like all of us. Right? Well, then for all of us, Mondays are the spot you gotta tune in. Maybe you can't join us every single weekday, Monday through Friday live at 7:30 AM when our livestream goes out, but maybe you can schedule out some time on Monday. Right? You don't have to spend hours each and every day trying to track everything that's happening and say, hey. How does this apply to me, my company, my career? That's what we do for you on Mondays. Alright.
Jordan Wilson [00:01:45]:
So, if you're new here, thank you for tuning in. Make sure if you have not already, please go to your everydayai.com. Sign up for that free daily newsletter. Alright? You know, it's written by me, a human. Alright? So I think there's just so much AI out there, and hey. Who are the humans trying to help us make sense of it? That's me. So make sure you go sign up for our newsletter. Alright.
Jordan Wilson [00:02:07]:
Enough chitchat. Let's get straight into the AI news that matters for the week of October 7th. Here's a big first one y'all. Didn't see this one coming. Meta has unveiled Moviegen, a new AI video competitor. So Meta has just announced MovieGen, a cutting edge generative AI video model that promises to change how we all create videos. So this model is not yet available to the public. And is now only released with some samples in a research paper similar to how OpenAI's Sora video model was teased about 8 months ago, yet we still don't have access to that one.
Jordan Wilson [00:02:50]:
So Moviegen looks to be a groundbreaking AI model that allows users to create, edit, and personalize high definition video and soundtracks using simple text inputs and images, setting a pretty high standard at least according to these samples, for AI video. So Moviegen, well, whenever they release it, will allow users to generate 16 second video snippets from a single text prompt, and also personalize them using just one photo. So some pretty, advanced capabilities. So the system offers precise editing features allowing users to replace objects in videos, that's huge, such as transforming a lantern into a bubble, or swapping a VR headset for steampunk goggles. Yeah. Think of the the applications for your company with this. If you sell a bunch of different products, right, shooting one video, swapping out, you know, images or just running a bunch of different videos with just, you know, slightly different objects in them. So many different applications for this.
Jordan Wilson [00:03:56]:
So the unveiling of MovieGen comes amid a growing trend in generative AI for video creation with competitors like Open AI's Sora, Adobe's Firefly, Luma's Dream Labs, Runway, Pika Labs, Cling, so many others. Right? Every day we wake up, there's a new AI video tool that looks pretty good. But despite its promising features, Moviegen is not yet available to the general public, and Meta has not yet provided a specific timeline for its consumer release. I do believe a lot of these AI video tools, kinda waiting until after the US election, so we might see either late, 2024 or early 2025, for a more tiered roll off. Alright. So safety concerns as well. We gotta talk about that. Right? So the development of all these tool raises a lot of safety concerns about the broad release of such powerful tools, and as they can require some significant processing power, to lead to potential misuse.
Jordan Wilson [00:04:59]:
Right? Meta's research right now on Meta Movie Gen is documented in a comprehensive 90 page paper. So, yeah, we'll link to that in our newsletter. So you can go read that or maybe just use notebook l m to, create a short little podcast on it. Alright. So, hey, for our livestream audience, I didn't say I didn't say what's up to y'all. Thanks for being there. But let me know if you can see these, examples on screen now, and thanks to everyone, for joining us. You know, Tara, Harvey, sabbatical life, Marie, Zane, Fred, Tara, everyone.
Jordan Wilson [00:05:31]:
Thanks for joining. I'm looking at these now. So podcast audience will leave, a link to these, but very impressive. Right? So we have what looks here to be a little girl running on the beach, with a kite. We have another video here. Someone's sipping, coffee based off of a photo. Right? Here's the wild thing. Being able to upload a photo and then being able to generate AI videos in different scenes.
Jordan Wilson [00:05:59]:
Right? So, a lot of these earlier AI video generators were a little limited on features. Right? Because the technology wasn't there. It takes time to catch up, but essentially, you would start with a simple text prompt, maybe upload an image, and you would get a 4 second video. Now with this new meta movie gen, once it does come out, the ability to upload a photo and be able to generate AI video in different scenes, it's mind boggling. Right? Even if you would have told me 6 months ago that this technology might exist, I would have said probably not. That's probably 3 or 4 years down the line, but here we are. So, yeah. Live stream audience, what do you think of these images? Right? So here's another very impressive one.
Jordan Wilson [00:06:44]:
This one looks cinematic. So it looks like it's almost set in a deep canyon. There's a storm brewing. Very, very cinematic. It almost looks like, you you know, a drone is is coming up. So many of these shots. Right? 6 months ago, 9 months ago, you like, you would not think these kind of shots would be possible. Right? I'm I'm almost speechless looking at them.
Jordan Wilson [00:07:09]:
Here's the example of, you know, being able to swap out, different elements in a video just with text prompts. So there's one with a child kind of floating a lantern up in the air, and then the exact same video side by side, but then a bubble floating up in the air. Right? Not having to reshoot, not having to reprompt. This really changes what's possible in terms of, creativity, marketing, and also what smaller companies can do. Right? Big companies, the biggest in the world, you know, your Nikes, your Coca Cola, your your P and G brands. Right? All these big companies that for many decades have been able to bring us, you know, these these great storytelling visuals to help sell their products and services. I think we're finally gonna start to see small and medium sized businesses be able to compete at a big level, thanks to these AI video generators. So, yes, although this is not probably the best video you've ever seen in your life, It's pretty good.
Jordan Wilson [00:08:09]:
Right? It's pretty good, and it does level, the playing field, I think. Jay said, yeah. It even added little additional bubbles. Yeah. Pretty impressive. So, yeah, let me let me know, you know, as we go along. What are your thoughts on this one? I was pretty impressed, but we'll see when or if, right, we actually get access to this and and the cost. Right? Because here we are, like I said, 8 months after Sora from OpenAI was teased.
Jordan Wilson [00:08:36]:
And all we've really seen is research paper and, you know, some filmmakers getting access. But I do think OpenAI's Sora from a visual perspective is probably still the leader, but Cling with their recent 1.5, updates, Runway with their, Gen 3 models, and now Meta with Moviegen, it looks like from a quality perspective, pretty pretty close. I mean, we'll have to see once people start getting access, but, you know, personally, I was fairly impressed. Alright. Let's move on and talk about money. $10,000,000,000 to be exact. So OpenAI has just secured an additional $4,000,000,000 credit line, bringing its total liquidity up to $10,000,000,000 just over the last, like, week. So OpenAI's recent financial maneuvers have reflected its rapid growth and ambitious plans in the AI sector.
Jordan Wilson [00:09:36]:
So OpenAI has just last week closed a significant 6.6 record breaking $6,600,000,000 funding round, putting its valuation to a $157,000,000,000 So just days later and just days ago, OpenAI announced a $4,000,000,000 revolving line of credit, which when combined with that $6,600,000,000 fundraising round, which broke records, brings its total liquidity to over $10,000,000,000 So this new line of credit comes in partnership with JPMorgan Chase, Citi, Goldman Sachs, and some other major financial institutions. So OpenAI plans to use this capital to invest in research, it, expand its infrastructure, and attract and retain top talent to maintain its competitive edge. Yeah. OpenAI has been losing a lot of its top talent recently to Anthropic as well as some other competitors. So like I said, this is back to back. This happened about 3 days. So right before OpenAI closed a record breaking, so the largest fundraising round ever at $6,600,000,000 with support coming from big names such as Thrive Capital, Microsoft, NVIDIA, Fidelity, and others. So, amidst all this big funding talk, OpenAI kind of released some more numbers or came out in various, pieces of reporting that their revenue has increased more than 1700% year over year, now generating up to $300,000,000 in monthly revenue with projected sales of $11,600,000,000 for next year.
Jordan Wilson [00:11:22]:
But despite this revenue growth, certain reports have showed OpenAI is maybe losing up to $5,000,000,000 this year due to high training and inference costs, and, GPUs aren't getting very much cheaper because they're getting more powerful, so the prices are going up and being able to retain top talent. So, definitely worth keeping an eye on what OpenAI does with this new $10,000,000,000 in liquidity. Doesn't that sound nice having $10,000,000,000 in liquidity? I mean, sometimes businesses would love to just have, you know, 10,000 or maybe a $100,000 in liquidity. But, you know, over here, OpenAI is, you know, flexing with a big 10:10 Billy. Jeez.
Jordan Wilson [00:13:01]:
Alright. New stuff coming to Windows y'all. So we are now getting more and more updates on the new Windows 11 2024 updates that will bring some new AI innovations to the new line of Copilot Plus PCs. Alright.
Jordan Wilson [00:13:37]:
So the new upcoming Windows 11 2024 update is generating a lot of buzz due to its new AI features. Particularly, some of these are specially designed for those with the new and more powerful Copilot Plus PCs, which utilize advanced neural processing units or on device NPUs. Yeah. We went from CPUs to graphic processing unit GPUs to now NPUs, which helps specifically right now, you know, Microsoft and Windows are using these for on device AI or Edge AI for its new Windows Copilot. So the update will introduce new AI capabilities gradually beginning with a, members of the Windows Insider Program. So we'll probably see some demos rolling out of these new features, coming soon. Starting in October and expanding to other pieces of hardware, those, Copilot Plus PCs with Intel Core Ultra and AMD Ryzen. I think that's how you say it.
Jordan Wilson [00:14:46]:
I'm not a PC person, but I I'm probably gonna be soon. But the Intel Core and the AMD, Copilot Plus PCs should start to get these a, these software updates starting in November. So one of the first features to debut will be a pretty controversial, but very highly anticipated, feature called recall. So recall essentially was delayed, was supposed to roll out a couple of months ago, was, delayed because of some regulatory and privacy concerns, but it aims to give PCs a an AI powered photographic memory. So this was initially planned for a release in June, but it was delayed due to, security concerns, and now it will be off by default. So those are 2, you you know, kind of 2 big changes. It was delayed by about 5 or 6 months, and now it will require users to, opt in, which, hey, I would opt in for a computer to remember literally everything that happens on it. Pretty impressive.
Jordan Wilson [00:15:49]:
Right? I think that's the future of AI computing. So some other new features rolling out in these new Windows 11 2024 updates for Copilot Plus PCs. So it has the click to do feature, which aims to enhance productivity by allowing users to perform AI based actions directly from a shortcut menu that appears over images or text. So essentially, the ability to, execute different AI, actions just with a simple click. So not even having to open anything or save anything or go into Copilot. Some other ones. Improvements to Windows search, and will allow users to describe what they are looking for without needing to remember specific file names or search syntax. So that's a pretty big one.
Jordan Wilson [00:16:38]:
So as an example, let's say whether you're looking for something for, business or, you know, your personal life on your PC, being able to search for something and it being able to understand the contents of the file better, including photos. Right? So, yeah, maybe you're at a big conference and you're searching, for certain photos from that conference. You know, before, it just might be a random file name, and it might be kinda hard. But now with Copilot, this new update coming out to Copilot Plus PCs, it will be able to understand the contents and the context, of your photos and files a little easier with this new, update. Also speaking of photos, the photos app inside of Microsoft Windows will introduce a super resolution feature that can enhance low resolution images to up to 8 times, utilizing the NPU for quick processing. Also, Paint. Hey, good old Paint. Paint's getting updated, y'all.
Jordan Wilson [00:17:36]:
It's not that, you know, late nineties Paint, because now it's even getting some Adobe esque, generative AI features such as generative fill, inside Paint. Also, the Copilot itself. Right? So the biz chat and the Copilot chat are getting some updates to the Copilot experience, promising a more personal AI assistant. Some more on that here in a minute. But we gotta talk about what I think will probably be the 2 biggest features. Well, 3 if you count recall, but 2 experimental features, Copilot Vision and Think Deeper. So these will allow users to interact with the web, in with web content and tactical complex questions. So these two features, I think, are going to be showstoppers once they roll out, and assuming they're working.
Jordan Wilson [00:18:24]:
So let's talk about them 1 by 1. So Copilot Vision is essentially co browsing inside of Microsoft Edge with an AI agent. Right? So being able to click the Copilot button within Microsoft Edge, and then this new AI assistant, which I'm gonna talk about here in a second, a live AI agent you can talk to. Right? And say, hey. I'm planning out, you know, this big conference. You know, I keep dropping like, oh, what's this big conference? You know, maybe you're planning out the, Microsoft build conference that's gonna be here in Chicago, in about 6 weeks. And you're saying, hey. I'm planning this out.
Jordan Wilson [00:19:01]:
What do you think of this? Or, hey. Is my outlining is my outline for this session covering everything that our customers need? So, essentially, it's like having a coworker who has access to all of your information over your shoulder and can see what you're doing as you browse the web. So pretty cool there with Copilot Vision. And then Think Deeper. So Think Deeper is essentially OpenAI's o one reasoning model. So the strawberry model, is coming to Copilot. So we don't know when these two new features, Copilot Vision and Think Deeper, will be rolled out. But, at Microsoft's event 2 weeks ago, they did say it would be rolling out soon.
Jordan Wilson [00:19:43]:
So I do believe that some users will get access, here yet this fall. So, like I said, those two features have not fully rolled out, but one feature that has. Yeah. Michael, thanks for the, comment here. When do they roll out? You know, we don't know. Sometime this fall and, you know, it's it's gonna start rolling out to Windows Insiders first, and then it's going to be slowly tripping out. Some of these features are going to be going to people who have a Copilot Pro license first. You know, so if your company has Microsoft 365 Enterprise, you might be getting these, you know, before the rest of the world if you are just on a free, Copilot account.
Jordan Wilson [00:20:27]:
So yeah. A lot of this, you know, depends on what country. I believe US users are part of that first wave. And if you have a paid account or if you're just using, the free version of Copilot that, that is going to impact when you are going to get these features. And again, slow rollout. So I assume that, you know, we saw a lot of these releases, that I'm gonna talk about now start to be rolled out last week. Still not fully rolled out. So I would assume I would think, that Microsoft wants these features probably out in the wild before the build conference.
Jordan Wilson [00:21:01]:
Like I said, which is November 19th here in Chicago. So we'll see. But two things that did roll out. Well, Microsoft debuted a new Copilot, what people are calling the v two interface, as well as its new advanced Copilot voice chat. So we did a couple of dedicated reviews on our YouTube channel. So if you didn't catch those, we'll share those in the newsletter today. So 2 pretty new features here. So for our livestream audience, I have them, screenshots of them left to right.
Jordan Wilson [00:21:35]:
So the first one you'll notice is a completely revamped and new Copilot interface. So if this looks kind of similar to Inflection AI, well, it kind of is. So Microsoft essentially aqua hired, so they, you know, did it wasn't a full acquisition of Inflection AI and Pie, but they essentially hired all of their top, people. And now we're seeing this similar kind of card layout. So the new Copilot has a much simplified layout. So I don't know if this is going to be rolling out to Microsoft 365 enterprise users as well. That's a big question I have. I haven't seen anything specific, from Microsoft.
Jordan Wilson [00:22:20]:
I know I got a lot of people from Microsoft, listening and watching, so please let me know if this is going to be rolling out to the new biz chat. But right now as an example, I am a, you know, $20 a month Copilot Plus user, so I have this new, kind of minimal simplified inflection esque, kind of card view of, Copilot chat that has rolled out there on the left hand side. And then also, the new advanced Copilot Copilot voice chat. So again, I don't know if, you you know, this is actually called v 2 or if it's just the new refreshed user interface, user experience. I'm not sure if there's a name for these new updates, but they're pretty sizable. Right? And also, the new kind of live voice assistant is out as well. Again, this is a slow rollout. You might not have access to it.
Jordan Wilson [00:23:11]:
I just got access to it last week, so did a couple of videos. So this new kind of advanced Copilot voice chat is very similar, at least in theory, to OpenAI's new advanced mode. So similar similarly, it does not have access to real time information. That one is important. So if you think that you're going to be talking to this new advanced Copilot voice chat about, like, hey. What's the weather in Chicago? What's the latest, you know, tech news for this week? Talk to me about, you know, Q4 trends in manufacturing and shipping, you know, in the Midwest. Right? You can't do those things right now. I do hope that both, Microsoft Co pilot's new advanced, voice chat and, OpenAI's advanced voice mode are going to get access to these tools.
Jordan Wilson [00:24:05]:
So it's it's kind of we don't have the best of either worlds right now. So we have these old and slow, dumb voice assistants. Sorry. Like Siri and Alexa that feel very archaic. The voice is not that good. It it sounds like a robot. It's very delayed. And then we have these new versions.
Jordan Wilson [00:24:25]:
So we have, you know, Gemini Live. If you're an Android user, you have access to that. But otherwise, you might have access to, OpenAI's new advanced voice mode or this new advanced copilot voice chat, which is very neural. It sounds like a real person. Low latency, so there's not these awkward 3 to 5 second pauses, except right now, it doesn't have access to real time information. So hopefully, we'll see that rolling out soon. Right now, free users have limited access to these new, Copilot updates, but you should have access once it does roll out. Alright.
Jordan Wilson [00:25:03]:
Mister Strawberry in the house. Love to see it. Alright. Let's keep going. Here's one that I didn't personally have on my bingo card. NVIDIA, yes, NVIDIA has unveiled a pretty impressive open source AI model called NVLM 1 point o. Yeah. That's a mouthful.
Jordan Wilson [00:25:25]:
So NVIDIA has made headlines by releasing its NVLM 1 point o family of large multimodal language models, which includes the 72,000,000,000 parameter NVLMD 72 b. Y'all can hey. Maybe I should reach out to my people at NVIDIA. Can we give this one a name? That's so hard to say. Right? Like, hey, GPT 4, GPT 4 o. That that's even not the best. That's a mouthful Nvidia, but hey. I guess they make up for it because this is an open source model that is very impressive.
Jordan Wilson [00:26:02]:
And this model is designed to also compete with the proprietary closed models, from companies such as OpenAI, Google, and Anthropic, as well as competing with kind of the leader right now in the open source, open weights, whichever one you wanna call it, model from Meta. So this new n v l m d seventy two b, my gosh, showcases exceptional performance in both vision and language tasks, significantly enhancing its text only reasoning after its multimodal training. So researchers are reporting that this new multimodal variation achieves an average accuracy increase of 4.3 points on key text benchmarks compared to its predecessors. So pretty big, pretty big uptick here. Also, by making the model weights, that's huge, by making the model weights publicly available and promising the release of training code, NVIDIA is breaking away from the trend of kind of these closed advanced AI systems offering pretty unprecedented access to this cutting edge technology for both researchers and developers. Alright. So the open access to these powerful AI models, though though also, like I talked about earlier, it does raise concern. Right? So it's also great, I think, for for researchers, developers, and for the general public to have access to the to the weights of this model.
Jordan Wilson [00:27:26]:
Right? But, also, when you do that, you also give all the bad guys and all the people that don't have good intentions. Right? Because I think especially big companies like Meta and NVIDIA do have good intentions, right, when building these open source models with NVIDIA here announcing that they're, releasing the the the weights in the training data. That's great. But there's also a downside, an ugly downside to this is you give all of these people with bad intentions access to the most cutting edge models in the world. Right? So, I don't think that can be overlooked, but, you know, it's it's it's really a much deeper conversation that I'm not gonna get into today when we just bring you the AI news that matters. We're here to report the facts, but regardless, huge, pretty unexpected news from NVIDIA, not just saying, hey, we're entering the large language model in the multimodal. Right? It's having vision capabilities. So they're not just saying, hey, we're here to enter, the arena so to speak, but they're coming in the open source, open weights, and it's benchmarking not above, but it's benchmarking similar to some of the heavyweight models such as, Google's 1 point, Gemini 1.5, such as, OpenAI's GPT 4, such as Claude 3.5 SONNET.
Jordan Wilson [00:28:46]:
So it's benchmarking below these, but already in that class, which is extremely impressive from NVIDIA. Alright. Let's keep going. We still got even more news. So we have this whole thing called Dev Day from OpenAI. Right? So, OpenAI unveiled some major API updates to enhance developer experience and also reduce cost. So OpenAI had their dev day last week, much smaller than the livestream, much hyped dev day of last year, where we saw the introduction of GPTs in the GPT store. So this year, OpenAI introduced a couple of, I'm not gonna say these are dorky things, but they're a little dorkier.
Jordan Wilson [00:29:30]:
So for the everyday person, this might not be, applicable for you. But if you are someone that goes into, OpenAI's kind of playground, their back end, and you're, playing around with their API, their, assistance API, some pretty big news here. So number 1, they introduced model distillation which allows developers to enhance smaller models, as an example, like gbt4 o Mini, by fine tuning them with outputs from larger models. So a lot of companies, specifically Meta, have introduced this, but essentially using these big powerful models, to help train or distill, and create better smaller models. So that's pretty big there. Also, OpenAI said that they're gonna make it rain on everyone, just raining tokens. So to support developers, OpenAI announced that they're offering 2,000,000 free training tokens daily. Jeez.
Jordan Wilson [00:30:23]:
On GPT 4 O Mini, and 1,000,000 on GPT 4 O until October 31st. That's spooky good. So making it easier to get started with model distillation. Y'all, you you you got 3 weeks for essentially free. Right? It's it's unless you're an enterprise company, it is hard to get to that 2,000,000 or 1,000,000, free daily tokens. Also, following Anthropic's lead, this one. Right? You gotta tip the cap to them, but Open Eye OpenAI did announce the a new prompt caching feature, which enables developers to reuse common prompts at a reduced cost, applying a 50% discount for prompts essentially that share lengthy prefixes prefixes which could lead to pretty significant savings, especially for big enterprise clients. More, OpenAI has expanded capabilities with its vision fine tuning, allowing developers to enhance GPT 4 o's understanding of images, which can improve applications in visual search, object detection, and medical image analysis.
Jordan Wilson [00:31:24]:
So, yes, being able to fine tune with, gbt 4 o on the vision side, pretty big. But the biggest of all, I think the biggest announcement from OpenAI's Dev Day was the introduce was the introduction of the real time API. Okay. So this allows for more efficient right now, more efficient speech to speech applications, eliminating the need for multiple processing steps and reducing latency. So essentially what this means, we saw this, new advanced voice mode finally released about 10 days ago to the general public, but that was just within OpenAI system. So now, this is rolling out to developers. So what that means for all of you out there, there's probably, whether you know it or not, 100 of a big SaaS products, big enterprise tools that run off of OpenAI's GPT for technology. Right? Probably 1,000, but I would say there's 100 of big brand name, companies out there, software as a service companies that run off of GPT 4.
Jordan Wilson [00:32:27]:
This changes everything. So you know how just about every single company in the world now has a GPT 4 enabled chatbot on their website? Now think of that with real time, with the real time voice assistant. Okay. So as an example, being able to to log on to your, you know, your Mailchimp account or your, you know, your analytics dashboard, your banking account. And instead of, you know, being able to type with a GPT 4, powered, assistant, being able to talk to one. So this obviously has enormous, enormous, implications on customer service and customer experience. Right? Because it is, in theory, not that hard to do, and it's in theory not that far off to where most of the services that you use are probably going to be using this real time API fairly soon. So businesses, you need to adapt.
Jordan Wilson [00:33:26]:
You know, you also need to see is you need to have the conversations now. Is this something that we want to bring to our customers? Essentially bringing your data, right, so with some rag, fine tuning one of these models, and using this real time API assistant. Also something key here, this wasn't a real time voice assistant either. You know, so maybe this was a a nuance or I'm reading between the lines here, but this real time API assistant is just real time. It's not talking about just voice. So in theory, I do think this lays the groundwork for real time video in the future as well. We're still waiting for that from OpenAI. They did demo that back at their, spring event.
Jordan Wilson [00:34:09]:
What was that? Back in, I think April or May, but we have still yet to seen that be released. So I do think that OpenAI, very strategic play here, releasing this real time API to bring this voice to voice, real time voice mode. Right? So like we just talked about earlier, you you know, with what Copilot has in Microsoft, the advanced voice mode right now, extremely impressive right now inside of chat gbt minus the fact that it doesn't have access to your data, real time information, and other tools that large language models need. But this is an agentic future y'all, where I think it is going to be very common. Whatever services, that you use on a a daily basis, there's a good chance that they're going to be using this real time API. So, it's definitely worth keeping an eye on. Alright. Also worth keeping an eye on, Google, our next piece of AI news.
Jordan Wilson [00:35:03]:
So Google has developed some new AI reasoning models to compete with OpenAI's strawberry or o one model. So according to reports from the Wall Street Journal, Google is advancing its artificial intelligence efforts to match the reasoning abilities demonstrated by OpenAI's latest model, o one, previously, called strawberry, previously called QStar. Whatever you call it, it is raising the stakes in this kind of agentic AI race now. So according to reports, multiple teams at Google's parent company, Alphabet, are reportedly making strides in developing AI reasoning software that can effectively tackle complex, multistep problems in areas such as mathematics and programming. So Google is utilizing a common method called chain of thought prompting, but on the front end, to bring more, kind of agentic or reasoning capabilities, to the Gemini series of models. So Google has faced pressure to innovate quickly, especially since the launch of OpenAI's ChatGPT and all of these new features we've been talking about, which kind of raises concerns among investors about Google's dominance. Just not just in with its Gemini model, but also in search. Right? Because as you bring all of these real time product, products, I mean, you have also OpenAI is another product they floated out there that we're still waiting for is, search GPT.
Jordan Wilson [00:36:36]:
Right? But a lot of pressure on Google. And despite moving cautiously due to ethical considerations and public trust, Google has made some notable advancements in this field, including the introduction of alpha proof in alpha geometry 2, which excelled at the IMO or the International Mathematical Olympiad, which is, usually a benchmark for reasoning models or, you know, kind of this chain of thought thinking, that models are now starting to do. At a recent developer conference, Google previewed an AI assistant also called Astra capable of using a phone's camera to interact with the environment and answer questions, hinting at future integrations with maybe this new reasoning model, which would be a pretty big pretty big piece of news. Another big piece of news from Google, Google has unveiled a major, change to how its new AI search works. So Google has announced some pretty significant updates to its search capabilities, leveraging advanced AI technologies to improve user experience and information discovery. So let's first talk about Google Lens. So Google Lens is used for nearly 20, according to Google, 20,000,000,000 visual visual searches each month, showcasing the growing reliance on this kind of computer vision visual search technology. Also there's a new, the introduction of a generative AI search in Lens, which allows users to ask questions by pointing their cameras resulting in AI generated overviews that provide relevant information and links.
Jordan Wilson [00:38:19]:
Also with the new AI powered lens search, users can now search with video. Wild. Enabling them to record moving objects and ask questions about them, enhancing this kind of interactive search experience. Voice input, obviously. Right? If you have video input, you think voice input Yeah. So voice input will also be available on Lens in the new Lens search or Lens feature within the new AI search, allowing users to ask questions verbally while taking photos, making the tool more intuitive and user friendly. So Lens has also been updated to facilitate shopping. Yeah.
Jordan Wilson [00:39:00]:
That's a big one. Providing detailed product information, reviews, and pricing from order over 45,000,000,000 items in Google's shopping graph. That's not all. So finally, we're seeing some updates to the new AI powered circle to search feature, which allows users to identify also songs they hear. So it's not just for products on Google Shopping, but coming for Shazam now so you can, identify songs you hear without switching apps, expanding the functionality of Google's mobile apps. Also, Google is rolling out AI organized search results page, starting with recipes and meal ideas, which aim to present diverse content formats and perspectives more efficiently. So, yeah, this seems very similar to what Microsoft Copilot is, unveiling and what with its pay Copilot pages, and we've also seen that with Perplexity pages as well. Oh, a lot there.
Jordan Wilson [00:39:53]:
We're not even done. So the updated design for AI overviews also, we talked about this on the show last week, but worth recapping, includes more prominent links to supporting web pages, improving traffic to these sites, and making it easier for users to find relevant information. One more thing, Google is also testing ads within AI overviews. Yeah. People are always wondering how are they gonna make money. Right? If they're not driving clicks, if people are just getting answers, we've known this for a while, but Google is now starting to test more widely ads inside of its new feature there. Last but not least y'all, here we go. Canvas.
Jordan Wilson [00:40:36]:
We saved some of the juiciest stuff for last. I had to take a sip there, y'all. I'm actually double fisting right now. I got I got coffee and water, because there's been so much news. But we're saving probably, I think, the biggest one for last, and I think people are really overlooking this. Big sip of water. Yeah. This is unedited, unscripted y'all.
Jordan Wilson [00:41:03]:
I like to say and people have said this is the realest thing in artificial intelligence. Everything else is so edited, polished, scripted. Yeah. You hear me slurping. That's me. Alright. Here we go. OpenAI has introduced Canvas, a new tool for coding and content creation.
Jordan Wilson [00:41:19]:
So OpenAI has just unveiled Canvas, a significant update to ChatGPT that enhances coding and content capabilities right within the editor. And this is kind of a direct competitor to Anthropic Claude's popular artifacts feature. I'll tell you in in ways that it is and ways that it's not. So right now, Canvas is available to all paid ChatGPT plus users, and is its own dedicated mode when you start a new chat. That's important to know. You will have a new feature that says GPT 4 o with Canvas. So if you want to take advantage of this, you have to be in that dedicated mode. So Canvas allows users to convert code from one programming language to another with just a few clicks, making it easier for developers to work across languages such as JavaScript, PHP, TypeScript, Python, C plus plus and Java as well as others.
Jordan Wilson [00:42:18]:
So the new feature includes a lot of things for developers, coders, whatever you wanna call it, including some tools for reviewing code, adding debugging logs, inserting comics, fig comments, fixing bugs, and a lot more which obviously improves things in programming. Also, users can highlight specific sections of ChatGPT's text responses to focus and enable the AI to provide inline feedback and suggestions. That piece is huge. Alright. So OpenAI has developed some new core behaviors for its g b t four o model to support Canvas, including generating specific contact types and making targeted edits. So OpenAI has emphasized that Canvas represents the 1st major major visual interface updates since Chad GPT's launch 2 years ago. I don't agree with that. I'll tell you why here in a second.
Jordan Wilson [00:43:17]:
With plans for ongoing improvements based on user feedback. So like I said, OpenAI has already rolled out this to most paid users, but also some free users may get very limited, access to this new canvas feature once it exits beta. Alright. But here's the thing that I think people are sleeping on. I think people are really just focusing on the coding aspect of this and just comparing it to Anthropic Claude's or sorry, Anthropic Claude's artifacts feature. But I honestly think they're 2 different things. And let me also call this out. People are just like, oh, you know, ChatGPT is just copying artifacts from, Claude.
Jordan Wilson [00:43:59]:
Well, not necessarily. I think there's huge benefits to Claude artifacts that we don't see right now in this new Canvas feature. Number 1, artifacts can render code. Right now, Canvas cannot. Alright? That's that's huge. Right? So essentially, you can, you know, create a website inside artifacts and actually render it. You can see what it looks like. Right? You can say, hey.
Jordan Wilson [00:44:25]:
Make me a, you know, website with HTML, CSS, and JavaScript. I wanted to do this, this, and this. Here's a screenshot of it. Go build it, and you can actually render it. You can see how it looks. So you don't have that right now in Canvas. So that is, I think, one of the biggest differentiators, at least that people are saying, oh, this is just for coders. But I don't think, Number 1.
Jordan Wilson [00:44:45]:
And also, let's be honest. This is not I don't think. I don't think this is OpenAI just straight up copying Anthropic Cloud, because guess what? Guess what OpenAI has had essentially now for a year. They've already had this interface, this kind of split screen interface where you talk to an AI on the left and it builds you something on the right. So this was not a new feature from, Anthropic with their new artifacts. The only thing that was new in that, which is super impressive, is the ability to render code. Right? But this whole kind of chatting on the left hand side and, the AI building something on the right hand side, that was open AI first with their GPT builder. People are overlooking that.
Jordan Wilson [00:45:30]:
Right? You can chat with the GPT builder on the left side of the screen all the way back to November of 2023, and it will render it on the right side of the screen. It'll make your GPT for you. So people are just saying, oh, OpenAI blindly copied. Not really. And I do think that it's different. But let me just go ahead and dive into some of these, I think, bigger differentiators right now because, yes, I do think that, anthropic in the artifacts feature has some huge benefits that you don't have inside of Canvas, mainly, being able to render code. But this is for so much more than coders. I think I'm actually gonna have a dedicated show on this tomorrow.
Jordan Wilson [00:46:12]:
So, livestream audience, podcast people, if if you're in the email newsletter, I'll probably put out a poll, but do you want this show tomorrow to be on Canvas? Because I think there's a lot of tips and tricks that people are overlooking. So number 1 is inline. Right? So being able to, essentially inside ChatGPT, we haven't had this yet, and you can't really do this in, in artifacts either. But you can now work in line. So what that means on the left hand side, and I have a screenshot here as an example, you can still talk to ChatGPT on the left hand side. Then on the right hand side as an example, you can highlight a big chunk of text and you can tell ChatGPT to like, yeah, like, hey, make this more attention grabbing. Right? I did this as an example. I highlighted an intro paragraph.
Jordan Wilson [00:46:57]:
I said make it more attention grabbing, and then it's going to update the text right there in line. The other thing is it is much more like a Word document, like a Word editor. You can even go in there and start typing in that in line editor, which again, you do not have the ability to do that inside of artifacts. That is huge, y'all. So people don't know this. A similar feature to this, like highlighting and replying to the chat has actually already been available for like 5 or 6 months, but this just makes it much more intuitive in the canvas editor. And again, being able to integrate and interact and build with chat gbt, the canvas mode in real time, that is the differentiator y'all. This is, I think, where you start to say, okay, this is more of an augmented intelligence.
Jordan Wilson [00:47:46]:
Right? We're always talking about this difference between, you know, human intelligence, AI, you know, artificial intelligence, and then augmented. Well, I think this is a great step in the right direction forward, which is collaborating. Because before, even with artifacts, you couldn't truly collaborate with the AI. Right? You just have to have it regenerate something, render the code, etcetera. But now, you can collaborate in real time with ChatGPT. There's a lot of other, I think, pretty cool features here. You know, on the left, I I don't really like the add emojis button, but when you do highlight something, you have, some new, kind of, interface items that you can work with inside the canvas editor. So emojis, there's a button to add polish, you can adjust the reading level which that is huge.
Jordan Wilson [00:48:31]:
You can adjust the length with an easy slider making it longer or shorter. You can have click a button to have ChatGPT suggest edits, and then you can apply those edits. That's huge y'all. So kind of, you know, going back to this, you know, Copilot vision and having an AI kind of working with you side by side. Right? If you and ChatGPT collaborate on something, and then click this suggest edits, I love that. And then also the thing that I think people literally aren't talking about here, is ChatGPT right here just killed. Killed a bunch of rappers. Right? A lot of these rappers, you know, essentially are thinly designed, kind of variations of chat gbt.
Jordan Wilson [00:49:14]:
And I've said this, go back in the archives. I've been doing this show for 18 months, probably oh, no. More than that now. But I've been saying, once one of these big companies, OpenAI, Google, Quad, whoever, once they actually bring in this in page editing, this in line editing, I'm like, that's the end of so many of these, you know, pretty popular and I would say useful, GPT wrappers essentially, because I think one of the main advantages of them is you can use them like a word doc. Right? You can say, okay. I'm gonna write this paragraph. Okay. Now chat gbt, you know, you you take the next 2, and okay, we're gonna collaborate on this last paragraph.
Jordan Wilson [00:49:57]:
You haven't had that, which is kinda crazy to think, right? I guess, you know, Microsoft just announced that with their wave 2 copilot, features. Kind of this, more collaborative too which I'm super excited about being able to collaborate and do this in real time with teammates with copilot pages, but now we have that with Chat GbT. And I think this is the one biggest feature that people number 1, they're sleeping on because they're looking at this new canvas just for just for editors, just for coders, right, just for developers, which, yes, there's some great features built in there, for, you know, software engineers, for people who are in development, people who are coders, you know, for lack of a better term. But I think this is for everyone. This ability to work in line with ChatGPT and still take advantage. That's the other thing. This is the GPT 4 o model. So you still have advantage to all these other tools.
Jordan Wilson [00:50:54]:
Browse with Bing, advanced, data analysis, all of these other things where, you know, if you're using the o one, the agentic model, you don't have access to these things. So this is huge, and I don't think it can be overlooked. Alright. This was a long one y'all. Thank you thank you for tuning in. I'm gonna do the world's fastest recap on the AI news that matters here. So, number 1, meta unveiled movie gen. Very impressive and it looks to be on the same quality as Sora, not released yet.
Jordan Wilson [00:51:24]:
Our next news story, OpenAI secured a $4,000,000,000 credit line, upping its current liquidity to $10,000,000,000 after its record breaking $6,600,000,000 fundraising round last week. Windows 11 getting some 2024 Copilot updates, especially some new, updates for those with the Copilot Plus PCs, such as the Rewind feature, Copilot Vision, Think Deeper, a lot of other, kind of new features. Then we also saw the new kind of Copilot V2 interface, and advanced Copilot voice chat, for those, Copilot Plus or sorry, Copilot Pro users. Out of nowhere, we saw NVIDIA unveil an open source model in its NVLM 1 point o. So not just an open source model, but releasing the weights pretty wild. OpenAI Dev Day. So tons of new updates there from model distillation, free training through October 31st, vision fine tuning, but I think the biggest one there is the real time API. Then we talked about Google, is reportedly working on a reasoning model according to the Wall Street Journal to more closely compete with OpenAI's new model, 01 preview, 01 mini, kind of the strawberry model.
Jordan Wilson [00:52:45]:
Then Google, so many new AI enhancements to its search, functionality and inside the Google app. And then last but not least, OpenAI introduced Canvas, a new tool not just for coding but for content creation. And I'm excited about that one y'all. Alright. That's it. Thank you all for sticking around. You know, Marie and Tara, David, everyone else, Philip, thanks for sticking around. So much going on in the AI news.
Jordan Wilson [00:53:12]:
We're gonna be recapping it all in our newsletter. If you haven't already, please make sure to go to your everydayai.com. Sign up for that free daily newsletter if this was helpful. I know this was a lot y'all. A 50 minute podcast. Yeah. But we just saved you hours every single day. You don't gotta drown every single day trying to figure out what's going on.
Jordan Wilson [00:53:30]:
Just tune in on Mondays. We do this almost every single Monday, bringing you the AI news that matters. If this was helpful, if you're listening on the podcast, please subscribe whether you're listening on Spotify or Apple. Please leave us a rating and subscribe, and please join us tomorrow and every day for more everyday AI. Thanks y'all.
