Resources:
Join the discussion: Ask Jordan questions on OpenAI
Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineup
Connect with Jordan Wilson: LinkedIn Profile
Try Our Free AI Prompting Course: Register for our free Prime, Prompt, and Polish AI Course!
Embracing the Future of AI with OpenAI's GPT-4.5
In the ever-evolving world of artificial intelligence, OpenAI's GPT-4.5 stands out as a significant advancement in large language models. As heralded in the latest episode of "Everyday AI," the new model isn't just about improved technical benchmarks—it's about bridging the gap between technology and humanity with increased emotional intelligence. For business owners and decision-makers, understanding these advancements could be pivotal for strategic development.
A Shift Toward Relatability and Reliability
OpenAI's GPT-4.5 revolutionizes the landscape by being more relatable and reliable than its predecessors. It enriches human-like interactions through enhanced emotional intelligence, which, while difficult to quantify, is vital for personal and business communications. This makes GPT-4.5 an invaluable tool for companies looking to enhance customer service, streamline internal communications, or even build AI-driven content that resonates more deeply with audiences.
Leveraging Emotional Intelligence
One of the standout features of GPT-4.5 is its advanced emotional intelligence, which allows it to understand and interact with human nuances in unprecedented ways. For businesses, this means the potential to use AI as a more empathetic tool in roles traditionally reliant on human intuition, such as HR communications or customer service emails. The AI's ability to craft messages with a human touch was notably demonstrated by comparing responses that showcased empathy and understanding—qualities every leader strives to emulate in business.
A Competitive Advantage in Communication
GPT-4.5 shines in its capacity to produce content that aligns with human expectations and emotional cues. This opens up new avenues for businesses to engage with both internal teams and external clients. By crafting messages that reflect human sensitivity, companies can foster stronger relationships and nurture a supportive corporate environment. Whether you're writing motivational messages or addressing employee concerns, GPT-4.5 offers a sophisticated approach that can enhance corporate culture and employee satisfaction.
Adapting to New Benchmarks
Despite not toppling every technical benchmark, GPT-4.5 has quickly become a preferred model due to its enhanced human interaction capabilities. For business owners, adapting to these new benchmarks involves recognizing and integrating these human-like qualities into their strategies. This could mean revamping client communications, developing AI-assisted content strategies, or using AI for strategic planning. The acceptance and preference for GPT-4.5 among users highlight a shift toward valuing relatable and intuitive AI interactions.
Preparing for the Future with AI
As businesses look to the future, embracing tools like GPT-4.5 can be crucial for staying competitive. The focus on emotional intelligence and relatability signals a significant shift in how AI can be used in everyday business scenarios. Companies that effectively integrate these tools will likely lead in innovation, creating environments where technology enhances human potential and drives business success.
In summary, GPT-4.5 isn't just another upgrade in a series of AI advancements. It's a step toward a future where AI complements human interaction, offering reliability and relatability that businesses can leverage for growth and innovation. For decision-makers aiming to harness the full potential of AI, understanding and utilizing the emotional intelligence of GPT-4.5 will be key.
Topics Covered in This Episode
1. Overview of GPT-4.5
2. Comparison of Models
3. Tools Available in GPT-4.5
4. Demonstration and Analysis of GPT-4.5
Podcast Transcript
Jordan Wilson [00:00:17]:
GPT-4.5 is officially the world's best large language model. But how do we use it? Right? This is something that a lot of people have been talking about for the last few days since OpenAI released its first big updated base model in more than two years. Right? Because two things that we're gonna talk about today and we're gonna do a a live demo as well about how this new model, it's more relatable and more reliable. But that's really just begged the question. Okay. What does that mean for actually how we use it? What is it actually better at? Right? And when we talk about a model excelling in emotional intelligence, I mean, you can't really benchmark that. So how do you actually know when you might want to really take advantage of this new GPT-4.5 model? Alright. We're gonna be answering, hopefully, those questions and a lot more today on Everyday AI.
Jordan Wilson [00:01:25]:
What's going on y'all? My name is Jordan Wilson, and I'm the host. And Everyday AI, it's for you. It is your daily livestream podcast and free daily newsletter helping us all not just understand AI, but how we can use it to actually grow our companies and grow our career. So, yeah, when all these, new models come out seemingly every single week, you might be wondering, is this for my department? Is this for my company? Well, hopefully, at least after today's show, you'll have a little bit better of an idea at least when it comes to GPT four point five, OpenAI's newest model. Alright. So, if you're new here, thank you for tuning in. Is this thing's live? It's unscripted. It's unedited.
Jordan Wilson [00:02:04]:
So, you know, I try to bring you something real in artificial intelligence, which I think is rare nowadays. Right? Everyone's creating all these super polished rundowns of of, you know, models and, you know, using AI, even avatars even. Right? Like, this is real. So, you know, if if you are listening on the podcast, thank you for tuning in. Please make sure to subscribe to the show. Leave us a rating. That would be great. And join the livestream.
Jordan Wilson [00:02:28]:
Right? Yeah. We have real other real humans, you know, hanging out here in the live stream. So thanks for everyone, joining us. Max holding it down, in Chicago. Big bogey face on, on the YouTube machine. Douglas, Woozy, Sandra. Let's see who else. Christopher, Brian, Marie.
Jordan Wilson [00:02:45]:
Thank you all for joining. Alright. So, I am gonna need your help. Alright. I don't know if these comments went on YouTube. Maybe they did, maybe they didn't. Alright. But I listed, 13 different questions.
Jordan Wilson [00:03:01]:
Alright. I want you all to go through. I have them on my screen here. I'll show them in a bit. They're super small. Just write down the number of which one you want. Alright. So, just say, hey.
Jordan Wilson [00:03:12]:
I wanna see number five. You know, I wanna see number seven. Alright. So, livestream audience, I need a little help from you guys. If you scroll through the comments, hopefully, you should see it. I do have a slide up here later. It's super small, but let me know which one, you actually want to hear. Alright.
Jordan Wilson [00:03:28]:
Let's get into it y'all. So if you want, the daily AI news, sorry, go check the newsletter. Didn't have time to put it all together this morning if I'm being honest because I was putting in a lot of work on this show. I don't sleep a lot if you haven't noticed, you know, live stream audience by looking at me. I don't sleep a lot. Yeah. I I I have AI go do a lot of my homework, but, it's a lot. So if you do want the AI news, we're gonna have that in the newsletter.
Jordan Wilson [00:03:53]:
Don't worry. And this is also as an FYI. This is part two. So I specifically you know, I noticed and I heard from a lot of you all. Right? Like, you're like, hey. These shows are too freaking long. Right? I don't need an hour ten minute podcast on a new AI model. So, we actually broke, a bigger show down into two separate parts.
Jordan Wilson [00:04:14]:
So if you want to hear the first part where we went over a lot more of the technical detail, we went over some benchmarks. We went over a lot of those other things for, OpenAI's new GPT four five model. So if you want that, go listen to episode four seventy two. So, you know, you can just scroll like two episode backs, two episodes back. That one is called OpenAI's, new GPT-4.5, what's new and who can benefit the most. But today, we're gonna be looking at a really a comparison. We're gonna be going in and actually, using this model hopefully with some feedback and suggestions, some suggestions, from our livestream audience. But let me just go over some of the basics.
Jordan Wilson [00:04:54]:
Some of these we already covered in the previous show, some things we didn't. So here's, kind of some talking points from OpenAI. So, they are reiterating that GPT-4.5 is a research preview. It is their largest and best model for chat. Alright. For chat. It is a step forward in scaling up pre training and post training. And by scaling unsupervised learning, GPT-4.5 improves its ability to recognize patterns, draw connections, and generate creative insights without reasoning.
Jordan Wilson [00:05:25]:
Alright. Early testers, early testing shows that interacting with GPT-4.5 feels more natural. Its broader knowledge base improved ability to follow user intent and greater EQ, emotional intelligence make it useful for tasks like improving, writing, programming, and solving practical problems, and they also expect it to hallucinate less. Alright. So this is kind of some of the, bullet points that I said on our first show. You know, this is the last non chain of thought model developed by OpenAI. So, OpenAI CEO Sam Altman did say, hey. Future models that are under this GPT five, kind of architecture, it's gonna be a hybrid model.
Jordan Wilson [00:06:06]:
Alright. So keep this in mind. But this is a new essentially a new base model. Alright. And and when we talk about reasoning models like o one and o3. Right? So we might not actually, you know, see an o four as an example. Right? Just like we might not, you know, see, you know, certain minis. So for many.
Jordan Wilson [00:06:26]:
Right? It just might all be under GPT five. We don't know yet. They might say it's GPT five and it uses o four reasoning. Right? But in the future, you're just gonna be working with one model, and that's why this is extremely important. I think a lot of people were kinda like upset. Right? And they're like, oh, this GPT-4.5 didn't break every single benchmark. Right? This GPT-4.5 is extremely expensive in the API. Yeah.
Jordan Wilson [00:06:49]:
I don't know any company that is going to be able to afford, to use this in the API, right, for, like, 75, input in a 50 output per million tokens, which is just, you know, 30 times more expensive, than than their previous g p t four models. So, but I did say after the first show, I said humans are really gonna like it. Right? Because this is, I think, I've always you know, anytime you go chat, quote, unquote chat with any AI model, I don't know. To me, it's never felt human. Feels like you're chatting with a machine. GPT-4.5 is the first time I felt different. Right? To me, it doesn't feel like I'm I'm chatting with a human. I don't know what it is, about me and how I tick as a person.
Jordan Wilson [00:07:42]:
Right? I I know. Right? I yeah. It's like, yeah, I'm talking to a computer still, but it felt like real human. Felt like a real human computer person. Right? I know a lot of people like reading people's first experiences. They're getting enthralled, in in in, GBT four point five's ability to carry on a conversation, and to show kind of these EQ, tones that make humans human. Right? People are always like, hey. What, you know, what separates humans from AI, from humans, from large language models and, widely those things are usually considered things that are EQ.
Jordan Wilson [00:08:18]:
Right? Emotional intelligence, being able to understand nuanced conversation. Right? And right now, at least from a tax standpoint, GPT-4.5 is doing amazingly well. And I did predict that humans are going to like it, and sure enough, humans loved it. Because even though GPT-4.5, and again, this is a preview, even though it did not break every single benchmark ever, right, which is what I think a lot of people were expecting or were hoping, from this model, What it did do in the, LM arena. Okay. So, I talk about this the the the easiest way to think about this is, you know, those, like, blind, like, Pepsi versus Coke tests, right, from I don't know when that was the late nineties, early '2 thousands. Right? Someone goes, there's no label, they drink both, and they say, oh, this one's better. Right? That that's kinda what, Elo rankings are, or the arena score from LM arena.
Jordan Wilson [00:09:16]:
Alright? So what this means is you put in a prompt, you get two different outputs. They're blind. You you know, so you don't know which one's which, and you choose which one is better. So this is at least in terms of how humans actually use a model. Right? Yes. We have dozens of benchmarks that test different things from from coding to, writing to, math to science. Right? So you have all of these kind of, you know, systemized and and organized and categorized benchmarks, but it's always like, well, what about humans? Right? Do humans care? Will humans notice? Well, with g p t four five, the answer is yes. Right? Because, it quickly shot up to the number one spot in the l m arena board.
Jordan Wilson [00:10:01]:
So think of this, you know, every single model out there, is is in this. Right? When you go into this blind taste test and GBT 4 point 5, immediately once once they, got enough votes to rank on the chart, they were number one. So the best model in the world. I do know, a couple hours later, you know, Grok's newest version came on, so I think it's technically, in a tie now. But still, you know, even without smashing every single, benchmark, this new model, g b t 4.5, just elevated itself to, I think, probably the most preferred general use case model in the world, which is extremely important because like I said, in the future, these reasoning models are gonna be built on top of this. Alright. So let's talk a little bit before we jump in live. In livestream audience, I see a couple of you voted.
Jordan Wilson [00:10:57]:
If you could go through, let me know which one, which one you wanna see. I agree with Douglas. Douglas said Jordan needs coffee. That was me sitting on the coffee. But let's talk about a little bit about the model itself and how it performs inside of chat g b t, inside of the chat g b t interface. Also, FYI, let me get this off my chest. Right? Because there's people online, you know, and they're like, oh, I tried this model. It's in you know? And I'm like, oh, okay.
Jordan Wilson [00:11:24]:
How'd you try it? Oh, via via a a third party. You can't do that. Alright. So I I I do have to talk about access because at least as of this hour, GPT-4.5 is only available for pro users inside of ChatGPT. Alright. That does cost $200 a month, But presumably, either by this week or early next week, that will be going to all paid ChatGPT subscribers. So even if you are on the $20 a month ChatGPT plus, you should be getting access, to g p t 4.5. So you might not have access to it now, but I highly encourage you.
Jordan Wilson [00:12:04]:
Yes. There's, you know, third party platforms. You know, if you're on paid versions of other, you know, perplexity or po or something like that, you can probably go use 4.5 in a limited capacity if you're on a paid plan, for one of those services, but that's not the best way to understand a model. Right? You should be using it in its natural environment. So yes, there's also the API that's available that's extremely expensive. Alright? So if you are trying to see what's best for your team and a lot of times, I don't understand why every single big Fortune 500 in America doesn't have at least a Teams or an enterprise ChatGPT account. It's mind boggling to me because, yes, your company can have an internal version, right, that maybe you use for customer support or sales or something like that, but you should all, you know, and I'm not just saying ChatGPT, but you should every single employee. If you are a CEO of a small to medium sized company, if you will if you are a, an IT leader, if you're a CMO, whatever, you need to be pushing for your entire team.
Jordan Wilson [00:13:03]:
Whatever your AI operating system of choice is, you need to also have a full team or enterprise license, whether that's, ChatGPT, Gemini, obviously Copilot. Right? If if if you're a Windows organization, Microsoft organization, Quad, whatever it is. But, because when you are using these models inside the chat interface, they come with a lot of tools. Right? I did a show probably about a year ago. I should update it. You know, kind of like, hey. What needs to happen, you know, for for us to get to something like artificial general intelligence? And and one of the things is number one, a model needs to have access to the Internet. And number two, it needs to also have tool use.
Jordan Wilson [00:13:47]:
Right? So this tool use, this occurs inside of the ChatGPT interface. And, yes, third parties sometimes have versions of some of these tools, but, I mean, the tools are best in the native interface. But right now at least, not all of the tools and features, work with GPT-4.5. So let's go over what does work and what is available versus what isn't because OpenAI didn't say this. I went through and tested it tested it all for you so you know. So, again, whether you have a pro account now or you're gonna be getting GPT-4.5 in the coming days or weeks, here's what's available out of the box. Okay. So projects, you can use GPT four four five in projects.
Jordan Wilson [00:14:29]:
You can use DALL E, you know, the AI image generator, which I don't know why anyone would. Right. It's not that good although it will be getting updated soon. Sora does have photo capabilities for early beta testers, just FYI. And also if you don't know projects, that's essentially where you can organize chats into one folder which is great, but you can also upload documents that that folders chats can access to as well as special instructions. So it's similar to GPTs, a little different. So, GPT four five does have access to projects. It does has it does have access to DALL E.
Jordan Wilson [00:15:03]:
It does have access to ChatGPT search, which is extremely important because actually the knowledge cut off for GPT-4.5 was rolled backwards. So its memory is a little worse or at least it's, the recency in the training data. So GPT four o is June 2024. GPT '4 point '5 is October 2023. So keep that in mind, and that's why it's important, that GPT-4.5 has access to ChatGPT search. You can upload files, to GPT four five, which is a a must. So glad that's there. Also, Canvas mode, one of the most underrated, I think tools or functionality of any, you know, AI, you you know, large language model company out there.
Jordan Wilson [00:15:45]:
Canvas is available. So right now, unavailable, and this is as of the time of, you know, when I checked, nineteen minutes ago. Okay. Right now, tasks do not work with g p GPT four five and GPTs. So those custom small versions, of ChatGPT that you can create, doesn't work right now with 4.5. Both of those things, both tasks and GPTs, those obviously still work with GPT four o. Alright. So let me just boil this out of two things.
Jordan Wilson [00:16:27]:
I wish OpenAI would just put these two words somewhere very large, on their GBT 4.5 page because a lot of people are asking two things that I think really separates the biggest difference between four or five. And this is in my experience so far. It is more relatable, so more human ask, right, that EQ and more reliable. So we went over the reliability a little bit in our first show, going over benchmarks, accuracy, lower hallucinations, etcetera. It just knows more. It knows way more. Actually, there was a, you know, there's a website that does, e sorry, IQ scores, for large language models. And, GPT four five actually got the highest score for a non reasoning model, which is pretty impressive because that was the first time a non reasoning model, performed at the same IQ level as the average human.
Jordan Wilson [00:17:22]:
Right? Which is pretty big. Right? When you have a a reasoning model, it does way better because it uses more compute. But the fact that a non reasoning model in GPT four five scored this high on an offline IQ test. Right? So this is an an IQ test that is not in training data. It's pretty impressive. So it is definitely more reliable, but it is also much more relatable because the emotional intelligence. So, this is from OpenAI, but, more natural human like interactions than GPT four o. It's better at reading and responding to emotional cues, and it is preferred, by users as well against GPT four o.
Jordan Wilson [00:18:01]:
Alright. So we're gonna jump in. We're gonna jump in soon. Alright. I know this is small on the screen. I don't know if these comments posted to YouTube as well, but, I'm gonna go ahead and try to send them again. So livestream audience, I know a couple of, couple of you guys have already voted. I don't know if these comments are coming through.
Jordan Wilson [00:18:19]:
Hopefully, they are, but I have 13 essentially examples. Alright. And I wanna do these live. I wanna maybe do two or three. We'll see how long it takes. And I wanna show you the difference, the difference between, a query in four five and a query in four o. Full disclosure, haven't run any of these yet. Right? I run a ton of tests, but, I'd like to do this live.
Jordan Wilson [00:18:44]:
This is unedited, unscripted. Right? So livestream audience just put the number. Right? Try not to put anything else. Just put the number, and I'm gonna scroll through the comments here, on the right side of my screen, you know, bringing in comments from from LinkedIn, Twitter, YouTube, etcetera. So which one do you actually want to see? So I haven't done these and I'm gonna, read the prompt out. These are very short prompts. Right? They're supposed to be short. I'm not gonna go through the whole, like, prime, prompt, polish process, which if you want the best output, you should be doing, the basics of prompt engineering still.
Jordan Wilson [00:19:16]:
But I wanna show you just hopefully some, short prompts, the inputs and then the outputs. And we're gonna run this in, GPT four five and GPT, four o and talk a little bit about the differences and, you know, hopefully, we'll see the difference. Who knows? Maybe we won't. That's the downside of trying to do, unscripted, unedited demos and and examples, inside generative AI. So, maybe if you're brand new to to ChatGPT, large language models, generative AI, and you don't know a ton, that's fine. We try to keep it simple, but let me say this. Generative AI in large language models, they're generative. They're not deterministic.
Jordan Wilson [00:19:56]:
So what that means, as an example, if you go search for something on the Internet, search engines are, for the most part, deterministic. Right? Yes. There's some, some personalization and some localization, but for the most part, those search results are gonna be roughly the same every single time you put them in. A large language model is complete it is a roll of the dice. It is generated. Right? It is, you know, there's there's some next token prediction. So, you know, you could in theory put the same prop 10 times. You could get nine very different answers.
Jordan Wilson [00:20:25]:
You could get two very different answers. You could get five things that are pretty much the same, but just worded differently. So that's, another thing to keep in mind. Generative AI is generative. Right? Which is why sometimes these, these live demos are are super fun. Alright. I see some, I see some votes here. Alright.
Jordan Wilson [00:20:44]:
I'm looking through here. I'm seeing which are some of the most, some of the most voted ones. So I have 13 different examples on the screen. And and I really focused on, a couple things. So these prompts are supposed to, you know, rely on creativity and intuition. You know? So storytelling, being able to to think and write clearly, strong in design and creative tasks, but they're also really around these four categories where I think EQ shines in a large language model. Right? So, think. So if you are using ChatGPT as a personal or life coach, so some of these prompts are are more in line with that.
Jordan Wilson [00:21:21]:
If you're using it as a therapist, if you're or, you know, a work therapist even, right, to to work through tough problems, tough issues, How do I send an email? Right? Those things, content writer, business strategist, and creative partner. So that's where I think some of the categories where the everyday person is going to really see the benefits of four or five. So let me repeat that. If you're using this as a personal life business coach, therapist, content writer, business strategist, and creative partner. Yeah. There's other things that it's actually gonna perform really well in. You know, I know a lot of people are saying, oh, 4.5 isn't great at coding. It's actually really good, at coding across the mark like, across the board in the LM arena test.
Jordan Wilson [00:22:02]:
It, like, swept everyone in almost every single category. So it is, measurably better in almost every single category that you would use a large language model for. But I think, hopefully, we'll see the biggest improvements in some of these areas. Alright. So let's go ahead. Let's see if we can do this. Live stream audience as always. Please let me know when you can see my screen.
Jordan Wilson [00:22:27]:
We're gonna do this live podcast audience. I'm gonna try not to make this one too long. I'm gonna try to, be somewhat concise. Alright. So as a reminder, if you have a normal ChatGPT plus plan and you log on in today, you're not gonna see this 4.5. Alright? But when you do, I'm guessing within a couple of days to a week or two, this should be rolling out to most paid users. So ChatGPT plus, ChatGPT teams, Chat, e d u, as well as enterprise. I think enterprise might be a little after for, those those companies that hire us, to train their large teams.
Jordan Wilson [00:23:03]:
You might not be getting enterprise in your, PPP biz, you know, training, at least not March. Alright. So we need to select, GPT-4.5 in the drop down. Alright. So last year, Monet, let me know if you can see this. I'm looking to see which one some of our most, popular ones. Okay. Alright.
Jordan Wilson [00:23:27]:
I see quite a few people, and again, both on LinkedIn and YouTube, it looks like, and a couple on on Twitter. Okay. So I'm seeing a lot of threes. I'm seeing, some sixes and thirteens. Let me go up earlier. 512. Okay. 911712.
Jordan Wilson [00:23:52]:
Okay. So I see I see quite a few twelves. Alright. So let's go ahead and grab, number 12, shall we? Alright. Livestream audience, can we see? Can we see the the screen? Alright. Let's do number 12. Let's see what number 12 was. Okay.
Jordan Wilson [00:24:05]:
People people, people were going for humor. I wrote this one, and and chuckled. Alright. So here's what number 12 was. People wanted to see this. Alright. It's saying you work at Microsoft. Write a convincing memo on why Microsoft Zune should be reinstated.
Jordan Wilson [00:24:23]:
Y'all remember Microsoft Zune? Alright. So actually what we're gonna do is we're gonna first do this with four o. I'm gonna read the response, then we're gonna do the same with four five. Alright. And we'll see how long it is. Alright. Now I'll try to do my best to even describe what's going on. So, unfortunately, ChatGPT is being extremely slow right now.
Jordan Wilson [00:24:47]:
So this is gonna be fun. Alright. So GPT four o, is giving us a memo. It is emoji heavy, which I'm not a fan of. And also, I intentionally left these prompts kind of open ended. Right? Because I this I'm doing this for everyday people. Right? Everyday people write in a super short prompt, and they hope to get a lot out of it. And I think I think and I hope maybe, that's where GPT four five shines.
Jordan Wilson [00:25:10]:
Okay. So here's what GPT four o says. Here's a compelling internal MO advocating for the return of Microsoft Zoom. I really wanted to know Microsoft Zoom when it first came out. Alright. So like I said, there's a ton of emojis. It gives a date to Microsoft leadership team from your name, your position, subject, the case for reinstating Zoom. So it's pretty long so far.
Jordan Wilson [00:25:37]:
It's going, a little bit slow. So we might have to cut it off here if we wanna do two or three quick examples without turning this into an hour show. So it starts with an executive summary. It says the Microsoft Zune, despite its untimely discontinuation, was ahead of its time in fostering a seamless ecosystem of music, community, and content consumption. Given the resurgence of retro tech nostalgia, the rise of dedicated music devices, and the renewed focus on hardware specific integration, Microsoft is uniquely positioned to reintroduce Zune as a modern AI powered cloud connected audio experience. This memo outlines how a Zune revival aligns with current market trends in Microsoft's broader strategy in AI hardware and consumer engagement. So so far, executive summary from four o seems pretty pretty good. Little wordy.
Jordan Wilson [00:26:32]:
I don't know that we need to, you know, say things like this memo outlines, you know, depends on, you you know, how you would, communicate to to your coworkers internally. I wouldn't put something like that in an internal memo saying this memo outlines. But other than that, pretty good. So let's go through. So here's, number one, it says the market opportunity. Number two, it says the vision for Zune two point o. Again, we're, seeing seeing some emojis, which for the most part, I don't think we'd see on an internal memo, but that's fine. So, it's going over key features of the Zoom reboot included AI powered music curation, lossless and spatial audio support, Microsoft three sixty five and Xbox integration, cloud sync and offline playback in Zoom social two point o.
Jordan Wilson [00:27:18]:
I know I know we have some, some some listeners, from Microsoft. So, you know, let me know if if if this is gonna hit the hit the, the the shelves anytime soon. Three, it goes over competitive advantage. Four, financial, viability and go to market strategy. So pretty good. It's it has a launch plan, a marketing approach. Conclusion, a strategic appeal. Alright.
Jordan Wilson [00:27:41]:
So I'll just read the intro and the conclusion for both. So conclusion, the Zune brand still holds holds strong emotional value and with the right vision, it can be reintroduced as a powerful differentiator in Microsoft's consumer hardware and AI portfolio. By integrating AI superior audio technology and cloud capabilities, the next generation of Zune could redefine how people explore music and content. Let's bring Zune back, not as a competitor to smartphones, but as the intelligent music companion for the AI driven era. Alright. So not bad. You know, I I I wouldn't, ever, you know, use this necessarily. But, you know, I'm my background, it's it's in content writing.
Jordan Wilson [00:28:24]:
Right? I was a journalist and, you know, more or less, I've been getting paid to write for for two decades. So I wouldn't use this. I'm sure some people would look at this and be like, oh, good enough. Not necessarily. Alright. But let's go ahead and switch models. So, yes, if if you didn't know, at the bottom of any prompt, there's these little icons. Right? So, once you do get access to four five and you're trying to see which model should be my daily driver, I encourage you to run tests like this.
Jordan Wilson [00:28:51]:
You can also do this in the back end, of, ChattiPete's, kind of playground, but it will cost you, on the API side. So, I would probably just wait, if you don't wanna pay those, super expensive prices. Alright. So we're gonna do the same thing with $4.05 now. So I just clicked $4.05. Interestingly enough, okay, this is great. Four five was, I'm guessing, smart enough to know that I would probably want to use canvas mode. So it's, love this.
Jordan Wilson [00:29:27]:
Love this. Right? Super smart, FYI. So it it started answering it in canvas mode even though I didn't tell it to. Four o did not because it probably knows, right, as a memo, It probably knows I'm gonna wanna go back and forth kinda like how I said, I wouldn't use this. It it probably knows that, hey. You're gonna wanna update this. So let me go ahead and throw this in Canvas even though you didn't ask me to. Right? And then it has the, the option here to click to answer in the chat instead.
Jordan Wilson [00:29:58]:
Alright. So let's go ahead and, take a look. We're gonna read, both the intro and the summary much more concise. So, personally, I like that. Maybe it's because I use AI too much, and I think, you know, let me just go on a a small and short soapbox here. Large language models are trained to be helpful assistance. Right? It's in their system prompt. And usually what happens if you don't go through this, you know, what we teach like the refined queue process, right? And that's when you make a large model smarter, smaller, and more specific for all of you that have taken our our PPP course.
Jordan Wilson [00:30:40]:
Right? For the most part, a model doesn't necessarily know exactly everything that it would want to know in order to provide an answer. So this is something I've seen from four five, and I personally appreciate, but this might be a preference thing. GPT four five is much more concise, right, where four o and previous models, they just get super verbose, right? Like and and this is something that large language models do all the time. If you don't give it a lot of detail, if you don't go through a back and forth process of essentially, you know, you having a conversation with, you know, four o or four five about the Zune and here's my role, here's my reasoning, here's my rationale. A lot of times it just is gonna spit out a bunch of overly kind of generic content and it's gonna try to tackle thing from too many different areas, but it's not gonna be super strong. So that is something that I've noticed in my, I won't say heavy use of GPT four five, but I've I mean, I've used it, fairly often. I know it just came out like Thursday. Right? And today's Tuesday.
Jordan Wilson [00:31:43]:
So, you know, a big chunk of that was the weekend, and I was trying to sleep for the first time in, like, a week. But I've I've used it very heavily since it came out, and this is something I've noticed. Answers are much more concise, sometimes more confident, which can be a good or a bad thing. You really just have to have an eye on the output, but it's it's much shorter. Right? Another thing, I see no emojis, which I personally like. Alright. So let's go ahead. So here's the internal memo.
Jordan Wilson [00:32:10]:
So it says to senior leadership team from your name, strategic product development, date, subject. Alright. So let's at least read the overview, the conclusion, and then talk about what else, is in here. And, hey, livestream audience, let me know which one also you think is better. Alright. So overview. Given current market trends and the growing nostalgia for early digital media devices, I propose that Microsoft just strategically reinstates the Zune product line. Originally discontinued in 2012, the Zune has since garnered a devoted fan base in iconic status, becoming emblematic of a simpler digital age.
Jordan Wilson [00:32:52]:
Alright. So then it goes over market opportunity. It outlines strategic benefits, a lot of, similar things that we saw from '4 o. Right? So, brand strengthening, ecosystem expansion, talking about integrating it with, Microsoft three sixty five, Xbox Game Pass, Azure cloud services. That's cool. Differentiation, the implementation approach, going through initial market research and validation, modern product development, marketing and launch strategy as well, and then the conclusion. So, reinstating the Microsoft Zune is a strategic decision with significant potential upsides. It positions Microsoft at the forefront of a profitable retro tech revival, reinvigorates consumer interest, and strengthens brand affinity.
Jordan Wilson [00:33:38]:
I look forward to discussing this further and obtaining your insights. So, I don't know if this in this case, if there's one, you know, definitive winner and one definitive loser. For me, my personal taste would be the second one. Right? An internal memo I mean, four o's version, super long. Right? So if I, if I x out of canvas here, if you didn't know this at the bottom, you'll see now one of two and two of two because I use that little model switching icon. So I can just toggle left and right and compare, the different outputs. So, for me, I much prefer the second version. Alright.
Jordan Wilson [00:34:19]:
Let's see. It looks like some other people wanted, number one. So let's do number one. I think this one might be a good example as well. Alright. So we're gonna launch a new chat. We're gonna start in four o. We're gonna do four o first and then we're gonna do four five second.
Jordan Wilson [00:34:35]:
Alright. So this one, a couple of you voted for write an email to a colleague who missed a deadline due to a family emergency. Alright. So we're doing GPT four o first. So subject line, hope you're doing okay. Hi, colleague's name. I just wanted to check-in and see how you're doing. I heard about the family emergency, and I hope everything is okay on your end.
Jordan Wilson [00:35:00]:
If there's anything I can do to help or support you, please don't hesitate to reach out. I understand that the deadline project name was affected, and I completely get it. Family comes first. Let me know if you need any assistance in catching up or if we should discuss a plan to move things forward. Take care, and looking forward to hearing from you when you are ready. Fast your name. Pretty good. I don't necessarily see, again, I'm not, let me be honest.
Jordan Wilson [00:35:28]:
I'm not the most empathetic person. Right? I'm not the, you know, I'll say, my my IQ is probably higher than my EQ, so I don't see anything wrong with this email. Maybe it's, I don't know, maybe it's it's it's, you know, setting off a bunch of red flags in in your mind. Alright. Let's go ahead. Let's switch over to that g p t four point five, and let's see, if we can spot differences in how four five would handle this. Presumably, it might be I would hope it would be better. Right? Since, you know, four five is supposed to be, much more relatable to humans.
Jordan Wilson [00:36:02]:
Alright. So same thing here. Four five launched this in Canvas, which I like because then you can use the Canvas features. Right? As an example, you can highlight something. You can suggest edits. You can adjust the length. You can change the reading level. You can add polish.
Jordan Wilson [00:36:17]:
You can add emojis. Right? Which I would never do. But launching it in Canvas, again, small thing, super smart, but four or five assumes. Right? There we saw it twice. The model says, oh, okay. Well, you're probably gonna wanna add some of your own insights. There's some placeholders in here. Right? If you wanna get work done, we're gonna launch this thing in Canvas.
Jordan Wilson [00:36:37]:
Alright. So subject line, checking in and offering support. Hi, colleague's name. I wanted to reach out and let you know I've been thinking of you. Okay. I see one difference right away. Right? Again, I don't know if this is this is good or bad. Right? But 45 is saying, like, been thinking of you.
Jordan Wilson [00:36:57]:
Alright? I'm trying to I'm trying to jot some notes because I actually can't, as easily compare them, in canvas mode. I heard about your family and see, and I'm truly sorry, for I'm truly sorry you're going through a difficult time. Okay? Please don't worry about the missed deadline. Okay. Interesting. Don't worry about the missed deadline. Your family and personal well-being are the priority right now. If there's anything I can do to support you during this period or assist with any pending task at work, please feel free to let me know.
Jordan Wilson [00:37:34]:
We're all here to help you manage through this. Take all the time you need, and please keep me updated when it's convenient for you. Okay. So, actually, now that I'm reading this, I'm seeing some some some nuances. Right? Some small things. Right? And, again, this is not my area of expertise. I'm not gonna be hired as anyone's HR, head anytime soon. But I do see a couple of things in four five's response, and I'm kind of bolding them, on my other screen right now, to and but but let me know, livestream audience, which one was better, and, you know, do you see anything that you liked in four o versus four five or four five versus, you know, four? Let me know if if if if one was preferable.
Jordan Wilson [00:38:26]:
Alright. So here's a couple things I picked up on. So let's go back to four o. So at least for me, four o starts with saying, right, again, I didn't think there was anything necessarily wrong with four o's response until I read four fives response, and I'm like, wait. Okay. Some things are better here. So four o's response, the first thing it says, wanted to check-in. Right? Yes.
Jordan Wilson [00:39:00]:
It says, wanted to check-in and see how you're doing, but I think even when you read, that email. Right? If if if you're in that situation of a family emergency and someone says, wanted to check-in, it sounds kind of business. Right? It sounds, I guess, a little cold. Whereas 45 says the first sentence, I wanted to reach out and let you know I've been thinking of you. K. There we I mean I mean, just that right there, I think you can hopefully see and realize the the bump in EQ. Right? And I think maybe that's where there's also I don't know. In my mind, I'm also, you you know, trying to trying to describe in real time the vibe.
Jordan Wilson [00:39:48]:
The vibe of the four o, letter I'm getting now that I'm reading it is sympathetic, you know, with a little bit of like, hey, let's get this project going forward. Where four five, I think is maybe a little more empathetic and talking about working together to move something forward. That's what I'm getting. 45 says, please don't worry about the missed deadline. Right? Where 4 o says, you know, I understand that the deadline was affected, where four five says, please don't worry about the missed deadline. Okay? Four o, you know, to kind of move the project forward, says, let me know if you need any assistance in catching up or if we should discuss a plan to move things forward. Okay? So, again, when I'm reading that by itself, I'm not necessarily like, oh, this is bad. Alright.
Jordan Wilson [00:40:53]:
And then 45 says, if there's anything I can do to support you during this period or assist with any pending task at work, please feel free to let me know. Alright. Where even just saying please feel free verse, you know, versus four o just says let me know essentially about these tasks. And then four five, again, it looks like showing a little more, empathy versus sympathy and maybe prioritizing, the family situation where at least now as I'm kind of comparing the two, you know, it looks like four o's just like wrapping up some sympathy and, like, yo, let's get this project going. Right? Which, I don't know. What do you guys think? Denny says four five sounds like the person really does care, and four o sounds like I need to write the this email to show I care. That's a great observation, from Denny. Max says either one would work.
Jordan Wilson [00:41:53]:
Four o is what I usually would expect from the regular office people. Four five is superior EQ and empathetic more than use more than usual office humans. Yeah. That's what I'm saying. Right? Like, when I first saw four o, I'm like, nothing wrong with this. Right? But then when I said four five, all of a sudden, I'm like, oh, okay. Yeah. I can see how on the, you know, on the human side, there's maybe some things that could have been improved, in this four o.
Jordan Wilson [00:42:20]:
Michael said I would prefer to receive four five. I feel like I would write something closer to four o. My gosh, Michael. We are the same. Right? We are the same. I'm reading these and I'm like, oh, yeah. I like that one. But I I totally I totally would have personally written, something like four o.
Jordan Wilson [00:42:38]:
Yeah. Sandra is saying four five is more empathetic. Jonathan. What's up, Jonathan? Jonathan says four o minus the hope you're well seems more HR friendly, while four five just seems friendly. Yeah. I agree with that. I, yeah, I I I do think maybe, you know, four o is maybe more in line, with, you know, maybe HR guidelines, and four five is probably something that is going to resonate with the human receiving it, I think, much more. Cecilia said four five version drops the need to move the deadline forward and puts it on a pause.
Jordan Wilson [00:43:13]:
Four o makes it clear that we will need to move things forward on the specific deadline missed yet. So I guess it all ultimately depends on what's the most pressing thing. Right? Is the company gonna shut down, if this deadline isn't met? Right. So, great observations. I love doing this live y'all because, you know, you guys, spotted a lot of things. Robert from Twitter saying, four five has more empathy in it. Alright. I think we can, do one more.
Jordan Wilson [00:43:40]:
So scrolling through here to see, one more we can do that people voted for. Okay. Let me bring this in. So, a lot of people wanted number six. So let's try that. Alright. So we're going now to a new chat. We're gonna start with four o.
Jordan Wilson [00:44:02]:
Here's number six. Alright. So it says, draft a daily motivational message that encourages personal growth with clear fact based tips. Not a lot of guidance. We'll see what happens. Alright. So first, we're running with four o. Lots of emojis.
Jordan Wilson [00:44:23]:
So if you're personally motivated by emojis, you can already tell four o is your model. Alright? That's something I've noticed as well. Not getting a lot of emojis out of four or five. Alright. So we can read this. So here's four o's daily motivational message that encourages personal growth with clear fact based tips. Alright. Livestream audience, let me know after I read both of these which one gets you pumped up.
Jordan Wilson [00:44:48]:
Which one of these is is, you you know, getting you ready to face the day. Alright. So daily motivation, small steps, big growth, rocking emoji. Alright. Every day is a chance to become 1% better. Personal growth isn't about massive leaps. It's about consistent small actions that add up over time. Fact based tip.
Jordan Wilson [00:45:06]:
Alright. We're even labeling it here. Studies show that setting up specific achievable goals increases your chances of success by up to 42%. Harvard Business Review. I'm obviously gonna Google that to see if that's real. Alright. And it is. Alright.
Jordan Wilson [00:45:25]:
Good. So, at least four zero did not hallucinate that part. Alright. So, studies show that setting specific achievable goals increases your chances of success by up to forty two percent. Instead of vague intentions like get healthier, try walk 10,000 steps daily, or read 10 pages before bed. It's actually a great tip. Actionable step. Pick one small habit to improve today, Whether it's drinking more water, writing down your thoughts, or practicing deep breathing for focus, small wins lead to big transformations.
Jordan Wilson [00:45:57]:
Keep moving forward. Your future self will thank you. Emojis, hashtags. Right? Sorry. I hate reading this. This is everything that's on social media right now. Right? I wanna vomit as a content writer. I hate this.
Jordan Wilson [00:46:15]:
I hate anything one shotted from any large language model because most of it always reads like this. Right? You can get writing chatty bitty writing or AI writing to sound extremely human, but you're you're I don't care. Right? Oh, people are like, oh, cloud three five or a quad three seven sonnet sounds so human. GPT four o, GPT four five sounds so human. No. For the most part. No. It sounds like this.
Jordan Wilson [00:46:39]:
This is what I read on social media, on LinkedIn, on Twitter, and I wanna vomit my mouth. There's nothing wrong with the actual tips. I just the the content writing is getting me woozy, like like woozy Rogers. Alright. Let's go four five. So, again, draft a daily motivational message that encourages personal growth with clear fact based tips. Alright. Interestingly enough, I already see, one emoji in four five.
Jordan Wilson [00:47:08]:
I think four had, like, let's see how many four had. One, technically, two, three, four, five. Alright. So we had four emojis and two hashtags, in four o. And then in four five, we had two emojis, no hashtags. So a little better. Alright. Ready.
Jordan Wilson [00:47:26]:
Here's four five. Today's motivation, embrace your growth. Growth doesn't happen overnight. It comes from small, consistent actions taken each day. Here's your simple fact based guide to grow personally and professionally today. Number one, learn something new, even just ten minutes. Research shows daily learning builds neural connections, improving memory and cognitive skills. Two, practice gratitude.
Jordan Wilson [00:47:53]:
Write down three things you're grateful for daily has been scientifically proven to enhance overall happiness and reduce stress. Number three, step slightly outside your comfort zone. Small challenges creates confidence through psychological resilience, enabling you to face bigger challenges more effectively. Take a small step forward today. You have everything it takes to grow, learn, and succeed. Alright. Which one, live stream audience? Which one has you pumped up and ready to run through a wall like the Kool Aid man? Oh, yeah. Do any of these? Jeep d four o, Jeep d four five.
Jordan Wilson [00:48:32]:
Let's talk quickly about some of the some of the differences. So, surprisingly, on this example, four o was a little more concise. It was a little too heavy on the emojis, a little too heavy, on the on the hashtags. Not a huge fan of this. One other thing is even if we're just looking at the quality of the content writing, I think four o was a little poor. There is no cadence, or switch up in the sentence structure. Yeah. I'm getting a little, you know, putting on my old, writing hat, right? You always want cadence in your written content.
Jordan Wilson [00:49:07]:
What that means, I try to throw cadence in my podcast. Right? I don't just always speak monotone. I don't always go in, you know, sentences that are, you know, 15 to 20 words. I try to pause. Sometimes I talk slowly. Sometimes I talk really fast, and I have these long sentences that go together, and there's no period, there's no punctuation, and I talk all excitedly. That's cadence. Right? So four o has no cadence.
Jordan Wilson [00:49:35]:
It actually falls into this this compound sentence. Right? So, yeah, we're talking about content writing now, but that's something that I think is significantly improved, in four five. Four o's, you you know, I know you're maybe not you you know, if you're listening on the podcast, maybe this doesn't worry, or matter as much. But four o is kind of the equivalent of watching paint dry when it comes to content structure. Yeah. I was a journalist. I wrote a lot. For the most part, most of these sentences are it looks between 12 to 20 words, and the majority of them are compound sentences with an em dash.
Jordan Wilson [00:50:16]:
Alright? So, yeah, all those people are like, oh, you know, an m dash is definitely, you know, a sign of of AI writing. Yeah. Not really. Right? I was using m dashes back when I was a a journalist at the Freeport Journal standard in twenty o two or twenty o3 or whatever. Right? Love em dashes. Love compound sentences, but huge overreliance on them here from GBT 4 o. So, let's see. One, two, three.
Jordan Wilson [00:50:44]:
So out of, like, the six sentences, three of them are compound sentences with m dashes. Not good. We only have let's see. We have zero sentences that I would consider short, which is five words or less. Alright. So if we look at, GPT four five, we only have one compound sentence, with an m dash, so that's better. Okay. We do at least have one short sentence.
Jordan Wilson [00:51:16]:
Alright. So a little better a little better in terms of content structure, you you know, some some cadence, some variance, but still nothing great if we're just looking at content. Right? I know this is more about the motivational message, but I did want to take a second to look at even just how the content is produced because I think that is, another small detail that four five actually has better. So yeah. Less, you know, like, oh, people are always like, oh, this is AI content. Right? You can't technically tell, although there's a lot of telltale signs. Right? Heavy emojis, double emojis and and headlines, you know, random hashtags, you know, like I said, an overreliance or a heavy percentage of sentence, of sentences that are compounding set, compound sentences separated by an m dash. So overall, the content writing, I think, is much, much better on four five.
Jordan Wilson [00:52:08]:
Alright. So what do you all think as we as we wrap up here? But like I said, these are the areas, and I think you saw it in probably that middle example, probably the best, the email example. How really where having a little bit of EQ, some emotional intelligence, and being relatable as a human. Right? A lot of you said the same thing. I said the same thing as well. I'm like, I I want to receive that second email that we talked about, the one that was from 04/05. It just felt more human. It was probably more human than something I would have written.
Jordan Wilson [00:52:46]:
Right? Which is pretty impressive. Right? It is pretty impressive. And I think that's one of the reasons, why this new model, GPT four five, when it comes to humans' preferences. Right? Yeah. You know, four or five didn't crush every single, LLM benchmark. It improved on almost all of the benchmarks from four o to four five. But, you know, people were like, oh, OpenAI has hit a wall. OpenAI is gonna go bankrupt.
Jordan Wilson [00:53:16]:
OpenAI is garbage. It didn't, you know, break every single benchmark out there. Right? I don't think most companies we saw the same thing with Claude three seven anthropic. Alright. Sonnet Sonnet three seven from from Claude. It didn't break every single benchmark out there. It really excelled and and just widen their lead in anything software development, anything on the on the dev side. Right? But I think now we're gonna see companies probably more focused on something like Elo scores.
Jordan Wilson [00:53:47]:
Right? On the on the chatbot arena. Right? And they're like, yeah. We hope our actual benchmark, you know, our our stem, our math, our our reason. Right? All these kind of, like, quote, unquote, more scientific research based, category based benchmarks improve. But I think, ultimately, we're past that. I think we're past that. And, right, and this is indicative. The fact that GPT four five did not crush every single benchmark on paper that people said, oh, these are important, but at the same time, instantly shot up to the number one model in the world preferred by humans that says something.
Jordan Wilson [00:54:24]:
Right? There is a human side to large language models that I think for the most part, you know, that we ignored before 2023. Right? Everything was about overfitting models to hit certain benchmarks. And I think over time we saw, okay, that's great for benchmarks, but it's not benchmarks using these models, it's humans. It's humans trying to solve real problems. It's humans trying to sell things to other humans, trying to improve customer relationships, trying to increase, accuracy and reliability, which are all things I think GPT four five does a great job of. So before you, listen to that random influencer online that is just spitting out these benchmarks and it's like, oh, OpenAI has hit a wall. I'd say the exact opposite. I'd say the exact opposite.
Jordan Wilson [00:55:12]:
I'd say if we're being honest, right, a lot of the things that we do on a day to day basis are creating communication for other humans. And as someone that's y'all have won national writing awards. I've I've done okay. I was a Pulitzer, fellow. Some of those emails, better than I would have written. Right? If I had to write some of those emails because it's thinking about the human. It is trying to be more relatable. It is really flexing its EQ skills, which I think is ushering in a new era, not just of how large language models are built, but how they ultimately should and could be used to strengthen relationships and connections between humans while also still, you know, hopefully excelling in all those benchmarks.
Jordan Wilson [00:56:06]:
But in the end, that's what it's all about. Alright. I hope this one was helpful y'all. If it was, please go to our website. Go to youreverydayai.com. Sign up for that free daily newsletter. Also share this. Right? I know a lot of people tell me, oh, Jordan, I'm not gonna tell anyone about this.
Jordan Wilson [00:56:26]:
Right? Everyone at my company thinks I'm a genius. Right? I've gotten so many so many messages. I love these. Reach out if if, you know, if if you have a story like this. I always love hearing it. It makes it makes the long nights and early mornings, really worth it. I love hearing from people that are like, hey. I just got a job.
Jordan Wilson [00:56:44]:
My first job in AI. Thanks to thanks to, you know, your your podcast. Thanks to these guests you bring on. Right? And people tell me, like, I'm not telling anyone about this. This is my cheat code. This is my secret. Share it, please. People are always like, how can I help? How are you making all this information free? It's because of those of you that actually do share this.
Jordan Wilson [00:57:02]:
So if you're listening on the podcast, thank you. I appreciate it. Please subscribe. Please leave us a rating. That would be great on the podcast, and also go to youreverydayai.com. Sign up for the free daily newsletter. Read the read the daily newsletter as well. Each and every day, we break down exclusive insights that you didn't hear from the podcast.
Jordan Wilson [00:57:19]:
We're gonna take this a step further, as well as keeping you up to date with everything else you need to know in AI. So thank you for tuning in. Hope to see you back tomorrow in everyday for more everyday AI. Thanks, y'all.
