Everyday AI Podcast – An AI and ChatGPT Podcast - EP 495: Gemini 2.5 Pro Unlocked: Exploring everyday use cases

Episode Date: April 2, 2025

Wait. Gemini 2.5 can do whhhaaaaat? 😱Finally coming up for air after generating 1,459 images in ChatGPT's new image generator? Just in time. After we wrapped up our Part 1 overview of Gemini 2....5 Pro Unlocked, we're going all-in on use-cases for Pt. 2. We're going to tackle some everyday business use-cases for Gemini 2.5, some wacky ones, and creative ones. All LIVE. (And we'll probably break things along the way.) Join Us. Gemini 2.5 Pro Unlocked: Exploring everyday use cases (Pt 2 of 2)Newsletter: Sign up for our free daily newsletterMore on this Episode: Episode PageJoin the discussion: Thoughts on this? Join the conversation.Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineupWebsite: YourEverydayAI.comEmail The Show: info@youreverydayai.comConnect with Jordan on LinkedInTopics Covered in This Episode:Innovative Uses of Gemini 2.5 ProAnalyzing Gemini's Built-in ThinkingReal-World Business ApplicationsLive Demos of Gemini FeaturesGemini's PDF and Image AnalysisInteractive Content with Gemini CanvasGemini's Onboarding and Training ToolsGemini's Unmatched Processing PowerTimestamps:00:00 "Google's Gemini 2.5 Boosts Business"04:51 Future Non-Coders Transform with AI07:35 "Top Tech Podcast Episode: AI Demos"10:44 Gemini 2.5 Pro Experimentation13:26 Google Gemini Podcast Musings16:21 Exploring Google Gemini's Tool Usage20:23 Testing AI Podcast Analysis25:17 "Efficient AI News Searching"27:52 "Canvas and Claude: HTML Challenge"29:31 PDF Extraction for Business Automation34:11 Canvas Integration: Google Doc Alternative36:36 Google Gemini: Live Rendering & HTML Pasting39:45 "AI Learning Tools: Pros and Cons"43:11 OpenAI's $40B Funding Round46:45 IBM Onboarding SOP Manual50:47 "Revitalize Business with Google Gemini"52:19 Effective Use of Large Language ModelsKeywords:Gemini 2.5, Google Gemini 2.5 pro, Google's new Gemini, Google's Gemini update, Gemini 2.5 pro unlocked, Large language model, Multimodal, AI Studio, Chain of Thought model, Chain of Thought reasoning, 1,000,000 token context window, Transformer model, Advanced coding, AI Studio vs. Gemini, AI Studio can’t turn off data training, Human preference, Benchmark scores, Self-help coding, AI-powered coding, Business use case, Coding capabilitiesSend Everyday AI and Jordan a text message. (We can't reply back unless you leave contact info) Start Here ▶️Not sure where to start when it comes to AI? Start with our Start Here Series. You can listen to the first drop -- Episode 691 -- or get free access to our Inner Cricle community and all episodes: StartHereSeries.com Also, here's a link to the entire series on a Spotify playlist. 

Transcript
Discussion (0)
Starting point is 00:00:00 This is the Everyday AI Show, the everyday podcast where we simplify AI and bring its power to your fingertips. Listen daily for practical advice to boost your career, business, and everyday life. Meet Firefly AI Assistant, now live in Adobe Firefly, the all-in-one creative AI studio. Just describe what you want to create and the assistant handles the rest, orchestrating multi-step workflows across Photoshop, Premiere Express, and more in one conversational interface. You direct the outcome. The assistant accelerates execution. All right.
Starting point is 00:00:46 So if you haven't heard Google's new large language model update in Gemini 2.5 Pro is good. It's like really good. As in benchmarks, the best, human preference, the best. But what can it actually do for your business? I think this is something that we're always thinking. about. I think early on, you know, in the chat GPT days, we got into this kind of rut, right? When large language models first came out and we thought, okay, well, they're just for creating content, right? This is to help me write a blog post or a large language model is to help me,
Starting point is 00:01:28 you know, write something for LinkedIn or improve an email to a colleague. Yes, large language models are good for these things. But what about when we talk about state of the art, multi-modal large language models like Gemini's, sorry, like Google's new Gemini 2.5 Pro. So today I thought we'd have a little bit of fun and maybe a little bit of chaos as we go over in part two, Gemini 2.5 Pro unlocked exploring everyday use cases. All right, I'm excited for this one. I hope you are too. What's going on, y'all? My name's Jordan Wilson. If you're new here, Thank you for joining us. This is Everyday AI.
Starting point is 00:02:10 This is your daily live stream podcast and free daily newsletter, helping us all not just keep up with AI, but how we can all actually use it to get ahead to grow our companies and to grow our careers. Is that personal? Is that you? Is that what you're trying to do? If so, this is step one, listening to this podcast or live stream.
Starting point is 00:02:28 Step two is going to our website at Your Everyday AI.com. Here's what we do. Two main things in our free daily newsletter. One is we recap and sometimes summarize. you know, the episode for today. You know, sometimes I have guests on today. It's just me talking about Gemini 2.5. So we give you what you really needs to know
Starting point is 00:02:48 and pull the valuable insights from each day's episode as well as keeping you up to date with everything else happening in the world of AI. So make sure you go to our website, Your Everyday AI.com. Sign up for the free daily newsletter there. All right. So normally we go over the AI news and all that in the beginning of the live stream.
Starting point is 00:03:06 This one could be a longer one, and I'm trying not to. So if you do want the EI News, we're going to have that in the newsletter. All right. So I'm excited, and I hope I can get a little bit of help from our live stream audience today. So thank you for tuning in. Dennis, joining us from New York City. Yeah, where are you all from? I should ask this more, right?
Starting point is 00:03:25 I like to know where our live stream audience is from. Brian joining us from Minnesota. Kyle, thanks for tuning in. Michelle, Big Bogey, Sandra, J, everyone else. Thank you. But I might be asking some help from you all today. All right. But let's just get caught up, right?
Starting point is 00:03:43 So I did an entire episode yesterday on what's new in Google Gemini 2.5. So if you do want to know, just scroll back one episode. Maybe you're listening to this on the podcast. It's episode 494, where we just went over the basics of Google Gemini 2.5. But as the world's fastest recap here, here's here's got to the super simplified version of what's new, okay, in Gemini 2.5. So it has built in thinking. That's the biggest one. It is technically a hybrid model.
Starting point is 00:04:20 It combines the kind of old school, quote unquote, transformer model with a reasoning or chain of thought model. So you'll see that as we do some live demos here. And it's gotten very impressive scores, not just on traditional benchmarks, but also some newer benchmarks like humans, humans last exam, humanity's last exam, where it scored way better than any other large language model. It does have a enormous one million token context window. So that is more than 1,500 pages, as an example, 30,000 lines of code before Google Gemini 2.5 Pro begins to forget things. I will let you know and we'll probably see here live. That is when you are using it in AI studio versus the front end of Google Gemini,
Starting point is 00:05:07 more on that a minute. Probably one of the biggest leaps in terms of capabilities, and maybe this will be applicable for your business, maybe not, is the advanced coding. So Gemini 2.5 is very, very good at coding. You might be thinking, all right, Jordan, that's not me. I'm not a software engineer, right?
Starting point is 00:05:25 Okay. If you listen to our 2025 AI roadmap and prediction series, I said in 2025, everyday non-technical people are going to be using large language models to spin up their own apps, to spin up their own, you know, I don't know, Chrome extensions, their own desktop apps that help them do things better. We're not there yet, but I think we will be there very soon. So keep that in mind, you know, just because you're not a current coder or developer or software engineer, you should still, I think, really pay attention to this Google Gemini 2.5, the big leap in coding.
Starting point is 00:05:59 And maybe some of our use case examples will show that. You know, number one benchmark ranking, that's huge. So the biggest thing we talk about, there's all these benchmarks that I think sometimes, you know, AI labs overfit for. But when it comes to ELO scores inside the CHAP, the LM chatbot arena, that's human preference. right so people put in all kinds of prompts you know write a blog post create you know write code for this you know generate a creative outline for you know or strategy for x right um and you get two responses you don't know who they are you choose which one is the winner uh and jemini 2.5 pro has literally broken the record for the biggest leap into the number one spot normally when a new model comes out
Starting point is 00:06:44 you know from open ai from claude etc it'll usually get the number one spot right because generally it is you know between two to six months between big models, especially earlier in 2024. So, you know, usually the top model would come in a couple points higher, a couple preference points higher. Gemini came in at 39 points higher than what was in second place. Now, what's in second place right now is GPT40. The other big recap here, it's free. Was not expecting this.
Starting point is 00:07:14 So Google did not even announce this in their initial Gemini 2.5 announcement. They quietly put it out in a tweet over the, weekend, right? But even if you don't have a paid account of Google Gemini, you do have access to Gemini 2.5 Pro for free. The limits are a little more restrictive. All right. So one more thing, one or two more things before we get started here. So in live stream audience, you know, if you have any things you want to try live, let me know. Maybe I don't know in your comment, I should have thought about this, but like beforehand, I don't know, maybe put two stars. All right. And then I'll see if I can maybe copy and paste it.
Starting point is 00:07:52 I don't know if I'm able to, but I'll try. Or at least I can try to, you know, get the gist of what you want to see. But before we get started, a couple things to keep in mind. Our podcast audience, thank you for tuning in. Y'all are awesome. I never would have thought when I started this thing. This would be a top 10 tech podcast. But this is one of those.
Starting point is 00:08:10 You might want to check out the newsletter so you can come and watch the video. You can always rewatch it on our website on YouTube, on LinkedIn. I'm going to try my best to verbally describe, what's going on with this is unfortunately going to be a very verbal or sorry a very visual episode and this is something uh this is always the number one request we get right do more live demos do more live demos uh so you know podcast audience uh i'm going to try my best but this is one you might want to come watch the video on another thing to keep in mind a i studio versus gemini okay jemini is the front end chat bot for google a i studio is kind of a sandbox for developers although
Starting point is 00:08:49 So it's not as hard as you may think, right? You know, there's some initial setup, but then after that, it's pretty easy. You know, if you're on a paid plan of the front end Gemini chatbot, you can turn off model training, which is important, right? Because you should never be sharing, you know, proprietary, sensitive, PHI, right, like private health information into a chat bot. But if you are using AI studio, there's no turning off data training. So AI Studio is free. That is actually where you get the more powerful version of Gemini 2.5 because you get the entire context window in some other controls that you don't get on the front end of the Gemini chat bot. And hopefully I'll be able to demo that here in a minute.
Starting point is 00:09:34 But just keep in mind, Google's AI studio is free, but you cannot turn off data training. If you are on a paid plan of Google Gemini on the front end on the chat bot, you can turn off training. All right. The other thing, I'm doing this live, y'all. all right. So bear with me. But I think it's actually important, right? Because if you go watch anything online, you know, there's some great creators out there, you know, who put together, you know, demo videos and all that.
Starting point is 00:09:59 I know a lot of these people. I talk to them and I know how long these videos take, right? So sometimes to put together a couple demo use cases of something like Gemini 2.5, it might take them five hours of recording for a 20 minute video. Okay. and a lot of editing to make sure it looks right. I don't like that. You know, people are always roasting,
Starting point is 00:10:23 roasting me on our YouTube channel because it's like, oh, your production quality stinks and you have all these mistakes. And sometimes you stutter or say the wrong word. I'm a human, right? This is live. This is unscripted. This is unedited. This is just, you know, but I think it's important because I think so much of these,
Starting point is 00:10:40 of these demos of all large language models that you see, all AI tools are overly polished. They're manufactured. You know, in some cases, you know, they're being artificially pumped and promoted on the back end to make you think there's something that they're not. This is real. This is live. This is unedited. All right. So keep that in mind. Live demos with generated AI are a terrible idea, right? But you all like them. You all want to see them. So we're going to do them. And so far my takes right now with Google Gemini, it has an extremely high ceiling, but a finicky floor. All right. Let me let me kind of describe what I mean by that. So here's, here's an example. And I put this out on Twitter and I'm going to ask the Google team about this. But it's, it's keep in mind, Gemini 2.5 Pro is experimental.
Starting point is 00:11:27 All right. Very experimental because sometimes you're going to get a weird, a weird result like this. Right. I always have a series of prompts that I use to, especially for internet connected models. So I can make sure that they're correctly pulling information, right? When we talk about the role of human in the loop, it's very important. And as large language models get more. powerful, more robust, more features like Gemini 2.5, I think us humans think, oh, we can sit back
Starting point is 00:11:52 and relax. We actually have to be more vigilant. The more that we hand off to large language models, the more that we have to, I like to think of it as expertise in the loop, right? Not human in the loop. Human in the loop just thinks like, okay, you know, I'm going to blindly, you know, do my human job here. This looks good click. No, you have to apply your expertise. This is a simple example, right? But I said, what's the latest episode of the everyday AI show by Jordan Wilson, right? I want to see if Google Gemini 2.5 can get my episode from yesterday, right? And in this example, uh, you know, because it is a hybrid model, I can even see the thinking and it says the user is asking for weather forecasts in Chicago, Illinois for today,
Starting point is 00:12:30 April 1st, 2025. I should use a weather tool to get the current weather and forecast for Chicago. Number one, not true, right? It didn't. Number two, uh, not surprisingly, right, it picked up my location without me telling it. All right. So keep that in mind. It's finicky. It's experimental. But when it works, I am very impressed. I am very impressed. All right. Let's get wild.
Starting point is 00:12:54 Let's get wild, y'all. Please, live through your audience, can someone tell me if you can see, see the screen here? I'm going to be jumping in between some tabs here. But if you could, let me know, because I don't want to do another 25 minutes of the show and bringing you guys these live demos. And you're like, oh, Jordan, you weren't sharing your screen at all. Kimberly says we need to see more bloopers too. It's a part of life. Yeah.
Starting point is 00:13:19 I think that's how you learn generative AI. That's how you get better at large language models. You try them, right? No one's an expert, right? I won't say no one. There's very few people that have been working in large language models since, you know, for 10 years. There's a couple people, right? But most of us, you know, you have to learn on the fly and you learn by failing and you learn by making it better.
Starting point is 00:13:40 All right. Dennis, thanks, Dennis. Dennis said AI is cool, but we love here. Humans more Jordan. Okay, cool. All right. Thank you, Nicole and Kimberly for let me know in Charles that you can see the screen. Cool. Let's let's get after it. All right. I'm going to be jumping around a little bit here, y'all. And I apologize if you hear like a lot of clicking. All right. That's my mouse. I should probably figure out how to, you know, not pick that up in the podcast. All right. So I'm going to go in and upload a file. So first, I am right now. I'm on the front end of Google Gemini. in your drop down, you have 2.5. One thing to keep in mind, and maybe this is a hack for chat GPT, there's no model switching,
Starting point is 00:14:24 which I wish there was in Google Gemini on the front end. So as an example, if I start in 2.0 flash, you know, I'm just going to say sup. All right. Now, if I want to model switch or start working in 2.5, I can't.
Starting point is 00:14:40 It refreshes that chat. So why does that matter? Why is important? Well, as an example, I'd love to like use deep research. So deep research inside Google Gemini has been upgraded to Gemini 2.0. It's actually amazingly good. But so if I wanted to, you know, do something in deep research and then go over to 2.5
Starting point is 00:14:59 pro, you can't do that. Whereas with chat chbt, you can't. I think that's like such an underrated hack is just model switching inside chat gbt. But, you know, before we get started, it's worth pointing out. All right. So I am on the gemini.com. I have a paid account, FYI, but even if you have a free account, you should be able to do this. Live stream audience, if you want to file them along, you know, you can do that as well.
Starting point is 00:15:22 All right. So I'm selecting 2.5 Pro experimental from the drop-down menu, and I'm going to add a file here. All right. So I'm going to add a PDF here. Adobe just introduced an entirely new way to create, bringing the power and precision of its creative suite into one conversational experience. Meet Firefly AI Assistant, now live in the Adobe Firefly app, the all-in-one creative AI studio. Powered by Adobe's Creative Agent, Firefly AI Assistant lets you start with your vision, just describe what you want, and shape the outcome as it takes form with the Assistant.
Starting point is 00:16:07 The Assistant orchestrates multi-step workflows, drawing on 60-plus pro-grade tools across Adobe Creative Cloud apps, including Photoshop, Illustrator, Premier, Lightroom Express, and more to help bring your ideas to life. You can also get started with creative skills, a growing library of pre-built workflows for common creative tasks like batch editing photos, creating mood boards, portrait retouching, and creating social variations. Every step the assistant takes is visible so you can refine, redirect, or take over at any time. You stay in the driver's seat as the creative director. Adobe Firefly AI assistant now in public beta. See it today at firefly.adopi.com. So I'm going to describe what's going on as this happens.
Starting point is 00:16:58 I'm going to say, please. So I'm just saying I'm uploading a PDF and I'm saying please transcribe every word of this. So this is about a, let me see how many pages this is. It's probably about a 15 page PDF here. So these are, you know, people reach out and they're like, hey, I want to, you know, advertise on the everyday AI podcast. So I have this little deck that I send potential advertisers sometimes. So hey, if you do want to reach.
Starting point is 00:17:24 one of the largest audiences in artificial intelligence, you know, on our podcast, you know, make sure to reach out to me. But the thing is, most large language models cannot read this because, I mean, number one, I made it in Canva. So, you know, most large language models when they're using computer vision, when they're using sometimes OCR technology, they all work a little bit different. They really struggle with this because it's all essentially images, right? It's not like I build this in Word and it's a bunch of texts.
Starting point is 00:17:53 This is very visual, right? There's backgrounds. There's tons of images on each page. Right. It's a lot going on. So, you know, even to pull all of these words, I mean, we'll see. I've done some of these so far. Some I haven't.
Starting point is 00:18:06 So let's see how Gemini 2.5 does. So I can click show thinking, right? And I'm not going to be able to do this for every single one. But it says, I need to get the relevant content to answer each user's questions. The user wants a transcription of the entire PDF document. I have the extracted text from the document provided by the content fetcher tool. So I'm going to spend a little bit more time on looking at the chain of thought. And y'all, this is huge, right?
Starting point is 00:18:36 The thing I love about Google Gemini's chain of thought is you can see their tool usage, all right, which is going to help you get more out of the tool if you know, because you can start to speak Google's language. And hopefully that will, you know, become a little more clear here when I try a, another prompt here. So anyways, let me just go ahead and scroll down. And you'll see right away, it's breaking it down. Page one, here we go.
Starting point is 00:19:01 Everyday AI sponsorship opportunities, daily podcast, live stream newsletter, perfect. It's got the website. Great. Page two, it's got it all. Okay? This is really,
Starting point is 00:19:12 really good. I haven't seen this out of a large language model yet. And it's formatted. It fixes, you know, sometimes the fonts look a little weird, you know, but it crushed it. All right. This is,
Starting point is 00:19:25 this is impressive, y'all. All right. So I'm going back. So at the bottom, I have trusted by leaders from, right? Because we have all these people from big companies that have,
Starting point is 00:19:36 you know, that read our email newsletter, that reach out to me, that have given us testimonials, you know, from Google, Amazon, Nvidia, Microsoft, etc. Right?
Starting point is 00:19:43 We have a lot of listeners. Yeah, if you want to reach them. So not only did it get the text, but Google Gemini here, very impressive. use computer vision and gave me just the names right i didn't put the name google the name invidia the name ibn those were multiple images mind-blowingly impressive all right page three you know partnership opportunities uh so good so good uh so i'm i'm actually curious uh and again i'm doing a lot of this
Starting point is 00:20:16 live uh did you guys know a live stream audience did you guys know this i i i i i I didn't even know that it was going to look at the images in this deck. You know, I've tried this a lot with Chad GBT. I've tried it a lot with Claude. I haven't tried it with the updated version of 4-0. That was just rolled out a couple of days ago. So maybe it'll do better. This is very impressive, right?
Starting point is 00:20:38 So I'm curious if it's even going to pull some of these stats. So I have like our ad channel overview. And there's like text within screenshots of this image. So, you know, I'm curious. I'm just going to scroll down to that page. Let's see here. Add channel overview. Okay, it didn't pull it in, but that's fine. The text was probably too small. But it literally crushed it. My gosh, it even created I have a chart. This is so, so good. I have a chart and it converted my little chart, which is just I made in Canva, right? So not only was it able to pull all that because a lot of it is images. It created a chart for me that I can export to sheets. So I can click that export to sheets. And then, Open and sheets, bam, there it all is, our little breakdown. Just that right there is wild, y'all.
Starting point is 00:21:28 How many times when we talk about business use cases, right? I don't know, but you guys, I read a lot of PDFs, right? Or a lot of documents. Sometimes you may not have the version that you need, right? It's like, oh, my gosh, this was from Bill. He left two years ago. I have to redo this entire thing. Well, you can upload it into Google Gemini 2.5 Pro.
Starting point is 00:21:47 It's going to transcribe the whole thing. If there's charts and graphs in there, it's going to reclass. create them. You can open them in Google sheets. This one use case alone. Wow. Wow. Very, very, very, very good. Uh, all right. Hey, cool. Sandra says she's doing it along on her computer. All right. Let's do another one. Uh, and this is where I think we're going to get some things that go wrong, but, uh, let's try it anyways. All right. Because like I said, I did try some of these, uh, some of them I did not. So I'm saying, find the 20 latest episodes of the Everyday AI podcast and give me a brief summary of each one,
Starting point is 00:22:29 then find five trends between episodes. All right. So think, what's your business use case? What are you following? And think, you know, obviously Google Gemini 2.5 is connected to Google. So one of the reasons I'm doing this, I think it's going to fail. All right, here we go. Hey, we got a live hallucination, y'all.
Starting point is 00:22:46 All right. So it says the user is asking for the date of Easter in 2025. Strangely enough, this is the exact same. hallucination I got the first time I tried it. So I'm just going to add one more, one more thing. I'm going to put my name by Jordan Wilson. I don't think. So last night, I did get this to work correctly.
Starting point is 00:23:04 But I did get some interesting, some interesting insights by looking at the chain of thought, by looking at the different tools that Google is using under the hood to pull this information. All right. So now on the second time, it got it right. It didn't tell me the dates of Easter, which I don't know why I did it. All right. So it's breaking this down. So it says, this requires multiple.
Starting point is 00:23:23 multiple steps. One, oh, it just shrunk that. Okay. Did you guys see that live? It was working correctly. Everything was good. And then it says the user is asking for the top five rock songs released in 1977. Y'all, this is why I said. I said this ahead of time. The ceiling is so high. The floor, so finicky. at least right now on the front end of Gemini 2.5 Pro. So what we could do, I wasn't planning on doing this, but let's just do it anyways, y'all. Let's go into AI Studio. All right, so AI Studio, it is more of a developer tool or a sandbox,
Starting point is 00:24:04 but it's actually very easy once you get it set up. All right. So you can click the create prompt button right here. You can choose the different models over on the right hand side. So a little different. I'm going to try the same thing. Let's go to Gemini 2.5 pro experimental. I'm going to turn the temperature down on this.
Starting point is 00:24:23 Okay, the default is one for creativity. I want facts. All right. And then I'm going to go ahead and turn on. So you can turn on and off different features. This isn't a full-blown AI studio tutorial. I just want to see if this will work. But I'm turning on grounding with Google search.
Starting point is 00:24:39 So I have found when I get some weird little hallucinations like you just saw on the front end of Google Gemini. usually when I try it inside AI studio, it works a little better. All right. So now I can expand to see the chain of thought. So it's saying the user wants a list of the 20 latest episodes of the Everyday AI podcast, identify five trends. So it's looking up search queries. These are the search queries. What are the latest episodes of the Everyday AI podcast?
Starting point is 00:25:06 Everyday AI podcast latest episodes list, right? It developed a plan. And then it says, here are the 20 latest episodes of the Everyday AI podcast. All right. I spoke too soon. I did not think Google Gemini was going to get this correct. We saw when we used the front end Google Gemini chatbot, it went off the rails.
Starting point is 00:25:26 It's experimental, y'all. It's going to do that, right? But inside Google AI Studio, very good job. So interestingly enough, it got this 100% right. So we got our latest episode, which was 494 from less than 24 hours ago. So it did a good job. And then it got the most recent 20. Fantastic. Now it says five trends between episodes. So it says there's been a focus on major
Starting point is 00:25:54 AI players and models, correct? Rise of AI agents and automation. Yep. Industry-specific AI applications, impact on work and productivity, hardware and infrastructure importance. Great. So it did a good job of picking up, you know, some kind of some common trends over the last 20 episodes. So even though Google Gemini, the chatbot got a big fat failure. The Google AI studio, very impressive job. I've done similar prompts like that between all the internet connected large language models about six months ago and none of them handled them the way that Google's AI Studio just did. All right, let's try another prompt here.
Starting point is 00:26:31 Here's what we're doing. This one's a little tricky. All right. I'm saying summarize this page and I'm giving it a Boolean search URL. All right. I'll explain what that is. But the reason I want to do this is to look at the tool use, right? So look at the chain of thought.
Starting point is 00:26:47 can click show thinking when you're using Google Gemini 2.5. And it says the user wants me to summarize the content of the Google search results page. And then it says the browse tool can be used to extract information from a specific web page URL. However, the URL provided is a Google search results page. The browse tool description explicitly states not to use it for Google search result URL. So instead is saying, I can use the Google search tool. And this is a huge, I'm not going to say cheat code, but this is going to save you so much time once, you know, this Gemini 2.5 Pro on the front end becomes a little bit more stable because now by looking at the chain of thought, you will know what exact tool that you need to call because Google doesn't necessarily tell you. Right. So just in case you're curious, this, this Boolean URL, it's essentially like a Boolean search operator that I use.
Starting point is 00:27:43 I do this every day when I go and see what's the most important AI news, right? But it's just search results for certain companies, you know, Open AI, Apple, Nvidia, Microsoft, Amazon, Anthropic, etc. The latest news. So it's the last 24 hours, just AI news from those companies. So let's see what Google Gemini ultimately did. So I just said essentially summarize it. Did a good job.
Starting point is 00:28:08 Did a good job. So it says key trends. major players are rapidly releasing enhanced AI models like Google Gemini 2.5, OpenAIs, GPD45, Anthropic Claude 3-7, IBM's Granite 3-2. It did a really good job, right? Even though I can't see exactly, oh, did it go to all of these pages? Did it just look at the headline and the meta description? It did a really good job. So think business use case.
Starting point is 00:28:37 I love Boolean search terms or, you know, Boolean operators, right? that for a Google search for what you care about, right? Maybe it's, it's market research, maybe it's logistics, right? Put in your competitor names, whatever. I think there's so much utility for using just Boolean search and AI tools to quickly get you caught up, that on things instantly that would normally take a very long time. All right, let's keep this train moving, chew, chew. All right, here's one I really wanted to do, but we're not going to have time. All right. So I'll move on to another one here. All right. Let's do this one. I'm saying, so for this one, I'm going to use canvas. So this is another kind of update to the update.
Starting point is 00:29:20 So Google Gemini 2.5 Pro was just released less than a week ago. And then over the weekend, Google did a lot of other updates to Gemini 2.5 Pro. Number one, they said, all right, it's free for everyone. Number two, they rolled out canvas just about a day ago. So canvas, it's kind of similar. I actually think it brings the best of both worlds between open AIs canvas, which is more of like an interactive document editor that can render some code, along with Claude's artifacts feature, which can render just like any programming language. All right.
Starting point is 00:29:55 So in this instance, I'm saying I'm enabling Canvas and I'm saying create an HTML clone of Wikipedia, but give it heavy Chicago vibes, make it fully featured including clickable links and multiple pages that work. Make sure to include the most important Chicago things, right? I'm trying to have a little fun here, y'all. So let's see if this works. So first, it is writing the code. So like I said, it's great at coding.
Starting point is 00:30:26 All right. Fantastic. All right. So once it's done, which I don't think it should take very long, there's a preview tab as well. So when I start this canvas mode, it kind of takes up the full screen, but I can minimize it if I want. I'm going to pull this over a little bit so I can see.
Starting point is 00:30:48 All right. It should be done here pretty quickly as I take a sip on the coffee. And I'm scrolling through the live stream comments here, y'all. I'm going to see if there's any questions. All right, Josh said, look what I created this morning. Go check out what Josh created. Charles says, why don't you use chat GPT for news? I do.
Starting point is 00:31:08 I do as well. So that same URL, I did a whole entire show on how I did this, Charles, using chat GPD task. Monica says, what do you think are some of the best business use cases for this model? I still have a couple here, Monica, but I think one of the best ones working with PDFs, right? This has been getting accurate information extracted from PDFs and then being able to use that as a baseline, right? Because now I have all that text and maybe I'll do something with it that I extracted from a PDF. That's a simple no-brainer.
Starting point is 00:31:42 Everyone's working with PDFs and, you know, essentially extracting any information from a PDF if you need to recreate it, if you need to grab some information from there and use that as a start for, you know, creating content, right? So in my example, I had, you know, our everyday AI kind of sponsorship kit. I could then use that, copy and paste some of that information, go into deep research and say, hey, are these rates accurate according to 2025 popular podcasts or something like that? So that's, that's one small thing. One small thing I can do.
Starting point is 00:32:13 All right. Let's look at this. I'm going to zoom out. All right. Here we go. So we have our, let's let's see. How can I do this full screen here? I had this last night.
Starting point is 00:32:31 I thought I could. All right. Anyways, we have our Chicago. Wikipedia, literally one shot. All right. So it says, welcome to Shaikopedia. Your go-to source for all things,
Starting point is 00:32:45 Chicago from a Chicagoist point of view. Forget the end, it's likeopedia. This is where the real info is at. So this is a fully functioning Wikipedia clone, right? I can click, oh my gosh, it works. There's multiple pages on here. It's interlinked.
Starting point is 00:33:00 So I can click Deep Dish Pizza, right? And then I can, you know, at the bottom, it says, see also Chicago hot dog. I can click Chicago hot dog. The Chicago hot dog, also known as Chicago Red Hot, is a culinary masterpiece in a bun, right? No ketchup. All beef, right? This is so good.
Starting point is 00:33:20 This is so good. It literally created a very small version of Wikipedia, but Chicago style. And then the good thing is, I can go in, I can go in and change anything with natural language, right? And I can just say, you know, make it, make it way more Chicago and more 90s bulls references, right? Whatever. All right. So we're going to come back to that one here in a second and go on to our next use case. That one was fun.
Starting point is 00:33:51 What did you guys think? Pretty impressive, I thought. All right. Let's do this next one here. Okay, here we go. All right. I might not even have time to read this because it's a little bit of a longer problem. But I am essentially saying, you know, you're an analytics and research expert using Gemini 2.5.
Starting point is 00:34:14 Analyze the sentiment of online mentions of Apple over the past 30 days. And I'm giving it kind of step by step instructions. You know, I'm saying essentially look at all of the information on the open web that people are talking about Apple, right? Then identify five recurring themes or issues based on sentiment analysis, right? provide actionable recommendations for Apple's PR team to address any negative sentiment. And then ultimately, I'm using Canvas for this. And then I say create an interactive dashboard that displays your findings. Make sure to go into insane detail, ensuring accuracy and depth.
Starting point is 00:34:56 All right. So I actually did this one previously. And my first version, okay, let's see if it does tooth. Okay, look at this. Gemini was was a step ahead of me y'all it actually created two different canvas files within the same within the same kind of response here so okay so it's building our sentiment dashboard cool all right so first here's the sentiment analysis over the past 30 days so I want to again human in the loop look for accuracy this is correct right it says AI strategy execution concerns all right
Starting point is 00:35:37 So this is good. It gave us a good text-based report. It gave us actionable recommendations for Apple PR based on real-time, up-to-date information. It gave us five recurring themes, right? Vision pros lackluster reception. Oh, weird. If only someone would have told you that six months before it came out.
Starting point is 00:35:54 Oh, wait, I did. All right. So it gave us a great text document in Canvas. So one thing you'll notice about the Canvas, if you haven't used it, it does have some of those great chat, GPT, UI, UX, features where you can just change the length. You can change the tone. You can suggest edits. So I can just type live, right? So it's literally like a Google Doc, which is very impressive, right? Even just the canvas integration from the text-based perspective is extremely useful
Starting point is 00:36:24 for any business use case because right away, I can export this to Docs or I can just continue to type and work with it here. But it created two different canvases for me. Let's see how the other one turned out. Bam, love it. It actually turned out not as good as my first. I did demo this one first, but it gave me a very nice looking, kind of interactive dashboard, you know, nice colors. It says overall sentiment, mixed slash cautious. It says, while investor metrics, you know, example, alt index score of 64 out of 100 show underlying positivity. Recent public discourse reveals significant caution, primarily due to AI strategy concerns and competitive pressures. So a great job of just understanding overall sentiment over the last month of what people are
Starting point is 00:37:13 talking about Apple. Is it good or bad? Right. It gave us a green column and a red column. Key positive sentiments, key negative sentiments, top five recurring themes. That's great. I'm going to try just one more thing here. I'm going to zoom out and I'm going to say, I'm going to say make, let me copy this. and I'm going to say make this more interactive and visual. All right. We'll come back to that. Let's go back and see if our Chicagopedia got even more, got even more Chicago.
Starting point is 00:37:53 Let's see. It did. Fantastic. Now we have a dedicated sidebar column for the teams, bulls and bears, doubles. All right. Yeah, I'm from Chicago. I love this. This screams out like, you know, 90s Chicago. I love it. It says tall buildings and stuff. The lake, dibs. Dibs, you know the rules? Yeah. Throw your chair out. Reserve your parking spot on the street.
Starting point is 00:38:16 This Chicagopedia, I love it, right? And the cool thing is, if you didn't know, the code is all here, right? So, yes, you can render everything live inside Google Gemini 2.5 Pro in the canvas feature. But if you did want to take this offline, you can copy and paste this. Sometimes it won't work just copy and paste because you might need to install some certain libraries. Sometimes it will. It depends on kind of what languages are being used. This is strictly HTML. So I think in theory, I could just copy and paste this, put it on a website and it would be good to go. And y'all should I publish this ChicagoPedia?
Starting point is 00:38:53 This ChicagoPedia. I don't know. This one's kind of fun. I like this one. All right. Sandra already says she's going to rewatch this episode. So let's see. Jackie is asking great question.
Starting point is 00:39:10 Jackie, can it get past logins on social platforms? No. So all we can do with, you know, those different tools that Google was using to look at the web, that's the open web, right? So anything on social media for the most part is closed web. So even on Twitter, right, you're like, oh, everything's public. Well, you have to be logged in because, you know, certain, there's certain restrictions specifically on social media that a lot of scraping sites or, you know, tool use, kind of tool use or internet use tools from AI large language models cannot pick up that information. Great question, though.
Starting point is 00:39:50 Love ChicagoPedia. Yeah, I do too. All right. I have so many examples, y'all. And I'm surprised that many of them are working. So let me scroll through here. And I'm going to try to find maybe something that's a little more impressive. Okay, here.
Starting point is 00:40:05 Here's one. I think this could be good. All right. So I'm saying let's go ahead and launch a new window here in Gemini 2.5 Pro. All right. I got way zoomed out. So I'm saying create a visual memory game or interactive quiz that will help me learn and memorize this content. All right.
Starting point is 00:40:26 So then what I'm going to do is I'm going to go to the. your everyday AI page. I'm going to click on episodes. I did mention this, but you can go read, watch and listen to anything on our website. So, you know, I'm going to our episode from Monday where we did the AI News That Matters, right? If you didn't know, you can listen to the podcast on the website for free. You can watch the video for free. We have a little write-up from some of the key points, you know, and then we have a complete transcript as well. All right, so all I'm do i'm going to copy and paste all of this information all right i'm going back into uh google gemini i'm just pasting this and i'm saying create a visual memory game or interactive quiz that will help me
Starting point is 00:41:07 learn and memorize this content all right i'm going to click enter uh and let's see what happens all right so talk about business use cases right uh how about making an onboarding fun you know you have all these long boring uh onboarding docks right make a fun game out of it, right? So this is what I'm doing. I love finding new ways to learn. I love learning with notebook LM, the audio overviews. I love notebook LM's new mind map feature,
Starting point is 00:41:36 but I'm always finding new ways to learn. One problem with AI is making it harder for me to retain information. I learn way more per day than I did pre-LLMs, but I also, that means I forget more. So I'm always looking for new and better ways to learn and retain important information. So again, think, You can use Gemini 2.5 pro to automatically curate, you know, certain information that you might want to, that you might want to learn. In this case, I'm just using a transcript from a podcast.
Starting point is 00:42:09 All right. So let's look, see what it did. It's done. Oh, gosh, this is going to be embarrassing. All right. So it created a quiz. There's 15. 15.
Starting point is 00:42:18 Hey, you guys want to do the first couple questions, live stream audience? All right. Let's just do a couple questions together. See if you tuned in. Let's see if you tuned in Monday. So it says AI News quiz. And just for you all, this looks pretty good. It's got this kind of purplish background, very like web 2.0.
Starting point is 00:42:38 There's hover animations. It's pretty slick. It looks nice. It's not some ugly, janky, you know, 1990s looking quiz. It looks really good. All right. So, live through your audience, let's play along. We'll just do a couple questions.
Starting point is 00:42:51 So it says the deterministic aspect mentioned in Microsoft's agent. flows aims to reduce issues like is it high cost hallucinations uh language translation errors or slow processing speed what do you guys think i'm going to take a sip all right i'm going to guess a i hallucinations yay it said correct cool all right so it works that's the thing i just one-shotted an interactive quiz based on i don't know couple thousand words and it took like a minute if this doesn't change how you think you and your team can interact with you, even your own internal docs. I don't know what else to say. Next question.
Starting point is 00:43:34 Live stream audience. Who's going to get it first? All right. This is meta, but not meta like Facebook. Meta as in we're using Gemini 2.5 Pro to ask about Gemini 2.5 pro. What key feature allows Gemini 2.5 pro to process extremely large amounts of text, audio, images and code. Ooh.
Starting point is 00:43:57 Ooh, this one's, this one's a little tricky. So cross layer transcoder, deep reasoning agents, deterministic logic, or one million token context window. This one's very actually interesting because it didn't just make things up. The wrong answers, which hopefully I get this right, live stream audience gets your vote in. The wrong answers are actually key terms from other announcements from clause. and from Microsoft, but we're asking about, we're asking about Gemini 2.5 Pro,
Starting point is 00:44:30 I believe it's the 1 million token context window. Oh, good, I got it right. All right, let's do one more. All right. It says Open AI is reportedly nearing a funding round of what massive amount potentially led by SoftBank. Okay, this one's actually a little tricky because there's a total amount.
Starting point is 00:44:48 So is it $40 billion, $10 billion, $20 billion or $33 billion? There's actually a total amount of funding. And then there's a certain amount of funding. And then there's a certain amount of funding that SoftBank is reportedly on the line for. But that's actually two different amounts. One amount is if OpenAI does successfully transition from a nonprofit to a for-profit. And the other amount is if they don't. So there's technically three terms, a total fundraising term, SoftBank, A, if they do convert to for-profit,
Starting point is 00:45:13 B, if they don't. So the question is, Open AI is reportedly nearing a funding round of what massive amount. The amount of the funding round is $40 billion. All right. The cool thing is I can say something like make it more, you know, make it even more interactive and detailed, maybe some slight animations, make it look and function better, right? That's the coolest thing. I didn't write a single line of code. I don't need to.
Starting point is 00:45:50 I can control this with just natural language. Like, yo, LLM, make this better. Make it shinier. Make it blue. Make it harder. Make it easier. Make it for pros. Make it for amateurs, right?
Starting point is 00:46:04 Create a graduated model. Right? First, you know, give me 10 questions that are much easier. Then, you know, help me level up or, you know, turn it into more of a video game, right? There's so many things that you can do. All right. I'm going to give this a second to finish. Let's check in.
Starting point is 00:46:20 Oh, my gosh. Look at this, y'all. So our Apple sentiment analysis, remember, I just in natural language, what did I say? I just said make it more interactive and visual. It improved it by a lot. So there's some things that didn't fully render, right? So there's some code that says like more rounding. But overall, it made this look much, much better.
Starting point is 00:46:46 It gave it kind of these, these gauges and barometers. with certain, you know, filling. It just made it look much better. So these are toggles, little toggles, even though there's not a lot of information in them. So really good, really good. All right, let's see. All right.
Starting point is 00:47:06 It's already done. Our news quiz is done. It added a status indicator. Okay, now it's actually hard. I don't know. What episode number and date was featured in the I knew summary. Oh, gosh, without looking this up, what was it?
Starting point is 00:47:23 I think it was $4.93. Oh, good. I got it right. Okay. So, okay, unfortunately, the status indicator did not light up, but I could change that. All right. So very impressive. Should we do one more, y'all?
Starting point is 00:47:37 Should we do one more? Should we wrap this up? Let me know. You guys, you guys all got this right. I'm looking at our live, at our live comments. You guys got it right. You must have all watched this episode. All right.
Starting point is 00:47:51 All right. You guys said one more. Let me just go ahead. Let me see if I can get something that I think is maybe impressive. Okay, cool. Let's do this. We're going to do one more quick one here. So this, you know, we talk about use cases.
Starting point is 00:48:17 I just randomly threw one out. I'm like, how about you make your internal documents a little better, a little more fun? Right. So here I'm saying, essentially, you're an. HR expert using Gemini 2.5 Pro, you know, hey, you, you work at IBM, create a manual with standard operating procedures for new employees. So essentially, I'm saying create an onboarding form for new employees at IBM. And then also, you know, an eight question quiz that covers key SOP elements, ensure all recommendations are based on real IBM training methodologies.
Starting point is 00:48:58 So I'm wondering if it's actually going to go pull this and find this information from the web. I guess I'll have to verify this later, right? Just because human and expert in the loop doesn't mean I need to do that live, right? I'm not going to post this and say it's perfect and working. But you'll see already, one thing I love that's a little different with the canvas inside Google Gemini versus some canvas features or functionality in open AI or Anthropics artifacts is it can create multiple kind of canvases. Is it canvases or canv? Right? I think it's canvases.
Starting point is 00:49:30 Multiple canvases at once. So the first one is just this onboarding material. Okay. So it's creating a an SOP with pre-boarding day one, week one, roll clarity, compliance and ethics. Right. So it's doing the kind of like boring, right. All right. Here's your here's your text base, you know, content. And then this probably won't be done yet. But it's, let's see, there we go. It's already done. Okay. It created an interactive version of this simple, you know, 10 step SOP onboarding
Starting point is 00:50:07 for new hires at IBM, right? So it has our onboarding SOP. It's interactive. It has these tabs. I can click weekly tasks, weekly schedule and tasks. And there's toggles with drop downs. Here's week one foundations and. setup, week two, role clarity and tools. This is really good. It's all interactive. It works.
Starting point is 00:50:30 And here's a quiz. Is this quiz? I don't think the quiz is going to work. Let's see. What is the primary focus during the first week of onboarding at IBM, leading a major product, completing essential compliance training and initial setup, presenting a strategy report senior leadership? I'm going to guess it's the middle one. All right. So it doesn't say unless I have to click, okay, there is a thing to submit the quiz. So I'm just going to click one. I'm wondering if it's going to tell me what's right and wrong or give me a score. This would be very impressive. A multi-step quiz embedded inside an accordion. All right. So it didn't tell me which ones were right or wrong on this one, probably because
Starting point is 00:51:10 there's no database. And then it has a checklist as well. This is very cool. So this is my onboarding milestones checklist, right? And when I check it, it says two of 10. I check one more, three of 10. let's see what happens when I finish it out. Boom, says 10 at 10. Very impressive, y'all. All right. We covered a lot. I know this episode was all over the place when we talk about different use cases for Gemini 2.5 Pro.
Starting point is 00:51:36 So I'll say this. It's not perfect, right? It's not perfect. The ceiling is high. The floor is finicky, right? But as long as you, the human, you the business leader out there are paying attention. are being patient, are properly prompting Gemini 2.5. And you know, you might have to dip your finger a little bit into Google's AI studio.
Starting point is 00:52:01 Extremely, extremely powerful state of the art, multimodal, multifaceted large language model in Gemini 2.5 pro. And the use cases are tremendous, right? It's actually baffling how many, even just what we went over here live, right? I didn't really plan these. I didn't refine them. I wanted to give you guys just the nitty gritty, right? Let's see some mistakes.
Starting point is 00:52:31 Let's try to improve it a little bit. But if your brain isn't churning, if one of these didn't hit home, you got to check, you got to check for a pulse, y'all. Because what we just showed in this one quick episode podcast audience, I'm sorry. I know this one was a little bit more visual. I know I didn't do a great job at, you know, describing everything, but, you know, make sure you go watch this one. But if you didn't get at least one idea on how your business, how your role, how your
Starting point is 00:52:58 department can fundamentally change by using Google Gemini 2.5, you got to rewatch this because it's in there, right? So think, what public data do you have? How can you make old documents? How can you bring them to life, right? In the same way that we talk about large language. models becoming multimodal, right? I think businesses also need to start taking that same approach, even for their own
Starting point is 00:53:25 internal document. We don't live in, like, we don't live in a text-based world anymore, right? We can create games. We can create interactive quizzes. We can create, you know, visualizations and business dashboards now with zero coding knowledge, right? Before, you might have to have a team of developers, some people in BI, right? Now you can just copy and paste.
Starting point is 00:53:47 That was one of the things I wanted to do, but we ran out of time. Copy and paste a bunch of data. Create a business dashboard, right? And you're off to the races, right? You already have ways that you can instantly use generative AI to grow your company and your career. That's what it's all about. All right. I hope this one was helpful, y'all.
Starting point is 00:54:06 Part two. Again, maybe you just listen to this one for the first time. Make sure to go back one episode. Listen to part one where we go over more of the details, the bullet points, everything that's kind of under the hood, how the model works, all that. But hopefully in this live demo example, sometimes they work. Sometimes they don't. I hope this was helpful.
Starting point is 00:54:24 And I hope this is sparking some, some ideas in your brain on how you can use, not just Gemini 2.5 Pro, but just large language models in general, right? If you're not already using generative AI in large language models, day to day for every aspect of your business, you've got to rethink how you are working. You need to rethink your role, rethink your department, rethink your company, rethink your Rethink what it means to be a knowledge worker. That's what all of us are. All right.
Starting point is 00:54:49 So it starts here, but you need to go to your everyday AI.com. Sign up for the free daily newsletter. We're going to be recapping. Today's post, you know, if some of y'all shared some examples, maybe I'll throw one in the newsletter as well. So thanks for tuning in. Hope to see you back tomorrow. And every day for more everyday AI.
Starting point is 00:55:07 Thanks, y'all. Meet Firefly AI assistant. Now live in Adobe Firefly, the Allman One Creative AI Studio. Just describe what you want to create in your own words, and the assistant handles the rest, orchestrating multi-step workflows across Adobe Creative Cloud apps, including Photoshop, Premiere Express, and more in one conversational interface. You direct the outcome while the assistant accelerates execution.
Starting point is 00:55:36 Stand control with the ability to step in and refine at any time. See it today at firefly.adop.com. And that's a wrap for today's edition of Everyday AI. Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rate. It helps keep us going. For a little more AI magic, visit your everyday AI.com and sign up to our daily newsletter so you don't get left behind. Go break some barriers and we'll see you next time.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.