, ,

Google I/O impresses with immediate Gemini response to GPT-4o

On May 13, 2024, OpenAI showcased its new model GPT-4o (o for omni) a multi-modal model with vision, voice, text, and video.

Not to be outdone, Google showcased its upgraded Gemini model 1.5 at its I/O conference a day later, along with other applications. This shows that Google has clearly upped its AI game to compete with OpenAI.

More advanced Search with multi-step reasoning and AI overviews:

@verge

The new “AI Overviews” from Google is rolling out today with more multi-step reasoning that can handle complex search questions. #google #googleio #longervideos #techtok

♬ original sound – The Verge

With vision, Google can analyze and search video:

@google

You’ll soon be able to ask your questions with a video, right in Google Search. Coming soon to Search Labs. #GoogleIO

♬ original sound – Google – Google

Project Astra can be your universal AI agent or assistant:

@google

How many iconic landmarks can Project Astra recognize in our travel journal? 📍Shot in one take in real time.

♬ original sound – Google – Google

Google’s text to video generator called Veo (not to be confused with Veoh, the video sharing platform):

@google

Veo can create high-quality, high-def videos from a text, image and video prompt. #GoogleIO #AI

♬ original sound – Google

Google’s ad reintroducing itself (sounds like Jay Z)

@google

The Gemini era is here — helping you do more with the magic of Google AI

♬ original sound – Google – Google

Leave a Reply


Discover more from Chat GPT Is Eating the World

Subscribe now to keep reading and get access to the full archive.

Continue reading