Dozens of New Features Announced at Google I/O 2024
- VoxPop/Misrohatun Hasanah
Jakarta – Google I/O is an annual conference for developers held by Google, where the company regularly introduces the latest version of its Android operating system.
Additionally, the American tech giant often releases new devices from its Google Pixel smartphone line at this event.
This year, artificial intelligence (AI) will be the main theme of Google I/O. The conference officially took place on Wednesday.
Google CEO Sundar Pichai announced several innovations and projects that will shape the future of technology. One of the most highlighted topics is Gemini.
Kepala Eksekutif Google Sundar Pichai.
- Twitter/@sundarpichai
Since its announcement at Google I/O 2023, Gemini has continued to evolve. Two months ago, Google introduced Gemini 1.5 Pro, which can handle 1 million tokens in a single query.
"Google is fully in the Gemini era. We have also brought Gemini's breakthrough capabilities across our products in a powerful way. We will showcase examples in Search, Photos, Workspace, Android, and more," said Sundar Pichai, quoted from Google's YouTube channel.
So, here are new AI features were showcased by Google.
Project Astra
Google DeepMind unveiled Project Astra, which aims to revolutionize the future of AI assistants with video comprehension capabilities.
Project Astra aims to develop a universal AI agent that can assist in everyday life.
During the demonstration, this research model showed its ability to identify objects producing sound, provide creative alliteration, explain code on a monitor, and find misplaced items.
Project Astra also demonstrated its potential in wearable devices, such as smart glasses, where it can analyze diagrams, suggest repairs, and generate intelligent responses to visual stimuli.
In the future, Gemini will use Project Astra's video comprehension capabilities to shape the future of AI assistants.
Veo
Veo (text-to-video generator) can produce high-quality 1080p resolution videos lasting more than one minute.
The model can better understand natural language to create videos that better represent the user's vision, according to Google.
