Google’s latest AI model, Gemini Robotics 2, is set to revolutionize the way robots interact with humans and their environment. Combining vision language models (VLM) for understanding visuals, two VLA models for movement, this system allows robots to perform complex tasks autonomously.
The potential applications are vast: from household chores to industrial automation. In demonstrations, Apollo 2, a robot equipped with hands from Sharpa, showcased its ability to organize shelves. Such advancements hint at the future where AI isn’t just digital but physical too.
However, alongside these exciting developments come safety concerns. Parada warns of unexpected and potentially dangerous actions by frontier AI models. Google is addressing this through a multi-layered approach including new benchmarks like ASIMOV-Agentic to ensure safety.
The journey towards full physical AGI (Artificial General Intelligence) is fraught with challenges but also immense promise. Google hopes to develop an AI ecosystem akin to Android, suitable for myriad robots. Yet, the road ahead requires careful navigation as we integrate AI into the fabric of our daily lives.







