By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google DeepMind AI Controls Entire Robot Body
Google DeepMind announced on Thursday that its latest AI model, Gemini Robotics 2, is now capable of controlling the entire body of a humanoid robot. This represents a significant advancement from its predecessor, which was limited to controlling only the upper body. The new model's capabilities extend to "whole-body motions," encompassing movements from the robot's feet to its fingertips. This enhanced control is expected to enable humanoid robots to perform a wider range of complex tasks that require coordinated full-body movements, such as walking, grasping objects with precision, and navigating varied terrains.
Gemini Robotics 2 builds upon the foundational work of earlier iterations, which focused on developing AI systems that could understand and interact with the physical world. The development of AI models capable of sophisticated robotic control is a key area of research for companies like Google DeepMind, aiming to bridge the gap between artificial intelligence and physical embodiment. Previous research in robotics has often focused on specific tasks or body parts, but a unified model that can manage the entire kinematic chain of a humanoid robot presents a more holistic approach to robotic autonomy. The ability to control the entire body allows for more nuanced and dynamic interactions with the environment, potentially leading to robots that are more adaptable and capable in real-world scenarios.
The implications of this development are far-reaching, particularly for industries that are exploring the use of humanoid robots for tasks in manufacturing, logistics, healthcare, and domestic assistance. By enabling more natural and comprehensive movement, Gemini Robotics 2 could accelerate the deployment of robots in environments that were previously too complex or dangerous for automated systems. For instance, robots equipped with this AI could potentially assist in disaster relief operations, perform intricate assembly line tasks, or provide support for elderly individuals in their homes. The advancement signifies a step towards more versatile and human-like robotic capabilities, moving beyond specialized functions to general-purpose physical interaction.
Google DeepMind has been at the forefront of AI research, with its Gemini family of models designed to be multimodal and capable of understanding and processing various types of information, including text, images, audio, and video. The application of this advanced AI to robotics highlights the company's strategy of integrating its cutting-edge AI research into practical applications. The development of Gemini Robotics 2 underscores the growing trend of AI systems not only processing information but also acting upon it in the physical world. This progress in robotic control is a critical component in the broader pursuit of artificial general intelligence (AGI), where AI systems possess human-like cognitive abilities across a wide range of tasks.
Original source — read the full reporting at the publisher:
Read on The VergeGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.