Top News

Industry Briefing

A single destination for timely, editor-curated robotics news from around the world.

Google DeepMind Launches Gemini Robotics 2: A Comprehensive Robot with Advanced Capabilities

Google DeepMind Launches Gemini Robotics 2: A Comprehensive Robot with Advanced Capabilities

On August 2, Google DeepMind introduced Gemini Robotics 2, the latest version of its visual-language-action model. This new model enhances full-body control, advanced dexterity, and multi-robot collaboration, marking a significant advancement from its predecessor, which only managed upper-body tasks. Gemini Robotics 2 allows the Apptronik Apollo 2 humanoid robot to autonomously execute complex tasks, such as placing a watering can into a designated bucket, requiring coordinated decision-making and balance. While there is room for improvement in movement speed, this development represents a crucial step towards handling more intricate real-world tasks. The model also introduces multi-robot collaboration, enabling different robots to communicate and work together on complex workflows. With the ASIMOV-Agentic benchmark, Gemini Robotics 2 excels in safety compliance and human proximity detection, ensuring safe operation. The Gemini Robotics On-Device 2 is designed for applications requiring offline functionality, adapting to new robotic forms with minimal data and time.

Humanoid Robots Robotics Technology AI Automation
Google DeepMind Introduces Gemini Robotics 2 for Enhanced Humanoid Control

Google DeepMind Introduces Gemini Robotics 2 for Enhanced Humanoid Control

Google DeepMind recently launched Gemini Robotics 2, an advanced version of its vision-language-action model. This iteration features intelligent whole-body control, enabling robots to perform complex tasks such as walking, crouching, and object manipulation. The model can adapt to new robotic bodies quickly, enhancing its versatility. The significance of Gemini Robotics 2 lies in its ability to expand physical AI capabilities, allowing robots to execute intricate movements and collaborate with other robots. This development is crucial for achieving the finesse required for real-world applications, such as household chores and workplace tasks. Looking ahead, the focus will be on further improving movement speed and dexterity. Gemini Robotics 2 represents a pivotal advancement in robotics, as it can control sophisticated end effectors like the SharpaWave hand, enabling delicate actions. No further timeline was disclosed at the time of publication.

Artificial Intelligence Artificial Intelligence / Cognition Design / Development End Effectors / Grippers Humanoids News
Introducing Google DeepMind's Gemini Robotics 2: A Leap in Robot Adaptability

Introducing Google DeepMind's Gemini Robotics 2: A Leap in Robot Adaptability

Google DeepMind has unveiled Gemini Robotics 2, marking a significant advancement in robotics. This new intelligence layer enables robots to achieve whole-body control, enhanced dexterity, and collaborative capabilities among multiple robots. The introduction of Gemini Robotics 2 is crucial as it represents a step towards creating truly adaptable robots that can perform complex tasks in dynamic environments. This technology could revolutionize various applications, from industrial automation to personal assistance. Looking ahead, the development of Gemini Robotics 2 will be closely monitored as it progresses. The potential for multi-robot collaboration and advanced dexterity opens up new possibilities in robotics, making it a key area of interest for future innovations.

Humanoid-robots Video-friday Quadruped-robots Robot-videos Drones Home-robots
Google DeepMind Unveils Gemini Robotics 2 AI Model for Humanoid Robots

Google DeepMind Unveils Gemini Robotics 2 AI Model for Humanoid Robots

On July 30, 2026, Google DeepMind announced the latest AI model for humanoid robots, Gemini Robotics 2. This model advances from upper-body control to full-body control, enabling walking, bending, and advanced dexterity, as well as collaborative tasks among multiple robots. Gemini Robotics 2 translates visual and language inputs into motion control, supporting full-body and dual-arm robot operations. It features enhanced physical reasoning capabilities with Gemini Robotics ER 2, which manages human communication and multi-step task planning, and Gemini Robotics On-Device 2, which adapts quickly to new robots with minimal data. The model demonstrated its capabilities using Apptronik's humanoid robot Apollo 2, performing tasks such as walking to a table, lifting objects, and accurately storing them. With a 22-degree-of-freedom hand, it can execute delicate tasks like tying knots. The introduction of the ASIMOV-Agentic benchmark emphasizes safety, allowing the robot to detect human presence and stop safely, marking it as the safest model to date.

Google DeepMind Launches Gemini Robotics 2 for Enhanced Autonomy in Humanoid Robots

Google DeepMind Launches Gemini Robotics 2 for Enhanced Autonomy in Humanoid Robots

Google DeepMind has introduced Gemini Robotics 2, a suite of AI models aimed at enhancing the autonomy of humanoid and other robots through advanced whole-body control and dexterity. This system is currently showcased on Apptronik's Apollo 2 humanoid robot, which can execute full-body autonomous movements such as walking and manipulating objects while reasoning through complex tasks in real time. The significance of this development lies in its ability to enable robots to move beyond pre-programmed tasks, adapting to dynamic environments. Gemini Robotics 2 serves as an 'intelligence layer' for robots, facilitating coordination from 'feet to fingertips' and improving dexterity with tasks like tying trash bags or unscrewing light bulbs. This advancement builds on the collaboration between Google DeepMind and Apptronik, which recently launched Robot Park to gather data for training future AI models. Looking ahead, the introduction of multi-robot collaboration capabilities allows different robot types to work together on complex workflows. Gemini Robotics On-Device 2 is designed for industrial settings with limited cloud connectivity, enabling rapid adaptation to new hardware. No further timeline was disclosed at the time of publication.

Artificial Intelligence Humanoids News Apptronik Apollo 2 embodied ai Gemini Robotics 2
Google DeepMind Launches Gemini AI to Enhance Robot Dexterity and Coordination

Google DeepMind Launches Gemini AI to Enhance Robot Dexterity and Coordination

Google DeepMind has introduced a new AI model named Gemini, designed to improve the dexterity of humanoid robots. This advancement enables robots to coordinate movements throughout their bodies, enhancing their ability to perform complex tasks. The introduction of Gemini AI is significant as it represents a step forward in enabling robots to reason, plan multi-step tasks, and adapt to human environments. This capability is crucial for the integration of robots into everyday settings, where they must interact seamlessly with humans and their surroundings. Looking ahead, the focus will be on how Gemini AI can be further developed and implemented in various robotic applications. No further timeline was disclosed at the time of publication.

Google Unveils Gemini Robotics 2 for Advanced Humanoid Robot Control and Reasoning

Google Unveils Gemini Robotics 2 for Advanced Humanoid Robot Control and Reasoning

Google has launched Gemini Robotics 2, a sophisticated robotics model that enables humanoid robots to perform full-body movements and complex tasks. This system allows for coordination among multiple robots and adapts to various robot designs with minimal training. The new capabilities extend beyond simple manipulation to include walking, bending, and balancing. The significance of Gemini Robotics 2 lies in its ability to facilitate multi-step reasoning and adaptability in unfamiliar environments, setting it apart from traditional robots. This model can control both humanoid robots and dual-arm systems, enhancing their functionality for household and industrial applications. Demonstrations showcased its ability to manage tasks such as picking up objects and performing intricate actions with dexterous robotic hands. Looking ahead, Google is also releasing Gemini Robotics ER 2 for high-level reasoning and Gemini Robotics On-Device 2 for local operation without cloud reliance. The introduction of multi-robot collaboration capabilities allows different robots to work together efficiently, marking a notable advancement in robotic technology. No further timeline was disclosed at the time of publication.

AI and Robotics
Google DeepMind Unveils Gemini Robotics 2 for Full-Body Control of Humanoid Robots

Google DeepMind Unveils Gemini Robotics 2 for Full-Body Control of Humanoid Robots

Google DeepMind has announced the launch of Gemini Robotics 2, an advanced AI model capable of controlling humanoid robots from 'feet to fingertips.' This new version enhances the robot's ability to perform a variety of actions, including walking, crouching, and manipulating objects, significantly expanding its functional capabilities. The importance of this development lies in its potential to enable humanoid robots to undertake more complex, real-world tasks that require whole-body coordination. With improved dexterity, Gemini Robotics 2 allows robots to perform intricate tasks such as sealing bags and unscrewing lightbulbs, marking a significant advancement in robotic manipulation. Looking ahead, Google DeepMind is also enhancing its Gemini Robotics ER model, which aids robots in analyzing their environment and executing multi-step tasks. This update aims to improve collaboration among different types of robots, enhancing safety features and operational efficiency. No further timeline was disclosed at the time of publication.

AI Google News Robot Tech
AgiBot WITA-Omni Achieves Top Score on DailyOmni Benchmark, Surpassing Major Competitors

AgiBot WITA-Omni Achieves Top Score on DailyOmni Benchmark, Surpassing Major Competitors

AgiBot WITA-Omni has achieved a score of 85.21 on the DailyOmni benchmark, securing first place in 6 out of 8 indicators. This performance surpasses notable competitors such as Google Gemini, ByteDance Doubao, and Alibaba Qwen in the realm of embodied cross-modal understanding. The significance of this achievement lies in the innovative Thinker-Talker-Actor architecture utilized by AgiBot WITA-Omni. This architecture effectively synchronizes speech, action, and expression on a single timeline, enhancing its capabilities in cross-modal understanding and interaction. Looking ahead, the performance of AgiBot WITA-Omni on the DailyOmni leaderboard may influence future developments in embodied AI technologies. No further timeline was disclosed at the time of publication.

Technology
Google rolls out Gemini Omni Flash for autonomous video creation across apps

Google rolls out Gemini Omni Flash for autonomous video creation across apps

Google has initiated the rollout of its latest multimodal AI model, Gemini Omni Flash, which is designed to enhance user interactions across various platforms. This launch began in October 2023 and aims to integrate advanced AI capabilities into Google's suite of services. The motivation behind this development is to provide users with a more seamless and efficient experience by allowing the model to process and respond to inputs in multiple formats, including text, images, and voice. By leveraging cutting-edge technology, Gemini Omni Flash is expected to significantly improve the way users engage with Google's applications, making interactions more intuitive and responsive. The rollout is part of Google's ongoing commitment to innovation in artificial intelligence, positioning the company at the forefront of the AI landscape.

Google’s Gemini Omni turns images, audio, and text into video — and that’s just the start

Google’s Gemini Omni turns images, audio, and text into video — and that’s just the start

Google has unveiled its latest innovation, Gemini Omni, a multimodal model designed to enhance user interaction by reasoning across various formats, including text, images, audio, and video. This cutting-edge technology allows users to generate and edit videos through straightforward conversational prompts, with the initial feature being Omni Flash. The launch of Gemini Omni marks a significant advancement in artificial intelligence, aiming to simplify content creation and editing processes for users. The model is trained on data available up to October 2023, ensuring it incorporates the most recent developments in AI capabilities. This initiative reflects Google's commitment to pushing the boundaries of technology and improving user experiences in digital content creation.

Media & Entertainment AI Google Veo google io 2026 google gemini omni
Google races to put Gemini at the center of Android before Apple’s AI reboot

Google races to put Gemini at the center of Android before Apple’s AI reboot

Google is leveraging its latest Android rollout to establish Gemini as the central artificial intelligence layer across a range of devices, including smartphones, Chromebooks, laptops, and vehicles. This strategic move aims to enhance user experience by integrating advanced AI capabilities into everyday technology. The rollout, which began in October 2023, reflects Google's commitment to innovation and its vision of a seamlessly connected ecosystem. By embedding Gemini across multiple platforms, the tech giant seeks to streamline operations and improve functionality, making AI more accessible and beneficial for users in their daily lives.

Boston Dynamics and Google Bring Gemini AI to Spot Robot for Smarter Facility Inspections

Boston Dynamics and Google Bring Gemini AI to Spot Robot for Smarter Facility Inspections

Boston Dynamics has announced the integration of Google's Gemini Robotics AI into its robotic platform, Spot. This advancement allows Spot to perform automated industrial gauge readings and conduct visual inspections of facilities. The collaboration aims to enhance operational efficiency and accuracy in industrial settings, addressing the growing demand for automation in various sectors. The integration is expected to streamline processes and reduce the need for human intervention in routine inspections, thereby improving safety and productivity. This development comes as part of Boston Dynamics' ongoing efforts to leverage cutting-edge technology to transform the capabilities of its robots, with the integration officially rolling out in late 2023.

Factory / Plant Maintenance
Google DeepMind Unveils Gemini Robotics-ER 1.6: A Leap in Spatial Reasoning and Industrial Utility

Google DeepMind Unveils Gemini Robotics-ER 1.6: A Leap in Spatial Reasoning and Industrial Utility

DeepMind has unveiled its latest and most sophisticated embodied reasoning model, which incorporates a cutting-edge "agentic vision" system designed specifically for industrial inspections. This new technology aims to improve the accuracy and efficiency of multi-view success detection in various industrial applications. The launch, which took place recently, marks a significant advancement in artificial intelligence capabilities, reflecting DeepMind's commitment to enhancing operational processes across industries. By leveraging advanced machine learning techniques, the model is expected to streamline inspection workflows and reduce the likelihood of errors, ultimately driving productivity and safety in industrial environments.

Google Gemini Gemini Robotics-ER 1.6 google-deepmind Boston Dynamics
Agile Robots and Google DeepMind Partner to Bring Gemini to the Factory Floor

Agile Robots and Google DeepMind Partner to Bring Gemini to the Factory Floor

Agile Robots, based in Munich, has entered into a strategic research partnership with Google DeepMind to enhance industrial robotics. This collaboration seeks to integrate Gemini Robotics foundation models with advanced industrial hardware, addressing the prevalent issue of data bottlenecks in the field. By leveraging a scalable AI flywheel, the partnership aims to improve the efficiency and effectiveness of robotic systems in various industries. The initiative highlights the growing intersection of artificial intelligence and robotics, as both companies work together to push the boundaries of technology and innovation.

DeepMind US Agile ONE Europe Google google-deepmind
Google DeepMind Gives Robots a 'Thinking' Brain with Agentic Gemini 1.5 Models

Google DeepMind Gives Robots a 'Thinking' Brain with Agentic Gemini 1.5 Models

Google DeepMind has introduced Gemini Robotics 1.5, an advanced AI framework aimed at enhancing the capabilities of robots. This new system allows robots to evolve from merely following commands to becoming 'physical agents' capable of reasoning, planning, and acquiring skills across various hardware platforms, including Apptronik's Apollo humanoid robot. The announcement marks a significant step in the development of intelligent robotics, reflecting the company's commitment to pushing the boundaries of artificial intelligence. By enabling robots to learn and adapt, DeepMind seeks to revolutionize the way machines interact with their environments and perform complex tasks. The unveiling of this framework comes as part of a broader trend in the tech industry to create more autonomous and versatile robotic systems.

ai-agents vla Gemini Apptronik google-deepmind robotics
Google DeepMind Unveils On-Device Gemini Robotics, Pushing AI Closer to Autonomous Dexterity

Google DeepMind Unveils On-Device Gemini Robotics, Pushing AI Closer to Autonomous Dexterity

Google DeepMind has introduced Gemini Robotics On-Device, a cutting-edge vision-language-action model that operates directly on robotic hardware. This innovative technology is designed to enhance the performance of autonomous robots by minimizing latency and increasing robustness across diverse environments. By enabling advanced dexterous manipulation, Gemini Robotics On-Device aims to equip a new generation of robots with the ability to quickly adapt to various tasks. The launch marks a significant step forward in the development of more efficient and capable robotic systems, reflecting DeepMind's commitment to pushing the boundaries of artificial intelligence in practical applications.

vla Apollo AI Humanoids Apptronik google-deepmind
RobotToday Initiative

Robotics needs a service framework.

RSF defines a common language for robot service capability, lifecycle operations, certification pathways, and service-provider networks.

inJoin the RobotToday community on LinkedIn

Daily robotics news, in-depth analysis, conference highlights, and discussions with professionals worldwide.