Google's DeepMind has unveiled the latest iteration of its AI model, Gemini Robotics 2, which is now capable of controlling humanoid robots to perform complex tasks such as screwing in lightbulbs and tying trash bags. This advancement is particularly significant for the architecture, engineering, construction, and manufacturing (AECM) industries, where automation and intelligent robotics are increasingly becoming essential.
What Happened
Google DeepMind's Gemini Robotics 2 integrates multiple AI models to control various robotic forms, including humanoids. This system combines a vision language model (VLM) and two vision language action (VLA) models to understand images, videos, and human commands, enabling robots to interact with their environment effectively. In demonstrations, robots like Apptronik’s Apollo 2, equipped with Sharpa's hands, autonomously tidied shelves, showcasing Gemini's capabilities in real-world applications.
The AI model is trained through a combination of human teleoperation, video examples, and simulations, though it still requires specific training for each task. This development underscores Google's commitment to pushing AI beyond digital confines, aiming for physical artificial general intelligence (AGI).
What This Means for Your Business
For AECM professionals, the introduction of Gemini Robotics 2 could revolutionize operations by integrating AI-driven robotics into the workforce. This technology holds potential for significant cost savings and efficiency improvements, particularly in repetitive or hazardous tasks. Contractors and manufacturers can leverage these robots to reduce human error and increase productivity, resulting in a more streamlined operation.
However, the integration of such advanced robots also demands attention to compliance and safety standards. As these robots become more prevalent, aligning with cybersecurity measures like the Cybersecurity Maturity Model Certification (CMMC) and NIST guidelines will be critical to ensure safe deployment and operation.
What US Operators Should Watch
US operators should monitor the development of safety benchmarks like Google's ASIMOV-Agentic, which is designed to measure the safety of AI systems in robotics. Understanding these safety standards will be crucial as more companies adopt AI-driven robotics to avoid potential risks associated with unpredictable AI behavior.
Additionally, stakeholders should stay informed about federal deadlines and funding opportunities that may arise with the growing adoption of AI and robotics in industry. Being proactive in these areas will ensure competitive positioning and optimal return on investment.
Source: https://www.wired.com/story/google-gemini-can-control-humanoid-robots/
Is your firm ready for what’s next?
VisioneerIT helps AECM and government contractors modernize operations, achieve compliance, and implement AI.
Explore VisioneerIT Solutions →