Google DeepMind Ships Three Physical AI Models For Whole Body Control, Dexterity And Multi Robot Collaboration
Google DeepMind released Gemini Robotics 2, comprising three models including a vision-language-action model for humanoid control and an on-device VLA that adapts to new robot bodies in hours. Only the embodied reasoning model ER 2 is publicly available, while one checkpoint operates both Apptronik Apollo 2 and Franka Duo robots.
Read full story →Thinking Machines debuts Inkling Small open source AI model nearing performance of predecessor at about 1/4 size
Thinking Machines released Inkling-Small, a 276-billion-parameter model that achieves near-parity with its 975-billion-parameter predecessor while using only 12 billion active parameters per token compared to 41 billion. The model accepts text, image, and audio inputs, supports a one-million-token context window, and comes with an Apache 2.0 license.
Google Ships Gemini Robotics ER 2 With Multi-Robot Teamwork
A new embodied-reasoning model called Gemini Robotics ER 2 plans multi-step robotic tasks lasting several minutes by processing video, images, audio, and text inputs. The model is available to developers through the Gemini API and Google AI Studio, with private-preview access on the Gemini Enterprise Agent Platform.
The Download: tricking LLMs, and reviving geothermal plants
A fundamental vulnerability in how large language models function makes them impossible to fully secure against attacks. Researchers have identified this inherent flaw prevents complete protection of LLMs from hacking attempts.