Gemini Robotics-ER 1.6
Gemini Robotics-ER 1.6 is a family of AI models developed by Google DeepMind to bring advanced multimodal intelligence into the physical world by enabling robots to perceive, reason, and act in real-world environments. Built on the Gemini 2.0 foundation, it extends traditional AI capabilities by adding physical action as an output modality, allowing robots to interpret visual input and natural language instructions and convert them directly into motor commands to complete tasks. It includes a vision-language-action model that processes images and instructions to execute tasks, as well as a complementary embodied reasoning model (Gemini Robotics-ER) that specializes in spatial understanding, planning, and decision-making within physical environments. These models enable robots to generalize across new situations, objects, and environments, allowing them to perform complex, multi-step tasks even if they were not explicitly trained for them.
Learn more
FLUX 3
FLUX 3 is a multimodal foundation model that jointly learns from images, video, and audio within one unified architecture, building a representation of how objects hold together, how things move, and how events sound. Built on the Self-Flow approach, it aligns multimodal generation and understanding in the same backbone so each modality constrains the others, sound matches impact, motion follows physical properties, and future events follow from the past. FLUX 3 can mix modalities and jointly generate images, video, and native audio from text prompts or references such as images, video, and audio. Its video capabilities include text-to-video, image-to-video animation, video-to-video transformation, generative video-and-audio continuation, keyframe-controlled transitions, multilingual dialogue, animated typography, diverse styles and aspect ratios, and agentic chaining into longer multi-shot sequences.
Learn more
Gemini Robotics 2
Gemini Robotics 2 is Google DeepMind’s intelligence layer for adaptable robots, bringing whole-body control, advanced dexterity, embodied reasoning, and multi-robot collaboration to physical AI. It includes three models. Gemini Robotics 2 is a vision-language-action model that converts visual and language input into motor control, enabling humanoids and bi-arm robots to act from feet to fingertips. It can coordinate walking, crouching, reaching, balancing, and object manipulation, while controlling five-fingered hands or standard grippers for delicate and precise tasks. Gemini Robotics ER 2 serves as the high-level brain, communicating with people, understanding its surroundings, planning multi-step tasks that last several minutes, coordinating actions with the VLA, tracking progress, self-correcting failures, and allowing different robots to work together.
Learn more
T-Plan Robot
T-Plan Robot automates scripted user actions for Test Automation or Robotic Process Automation (RPA) on Mac, Windows Linux & Mobile.
T-Plan develops and sells two main toolsets. 1) Test Automation and 2) Robotic Process Automation (RPA).
T-Plan Robot is a highly flexible, easy to use, image-based black box GUI automation tool that creates robust automated scripts and exercises applications in the same way as would an end-user.
T-Plan Robot is platform-independent (Java) and runs on, and automates all major systems such as Windows, Mac, Linux and Unix plus mobile platforms. We believe we have a solution for any environment.
GUI automation interacts with your business sponsor and development teams throughout the whole project lifecycle. Working intuitively at the screen level business analysts can help testers drive testable paths through the application, whilst at the same time combining with the development team to define repeatable actions to test code in continuous development.
Learn more