Tag: #vla
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 8 posts
Gemini Robotics 2 and the Robot Foundation Model — What Whole-Body Control Actually Changes
On July 30, 2026, Google DeepMind unveiled Gemini Robotics 2, announcing that a single vision-language-action model now controls a humanoid from its toes to its fingertips. Interestingly, the numbers released alongside i
2026-07-31 · 14 min read #ai#robotics#vla#foundation-models#evaluationThe 2026 Robotics Company Map — The Humanoid Showdown, VLA Models, and the Engineer's Way In
2026 is the inflection point where humanoid shipments jump about 7x year-over-year to a forecast of 50,000+ units. From Figure, valued around USD 39B with its in-house VLA model Helix, to Tesla Optimus Gen 3 entering mas
2026-07-09 · 7 min read #robotics#ai#vla#trends#careerRobots That Learn from Human Video — The Dream of Web-Scale Data
Can a robot learn from videos of people handling objects. We cover affordances and trajectories, the domain gap, representation learning and pre-training, one-shot imitation, web-video scale-up, and combining with robot
2026-06-29 · 19 min read #ai-papers#robotics#imitation-learning#representation-learning#videoRobot Foundation Models — One Policy for Many Jobs
We lay out the trend of robot foundation models that aim to do many jobs across many robots with a single policy. We cover the concept of a generalist policy, large-scale robot data such as Open X-Embodiment, cross-embod
2026-06-29 · 19 min read #ai-papers#robotics#foundation-model#generalist-policy#open-x-embodimentHow Robots Learn — Imitation Learning and Reinforcement Learning
An overview of the four ways robots acquire skills, followed by a deep look at imitation learning (teleoperation, behavioral cloning, DAgger) and reinforcement learning (rewards, policies, exploration): their principles,
2026-06-29 · 16 min read #ai-papers#robotics#imitation-learning#reinforcement-learning#robot-learningTwo Brains for a Humanoid — GR00T N1 and Helix
A VLA for humanoid robots must combine fast reflexes with slow deliberation. Centered on NVIDIA GR00T N1 and Figure AI Helix, we organize the dual-system architecture that combines fast low-level control (System 1) with
2026-06-27 · 15 min read #ai-papers#robotics#humanoid#groot-n1#helixRobots That See, Hear, and Move — A Review of VLA Models RT-2 and OpenVLA
Vision-Language-Action (VLA) models take camera images and natural-language instructions and output robot actions directly. Centered on RT-2, Open X-Embodiment, and OpenVLA, this post organizes the VLA paradigm: its idea
2026-06-27 · 15 min read #ai-papers#robotics#vla#rt-2#openvlaThe Complete Autonomous Driving & Robotics Tech Stack: From C++, ROS2, CUDA, TensorRT to VLM/VLA, Simulation, and Beyond
A comprehensive guide to the core technology stack behind autonomous driving and robotics. Covering Modern C++, ROS/ROS2, CUDA parallel programming, TensorRT optimization, model compression (quantization/pruning), sensor
2026-03-01 · 22 min read #autonomous-driving#robotics#ros2#cuda#tensorrt