1. X
  2. Yevgen Chebotar
Log inSign up
Yevgen Chebotar
47 posts
user avatar
Yevgen Chebotar
@YevgenChebotar
Robotic foundation models @NVIDIA 🤖 Previously @GoogleDeepMind (VLAs, RT-2, Offline RL) and @Figure_robot (Helix models)
Joined March 2017
359
Following
2,085
Followers
RepliesRepliesMediaMedia
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Don't miss what's happening
People on X are the first to know.
Log inSign up

New to X?

Sign up now to get your own personalized timeline!

Create account

By signing up, you agree to the Terms of Service and Privacy Policy, including Cookie Use.

  • user avatar
    Yevgen Chebotar
    @YevgenChebotar
    Jul 15
    Long visuomotor context is one of THE most important missing parts for robotic foundation models, there is no way around it if we want generic policies that remember more than a couple of camera frames. However, adding visual context naively significantly increases inference time
    user avatar
    Yunfan Jiang
    @YunfanJiang
    Jul 15
    We scaled robot policies to 8K timesteps of visuomotor context, orders of magnitude beyond current SoTAs, at constant inference latency. Introducing RoboTTT 🤖 With minutes of experience in context, our robots: 🎥 one-shot imitate human video demos 📈 improve themselves during
    00:00
    180K
  • user avatar
    Yevgen Chebotar
    @YevgenChebotar
    Feb 4
    We are moving towards world-model based backbones for robotic policies, pre-trained on web-scale videos and outputting robotic actions directly within the same diffusion model! There are new levels of transfer to unseen tasks and motions being unlocked, something that has been
    user avatar
    Joel Jang
    @jang_yoel
    Feb 4
    Introducing DreamZero 🤖🌎 from @nvidia > A 14B “World Action Model” that achieves zero-shot generalization to unseen tasks & few-shot adaptation to new robots > The key? Jointly predicting video & actions in the same diffusion forward pass Project Page: dreamzero0.github.io
    00:00
    1.8K
  • user avatar
    Yevgen Chebotar
    @YevgenChebotar
    Dec 10, 2025
    Excited to join the NVIDIA GEAR team to help build the next generation of open robotic foundation models!
    14K
  • user avatar
    Yevgen Chebotar
    @YevgenChebotar
    Feb 20, 2025
    We've made great progress on Vision-Language-Action Models for humanoids in our new Helix model! Check out the technical report for more details: figure.ai/news/helix
    user avatar
    Figure
    @Figure_robot
    Feb 20, 2025
    Meet Helix, our in-house AI that reasons like a human Robotics won't get to the home without a step change in capabilities Our robots can now handle virtually any household item:
    00:00
    8.1K
  • user avatar
    Yevgen Chebotar
    @YevgenChebotar
    Jun 17, 2024
    The path to VLAs lies through VLMs. A very nice intro for everyone interested in working with Vision-Language Models:
    arXiv logo
    arxiv.org
    An Introduction to Vision-Language Modeling
    Following the recent popularity of Large Language Models (LLMs), several attempts have been made to extend them to the visual domain. From having a visual assistant that could guide us through...
    6.3K