system idle
Robot arm Red ball Green cube Yellow cylinder Blue box Orange box
⚠ Couldn't load the 3D engineThree.js failed to load from the CDN. Check your connection and reload.

Vision-Language-Action (VLA) Robot Simulator

A minimal 3D stand-in for a Vision-Language-Action model: a vision encoder locates the objects on the table, a language encoder parses a typed command into a target object and an action, and an action decoder plans the pick-and-place trajectory the arm then carries out — all in one continuous pass, without any hand-coded per-object rules.