Rover Obstacle Target
Top-down arena — raycast sensors
Approaching
Reward curve

Object Capture 2D — Sensor-Guided Dwell Grasp

The 3D model trains a swarm of massless points to glide toward goal orbs through a potential field. This 2D companion instead drives one differential-drive rover with five raycast sensors: it has to steer around obstacles it can only sense as ray distances, brake as it nears the target, and hold still inside the capture ring for a set dwell time before the object actually counts as captured. Watch the reward curve tighten as the exploration noise (ε) decays episode over episode — the same training-convergence signature as the 3D dashboard, produced by a completely different rover-and-sensor mechanic.