Reward AI releases OM-1 cross-embodiment robot policy
Reward AI released OM-1, a general-purpose manipulation policy trained directly from natural human demonstrations and designed to run across industrial robot arms, mobile manipulators and humanoids.

Evidence notes
- OM-1 is trained without teleoperation or on-robot experience, using images, tactile signals, inter-finger proximity and hand-pose trajectories captured through the wearable Omnibody Hand.
- The policy outputs motion direction, speed, force and grasp timing while a separate reinforcement-learned control layer compensates for robot dynamics, disturbances and inference delay.
- Reward AI says a new task, including dynamic and long-horizon manipulation, can be learned from less than 30 minutes of demonstration data.
Company overview
Reward AI develops general-purpose robot intelligence and a human-demonstration data interface for cross-embodiment manipulation.