Research Publication · Sep 14, 2026

Reward AI releases OM-1 cross-embodiment robot policy

Reward AI released OM-1, a general-purpose manipulation policy trained directly from natural human demonstrations and designed to run across industrial robot arms, mobile manipulators and humanoids.

Reward AI company media
Company media · Reward AI
  • OM-1 is trained without teleoperation or on-robot experience, using images, tactile signals, inter-finger proximity and hand-pose trajectories captured through the wearable Omnibody Hand.
  • The policy outputs motion direction, speed, force and grasp timing while a separate reinforcement-learned control layer compensates for robot dynamics, disturbances and inference delay.
  • Reward AI says a new task, including dynamic and long-horizon manipulation, can be learned from less than 30 minutes of demonstration data.

Reward AI develops general-purpose robot intelligence and a human-demonstration data interface for cross-embodiment manipulation.