Data & Benchmarks
BEHAVIOR Benchmark
BEHAVIOR is a benchmark family from Stanford for embodied AI in household settings, whose current form, BEHAVIOR-1K, defines 1,000 everyday activities such as cleaning, cooking, and tidying, grounded in surveys of what people most want automated. Tasks are specified in a symbolic predicate language and simulated in OmniGibson atop NVIDIA Omniverse, with realistic homes, articulated objects, and fluid and thermal effects, and long-horizon mobile manipulation remains largely unsolved on it.
Why it matters for physical AI
Household benchmarks with rigorous success predicates expose how far current policies are from useful home autonomy, providing a common yardstick for long-horizon mobile manipulation progress.
Build physical AI
Put these concepts to work on real hardware
Axol is a dual-arm robot built for physical AI — teleoperate it, collect demonstrations, and deploy learned policies out of the box.