The Eval Index / Benchmarks / #61

robocurve/inspect-robots

by robocurve · Benchmarks · updated today

Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.

69
momentum
564
stars
60
forks
#61
rank
benchmarkembodied-aievaluationphysical-airerunroboticsvision-language-actionvla
View on GitHub →