RoboDojo 是一个统一的仿真与真实世界基准测试平台,用于评估通用机器人操作策略,涵盖 42 个仿真任务和 18 个真实世界任务,支持三种机器人形态。

Stars

311

7 天增长

暂无数据

Fork 数

27

开放 Issue

6

开源协议

MIT

最近更新

2026-07-28

AI 仓库情报摘要
FR-AI / ANALYSIS

为什么值得关注

它提供了一套全面、具有挑战性且可重复的评估框架,包含五个不同的能力维度、异构并行仿真以及公开维护的排行榜。

适合谁使用

  • 机器人学习研究人员
  • 策略开发与工程人员
  • 机器人基准测试社区
  • 评估操作策略的学术实验室

典型使用场景

  • 对策略的泛化能力和记忆能力进行基准测试
  • 测试长程操作能力
  • 比较仿真到真实的迁移性能
  • 向社区排行榜提交结果

项目优势

  • 统一的仿真与真实世界评估,涵盖 60 个任务
  • 五个能力维度(泛化、记忆、精度、长程、开放)深入探测不同技能
  • 异构并行仿真实现快速可扩展的反馈
  • 种子可控的可重复性与一键汇总排行榜

使用前须知

  • 仅提供评估功能,策略训练与集成需通过 XPolicyLab 处理
  • 需要 Isaac Sim 5.1 及 NVIDIA GPU 硬件
  • 非商业许可证限制商业用途
  • 真实世界评估依赖特定机器人平台

README 快速开始

RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies

Webpage | Document | Paper | Community | Leaderboard

https://private-user-images.githubusercontent.com/88101805/619409345-cc074c5d-4567-4418-8a29-1385aaba9d5b.mp4

✨ Highlights

Overview of RoboDojo. RoboDojo unifies efficient simulation evaluation and reproducible real-world testing for generalist robot manipulation, covering 42 simulation tasks, 18 real-world tasks, heterogeneous parallel simulation, RoboDojo-RealEval, XPolicyLab, and a continuously updated leaderboard.

RoboDojo is eval-only in this release: it provides the simulator client, benchmark tasks, asset/config validation, and result artifacts. Policy integration and policy servers are owned by XPolicyLab.

  • 🌐 Unified sim-and-real benchmark — 42 simulation tasks and 18 real-world tasks across 3 robot embodiments for generalist robot manipulation.
  • 🧭 Five capability dimensions — Generalization, Memory, Precision, Long-Horizon, and Open, designed to probe different skills rather than simple object or layout reskins.
  • 🧗 Challenging by design — intentionally hard, diverse, long-horizon tasks that expose failures hidden by simpler benchmarks.
  • Heterogeneous parallel simulation — runs different tasks, scenes, and processes concurrently on Isaac Sim for fast, scalable feedback.
  • 🧱 Physically grounded assets — rigid, articulated, and deformable objects in a single configuration-driven scene.
  • 🤖 Integrate once, evaluate everywhereXPolicyLab unifies 40+ policies behind one interface for both simulation and real-world runs.
  • 📊 Reproducible & leaderboard-ready — seed-controlled layouts and one-command summarize aggregation into a leaderboard table.

📚 Documentation

The RoboDojo documentation is the canonical reference. Key sections:

SectionDescription
Usage OverviewEnd-to-end walkthrough of the evaluation workflow.
Installation & Downloading (Assets and Data)Environment setup and downloading robot/object/layout assets/

相关仓库与替代方案

根据分类、Topic 和编程语言匹配的相似项目。

lopopolo
精选
lopopolo GitHub avatar

harness-engineering

Harness Engineering is a methodology for improving coding agent outputs by carefully crafting the environment around them—providing curated context, tools, and executable constraints that encode an organization’s nonfunctional requirements and cumulative lessons.

AI 与机器学习AI 智能体
2,390
slvDev
精选
slvDev GitHub avatar

esp32-ai

A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.

AI 与机器学习大语言模型
1,960
littledivy
精选
littledivy GitHub avatar

mimic

mimic captures traffic from any iOS or web app and automatically generates a Python client library that lets you call the app's API like a regular library.

AI 与机器学习
1,482