Benchmark / simulation

RLBench

RLBench is a manipulation benchmark built on CoppeliaSim and PyRep, with task variations and interfaces for imitation and reinforcement learning.

Source contextPrimary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Reviewed recordLast reviewed 1 sources ↗6 sourced fields

What the protocol tests

The default benchmark robot is Panda. The maintainers allow other arms for experimentation but do not promise that every task is solvable with them; changing the arm changes what a result means.

Source contextPrimary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Reading results

Task-set selection, variation, camera observations and whether privileged low-dimensional state is used affect comparisons. The repository explicitly distinguishes visual learning from low-dimensional evaluation.

Source contextPrimary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Methodology and reuse limits

Benchmark comparisons should retain the default robot and declared observation setting. Simulator and PyRep versions must be pinned.

Source contextPrimary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Evaluation protocol

Exact version / scope
RLBench task benchmark
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Environment
simulation
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Metric definition
Task-defined success conditions and success rate
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Task scope
Not publicly disclosed
Embodiment
Default Franka Panda
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Access & reuse

Access model
Public research documentation and evaluation resources
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

License
Not publicly disclosed

Additional disclosed fields

Category
Robot evaluation benchmark
Primary specReviewed
Sources 1
stepjam/RLBench official repository

stepjam · Publication date not disclosed · Source accessed 2026-09-28

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.

Reported benchmark results

Context before scores. Results from different benchmarks are not directly comparable. A simulation result does not establish real-world reliability, safety or commercial availability.
No result meets our complete revision and methodology requirements for this record yet. Inspect the original benchmark documentation before comparing published scores.

What this evidence does not establish

  • Benchmark comparisons should retain the default robot and declared observation setting. Simulator and PyRep versions must be pinned.

Relationships & deployments

Related reading is an editorial crosslink. Sourced connections describe relationships reported in the cited material. A link to a versioned profile does not establish compatibility with that version unless the connection note explicitly identifies it.

Record history & verified changes

A research review records when we checked a source. It does not mark a product launch or a new deployment.

Initial reviewed record. No subsequent field change has been recorded.

Inspect the evidence

Sources & evidence

md-rlbench
stepjam/RLBench official repository

stepjam · Repository · Publication date not disclosed · Source accessed 2026-09-28

Factual summary and attribution only; no upstream prose, images, weights, or dataset redistributed.

Official README inspected live. Repository code terms do not automatically cover model weights, data, or third-party assets.