Diffusion Policy
Columbia / Toyota Research Institute / collaborators
Visuomotor policy approach that represents robot behavior as a conditional denoising diffusion process.
Multi-modal action generation
Important policy baseline for comparing whether a dataset needs a larger VLA/WAM or a stronger task policy.
A model hub link is a declaration. Files, loadability, evaluation, and deployment are scored separately.
Model decision scorecard
Use-case scores and evidence confidence are separate; Unknown is not treated as failure.
Code, weights, and checkpoints
UsefulAn official model hub is linked; weight files and loadability are not verified.
Loading and training reproducibility
UnknownNo verified loading configuration is available.
Training data requirements
UsefulRequired signal categories are declared; exact tensor and action interfaces still need verification.
Evaluation evidence
UsefulEvaluation focus is declared, but metrics are not independently verified.
Deployment readiness
UnknownHardware, latency, dependencies, and runtime loading are not yet verified.
Artifact facts and provenance
No metadata-verified artifact facts yet. Source links remain declarations only.
Loop signal demand
Signals this model family needs for training, evaluation, or failure mining.
Observation / ego video
observation · Visual and low-dimensional state observations
Language intent / task phase
task success · Long-horizon task success
Action / robot state
actions · robot state · trajectory · Expert demonstrations with action sequences
Future state / dynamics
Needs future-state supervision or rollout structure to validate predictive dynamics.
Feedback / correction / failure
task success · Failure-aware evaluation splits for multi-modal action distributions · Long-horizon task success
Sim-real / embodiment metadata
robot state
Evaluation focus
- Multi-modal action generation
- Long-horizon task success
- Sensitivity to noisy or suboptimal demonstrations
Missing critical loop signals
Core signal demands are represented. Check quality, alignment, and access constraints.
Related catalog datasets
EgoWorld
Bimanual manipulation in LeRobot format
Dataset license restricts commercial use.
EgoStation GoPro Pick-and-Place
GoPro first-person pick-and-place trajectories
Dataset license restricts commercial use.
Egocentric Adjust Bottle
Apache-2.0 LeRobot bottle-adjustment task
Exact action dimensions, control frequency, normalization, and camera mapping require interface verification.
OpenBot notes
- Important policy baseline for comparing whether a dataset needs a larger VLA/WAM or a stronger task policy.
- OpenBot failure mining should identify when diffusion policies fail from missing feedback or contact signals.
