LIBERO-Plus
Table 2 · OOD success rate (%)| Method | Params. | PT | Camera | Robot | Language | Light | Background | Noise | Layout | Avg. |
|---|---|---|---|---|---|---|---|---|---|---|
| Without robot-policy pretraining | ||||||||||
| VLA-Adapter | 0.5 | ✗ | 36.2 | 37.9 | 74.6 | 70.6 | 76.1 | 58.0 | 69.7 | 60.4 |
| RoVLA | 2 | ✗ | 58.4 | 36.3 | 92.9 | 95.6 | 95.0 | 80.9 | 73.0 | 76.0 |
| ResVLA | 2 | ✗ | 49.8 | 59.9 | 88.5 | 90.5 | 94.9 | 76.8 | 79.0 | 77.1 |
| JEPA-WAM (Ours) | 0.5 | ✗ | 79.2 | 59.2 | 68.2 | 93.3 | 94.6 | 83.6 | 76.1 | 79.2 |
| With robot-policy pretraining | ||||||||||
| VLA-JEPA | 2 | ✓ | 63.3 | 67.1 | 85.4 | 95.6 | 93.6 | 66.3 | 85.1 | 79.5 |
| PokeVLA | 0.5 | ✓ | 84.7 | 46.1 | 84.8 | 94.6 | 82.6 | 89.8 | 77.2 | 80.0 |
| ABot-M0 | 4 | ✓ | 60.4 | 67.9 | 86.4 | 96.2 | 91.6 | 86.4 | 82.6 | 81.6 |
| Cosmos-Policy | 2 | ✓ | 75.8 | 63.3 | 81.7 | 96.5 | 88.9 | 92.7 | 82.2 | 83.0 |
| π0.5 | 3 | ✓ | 69.4 | 75.3 | 82.6 | 96.7 | 96.8 | 84.3 | 86.2 | 84.5 |
| Being-H0.7 | 3 | ✓ | 82.0 | 59.0 | 82.8 | 97.8 | 90.0 | 93.5 | 88.5 | 84.8 |
| π0.5 + JEPA Obj. (Ours) | 3 | ✓ | 66.0 | 82.0 | 86.5 | 96.8 | 96.0 | 88.3 | 88.3 | 86.3 |
Success rates after training on LIBERO demonstrations, evaluated on seven out-of-distribution shift categories without OOD fine-tuning.