--- base_model: lerobot/pi05_base datasets: K-vr/cube_stack3 library_name: lerobot license: apache-2.0 model_name: pi05 pipeline_tag: robotics tags: - lerobot - pi05 - robotics --- # Model Card for pi05 [π₀.₅ (Pi05)](https://www.physicalintelligence.company/blog/pi05) is a Vision-Language-Action model from Physical Intelligence designed for open-world generalization: it evolves π₀ to generalize to entirely new environments and situations that were never seen during training. The LeRobot implementation is adapted from their open-source OpenPI repository. This policy has been trained and pushed to the Hub using [LeRobot](https://github.com/huggingface/lerobot). Learn how to train and run it in the [LeRobot pi05 guide](https://huggingface.co/docs/lerobot/main/en/pi05), or browse the [full documentation](https://huggingface.co/docs/lerobot/index). --- ## Model Details - **License:** apache-2.0 - **Fine-tuned from:** [lerobot/pi05_base](https://huggingface.co/lerobot/pi05_base) - **Robot type:** `bi_so_follower_7dof` - **Cameras:** `left_cam_left`, `left_cam_scene`, `right_cam_right`, `right_cam_scene_realsense` ## Inputs & Outputs The policy consumes these observation features and produces these action features. **Inputs** | Feature | Type | Shape | | --- | --- | --- | | `observation.images.base_0_rgb` | VISUAL | `(3, 224, 224)` | | `observation.images.left_wrist_0_rgb` | VISUAL | `(3, 224, 224)` | | `observation.images.right_wrist_0_rgb` | VISUAL | `(3, 224, 224)` | | `observation.state` | STATE | `(32,)` | **Outputs** | Feature | Type | Shape | | --- | --- | --- | | `action` | ACTION | `(14,)` | ## Training Data Verification | Check | Value | | --- | --- | | Dataset repository | [K-vr/cube_stack3](https://huggingface.co/datasets/K-vr/cube_stack3) | | Resolved training root | `/workspace/datasets/lerobot/cube_stack3` | | Repo/root folder-name check | ✅ Match (`cube_stack3`) | | Dataset size | 120 episodes, 66920 frames at 30 FPS | | Episodes selected | all | | Task(s) | "Stack the three cubes with the 40 mm cube on the bottom, the 30 mm cube in the middle, and the 20 mm cube on top."
"Stack the three cubes with the largest cube on the bottom, the mid sized cube in the middle, and the smallest cube on top." | ## Training Configuration | Setting | Value | | --- | --- | | Training steps | 34000 | | Batch size | 16 | | Action representation | absolute | | Sample weighting | none | | Image augmentation | disabled | | Optimizer | adamw | | Learning rate | 2.5e-05 | | Seed | 1000 | | LeRobot version | 0.6.1 | --- ## How to Get Started with the Model New to LeRobot? These guides cover the full workflow: - **[Install LeRobot](https://huggingface.co/docs/lerobot/main/en/installation)** — set up the `lerobot` package. - **[Hardware setup](https://huggingface.co/docs/lerobot/main/en/hardware_guide)** — assemble, wire, and calibrate your robot and cameras. - **[Record data & train a policy](https://huggingface.co/docs/lerobot/en/il_robots)** — the end-to-end imitation-learning walkthrough. - **[CLI cheat-sheet](https://huggingface.co/docs/lerobot/main/en/cheat-sheet)** — quick reference for the `lerobot-*` commands. The short version to run and train this policy: ### Run the policy on your robot ```bash lerobot-rollout \ --strategy.type=base \ --robot.type=bi_so_follower_7dof \ --robot.port= \ --robot.cameras="{ : {type: opencv, index_or_path: , width: 640, height: 480, fps: 30}, : {type: opencv, index_or_path: , width: 640, height: 480, fps: 30}}" \ --policy.path=K-vr/cube_stack3_Pi05_abs \ --task="Stack the three cubes with the 40 mm cube on the bottom, the 30 mm cube in the middle, and the 20 mm cube on top." \ --duration=60 ``` Replace the remaining `<...>` placeholders with your own values: `--robot.port` and the camera names/indices are specific to your machine, and the camera names must match the observation keys this policy was trained on. When `--strategy.type=base` is used the script doesn't record the episodes. Skipping duration will make the policy run indefinitely. For more information look at [rollout documentation](https://huggingface.co/docs/lerobot/main/en/inference). ### Train your own policy This policy type is usually fine-tuned from the pretrained base model [lerobot/pi05_base](https://huggingface.co/lerobot/pi05_base): ```bash lerobot-train \ --dataset.repo_id=${HF_USER}/ \ --policy.path=lerobot/pi05_base \ --output_dir=outputs/train/ \ --job_name=lerobot_training \ --policy.device=cuda \ --policy.repo_id=${HF_USER}/ \ --wandb.enable=true ``` _Writes checkpoints to `outputs/train//checkpoints/`._ --- ## Evaluation _No evaluation results have been provided for this policy yet._ --- ## Citation If you use this policy, please cite the method linked in the description above, along with LeRobot: ```bibtex @misc{cadene2024lerobot, author = {Cadene, Remi and Alibert, Simon and Soare, Alexander and Gallouedec, Quentin and Zouitine, Adil and Palma, Steven and Kooijmans, Pepijn and Aractingi, Michel and Shukor, Mustafa and Aubakirova, Dana and Russi, Martino and Capuano, Francesco and Pascal, Caroline and Choghari, Jade and Moss, Jess and Wolf, Thomas}, title = {LeRobot: State-of-the-art Machine Learning for Real-World Robotics in Pytorch}, howpublished = "\url{https://github.com/huggingface/lerobot}", year = {2024} } ```