L7-Robotics/smolvla_so101_3cam_20fps_cube_tasks_merge_v1
Robotics Model
This is a fine-tuned SmolVLA vision-language-action model for robotic manipulation, trained on a merged dataset of cube tasks captured with three cameras at 20 fps, based on the lerobot/smolvla_base model.
Catalog metadata
- Repository owner
- L7-Robotics
- Source
- Hugging Face Models
- Task
- robotics
- Cameras
- 0
- License
- apache-2.0
Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.