L7-Robotics/smolvla_so101_3cam_20fps_cube_tasks_merge_v1

Robotics Model

This is a fine-tuned SmolVLA vision-language-action model for robotic manipulation, trained on a merged dataset of cube tasks captured with three cameras at 20 fps, based on the lerobot/smolvla_base model.

Catalog metadata

Repository owner
L7-Robotics
Source
Hugging Face Models
Task
robotics
Cameras
0
License
apache-2.0

Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.

Explore more robotics datasets and models