D-Robotics/SmolVLM2-256M-Video-Instruct-GGUF-BPU
Robotics Model
This is a quantized, GGUF and ONNX format version of the SmolVLM2-256M-Video-Instruct model, designed for image-text-to-text tasks, supporting video and language modalities, and optimized for deployment on D-Robotics hardware.
Catalog metadata
- Repository owner
- D-Robotics
- Source
- Hugging Face Models
- Task
- image text to text
- Cameras
- 0
- License
- apache-2.0
Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.