D-Robotics/SmolVLM2-256M-Video-Instruct-GGUF-BPU

Robotics Model

This is a quantized, GGUF and ONNX format version of the SmolVLM2-256M-Video-Instruct model, designed for image-text-to-text tasks, supporting video and language modalities, and optimized for deployment on D-Robotics hardware.

Catalog metadata

Repository owner
D-Robotics
Source
Hugging Face Models
Task
image text to text
Cameras
0
License
apache-2.0

Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.

Explore more robotics datasets and models