CurHarsh/sft_robotics_vlm_all_task_821_Qwen2-VL-7B-Instruct

Robotics Model

This is a fine-tuned Qwen2-VL-7B-Instruct vision-language model for robotics, supporting image-text-to-text tasks such as conversational interaction and visual question answering.

Catalog metadata

Repository owner
CurHarsh
Source
Hugging Face Models
Task
image text to text
Cameras
0

Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.

Explore more robotics datasets and models