CurHarsh/sft_robotics_vlm_all_task_821_Qwen2-VL-7B-Instruct
Robotics Model
This is a fine-tuned Qwen2-VL-7B-Instruct vision-language model for robotics, supporting image-text-to-text tasks such as conversational interaction and visual question answering.
Catalog metadata
- Repository owner
- CurHarsh
- Source
- Hugging Face Models
- Task
- image text to text
- Cameras
- 0
Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.