CurHarsh/sft_robotics_vlm_all_task_821_llava-v1.6-mistral-7b-hf

Robotics Model

This is a Hugging Face model for image-text-to-text tasks, specifically a LLaVA-Next variant fine-tuned for robotics, capable of conversational responses based on visual and textual input.

Catalog metadata

Repository owner
CurHarsh
Source
Hugging Face Models
Task
image text to text
Cameras
0

Browse the catalog to compare robots, tasks, sensors, and licenses. Sign in to open protected download and model links.

Explore more robotics datasets and models