llava
from Orchestra-Research/AI-research-SKILLs
Large Language and Vision Assistant. Enables visual instruction tuning and image-based conversations. Combines CLIP vision encoder with Vicuna/LLaMA language models. Supports multi-turn image chat, vi
v1.0.0MIT
305
Lines
937
Words
15
Code Blocks
Languages
bashpython
18-multimodal/llava/SKILL.md