Skip to content

llava

from Orchestra-Research/AI-research-SKILLs

Large Language and Vision Assistant. Enables visual instruction tuning and image-based conversations. Combines CLIP vision encoder with Vicuna/LLaMA language models. Supports multi-turn image chat, vi

v1.0.0MIT
305
Lines
937
Words
15
Code Blocks

Languages

bashpython
18-multimodal/llava/SKILL.md