Skip to content

clip

from Orchestra-Research/AI-research-SKILLs

OpenAI's model connecting vision and language. Enables zero-shot image classification, image-text matching, and cross-modal retrieval. Trained on 400M image-text pairs. Use for image search, content m

v1.0.0MIT
254
Lines
779
Words
8
Code Blocks

Languages

bashpython
18-multimodal/clip/SKILL.md