Agent skill · zechenzhangagi

llava

Large Language and Vision Assistant. Enables visual instruction tuning and image-based conversations. Combines CLIP vision encoder with Vicuna/LLaMA langua

Reference it in any coding agent with:

@skills zechenzhangagi/llava

Browse the @skills marketplace