llava
Generates grounded conversational responses from user-provided images using CLIP and LLaMA/Vicuna-style language pipelines.
npx skills add https://github.com/qcmuu/AI-Research-Skills --skill llava-qcmuu
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llava Source: https://github.com/qcmuu/AI-Research-Skills/tree/main/18-multimodal/llava Command: npx skills add https://github.com/qcmuu/AI-Research-Skills --skill llava-qcmuu