llava
Set up multimodal chat pipelines combining CLIP vision encoders with Vicuna/LLaMA language models for image-based VQA tasks.
npx skills add https://github.com/tadod12/fraud-detection-research --skill llava-tadod12
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llava Source: https://github.com/tadod12/fraud-detection-research/tree/main/.agent/skills/18-multimodal/llava Command: npx skills add https://github.com/tadod12/fraud-detection-research --skill llava-tadod12