blip-2-vision-language
Bridges frozen image encoders with LLMs for vision-language tasks like image captioning, VQA, and image-text retrieval.
npx skills add https://github.com/zhuangbiaowei/smart_bot --skill blip-2-vision-language-zhuangbiaowei
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: blip-2-vision-language Source: https://github.com/zhuangbiaowei/smart_bot/tree/main/skills/blip-2 Command: npx skills add https://github.com/zhuangbiaowei/smart_bot --skill blip-2-vision-language-zhuangbiaowei