vlm-visual-refiner

Refine and validate VLM-generated chart bounding boxes using local figure titles and safety margins.

Updated Mar 6, 2026
One-click install
npx skills add https://github.com/ghjghjghkimo/skills --skill vlm-visual-refiner
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: vlm-visual-refiner
Source: https://github.com/ghjghjghkimo/skills/tree/main/.gemini/skills/vlm-visual-refiner
Command: npx skills add https://github.com/ghjghjghkimo/skills --skill vlm-visual-refiner

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

專門用於視覺化與校準 VLM 生成的圖表邊界框 (Bounding Boxes)。當需要將相對座標轉換為標註圖片以供人工驗證或報告生成時使用。

Core Features & Use Cases

  • 提供對局部標題文字的辨識與邊界框對齊,降低幻覺標註與裁切風險。
  • 支援將邊界框與圖表內容關聯,以產出可追溯的校驗報告與高質量輸出。

Quick Start

Refine VLM chart bounding boxes by feeding an image and output calibrated boxes for validation.

Frequently Asked Questions about vlm-visual-refiner

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I calibrate VLM-generated chart bounding boxes to prevent mislabeling?

To calibrate VLM-generated chart bounding boxes, this Skill validates relative coordinates and enforces safety margins using local figure titles as anchor labels. It outputs a structured 2D layout JSON with validated coordinates to prevent improper cropping and hallucinated annotations in scientific charts.

Why does my VLM chart annotation include improperly cropped bounding boxes?

Improperly cropped VLM chart bounding boxes occur when relative coordinates lack safety margins. This Skill addresses that by aligning boundaries with local figure titles and validating the coordinates, outputting a structured 2D layout JSON that maintains content integrity for scientific reports.

What is the best way to validate relative coordinates for chart analysis bounding boxes?

The best way to validate relative coordinates for chart analysis is using local figure titles as anchor labels to enforce safety margins. This Skill refines VLM-generated boundaries and outputs a structured 2D layout JSON with validated coordinates for high-quality annotation and report generation.

Can I use anchor labels from scientific charts to refine VLM bounding boxes?

Yes, you can use local figure titles as anchor labels to refine VLM bounding boxes. This Skill applies these anchors to enforce safety margins, preventing mislabeling and improper cropping in scientific charts while producing a structured 2D layout JSON with validated coordinates.

Do I need a structured 2D layout JSON for visual quality assurance of chart annotations?

A structured 2D layout JSON is needed for visual quality assurance as it provides validated coordinates to prevent mislabeling. This Skill outputs exactly that by refining VLM-generated chart bounding boxes and enforcing safety margins using local figure titles as anchor labels.

Does vlm-visual-refiner support converting relative coordinates to annotated images for validation?

Yes, vlm-visual-refiner supports converting relative coordinates to annotated images for validation. It refines VLM-generated chart bounding boxes, enforces safety margins, and outputs a structured 2D layout JSON with validated coordinates for visual quality assurance and report generation.