agent-visual-feedback

Analyze Rerun canvases, video frames, and images to generate diagnostic feedback.

17|8|Updated Apr 7, 2026
One-click install
npx skills add https://github.com/nebius/nebius-physical-ai --skill agent-visual-feedback
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: agent-visual-feedback
Source: https://github.com/nebius/nebius-physical-ai/tree/main/skills/atomic/agent-visual-feedback
Command: npx skills add https://github.com/nebius/nebius-physical-ai --skill agent-visual-feedback

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill solves the challenge of interpreting complex visual data from simulation and robotics environments, allowing the agent to provide immediate, context-aware feedback on Rerun frames, videos, and images.

Core Features & Use Cases

  • Visual Critique: Provides actionable feedback on simulation rollouts, skeleton movements, and environment meshes.
  • Multimodal Analysis: Interprets diverse inputs including Rerun canvases, video frames, and data-pane JSON.
  • Use Case: A robotics researcher can trigger a visual critique of a failed sim-to-real rollout to identify specific task progress cues or defects in the policy execution.

Quick Start

Use the agent-visual-feedback skill to describe the current viewer content and provide actionable critique.

Frequently Asked Questions about agent-visual-feedback

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I analyze visual data from a robotics simulation rollout to find execution defects?

To critique failed robotics rollouts visually, provide Rerun canvases, video frames, and data-pane JSON to generate structured diagnostic reports identifying task progress cues and policy execution defects.

How does multimodal visual critique work for interpreting simulation environments?

Multimodal visual critique operates by processing Rerun canvases and video frames through vision-language models to return structured diagnostic reports assessing environment meshes and skeleton movements.

Does Rerun canvas visualization support automated visual feedback for robotics research?

Yes, Rerun canvas visualization supports automated visual feedback by operating on image artifacts and video frames to deliver immediate, context-aware operator feedback for robotics simulation environments.

What do I need to generate actionable operator feedback from image artifacts in a sim-to-real environment?

Generating actionable operator feedback from image artifacts requires integrating vision-language models and the NPA agent chat queue to process visual context and return structured diagnostic reports.

What are the limitations of using automated visual analysis for robotics simulation data?

Automated visual analysis for robotics simulation data requires the NPA agent chat queue and vision-language models to process visual context, limiting standalone usage without these specific integrations.