minimax-multimodal

Verify multimodal outputs and bind evidence to task/run IDs.

6|Updated Aug 26, 2025
One-click install
npx skills add https://github.com/POWERFULMOVES/PMOVES.AI --skill minimax-multimodal
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: minimax-multimodal
Source: https://github.com/POWERFULMOVES/PMOVES.AI/tree/main/.kilocode/skills/minimax-multimodal
Command: npx skills add https://github.com/POWERFULMOVES/PMOVES.AI --skill minimax-multimodal

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill provides a structured approach to validate PMOVES multimodal outputs (text, audio, visuals) and capture evidence tied to task/run IDs, enabling auditable results across deployments.

Core Features & Use Cases

  • Multimodal verification: perform text, audio, and VLM checks to validate outputs.
  • Evidence binding: attach evidence artifacts to task/run identifiers for traceability.
  • Operational integration: works with Jellyfin video frames, transcription pipelines, and Prometheus metrics to support end-to-end quality checks.

Quick Start

Run a multimodal verification workflow on a given task id to capture evidence and bind results.

Frequently Asked Questions about minimax-multimodal

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I verify multimodal outputs with verifiable evidence during task runs?

To verify multimodal outputs with verifiable evidence, run a verification workflow on a specific task ID to apply deterministic checks across text, audio, and visual outputs. This captures and binds evidence artifacts directly to task and run IDs for auditable traceability.

Can I use Jellyfin video frames and transcription pipelines for end-to-end audio and visual validation?

Yes, you can use Jellyfin video frames and transcription pipelines for end-to-end audio and visual validation. The skill integrates with these pipelines alongside VLM checks to support comprehensive quality validation across multimodal deployments.

What is evidence binding for task and run IDs in multimodal verification?

Evidence binding for task and run IDs is the process of attaching validation artifacts to specific identifiers during multimodal verification. It ensures text, audio, and visual checks are traceable and auditable across experiments and deployments.

Does multimodal verification integrate with Prometheus metrics for deployment quality checks?

Yes, multimodal verification integrates with Prometheus metrics to support end-to-end deployment quality checks. This operational integration allows you to monitor and validate text, audio, and visual outputs within existing monitoring infrastructure.

Are there limitations when applying deterministic checks to VLM and text-analysis outputs?

The metadata does not detail specific limitations when applying deterministic checks to VLM and text-analysis outputs. It focuses on the capability to perform these checks and bind evidence to task IDs for structured, auditable validation across deployments.