quality-evaluator

Evaluate translation fluency, naturalness, tone, and cultural fit.

2|Updated Jan 2, 2026
One-click install
npx skills add https://github.com/gonsoomoon-ml/Self-Correcting-Explainable-Translation-Agent --skill quality-evaluator
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: quality-evaluator
Source: https://github.com/gonsoomoon-ml/Self-Correcting-Explainable-Translation-Agent/tree/main/01_explainable_translate_agent/skills/quality-evaluator
Command: npx skills add https://github.com/gonsoomoon-ml/Self-Correcting-Explainable-Translation-Agent --skill quality-evaluator

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Evaluates translations for fluency, naturalness, tone/formality, and cultural fit to ensure translations feel authentic in the target language.

Core Features & Use Cases

  • Fluency & Naturalness Review: assesses whether a translation reads like native content.
  • Tone & Cultural Fit: verifies formality level and cultural appropriateness for the target audience.
  • Pairwise Candidate Comparison: compares multiple translations to select the best option and provide actionable feedback.

Quick Start

Provide the candidate translations and context, and the evaluator will return a native-speaker quality assessment with actionable feedback.

Frequently Asked Questions about quality-evaluator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate translation quality across multiple candidate versions?

Evaluate translation quality by comparing fluency, naturalness, tone, and cultural fit across candidate versions. Pairwise candidate comparison identifies which translation reads most naturally, providing structured scoring and native-speaker-style feedback.

What is the best way to assess if a translation sounds like native content?

Assess native-like fluency by checking whether the translation reads naturally for the target language. Review formality levels and cultural appropriateness to ensure the text feels authentic to native speakers in the target domain.

How do I compare two translations to select the best option?

Compare two translations using pairwise candidate comparison to evaluate fluency and cultural fit. The process provides structured scoring and actionable feedback to determine which version reads most naturally for the target audience.

Does this translation evaluator check cultural fit and tone formality?

Yes, the translation evaluator verifies tone and cultural fit by assessing formality levels and cultural appropriateness for the target audience. This ensures the translated content aligns with native-speaker expectations.

When do I need native-speaker-style translation feedback?

You need native-speaker-style translation feedback during translation QA scenarios when multiple candidate translations exist across languages. It provides structured scoring and actionable improvement recommendations to ensure authentic target-language output.

Can I use pairwise comparison for translation QA across different languages?

Yes, you can use pairwise comparison for translation QA across languages and domains. Provide candidate translations with context to receive structured scoring and native-speaker-style assessment with actionable feedback.