pdd-to-app-journeys-eval

Grades PDD-to-app journey documents with a 7-dimension LLM-as-Judge rubric for CommCare deployability.

1|2|Updated Apr 1, 2026
One-click install
npx skills add https://github.com/dimagi-internal/ace --skill pdd-to-app-journeys-eval
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: pdd-to-app-journeys-eval
Source: https://github.com/dimagi-internal/ace/tree/main/skills/pdd-to-app-journeys-eval
Command: npx skills add https://github.com/dimagi-internal/ace --skill pdd-to-app-journeys-eval

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill eliminates the risk of shipping un deployable CommCare apps by automatically grading PDD-derived journey sets against a rigorous rubric, catching gaps in validation, edge case coverage, and field-ready requirements that manual review often misses.

Core Features & Use Cases

  • 7-Dimension LLM-as-Judge Grading: Evaluates journey sets across persona specificity, archetype alignment, coverage completeness, happy-path voice, edge-case recoverability, pass-criteria measurability, and a hard-gated out-of-chain deployability fitness check.
  • Deployability Hard Gate: Automatically fails journey sets that lack critical field requirements like input validation, GPS accuracy gating, case write-back, and data quality enforcement, even if they are fully faithful to a thin PDD.
  • Use Case: ACE Connect opportunity teams building CommCare apps can use this skill to validate journey docs before app development, ensuring the final product meets field expert standards for real-world data collection.

Quick Start

Use the pdd-to-app-journeys-eval skill to grade the journey set for your current Connect opportunity and receive a deployability verdict with surfaced concerns.

Frequently Asked Questions about pdd-to-app-journeys-eval

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I validate a CommCare journey document for app deployability?

To validate a CommCare journey document for app deployability, grade it against a 7-dimension LLM-as-Judge rubric. This process checks persona specificity, archetype alignment, edge-case recoverability, and field-ready requirements to ensure the app meets field expert standards.

What is LLM-as-Judge grading for product design validation?

LLM-as-Judge grading for product design validation is an automated evaluation method that scores journey sets against a rigorous rubric. It assesses coverage completeness, happy-path voice, and pass-criteria measurability to catch gaps in validation that manual review often misses before app build initiation.

How do I check if my PDD journey sets cover edge cases and input validation?

To check if your PDD journey sets cover edge cases and input validation, apply a deployability hard gate evaluation. This automatically fails journey sets lacking critical field requirements like GPS accuracy gating, case write-back, and data quality enforcement, even if they are fully faithful to a thin PDD.

Can I use automated journey grading for ACE Connect opportunities?

Yes, you can use automated journey grading for ACE Connect opportunities. The evaluation validates product design documents by applying a 7-dimension rubric to journey sets generated from idea-to-PDD artifacts, ensuring the final CommCare product meets real-world data collection standards.

What are the limitations of using an LLM rubric for CommCare deployability checks?

A limitation of using an LLM rubric for CommCare deployability checks is that it applies a hard gate for out-of-chain requirements. Journey sets will automatically fail if they lack critical field requirements like input validation or GPS accuracy gating, regardless of their overall fidelity to the PDD.

Does the pdd-to-app-journeys-eval skill require any external dependencies?

No, the pdd-to-app-journeys-eval skill does not require any external dependencies. It operates independently to grade product design documents and journey sets against its built-in 7-dimension LLM-as-Judge rubric to deliver a deployability verdict.