What problem does it solve?
This Skill addresses the challenge of ensuring reliable, auditable, and source-grounded memory for AI agents by providing a structured framework for testing memory retrieval, policy enforcement, and knowledge grounding.
Core Features & Use Cases
- Evaluation Fixture Design: Create standardized test cases for memory recall, tenant isolation, and hierarchical retrieval.
- Quality Assurance: Validate that agent memory adheres to strict policy, provenance, and scope requirements.
- Use Case: When implementing a new retrieval strategy for code symbols, use this Skill to define positive recall cases and negative leakage tests to ensure the system remains auditable and secure.
Quick Start
Use the engram-eval skill to generate a new evaluation fixture for testing tenant isolation and workspace filtering.