What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts output tokens by roughly 65% by stripping articles, filler, pleasantries, and hedging while keeping every technical detail, code block, and error string exact. ## Core Features & Use Cases - Six Intensity Levels: Choose from lite, full (default), ultra, wenyan-lite, wenyan-full, and wenyan-ultra to control compression from light trimming to extreme Classical Chinese terseness. - Auto-Clarity Fallback: Automatically reverts to normal prose for security warnings, irreversible-action confirmations, and ambiguous multi-step sequences, then resumes compression afterward. - Language Preservation: Compresses style without switching languages, keeping technical terms, code, API names, and error strings verbatim. - Use Case: During a long debugging session, activate full mode so every explanation arrives as short fragments like "Bug in auth middleware. Token expiry check use < not <=", cutting reading time without losing precision. ## Quick Start Ask the AI to talk like caveman or invoke /caveman to enable compressed responses for the rest of the session.