What problem does it solve? Long AI responses burn context window tokens and slow down reading during extended build sessions. This Skill cuts output tokens by roughly 65% by stripping filler, articles, and pleasantries while keeping all technical substance intact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus three classical Chinese (wenyan) variants for maximum character reduction. - Auto-clarity fallback: Automatically reverts to full sentences for security warnings, destructive operations, and ambiguous multi-step sequences, then resumes compression. - Language preservation: Compresses style without switching languages, keeping code, error strings, and technical terms verbatim. - Use Case: During a long debugging session, activate full mode so every response arrives as terse fragments like "Bug in auth middleware. Token expiry check use < not <=.", saving context for actual work. ## Quick Start Say "caveman mode" or "use caveman" to activate compressed responses, and say "stop caveman" or "normal mode" to revert.