What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts token usage by roughly 75% by stripping articles, filler words, and pleasantries while keeping all technical substance intact. ## Core Features & Use Cases - Six Intensity Levels: Choose from lite, full (default), ultra, wenyan-lite, wenyan-full, and wenyan-ultra to control compression aggressiveness, including classical Chinese (文言文) modes. - Persistent Mode: Stays active across every response in a session until explicitly disabled with "stop caveman" or "normal mode". - Auto-Clarity Safety: Automatically drops the terse style for security warnings, irreversible action confirmations, and multi-step sequences where fragments could be misread. - Use Case: During a long debugging session, enable ultra mode so every explanation like "Inline obj prop → new ref → re-render. useMemo." costs minimal tokens while code blocks and error messages remain exact. ## Quick Start Tell the AI "use caveman mode" or invoke /caveman to start receiving ultra-compressed responses immediately.