What problem does it solve? Long, filler-heavy AI responses waste output tokens and slow down reading. This Skill cuts roughly 65-75% of output tokens by rewriting every response in a compressed caveman register while keeping code, error strings, and technical terms exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, wenyan-lite, wenyan-full, and wenyan-ultra, ranging from light filler removal to extreme Classical Chinese compression. - Session persistence: the mode stays active across all turns until explicitly stopped with "stop caveman" or "normal mode". - Auto-clarity fallback: automatically reverts to normal prose for security warnings, irreversible-action confirmations, and ambiguous multi-step sequences, then resumes compression. - Use Case: During a long debugging session, enable full mode so every explanation arrives as terse fragments like "Bug in auth middleware. Token expiry check use < not <=", preserving all technical detail at a fraction of the tokens. ## Quick Start Ask the assistant to enable caveman mode by saying "use caveman mode" or invoking /caveman, then continue the conversation normally.