What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts token usage by roughly 75% by stripping articles, filler words, and pleasantries while keeping all technical substance intact. ## Core Features & Use Cases - Six Intensity Levels: Choose from lite, full (default), ultra, wenyan-lite, wenyan-full, and wenyan-ultra to control compression aggressiveness, including classical Chinese (文言文) modes. - Persistent Mode: Stays active across every response in the session until explicitly disabled with "stop caveman" or "normal mode". - Auto-Clarity Fallback: Automatically reverts to clear, full language for security warnings, irreversible action confirmations, and multi-step instructions where fragments could cause misreading. - Use Case: During a long debugging session, activate caveman mode to get dense, scannable answers like "Bug in auth middleware. Token expiry check use < not <=." instead of paragraph-length explanations. ## Quick Start Say "caveman mode" or "use caveman" to activate terse responses, then switch intensity anytime with /caveman lite, full, or ultra.