What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts token usage by roughly 75% by stripping filler, articles, and pleasantries while keeping all technical substance, code, and error messages exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus wenyan-lite, wenyan-full, and wenyan-ultra for classical Chinese compression. - Persistent mode with safety overrides: Stays active across turns until told to stop, but automatically drops compression for security warnings, destructive operations, and ambiguous multi-step instructions. - Language preservation: Compresses style, not language — replies in Portuguese, Spanish, or Chinese stay in that language, and code, API names, and error strings remain verbatim. - Use Case: During a long debugging session, invoke /caveman to get terse answers like "Bug in auth middleware. Token expiry check use < not <=." instead of paragraph-length explanations. ## Quick Start Ask the assistant to use caveman mode to answer your technical questions briefly while keeping all code and error messages exact.