caveman

Compresses AI responses into terse caveman-style phrasing while preserving full technical accuracy.

Updated May 21, 2026
One-click install
npx skills add https://github.com/CagesThrottleUs/private-ai-harness --skill caveman-cagesthrottleus
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/CagesThrottleUs/private-ai-harness/tree/main/skills/caveman
Command: npx skills add https://github.com/CagesThrottleUs/private-ai-harness --skill caveman-cagesthrottleus

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts output tokens by roughly 65% (measured) by stripping articles, filler, pleasantries, and hedging while keeping every technical detail, code snippet, and error string exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus wenyan-lite, wenyan-full, and wenyan-ultra for classical Chinese compression, switchable mid-session with /caveman lite|full|ultra. - Language preservation: compresses the user's dominant language (Portuguese, Spanish, Chinese, etc.) instead of forcing English, and never alters code, API names, CLI commands, or error strings. - Auto-clarity override: automatically drops compression for security warnings, irreversible-action confirmations, and multi-step sequences where fragments could cause misreads. - Use Case: During a long debugging session, invoke /caveman so every diagnosis arrives as terse fragments like "Bug in auth middleware. Token expiry check use < not <=. Fix:" — then say "normal mode" to revert. ## Quick Start Ask the AI to enable caveman mode for the rest of the session so all answers come back ultra-terse with full technical accuracy.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce AI response token usage in chat?

Enable caveman mode by saying "caveman mode" or invoking /caveman. The assistant drops articles, filler words, and pleasantries, cutting output tokens by roughly 65% while keeping code, error strings, and technical terms exact.

How do I change or turn off caveman compression level?

Say /caveman lite, full, or ultra to switch intensity mid-session, or /caveman wenyan-lite/full/ultra for classical Chinese tiers. Say "stop caveman" or "normal mode" to revert to standard responses for the rest of the session.

Does caveman mode work in languages other than English?

Yes. The style compresses the user's dominant language, so Portuguese or Spanish replies stay in that language. Dedicated wenyan-lite, wenyan-full, and wenyan-ultra levels provide classical Chinese compression reaching 80-90% character reduction.

When does caveman mode stop compressing responses?

Compression automatically pauses for security warnings, irreversible-action confirmations, and multi-step sequences where fragments could cause misreading. Code blocks, commit messages, and PR bodies are always written in normal uncompressed prose.

Does terse compression lose technical accuracy?

No. Technical terms, code, API names, CLI commands, and exact error strings are always preserved verbatim. Only stylistic elements like articles, hedging, and pleasantries are removed, and invented abbreviations are forbidden because they save no tokens.