caveman

Compress AI chat responses into caveman-style fragments to reduce token consumption.

Updated Aug 27, 2026
One-click install
npx skills add https://github.com/tDalile/dotfiles --skill caveman-tdalile
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/tDalile/dotfiles/tree/main/agents/skills/caveman
Command: npx skills add https://github.com/tDalile/dotfiles --skill caveman-tdalile

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

High token consumption inflates costs and slows down AI conversations; this Skill compresses replies to a bare‑bones caveman style, drastically reducing token count while preserving technical accuracy.

Core Features & Use Cases

  • Intensity Levels: Choose from lite, full (default), ultra, and classical Chinese (wenyan) modes to balance brevity and detail.
  • Persistent Mode: Once activated, caveman style stays active across turns until explicitly turned off.
  • Auto‑Clarity Safeguard: Switches back to normal responses for security warnings, irreversible actions, or when clarification is requested.
  • Rule‑Based Compression: Drops articles, filler words, and hedging; uses fragment patterns and abbreviated syntax per intensity level.
  • Use Case Example: When answering technical troubleshooting questions, enable /caveman ultra to deliver concise, token‑efficient steps without redundant wording.

Quick Start

Activate caveman mode with /caveman full to receive terse, token‑saving replies.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce token usage and lower AI chat costs without losing technical accuracy?

You can reduce token usage by applying compressed caveman-style responses that drop articles and filler words. This technique slashes token consumption by up to 75% while preserving technical accuracy through rule-based syntax compression.

What intensity levels can I choose from when compressing AI responses for brevity?

Brevity compression supports lite, full, ultra, and classical Chinese (wenyan) intensity levels. Each mode balances detail and token efficiency differently, allowing you to select the exact compression strength needed for your conversation.

How do I keep concise replies active across multiple turns in a conversation?

Persistent mode keeps caveman-style compression active across multiple turns until explicitly turned off. Once activated with a command like /caveman full, the terse reply style stays engaged automatically throughout the session.

Does the caveman compression mode work safely for security warnings or irreversible actions?

An auto-clarity safeguard automatically switches back to normal responses for security warnings, irreversible actions, or clarification requests. This ensures compressed replies never obscure critical safety information during technical troubleshooting.

When should I avoid using terse replies for AI chat interactions?

You should avoid terse replies when handling security warnings, irreversible actions, or when clarification is requested. The auto-clarity safeguard detects these contexts and automatically restores normal, detailed responses to prevent misunderstandings.

What is the best way to get concise, token-efficient troubleshooting steps from AI?

The best way to get token-efficient troubleshooting steps is enabling an ultra compression mode. This delivers concise instructions by enforcing fragment patterns and dropping redundant wording while maintaining the required technical accuracy.