caveman

Compress AI responses to reduce token consumption while preserving technical accuracy.

1|Updated May 10, 2026
One-click install
npx skills add https://github.com/Avihusitton/gil-therapy --skill caveman-avihusitton
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/Avihusitton/gil-therapy/tree/main/.agents/skills/caveman
Command: npx skills add https://github.com/Avihusitton/gil-therapy --skill caveman-avihusitton

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill eliminates wasted token usage in AI responses caused by unnecessary filler language, pleasantries, and redundant phrasing, which drains context window space, increases processing costs, and slows down technical workflows.

Core Features & Use Cases

  • 6 Adjustable Intensity Levels: Choose from lite, full, ultra, wenyan-lite, wenyan-full, and wenyan-ultra to match your needed balance of brevity and clarity.
  • Auto-Trigger Support: Automatically activates when you request shorter responses, token efficiency, or use trigger phrases like "caveman mode" or "be brief".
  • Use Case: Perfect for developers debugging code, reviewing pull requests, or getting quick technical answers without sifting through fluff, while preserving exact formatting for code blocks, error messages, and API references.

Quick Start

Request a concise explanation of your current technical issue using caveman mode full intensity to get a direct, fluff-free response.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce AI token usage during code debugging and technical troubleshooting?

You can reduce AI token usage by applying ultra-compressed, fluff-free communication that eliminates unnecessary filler language while retaining full technical accuracy for code debugging and troubleshooting workflows.

What is the best way to get concise AI responses for pull request reviews without losing technical details?

Concise AI responses for pull request reviews are achieved by using compressed communication modes that strip pleasantries and redundant phrasing while preserving exact formatting for code blocks, error strings, and API references.

Can I adjust the intensity of AI response compression for different technical query scenarios?

Yes, you can adjust response compression intensity across 6 tiers—lite, full, ultra, wenyan-lite, wenyan-full, and wenyan-ultra—to match your needed balance of brevity and clarity for technical queries.

Does compressed AI communication still preserve exact formatting for error messages and API references?

Compressed AI communication preserves exact formatting for code blocks, error strings, and API references, ensuring technical accuracy is maintained even when response token consumption is reduced by approximately 75%.

How do I automatically trigger shorter AI responses for software engineering tasks?

Shorter AI responses auto-trigger when you request token efficiency, brevity, or use specific trigger phrases like "caveman mode" or "be brief" during software engineering tasks.

When should I not use ultra-compressed AI responses for technical explanations?

You should avoid ultra-compressed AI responses when maximum clarity is prioritized over token efficiency, or when your context requires extensive conversational detail rather than direct, fluff-free technical explanations.