token-budget-advisor

Estimate input tokens and present response depth options before answering.

1|1|Updated Mar 31, 2026
One-click install
npx skills add https://github.com/zardusai-cyber/zardus_setup --skill token-budget-advisor-zardusai-cyber
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: token-budget-advisor
Source: https://github.com/zardusai-cyber/zardus_setup/tree/main/ecc/skills/token-budget-advisor
Command: npx skills add https://github.com/zardusai-cyber/zardus_setup --skill token-budget-advisor-zardusai-cyber

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Intercepts the response flow to offer the user a choice about response depth before Claude answers, enabling proactive control over token usage and response length.

Core Features & Use Cases

  • Presents depth options before replying, enabling concise or detailed outputs.
  • Estimates input tokens and projects the potential output window based on prompt complexity.
  • Applies level-based responses to balance conciseness and depth across a variety of tasks and prompts.

Quick Start

Ask the assistant to set a specific response depth for the next reply (e.g., 'use 50% depth for this answer').

Frequently Asked Questions about token-budget-advisor

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I control response length and token budget before Claude answers?

To control response length and token budget, you can trigger depth selection before the assistant answers. The skill intercepts the response flow to offer a choice about response depth, enabling proactive token usage management.

Can I set a specific token budget or depth percentage for AI responses?

Yes, you can set a specific token budget or depth percentage for AI responses by requesting a defined depth level. For example, prompting the assistant to use 50% depth configures the output window to match that chosen response length.

What is token estimation and how does projecting the output window work?

Token estimation calculates input tokens and projects the potential output window based on prompt complexity. This mechanism evaluates the prompt to present suitable depth options, balancing conciseness and detail before generating the response.

How do I get a short version or detailed answer on demand from the assistant?

To get a short version or detailed answer on demand, use triggers like 'short version' or 'detailed answer' in your prompt. The skill intercepts these phrases to apply level-based responses and adjust the output depth accordingly.

Does this token budget approach work with various tasks or only specific prompts?

This token budget approach works with various tasks and prompts. It handles depth control triggers across different contexts, applying level-based responses to balance conciseness and depth regardless of the specific task type.