optimize

Identify API cost inefficiencies and performance bottlenecks across infrastructure dimensions.

4|Updated Jul 20, 2026
One-click install
npx skills add https://github.com/highflame-ai/ai-factory --skill optimize-highflame-ai
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: optimize
Source: https://github.com/highflame-ai/ai-factory/tree/main/skills/optimize
Command: npx skills add https://github.com/highflame-ai/ai-factory --skill optimize-highflame-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill addresses the challenge of identifying hidden infrastructure costs and performance bottlenecks in complex API-driven architectures.

Core Features & Use Cases

  • Multi-Dimensional Scanning: Simultaneously evaluates API costs, database performance, and latency metrics.
  • Actionable Reporting: Generates prioritized recommendations categorized by impact, effort, and risk.
  • Use Case: Use this when you suspect your LLM usage or database queries are becoming inefficient, allowing you to pinpoint specific endpoints that require optimization before they impact your budget or user experience.

Quick Start

Run the optimize skill to scan the entire project for cost and performance improvements.

Frequently Asked Questions about optimize

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I identify API cost inefficiencies and performance bottlenecks in my project?

To identify API cost inefficiencies and performance bottlenecks, you can scan your project using parallel agents that analyze architecture documentation and context. This provides prioritized optimization recommendations based on cost and latency assessments.

What is the best way to analyze infrastructure latency and database performance simultaneously?

The best way to analyze infrastructure latency and database performance simultaneously is through multi-dimensional scanning. This approach evaluates API costs, database queries, and latency metrics in parallel to generate actionable, prioritized reports.

When do I need to scan my architecture for LLM usage and database query optimization?

You need to scan your architecture for LLM usage and database query optimization when you suspect they are becoming inefficient. Pinpointing specific endpoints early prevents budget overruns and negative user experience impacts.

Do I need project-specific context files to perform accurate API cost and latency assessments?

Yes, you need project-specific context files to perform accurate API cost and latency assessments. The scanning agents require integration with your architecture documentation to evaluate the project context correctly.

How are optimization recommendations categorized to help prioritize infrastructure improvements?

Optimization recommendations are categorized by impact, effort, and risk. This prioritized reporting structure helps you decide which API cost and performance issues to address first for instant gains.