llm-rankings

Compare LLM benchmarks, pricing, and capabilities for task-specific recommendations.

Updated Aug 23, 2026
One-click install
npx skills add https://github.com/andaydvice/smart-rv-portal --skill llm-rankings
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: llm-rankings
Source: https://github.com/andaydvice/smart-rv-portal/tree/main/.claude/skills/llm-rankings
Command: npx skills add https://github.com/andaydvice/smart-rv-portal --skill llm-rankings

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Users struggle to choose the best LLM for their specific needs, often getting lost in benchmarks, pricing, and capabilities. This skill cuts through the complexity, providing clear, data-driven recommendations to help you select the optimal model for any task.

Core Features & Use Cases

  • Comprehensive LLM Comparison: Get benchmark-based rankings (MMLU, HumanEval), cost analysis, and real-world performance insights for models like GPT-4, Claude, Gemini, and Llama.
  • Task-Specific Recommendations: Find the ideal LLM for code generation, long context tasks, creative writing, complex reasoning, or cost-sensitive applications.
  • Use Case: "Which LLM is best for generating Python code, considering both performance and cost?"

Quick Start

Use the llm-rankings skill to compare GPT-4o, Claude Sonnet 4.5, and Gemini 1.5 Pro for long context document analysis, including their pricing and context window sizes.