llm-models

Access over 100 LLMs via inference.sh CLI and OpenRouter.

Updated Aug 27, 2026
One-click install
npx skills add https://github.com/RomainGRAS42/Procedio-AI --skill llm-models
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: llm-models
Source: https://github.com/RomainGRAS42/Procedio-AI/tree/main/.agents/skills/llm-models
Command: npx skills add https://github.com/RomainGRAS42/Procedio-AI --skill llm-models

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a unified interface to access a vast array of Large Language Models (LLMs) through the inference.sh CLI and OpenRouter, simplifying AI model selection and integration.

Core Features & Use Cases

  • Unified LLM Access: Interact with over 100 LLMs, including popular ones like Claude, Gemini, and Kimi, through a single API.
  • Cost & Performance Optimization: Features automatic fallback and cost optimization for model selection.
  • Use Case: Developers can easily experiment with different LLMs for tasks like code generation, content creation, or building AI agents without managing multiple API keys or SDKs.

Quick Start

Use the infsh app run openrouter/claude-sonnet-45 command to explain quantum computing.

Frequently Asked Questions about llm-models

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I access multiple LLMs like Claude and Gemini through a single API?

Use the inference.sh CLI with OpenRouter to access over 100 LLMs including Claude and Gemini through a single API. It supports diverse natural language processing tasks like code generation and content creation without managing multiple API keys.

What is the best way to run LLM inference with automatic fallback and cost optimization?

Run LLM inference with automatic fallback and cost optimization by using the inference.sh CLI integrated with OpenRouter. This approach manages model selection across 100+ language models to optimize costs for tasks like reasoning and AI agent building.

Can I use the inference.sh CLI for code generation and building AI agents?

Yes, the inference.sh CLI supports code generation and building AI agents by providing access to over 100 LLMs via OpenRouter. Developers can experiment with different models for content creation and agent workflows through a single API.

Do I need multiple API keys to use different language models for chat and content generation?

No, you do not need multiple API keys to use different language models for chat and content generation. The Skill integrates with multiple LLM providers through a single API via OpenRouter and the inference.sh CLI.

How do I run a specific language model like Claude Sonnet using the infsh app?

Run a specific language model like Claude Sonnet by executing the command `infsh app run openrouter/claude-sonnet-45` in the inference.sh CLI. This directly invokes the model for tasks such as explaining complex topics.

Does OpenRouter support natural language processing tasks for AI assistants and reasoning?

Yes, OpenRouter supports natural language processing tasks for AI assistants and reasoning. The integration provides access to over 100 LLMs tailored for chat, content generation, and complex reasoning workflows.