vllm-nvidia-hardware

Size NVIDIA vLLM deployments by hardware SKU, power, and facility readiness.

5|1|Updated Apr 19, 2026
One-click install
npx skills add https://github.com/air-gapped/skills --skill vllm-nvidia-hardware
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: vllm-nvidia-hardware
Source: https://github.com/air-gapped/skills/tree/main/.claude/skills/vllm-nvidia-hardware
Command: npx skills add https://github.com/air-gapped/skills --skill vllm-nvidia-hardware

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

vllm-nvidia-hardware provides a comprehensive, SKU-grounded reference for sizing and procuring NVIDIA hardware to run vLLM at scale, consolidating SKUs, power, cooling, and platform readiness into actionable guidance.

Core Features & Use Cases

  • Comprehensive SKU matrices for Hopper through Vera Rubin NVL72, including B200/B300, GB300, and Rubin roadmap.
  • Facility prerequisites and procurement considerations to inform data-center readiness and capital planning.
  • Use Case: sizing a GB300 NVL72 rack for a 70B FP4 model with 1M context windows.

Quick Start

Ask this skill to help size NVIDIA hardware for vLLM deployments across Hopper through Rubin.

Frequently Asked Questions about vllm-nvidia-hardware

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I size NVIDIA hardware for vLLM deployments across Hopper to Rubin?

Size NVIDIA vLLM hardware by evaluating SKU matrices, power budgets, and cooling prerequisites for Hopper through Rubin platforms. Apply facility readiness checks and platform-matrix references to guide procurement and risk assessment for enterprise deployments.

What NVIDIA SKUs are supported for vLLM hardware sizing?

Supported NVIDIA SKUs for vLLM hardware sizing include B200, B300, GB300 NVL72, and Vera Rubin roadmap platforms. SKU tables cover Hopper through Rubin architectures to inform deployment scale and context.

Can I use vllm-nvidia-hardware for GB300 NVL72 rack procurement planning?

Yes, use this reference to size a GB300 NVL72 rack deployment for vLLM, including facility prerequisites and capital planning. It applies SKU matrices and power constraints to procurement decisions for Dell, Lenovo, and HGX racks.

How do I calculate power and cooling requirements for a 70B FP4 vLLM deployment?

Calculate power and cooling requirements for a 70B FP4 vLLM deployment by applying facility prerequisites and platform-matrix references. Check data-center readiness and power budgets against specific GB300 NVL72 hardware SKUs for 1M context windows.

Does vllm-nvidia-hardware compare Dell, Lenovo, and HGX rack platforms?

Yes, it compares Dell, Lenovo, and HGX rack platforms for vLLM deployments using platform-matrix references. Assess hardware sizing, facility readiness, and risk across these platforms for Hopper through Rubin architectures.

When should I not use this skill for vLLM hardware sizing?

Avoid using this skill for vLLM hardware sizing when evaluating non-NVIDIA accelerators or platforms outside the Hopper through Rubin roadmap. It strictly relies on NVIDIA SKU tables, power constraints, and facility prerequisites.