What problem does it solve?
Nemotron 3 Nano users need fast, authoritative answers about the model’s architecture, training data, recipes, evaluation results, quantization, and deployment behavior without wading through a large repository.
Core Features & Use Cases
- Paper-first retrieval: resolves questions using the Nano3 tech report chunks for architecture, pretraining, SFT, RL (RLVR/GRPO/RLHF), evaluation, and safety/alignment construction.
- Public-recipe grounding: maps what the public repo’s Nano3 stage recipes expose, including where they match the paper “shape” versus where they do not reproduce proprietary mixtures.
- Model-card checkpoint guidance: answers which released checkpoints exist (Base/BF16/FP8), what they’re for, and how to interpret context and reasoning controls for deployment.
- Handoff boundary: when the user’s goal becomes procedural (build, fine-tune, reproduce pipelines, customize to hardware/data), the skill directs them to /nemotron-customize.
Quick Start
Use the nemotron-nano3 skill when you ask, in plain English, for facts about how Nemotron 3 Nano works (for example: “How many experts are activated per token, and what is the active-parameter count?”).