NVIDIA Nemotron 3 Nano 30B is a hybrid Mamba-2/MoE/Attention model with 3.5B active parameters, trained from scratch as a unified model for both reasoning and non-reasoning tasks. Reasoning traces can be toggled via a chat template flag, and the model supports tool use within a 1M-token context window, making it well suited for specialized agentic applications.
Get started by making your first request to FastRouter. Use your favorite SDK or make direct API calls.
<FASTROUTER_API_KEY> with your actual API keypip install openaiModel scores sourced from ArtificialAnalysis