Nemotron Nano 9B V2 API

NVIDIA Nemotron Nano 9B V2 is a compact yet capable language model built to balance performance, efficiency, and accessibility.
Output

How to use Nemotron Nano 9B V2 API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to .

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Nemotron Nano 9B V2 API Pricing

TypePrice
Input
Output

Nemotron Nano 9B V2 Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
7.4
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 22, 2026
Math
69.7
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 22, 2026
AIME 2025
72.1%
Competition mathematics (AIME), 2025SourceSeptember 22, 2026
GPQA Diamond
64.0%
Google-proof graduate science questions (hardest subset)SourceSeptember 22, 2026
LiveCodeBench
71.1%
Contamination-free competitive programming problemsSourceSeptember 22, 2026
Humanity's Last Exam
6.5%
Expert-level questions across many domainsSourceSeptember 22, 2026

Nemotron Nano 9B V2 vs other models

ModelInputOutputContextBest for
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$3.9 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Yes, Nemotron Nano 9B V2 can stream responses as they are generated.

Yes, Nemotron Nano 9B V2 supports both function calling and structured outputs.

Nemotron Nano 9B V2 was built by NVIDIA.

No. Its listed capabilities are function calling, streaming, and structured outputs, with no image input support mentioned.

Requests are sent to the v1/chat/completions endpoint.

Start building with Nemotron Nano 9B V2

Get API Key
1000+ models, one API.