// nvidia

Nemotron 3 Nano Omni (free)

nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free

reasoningtool callingaudio inputimage inputvideo input
OpenRouter + LiveBenchAll modelsFull leaderboard

Nemotron 3 Nano Omni (free) is free to call with a 256K-token context window.

A reasoning, tool-calling model from nvidia, released Apr 28, 2026. LiveBench hasn't published a run for it, so this page carries measured price and specs only — no benchmark numbers.

Input / 1M
Free
Output / 1M
Free
Context
256K
LiveBench overall
Not evaluated

Specs and pricing

Input price

USD per 1M prompt tokens

Free
Output price

USD per 1M completion tokens, reasoning included

Free
Blended price

3:1 input:output — #1 cheapest of 343 listed models

Free
Cached input read

Blank means not priced separately — not free

Cache write
Context window256,000 tokens
Max output65,536 tokens
Input modalitiestext, audio, image, video
Tool callingYes
Extended reasoningYes
Knowledge cutoff
ListedApr 28, 2026
All NVIDIA API pricing

What Nemotron 3 Nano Omni (free) costs to run

Monthly list cost across five workload shapes, using the published cached-input rate where there is one. A floor, not a quote — batch discounts and cache writes aren't included.

WorkloadPer month
Support chatbot

1.2K in / 400 out × 200K requests

$0/mo
RAG assistant

8K in / 600 out × 100K requests

$0/mo
Coding agent

40K in / 4K out × 20K requests

$0/mo
Document extraction

20K in / 1.5K out × 50K requests

$0/mo
Bulk classification

500 in / 20 out × 5M requests

$0/mo
Price your own workload

Compare Nemotron 3 Nano Omni (free) with

Benchmarked models at a similar price.

Nemotron 3 Nano Omni (free) FAQ

How much does Nemotron 3 Nano Omni (free) cost?

Nemotron 3 Nano Omni (free) is listed as free on OpenRouter, typically with tighter rate limits than paid models.

What is the context window of Nemotron 3 Nano Omni (free)?

Nemotron 3 Nano Omni (free) accepts up to 256,000 tokens of context and can return up to 65,536 output tokens per response.

Does Nemotron 3 Nano Omni (free) support tool calling?

Yes — Nemotron 3 Nano Omni (free) supports tool (function) calling and exposes extended reasoning. It accepts text, audio, image, video input.

How does Nemotron 3 Nano Omni (free) score on benchmarks?

LiveBench has not published a run for Nemotron 3 Nano Omni (free) yet. Until it does, compare it on price and specs, or run your own evaluation on your own tasks.

What is the API model ID for Nemotron 3 Nano Omni (free)?

On OpenRouter the model ID is "nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free". It was listed on Apr 28, 2026.

More from nvidia

How these numbers are produced

  • Price and specs — provider list data from OpenRouter, refreshed every 15 minutes. “Blended” is a 3:1 input:output mix.
  • Scores and ranksLiveBench release 2026-06-25. Ranks count only models that have a published run; a blank means “not evaluated”, never “bad”.
  • Cost per point — the measured dollars LiveBench spent on the run, divided by the score it earned.

Published benchmarks rank models on someone else's tasks. Before committing, see LLM & agent evaluation for building an eval on your own.