---
title: "Nemotron 3 Ultra AI model: facts and benchmarks | Kendr"
canonical: "https://kendr.org/models/nvidia-nemotron-3-ultra-550b-a55b"
date_modified: "2026-08-12"
profile_kind: "reference"
---

# Nemotron 3 Ultra model profile

Nemotron 3 Ultra is a NVIDIA model profile with a 512,288-token context window. The dated catalog snapshot records text modalities and interfaces for prompt caching, reasoning controls, structured output, text generation, tool calling.

Nemotron 3 Ultra is an independent knowledge profile. It does not claim that Kendr hosts the model, quote a Kendr customer price, promise API availability, or advertise a Kendr routing receipt.

Nemotron 3 Ultra accepts text and returns text. The snapshot records prompt caching, reasoning controls, structured output, text generation, and tool calling as capabilities or interfaces.

The dated reference snapshot lists $0.6 input and $3.6 output per million tokens. Cached input is $0.2 per million tokens.

The 2026-08-12 snapshot includes 3 Artificial Analysis indexes and 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable.

Nemotron 3 Ultra ranked #66 in the cited trailing-7-day third-party catalog usage dataset as of 2026-08-12. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.

Profile data was reviewed for the 2026-08-12 snapshot. Claims retain their source dates, and unavailable fields are shown as unavailable instead of estimated.

## Profile facts

- **Profile type:** Model research profile
- **Provider or publisher:** NVIDIA
- **Context window:** 512,288 tokens
- **Snapshot date:** 2026-08-12
- **Reference catalog ID:** nvidia/nemotron-3-ultra-550b-a55b

## Context, modalities, and identity

Nemotron 3 Ultra accepts text and returns text. The snapshot records prompt caching, reasoning controls, structured output, text generation, and tool calling as capabilities or interfaces.

- **Provider or publisher:** NVIDIA
- **Context window:** 512,288 tokens
- **Maximum output:** No verified numeric limit
- **Input modalities:** text
- **Output modalities:** text
- **Knowledge cutoff:** Not disclosed

## Capabilities and supported controls

Capabilities and parameters are reported from the dated catalog and source records; their presence does not guarantee identical behavior across every provider route.

- **Supported parameters:** frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p

- prompt caching
- reasoning controls
- structured output
- text generation
- tool calling

## Dated pricing snapshot

The dated reference snapshot lists $0.6 input and $3.6 output per million tokens. Cached input is $0.2 per million tokens.

These are dated third-party reference prices, not Kendr prices or an availability offer. Routes, tiers, caching, region, and provider policy can change the landed price.

- **Price date:** 2026-08-12
- **Input:** $0.6 per 1M tokens
- **Cached input:** $0.2 per 1M tokens
- **Output:** $3.6 per 1M tokens

## Benchmark evidence and limitations

The 2026-08-12 snapshot includes 3 Artificial Analysis indexes and 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable.

Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.

- **Intelligence Index:** 38.3
- **Coding Index:** 49.3
- **Agentic Index:** 27.5

| Design Arena category | Elo | Win rate | Rank |
| --- | --- | --- | --- |
| 3d | 1186 | 41.2% | #53 |
| asciiart | 1113 | 36.8% | #51 |
| codecategories | 1155 | 36% | #72 |
| dataviz | 1152 | 37.4% | #71 |
| gamedev | 1166 | 37.8% | #64 |
| svg | 1126 | 37.6% | #50 |
| uicomponent | 1173 | 38.4% | #62 |
| website | 1135 | 33.3% | #82 |

## Popularity and market context

Nemotron 3 Ultra ranked #66 in the cited trailing-7-day third-party catalog usage dataset as of 2026-08-12. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.

- **Third-party catalog rank:** #66
- **Global market share:** Not inferred
- **Observation date:** 2026-08-12

## Frequently asked questions

### What is Nemotron 3 Ultra?

Nemotron 3 Ultra is a NVIDIA model profile with a 512,288-token context window. The dated catalog snapshot records text modalities and interfaces for prompt caching, reasoning controls, structured output, text generation, tool calling.

### What context window does Nemotron 3 Ultra have?

The dated profile lists 512,288 tokens of context. Provider routes, variants, and runtime configuration can impose lower effective limits.

### What benchmark evidence is available for Nemotron 3 Ultra?

The 2026-08-12 snapshot includes 3 Artificial Analysis indexes and 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable. Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.

### Is Nemotron 3 Ultra available through Kendr?

This is an independent knowledge profile, not a Kendr-hosted availability claim. Check Kendr's live public model API for currently enabled Kendr aliases.

### Does the popularity rank represent Nemotron 3 Ultra's global market share?

Nemotron 3 Ultra ranked #66 in the cited trailing-7-day third-party catalog usage dataset as of 2026-08-12. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share. No global market-share percentage is inferred when the source does not publish one.

## Sources and evidence dates

- [Nemotron 3 Ultra third-party catalog record](https://openrouter.ai/nvidia/nemotron-3-ultra-550b-a55b) — catalog, checked 2026-08-12
- [Third-party model catalog methodology](https://openrouter.ai/docs/guides/overview/models) — methodology, checked 2026-08-12
- [NVIDIA official model documentation](https://build.nvidia.com/models) — primary, checked 2026-08-12
- [Artificial Analysis capability indices methodology](https://artificialanalysis.ai/methodology/capability-indices) — benchmark, checked 2026-08-12
- [Design Arena leaderboard and methodology](https://www.designarena.ai/leaderboard) — benchmark, checked 2026-08-12

Live Kendr operational availability is published separately at https://api.kendr.org/api/public/models.
