Unlimited Open Source Models

Get Plan
Skip to main content

NVIDIA: Nemotron 3 Ultra (batch)

NVIDIA: Nemotron 3 Ultra (batch)

Choose a prompt below to get started or type your own message

About NVIDIA: Nemotron 3 Ultra (Batch)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

Technical Specifications

Model ID
nvidia-nemotron-3-ultra-550b-a55b-batch
Category
LLM Models
Task
Text Generation
Price
$1.05 per million tokens
Added
August 6, 2026

Key Features

  • Chat completion and multi-turn conversation API
  • Streaming response with token-by-token output
  • Function calling and tool use support
  • System prompts and role-based messaging
  • JSON mode and structured output

Quick Start

Integrate NVIDIA: Nemotron 3 Ultra (Batch) into your application with a single API call. Get your API key from the pricing page to get started.

import requests
import json
url = "https://modelslab.com/api/v7/llm/chat/completions"
headers = {
"Content-Type": "application/json"
}
data = {
"model_id": "nvidia-nemotron-3-ultra-550b-a55b-batch",
"messages": [
{
"role": "user",
"content": "Hello!"
}
],
"max_tokens": 1000,
"key": "YOUR_API_KEY"
}
try:
response = requests.post(url, headers=headers, json=data)
response.raise_for_status() # Raises an HTTPError for bad responses (4XX or 5XX)
result = response.json()
print("API Response:")
print(json.dumps(result, indent=2))
except requests.exceptions.HTTPError as http_err:
print(f"HTTP error occurred: {http_err} - {response.text}")
except Exception as err:
print(f"Other error occurred: {err}")

Pricing

NVIDIA: Nemotron 3 Ultra (Batch) API costs $1.05 per million tokens. Pay only for what you use with no minimum commitments. View pricing plans

Use Cases

  • AI chatbots and virtual assistants
  • Code generation and developer tools
  • Content writing and copywriting automation
  • Data analysis, summarization, and extraction

NVIDIA: Nemotron 3 Ultra (Batch) FAQ

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

You can integrate NVIDIA: Nemotron 3 Ultra (Batch) into your application with a single API call. Sign up on ModelsLab to get your API key, then use the model ID "nvidia-nemotron-3-ultra-550b-a55b-batch" in your API requests. We provide SDKs for Python, JavaScript, and cURL examples in the API documentation.

NVIDIA: Nemotron 3 Ultra (Batch) costs $1.05 per million tokens. ModelsLab plans start at $21/month (Basic), and the $149/month Open Source plan includes unlimited generation on open-source models.

The model ID for NVIDIA: Nemotron 3 Ultra (Batch) is "nvidia-nemotron-3-ultra-550b-a55b-batch". Use this ID in your API requests to specify this model.

Yes. ModelsLab is subscription-based with no free tier — plans start at $21/month (Basic, 3,250 API calls). The $149/month Open Source Unlimited plan includes unlimited generation across open-source models; premium third-party models are billed per call from your wallet.