🔥 20% OFF All Kling Models

Grab the Offer
Skip to main content

NVIDIA: Nemotron Nano 12B 2 VL (free)

NVIDIA: Nemotron Nano 12B 2 VL (free)

Choose a prompt below to get started or type your own message

Open Source Alternatives

Similar models we host ourselves — unlimited on the $149/month Open Source Unlimited plan, where this model is billed per call from your wallet

View all open source models

Looking for something different?

Models that trade off differently against the one you are viewing

Related Models

Discover similar models you might be interested in

View all LLM Models

About NVIDIA: Nemotron Nano 12B 2 VL (Free)

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

Technical Specifications

Model ID
nvidia-nemotron-nano-12b-v2-vl-free
Category
LLM Models
Task
Text Generation
Added
February 20, 2026

Key Features

  • Chat completion and multi-turn conversation API
  • Streaming response with token-by-token output
  • Function calling and tool use support
  • System prompts and role-based messaging
  • JSON mode and structured output

Quick Start

Integrate NVIDIA: Nemotron Nano 12B 2 VL (Free) into your application with a single API call. Get your API key from the pricing page to get started.

import requests
import json
url = "https://modelslab.com/api/v7/llm/chat/completions"
headers = {
"Content-Type": "application/json"
}
data = {
"model_id": "nvidia-nemotron-nano-12b-v2-vl-free",
"messages": [
{
"role": "user",
"content": "Hello!"
}
],
"max_tokens": 1000,
"key": "YOUR_API_KEY"
}
try:
response = requests.post(url, headers=headers, json=data)
response.raise_for_status() # Raises an HTTPError for bad responses (4XX or 5XX)
result = response.json()
print("API Response:")
print(json.dumps(result, indent=2))
except requests.exceptions.HTTPError as http_err:
print(f"HTTP error occurred: {http_err} - {response.text}")
except Exception as err:
print(f"Other error occurred: {err}")

Use Cases

  • AI chatbots and virtual assistants
  • Code generation and developer tools
  • Content writing and copywriting automation
  • Data analysis, summarization, and extraction

NVIDIA: Nemotron Nano 12B 2 VL (Free) FAQ

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

You can integrate NVIDIA: Nemotron Nano 12B 2 VL (Free) into your application with a single API call. Sign up on ModelsLab to get your API key, then use the model ID "nvidia-nemotron-nano-12b-v2-vl-free" in your API requests. We provide SDKs for Python, JavaScript, and cURL examples in the API documentation.

The model ID for NVIDIA: Nemotron Nano 12B 2 VL (Free) is "nvidia-nemotron-nano-12b-v2-vl-free". Use this ID in your API requests to specify this model.

Yes. ModelsLab is subscription-based with no free tier — plans start at $21/month (Basic, 3,250 API calls). The $149/month Open Source Unlimited plan includes unlimited generation across open-source models; premium third-party models are billed per call from your wallet.