Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

Model Services/AWS Bedrock/NVIDIA Nemotron Nano 12B v2 VL BF16
aws_bedrock

NVIDIA Nemotron Nano 12B v2 VL BF16

Active

Nemotron multimodal model for visual reasoning and agentic AI workflows

View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Go to pricing page →
Input / 1M tokens
$0.20
Output / 1M tokens
$0.60
Record updated 2026-10-05
Modality:LLM
Context:131,072 tokens
Release:Oct 28, 2025

Pricing

Platform
Scope
Input
Output
Details
Aws_bedrock
Standard
$0.20
$0.60
131k context · 8k max output

How much are you overpaying?

Enter your monthly spend on this model to see what the same workload could cost on cheaper alternatives.

$
Enter a monthly spend to see the listed-rate comparison.

Latest changes

No recent changes.

Benchmarks

No benchmark data available.

⚡NVIDIA Nemotron Nano 12B v2 VL BF16 Comparisons & Scenario Guides

❓NVIDIA Nemotron Nano 12B v2 VL BF16 Frequently Asked Questions

How much does NVIDIA Nemotron Nano 12B v2 VL BF16 cost per 1M tokens?▼

NVIDIA Nemotron Nano 12B v2 VL BF16 is priced at $0.20 per 1M input tokens and $0.60 per 1M output tokens on aws_bedrock. Check provider terms before deployment.

What is the context window for NVIDIA Nemotron Nano 12B v2 VL BF16?▼

NVIDIA Nemotron Nano 12B v2 VL BF16 supports up to 131,072 tokens context with a max completion limit of 8,192 tokens.

What are cost-saving alternatives to NVIDIA Nemotron Nano 12B v2 VL BF16?▼

ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.

AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.