Tech Meridian← ALL MODELS
PROMY MERIDIAN

NVIDIA · MODEL RELEASE TRACKER

NVIDIA Nemotron-3 Nano 30B

NVIDIA Nemotron-3 Nano 30B is presented as a generative AI model that can be deployed to Amazon SageMaker AI endpoints; it is used in an automated concurrency-sweep workflow to measure endpoint throughput and latency.

CURRENT SNAPSHOT2/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING

Not established from the available sources.

CONTEXT WINDOW

Not established from the available sources.

MODALITIES

Not established from the available sources.

BENCHMARKS
Benchmarking (concurrency sweeps)The model is used in an automated concurrency-sweep workflow that measures throughput (tokens per second) and latency to identify endpoint saturation and right-size capacity.
AVAILABILITY
DeployabilityCan be deployed to Amazon SageMaker AI endpoints using the native vLLM container.

VERIFIABLE FACTS

Every value stays attached to a source and date.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport created3 facts