← All models
nvidia
API MODEL PROFILE / nvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

nvidia/nemotron-3.5-asr-streaming-multilingual-0.6b
View on OpenRouter
Catalogue observation · 2026-09-22T10:13:34.498Z · Source: OpenRouter
transcriptionCatalogue only · output not supported by the text API

About this API

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Provider-supplied description, not an independently validated performance claim.
Input / 1 million tokens$3.33Published prompt tariff
Output / 1 million tokens$0Published completion tariff
Context windowNot listedPublished token capacity

These are the catalogue’s token tariffs. Image, audio, video, search and other charges may use different billing units. Zero token rates alone do not mean those services are free. See the complete pricing fields below and the source model page.

Pros and useful characteristics
  • Published input token pricing is at or below the median of catalogue entries with the same output modalities and known tariffs.
  • Published output token pricing is at or below the same-modality catalogue median.
  • No separate per-request fee is listed in this catalogue observation. Other modality charges may apply.
Cons and practical limits
  • No positive token context window is published for this model.
  • Catalogue only · output not supported by the text API. View the source model page for its API requirements.
  • No independent latency, uptime, reasoning-quality or coding benchmark has been measured by this application.

Assessment uses published fields and unweighted price medians across 21 entries with the same output modalities and known tariffs. It is a cost-and-capability comparison, not a quality ranking or recommendation.

Observed price history

No observations yet. The chart starts with the first recorded data point.

No observations yet. The chart starts with the first recorded data point.

One catalogue observation per hour. History begins when this deployment first observes the model. No backfilled or interpolated prices.

Published model tariffsUSD / 1M tokens
InputOutput
NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B
$3.33
$0.00

Zero baseline. Published token tariffs, not total multimodal cost or performance scores.

API availability

Catalogue only · output not supported by the text API.

Trident currently serves text Chat Completions with known token prices. This profile is available for discovery; use the upstream model’s documentation for other API formats and billing units.

Read the catalogue guide ↗
Technical profile
Canonical model
nvidia/nemotron-3.5-asr-streaming-multilingual-0.6b-20260813
Separate request tariff
$0
Upstream output limit
Not published
Gateway output limit
Not supported
Upstream inputs
audio
Upstream outputs
transcription
Knowledge cutoff
Not published

Advertised upstream parameters

Not published

Upstream capabilities are separate from the text subset exposed by this gateway.

View source model page

Complete published pricing

Source fields are retained exactly, including extra charges and tier overrides. Negative values indicate variable pricing, not rebates. An absent field is not evidence that a service is free.

View OpenRouter pricing fields
{
  "prompt": "0.00000333",
  "completion": "0"
}
View nvidia profile ↗