Expanding the Pareto Frontier of STT with Nari Qwen3-ASR
Nari Qwen3-ASR is now available in Free Public Beta. Fast and Standard endpoints are free for a limited time.
Ship multimodal models with lower latency, higher efficiency and less infrastructure work.
Model APIs
Optimized multimodal model endpoints built to be fastest in production: starting with speech.
Learn MoreLatencyElevenLabs vs Nari FastTTFA
CostElevenLabs vs Nari StandardUSD / 1M characters
LatencyDeepgram vs Nari FastTTFS
CostDeepgram vs Nari StandardUSD / hour of audio
Based on preliminary independent measurements of publicly available TTS and STT endpoints from US East. Results may vary by region, connection, and workload. Free endpoints are best effort and carry no latency or uptime SLO. Free Public Beta is available for a limited time; Early Access pricing is shown. Pricing checked Sep 8, 2026 from Artificial Analysis.
Model library
Sub 50 ms server-side TTFA$10 / 1M characters
Text to SpeechFree Public Beta$5 / 1M characters
Text to SpeechFree Public BetaSub 40 ms server-side TTFS$0.12 / hour
Speech to TextFree Public Beta$0.06 / hour
Speech to TextFree Public BetaInference
Model and workload-specific runtimes deliver predictable latency and production-scale serving across managed, dedicated, and private infrastructure.
Learn MoreModel-specific optimizations
Kernels, batching, quantization, and routing tuned to your workload.
Proven on our stack
The same technology we used to achieve the Pareto frontier and #1 on Benchmarks.
Deployment flexibility
Managed, dedicated, or private infrastructure.
Training
Fine-tune multimodal models on data that matches your use-case and deploy using our optimized stack. Work with the team behind Dia and Narvatar.
Learn More
Open-source dialogue TTS
Dia
Our dialogue TTS model series with 2M+ downloads and 20K+ GitHub stars.
In-house avatar system
Narvatar
Our best-in-class human-centric avatar model under 5B, available in both streaming and non-streaming variants. Optimized for realtime performance and low costs.
Coming next
Your Model
Expand the frontier by working with us
What's new
Nari Qwen3-ASR is now available in Free Public Beta. Fast and Standard endpoints are free for a limited time.
Nari Qwen3-TTS Free Public Beta. Fast and Standard endpoints are free for a limited time.
Preliminary independent measurements put our TTS endpoint first in the benchmark snapshot.
Use our pre-built endpoints, bring your workload, or fine-tune a model with us. All served on our specialized inference stack.