How AI is applied across API Evangelist and APIs.io. Read my AI disclosure →
API Evangelist API Evangelist
Discovery
Learnings
Guidance
Toolbox
Alignment
API Evangelist LLC

Building Voice Agents with NVIDIA Open Models

calendar_today January 6, 2026 person Kwindla Hultman Kramer domain daily-co

How to Build Ultra-low-latency Voice Agents With NVIDIA Cache-aware Streaming ASRThis post accompanies the launch of NVIDIA Nemotron Speech ASR on Hugging Face. Read the full model announcement here.In this post, we’ll build a voice agent using three NVIDIA open models:The new Nemotron Speech ASR modelNemotron 3 Nano LLMA preview checkpoint of the upcoming NVIDIA Magpie text-to-speech modelThis voice agent leverages the new streaming ASR model, Pipecat’s low-latency voice agent building blocks, and some fun code experiments to optimize all three models for very fast response times.

open_in_new Read original post