Deepdub extended NVIDIA Nemotron 3.5 ASR to do more than transcribe speech, detect speaker gender, and detect end-of-turn beyond silence.
Public source
Publisher name
Public post
Shipping real-time voice AI at scale means squeezing out every millisecond and every GPU. 🔥 That's why we extended NVIDIA Nemotron 3.5 ASR to do more than transcribe. W…
Company
Deepdub
The voice layer for AI in the real world. Built in production. Deployed at scale.
- Industry
- Entertainment Providers
- Location
- Plano, US
- Company size
- 201–500 employees
About Deepdub
Deepdub is the enterprise voice infrastructure powering AI in production. We earned our credibility in the most demanding voice environments in the world: Hollywood studios, global broadcasters, and premium content pipelines where voice quality, emotional accuracy, and reliability are non-negotiable. That production-grade foundation now powers AI agents and real-time systems with complex, multilingual conversations where latency, control, and trust matter as much as intelligence. Built on proprietary eTTS™ foundational models, Deepdub enables zero-shot voice cloning, voice-to-voice, ADR, accent control, and ultra-low latency delivery designed for systems that operate live, at scale, and in front of real customers. Get started here> https://deepdub.ai/
See moreLatest activity
Latest activity from Deepdub
8 signals
Presence & Recognition
Deepdub announced that CBDO Oz Krakowski will be joined on stage by Gilberto Castañon, Director of Content Distribution at Grupo Globo, to discuss agentic orchestration at scale at the IBC 2026 session on Sept 13 at the Future Tech Stage.
Presence & Recognition
Deepdub will be exhibiting at the IBC booth in Hall 14, Stand 14.C51 from September 11 to 14.
Products & Services
Deepdub built the Phantom Z 3.4 voice AI model for real-time conversations with 150ms to first audio, 48 kHz, multilingual support, and full 48 kHz.
Discover more
Similar signals
Similar public activity from other companies.
Products & Services
3Play Media
3Play Media includes human reviewers checking timing, accuracy, and audio quality before anything gets delivered in its AI dubbing.
Products & Services
Deepgram
Deepgram released Frontier research in voice generation on the Flux TTS launch model trained on real restaurant vocabulary.
Products & Services
Soniox
Soniox STT provides native semantic endpointing built directly into real-time transcription across all 60+ supported languages.
Products & Services
Colossyan
Colossyan built dubbing that allows users to speak German without ever having spoken German, keeping their voice intact.
Products & Services
pyannoteAI