This is ZeroRuntime's home on Hugging Face for the models we build in house and the research behind them. Everything here is designed to give developers and researchers production-grade starting points for real-time voice, backed by the same work that powers the ZeroRuntime platform.
An AI voice agent infrastructure platform for building, deploying and running production voice AI agents across web, mobile, telephony and physical devices.
| Runtime | Speech, models and audio in one system. Up to 70% lower latency |
| SDKs | Voice agents for web, mobile, telephony and devices |
| Cloud | Managed infrastructure, agents start on demand |
| Playground | Test live, switch models, inspect latency, tune prompts |
We build proprietary models when off-the-shelf solutions fall short on speed or accuracy.
Knowing when a person has actually finished talking is the difference between a conversation and an interruption. Silence timers guess. Echo does not.
| Model | Reads | Built for |
|---|---|---|
| Echo Small | Transcript | The lowest latency |
| Echo Large | Transcript | Higher accuracy |
| Echo Omni | Audio + transcript | Multimodal accuracy across 27 languages |
Docs · Code samples · Pricing · Blog · Community
San Francisco · Surat