# Models for continuous inference.

Models available on .wave, and the API that serves each.

- [Nemotron VoiceChat 11B](https://dotwave.ai/docs/models/nvidia/nemotron-voicechat-11b/) (Speech ↔ speech · full duplex): `nemotron-voicechat` Live API Listen and respond in the same live session, with audio plus user and assistant transcript events.
- [Nemotron 3.5 ASR Streaming 0.6B](https://dotwave.ai/docs/models/nvidia/nemotron-asr-streaming-0-6b/) (Speech → text · streaming): `nemotron-asr-streaming` Realtime API · Deepgram-compatible API Stream live audio and receive punctuated text as people speak, in 32 languages or with automatic language detection.
