Real-time speech and talking-head models Collection Open models for real-time voice pipelines: VAD, diarization, streaming ASR, TTS, full-duplex speech, and talking-head animation. • 46 items • Updated 4 days ago • 1