0:22Microsoft's open-source VibeVoice-ASR: one pass over a 60-minute meeting yields speakers, timestamps and content; 54k GitHub stars@axichuhai · 106 views · 2026-09-30Voice InteractionASROpen Source
0:28Audio8 ASR Infinite: rolling KV cache enables 24/7 drift-free streaming ASR (Apache 2.0)@_akhaliq · 137 views · 2026-09-30Audio8 ASR InfiniteStreaming ASRRolling KV Cache
0:59Fish Audio upgrades its ASR model: speaker identification and inline emotion cues such as [laughter]@FishAudio · 338 views · 2026-09-28ASRSpeech RecognitionEmotion Recognition
0:54Nemotron 3 Diarization tested: 100M-param real-time speaker diarization, better than expected locally@ouchi · 211 views · 2026-09-27NVIDIANemotron 3Speaker Diarization
0:10JEV-Speech: Same 24-Layer Encoder, 2.18x Faster Inference@nrol_ling · 260 views · 2026-09-22Speech RecognitionInference OptimizationLatency