Reverb ASR+Diarization: Best open-source ASR for long-form audio
WHY IT MATTERS
Reverb is an open-source automatic speech recognition system with integrated speaker diarization, optimized for long-form audio. Its Show HN post has 12 points.
Reverb, an open-source automatic speech recognition system with integrated speaker diarization optimized for long-form audio, was posted on HackerNews as a Show HN with 12 points. This provides a free alternative to commercial ASR services for transcription and meeting analysis, particularly relevant for agent-based voice workflows.
For builders of voice agents and meeting analysis tools, Reverb removes the per-minute cost of cloud ASR APIs and enables on-premise deployment, reducing both latency and privacy exposure. The integrated diarization eliminates the need to chain separate speaker identification services, simplifying pipeline architecture. A second-order effect is increased feasibility of running real-time transcription for long, multi-speaker conversations in edge or air-gapped environments, which previously required expensive or closed-source solutions. This shifts the cost baseline for voice agents from variable API fees to fixed compute overhead.
SOURCE
HackerNews
SHARE
MORE FROM STUFFINSIDER
Open-sourced prompt-injection detector (regex gate + quantised DeBERTa-v3 ONNX)
Jul 20OPEN SOURCEReverb ASR+Diarization: Open Source Long-Form Audio ASR
Jul 11OPEN SOURCETRACE: Open-Source Hierarchical Memory for LLM Agents
Jul 7OPEN SOURCEaudio.cpp: 12 audio models with 5x TTS speedup
Jun 26