Tag
6 articles
Learn how to use S1-mini, a 462 MB open-weight text normalizer that cleans up raw ASR transcripts by removing fillers and self-corrections.
In 2026, the open-source speech recognition landscape has diversified beyond Whisper's dominance, with several models now competing closely on performance metrics. A detailed comparison reveals nuanced trade-offs in accuracy, language support, and latency.
Learn how a new AI model called diffusion-gemma-asr-small uses a 'diffusion' approach to transcribe speech in six languages more efficiently than traditional methods.
NVIDIA's Canary-1B-v2 model enables developers to build multilingual ASR and translation pipelines with automatic SRT subtitle export, showcasing advancements in AI-powered speech processing.
IBM has launched two new Granite Speech 4.1 2B models — one autoregressive for high-accuracy speech recognition with translation, and one non-autoregressive for fast inference.
Cohere AI has released Cohere Transcribe, a state-of-the-art automatic speech recognition model designed to transform audio into actionable text for enterprise use cases.