Speech To Text
Speaker Diarization Explainer: How It Works and Key Use Cases
Understand how speaker diarization partitions audio by speaker, its challenges, and use cases in transcription, analytics, and AI-powered applications.
Google Launches Gemini 3.5 Transcribe for Smarter Speech-to-Text
Google unveils Gemini 3.5 Transcribe, its most accurate speech-to-text model yet, with features like real-time streaming and multi-speaker attribution.
Together AI Claims Fastest Speech-to-Text Stack with Parakeet v3
Together AI unveils its fastest ASR stack, leveraging NVIDIA Parakeet v3 and Whisper for real-time, low-latency transcription. Details on the tech and market impact.
Mistral AI Launches Voxtral Transcribe 2 With Sub-200ms Latency
Mistral releases Voxtral Transcribe 2 with real-time streaming at $0.003/min, undercutting competitors while matching accuracy. Open weights under Apache 2.0.
ElevenLabs Introduces Scribe v2 Realtime for Enhanced Speech-to-Text Capabilities
ElevenLabs launches Scribe v2 Realtime, offering low-latency speech-to-text transcription in under 150 ms across multiple languages, enhancing live voice applications.
AssemblyAI Expands Speech-to-Text Capabilities with 99 Languages
AssemblyAI enhances its speech-to-text services by introducing support for 99 languages, offering advanced features at a single price point. Explore the latest developments in AI-driven language recognition.
AssemblyAI's Universal-2 Model Expands Language Coverage and Features
AssemblyAI's Universal-2 model now supports 99 languages, offering advanced features at a single price, enhancing its speech-to-text capabilities and leading in English, German, and Spanish.
AssemblyAI Unveils Advanced Speech-to-Text Technology
AssemblyAI introduces Universal-Streaming, enhancing speech-to-text capabilities with faster transcripts, improved accuracy, and scalable pricing.
AssemblyAI Unveils Advanced Streaming Speech-to-Text Solutions
AssemblyAI introduces cutting-edge streaming speech-to-text technology, offering ultra-fast and accurate solutions for AI voice agents, enhancing real-time transcription capabilities.
AssemblyAI Enhances Speech-to-Text Technology with New Features
AssemblyAI's latest newsletter highlights advancements in speech-to-text models, including improved speaker diarization and new billing alerts, showcasing their commitment to innovation.