Amazon Launches New ASR System Supporting Over 100 Languages
8.9K views
Amazon has released a next-generation ASR system that covers over 100 languages, providing comprehensive automatic speech recognition services. The speech foundation model improves accuracy by 20% to 50%, with enhancements of 30% to 70% in challenging areas such as telephone speech. The system supports multiple features, including automatic punctuation, custom vocabulary, automatic language identification, and speaker separation. Thousands of businesses are leveraging Amazon Transcribe to unlock insights from audio content, enhancing accessibility and discoverability.
Related
AliQwen-Audio-3.0 Series Speech Models Launched on Tongyi AI Platform, Winning a Grand Slam in International Evaluations
17.0K
Qwen Launches a Major Upgrade: Real-Time Speech Recognition Model Fun-ASR-Realtime Officially Released
16.2K

iFlytek Launches AI Hardware and Software Integrated Solution: Accurate Recognition Even in 90 Decibel Noise
14.6K
Alibaba Launches Revolutionary Speech Recognition Model FunAudio-ASR with Remarkable Noise Reduction
20.5K
OpenAI Evals Adds Native Audio Input and Evaluation Features
12.0K

Chinese Visual and Speech Open Source Model VITA-1.5 Released with GPT-4o Level Advanced Speech and Visual Capabilities
14.6K
