Modèles de reconnaissance vocale
Le catalogue des modèles de reconnaissance vocale (speech-to-text), classés par éditeur. Chaque fiche détaille le taux d'erreur de mots (WER) et les usages.
Modèles de reconnaissance vocale
24 modèles de reconnaissance vocale (speech-to-text), classés par éditeur.
| Éditeur | Modèle | WER | Licence | Date de sortie |
|---|---|---|---|---|
| xAI | xAI: Grok STT 1.0 | n.d. | ▫ n.d. | 23 juil. 2026 |
| Grok Speech to Text, xAI | 4,03 % | ▫ n.d. | 18 avr. 2026 | |
| Deepgram | Nova 3 | 5,18 % | 🔒 Propriétaire | 15 juil. 2026 |
| Smallest.ai | Smallest AI Pulse Pro | 2,43 % | ▫ n.d. | 16 juin 2026 |
| Soniox | Soniox v5 Async | 3,81 % | 🔒 Propriétaire | 11 juin 2026 |
| Gladia | Solaria-3, Gladia | 3,23 % | 🔒 Propriétaire | 10 juin 2026 |
| Microsoft | Microsoft: MAI-Transcribe 1.5 | 2,38 % | ▫ n.d. | 2 juin 2026 |
| NVIDIA | NVIDIA: Parakeet TDT 0.6B v3 | n.d. | ▫ n.d. | 27 mai 2026 |
| Mistral AI | Mistral: Voxtral Mini Transcribe | 3,53 % | ▫ n.d. | 15 mai 2026 |
| Voxtral Small | 2,77 % | ▫ n.d. | 15 juil. 2025 | |
| Google: Chirp 3 | n.d. | ▫ n.d. | 5 mai 2026 | |
| OpenAI | OpenAI: GPT-4o Mini Transcribe | 4,47 % | 🔒 Propriétaire | 1 mai 2026 |
| Whisper Large V3 Turbo | n.d. | 🟢 Ouvert | 1 mai 2026 | |
| OpenAI: GPT-4o Transcribe | 3,96 % | 🔒 Propriétaire | 27 avr. 2026 | |
| Whisper Large v2 | 4,06 % | 🔒 Propriétaire | 8 déc. 2022 | |
| Qwen | Qwen3.5 Omni Plus | 3,55 % | ▫ n.d. | 29 mars 2026 |
| Cohere | transcribe-03-2026 | 4,57 % | ▫ n.d. | 26 mars 2026 |
| AssemblyAI | AssemblyAI U3 Realtime Pro - Max Accuracy | 3,02 % | 🔒 Propriétaire | 3 mars 2026 |
| Universal-3 Pro | 3,12 % | 🔒 Propriétaire | 3 févr. 2026 | |
| ElevenLabs | Scribe v2 | 2,18 % | 🔒 Propriétaire | 9 janv. 2026 |
| Replicate | Canary Qwen 2.5B, NVIDIA | 4,27 % | ▫ n.d. | 17 juil. 2025 |
| Rev AI | Rev AI | 5,92 % | 🔒 Propriétaire | 24 oct. 2019 |
| Amazon Bedrock | Amazon Transcribe | 4,12 % | 🔒 Propriétaire | 4 avr. 2018 |
| Speechmatics | Speechmatics Enhanced | 4,05 % | 🔒 Propriétaire | — |