Speech-to-Text
Also called: ASR, automatic speech recognition
Turning spoken audio into written text.
Modern speech-to-text is accurate enough across most accents and noisy conditions to be used unsupervised for drafts. It underpins meeting notes, subtitles, and voice interfaces. Accuracy still drops on domain jargon, overlapping speakers, and under-represented languages.
In practice: A one-hour call transcribed and summarised before you leave the room.
Where this comes up
- AI Course for Virtual Assistants: Enhance Your Skills and Efficiency
- Descript Alternatives in 2026: How to Compare CapCut, Premiere Pro, DaVinci Resolve, and VEED
- ElevenLabs Alternatives in 2026: How to Compare Murf, PlayHT, Speechify Studio, and Descript
- How to Use Copilot in Microsoft Teams: Meetings, Chats, and Recaps
- Otter AI Alternatives in 2026: How to Compare Fireflies, Fathom, Avoma, and Zoom AI Companion
- Will AI Replace Nurses? A Comprehensive Overview