Skip to content

Saturday, October 10, 2026

Live
Loading the latest AI news…

Glossary · Images, voice and video

Speech recognition

Also called Speech-to-text

Speech recognition turns spoken words into text. Modern systems use deep learning and handle many languages and accents, though accuracy still drops with background noise, strong accents and specialised vocabulary.

In one line, for a 12-year-old

Speech recognition is the AI typing what you say.

An example

Dictating a text message, or live captions on a video call.

Why it matters to people

Captions make meetings and video accessible to people who are deaf or hard of hearing. Recording others' voices still needs their knowledge, and in many settings their consent.