Speech recognition
automatic conversion of spoken language into text

Speech recognition (automatic speech recognition (ASR), computer speech recognition, or speech-to-text (STT)) is a sub-field of computational linguistics concerned with methods and technologies that translate spoken language into text or other interpretable forms.
Speech recognition applications include voice user interfaces, where the user speaks to a device, which "listens" and processes the audio. Common voice applications include interpreting commands for calling, call routing, home automation, and aircraft control. These applications are called direct voice input. Productivity applications include searching audio recordings, creating transcripts, and dictation.
Speech recognition can be used to analyse speaker characteristics, such as identifying native language using pronunciation assessment.
Voice recognition (speaker identification) refers to identifying the speaker, rather than speech contents. Recognizing the speaker can simplify the task of translating speech in systems trained on a specific person's voice. It can also be used to authenticate the speaker as part of a security process.
History
Applications for speech recognition developed over many decades, with progress accelerated due to advances in deep learning and the use of big data.
“Speech recognition” enters the record as automatic conversion of spoken language into text. Crown Archives preserves that source wording while asking what Speech, recognition and automatic can confirm, complicate or overturn.
Why this record matters
“Speech recognition” is worth following because a concise public description often conceals a longer documentary argument. Here, Speech, recognition and automatic provides the most credible route into that argument.
Vocabulary and entity names are the principal evidence signals here, because they determine the precision of every later search. The source revision retrieved here is dated Sep 19, 2026. The linked authority identifier is Q189436. The Library of Congress control number is sh85010109. 1 of 1 selected statements include explicit references; 0 carry qualifiers and 0 use preferred rank.
Overview language is designed for orientation and should not be treated as a substitute for the evidence cited beneath it. The lead is largely declarative, so disagreement and counter-evidence require a deliberate search beyond the opening account. Authority statements aid reconciliation but still require their own references, qualifiers and ranks to be checked.
How to read it
Use the entry as an orientation point, then follow its citations and revision history. Names, dates and institutional relationships should be checked against the original record.
- Subject orientation
- Search vocabulary
- Locating named sources
The closest primary source, responsible institution and strongest cited specialist reference.
Three-step research path
- Establish the record: confirm the title “Speech recognition”, its source revision and the description used here.
- Expand the search: follow Speech recognition primary sources, Speech recognition archive and Speech research across catalogues and specialist indexes.
- Test the account: compare the strongest cited source with the responsible institution’s current record and note any disagreement.
Questions for further research
- Which source most directly establishes the central claim about “Speech recognition”?
- What terminology or title could unlock a more precise catalogue search?
- Which institution is responsible for the underlying evidence?
Search terms from this dossier
This entry incorporates text from “Speech recognition” on English Wikipedia. Contributors are listed in the page history. Text is available under the Creative Commons Attribution-ShareAlike 4.0 License. Selected authority identifiers and statements are retrieved from Wikidata under CC0; their references and qualifiers remain part of the verification path.