456
Dictionary

Transcription (ASR)

Converting speech audio into text.

1 min readupdated 2026-07-04

/ quick answer

Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise. Converting speech audio into text.

Converting speech audio into text. Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise. In practice: A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain. This dictionary node is part of the Onexial knowledge graph and links to related concepts, workflows and tools below.
Definition
Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise.
Example
A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain.
Related Workflows
/ frequently asked

What is Transcription (ASR)?

Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise.

What is an example of Transcription (ASR)?

A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain.

Why does Transcription (ASR) matter for AI and automation?

Converting speech audio into text. It connects to the workflows, prompts and tool stacks linked on this page, so you can move from definition to execution without leaving Onexial.

/ topics#ai#audio