563
Dictionary

Transcription (ASR)

Converting speech audio into text.

1 min readupdated 2026-07-04

/ quick answer

Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise. Converting speech audio into text.

Converting speech audio into text. Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise. In practice: A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain. This dictionary node is part of the Onexial knowledge graph and links to related concepts, workflows and tools below.
Definition
Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise.
Example
A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain.
Related Workflows
/ frequently asked

What is Transcription (ASR)?

Automatic Speech Recognition (ASR) turns audio into a timestamped transcript. Whisper, Deepgram, and AssemblyAI are the common choices; accuracy varies by accent, jargon, and noise.

What is an example of Transcription (ASR)?

A podcast pipeline transcribes each episode with Whisper, then feeds the text into a repurposing chain.

Why does Transcription (ASR) matter for AI and automation?

Converting speech audio into text. It connects to the workflows, prompts and tool stacks linked on this page, so you can move from definition to execution without leaving Onexial.

/ topics#ai#audio

/ continue exploring

Related concepts

The vocabulary this page depends on.

  • Automatic Speech Recognition (ASR)

    Automatic Speech Recognition (ASR) is a technology that converts spoken language into written text, acting as a core component for voice assistants, dictation software, and transcription services. It enables machines to understand human speech.

  • Text-to-Speech (TTS)

    Generating natural-sounding audio from text.

all dictionary

Related workflows

Turn this into a repeatable process.

all workflows

Related tool stacks

The tools that run it in production.

all tool stacks

Comparisons & alternatives

Pick between the options.

all comparisons

Long-form guides on this topic