Skip to main content
Home

Transcribing Audio and Video Files with Automated Speech Recognition on Galaxy

Audio and media files are a rich source in the social sciences and the humanities. But how can you make your audio and visual material accessible for structured analysis?
You will need to transcribe the media content into machine-readable text first. This tutorial shows how you can do this by using the data analysis platform Galaxy - all from within your Browser. The platform contains several tools for Automatic Speech Recognition (ASR). From uploading and converting to suitable file types to transcriptions and post-processing, Galaxy has you covered.

If you are new to Galaxy, we recommend you take a look at our Introduction to Galaxy first.

Learning outcomes:

After completing this training module, learners will be able to:

  • Use WhisperX in Galaxy to transcribe their media to machine-readable text
  • Perform text-cleaning
  • Use Regular Expressions (RegEx) to extract meaningful passages
Domain
Social Sciences and Humanities
Language
English
Published to DARIAH-Campus
08/10/2026
Originally published
08/04/2026
License
CC BY 4.0
Sources
DARIAH

Cite as

Daniela Schneider (2026). Transcribing Audio and Video Files with Automated Speech Recognition on Galaxy. Version 1.0.0. Galaxy Training Network [Training module]. https://gxy.io/GTN:T00577

Reuse conditions

Resources hosted on DARIAH-Campus are subjects to the DARIAH-Campus Training Materials Reuse Charter.