mydubly

mydubly Blog

Explainers on the technology behind translating, dubbing and transcribing video and audio, written to be useful whichever tool you use.

About this collection

Articles are grouped by topic. Each one explains how the technology works, what affects quality, its honest limitations and, where relevant, how mydubly handles it.

How the blog is organised

The articles are grouped into topics, from how the technology works to the jobs people use it for. Roughly, the topics fall into four bands. The first explains the technology: AI video translation, speech recognition, machine translation, AI voices, and how mydubly itself is built. The second is about quality: translation quality in depth, difficult recordings, audio engineering for speech, video engineering and media formats. The third is about doing the work: preparing files, workflows, repurposing recordings, troubleshooting, subtitles and captions in depth, and accessibility. The fourth covers who uses it and why: creators, education and research, business video, language learning, localization and international content strategy.

Each article owns one question and links to its neighbours rather than repeating them, so following links from any article is a reasonable way to read a topic.

Where to start

How articles treat mydubly

Most articles explain a subject in general terms first, the way a textbook chapter would, and only then say where mydubly fits. When the tool doesn't do something an article discusses, such as cloning a voice, matching lip movements, labelling speakers or translating text in the picture, the article says so plainly and describes what people do instead. Prices are always quoted from the published per-minute rates, and articles avoid invented statistics: where a number can't be sourced, you'll find a method for measuring it on your own files instead.

New articles are added to an existing topic rather than as standalone posts, so a topic's section on this page stays the best entry point for that subject.

Articles, guides and tool pages

Articles explain why things work the way they do and what affects the result. When you're ready to do a specific task, the guides walk through it step by step, and the tool pages, such as the video translator or audio to text, are where the work actually happens.

AI video translation

What AI video translation is, how the pipeline works, and what affects speed and accuracy.

Speech recognition & transcription

How speech-to-text models such as Whisper turn audio into timestamped text, and where they struggle.

AI translation

How neural machine translation works, how quality is measured, and how to fix common errors.

AI voice & dubbing

Text-to-speech, voice generation and dubbing: how synthetic voices are made and what makes them sound natural.

Video localization

Planning, adapting and measuring video for audiences in other languages and cultures.

Creators, YouTube & podcasts

Workflows for creators who want one recording to reach audiences in several languages.

Education & research

Transcripts, captions and translation for teaching, studying and research.

Privacy & architecture

Where your media goes during processing, and why keeping video on the device matters.

Comparisons

Side-by-side explanations of terms and approaches that are easy to confuse.

How it's built

The engineering behind long-file processing, timing alignment and in-browser video assembly.

Translation quality in depth

The hard cases for translated video: idioms, register, gender, numbers, lyrics, and how to check a translation you can't read.

Difficult recordings

Fast, quiet, clipped, windy, phone-line and archival audio: what each problem does to transcription and how to deal with it.

Preparing your files

Recording, extracting, splitting, converting and checking media before you transcribe or translate it.

Workflows

Step-by-step ways to get from a recording to the finished deliverable, from editing SRT files to multi-language releases.

Repurposing recordings

Turning transcripts of videos, meetings and interviews into articles, newsletters, chapters, clips and help content.

Language learning

Methods for learning languages with transcripts, podcasts, dictation and listening practice.

Business video

Translating support, sales, safety, healthcare and research videos, and choosing who does the work.

Audio & video formats

Codecs, containers, sample rates, bitrates, channels and loudness, explained for people working with speech.

Troubleshooting

Common problems with sync, subtitles, transcripts, AI voices and media files: causes, diagnosis and fixes.

Industry & future

Where video translation, dubbing and accessibility technology came from and where it is heading.

Accessibility & inclusive video

Making videos usable for viewers who are deaf, hard of hearing, blind, low vision or neurodivergent, from captions to players and description.

Subtitles & captions in depth

Reading speed, timing rules, segmentation, positioning, formats and quality control for subtitle and caption files.

Audio engineering for speech

Bit depth, dynamics, room acoustics, EQ and microphone technique for speech that will be transcribed, translated or dubbed.

Video engineering & media pipelines

Demuxing, timestamps, keyframes, frame rates, bitrates and decoding: how video files and processing pipelines work.

Knowledge work with recordings

Organizing, searching, summarizing and checking information captured in meetings, interviews and video archives.

International content strategy

Governance, team structure, journeys, search and website delivery for content published in several languages.

Research & academic workflows

Transcription conventions, qualitative analysis, fieldwork, data management and ethics for research recordings.

Creator content operations

Assets, procedures, stems, licensing and hand-offs that keep a multilingual channel running.