Knowledge work with recordings

How to Organize a Growing Collection of Audio Files

To organize audio files, gather them in one place, give each a name that starts with the recording date, keep an untouched master copy separate from the copies you edit, record a few metadata fields in a catalog, and back everything up in more than one location. Then store a transcript beside each recording with the same name, so the words inside your audio become searchable text.

8 min read · Updated

Start by gathering everything in one place

Most audio collections are not disorganized so much as scattered: voice memos on a phone, interviews on a field recorder's memory card, podcast sessions on a laptop, old projects on an external drive in a drawer. Before you design any system, copy everything into one location, even if it is a single messy folder called "inbox".

Copy, do not move, until the new home is backed up. Then deal with duplicates. Many collections hold the same recording three times: the phone original, a copy emailed to yourself and an edited export. Keep the highest-quality original as the master, and decide whether the others are working copies worth keeping or clutter to delete.

Choose a folder structure you will keep using

The right structure is the one you will still follow in two years. Two layouts work for most people:

By year, then project
Archive/2026/client-interviews/, Archive/2026/podcast-s3/. Good when recordings belong to clear projects.
By project, then year
Archive/podcast/2025/, Archive/podcast/2026/. Good for long-running series that span years.
Flat with strong names
One folder per year and nothing deeper. Works surprisingly well when file names carry the detail.

Keep the hierarchy shallow, around two or three levels. Deep trees hide files, and a recording that belongs to two projects forces an arbitrary choice. When in doubt, put more information in the file name and less in the folder path. Names travel with the file when it is copied, emailed or uploaded; folder paths do not.

A file naming convention that sorts itself

A good name answers when, what and which version at a glance, and sorts correctly in any file browser. A widely used pattern:

  • 2026-03-14_podcast-s3e07_guest-lopez_raw.wav
  • 2026-03-14_podcast-s3e07_guest-lopez_edit-v2.wav
  • 2026-03-14_podcast-s3e07_guest-lopez_master.mp3

Rules that keep names useful:

  • Date first, written year-month-day, so alphabetical order is chronological order.
  • Lowercase, no spaces, hyphens inside a part and underscores between parts. Some tools and servers handle spaces and special characters badly.
  • A fixed set of status words such as raw, edit, final, and a version number rather than "final-final-2".
  • No personal details in names that might be shared. A participant code is safer than a full name when the content is sensitive.

Renaming hundreds of files by hand is tedious. Finder on macOS has a built-in batch rename, Windows users often use PowerToys PowerRename, and many audio librarian apps rename from tags. Test on a copy first.

Metadata: what to record about each recording

A file name cannot hold everything. Metadata fills the gap: the facts you will want later but cannot hear in the audio. Keep the list short enough that you actually fill it in.

  • Title and a one-line description of the content
  • Date and place recorded
  • People who speak, or participant codes
  • Project, series or episode
  • Rights and permissions: who owns it, what consent was given, any restrictions on use
  • Technical notes: recorder, microphone, format, sample rate, number of channels
  • Status: raw, edited, published, transcribed, checked

There are two places to keep it. Embedded metadata lives inside the file: ID3 tags in MP3, tags in M4A, Vorbis comments in FLAC and OGG, and broadcast-style chunks in WAV. It travels with the file but support varies between apps, and some editors drop tags on export. A sidecar catalog, usually a spreadsheet with one row per recording and the file name as the key, is easier to sort, filter and share, and it does not depend on any app reading tags correctly. Many people use both: a few basic embedded tags for players, and the full record in the catalog.

Master copies and working copies

The single most protective habit in audio archiving is never editing the original. Keep the file exactly as it came off the recorder as the master, ideally in its original lossless format such as WAV or FLAC, and mark the folder read-only if your system allows. Every edit, trim, noise reduction pass or compressed MP3 for sharing is a working copy made from it.

This matters because processing cannot be undone. A noise reduction pass that sounded fine on laptop speakers may have smeared consonants, and a low-bitrate MP3 has permanently discarded detail. If the master is intact, you can always start again. The trade-offs between compressed and uncompressed formats for transcription and for archives are covered in MP3 versus WAV for transcription.

Some archivists also store a checksum for each master: a short fingerprint calculated from the file's contents. If a later copy produces a different checksum, the file has changed or been corrupted. Many backup and file-management tools can generate and verify checksums for you.

Backups that survive a bad day

An archive with one copy is a collection waiting to be lost. Drives fail, laptops are stolen and cloud accounts get locked. A commonly cited guideline is the 3-2-1 rule: three copies of everything, on at least two different kinds of storage, with one copy kept somewhere else. In practice that might be the working drive, an external drive at home and a cloud backup service.

Backups also need checking. Once or twice a year, restore a few random files from each backup and play them. A backup you have never restored from is a hope, not a plan. If your collection includes digitized tapes, the originals are often irreplaceable, so protect those masters with particular care; capture and digitizing advice is in transcribing cassette tapes.

Transcripts as a text index beside the audio

Names and metadata tell you which recording is which. They do not tell you what was said. A transcript stored beside each recording, with the same base name, turns the content of your archive into text that any desktop search, cloud drive or command-line tool can search:

  • 2026-03-14_podcast-s3e07_guest-lopez_raw.wav
  • 2026-03-14_podcast-s3e07_guest-lopez_transcript.txt
  • 2026-03-14_podcast-s3e07_guest-lopez_timestamped.txt

The plain text file is for reading and searching. The timestamped version tells you where in the audio a passage sits, so a search hit leads straight to the right minute. You do not need to correct every transcript to make it useful as an index; a raw machine transcript is usually good enough to find a topic, as long as you check the audio before quoting. The same approach applied to recorded meetings is in searching meeting recordings.

A journalist's interview drive

A hypothetical freelance reporter has around 400 interview recordings from six years: M4A voice memos from a phone, WAV files from a field recorder and a few MP3s sent by producers. She copies everything to one drive, renames each file to date_outlet_subject_status, adds a spreadsheet catalog with columns for subject, story, consent notes and transcript status, and transcribes recordings as she revisits stories rather than all at once. A year later, when an editor asks about a quote on water pricing from 2023, a search across the transcript files finds three interviews in seconds, and the timestamps show where to listen.

Mistakes to avoid when organizing recordings

  • Designing an elaborate system before gathering the files. Look at what you actually have first.
  • Editing masters in place. Once overwritten, the original is gone.
  • Relying on embedded tags alone. An export from the wrong app can strip them, and a spreadsheet catalog survives that.
  • Inconsistent dates, such as 03-04-26, which reads as March or April depending on who is looking.
  • Keeping sensitive recordings in a folder synced to every device you own. Decide who should have access, and limit sync accordingly.
  • Treating a transcript as a replacement for the audio. Machine transcripts contain errors and miss tone, so the recording stays the source of truth. Specialist archival practice for interviews, including consent and long-term formats, is covered in oral history transcription.

Adding transcripts with mydubly

mydubly produces the transcript layer for an archive. You choose a recording from your device, and it accepts MP3, WAV, M4A, AAC, OGG and FLAC audio as well as common video formats; .opus files are not accepted, so convert those first. Each file can be up to 2 hours long. The audio is decoded in your browser and sent as compressed audio chunks for recognition, and you download a plain transcript.txt, a timestamped transcript with [m:ss] labels, and SRT and VTT subtitle files. The spoken language is detected automatically. Details are on the audio to text page, and the WAV to text page covers the format most field recorders produce.

mydubly is not an archive. It does not store, catalog, search or back up your recordings, and uploaded audio chunks and results are deleted within 30 minutes of a job finishing, so save the transcripts into your archive folder right away. Transcripts have no speaker labels, and only one audio track is used if a file has several. The browser tab must stay open while a job runs.

Transcription costs 1 credit per minute with a 5-credit minimum per file. A 40-minute interview is 40 credits (4¢), and a two-minute voice memo is charged the minimum, 5 credits (0.5¢), which makes very short memos relatively more expensive per minute; consider whether they need transcripts at all.

Next step: one folder, one pattern, one catalog

Pick the naming pattern, create the top-level folders, and set up the catalog spreadsheet with the fields you will genuinely fill in. Move this month's recordings into the system first, back them up, and transcribe the ones you expect to search. Work through older recordings in batches when you have time, starting with the ones you go back to most.

Frequently asked questions

Should I organize audio files by date or by project?

Put the date in every file name either way, because it sorts reliably. Then use folders for whatever you most often browse by: projects for client or series work, years for collections that do not split cleanly into projects. If a recording could belong to two folders, the hierarchy is probably too detailed.

Which format should I keep for master copies?

Keep the original format the recorder produced. If that was an uncompressed or lossless format such as WAV or FLAC, keep it that way; converting an MP3 or M4A to WAV does not restore lost quality, it just makes the file bigger. Make compressed copies for sharing and listening.

Is embedded metadata or a spreadsheet catalog better?

They solve different problems. Embedded tags travel with the file and show up in players, but support is uneven and some apps strip them. A spreadsheet catalog is easier to sort, filter and back up. Many archives keep a few embedded tags and the full record in the catalog.

Do I need to correct transcripts before using them as an index?

Not for finding things. A raw machine transcript usually contains enough correct words to locate a topic, and the timestamp takes you to the audio. Correct a transcript when you are going to quote it, publish it or rely on its exact wording.

How do I handle recordings with sensitive content?

Restrict access to their folders, avoid personal names in file names, record the consent and any restrictions in your catalog, and be careful about which devices sync the archive. Decide on a retention period and delete what you no longer need. Check your organization's policies and local law for anything involving other people's personal information.

What should I do with .opus voice notes from messaging apps?

mydubly does not accept .opus files, so convert them to a supported format such as M4A, MP3 or WAV with an audio converter or a command-line tool like ffmpeg before transcribing. Keep the original .opus file as the master in your archive and treat the converted copy as a working file.