Uploading voice files and downloading transcripts

How to upload audio files and download the resulting transcripts.

Skimle converts audio and video recordings into formatted text documents with automatic speaker identification, then analyses them alongside the rest of your project. Upload your interview recordings, focus groups, or other recordings as an Audio or video data source and Skimle transcribes them for you.

Supported formats

You can upload audio files in the following formats:

  • WAV
  • MP3
  • MPEG
  • MP4 / M4A / AAC
  • OGG
  • WebM
  • FLAC

Files can be up to 500 MB each, and you can upload up to 20 files at once.

Adding an audio or video data source

  1. In your project's Data sources section, click New data source and choose Audio or video.
  2. Drag and drop your recordings into the upload area, or click to browse your files.
  3. Your files appear as pending.

At this stage, no credits are used. You can review your uploads before committing.

Confirming transcription

Each pending file shows its duration and the credit cost for transcription. Credits are based on the length of the audio; see pricing for how credits work.

  1. Select the files you want to transcribe.
  2. Review the total credit cost.
  3. Click Confirm to start transcription.

If you do not have enough credits, you will see a warning before confirming. Confirmed files are transcribed in the background — you can close the page and come back later.

The transcript

Skimle produces a formatted transcript for each recording, which becomes a document in your project. It includes:

  • Speaker labels — each speaker is automatically identified (Speaker 1, Speaker 2, etc.).
  • Timestamps — each block of speech is marked with its start and end time.
  • Formatted text — consecutive utterances by the same speaker are grouped into paragraphs for readability.

For example:

Speaker 1 [00:00:12 - 00:01:03]: I think the main challenge we faced was...

Speaker 2 [00:01:05 - 00:01:45]: That's interesting. Could you elaborate on...

Reviewing transcripts

Turn on Review transcripts on the data source if you want to check and correct the automatic transcripts before they are analysed. Confirmed recordings then wait in a Review step where you can fix speaker labels or wording; once you approve them, they continue to analysis. With review off, transcripts flow straight through.

Analysis and privacy

Transcripts are analysed in place, alongside every other document in the project — there is no separate download step. You can still export the finished analysis, including the transcripts, at any time.

Please note that audio files are deleted from the system immediately after transcription is complete. Audio recordings are generally personal biometric data and we do not allow users to download them.

Once your recordings are transcribed, add an analysis to code them into themes and categories.