---
title: "File formats"
description: "MP3, WAV, M4A, FLAC, AAC, MP4, MOV, AVI and more upload directly, alongside TXT, DOCX and PDF, plus PNG and JPEG images read by OCR."
---

> Documentation Index
> Fetch the complete documentation index at: https://docs.speakai.co/llms.txt
> Use this file to discover all available pages before exploring further.

# File formats

## Supported formats

Speak AI supports most common audio and video formats:

### Audio

- **MP3** - Most common audio format
- **WAV** - Uncompressed audio (highest quality)
- **M4A** - Apple audio format
- **FLAC** - Lossless compressed audio
- **OGG** - Open-source audio format
- **WMA** - Windows Media Audio
- **AAC** - Advanced Audio Coding
- **WEBM** - Web audio format

### Video

- **MP4** - Most common video format
- **MOV** - Apple QuickTime video
- **AVI** - Windows video format
- **MKV** - Matroska video
- **WEBM** - Web video format
- **WMV** - Windows Media Video

### Other

- **Text files** (TXT, CSV) - For text-based analysis
- **URLs** - YouTube links, direct media URLs

## File limits

- **Maximum duration:** Up to 4 hours per file
- **File size:** On the free plan, each file can be up to 2GB. Paid plans can upload larger files. Compressed formats (MP3, M4A) allow longer recordings within size limits.

## Tips for best results

- **MP3 at 128kbps** is the sweet spot for most recordings: small file size with good speech clarity
- **WAV files** give the highest transcription accuracy but are much larger
- If your file is too large, compress the audio bitrate (64kbps still works well for speech)
- For video, Speak extracts the audio track automatically. Video quality doesn't affect transcription accuracy.

## Converting files

If your file is in an unsupported format, you can convert it using free tools:

- [HandBrake](https://handbrake.fr/) for video conversion
- [Audacity](https://www.audacityteam.org/) for audio conversion
- Online converters like CloudConvert or Zamzar

Having trouble with a specific format? Send us a message and we can help.

## Supported audio and video file formats

## Supported formats

Speak AI supports most common audio and video formats:

### Audio

- **MP3** - Most common audio format
- **WAV** - Uncompressed audio (highest quality)
- **M4A** - Apple audio format
- **FLAC** - Lossless compressed audio
- **OGG** - Open-source audio format
- **WMA** - Windows Media Audio
- **AAC** - Advanced Audio Coding
- **WEBM** - Web audio format

### Video

- **MP4** - Most common video format
- **MOV** - Apple QuickTime video
- **AVI** - Windows video format
- **MKV** - Matroska video
- **WEBM** - Web video format
- **WMV** - Windows Media Video

### Other

- **Text files** (TXT, CSV) - For text-based analysis
- **URLs** - YouTube links, direct media URLs

## File limits

- **Maximum duration:** Up to 4 hours per file
- **File size:** On the free plan, each file can be up to 2GB. Paid plans can upload larger files. Compressed formats (MP3, M4A) allow longer recordings within size limits.

## Tips for best results

- **MP3 at 128kbps** is the sweet spot for most recordings: small file size with good speech clarity
- **WAV files** give the highest transcription accuracy but are much larger
- If your file is too large, compress the audio bitrate (64kbps still works well for speech)
- For video, Speak extracts the audio track automatically. Video quality doesn't affect transcription accuracy.

## Converting files

If your file is in an unsupported format, you can convert it using free tools:

- [HandBrake](https://handbrake.fr/) for video conversion
- [Audacity](https://www.audacityteam.org/) for audio conversion
- Online converters like CloudConvert or Zamzar

Having trouble with a specific format? Send us a message and we can help.

Evaluating Speak AI for a team? [Book a demo and see your own recordings analyzed](https://calendly.com/speak-ai/demo?utm_source=docs&utm_campaign=book-demo).

---
Related: [Uploads](/help/uploads/) · [CSV import](/help/uploads/csv-import/)

Source: https://docs.speakai.co/help/uploads/formats/index.mdx
