# Translate Audio > To translate audio means turning speech in one language into speech in another. Older tools chain speech recognition, text translation and text-to-speech; newer ones like Meta's SeamlessM4T do it in one model, and some keep the original voice. On Kenerate you upload an audio file — MP3, WAV, M4A, AAC, OGG, FLAC or Opus up to 50 MB — or record in the browser, pick one of 33 languages, and the ElevenLabs dubbing engine returns it spoken in that language in the speaker's voice. The source language is auto-detected from 57. It works on the free plan. URL: https://kenerateai.com/translate-audio Publisher: Kenerate AI (https://kenerateai.com) Last updated: 27 September 2026 ## Key facts - Target languages: 33 - Audio files: Up to 50 MB - Longest audio: 15 min - Free plan: Yes - Translate audio (ElevenLabs): Audio file or recording, up to 15 min → Speech in 33 languages. Voice notes, podcasts, lectures. - Video dubbing (ElevenLabs): Video up to 15 min / 500 MB → Dubbed video. YouTube, courses, ads. - Voice Studio (ElevenLabs, MiniMax, Qwen): Text → Speech in up to 40 languages. Reading a written translation aloud. - Subtitles (VEED): Video up to 4 min → Captions in the spoken language. Reading along. ## Steps 1. Open AI dubbing and choose audio — Sign in — it works on the free plan. 2. Upload or record — MP3, WAV, M4A, AAC, OGG, FLAC, Opus or WebM, up to 50 MB and 15 minutes — or record in the browser. 3. Pick the language — 33 targets; the source is auto-detected. 4. Download — The translated speech, in the speaker's voice. ## In depth ### How speech-to-speech translation works The classic pipeline is three steps: speech recognition, text translation, then text-to-speech. Meta's SeamlessM4T replaces that chain with one model covering about 100 input languages. Voice-preserving tools add a clone of the speaker so the translation still sounds like them. ### When to use which Live conversation with someone in front of you: a phone app's conversation mode. A recording you want to share — a podcast, lecture or voice note — in another language: translate the file and keep the voice. ## FAQ Q: How do I translate audio? A: Upload the file (or record it), pick one of 33 languages, and download the translated speech — the source language is auto-detected. Q: Can it keep my own voice? A: Yes — the ElevenLabs dubbing engine re-voices each speaker with a clone of their voice. Q: Is there a Spanish voice translator? A: Yes — translate any audio into Spanish, or Spanish audio into any of the other 32 languages. Q: Can I translate Chinese audio to English? A: Yes — Chinese is auto-detected as a source and English is a target, and the reverse works too (see the demo). Q: What file types are supported? A: MP3, WAV, M4A, AAC, OGG, FLAC, Opus and WebM audio, up to 50 MB and 15 minutes — or record in the browser. Q: Is this different from Google Translate conversation mode? A: Google's conversation mode speaks translations aloud in a live chat. This tool translates a recording and gives you the file, in the speaker's own voice. Q: Is it free? A: It works on the free plan with sign-up credits; longer audio uses one-time credit packs that never expire. ## Related pages - AI video translator: https://kenerateai.com/ai-video-translator-and-dubbing - AI lip sync: https://kenerateai.com/ai-lipsync-generator - Talking photo AI: https://kenerateai.com/talking-photo-ai - AI video upscaler: https://kenerateai.com/ai-video-upscaler ## Sources - Kenerate video studio: https://kenerateai.com/app/video - Meta AI — SeamlessM4T: https://ai.meta.com/blog/seamless-m4t/ - Google Translate Help — conversations: https://support.google.com/translate/answer/6142474 - ElevenLabs — Dubbing docs: https://elevenlabs.io/docs/capabilities/dubbing