Polyglot Voice

You're on a Polyglot Voice page — free tools, examples, and our platform for speech-to-text, translation, and subtitles.

Sign in or create a free account to process full files in the studio (about a minute). The studio supports 98 languages; you can change direction and models anytime after upload.

Extract audio from video

Pull the speech track from a video file so it can be transcribed, translated, archived, edited or repurposed in audio-first workflows.

Useful before transcription, translation and subtitle generation

Helps repurpose interviews, podcasts, lectures and creator media

Makes audio-only workflows easier after video capture

Pricing and limits

Free tier, minute and TTS character packs, team plans. Pay after registration — no hidden subscriptions.

Platform highlights

  • 98+ languages for recognition and translation
  • 5 free browser tools
  • REST API for automation

Try the workflow before paying

From recording to text, translation or voiceover in a few steps

Give visitors a clear path: upload a file or record speech, get text, then translate, dub or export the result. The free start is enough to understand the quality.

AI speechTranslationSubtitles

Quick price estimate

10 min
10,000

Transcription from: 490

Voiceover approx.: 10

This is an estimate: final cost depends on quality mode, pack and tariff.

How it works

  1. Upload audio/video, paste a link or record speech in the browser.
  2. Choose language, quality and the target result: text, translation, subtitles or voiceover.
  3. Get the result in the studio and add minutes, TTS characters or API access when needed.

Example result

Before

Before: a lecture recording, interview or video in another language.

After

After: structured text, translation, subtitles and a base for voiceover or clips.

Why users can trust it

  • Source files are removed from disk about 1 hour after processing.
  • Start for free and pay only for the minutes, attempts or TTS characters you need.
  • Use the web studio for one-off jobs or API access for automation.

What plans include

  • Transcription minutes/attempts
  • TTS characters for voiceover
  • Task history, export, API keys and webhooks

Honest limitations

  • Demo is limited to 60 seconds
  • Quality depends on noise and speech clarity
  • Long files and heavier models require a plan

How it works

Upload a file

Audio, video, link (YouTube, RuTube, VK Video, OK, Mail.ru), or microphone.

Choose a workflow

Transcription, translation, subtitles, or voiceover.

AI processing

Speech recognition and translation to your target language.

Export

TXT, SRT, VTT, ASS, CSV, ZIP, or media file.

Video links (5 platforms)

Paste a link from YouTube, RuTube, VK Video, OK, Mail.ru, or a direct media URL — no device upload required.

YouTubeRuTubeVK VideoOKMail.ru

Supported formats and inputs

MP4MOVWebMM4VYouTube linksRuTube linksVK Video linksOK linksMail.ru linksdirect media URL

Best for

  • video to mp3 workflow for creators
  • speech extraction before transcription
  • podcast repurposing from recorded video
  • lecture audio export for education

FAQ

Why extract audio from video?

It simplifies transcription, translation, podcast repurposing and audio editing workflows.

Can extracted audio be used for subtitles later?

Yes, extracted audio can become the source for speech-to-text and subtitle generation.

When should I extract audio before transcription?

It is useful when you want a lighter file, an audio-only workflow or a simpler input for speech recognition and translation.

Polyglot Voice

You're trying the Polyglot Voice demo — our platform for turning speech into text and translating it.

To transcribe or translate full files beyond this demo, sign in or create a free account (it only takes a minute). The studio supports 98 languages; you can change direction and models anytime after upload.

Related guides

Related pages