StemVocalCredits

Watch. Read. Refine.

Meet your transcription studio

Built for the work after transcription

A transcript is most useful when you can check it. StemVocal pairs automatic speech recognition with the original recording, word timestamps, a text editor and practical exports. You can listen to a phrase, correct it, rename a speaker and save a document or subtitle file without uploading the recording to a second service.

One StemVocal account

Our audio studio, video studio and music separator use the same credit wallet and daily free allowance. Sign in with the same email on each site. Browser sessions are separate for each domain; signing in on a new domain does not create a new wallet.

Models and limitations

Transcription uses the MIT-licensed Whisper large-v3-turbo model through faster-whisper. Optional speaker grouping uses SpeechBrain ECAPA speaker embeddings under Apache 2.0. These tools estimate words and speaker changes. They do not verify the facts in a recording or identify people. Check any important transcript against the audio.

The example recording is LibriSpeech excerpt 1272-128104-0000, prepared by Vassil Panayotov, Daniel Povey and collaborators from LibriVox readings. LibriSpeech is available under CC BY 4.0. We converted the audio to MP3, added a simple visual for the video example, and generated its editable transcript.

Contact

For help with a file, credit balance or purchase, email help@acapellaextractor.com. Describe the problem without attaching sensitive recordings. We will explain what information is needed to investigate.