login
quiescence.eu

AI Audio Transcription

Server-side speech-to-text powered by OpenAI Whisper. Transcribe, translate and export subtitles for any audio file.

Upload Audio

🎤 Drop your audio file here

or click to browse — MP3, WAV, FLAC, OGG, AAC, M4A…

Vocal Isolation

Isolate vocals with htdemucs (2-stem) before transcribing — usually improves accuracy on music, but adds processing time and cost. Turn it off to transcribe the original audio directly.

1

Higher = cleaner vocal isolation, slower. 1 (fast) to 5.

Shorter segments use less memory but may introduce more boundary artefacts.

Higher overlap reduces boundary artefacts at the cost of extra processing time.

Text to speech

🤖 faster-whisper-small 🌏 auto-detect 📈 beam = 1 📢 VAD on

Output Format

Plain text transcript — clean, easy to copy.