HomeAI & Machine LearningSpeech Recognition & Synthesis Pipeline Simulator

🎙️ Speech Recognition & Synthesis Pipeline Simulator

Watch raw audio turn into a mel spectrogram, flow through transformer encoder layers, and come out as decoded text (ASR, Whisper-style) or resynthesized speech (TTS, ElevenLabs/Bark-style) — live in your browser.

AI & Machine Learning3DModerate60 FPS
speech-recognition-synthesis-pipeline-simulator ↗ Open standalone
⚙ Under the hood

Watch raw audio become a mel spectrogram, flow through transformer encoder layers, and decode into text (ASR, Whisper-style) or resynthesize into speech (TTS, ElevenLabs/Bark-style). Adjust pitch, speech rate and noise and watch confidence and real-time factor change live.

Three.jsspeech synthesisspeech recognitionWhispertransformerspectrogramAI

3D · Three.js / WebGL renderer · 60 FPS target · runs fully client-side, no install

What did you find?

Add reproduction steps (optional)