TTSLab
A voice AI agent and TTS lab running in the browser via WebGPU
Details
- External ID
- 47123980
- Source
- HN
- Company
- —
- Product
- TTSLab
- Website domain
- ttslab.dev
- Launched
- Feb. 23, 2026
- Cohort
- —
- Upvotes
- 5
- Upvotes percentile
- 0.10512129380053908
- Tags
- —
- Fetched at
- Sept. 7, 2026, 9:25 p.m.
- Updated at
- Sept. 7, 2026, 9:25 p.m.
Description
I built TTSLab — a free, open-source tool for running text-to-speech and speech-to-text models directly in the browser using WebGPU and WASM.No API keys, no backend, no data leaves your machine.When you open the site, you'll hear it immediately — the landing page auto-generates speech from three different sentences right in your browser, no setup required.You can then try any model yourself: type text, hit generate, hear it instantly. Models download once and get cached locally.The most experimental feature: a fully in-browser Voice Agent. It chains speech-to-text → LLM → text-to-speech, all running locally on your GPU via WebGPU. You can have a spoken conversation with an AI without a single network request.Currently supported models: - TTS: Kokoro 82M, SpeechT5, Piper (VITS) - STT: Whisper Tiny, Whisper BaseOther features: - Side-by-side model comparison - Speed benchmarking on your hardware - Streaming generation for supported modelsSource: https://github.com/MbBrainz/ttslab (MIT)Feedback I'd especially like: 1. How does performance feel on your hardware? 2. What models should I add next? 3. Did the Voice Agent work for you? That's the most experimental part.Built on top of ONNX Runtime Web (https://onnxruntime.ai) and Transformers.js — huge thanks to those communities for making in-browser ML inference possible.
Enrichment
- Theme
- voice AI agents and infrastructure
- Vertical
- Horizontal
- Function
- Agent / copilot
- Audience
- Prosumer
- AI stance
- AI-native
- Project type
- Hobby / open-source project
- Normalized one-liner
- voice ai agent and text-to-speech lab in browser
- Manually corrected
- False
Could you build this?
Yes It wraps open-source ONNX/WebGPU inference libraries (like Transformers.js or ONNX Runtime Web) into a web front-end to run client-side models.
Discussion
3 comments analyzed.
Competitors mentioned: Silero VAD, Whisper (STT), Kokoro/SpeechT5 (TTS)
Concerns raised: Latency not yet optimized, Model download time is slow
Competitors
Other products that read as similar to this one — 170 launches clear the similarity bar, closest 8 shown.
Attention rank: #151 of 171 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 108 days after the earliest competitor.
- TTSForge · ph · 2026-09-29 · 2 upvotes · similarity 0.54
- Audio AI had a wild day · hn · 2026-01-23 · 5 upvotes · similarity 0.51
- Nari Qwen3-TTS and Qwen3-ASR · hn · 2026-09-14 · 90 upvotes · similarity 0.50
- Free Text-to-Speech Tool · hn · 2026-01-31 · 5 upvotes · similarity 0.50
- Three new Kitten TTS models · hn · 2026-03-19 · 561 upvotes · similarity 0.50
- Inflect TTS v2+ONNX, 9M/4M text-to-speech models running in the browser · hn · 2026-07-26 · 7 upvotes · similarity 0.49
- Speechable · hn · 2026-01-03 · 5 upvotes · similarity 0.48
- Lark Text to speech · ph · 2026-09-22 · 1 upvotes · similarity 0.48
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a agent / copilot tool for Agriculture yet.