Qwen3-TTS
Voice & SubtitlesText-to-Speech
Open-source TTS from Alibaba Qwen team. Supports expressive streaming speech, voice design, and voice cloning.
Speech recognition, transcription, and offline text-to-speech engines.
9 tools in this category
Text-to-Speech
Open-source TTS from Alibaba Qwen team. Supports expressive streaming speech, voice design, and voice cloning.
Speech to Subtitles
Fast Whisper-based transcription for turning speech into timed subtitles locally.
Aligned Transcription
Whisper transcription with word-level timestamps and speaker diarization support.
Speech Recognition
Automatic speech recognition pack for converting audio into text and subtitles.
Streaming Transcription
Dockerized streaming ASR for real-time speech-to-text on local hardware.
Edge Speech Toolkit
All-in-one edge speech stack: ASR, TTS, VAD, and speaker ID with ONNX runtime.
Offline Speech Synthesis
Fast, lightweight offline text-to-speech engine with many voice models.
Offline Speech Recognition
Open-source offline speech recognition toolkit for 20+ languages on CPU.
Lightweight Offline TTS
Compact ONNX-based TTS model for low-resource offline speech synthesis.