| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Open-source audio intelligence.
Documentation · HuggingFace (Apple · ONNX & LiteRT) · Blog
📖 English · 中文 · 日本語 · 한국어 · Español · Deutsch · Français · हिन्दी · Português · Русский · العربية · Tiếng Việt · Türkçe · ไทย
speech-swift — AI speech models for Apple Silicon. ASR, TTS, speech-to-speech, VAD, diarization, and speech enhancement — all running locally via MLX and CoreML. No cloud, no API keys.
speech-android — On-device speech SDK for Android. ASR, TTS, VAD, and noise cancellation powered by ONNX Runtime with Qualcomm NNAPI acceleration.
speech-core — On-device VAD, streaming STT, TTS, and diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline state machine. Linux, Windows, Android.
speech-studio — Open-source desktop voice-cloning studio for creators. Tauri + Qwen3-TTS on Apple Silicon.
soniqo.audio covers setup, usage, and architecture for all SDKs:
Join our Discord → — questions, support, model requests, and updates.
Integrating on-device speech into your app, need support, or want your model to be supported?
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.
On-device speech SDK for Android — ASR, TTS, VAD, and noise cancellation powered by ONNX Runtime with Qualcomm NNAPI acceleration
Open-source desktop voice-cloning studio for creators — clone a voice, script lines with emotion markers, synthesize on-device. Tauri + VoxCPM2, runs on macOS, Windows, and Linux.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |