FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

speech-synthesis · GitHub Topics · GitHub

#

speech-synthesis

Here are 2,059 public repositories matching this topic...

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

  • Updated Aug 16, 2024
  • Python

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

  • Updated Aug 12, 2026
  • Python

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

  • Updated Aug 19, 2026
  • Python

State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.

  • Updated Aug 12, 2024
  • Jupyter Notebook

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

  • Updated Jul 24, 2026
  • Swift

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

  • Updated Aug 12, 2026
  • Python

Build voice agents with open-source models

  • Updated Aug 19, 2026
  • Python

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

  • Updated Jul 13, 2026
  • Python

Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key

  • Updated Mar 22, 2026
  • Python

A fast, local neural text to speech system

  • Updated Aug 26, 2025
  • C++

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.

  • Updated Mar 25, 2026
  • Python

so-vits-svc fork with realtime support, improved interface and more features.

  • Updated Aug 20, 2026
  • Python

EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine

  • Updated Aug 13, 2024
  • Python

VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech

  • Updated Dec 6, 2023
  • Python

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

  • Updated Aug 17, 2026
  • Python

eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.

  • Updated Aug 19, 2026
  • C

StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models

  • Updated Aug 10, 2024
  • Python

Silero Models: pre-trained text-to-speech models made embarrassingly simple

  • Updated Jul 31, 2026
  • Jupyter Notebook

Improve this page

Add a description, image, and links to the speech-synthesis topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the speech-synthesis topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL