OPEN SOURCE, OPEN TO EVERYONE

Open source. Open possibilities.

Discover quality open-source projects, submit projects anonymously, and claim and edit your own project.

Human-curated · Discover open source4100discovered

A little curiosity. A world of open source.

THE FIRST COLLECTION
Topic: speech-synthesis清除
voxcpmOpenBMB
ADDED

VoxCPM2 is a tokenizer-free text-to-speech system with a 2B diffusion autoregressive model supporting 30 languages, voice design from text descriptions, controllable voice cloning, and 48kHz audio output.

AI & MLMusicMultimodal AI
speechNVIDIA-NeMo
ADDED

NVIDIA NeMo Speech is an Apache-2.0 PyTorch framework for building, customizing and deploying speech AI models, covering automatic speech recognition, text-to-speech and speech LLMs, with pre-trained checkpoints and Docker/uv install paths.

AI & MLMusicModel runtime
espnetespnet
ADDED

ESPnet is an end-to-end speech processing toolkit built on PyTorch, covering ASR, TTS, speech translation, enhancement, diarization, and more, with reproducible recipes and pretrained models.

AI & MLDeveloper toolsModel runtime
funttsfarfarfun
ADDED

FunTTS is a text-to-speech library with a unified interface that supports seamless switching between various open-source TTS engines, offering features like subtitle generation and multi-role dialogue.

Developer toolsMusicTemplates & starters
speech-to-speechhuggingface
ADDED

A low-latency, modular voice-agent pipeline (VAD -> STT -> LLM -> TTS) that implements the OpenAI Realtime event set over WebSocket and WebRTC.

AI & MLDeveloper toolsAI assistants