Skip to content
#

speech-synthesis-api

Here are 22 public repositories matching this topic...

Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.

  • Updated May 26, 2026
  • Python

Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cloning, dialogue generation, and GPU/CPU execution.

  • Updated Mar 28, 2026
  • Python

A curated list of AI audio generation APIs, SDKs, and tools including text-to-speech, speech synthesis, music generation, voice cloning, sound design, and generative AI platforms. Covers commercial services, open source models with APIs, and production-ready infrastructure for developers building audio applications.

  • Updated Jun 17, 2026

Real-time ASL sign language recognition system that uses computer vision and machine learning to convert hand gestures into text, automatically form words and sentences, and translate them into spoken audio through a desktop application.

  • Updated Jun 22, 2026
  • Python

Improve this page

Add a description, image, and links to the speech-synthesis-api topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the speech-synthesis-api topic, visit your repo's landing page and select "manage topics."

Learn more