Python Text to Speech Module

Qwen TTS Ships Local TTS Under Apache 2.0 : Voice Cloning in 3 Seconds

Qwen TTS focuses on on-device processing with no external API; emotion control relies on precise prompts, shaping output ...

IEEE

An Automated Method to Correct Artifacts in Neural Text-to-Speech Models

Abstract: Recent advances in deep learning technology have enabled high-quality speech synthesis, and text-to-speech models are widely used in a variety of applications. However, even state-of-the-art ...

How Large Scale Speech Models Will Impact Voice AI

A duplex speech-to-speech model changes the premise: The intelligence layer consumes audio and produces audio directly. The model can attend to what was said and how it was said—content and delivery ...

Tom Pelphrey Takes on Task of Playing the Most Famous Man to Ever Live — Jesus Christ (Exclusive)

The Emmy nominee segues from ‘Task’ to ‘The Christ,’ a Faith Radio Network podcast starring Pelphrey as Jesus alongside David ...

GitHub

Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching

Small and fast: only 123M parameters. High-quality voice cloning: state-of-the-art performance in speaker similarity, intelligibility, and naturalness. Multi-lingual: support Chinese and English.

KittenTTS Nano AI Small Text-to-Speech LLM Runs on CPUs Without a GPU

KittenTTS brings small text to speech models to edge devices; the Nano 8-bit model is about 25 MB, local playback is possible.

IEEE

Aligning Speech-Text Representations via Contrastive Modality Translation

Abstract: Recent advances in automatic speech recognition (ASR) have led to substantial improvements in system accuracy and robustness, particularly in converting speech signals into text sequences.

GitHub

mcp-tool-shop-org/original_voice-soundboard

mkdir models && cd models curl -LO https://github.com/thewh1teagle/kokoro-onnx/releases/download/model-files-v1.0/kokoro-v1.0.onnx curl -LO https://github.com ...

devdiscourse

BharatGen: India's Sovereign Multilingual AI Engine to Complete Text Modules Soon

The government's BharatGen AI engine is set to complete text-based services in 22 official languages by month-end, with 15 also having speech and vision modules. BharatGen aims to develop foundational ...

Some results have been hidden because they may be inaccessible to you

Show inaccessible results