1 रिपॉजिटरी
Tools for preparing raw text specifically for high-quality audio synthesis.
Distinct from Natural Language Formatters: Focuses on speech-specific expansion of text rather than general natural language formatting or querying.
Explore 1 awesome GitHub repository matching development tools & productivity · Speech Synthesis Formatters. Refine with filters or upvote what's useful.
KittenTTS is a neural text-to-speech engine and text-to-audio synthesis tool that converts written text into spoken audio using lightweight neural network models. It functions as both a speech synthesizer and an audio file generator, producing spoken audio for offline playback. The system includes a text normalization processor that expands numbers and abbreviations into full spoken words to improve the naturalness of the synthesized speech. It supports diverse voice options and provides the ability to adjust playback speed.
Prepares raw text by expanding abbreviations and numbers to ensure high-quality synthesized speech.