awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
rany2 avatar

rany2/edge-tts

0
View on GitHub↗
10,041 نجوم·959 تفرعات·Python·other·9 مشاهداتpypi.org/project/edge-tts↗

Edge Tts

edge-tts is a command line interface and text-to-speech engine that converts written text into audio files using the Microsoft Edge online synthesis service. It functions as a client for generating high-quality speech and managing the conversion of text to audio.

The project provides utilities for generating synchronized SRT subtitle files by tracking word and sentence boundaries during synthesis. It also includes a voice profile discovery system to browse a catalog of available synthetic voices based on gender and personality traits.

Users can customize vocal characteristics by adjusting the pitch, rate, and volume of the output. The system supports real-time audio streaming for immediate playback without requiring local file storage, allowing for the simultaneous streaming of synthesized speech and synchronized text captions.

Features

  • Text-to-Speech - Converts written text into high-quality audio files using the Microsoft Edge online synthesis service.
  • Speech Synthesis Gateways - Offloads text-to-speech conversion to a remote cloud engine via a synthesis gateway.
  • Speech Synthesis & TTS - Provides a command line interface for performing text-to-speech synthesis via Microsoft Edge.
  • Websocket Connection Managers - Maintains persistent duplex WebSocket connections for sending text and receiving binary audio streams.
  • Voice Profile Managers - Provides a system for querying catalogs and selecting specific vocal identity profiles.
  • Vocal Characteristic Adjustments - Allows customization of pitch, rate, and volume to create specific vocal characteristics.
  • Voice Discovery Interfaces - Provides a discovery interface to browse a catalog of synthetic voices based on gender and traits.
  • Speech Parameter Configuration - Allows adjustment of voice profiles, playback rates, volume, and pitch.
  • Vocal Tone Customization - Provides tools to modify speaker characteristics and tone to change synthesized audio quality.
  • Timestamped Subtitle Generators - Generates synchronized SRT files by parsing timing metadata for word and sentence boundaries.
  • Audio Playback Engines - Streams synthesized audio for immediate playback without requiring the audio to be saved to disk.
  • Synthesis Playback Clients - Functions as a client for playing synthesized speech in real time without local file storage.
  • Instant Streaming Playback - Streams synthesized audio directly to a media player for instant listening without disk storage.
  • Real-Time Audio Streaming Buffers - Processes raw audio data packets in real time via streaming buffers for immediate playback.
  • Speech Processing - Library for using Microsoft Edge's TTS service.
  • Speech Synthesis - Wrapper for Microsoft Edge's text-to-speech service.

سجل النجوم

مخطط تاريخ النجوم لـ rany2/edge-ttsمخطط تاريخ النجوم لـ rany2/edge-tts

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ Edge Tts

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع Edge Tts.
  • openbmb/voxcpmالصورة الرمزية لـ OpenBMB

    OpenBMB/VoxCPM

    29,985عرض على GitHub↗

    VoxCPM is a multilingual speech synthesis system and text-to-speech inference server. It functions as an AI voice cloning tool and a synthetic voice designer, capable of generating natural speech across global languages and regional dialects using a GPU-accelerated audio generator. The project features a speech model fine-tuning framework that supports both full parameter updates and low-rank adaptation for customizing voice characteristics. It enables high-fidelity voice cloning from reference audio, including cross-lingual voice transfer and acoustic environment mimicry, as well as the crea

    Pythonaudiodeeplearningminicpm
    عرض على GitHub↗29,985
  • hexgrad/kokoroالصورة الرمزية لـ hexgrad

    hexgrad/kokoro

    5,729عرض على GitHub↗

    Kokoro is a lightweight neural text-to-speech engine that converts written text into spoken audio using a compact model designed for fast inference. It supports multiple languages through language-specific grapheme-to-phoneme conversion pipelines, and offers voice profile selection to change the character of the generated speech. The engine provides GPU acceleration on Apple Silicon hardware by setting a single environment variable, enabling faster inference on Mac M-series machines. It also includes pattern-based text segmentation, allowing input text to be split at user-defined delimiters t

    JavaScript
    عرض على GitHub↗5,729
  • k2-fsa/sherpa-onnxالصورة الرمزية لـ k2-fsa

    k2-fsa/sherpa-onnx

    13,017عرض على GitHub↗

    Sherpa-ONNX is an ONNX-based speech processing toolkit that provides a local speech recognition engine, an on-device voice synthesis tool, and a speaker identification framework. It is designed as a cross-platform speech API that enables speech-to-text, text-to-speech, and speaker verification tasks to be executed locally on a device without requiring network access. The project is distinguished by its ability to perform zero-shot voice cloning and speaker diarization on-device. It supports a wide range of hardware accelerations, including GPU and various NPU architectures, and provides a Web

    C++aarch64androidarm32
    عرض على GitHub↗13,017
  • getstream/vision-agentsالصورة الرمزية لـ GetStream

    GetStream/Vision-Agents

    6,029عرض على GitHub↗
    Pythonagentic-aiagentsai
    عرض على GitHub↗6,029
عرض جميع البدائل الـ 30 لـ Edge Tts→

الأسئلة الشائعة

ما هي وظيفة rany2/edge-tts؟

edge-tts is a command line interface and text-to-speech engine that converts written text into audio files using the Microsoft Edge online synthesis service. It functions as a client for generating high-quality speech and managing the conversion of text to audio.

ما هي الميزات الرئيسية لـ rany2/edge-tts؟

الميزات الرئيسية لـ rany2/edge-tts هي: Text-to-Speech, Speech Synthesis Gateways, Speech Synthesis & TTS, Websocket Connection Managers, Voice Profile Managers, Vocal Characteristic Adjustments, Voice Discovery Interfaces, Speech Parameter Configuration.

ما هي البدائل مفتوحة المصدر لـ rany2/edge-tts؟

تشمل البدائل مفتوحة المصدر لـ rany2/edge-tts: openbmb/voxcpm — VoxCPM is a multilingual speech synthesis system and text-to-speech inference server. It functions as an AI voice… hexgrad/kokoro — Kokoro is a lightweight neural text-to-speech engine that converts written text into spoken audio using a compact… k2-fsa/sherpa-onnx — Sherpa-ONNX is an ONNX-based speech processing toolkit that provides a local speech recognition engine, an on-device… getstream/vision-agents. abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… vocodedev/vocode-core — Vocode-core is a framework for building real-time conversational AI voice agents. It serves as a conversational…