awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

4 个仓库

Awesome GitHub RepositoriesSpeech Recognition Toggles

Controls for enabling or disabling the automatic speech recognition process during a session.

Distinct from Automatic Speech Recognition: Focuses on the control logic (on/off) of the ASR system rather than the transcription model itself.

Explore 4 awesome GitHub repositories matching artificial intelligence & ml · Speech Recognition Toggles. Refine with filters or upvote what's useful.

Awesome Speech Recognition Toggles GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • duixcom/duix-mobileduixcom 的头像

    duixcom/Duix-Mobile

    8,093在 GitHub 上查看↗

    Duix-Mobile is a software development kit for deploying real-time conversational AI characters on mobile devices. It enables the creation of interactive digital humans capable of fluid voice-to-voice interactions, featuring low-latency speech recognition and synchronized lip movements. The project distinguishes itself through the ability to integrate custom external language models and speech providers to define an avatar's intelligence and voice. It supports the generation of real-time multilingual subtitles and provides mechanisms to track the training status of newly created digital charac

    Toggles the ability to capture and process user voice input via automated speech recognition.

    C++ai-avatarsai-boyfriendai-companion
    在 GitHub 上查看↗8,093
  • talater/annyangTalAter 的头像

    TalAter/annyang

    6,814在 GitHub 上查看↗

    Annyang is a speech recognition library and web speech API wrapper that enables the integration of voice command interfaces into websites. It functions as a browser-based voice controller, mapping spoken phrases and regular expressions to specific JavaScript functions to trigger application actions. The library provides mechanisms for voice command mapping and simulation, allowing developers to associate spoken text with executable callbacks. It includes tools for command variable extraction using regular expression capture groups, which allows specific words from a spoken phrase to be passed

    Provides controls to start, pause, resume, and stop the speech recognition engine and microphone access.

    TypeScript
    在 GitHub 上查看↗6,814
  • julius-speech/juliusjulius-speech 的头像

    julius-speech/julius

    1,927在 GitHub 上查看↗

    Julius is a high-performance, open-source speech recognition engine designed for large vocabulary continuous speech recognition. It functions as a comprehensive framework utilizing Hidden Markov Model-based acoustic modeling and N-gram language models to convert live or recorded audio into text. The engine is built to support real-time streaming and provides a network-accessible service that allows external applications to manage recognition sessions and receive transcription results through programmatic commands. The engine distinguishes itself through its modular architecture and support fo

    Allows pausing, resuming, or terminating recognition tasks to manage engine activity and handle incoming audio streams.

    Caudio-processingrecognitionspeech
    在 GitHub 上查看↗1,927
  • jamsch/expo-speech-recognitionjamsch 的头像

    jamsch/expo-speech-recognition

    541在 GitHub 上查看↗

    Expo Speech Recognition is a cross-platform mobile module that converts live microphone audio and pre-recorded files into text using native speech engines. It provides offline speech recognition capabilities by downloading and verifying local speech models to enable on-device processing without an active network connection. The library includes session lifecycle management to start, stop, or abort recording, alongside real-time spoken language detection with confidence scoring. It emits volume change events for metering interfaces, handles audio session configuration and routing, and persist

    Controls active speech recognition session lifecycles by starting, stopping, and dispatching state events.

    TypeScriptexporeact-nativespeech-recognition
    在 GitHub 上查看↗541
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Speech Processing
  5. Automatic Speech Recognition
  6. Speech Recognition Toggles