awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

3 个仓库

Awesome GitHub RepositoriesVoice-Driven Interfaces

Capabilities for interacting with software using spoken language input.

Distinct from Speech Recognition: Focuses on the UI interaction layer using voice, rather than the underlying speech-to-text models.

Explore 3 awesome GitHub repositories matching artificial intelligence & ml · Voice-Driven Interfaces. Refine with filters or upvote what's useful.

  1. Home
  2. Artificial Intelligence & ML
  3. Speech Recognition
  4. Voice-Driven Interfaces

Awesome Voice-Driven Interfaces GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • basedhardware/omiBasedHardware 的头像

    BasedHardware/omi

    12,869在 GitHub 上查看↗

    Omi is an open-source wearable AI platform that captures audio and screen data to provide real-time conversational assistance and memory. It integrates a wearable hardware development kit with a vector memory database and large language model capabilities to create a persistent digital record of user interactions. The platform is distinguished by its BLE audio streaming pipeline, which transmits raw audio from wearable hardware for real-time transcription and speaker identification. It utilizes a plugin-based agent tool framework that allows AI assistants to autonomously invoke custom functio

    Enables the extraction of action items and management of calendars and issue trackers through natural language voice input.

    Dartaiappbci
    在 GitHub 上查看↗12,869
  • microsoft/vscode-copilot-chatmicrosoft 的头像

    microsoft/vscode-copilot-chat

    9,493在 GitHub 上查看↗

    This project is an AI-powered IDE extension and LLM coding assistant that provides a conversational interface for generating, refactoring, and debugging code. It functions as an AI agent framework and a Model Context Protocol client, connecting AI models to external data sources and tools to automate complex development tasks. The system is distinguished by its use of autonomous AI agents capable of multi-step task execution, including the ability to read files, modify code, and run terminal commands iteratively. It supports recursive agent orchestration through subagent delegation and employ

    Allows users to trigger a conversational AI interface at the cursor using speech recognition.

    TypeScript
    在 GitHub 上查看↗9,493
  • moonshine-ai/moonshinemoonshine-ai 的头像

    moonshine-ai/moonshine

    8,527在 GitHub 上查看↗

    Moonshine is a complete on-device voice interface toolkit that provides speech recognition, text-to-speech synthesis, phonetic processing, speaker diarization, and intent recognition, all running locally on edge hardware without any cloud dependency. It executes quantized neural networks for speech and language tasks directly on the device, enabling fully offline conversational AI capabilities. The toolkit distinguishes itself by orchestrating multi-turn spoken exchanges through a conversational flow manager that maintains context across interactions and manages branching dialog flows. It inc

    A set of tools for building voice-driven applications with intent recognition, dialog management, and audio processing.

    C++
    在 GitHub 上查看↗8,527