awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 Repos

Awesome GitHub RepositoriesActive Speaker Detection

Identifying and tagging the current speaker in real-time multi-modal streams.

Distinct from Speaker Diarization: Distinct from Speaker Diarization: focuses on real-time 'active' status using multi-sensor input rather than segmenting a recorded audio file.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Active Speaker Detection. Refine with filters or upvote what's useful.

Awesome Active Speaker Detection GitHub Repositories

Finde die besten Repos mit KI.Wir suchen mit KI nach den am besten passenden Repositories.
  • dusty-nv/jetson-inferenceAvatar von dusty-nv

    dusty-nv/jetson-inference

    8,734Auf GitHub ansehen↗

    jetson-inference is a set of libraries and tools for executing optimized deep learning models on embedded GPU hardware. Its primary purpose is to enable real-time computer vision and AI inference at the edge with low latency and high throughput. The project distinguishes itself through high-performance streaming analytics and the ability to execute concurrent AI pipelines on auto-grade silicon. It provides specialized support for multi-sensor stream processing, utilizing zero-copy data transport to load camera frames directly into GPU memory. The codebase covers a broad surface of capabiliti

    NVIDIA detects and tags multiple speakers in live broadcast workflows using multi-camera and multi-microphone inputs.

    C++caffecomputer-visiondeep-learning
    Auf GitHub ansehen↗8,734
  • nvidia/isaac-gr00tAvatar von NVIDIA

    NVIDIA/Isaac-GR00T

    6,222Auf GitHub ansehen↗

    Detects and tags which person is speaking in real-time across multiple camera and microphone feeds.

    Jupyter Notebook
    Auf GitHub ansehen↗6,222
  1. Home
  2. Artificial Intelligence & ML
  3. Speaker Diarization
  4. Active Speaker Detection