awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectAboutHow we rankPressMCP server
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
X

xzf-thu/Mega-ASR

0
View on GitHub↗
0 stars·0 forks·1 view

Mega ASR

Features

  • Speech Processing - High-performance automatic speech recognition.
  • Speech Recognition - Large-scale speech recognition model.

Star history

Star history chart for xzf-thu/mega-asrStar history chart for xzf-thu/mega-asr

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What are the main features of xzf-thu/mega-asr?

The main features of xzf-thu/mega-asr are: Speech Processing, Speech Recognition.

What are some open-source alternatives to xzf-thu/mega-asr?

Open-source alternatives to xzf-thu/mega-asr include: openai/whisper — This project is a speech recognition and translation engine that utilizes a sequence-to-sequence transformer… stepfun-ai/step-audio2. facebookresearch/omnilingual-asr — Omnilingual-ASR is a multilingual automatic speech recognition framework and toolkit designed to transcribe audio… kyutai-labs/delayed-streams-modeling — Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework. qwenlm/qwen3-asr. bytedance/megatts3 — MegaTTS3 is a bilingual speech synthesis system that generates natural-sounding speech in Chinese and English,…

Open-source alternatives to Mega ASR

Similar open-source projects, ranked by how many features they share with Mega ASR.
  • kyutai-labs/delayed-streams-modelingkyutai-labs avatar

    kyutai-labs/delayed-streams-modeling

    2,955View on GitHub↗

    Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.

    Python
    View on GitHub↗2,955
  • openai/whisperopenai avatar

    openai/whisper

    102,828View on GitHub↗

    This project is a speech recognition and translation engine that utilizes a sequence-to-sequence transformer architecture to convert audio into text. It is built upon a weakly supervised learning framework, which leverages large-scale, unlabelled audio-transcript data to create generalized speech representations capable of performing simultaneous transcription, language identification, and translation. The system distinguishes itself through a unified multi-task modeling approach that shares token sequences across different objectives, allowing it to handle diverse languages and vocabularies

    Python
    View on GitHub↗102,828
  • facebookresearch/omnilingual-asrfacebookresearch avatar

    facebookresearch/omnilingual-asr

    2,671View on GitHub↗

    Omnilingual-ASR is a multilingual automatic speech recognition framework and toolkit designed to transcribe audio across 1,600 languages. It provides a complete pipeline for converting speech to text, including a toolkit for fine-tuning pre-trained speech models to specific languages or datasets using custom training recipes. The system supports zero-shot speech recognition, allowing the model to predict text in unseen languages without extensive training data. It further enables few-shot language guidance through in-context examples and uses language codes to constrain transcription output t

    Python
    View on GitHub↗2,671
  • qwenlm/qwen3-asrQwenLM avatar

    QwenLM/Qwen3-ASR

    1,603View on GitHub↗
    Python
    View on GitHub↗1,603
See all 30 alternatives to Mega ASR→