awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
S

stepfun-ai/Step-Audio2

0
View on GitHub↗

Step Audio2

Features

  • Speech Processing - Advanced speech-to-text and audio processing.
  • Speech Recognition - Multimodal audio processing and recognition model.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI
0 stars·0 forks·9 views

Star history

Star history chart for stepfun-ai/step-audio2Star history chart for stepfun-ai/step-audio2

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

Frequently asked questions

What are the main features of stepfun-ai/step-audio2?

The main features of stepfun-ai/step-audio2 are: Speech Processing, Speech Recognition.

Which projects share features with stepfun-ai/step-audio2?

Projects with overlapping indexed features include: openai/whisper — This project is a speech recognition and translation engine that utilizes a sequence-to-sequence transformer… xzf-thu/mega-asr. facebookresearch/omnilingual-asr — Omnilingual-ASR is a multilingual automatic speech recognition framework and toolkit designed to transcribe audio… kyutai-labs/delayed-streams-modeling — Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework. qwenlm/qwen3-asr. bytedance/megatts3 — MegaTTS3 is a bilingual speech synthesis system that generates natural-sounding speech in Chinese and English,…

Projects sharing features with Step Audio2

These projects share indexed features with Step Audio2. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • kyutai-labs/delayed-streams-modelingkyutai-labs avatar

    kyutai-labs/delayed-streams-modeling

    2,955View on GitHub↗

    Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.

    Python
    View on GitHub↗2,955
  • openai/whisperopenai avatar

    openai/whisper

    102,828View on GitHub↗

    This project is a speech recognition and translation engine that utilizes a sequence-to-sequence transformer architecture to convert audio into text. It is built upon a weakly supervised learning framework, which leverages large-scale, unlabelled audio-transcript data to create generalized speech representations capable of performing simultaneous transcription, language identification, and translation. The system distinguishes itself through a unified multi-task modeling approach that shares token sequences across different objectives, allowing it to handle diverse languages and vocabularies

    Python
    View on GitHub↗102,828
  • facebookresearch/omnilingual-asrfacebookresearch avatar

    facebookresearch/omnilingual-asr

    2,671View on GitHub↗

    Omnilingual-ASR is a multilingual automatic speech recognition framework and toolkit designed to transcribe audio across 1,600 languages. It provides a complete pipeline for converting speech to text, including a toolkit for fine-tuning pre-trained speech models to specific languages or datasets using custom training recipes. The system supports zero-shot speech recognition, allowing the model to predict text in unseen languages without extensive training data. It further enables few-shot language guidance through in-context examples and uses language codes to constrain transcription output t

    Python
    View on GitHub↗2,671
  • qwenlm/qwen3-asrQwenLM avatar

    QwenLM/Qwen3-ASR

    1,603View on GitHub↗
    Python
    View on GitHub↗1,603
Compare all 30 related projects→