awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to freewym/espresso

Open-source alternatives to Espresso

30 open-source projects similar to freewym/espresso, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Espresso alternative.

  • mravanelli/pytorch-kaldiAvatar de mravanelli

    mravanelli/pytorch-kaldi

    2,398Voir sur GitHub↗

    pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label computation, and decoding are performed with the kaldi toolkit.

    Pythonasrdeep-learningdeep-neural-networks
    Voir sur GitHub↗2,398
  • speechbrain/speechbrainAvatar de speechbrain

    speechbrain/speechbrain

    11,624Voir sur GitHub↗

    SpeechBrain is an all-in-one deep learning toolkit designed for speech and audio processing. Built as a modular library, it provides a structured environment for developing, training, and deploying neural network models across a wide range of tasks, including automatic speech recognition, speaker identification, and audio enhancement. The framework distinguishes itself through a configuration-driven approach that separates model architecture and training hyperparameters from application logic. By utilizing externalized configuration files and standardized recipes, it enables reproducible rese

    Pythonasraudioaudio-processing
    Voir sur GitHub↗11,624
  • pytorch/audioAvatar de pytorch

    pytorch/audio

    2,886Voir sur GitHub↗

    Data manipulation and transformation for audio signal processing, powered by PyTorch

    Python
    Voir sur GitHub↗2,886
  • awni/speechA

    awni/speech

    0Voir sur GitHub↗
    Voir sur GitHub↗0

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Find more with AI search
  • mindslab-ai/voicefilterAvatar de mindslab-ai

    mindslab-ai/voicefilter

    1,213Voir sur GitHub↗

    Unofficial PyTorch implementation of Google AI's VoiceFilter system

    Python
    Voir sur GitHub↗1,213
  • mozilla/ttsAvatar de mozilla

    mozilla/TTS

    10,151Voir sur GitHub↗

    This project is a comprehensive suite for neural speech synthesis, featuring a deep learning text-to-speech engine, a neural speech synthesis trainer, and a voice cloning toolkit. It provides a system for synthesizing human-like speech from text using neural network models and high-fidelity vocoders. The suite includes a speech model conversion utility to transform deep learning models between different formats for deployment across various hardware runtimes. It also provides a self-contained HTTP server to expose pre-trained text-to-speech models as a remote audio API. Capabilities include

    Jupyter Notebookdataset-analysisdeep-learninggantts
    Voir sur GitHub↗10,151
  • google/uis-rnnAvatar de google

    google/uis-rnn

    1,589Voir sur GitHub↗

    This is the library for the Unbounded Interleaved-State Recurrent Neural Network (UIS-RNN) algorithm, corresponding to the paper Fully Supervised Speaker Diarization.

    Pythonclusteringmachine-learningspeaker-diarization
    Voir sur GitHub↗1,589
  • facebookresearch/loopF

    facebookresearch/loop

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • pyannote/pyannote-audioAvatar de pyannote

    pyannote/pyannote-audio

    9,203Voir sur GitHub↗

    Pyannote.audio is a PyTorch toolkit for speaker diarization, speaker identification, and speech activity detection. Its primary purpose is to partition audio recordings into segments and assign each segment to a specific speaker identity to determine who spoke when. The project includes a framework for classifying speaker identities and a pipeline for distinguishing human speech from background noise. It provides specialized tools for handling symmetric-overlap speech, where multiple speakers talk simultaneously, and employs learnable band-pass filters for raw waveform feature extraction. Th

    Jupyter Notebookoverlapped-speech-detectionpretrained-modelspytorch
    Voir sur GitHub↗9,203
  • nvidia/nemoAvatar de NVIDIA

    NVIDIA/NeMo

    17,394Voir sur GitHub↗

    NeMo is a multimodal AI framework and toolkit designed for the development, training, and scaling of large language models, generative AI systems, and speech-based models. It functions as an automatic speech recognition toolkit, a text-to-speech engine, and a framework for building models that process and generate combinations of text, image, and audio data. The project serves as a conversational AI orchestrator capable of managing real-time, interruptible voice interactions. It provides specialized workflows for speech translation, converting spoken audio from one language into text or speec

    Python
    Voir sur GitHub↗17,394
  • soobinseo/tacotron-pytorchS

    soobinseo/Tacotron-pytorch

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • vincentherrmann/pytorch-wavenetAvatar de vincentherrmann

    vincentherrmann/pytorch-wavenet

    1,025Voir sur GitHub↗

    An implementation of WaveNet with fast generation

    Jupyter Notebook
    Voir sur GitHub↗1,025
  • sanyam5/skip-thoughtsAvatar de sanyam5

    sanyam5/skip-thoughts

    223Voir sur GitHub↗

    The first public PyTorch implementation of Skip-Thought Vectors

    Python
    Voir sur GitHub↗223
  • huggingface/torchmojiAvatar de huggingface

    huggingface/torchMoji

    922Voir sur GitHub↗

    😇A pyTorch implementation of the DeepMoji model: state-of-the-art deep learning model for analyzing sentiment, emotion, sarcasm etc

    Pythondeep-learningmachine-learningnatural-language-processing
    Voir sur GitHub↗922
  • facebookresearch/fairseq-pyF

    facebookresearch/fairseq-py

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • facebookresearch/museAvatar de facebookresearch

    facebookresearch/MUSE

    3,245Voir sur GitHub↗

    A library for Multilingual Unsupervised or Supervised word Embeddings

    Python
    Voir sur GitHub↗3,245
  • eladhoffer/seq2seq.pytorchE

    eladhoffer/seq2seq.pytorch

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • espnet/espnetAvatar de espnet

    espnet/espnet

    9,861Voir sur GitHub↗

    ESPnet is a comprehensive speech processing toolkit and PyTorch-based trainer designed for building end-to-end speech recognition, synthesis, and translation models. It provides a structured framework for developing automatic speech recognition systems using transducer and encoder-decoder architectures, alongside engines for text-to-speech synthesis and speech translation pipelines. The project distinguishes itself through a recipe-based workflow execution system that ensures experimental reproducibility by running standardized sequences of scripts for data preparation and model training. It

    Python
    Voir sur GitHub↗9,861
  • facebookresearch/xlmAvatar de facebookresearch

    facebookresearch/XLM

    2,930Voir sur GitHub↗

    PyTorch original implementation of Cross-lingual Language Model Pretraining.

    Python
    Voir sur GitHub↗2,930
  • alibaba-edu/simple-effective-text-matching-pytorchA

    alibaba-edu/simple-effective-text-matching-pytorch

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • alexsergivan/transliteratorA

    alexsergivan/transliterator

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • abosamoor/polyglotAvatar de aboSamoor

    aboSamoor/polyglot

    2,367Voir sur GitHub↗

    Multilingual text (NLP) processing toolkit

    Python
    Voir sur GitHub↗2,367
  • argilla-io/argillaAvatar de argilla-io

    argilla-io/argilla

    5,015Voir sur GitHub↗

    Argilla is a collaborative AI feedback tool and data curation management system. It serves as a human-in-the-loop dataset platform designed to coordinate workforce annotators and domain experts in labeling, rating, and refining data samples for machine learning projects. The platform focuses on large language model dataset curation and reinforcement learning from human feedback workflows. It provides a shared workspace for integrating human expertise into AI development to validate model outputs and correct data errors. The system manages the end-to-end machine learning data pipeline, includ

    Python
    Voir sur GitHub↗5,015
  • alexrozanski/llamachatAvatar de alexrozanski

    alexrozanski/LlamaChat

    1,510Voir sur GitHub↗

    Chat with your favourite LLaMA models in a native macOS app

    Swiftaillamallamacpp
    Voir sur GitHub↗1,510
  • arc53/docsgptAvatar de arc53

    arc53/DocsGPT

    17,939Voir sur GitHub↗

    DocsGPT is a retrieval-augmented generation platform and private knowledge base used to build AI agents that perform grounded search and analysis. It functions as a multi-model AI orchestrator and enterprise agent builder, allowing for the integration of various local and cloud language models to customize reasoning and text generation. The project provides a visual environment for developing automated assistants using conditional logic and third-party API connectivity. It enables the creation of private AI agents capable of performing enterprise search and detailed document analysis using pr

    Pythonagent-builderagentsai
    Voir sur GitHub↗17,939
  • arongdari/python-topic-modelAvatar de arongdari

    arongdari/python-topic-model

    374Voir sur GitHub↗

    Implementation of various topic models

    Jupyter Notebook
    Voir sur GitHub↗374
  • arongdari/topic-model-lecture-noteAvatar de arongdari

    arongdari/topic-model-lecture-note

    22Voir sur GitHub↗

    lecture notes for probabilistic topic models using ipython notebook

    Voir sur GitHub↗22
  • artidoro/qloraAvatar de artidoro

    artidoro/qlora

    10,929Voir sur GitHub↗

    This project is a quantized fine-tuning framework for large language models. It implements a low-rank adaptation library and a four-bit quantizer to reduce the GPU memory requirements needed to train large models. The framework utilizes four-bit quantization and low-rank adapters to enable model training on consumer-grade hardware. It further reduces the memory footprint through double quantization and a paged optimizer that offloads states to system RAM. The system supports distributed training across multiple GPUs to handle larger parameter scales and includes utilities for custom dataset

    Jupyter Notebook
    Voir sur GitHub↗10,929
  • artificiai/multilingual-latent-dirichlet-allocation-ldaAvatar de ArtificiAI

    ArtificiAI/Multilingual-Latent-Dirichlet-Allocation-LDA

    83Voir sur GitHub↗

    A Multilingual Latent Dirichlet Allocation (LDA) Pipeline with Stop Words Removal, n-gram features, and Inverse Stemming, in Python.

    Pythonclusteringenglishfrench
    Voir sur GitHub↗83
  • anujvyas/natural-language-processing-projectsAvatar de anujvyas

    anujvyas/Natural-Language-Processing-Projects

    254Voir sur GitHub↗

    This repository consists of all my NLP Projects

    Jupyter Notebook
    Voir sur GitHub↗254