awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 Repos

Awesome GitHub RepositoriesEncoder-Decoder Model Integrations

Connectors for integrating encoder-decoder architectures like T5 into workflows.

Distinct from Language Model Integrations: Focuses specifically on encoder-decoder (T5) architectures rather than generic language model adapters.

Explore 5 awesome GitHub repositories matching artificial intelligence & ml · Encoder-Decoder Model Integrations. Refine with filters or upvote what's useful.

Awesome Encoder-Decoder Model Integrations GitHub Repositories

Finde die besten Repos mit KI.Wir suchen mit KI nach den am besten passenden Repositories.
  • ml-explore/mlx-examplesAvatar von ml-explore

    ml-explore/mlx-examples

    8,254Auf GitHub ansehen↗

    This repository provides a collection of reference implementations and code examples for training and deploying machine learning models using the MLX framework. It serves as a practical guide for executing distributed training, fine-tuning large language models, converting model weights, and implementing multimodal generative workflows. The project distinguishes itself through specialized examples for local hardware execution, featuring weight quantization to reduce memory usage and low-rank adaptation for parameter-efficient fine-tuning. It also includes scripts for transforming external mod

    Implements T5 model integration using task-specific prefixes for various natural language tasks.

    Pythonmlx
    Auf GitHub ansehen↗8,254
  • datawhalechina/so-large-lmAvatar von datawhalechina

    datawhalechina/so-large-lm

    7,400Auf GitHub ansehen↗

    This project is a comprehensive educational curriculum and structured learning path covering the full lifecycle of large language models. It provides a guided progression through the theory, architecture, training, and deployment of these models. The curriculum includes specialized guides on transformer architecture, model training tutorials, and frameworks for designing autonomous agents. It also provides dedicated resources for studying model safety and ethics. The material covers a wide range of technical capabilities, including distributed training strategies, parameter-efficient fine-tu

    Explains training methods for sequence-to-sequence encoder-decoder architectures.

    Auf GitHub ansehen↗7,400
  • google/gemma_pytorchAvatar von google

    google/gemma_pytorch

    5,697Auf GitHub ansehen↗

    The official PyTorch implementation of Google's Gemma models

    Processes input through separate encoder and decoder stages to produce outputs requiring deep contextual understanding.

    Pythongemmagooglepytorch
    Auf GitHub ansehen↗5,697
  • bentrevett/pytorch-seq2seqAvatar von bentrevett

    bentrevett/pytorch-seq2seq

    5,697Auf GitHub ansehen↗

    This is a collection of educational Jupyter Notebook tutorials that teach sequence-to-sequence modeling using PyTorch and TorchText, focused on neural machine translation. The project provides hands-on guides for building and training encoder-decoder architectures with recurrent neural networks like LSTM and GRU, implementing attention mechanisms that allow the decoder to focus on relevant input tokens during sequence generation. The tutorials cover the full pipeline of machine translation, from tokenizing multilingual text using language-specific tokenizers to training multi-layer encoder-de

    Teaches building and training multi-layer LSTM/GRU encoder-decoder architectures for machine translation.

    Jupyter Notebookattentioncnn-seq2seqencoder-decoder
    Auf GitHub ansehen↗5,697
  • qwenlm/qwen2.5-omniAvatar von QwenLM

    QwenLM/Qwen2.5-Omni

    4,026Auf GitHub ansehen↗

    Qwen2.5-Omni ist ein multimodales Large Language Model für Omnichannel-Anwendungen, das Inhalte über Text, Audio, Bild und Video verarbeiten und generieren kann. Es fungiert als Echtzeit-Sprach-KI und nutzt eine End-to-End-Architektur, um synchrone Sprachkonversationen mit geringer Latenz zu ermöglichen. Das Projekt betont Effizienz durch quantisierte Edge-Modelle, die eine lokale Inferenz auf mobiler Hardware und ressourcenbeschränkten Geräten ermöglichen. Es verwendet 4-Bit-Gewichtungsquantisierung, CPU-basiertes Process-Offloading und On-Demand-Gewichtungsladung, um den GPU-Speicherbedarf zu senken. Das System integriert spezialisierte Encoder zur Analyse multimodaler Datenströme und verfügt über einen Streaming-Decoder für die Echtzeit-Sprachgenerierung. Es enthält zudem Funktionen zur Anpassung der Sprachausgabe, um die tonalen und geschlechtsspezifischen Eigenschaften des Audiosignals zu modifizieren.

    Integrates specialized encoders to convert raw audio and visual streams into high-level conceptual representations.

    Jupyter Notebook
    Auf GitHub ansehen↗4,026
  1. Home
  2. Artificial Intelligence & ML
  3. Artificial Intelligence Tooling
  4. Encoder-Decoder Model Integrations

Unter-Tags erkunden

  • Encoder-Decoder Training MethodsTraining procedures for sequence-to-sequence models that encode input and decode output sequences. **Distinct from Encoder-Decoder Model Integrations:** Distinct from Encoder-Decoder Model Integrations: focuses on the training process itself, not integration connectors.
  • Multimodal EncodersSpecialized components that convert raw audio and visual signals into conceptual representations for a central model. **Distinct from Encoder-Decoder Model Integrations:** More specific than general encoder-decoder integrations; focuses on multimodal signal conversion.