awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

37 个仓库

Awesome GitHub RepositoriesSequence Decoders

Components for generating output sequences conditioned on input context.

Distinct from Sequence Decoding Models: Distinct from general sequence decoding models: focuses on the decoder component logic rather than the full model architecture.

Explore 37 awesome GitHub repositories matching artificial intelligence & ml · Sequence Decoders. Refine with filters or upvote what's useful.

Awesome Sequence Decoders GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • azl397985856/leetcodeazl397985856 的头像

    azl397985856/leetcode

    55,758在 GitHub 上查看↗

    This project is a curated educational resource and solution repository for algorithmic challenges, specifically focused on LeetCode problems. It serves as a technical reference for common data structures and algorithmic patterns, providing verified code implementations across multiple programming languages alongside detailed logic and complexity analysis. The repository functions as a comprehensive study guide for competitive programming and technical interview preparation. It includes specialized learning tools such as an Anki flashcard dataset for spaced repetition and a browser extension t

    The project computes the total ways to decode a numeric string into letters using dynamic programming.

    JavaScriptalgoalgorithmalgorithms
    在 GitHub 上查看↗55,758
  • sgl-project/sglangsgl-project 的头像

    sgl-project/sglang

    29,079在 GitHub 上查看↗

    Sglang is a high-performance inference engine and serving system designed for large language and multimodal models. It provides a programmable interface for orchestrating complex generation workflows, enabling developers to coordinate multi-turn dialogues, tool invocations, and reasoning chains through a domain-specific language. The platform is built to support production-scale deployments, offering an OpenAI-compatible API that allows for integration with existing application ecosystems. The system distinguishes itself through a disaggregated architecture that separates compute-intensive pr

    Separates compute-intensive prefill and memory-intensive decoding phases across distinct hardware nodes to maximize throughput.

    Pythonattentionblackwellcuda
    在 GitHub 上查看↗29,079
  • d2l-ai/d2l-end2l-ai 的头像

    d2l-ai/d2l-en

    29,001在 GitHub 上查看↗

    This project is an educational platform and research toolkit designed to teach deep learning through a combination of mathematical theory, visual diagrams, and executable code. It provides a comprehensive environment for building, training, and evaluating neural networks, grounding complex concepts in interactive computational notebooks that allow for hands-on experimentation. The framework distinguishes itself by interleaving theoretical foundations—including linear algebra, calculus, and probability—with practical implementations across multiple industry-standard libraries. It supports flex

    Integrates attention mechanisms into sequence-to-sequence models by dynamically updating context variables.

    Pythonbookcomputer-visiondata-science
    在 GitHub 上查看↗29,001
  • arendst/tasmotaarendst 的头像

    arendst/Tasmota

    24,502在 GitHub 上查看↗

    Tasmota is a universal firmware platform for ESP8266 and ESP32 microcontrollers, designed to provide local control and management of smart home hardware. It functions as an event-driven automation controller that replaces proprietary factory firmware, allowing users to manage relays, sensors, and lighting systems without relying on external cloud services. The system is built on a modular driver architecture that enables dynamic hardware configuration and peripheral support through a web-based management interface. The platform distinguishes itself through a template-driven hardware mapping s

    Translates raw sensor data into structured messages using custom decoder files to simplify integration with local automation systems.

    Carduinoautomationesp32
    在 GitHub 上查看↗24,502
  • microsoft/unilmmicrosoft 的头像

    microsoft/unilm

    22,030在 GitHub 上查看↗

    This project is a comprehensive framework and toolkit for developing, optimizing, and deploying transformer-based models across multimodal, document intelligence, and natural language processing tasks. It provides a unified neural architecture that processes text, vision, audio, and document layout data through a shared set of weights, enabling researchers and developers to build foundational models that align cross-modal representations. The platform distinguishes itself through advanced training and inference strategies designed for large-scale deep learning. It incorporates specialized mec

    Predicts multiple tokens simultaneously during sequence generation to reduce decoding steps.

    Pythonbeitbeit-3bitnet
    在 GitHub 上查看↗22,030
  • deepseek-ai/flashmladeepseek-ai 的头像

    deepseek-ai/FlashMLA

    12,706在 GitHub 上查看↗

    FlashMLA is an LLM attention kernel library and inference acceleration library providing a collection of high-performance CUDA kernels. It implements multi-head latent attention mechanisms designed to reduce memory overhead and increase throughput during the forward and backward passes of large language model inference. The library utilizes quantized cache attention kernels to improve computation efficiency across both sparse and dense token processing. It specifically optimizes the prefill and decoding phases of model inference through these latent attention implementations. The project cov

    Implements distinct computational paths to optimize the transition between prefill and decoding phases.

    C++
    在 GitHub 上查看↗12,706
  • sapientinc/hrmsapientinc 的头像

    sapientinc/HRM

    12,546在 GitHub 上查看↗

    HRM is an automated reasoning engine and language framework designed to execute complex, multi-scale problem solving. It functions as a reinforcement learning agent that continuously updates internal knowledge representations to improve task performance based on incoming data streams. The system distinguishes itself through a hierarchical architecture that coordinates abstract, long-term planning with granular, low-level logic. By integrating evolutionary algorithms and reinforcement learning, the framework refines model parameters and weights over successive generations, ensuring that intern

    Decodes high-dimensional latent representations into structured natural language sequences.

    Pythonbrain-inspired-aideep-learninglarge-language-models
    在 GitHub 上查看↗12,546
  • ludwig-ai/ludwigludwig-ai 的头像

    ludwig-ai/ludwig

    11,717在 GitHub 上查看↗

    Ludwig is a multimodal machine learning platform and low-code framework designed for building, training, and deploying neural networks. It enables the construction of models that process text, images, audio, and tabular data through a unified interface using declarative configuration files rather than custom code. The system features a specialized low-code framework for large language models, supporting supervised fine-tuning, preference alignment, and a constrained decoding tool to force structured data output via logit extraction. It also includes an automated model architecture search to i

    Forces large language models to produce structured data using logit extraction and constrained decoding.

    Pythoncomputer-visiondata-centricdata-science
    在 GitHub 上查看↗11,717
  • idea-research/groundingdinoIDEA-Research 的头像

    IDEA-Research/GroundingDINO

    9,738在 GitHub 上查看↗

    GroundingDINO is a deep learning vision model and open-vocabulary object detector designed to map natural language prompts to spatial coordinates. It functions as a text-to-bounding-box framework that enables zero-shot image localization, allowing the system to identify and locate arbitrary objects without requiring predefined classes or specific training for those categories. The project distinguishes itself by matching visual features to natural language descriptions to achieve open-set visual recognition. It supports text-guided image localization and the isolation of specific objects base

    Employs a sequence decoder to predict spatial coordinates and class labels from multimodal features.

    Pythonobject-detectionopen-worldopen-world-detection
    在 GitHub 上查看↗9,738
  • infrasys-ai/aiinfraInfrasys-AI 的头像

    Infrasys-AI/AIInfra

    7,414在 GitHub 上查看↗

    Implements separation of prefill and decode phases to avoid resource contention.

    Jupyter Notebookaiinfraaisystem
    在 GitHub 上查看↗7,414
  • harvardnlp/annotated-transformerharvardnlp 的头像

    harvardnlp/annotated-transformer

    7,325在 GitHub 上查看↗

    The Annotated Transformer is an educational resource that provides annotated code implementations of the Transformer architecture for sequence-to-sequence tasks, built with PyTorch. It serves as a learning tool for understanding attention mechanisms, multi-head parallel attention, and scaled dot-product attention through executable examples that walk through each component of the model. The project covers the full Transformer pipeline, including stacked encoder-decoder layers with residual connections and layer normalization, sinusoidal positional encoding for order-aware representation, and

    Generates an output sequence token by token using masked self-attention and encoder-decoder attention.

    Jupyter Notebookannotatednotebookpython
    在 GitHub 上查看↗7,325
  • princewen/tensorflow_practiceprincewen 的头像

    princewen/tensorflow_practice

    7,009在 GitHub 上查看↗

    This repository is a collection of practical deep learning implementations and examples built using the TensorFlow framework. It provides a variety of neural network architectures focusing on natural language processing, recommendation systems, reinforcement learning, and time series prediction. The project features a range of specialized models, including sequence-to-sequence and transformer architectures for text processing, and factorization machines for personalized ranking and retrieval. It also includes implementations of reinforcement learning agents using actor-critic and policy gradi

    Implements decoder components for generating output sequences conditioned on input context via attention.

    Python
    在 GitHub 上查看↗7,009
  • lmcache/lmcacheLMCache 的头像

    LMCache/LMCache

    6,909在 GitHub 上查看↗

    LMCache is a distributed key-value cache manager and tiering system designed to accelerate large language model inference. It functions as a tiered storage layer that offloads tensors from GPU memory to CPU RAM, local disks, or remote object stores, enabling the reuse of cached prefixes across different inference sessions and serving engines. The system differentiates itself through a disaggregated prefill-decode model, which separates prompt processing from token generation by transferring caches between distributed compute nodes. It utilizes peer-to-peer orchestration to share and retrieve

    Implements an architecture that separates prompt processing from token generation by transferring KV caches across compute nodes.

    Pythonamdcudafast
    在 GitHub 上查看↗6,909
  • ericlbuehler/mistral.rsEricLBuehler 的头像

    EricLBuehler/mistral.rs

    6,597在 GitHub 上查看↗

    mistral.rs is an inference engine for large language models that runs locally and exposes models behind OpenAI and Anthropic-compatible APIs. It serves as a multi-model serving platform, capable of loading several models in a single server process with per-request routing and on-demand loading and unloading. The engine supports multimodal inference, processing text alongside images, video, audio, and speech inputs, and includes a quantized model deployment runtime that reduces memory use and speeds up inference on consumer hardware. The project distinguishes itself through an agentic tool exe

    Enforces JSON Schema on tool call arguments during decoding to prevent malformed output.

    Rustllmrustuqff
    在 GitHub 上查看↗6,597
  • facebookresearch/wav2letterfacebookresearch 的头像

    facebookresearch/wav2letter

    6,444在 GitHub 上查看↗

    wav2letter is an automatic speech recognition toolkit and deep learning framework designed to convert audio speech signals into written text. It functions as a distributed training system and an inference engine for building and deploying neural network architectures. The system enables the training of large-scale speech models across multiple compute nodes using custom architecture files and structured recipes. It includes an inference engine that allows these trained models to be executed within Python workflows to transform audio sequences into text. The framework covers the full speech r

    Implements sequence decoders to determine the most accurate sequence of words for a given audio input.

    C++
    在 GitHub 上查看↗6,444
  • ai-dynamo/dynamoai-dynamo 的头像

    ai-dynamo/dynamo

    6,112在 GitHub 上查看↗

    Dynamo is a distributed inference orchestration platform designed for large language models. It functions as a system to coordinate prefill and decode phases across GPU nodes, utilizing a multi-backend runtime adapter to connect engines like vLLM and TensorRT-LLM through a unified block-oriented memory interface. An OpenAI-compatible API server provides the frontend for integration with existing tools and clients. The project is distinguished by its disaggregated serving architecture, which separates prompt processing and token generation onto independent GPU pools to optimize throughput and

    Separates prompt processing and token generation onto independent GPU pools to optimize throughput and memory.

    Rust
    在 GitHub 上查看↗6,112
  • neuphonic/neuttsneuphonic 的头像

    neuphonic/neutts

    6,007在 GitHub 上查看↗

    Neutts is a neural text-to-speech engine designed for real-time streaming output on edge devices such as phones and laptops. It supports voice cloning from short audio references, enabling zero-shot reproduction of a target speaker's voice, and can be fine-tuned or retrained from scratch for custom voices and styles. The system distinguishes itself through a decoder-only architecture that halves memory and accelerates generation on constrained hardware, combined with quantized model inference for reduced memory footprint. Its streaming decoder loop interleaves synthesis with playback, deliver

    Loads only the decoder portion of the speech model during inference to minimize memory and computation.

    Python
    在 GitHub 上查看↗6,007
  • google/gemma_pytorchgoogle 的头像

    google/gemma_pytorch

    5,697在 GitHub 上查看↗

    The official PyTorch implementation of Google's Gemma models

    Loads and executes a decoder-only transformer to generate text completions from a prompt on CPU, GPU, or TPU.

    Pythongemmagooglepytorch
    在 GitHub 上查看↗5,697
  • google/seq2seqgoogle 的头像

    google/seq2seq

    5,621在 GitHub 上查看↗

    This is a TensorFlow-based encoder-decoder framework and model library used for mapping input sequences to output sequences. It functions as a deep learning sequence mapper designed to transform sequential data from one domain to another. The library provides tools for implementing sequence-to-sequence modeling across multiple domains, including neural machine translation, automatic text summarization, and image captioning generation. The framework incorporates recurrent neural networks and utilizes attention-based contextualization to weight input sequences. It supports multiple decoding st

    Includes a greedy decoding strategy that selects the highest probability token at each step.

    Pythondeeplearningmachine-translationneural-network
    在 GitHub 上查看↗5,621
  • google-ai-edge/litert-lmgoogle-ai-edge 的头像

    google-ai-edge/LiteRT-LM

    5,619在 GitHub 上查看↗

    LiteRT-LM is a high-performance inference framework designed to execute large language models locally on mobile, desktop, and IoT hardware. It serves as an on-device model runtime that utilizes CPU, GPU, and NPU acceleration to provide low-latency processing. The framework is distinguished by its ability to process text, vision, and audio inputs through a single multi-modal inference engine. It features a local HTTP server that emulates OpenAI-compatible API endpoints and a WebGPU-based runtime for executing models directly within a web browser. To ensure output reliability, it includes a con

    Provides constrained decoding to ensure model outputs follow specific structured formats via logit manipulation.

    C++
    在 GitHub 上查看↗5,619
上一个12下一个
  1. Home
  2. Artificial Intelligence & ML
  3. Sequence Decoding Models
  4. Sequence Decoders

探索子标签

  • Combinatorial Sequence DecodingCalculates the total number of ways to decode a sequence into valid interpretations using state transition equations. **Distinct from Sequence Decoders:** Focuses on counting decoding possibilities via dynamic programming, unlike sequence decoders that generate specific output sequences.
  • Conditional Random Fields3 个子标签Probabilistic graphical models used for sequence labeling by considering transition constraints between labels. **Distinct from Sequence Decoders:** Provides structured sequence decoding via CRFs, distinct from generic sequence decoders.
  • Constrained Decoding1 个子标签Techniques for forcing model outputs into specific structured formats via logit manipulation. **Distinct from Sequence Decoders:** Distinct from general sequence decoding by focusing on the constraint/forcing of specific output formats.
  • Decoder-Only InferenceExecution modes optimized for models that only use a decoder for autoregressive generation. **Distinct from Sequence Decoders:** Focuses on general text generation using decoder-only architectures, whereas the parent is a general component logic.
  • Decoder-Only Inference ModesInference configurations that load only the decoder portion of a speech model to reduce memory and computation. **Distinct from Prefill-Decode Disaggregation:** Distinct from Prefill-Decode Disaggregation: focuses on using only the decoder part, not separating prefill and decode phases.
  • Greedy Decoding StrategiesDecoding methods that select the most likely token at each step to minimize computational cost. **Distinct from Sequence Decoders:** Distinct from general sequence decoders: focuses on the greedy selection strategy specifically.
  • LoRaWan Decoders1 个子标签Translates raw LoRaWan sensor payloads into structured data formats. **Distinct from Sequence Decoders:** Distinct from general sequence decoders: focuses on protocol-specific payload translation for IoT sensors.
  • Prefill-Decode Disaggregation1 个子标签Separation of compute-intensive prefill and memory-intensive decoding phases into distinct engine instances. **Distinct from Sequence Decoders:** Distinct from general sequence decoders: focuses on the architectural separation of inference phases.