awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
NVIDIA avatar

NVIDIA/Model-Optimizer

0
View on GitHub↗

Model Optimizer

Model-Optimizer is a deep learning toolkit and framework dedicated to compressing, pruning, quantizing, and optimizing neural network architectures. It provides methodologies covering weight quantization, model distillation, and speculative decoding for efficient text generation, alongside automated neural architecture search for discovering optimal network structures.

The library implements post-training quantization pipelines that convert high-precision neural network weights into lower-bit formats using calibration data. Additional optimization techniques include teacher-student knowledge distillation, magnitude-based post-training sparsification, and pruning utilities that remove redundant weights and connections. It also supports speculative decoding acceleration using lightweight draft models or auxiliary heads.

For operational workflows, the toolkit includes sparse model checkpointing mechanisms to persist masks and metadata alongside weights, as well as export utilities that serialize compressed models into standard formats for downstream inference frameworks.

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI
nvidia.github.io/Model-Optimizer
↗

Features

  • Deep Learning Optimization - Provides a toolkit for compressing deep neural networks through quantization, pruning, and knowledge distillation to accelerate inference performance.
  • Deep Learning Frameworks - Removes redundant network connections and weights to decrease model size without sacrificing accuracy.
  • Quantization Toolkits - Reduces the numerical precision of model weights to lower memory usage and speed up hardware execution.
  • Model Checkpointing - Persists sparse model checkpoints along with necessary masks and metadata, restoring them onto base network architectures for downstream usage.
  • Model Pruning - Removes redundant weights and connections from deep neural networks to decrease model size and improve inference performance.
  • Model Sparsification - Transforms pre-trained dense neural network models into sparse variants using magnitude-based thresholding or data-driven calibration without retraining.
  • Architecture Quantization Pipelines - Converts high-precision neural network weights into lower-bit formats using calibration data to reduce memory usage and accelerate hardware inference.
  • Weight Quantization - Reduces the precision of neural network weights to lower memory usage and accelerate inference performance on specialized hardware accelerators.
  • Teacher-Student Distillation - Transfers knowledge from a larger teacher model to a smaller student model to maintain accuracy while reducing size and computational cost.
  • Neural Architecture Search - Automates the discovery of optimal network structures to balance execution speed and accuracy on target hardware.
  • Speculative Decoding - Configures draft models and auxiliary heads to propose extra tokens for fast verification during text generation.
  • Model Export Formats - Saves compressed models in standard formats compatible with downstream inference engines and deployment frameworks.
2,975 स्टार्स·455 फोर्क्स·Python·Apache-2.0·11 व्यूज़

स्टार हिस्ट्री

nvidia/model-optimizer के लिए स्टार हिस्ट्री चार्टnvidia/model-optimizer के लिए स्टार हिस्ट्री चार्ट

Model Optimizer के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो Model Optimizer के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • pytorch/torchtunepytorch का अवतार

    pytorch/torchtune

    5,774GitHub पर देखें↗

    Torchtune is a PyTorch-native library for fine-tuning, aligning, and quantizing large language models. It provides a configurable training pipeline orchestrated through YAML recipes, with CLI overrides and component swapping, distributed training via FSDP2, memory optimizations, and parameter-efficient fine-tuning methods like LoRA, DoRA, and QLoRA. The library distinguishes itself through its YAML-driven configuration system that defines all training parameters and instantiates components from config files, with full CLI override capability for any field or component at launch time. It suppo

    Python
    GitHub पर देखें↗5,774
  • deci-ai/super-gradientsDeci-AI का अवतार

    Deci-AI/super-gradients

    5,041GitHub पर देखें↗

    Super-Gradients is a PyTorch computer vision framework and training library designed for the full lifecycle of vision models. It functions as a deep learning model optimizer and a deployment toolkit for training and fine-tuning models across image classification, object detection, semantic segmentation, and pose estimation tasks. The project provides specific tools for model optimization, including teacher-student knowledge distillation and numerical precision compression to reduce memory and computational requirements. It also includes the implementation of the Yolo-NAS architecture for high

    Jupyter Notebook
    GitHub पर देखें↗5,041
  • tencent/pocketflowTencent का अवतार

    Tencent/PocketFlow

    2,914GitHub पर देखें↗

    PocketFlow is an integrated toolkit for deep learning model compression, distributed training, and mobile format optimization. It provides a system for reducing the size and complexity of neural networks to improve inference efficiency, featuring a dedicated engine for knowledge distillation and a mobile model optimizer. The framework differentiates itself through an automated hyperparameter tuning system that uses reinforcement learning and statistical models to determine optimal compression ratios and layer-wise bit allocation. It also includes a distributed training system that utilizes mu

    Pythonautomlcomputer-visiondeep-learning
    GitHub पर देखें↗2,914
  • meta-pytorch/torchtunemeta-pytorch का अवतार

    meta-pytorch/torchtune

    5,774GitHub पर देखें↗

    Torchtune is a PyTorch-native library for fine-tuning, aligning, and quantizing large language models. It provides a config-driven system for instantiating components, orchestrating distributed training, and managing parameter-efficient fine-tuning with quantization support, all through YAML-based configurations and command-line overrides. The library distinguishes itself through its comprehensive post-training workflow orchestration, combining supervised fine-tuning, preference optimization (DPO, PPO, GRPO), knowledge distillation, and quantization-aware training in a single configurable pip

    Python
    GitHub पर देखें↗5,774
Model Optimizer के सभी 30 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

nvidia/model-optimizer क्या करता है?

Model-Optimizer is a deep learning toolkit and framework dedicated to compressing, pruning, quantizing, and optimizing neural network architectures. It provides methodologies covering weight quantization, model distillation, and speculative decoding for efficient text generation, alongside automated neural architecture search for discovering optimal network structures.

nvidia/model-optimizer की मुख्य विशेषताएं क्या हैं?

nvidia/model-optimizer की मुख्य विशेषताएं हैं: Deep Learning Optimization, Deep Learning Frameworks, Quantization Toolkits, Model Checkpointing, Model Pruning, Model Sparsification, Architecture Quantization Pipelines, Weight Quantization।

nvidia/model-optimizer के कुछ ओपन-सोर्स विकल्प क्या हैं?

nvidia/model-optimizer के ओपन-सोर्स विकल्पों में शामिल हैं: pytorch/torchtune — Torchtune is a PyTorch-native library for fine-tuning, aligning, and quantizing large language models. It provides a… deci-ai/super-gradients — Super-Gradients is a PyTorch computer vision framework and training library designed for the full lifecycle of vision… tencent/pocketflow — PocketFlow is an integrated toolkit for deep learning model compression, distributed training, and mobile format… meta-pytorch/torchtune — Torchtune is a PyTorch-native library for fine-tuning, aligning, and quantizing large language models. It provides a… pytorchlightning/pytorch-lightning — PyTorch Lightning is a high-level deep learning framework for PyTorch that automates training loops and removes… timdettmers/bitsandbytes — bitsandbytes is a quantization library for large language models that reduces memory footprints using k-bit…

Model Optimizer को शामिल करने वाली क्यूरेटेड खोजें

चुनिंदा कलेक्शन जहाँ Model Optimizer दिखाई देता है।
  • LLM optimization framework
  • LLM क्वांटाइज़ेशन ऑप्टिमाइज़ेशन टूल्स
  • स्पेक्युलेटिव डिकोडिंग फ्रेमवर्क्स