awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

8 مستودعات

Awesome GitHub RepositoriesModular Vision Pipelines

Architectures that decouple image processing, feature detection, and analysis stages into configurable, independent components.

Explore 8 awesome GitHub repositories matching artificial intelligence & ml · Modular Vision Pipelines. Refine with filters or upvote what's useful.

Awesome Modular Vision Pipelines GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • paddlepaddle/paddleocrالصورة الرمزية لـ PaddlePaddle

    PaddlePaddle/PaddleOCR

    82,412عرض على GitHub↗

    PaddleOCR is a comprehensive optical character recognition framework designed for detecting and transcribing text from images and documents into structured, machine-readable formats. It provides a modular computer vision pipeline that decouples image preprocessing, text detection, and character recognition into independent, configurable stages. This architecture supports automated document digitization and multilingual text recognition, capable of identifying text in over one hundred languages across diverse environments ranging from scanned documents to industrial scenes. The framework disti

    Separates image preprocessing, detection, and recognition into independent, swappable components for custom analysis workflows.

    Pythonai4sciencechineseocrdocument-parsing
    عرض على GitHub↗82,412
  • facefusion/facefusionالصورة الرمزية لـ facefusion

    facefusion/facefusion

    28,806عرض على GitHub↗

    Facefusion is a modular framework designed for automated image and video manipulation, specializing in tasks such as face swapping, enhancement, and restoration. It functions as a computer vision processing pipeline that chains independent machine learning modules to perform complex transformations, including facial animation, age modification, and lip synchronization. The system is built to handle both real-time interactive feeds and large-scale batch processing tasks. The platform distinguishes itself through a highly extensible architecture that supports custom processing modules and inter

    Decouples image processing, feature detection, and analysis stages into configurable, independent components.

    Pythonaideep-fakedeepfake
    عرض على GitHub↗28,806
  • paddlepaddle/paddledetectionالصورة الرمزية لـ PaddlePaddle

    PaddlePaddle/PaddleDetection

    14,243عرض على GitHub↗

    PaddleDetection is an object detection framework designed for the end-to-end development, training, and deployment of computer vision models. It provides a comprehensive library of modular neural network architectures and pipelines that support object detection, instance segmentation, and multi-object tracking tasks. The project distinguishes itself through a configuration-driven approach that decouples model components like backbones and heads, allowing for the flexible assembly of custom vision workflows. It incorporates advanced techniques such as anchor-free detection logic, joint detecti

    Decouples model components like backbones and heads into declarative files to enable flexible assembly of custom computer vision workflows.

    Pythonblazefacedeepsortdetr
    عرض على GitHub↗14,243
  • deci-ai/super-gradientsالصورة الرمزية لـ Deci-AI

    Deci-AI/super-gradients

    5,041عرض على GitHub↗

    Super-Gradients هو إطار عمل للرؤية الحاسوبية لـ PyTorch ومكتبة تدريب مصممة لدورة حياة نماذج الرؤية بالكامل. يعمل كمحسن لنماذج تعلم الآلة ومجموعة أدوات نشر لتدريب وضبط النماذج عبر مهام تصنيف الصور، واكتشاف الكائنات، والتجزئة الدلالية، وتقدير الوضع. يوفر المشروع أدوات محددة لتحسين النماذج، بما في ذلك تقطير المعرفة (Knowledge Distillation) وضغط الدقة الرقمية لتقليل متطلبات الذاكرة والحوسبة. كما يتضمن تنفيذ بنية Yolo-NAS لاكتشاف الكائنات عالي الأداء. يغطي إطار العمل سطح قدرات واسع بما في ذلك التدريب الموزع على GPU، وخطوط أنابيب الرؤية المعيارية، وأتمتة عمليات التدريب عبر تكوينات الوصفات المهيكلة. كما يدير تحميل البيانات، وتعزيز الصور، وتصدير الأوزان المدربة إلى تنسيقات عالمية لمسرعات الأجهزة الإنتاجية.

    Implements architectures that decouple image processing, feature detection, and analysis stages into configurable, independent components.

    Jupyter Notebook
    عرض على GitHub↗5,041
  • open-mmlab/mmtrackingالصورة الرمزية لـ open-mmlab

    open-mmlab/mmtracking

    3,881عرض على GitHub↗

    mmtracking is a PyTorch video perception framework designed for training and deploying computer vision models that analyze sequential image data. It provides specialized tools for multi-object tracking, video instance segmentation, and a configuration-driven system for managing deep learning models. The project utilizes a deep learning model registry and a configuration-driven pipeline to swap model backbones and detectors without modifying the core codebase. This modular approach allows for the development of custom perception architectures by combining various components and configurations.

    Implements vision workflows that decouple model components into declarative configuration files for flexible assembly.

    Pythonmulti-object-trackingsingle-object-trackingtracking
    عرض على GitHub↗3,881
  • sharpai/deepcameraالصورة الرمزية لـ SharpAI

    SharpAI/DeepCamera

    2,858عرض على GitHub↗

    DeepCamera is an open-source AI video surveillance and network video recorder platform powered by local vision language models and hardware-accelerated processing. It integrates live feeds from network cameras, webcams, and mobile devices to monitor physical spaces while running local edge vision inference without relying on cloud servers. The platform incorporates privacy-preserving video anonymization that converts raw video frames into abstract depth maps in real time, retaining motion tracking while protecting personal identity. Its modular architecture supports pluggable AI scripts and

    Decouples video analysis stages into independent components using extensible modular pipelines.

    JavaScriptaiai-cameraai-nvr
    عرض على GitHub↗2,858
  • fafa-dl/awesome-backbonesالصورة الرمزية لـ Fafa-DL

    Fafa-DL/Awesome-Backbones

    1,945عرض على GitHub↗

    Awesome-Backbones is a modular deep learning framework designed for the end-to-end lifecycle of computer vision models. It provides an integrated platform for training, benchmarking, and deploying convolutional and transformer-based neural network architectures for image classification tasks. The framework distinguishes itself through a configuration-driven approach to model assembly, allowing users to define backbone, neck, and head components externally. It includes a specialized toolkit for model interpretability, utilizing gradient-based visualization techniques to generate class activati

    Assembles neural network models by dynamically linking backbone, neck, and head components through external configuration files.

    Pythoncnndeep-learningimage-classification
    عرض على GitHub↗1,945
  • vincentqyw/image-matching-webuiالصورة الرمزية لـ Vincentqyw

    Vincentqyw/image-matching-webui

    1,283عرض على GitHub↗

    This project is a web-based platform designed for benchmarking, visualizing, and evaluating computer vision algorithms focused on image feature extraction and matching. It provides a unified interface to compare the performance and accuracy of different models by processing image pairs or live video streams. The system distinguishes itself through a modular architecture that allows users to define custom processing pipelines and register external algorithms via configuration files. It incorporates geometric verification techniques to refine visual data and improve the precision of detected co

    Enables the construction of modular image processing workflows by integrating custom extractors and matching logic.

    Pythonaspanformerdeep-learningfeature-matching
    عرض على GitHub↗1,283
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Frameworks
  5. Computer Vision
  6. Modular Vision Pipelines

استكشف الوسوم الفرعية

  • Configuration-Driven PipelinesVision workflows that decouple model components into declarative configuration files for flexible assembly. **Distinct from Modular Vision Pipelines:** Distinct from Modular Vision Pipelines: focuses on the configuration-driven assembly of components rather than just modularity.