awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

8 रिपॉजिटरी

Awesome GitHub RepositoriesModular Vision Pipelines

Architectures that decouple image processing, feature detection, and analysis stages into configurable, independent components.

Explore 8 awesome GitHub repositories matching artificial intelligence & ml · Modular Vision Pipelines. Refine with filters or upvote what's useful.

Awesome Modular Vision Pipelines GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • paddlepaddle/paddleocrPaddlePaddle का अवतार

    PaddlePaddle/PaddleOCR

    82,412GitHub पर देखें↗

    PaddleOCR is a comprehensive optical character recognition framework designed for detecting and transcribing text from images and documents into structured, machine-readable formats. It provides a modular computer vision pipeline that decouples image preprocessing, text detection, and character recognition into independent, configurable stages. This architecture supports automated document digitization and multilingual text recognition, capable of identifying text in over one hundred languages across diverse environments ranging from scanned documents to industrial scenes. The framework disti

    Separates image preprocessing, detection, and recognition into independent, swappable components for custom analysis workflows.

    Pythonai4sciencechineseocrdocument-parsing
    GitHub पर देखें↗82,412
  • facefusion/facefusionfacefusion का अवतार

    facefusion/facefusion

    28,806GitHub पर देखें↗

    Facefusion is a modular framework designed for automated image and video manipulation, specializing in tasks such as face swapping, enhancement, and restoration. It functions as a computer vision processing pipeline that chains independent machine learning modules to perform complex transformations, including facial animation, age modification, and lip synchronization. The system is built to handle both real-time interactive feeds and large-scale batch processing tasks. The platform distinguishes itself through a highly extensible architecture that supports custom processing modules and inter

    Decouples image processing, feature detection, and analysis stages into configurable, independent components.

    Pythonaideep-fakedeepfake
    GitHub पर देखें↗28,806
  • paddlepaddle/paddledetectionPaddlePaddle का अवतार

    PaddlePaddle/PaddleDetection

    14,243GitHub पर देखें↗

    PaddleDetection is an object detection framework designed for the end-to-end development, training, and deployment of computer vision models. It provides a comprehensive library of modular neural network architectures and pipelines that support object detection, instance segmentation, and multi-object tracking tasks. The project distinguishes itself through a configuration-driven approach that decouples model components like backbones and heads, allowing for the flexible assembly of custom vision workflows. It incorporates advanced techniques such as anchor-free detection logic, joint detecti

    Decouples model components like backbones and heads into declarative files to enable flexible assembly of custom computer vision workflows.

    Pythonblazefacedeepsortdetr
    GitHub पर देखें↗14,243
  • deci-ai/super-gradientsDeci-AI का अवतार

    Deci-AI/super-gradients

    5,041GitHub पर देखें↗

    Super-Gradients एक PyTorch कंप्यूटर विजन फ्रेमवर्क और ट्रेनिंग लाइब्रेरी है जिसे विजन मॉडल के पूरे लाइफसाइकिल के लिए डिज़ाइन किया गया है। यह इमेज क्लासिफिकेशन, ऑब्जेक्ट डिटेक्शन, सिमेंटिक सेगमेंटेशन और पोज़ एस्टिमेशन कार्यों में मॉडल को ट्रेन और फाइन-ट्यून करने के लिए एक डीप लर्निंग मॉडल ऑप्टिमाइज़र और डिप्लॉयमेंट टूलकिट के रूप में कार्य करता है। यह प्रोजेक्ट मॉडल ऑप्टिमाइज़ेशन के लिए विशिष्ट टूल प्रदान करता है, जिसमें टीचर-स्टूडेंट नॉलेज डिस्टिलेशन और मेमोरी व कंप्यूटेशनल आवश्यकताओं को कम करने के लिए न्यूमेरिकल प्रिसिजन कम्प्रेशन शामिल है। इसमें उच्च-प्रदर्शन ऑब्जेक्ट डिटेक्शन के लिए Yolo-NAS आर्किटेक्चर का कार्यान्वयन भी शामिल है। फ्रेमवर्क डिस्ट्रीब्यूटेड GPU ट्रेनिंग, मॉड्यूलर विजन पाइपलाइन और स्ट्रक्चर्ड रेसिपी कॉन्फ़िगरेशन के माध्यम से ट्रेनिंग रन के स्वचालन सहित क्षमताओं की एक विस्तृत सतह को कवर करता है।

    Implements architectures that decouple image processing, feature detection, and analysis stages into configurable, independent components.

    Jupyter Notebook
    GitHub पर देखें↗5,041
  • open-mmlab/mmtrackingopen-mmlab का अवतार

    open-mmlab/mmtracking

    3,881GitHub पर देखें↗

    mmtracking is a PyTorch video perception framework designed for training and deploying computer vision models that analyze sequential image data. It provides specialized tools for multi-object tracking, video instance segmentation, and a configuration-driven system for managing deep learning models. The project utilizes a deep learning model registry and a configuration-driven pipeline to swap model backbones and detectors without modifying the core codebase. This modular approach allows for the development of custom perception architectures by combining various components and configurations.

    Implements vision workflows that decouple model components into declarative configuration files for flexible assembly.

    Pythonmulti-object-trackingsingle-object-trackingtracking
    GitHub पर देखें↗3,881
  • sharpai/deepcameraSharpAI का अवतार

    SharpAI/DeepCamera

    2,858GitHub पर देखें↗

    DeepCamera is an open-source AI video surveillance and network video recorder platform powered by local vision language models and hardware-accelerated processing. It integrates live feeds from network cameras, webcams, and mobile devices to monitor physical spaces while running local edge vision inference without relying on cloud servers. The platform incorporates privacy-preserving video anonymization that converts raw video frames into abstract depth maps in real time, retaining motion tracking while protecting personal identity. Its modular architecture supports pluggable AI scripts and

    Decouples video analysis stages into independent components using extensible modular pipelines.

    JavaScriptaiai-cameraai-nvr
    GitHub पर देखें↗2,858
  • fafa-dl/awesome-backbonesFafa-DL का अवतार

    Fafa-DL/Awesome-Backbones

    1,945GitHub पर देखें↗

    Awesome-Backbones is a modular deep learning framework designed for the end-to-end lifecycle of computer vision models. It provides an integrated platform for training, benchmarking, and deploying convolutional and transformer-based neural network architectures for image classification tasks. The framework distinguishes itself through a configuration-driven approach to model assembly, allowing users to define backbone, neck, and head components externally. It includes a specialized toolkit for model interpretability, utilizing gradient-based visualization techniques to generate class activati

    Assembles neural network models by dynamically linking backbone, neck, and head components through external configuration files.

    Pythoncnndeep-learningimage-classification
    GitHub पर देखें↗1,945
  • vincentqyw/image-matching-webuiVincentqyw का अवतार

    Vincentqyw/image-matching-webui

    1,283GitHub पर देखें↗

    This project is a web-based platform designed for benchmarking, visualizing, and evaluating computer vision algorithms focused on image feature extraction and matching. It provides a unified interface to compare the performance and accuracy of different models by processing image pairs or live video streams. The system distinguishes itself through a modular architecture that allows users to define custom processing pipelines and register external algorithms via configuration files. It incorporates geometric verification techniques to refine visual data and improve the precision of detected co

    Enables the construction of modular image processing workflows by integrating custom extractors and matching logic.

    Pythonaspanformerdeep-learningfeature-matching
    GitHub पर देखें↗1,283
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Frameworks
  5. Computer Vision
  6. Modular Vision Pipelines

सब-टैग एक्सप्लोर करें

  • Configuration-Driven PipelinesVision workflows that decouple model components into declarative configuration files for flexible assembly. **Distinct from Modular Vision Pipelines:** Distinct from Modular Vision Pipelines: focuses on the configuration-driven assembly of components rather than just modularity.