awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

10 रिपॉजिटरी

Awesome GitHub RepositoriesImage Encoder Embedding Extractions

Tools that process images to extract numerical vector representations for use in downstream machine learning tasks.

Explore 10 awesome GitHub repositories matching artificial intelligence & ml · Image Encoder Embedding Extractions. Refine with filters or upvote what's useful.

Awesome Image Encoder Embedding Extractions GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • comfyanonymous/comfyuicomfyanonymous का अवतार

    comfyanonymous/ComfyUI

    117,322GitHub पर देखें↗

    ComfyUI is a modular generative AI workflow orchestrator and node-based GUI for designing and executing complex diffusion model pipelines. It functions as both a visual interface for building generative logic graphs and a programmable backend API that exposes diffusion model operations for external integration. The system distinguishes itself through a graph-based execution model that supports differential workflow execution, re-running only modified nodes to reduce computation. It features dynamic model offloading to manage memory between system RAM and GPU VRAM and utilizes metadata-embedde

    Analyzes input images to use their conceptual elements as inspiration for creating new images.

    Python
    GitHub पर देखें↗117,322
  • facebookresearch/segment-anythingfacebookresearch का अवतार

    facebookresearch/segment-anything

    54,353GitHub पर देखें↗

    This project provides a deep learning architecture designed to identify and isolate distinct objects within images by generating precise pixel-level masks. It functions as a browser-based inference engine, enabling the execution of complex machine learning models directly within web environments without requiring server-side processing. The system distinguishes itself by utilizing hardware-accelerated execution and parallel processing to achieve real-time segmentation speeds. It supports prompt-based mask decoding, allowing users to generate spatial masks by providing specific points or boxes

    Transforms raw image inputs into compact vector embeddings suitable for downstream analysis and predictive tasks.

    Jupyter Notebook
    GitHub पर देखें↗54,353
  • rwightman/pytorch-image-modelsrwightman का अवतार

    rwightman/pytorch-image-models

    36,893GitHub पर देखें↗

    This project is a library of pretrained computer vision architectures and backbones for image classification and feature extraction. It serves as a comprehensive model zoo and collection of standardized image encoders, including ResNet, Vision Transformers, and EfficientNet, for use in visual analysis and as backbones for object detection and image segmentation. The library provides a framework for distributed training and evaluation of image models using advanced data augmentation and optimization scripts. It includes a dedicated toolset for converting trained PyTorch vision models into the

    Provides standardized image encoders that extract numerical vector representations to serve as backbones for detection and segmentation.

    Python
    GitHub पर देखें↗36,893
  • serengil/deepfaceserengil का अवतार

    serengil/deepface

    22,226GitHub पर देखें↗

    Deepface is a comprehensive deep learning library for facial recognition and demographic analysis. It provides a modular pipeline that handles the entire lifecycle of facial processing, including detection, geometric alignment, and the transformation of facial images into high-dimensional numerical vector embeddings for identity verification and similarity comparison. The library distinguishes itself through a model ensemble approach, which combines predictions from multiple pre-trained neural networks to improve classification accuracy and reduce bias. It also integrates advanced security fe

    Extracts multi-dimensional vector representations from facial images for downstream machine learning tasks.

    Pythonage-predictionarcfacedeep-learning
    GitHub पर देखें↗22,226
  • camel-ai/camelcamel-ai का अवतार

    camel-ai/camel

    17,253GitHub पर देखें↗

    This project is a comprehensive framework for building and managing autonomous agent systems. It provides a unified architecture for orchestrating multi-agent societies, where specialized agents collaborate through roleplay to decompose and solve complex tasks. The system integrates language models with external environments, enabling agents to perform real-world actions through a standardized tool-calling abstraction layer. The framework distinguishes itself through its focus on iterative reasoning and data reliability. It employs automated feedback loops to refine agent outputs and self-eva

    Converts visual inputs into numerical vector representations for downstream similarity and classification tasks.

    Pythonagentai-societiesartificial-intelligence
    GitHub पर देखें↗17,253
  • autogluon/autogluonautogluon का अवतार

    autogluon/autogluon

    9,997GitHub पर देखें↗

    AutoGluon is an automated machine learning framework and multimodal library designed to automate the end-to-end pipeline from data preprocessing to high-accuracy model training and validation. It functions as an automated model trainer for tabular, image, text, and time series data, as well as a tool for time series forecasting and foundation model finetuning. The project is distinguished by its ability to jointly process and fuse different data types, allowing for the construction of multimodal neural networks that integrate images, text, and structured tables. It supports zero-shot inferenc

    Converts images into feature vectors to enable the calculation of semantic similarity scores.

    Pythonautogluonautomated-machine-learningautoml
    GitHub पर देखें↗9,997
  • facebookresearch/dinov3facebookresearch का अवतार

    facebookresearch/dinov3

    9,613GitHub पर देखें↗

    This project is a self-supervised vision foundation model based on a vision transformer architecture. It is designed to learn dense visual representations from unlabeled images, serving as a general-purpose backbone for a wide variety of downstream vision tasks. The system is distinguished by its use of self-distillation and masked image modeling to extract semantic and geometric features. It also incorporates an image-text alignment model that maps visual embeddings to textual descriptions, enabling zero-shot image recognition, zero-shot segmentation, and cross-modal retrieval. The project

    Generates vector representations of images using pretrained backbones via standard model loaders.

    Jupyter Notebook
    GitHub पर देखें↗9,613
  • cubiq/comfyui_ipadapter_pluscubiq का अवतार

    cubiq/ComfyUI_IPAdapter_plus

    6,031GitHub पर देखें↗

    ComfyUIIPAdapterplus, ComfyUI के लिए एक नोड-आधारित एक्सटेंशन है जो संदर्भ छवियों का उपयोग करके इमेज जनरेशन को गाइड करने के लिए IPAdapter मॉडल्स को लागू करता है। यह एक इमेज प्रॉम्प्टिंग टूल और एक Stable Diffusion इमेज एडाप्टर के रूप में कार्य करता है, जो संदर्भ फ़ाइलों को शैली, संरचना और विषय पहचान को नियंत्रित करने के लिए विजुअल प्रॉम्प्ट के रूप में कार्य करने की अनुमति देता है। यह प्रोजेक्ट पोर्ट्रेट्स में चेहरे की पहचान और उच्च-निष्ठा सुविधाओं को बनाए रखने के लिए विशेष क्षमताएं प्रदान करता है। यह संदर्भ छवियों से विजुअल विशेषताओं और कलात्मक शैलियों के हस्तांतरण को सक्षम बनाता है, साथ ही नई पीढ़ियों में वस्तुओं की व्यवस्था को गाइड करने के लिए स्थानिक लेआउट के निष्कर्षण को भी सक्षम बनाता है।

    Uses pretrained CLIP vision models to extract numerical embedding representations from reference images.

    Python
    GitHub पर देखें↗6,031
  • idealo/imagededupidealo का अवतार

    idealo/imagededup

    5,642GitHub पर देखें↗

    imagededup एक Python लाइब्रेरी है जिसका उपयोग सटीक और लगभग-डुप्लिकेट छवियों को खोजने के लिए किया जाता है। यह इमेज फिंगरप्रिंट उत्पन्न करने, न्यूरल एम्बेडिंग की गणना करने और डिडुप्लीकेशन प्रक्रियाओं की सटीकता का मूल्यांकन करने के लिए उपयोगिताएं प्रदान करती है। टूल आकार या प्रारूप की परवाह किए बिना दृश्य रूप से समान फाइलों की पहचान करने के लिए परसेप्चुअल हैशिंग का उपयोग करता है और उच्च-सटीकता समानता खोज के लिए छवियों को वैक्टर में एन्कोड करने के लिए डीप लर्निंग मॉडल का उपयोग करता है। इसमें ज्ञात ग्राउंड ट्रुथ डेटासेट के खिलाफ परिणामों की तुलना करके इन प्रक्रियाओं की सटीकता और रिकॉल को मापने के लिए एक सिस्टम शामिल है।

    Uses deep learning models to encode images into vectors for high-accuracy similarity search.

    Python
    GitHub पर देखें↗5,642
  • lightly-ai/lightlylightly-ai का अवतार

    lightly-ai/lightly

    3,684GitHub पर देखें↗

    Lightly is a self-supervised learning framework and computer vision data curation tool designed to manage large image datasets and train models on unlabeled data. It functions as a PyTorch vision library and dataset management SDK, providing tools to convert raw images into high-dimensional vectors for similarity search, visualization, and feature extraction. The project implements a variety of self-supervised architectures, including MoCo, SimCLR, VICReg, Barlow Twins, and masked image modeling. It distinguishes itself by combining these learning frameworks with active learning capabilities,

    Converts raw image datasets into high-dimensional vectors for similarity search and visualization.

    Pythoncomputer-visioncontrastive-learningcontributions-welcome
    GitHub पर देखें↗3,684
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Infrastructure
  5. Domain-Specific Processing Pipelines
  6. Image Encoder Embedding Extractions

सब-टैग एक्सप्लोर करें

  • Concept ExtractionAnalyzing images to extract high-level conceptual elements for use as generative inspiration. **Distinct from Image Encoder Embedding Extractions:** Focuses on conceptual inspiration for new images rather than raw numerical vector extraction