awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 रिपॉजिटरी

Awesome GitHub RepositoriesPhrase-Specific Isolation

Isolating specific objects by matching visual regions to high-similarity scores for individual words within a phrase.

Distinct from Object Detection: Moves beyond general object detection to target specific sub-phrases for precise isolation.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Phrase-Specific Isolation. Refine with filters or upvote what's useful.

Awesome Phrase-Specific Isolation GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • idea-research/groundingdinoIDEA-Research का अवतार

    IDEA-Research/GroundingDINO

    9,738GitHub पर देखें↗

    GroundingDINO is a deep learning vision model and open-vocabulary object detector designed to map natural language prompts to spatial coordinates. It functions as a text-to-bounding-box framework that enables zero-shot image localization, allowing the system to identify and locate arbitrary objects without requiring predefined classes or specific training for those categories. The project distinguishes itself by matching visual features to natural language descriptions to achieve open-set visual recognition. It supports text-guided image localization and the isolation of specific objects base

    Extracts precise object locations by targeting the highest text similarity scores for specific words within a sentence.

    Pythonobject-detectionopen-worldopen-world-detection
    GitHub पर देखें↗9,738
  • facebookresearch/multimodalfacebookresearch का अवतार

    facebookresearch/multimodal

    1,723GitHub पर देखें↗

    Multimodal is a machine learning library built on PyTorch for training large-scale models that combine text, image, audio, and video data streams. It functions as a deep learning framework dedicated to generative diffusion models, multi-task training, and vision-language tasks. The library supplies modular building blocks, discrete latent codebook quantization, shared-space embeddings, and stackable adapter layers to handle diverse conditional inputs during training and inference. The framework supports specific architectures for diffusion models, text-to-video generation, image-text retrieva

    Locates and boxes specific regions in an image corresponding to noun phrases found in text queries.

    Python
    GitHub पर देखें↗1,723
  1. Home
  2. Artificial Intelligence & ML
  3. Computer Vision Systems
  4. Computer Vision
  5. Object Detection and Tracking
  6. Object Detection
  7. Phrase-Specific Isolation