awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 dépôts

Awesome GitHub RepositoriesPhrase-Specific Isolation

Isolating specific objects by matching visual regions to high-similarity scores for individual words within a phrase.

Distinct from Object Detection: Moves beyond general object detection to target specific sub-phrases for precise isolation.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Phrase-Specific Isolation. Refine with filters or upvote what's useful.

Awesome Phrase-Specific Isolation GitHub Repositories

Trouvez les meilleurs dépôts grâce à l'IA.Nous recherchons les dépôts les plus pertinents grâce à l'IA.
  • idea-research/groundingdinoAvatar de IDEA-Research

    IDEA-Research/GroundingDINO

    9,738Voir sur GitHub↗

    GroundingDINO is a deep learning vision model and open-vocabulary object detector designed to map natural language prompts to spatial coordinates. It functions as a text-to-bounding-box framework that enables zero-shot image localization, allowing the system to identify and locate arbitrary objects without requiring predefined classes or specific training for those categories. The project distinguishes itself by matching visual features to natural language descriptions to achieve open-set visual recognition. It supports text-guided image localization and the isolation of specific objects base

    Extracts precise object locations by targeting the highest text similarity scores for specific words within a sentence.

    Pythonobject-detectionopen-worldopen-world-detection
    Voir sur GitHub↗9,738
  • facebookresearch/multimodalAvatar de facebookresearch

    facebookresearch/multimodal

    1,723Voir sur GitHub↗

    Multimodal is a machine learning library built on PyTorch for training large-scale models that combine text, image, audio, and video data streams. It functions as a deep learning framework dedicated to generative diffusion models, multi-task training, and vision-language tasks. The library supplies modular building blocks, discrete latent codebook quantization, shared-space embeddings, and stackable adapter layers to handle diverse conditional inputs during training and inference. The framework supports specific architectures for diffusion models, text-to-video generation, image-text retrieva

    Locates and boxes specific regions in an image corresponding to noun phrases found in text queries.

    Python
    Voir sur GitHub↗1,723
  1. Home
  2. Artificial Intelligence & ML
  3. Computer Vision Systems
  4. Computer Vision
  5. Object Detection and Tracking
  6. Object Detection
  7. Phrase-Specific Isolation