awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to wkentaro/labelme

Open-source alternatives to Labelme

30 open-source projects similar to wkentaro/labelme, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Labelme alternative.

  • opencv/cvatopencv avatar

    opencv/cvat

    16,086View on GitHub↗

    CVAT is an open-source computer vision annotation tool and visual dataset management platform. It provides a self-hosted interface for labeling images, videos, and 3D data to create datasets for vision AI models. The platform features AI-assisted data labeling to automate the creation of masks and bounding boxes, utilizing a plug-in system to connect external machine learning models. It includes a consensus-based quality assurance system that verifies label accuracy by comparing independent annotations. The system covers collaborative team management, project organization through task decomp

    Python
    View on GitHub↗16,086
  • tzutalin/labelimgtzutalin avatar

    tzutalin/labelImg

    25,012View on GitHub↗

    labelImg is a desktop image annotation tool and dataset preparation utility used to create labeled datasets for computer vision training. It provides a graphical interface for drawing bounding boxes around objects in images and assigning them class labels to build ground truth data for machine learning models. The software specifically supports the Pascal VOC XML annotation format, exporting image coordinates and class names into standard XML or text structures. It allows users to load predefined class lists from text files to standardize naming across an entire project. Beyond initial label

    Python
    View on GitHub↗25,012
  • cvhub520/x-anylabelingCVHub520 avatar

    CVHub520/X-AnyLabeling

    8,193View on GitHub↗

    X-AnyLabeling is an AI-assisted annotation platform and computer vision labeling tool. It provides an interface for annotating images and videos using polygons and rectangles to create training sets for machine learning models. The project distinguishes itself through the integration of external AI models via a plugin-based inference backend, allowing for automated generation of candidate labels and the execution of specialized tasks like pose estimation and object detection. It also functions as an optical character recognition tool for extracting text and layout information from document im

    Pythonartificial-intelligenceclipcomputer-vision
    View on GitHub↗8,193

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • heartexlabs/label-studioheartexlabs avatar

    heartexlabs/label-studio

    27,626View on GitHub↗

    Label Studio is a multi-type data labeling tool and data annotation workspace designed to prepare datasets for machine learning training. It functions as a cloud-integrated data pipeline that imports raw data from storage, manages the annotation process, and exports labels into standardized formats. The platform features a machine learning model integration framework that connects to external model servers. This enables model-assisted annotation and active learning, allowing the system to perform pre-labeling and refine predictions based on human feedback. The software provides project manag

    TypeScript
    View on GitHub↗27,626
  • labelbox/labelboxL

    Labelbox/Labelbox

    0View on GitHub↗
    View on GitHub↗0
  • cocodataset/cocoapicocodataset avatar

    cocodataset/cocoapi

    6,377View on GitHub↗

    This project is a toolkit and API designed for parsing, manipulating, and visualizing image annotations for computer vision tasks. It provides a programming interface to load and organize Common Objects in Context annotations, specifically for object detection, image segmentation, and keypoint estimation. The library includes tools for converting formatted JSON files into data structures that support the analysis of pixel-level masks and skeletal markers. It enables the visual verification of ground truth accuracy by rendering bounding boxes, segmentation masks, and keypoint markers directly

    Jupyter Notebook
    View on GitHub↗6,377
  • roboflow/supervisionroboflow avatar

    roboflow/supervision

    44,437View on GitHub↗

    Supervision is a computer vision toolset for normalizing model outputs, managing datasets, and visualizing annotations. It provides a framework to convert predictions from various classification and detection models into a standardized data format to ensure interoperability across different computer vision pipelines. The library features a post-processor for filtering, counting, and tracking detected objects across image frames and video streams. It includes capabilities for large image tiling to improve the detection of small objects and tools for assigning persistent identities to objects t

    Pythonclassificationcococomputer-vision
    View on GitHub↗44,437
  • microsoft/vottmicrosoft avatar

    microsoft/VoTT

    4,427View on GitHub↗

    VoTT is a computer vision annotation software and machine learning dataset preparation tool. It is a desktop application designed for drawing bounding boxes and assigning tags to objects in images and videos to create training datasets for object detection models. The application utilizes a cross-platform desktop interface to manage image and video assets. It features a local-first storage integration to handle large media assets directly from the host machine's file system and includes frame-rate controlled video sampling to extract specific images from video streams for labeling. The softw

    TypeScript
    View on GitHub↗4,427
  • microsoft/computervision-recipesmicrosoft avatar

    microsoft/computervision-recipes

    9,866View on GitHub↗

    This project is a collection of educational resources and implementation frameworks providing deep learning model recipes, code samples, and step-by-step guides for computer vision tasks. It organizes complex workflows into modular recipes and implementation guides to facilitate the building of image and video analysis models. The framework focuses on specialized vision capabilities, including an image similarity framework for fast retrieval and re-ranking, human pose estimation, and video action recognition. It also provides specific tools for crowd density estimation and document image clea

    Jupyter Notebookartificial-intelligenceazurecomputer-vision
    View on GitHub↗9,866
  • abreheret/pixelannotationtoolabreheret avatar

    abreheret/PixelAnnotationTool

    1,454View on GitHub↗

    Annotate quickly images.

    C++
    View on GitHub↗1,454
  • cvat-ai/cvatcvat-ai avatar

    cvat-ai/cvat

    15,317View on GitHub↗

    CVAT is an open-source, web-based platform designed for annotating images, videos, and 3D point clouds to create high-quality training datasets for machine learning. It functions as a containerized server that orchestrates the entire lifecycle of computer vision data, from initial task creation and manual labeling to quality assurance and final dataset export. The platform distinguishes itself through deep integration with machine learning models, allowing users to deploy custom AI models as serverless functions for automated object detection, tracking, and skeleton annotation. It supports co

    Pythonannotationannotation-toolannotations
    View on GitHub↗15,317
  • simular-ai/agent-ssimular-ai avatar

    simular-ai/Agent-S

    11,855View on GitHub↗

    Agent-S is a multimodal AI agent and LLM desktop automation framework designed to control operating systems through graphical user interface interactions. It functions as a computer use interface, utilizing vision-language grounding to translate natural language goals into precise screen coordinates and system actions. The project differentiates itself by combining structured accessibility tree inspection with vision-based element localization. It manages cross-application workflows by mapping conceptual descriptions to physical pixels and simulating low-level keyboard and mouse events to mov

    Pythonagent-computer-interfaceai-agentscomputer-automation
    View on GitHub↗11,855
  • carla-simulator/carlacarla-simulator avatar

    carla-simulator/carla

    14,072View on GitHub↗

    CARLA is an autonomous driving simulator and research environment designed for developing and validating self-driving software. It functions as an urban traffic simulator that generates realistic vehicle and pedestrian behavior and as a synthetic sensor data generator producing LiDAR, Radar, and camera data. The platform distinguishes itself through its deep integration with robotics frameworks, specifically providing native connectivity to ROS2 nodes for robotic control and data processing. It supports the training of driving models via imitation and reinforcement learning within a controlle

    C++
    View on GitHub↗14,072
  • phillipi/pix2pixphillipi avatar

    phillipi/pix2pix

    10,644View on GitHub↗

    pix2pix is a framework for image-to-image translation using conditional generative adversarial networks. It functions as a supervised trainer and visual domain mapper designed to learn a mapping between input and output images for style and domain transfer. The system utilizes a U-Net encoder-decoder architecture combined with a PatchGAN local discriminator to enforce high-frequency local consistency. It employs L1 loss regularization to ensure generated outputs remain structurally close to the ground truth. The project covers a broad range of computer vision capabilities, including semantic

    Lua
    View on GitHub↗10,644
  • sillytavern/sillytavernSillyTavern avatar

    SillyTavern/SillyTavern

    29,463View on GitHub↗

    SillyTavern is a comprehensive interface and orchestration platform designed for immersive AI roleplay and interactive chat experiences. It functions as a unified gateway that connects users to a wide array of local and cloud-based large language models, providing a centralized environment to manage complex character personas, narrative context, and model-driven interactions. The platform distinguishes itself through its advanced prompt engineering and automation capabilities. It utilizes a sophisticated macro-based templating engine and vector-database retrieval to dynamically inject lore, c

    JavaScriptaichatllm
    View on GitHub↗29,463
  • flameshot-org/flameshotflameshot-org avatar

    flameshot-org/flameshot

    30,209View on GitHub↗

    This project is a desktop screen capture and annotation utility designed for Linux environments. It provides an interactive graphical overlay that allows users to select specific screen regions, apply visual annotations such as shapes, text, and pixelation, and manage the resulting images through a configurable post-capture pipeline. The application distinguishes itself through deep system integration and automation capabilities. It operates as a persistent background daemon that monitors global hotkeys and supports inter-process communication via a system message bus, enabling users to trigg

    C++capturecross-platformfree-software
    View on GitHub↗30,209
  • jveitchmichaelis/deeplabeljveitchmichaelis avatar

    jveitchmichaelis/deeplabel

    214View on GitHub↗

    A cross-platform desktop image annotation tool for machine learning

    C++
    View on GitHub↗214
  • hitachi-automotive-and-industry-lab/semantic-segmentation-editorHitachi-Automotive-And-Industry-Lab avatar

    Hitachi-Automotive-And-Industry-Lab/semantic-segmentation-editor

    1,961View on GitHub↗

    Web labeling tool for bitmap images and point clouds

    JavaScript
    View on GitHub↗1,961
  • kinhong/openlabelerkinhong avatar

    kinhong/OpenLabeler

    128View on GitHub↗

    OpenLabeler is an open source desktop application for annotating objects for AI appplications

    Java
    View on GitHub↗128
  • bit-bots/imagetaggerbit-bots avatar

    bit-bots/imagetagger

    277View on GitHub↗

    An open source online platform for collaborative image labeling

    HTML
    View on GitHub↗277
  • open-mmlab/mmdetectionopen-mmlab avatar

    open-mmlab/mmdetection

    32,756View on GitHub↗

    This project is a modular research toolkit designed for developing, training, and evaluating deep learning models for object detection, segmentation, and video instance tracking. It provides a flexible training engine that manages complex neural network execution, including distributed training, custom lifecycle hooks, and weight optimization. The framework is built around a hierarchical configuration system that allows users to define architectures, data pipelines, and training hyperparameters through composable, inheritable files. The project distinguishes itself through its highly modular

    Pythoncascade-rcnnconvnextdetr
    View on GitHub↗32,756
  • recogito/annotoriousrecogito avatar

    recogito/annotorious

    846View on GitHub↗

    Add image annotation functionality to any web page with a few lines of JavaScript.

    TypeScript
    View on GitHub↗846
  • catmaid/catmaidcatmaid avatar

    catmaid/CATMAID

    200View on GitHub↗

    Collaborative Annotation Toolkit for Massive Amounts of Image Data

    JavaScript
    View on GitHub↗200
  • alturosdestinations/alturos.imageannotationAlturosDestinations avatar

    AlturosDestinations/Alturos.ImageAnnotation

    70View on GitHub↗

    A collaborative tool for labeling image data for yolo

    C#
    View on GitHub↗70
  • kyamagu/js-segment-annotatorkyamagu avatar

    kyamagu/js-segment-annotator

    526View on GitHub↗

    JS Segment Annotator

    JavaScript
    View on GitHub↗526
  • google/mediapipegoogle avatar

    google/mediapipe

    35,673View on GitHub↗

    MediaPipe is a cross-platform machine learning framework designed for building and deploying pipelines that process live and streaming media. It provides a system for connecting processing components into custom machine learning chains to analyze real-time audio and video streams. The framework includes a suite of pre-trained models for tasks such as hand, face, and pose tracking, along with tools for retraining and customizing these models with specific datasets. It also features a dedicated benchmarker for measuring the execution speed and accuracy of machine learning models directly within

    C++
    View on GitHub↗35,673
  • napari/naparinapari avatar

    napari/napari

    2,674View on GitHub↗

    napari: a fast, interactive, multi-dimensional image viewer for python

    Python
    View on GitHub↗2,674
  • naturalintelligence/imglabNaturalIntelligence avatar

    NaturalIntelligence/imglab

    1,020View on GitHub↗

    To speedup and simplify image labeling/ annotation process with multiple supported formats.

    HTML
    View on GitHub↗1,020
  • jsbroks/coco-annotatorjsbroks avatar

    jsbroks/coco-annotator

    2,276View on GitHub↗

    :pencil2: Web-based image segmentation tool for object detection, localization, and keypoints

    Vue
    View on GitHub↗2,276
  • l3p-cv/lostl3p-cv avatar

    l3p-cv/lost

    577View on GitHub↗

    Label Objects and Save Time (LOST) - Design your own smart Image Annotation process in a web-based environment.

    Python
    View on GitHub↗577