How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
Doccano is a collaborative labeling platform and text annotation tool designed to create training data for machine learning. It provides a specialized interface for performing sequence labeling and text classification on natural language datasets. The system functions as a supervised learning dataset manager, allowing multiple users to coordinate within a shared workspace to label datasets for natural language processing tasks. It supports the preparation of raw text data for model training by converting unstructured documents into structured labeled examples. The platform includes capabilit
This project provides a high-resolution face dataset consisting of 70,000 human face images in PNG format. It serves as a curated library of aligned images and facial landmark data designed for generative model training, facial recognition, and image synthesis research. The dataset includes machine-readable metadata that pairs images with precise facial coordinate points, source URLs, and copyright information. This coordinate data enables the transformation of raw photos into a standardized 1024x1024 pixel resolution through landmark-based alignment and cropping. The repository includes aut
UI-TARS is an LLM GUI automation framework and multimodal action grounding system. It functions as a GUI agent orchestrator and cross-platform device controller that uses large language models to interpret graphical interfaces and execute actions across desktop and mobile operating systems. The system translates model-generated coordinates into precise screen positions to interact with visual user interface elements. It employs a multimodal approach to interpret screen layouts and decomposes complex goals into multi-step trajectories through reasoning and error correction. The project provid
This project is a large-scale dataset of hand-drawn sketches, providing millions of timestamped vector drawings and bitmaps for training machine learning models. It serves as a computer vision training corpus and a neural network dataset, consisting of categorized human sketches used to develop image classification and recognition algorithms.
The main features of googlecreativelab/quickdraw-dataset are: Computer Vision Datasets, Vector Drawing Dataset Downloads, Coordinate Normalization Utilities, Hand-Drawn Sketch Datasets, Neural Network Training Datasets, RNN Dataset Access, Sketch-Based Machine Learning, Sketch Classifiers.
Projects with overlapping indexed features include: chakki-works/doccano — Doccano is a collaborative labeling platform and text annotation tool designed to create training data for machine… nvlabs/ffhq-dataset — This project provides a high-resolution face dataset consisting of 70,000 human face images in PNG format. It serves… bytedance/ui-tars — UI-TARS is an LLM GUI automation framework and multimodal action grounding system. It functions as a GUI agent… bupt-ai-cz/llvip. openimages/dataset — This project is a computer vision dataset and image annotation repository designed for training and evaluating machine… zalandoresearch/fashion-mnist — This project is a computer vision benchmark and image classification dataset used to measure and compare the accuracy…