30 open-source projects similar to 6/stopwords-json, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Stopwords Json alternative.
VFHQ-downloader is a Python-based utility designed for the easy downloading and processing of videos from the VFHQ dataset.
CelebV-HQ: A Large-Scale Video Facial Attributes Dataset Hao Zhu\, Wayne Wu\, Wentao Zhu, Liming Jiang, Siwei Tang, Li Zhang, Ziwei Liu, and Chen Change Loy In ECCV 2022. (Equal contribution) Demo Video | Project Page | Paper | Annotations
This is a movie review dataset in the Korean language. Reviews were scraped from Naver Movies.
The Replica Dataset is a dataset of high quality reconstructions of a variety of indoor spaces. Each reconstruction has clean dense geometry, high resolution and high dynamic range textures, glass and mirror surface information, planar segmentation as well as semantic class and instance…
This Tutorial contains installation instructions for the packages released along with Ford Multi AV Dataset. For more details please visit the website.
This repository contains State of the Art Language models and Classifier for Hindi language (spoken in Indian sub-continent).
Download the paper from here
pip install medmnist 18x Standardized Datasets for 2D and 3D Biomedical Image Classification
This a summary dataset. You can train abstractive summarization model using this dataset. It contains 3 files i.e. train, test and val. Data is in jsonl format.
أكبر قائمة لمستبعدات الفهرسة العربية على جيت هاب
CropHarvest is an open source remote sensing dataset for agriculture with benchmarks. It collects data from a variety of agricultural land use datasets and remote sensing products.
Alphabetical list of free/public domain datasets with text data for use in Natural Language Processing (NLP). Most stuff here is just raw unstructured text data, if you are looking for annotated corpora or Treebanks refer to the sources at the bottom.
The Matterport3D V1.0 dataset contains data captured throughout 90 properties with a Matterport Pro Camera.
State-of-the-Art Language Modeling and Text Classification in Hindi Language
Github of the FaceForensics dataset
In recent years, millions of people are into social media networking, blogs and review sites, where they express opinions of various entities including movies, restaurants, products etc. These websites provide myriad amount of information that are not only useful to the creators of these…
This repository contains a pytorch implementation for the Interspeech 2024 paper, MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset. MultiTalk generates 3D talking head with enhanced multilingual performance.
Research datasets regularly disappear, change over time, become obsolete or come without a sane implementation to handle the data format reading and processing.