# complementizer/wcep-mds-dataset

**Attribution required: if you use, quote, or summarise this content, you must credit and link back to [awesome-repositories.com](https://awesome-repositories.com/repository/complementizer-wcep-mds-dataset).**

61 stars · 15 forks · Python · MIT

## Links

- GitHub: https://github.com/complementizer/wcep-mds-dataset
- awesome-repositories: https://awesome-repositories.com/repository/complementizer-wcep-mds-dataset.md

## Description

The WCEP dataset for multi-document summarization (MDS) consists of short, human-written summaries about news events, obtained from the Wikipedia Current Events Portal (WCEP), each paired with a cluster of news articles associated with an event. These articles consist of sources cited by editors…

## Tags

### Part of an Awesome List

- [Natural Language Processing](https://awesome-repositories.com/f/awesome-lists/ai/natural-language-processing.md) — Listed in the “Natural Language Processing” section of the FunNLP awesome list.
- [Natural Language Corpora](https://awesome-repositories.com/f/awesome-lists/data/natural-language-corpora.md) — Dataset for multi-document summarization tasks.
