CLUEbenchmark/CLUENER2020
CLUENER2020 is a Chinese fine-grained named entity recognition benchmark dataset with 10 entity categories.

Not currently ranked — collecting fresh signals.
star history
The repository provides a labeled dataset for fine-grained named entity recognition in Chinese text, covering 10 entity types including address, book, company, game, government, movie, name, organization, position, and scene. It includes baseline implementations using pre-trained language models such as BERT, RoBERTa, and ALBERT for sequence labeling tasks. The dataset serves as a benchmark for training and evaluating NER models.
Frequently asked
- What is CLUEbenchmark/CLUENER2020?
- CLUENER2020 is a Chinese fine-grained named entity recognition benchmark dataset with 10 entity categories.
- Is CLUENER2020 open source?
- Yes — CLUEbenchmark/CLUENER2020 is an open-source project tracked on heatdrop.
- What language is CLUENER2020 written in?
- CLUEbenchmark/CLUENER2020 is primarily written in Python.
- How popular is CLUENER2020?
- CLUEbenchmark/CLUENER2020 has 1.5k stars on GitHub.
- Where can I find CLUENER2020?
- CLUEbenchmark/CLUENER2020 is on GitHub at https://github.com/CLUEbenchmark/CLUENER2020.