← all repositories

CLUEbenchmark/CLUENER2020

CLUENER2020 is a Chinese fine-grained named entity recognition benchmark dataset with 10 entity categories.

1.5k stars Python Data ToolingLanguage Models
CLUENER2020
Not currently ranked — collecting fresh signals.
star history

The repository provides a labeled dataset for fine-grained named entity recognition in Chinese text, covering 10 entity types including address, book, company, game, government, movie, name, organization, position, and scene. It includes baseline implementations using pre-trained language models such as BERT, RoBERTa, and ALBERT for sequence labeling tasks. The dataset serves as a benchmark for training and evaluating NER models.

Frequently asked

What is CLUEbenchmark/CLUENER2020?
CLUENER2020 is a Chinese fine-grained named entity recognition benchmark dataset with 10 entity categories.
Is CLUENER2020 open source?
Yes — CLUEbenchmark/CLUENER2020 is an open-source project tracked on heatdrop.
What language is CLUENER2020 written in?
CLUEbenchmark/CLUENER2020 is primarily written in Python.
How popular is CLUENER2020?
CLUEbenchmark/CLUENER2020 has 1.5k stars on GitHub.
Where can I find CLUENER2020?
CLUEbenchmark/CLUENER2020 is on GitHub at https://github.com/CLUEbenchmark/CLUENER2020.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.