← all repositories

yaoguangluo/Deta_Parser

Chinese word segmentation library processing over 16 million characters per second using HMM and neural network algorithms.

478 stars Java Data ToolingLanguage Models
Deta_Parser
Not currently ranked — collecting fresh signals.
star history

This is a high-speed Chinese text segmentation library for NLP tasks. It implements word segmentation, part-of-speech tagging (POS), and text mining using Hidden Markov Models (HMM) and neural network approaches. The project claims peak performance of 16.3 million Chinese characters per second and supports multi-language mixed text processing across dozens of languages.

Frequently asked

What is yaoguangluo/Deta_Parser?
Chinese word segmentation library processing over 16 million characters per second using HMM and neural network algorithms.
Is Deta_Parser open source?
Yes — yaoguangluo/Deta_Parser is open source, released under the GPL-2.0 license.
What language is Deta_Parser written in?
yaoguangluo/Deta_Parser is primarily written in Java.
How popular is Deta_Parser?
yaoguangluo/Deta_Parser has 478 stars on GitHub.
Where can I find Deta_Parser?
yaoguangluo/Deta_Parser is on GitHub at https://github.com/yaoguangluo/Deta_Parser.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.