ivanhe/opencap
The NYU Open CAP dataset consists of 50 patents related to the technology field of speech recognition, in which citations to scientific articles are annotated with CoNLL style B/I/O tags.
ansj分词.ict的真正java实现.分词效果速度都超过开源版的ict. 中文分词,人名识别,词性标注,用户自定义词典
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
The NYU Open CAP dataset consists of 50 patents related to the technology field of speech recognition, in which citations to scientific articles are annotated with CoNLL style B/I/O tags.
中文自然语言处理工具包 Toolkit for Chinese natural language processing (formerly FudanNLP)
AIDA Named Entity Disambiguation by the Databases and Information Systems Group at the Max Planck Institute for Informatics.
Public repository.