11 repos
monologg/KoBERT-NER
NER Task with KoBERT (with Naver NLP Challenge dataset)
99
27 commits
zjunlp/DeepKE
[EMNLP 2022] An Open Toolkit for Knowledge Graph Extraction and Construction
4,481
1,691 commits
CLUEbenchmark/CLUEDatasetSearch
搜索所有中文NLP数据集,附常用英文NLP数据集
4,454
36 commits
arthrod/gliner-looong
Long-context experiments with GLiNER for documents that exceed 512-token windows.
0
592 commits
arclabs561/anno
Text annotation and entity extraction
11
520 commits
arthrod/gliner2_finetune
Fine-tuning workflows for GLiNER 2 on domain-specific entity recognition.
1
24 commits
hamza-aziz-ai/CrossLingual-IE-MT
Cross-lingual NLP pipeline (English <-> Hindi/Tamil/Marathi/Bengali) translating text with NLLB-200…
3 commits
ramisa2108/Bangla-Complex-Named-Entity-Recognition-Challenge
Winning Solution for the Bangla Complex Named Entity Recognition Challenge - BDOSN NLP Hackathon…
7
26 commits
thunlp/Few-NERD
Code and data of ACL 2021 paper "Few-NERD: A Few-shot Named Entity Recognition Dataset"
400
esbatmop/MNBVC
MNBVC(Massive Never-ending BT Vast Chinese corpus)超大规模中文语料集。对标chatGPT训练的40T数据。MNBVC数据集不但包括主流文化,也包括各…
4,275
305 commits
aplmikex/deduplication_mnbvc
文本去重
76
10 commits