3
stars
7
commits
1
linked in READMEs
Feb 19, 2025
updated
用于中文拼写和语法纠错大规模数据集,共计1568885条数据集,含有Lang8、HSK数据集,完整项目代码:https://github.com/TW-NLP/ChineseErrorCorrector
7 commits
twnlp/csc_data
4
twnlp/ChinseseErrorCorrectData
2
shibing624/chinese_text_correction
15
shibing624/CSC
37
twnlp/ChineseErrorCorrector4-4B
karl-wang/MuChin1k
Linly-AI/Chinese-pretraining-dataset
44
FreedomIntelligence/evol-instruct-chinese
10
TW-NLP/ChineseErrorCorrector
一个面向中文文本纠错任务的综合平台,集学术研究、模型训练、模型评测和推理部署于一体,文本纠错新Sota。( 2026 ACL Main Oral )
639