The ArabicCulture dataset was generated by gpt3.5 and contains 8000+ True and False questions.
The dataset contains questions from 58 different areas.
In the answers, "True" accounted for 59.62%, and "False" accounted for 40.38%
It contains 8000+ data, and we took 5 data from each area as few-shot data.
We asked two Arabs to judge 4000 of all the data for us, and we left data that two Arabs both thought were good. Finally, we got 2.4k data covering 9 areas.
We divided them into test sets and validation sets as above.
The ArabicCulture dataset was generated by gpt3.5 and contains 8000+ True and False questions.
The dataset contains questions from 58 different areas.
In the answers, "True" accounted for 59.62%, and "False" accounted for 40.38%
It contains 8000+ data, and we took 5 data from each area as few-shot data.
We asked two Arabs to judge 4000 of all the data for us, and we left data that two Arabs both thought were good. Finally, we got 2.4k data covering 9 areas.
We divided them into test sets and validation sets as above.