Dataset containing 216,930 Jeopardy questions, answers and other data.
The json file is an unordered list of questions where each question has 'category' : the question category, e.g. "HISTORY" 'value' : integer $ value of the question as string, e.g. "200" Note: This is "None" for Final Jeopardy! and Tiebreaker questions 'question' : text of question Note: This sometimes contains hyperlinks and other things messy text such as when there's a picture or video question 'answer' : text of answer 'round' : one of "Jeopardy!","Double Jeopardy!","Final Jeopardy!" or "Tiebreaker" Note: Tiebreaker questions do happen but they're very rare (like once every 20 years) 'show_number' : int of show number, e.g '4680' 'air_date' : string of the show air date in format YYYY-MM-DD
An example of 'train' looks as follows.
{
"air_date": "2004-12-31",
"answer": "Hattie McDaniel (for her role in Gone with the Wind)",
"category": "EPITAPHS & TRIBUTES",
"question": "'1939 Oscar winner: \"...you are a credit to your craft, your race and to your family\"'",
"round": "Jeopardy!",
"show_number": 4680,
"value": 2000
}
The data fields are the same among all splits.
category: a string feature.air_date: a string feature.question: a string feature.value: a int32 feature.answer: a string feature.round: a string feature.show_number: a int32 feature.| name | train |
|---|---|
| default | 216930 |
Thanks to @thomwolf, @lewtun, @patrickvonplaten for adding this dataset.
Dataset containing 216,930 Jeopardy questions, answers and other data.
The json file is an unordered list of questions where each question has 'category' : the question category, e.g. "HISTORY" 'value' : integer $ value of the question as string, e.g. "200" Note: This is "None" for Final Jeopardy! and Tiebreaker questions 'question' : text of question Note: This sometimes contains hyperlinks and other things messy text such as when there's a picture or video question 'answer' : text of answer 'round' : one of "Jeopardy!","Double Jeopardy!","Final Jeopardy!" or "Tiebreaker" Note: Tiebreaker questions do happen but they're very rare (like once every 20 years) 'show_number' : int of show number, e.g '4680' 'air_date' : string of the show air date in format YYYY-MM-DD
An example of 'train' looks as follows.
{
"air_date": "2004-12-31",
"answer": "Hattie McDaniel (for her role in Gone with the Wind)",
"category": "EPITAPHS & TRIBUTES",
"question": "'1939 Oscar winner: \"...you are a credit to your craft, your race and to your family\"'",
"round": "Jeopardy!",
"show_number": 4680,
"value": 2000
}
The data fields are the same among all splits.
category: a string feature.air_date: a string feature.question: a string feature.value: a int32 feature.answer: a string feature.round: a string feature.show_number: a int32 feature.| name | train |
|---|---|
| default | 216930 |
Thanks to @thomwolf, @lewtun, @patrickvonplaten for adding this dataset.