ARC (AI2 Reasoning Challenge)
The ARC (AI2 Reasoning Challenge) is a benchmark dataset and challenge designed to evaluate the reasoning abilities of artificial intelligence systems on grade-school level science questions. Developed by the Allen Institute for Artificial Intelligence, it aims to advance AI research in natural language understanding and reasoning.