Viewer
• Updated • 412k • 16.4k
• 117
Viewer
• Updated • 570k • 36.7k
• 96
Updated • 1.19k
• 6
Right for the Wrong Reasons: Diagnosing Syntactic Heuristics in Natural
Language Inference
Paper
• 1902.01007
• Published • 1
A large annotated corpus for learning natural language inference
Paper
• 1508.05326
• Published • 1
Benchmark
• Updated • 1.25k • 121k
• 513
Benchmark
• Updated • 12.1k • 199k
• 512
google-research-datasets/nq_open
Viewer
• Updated • 91.5k • 24.6k
• 35
Viewer
• Updated • 203k • 96.9k
• 326
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question
Answering
Paper
• 1809.09600
• Published • 5
Viewer
• Updated • 237k • 14k
• 83
LexGLUE: A Benchmark Dataset for Legal Language Understanding in English
Paper
• 2110.00976
• Published • 1
Viewer
• Updated • 1.59k • 7.87k
• 110
A Dataset of Information-Seeking Questions and Answers Anchored in
Research Papers
Paper
• 2105.03011
• Published • 2
Amod/mental_health_counseling_conversations
Viewer
• Updated • 3.51k • 1.61k
• 491
MuskumPillerum/General-Knowledge
Viewer
• Updated • 37.6k • 322
• 49
Viewer
• Updated • 98.2k • 209k
• 473
SQuAD: 100,000+ Questions for Machine Comprehension of Text
Paper
• 1606.05250
• Published • 6
Viewer
• Updated • 196k • 273k
• 189
SuperGLUE: A Stickier Benchmark for General-Purpose Language
Understanding Systems
Paper
• 1905.00537
• Published • 3
Viewer
• Updated • 11.9k • 246k
• 136
Viewer
• Updated • 7.79k • 606k
• 385
Think you have Solved Question Answering? Try ARC, the AI2 Reasoning
Challenge
Paper
• 1803.05457
• Published • 5
Viewer
• Updated • 914k • 136k
• 201
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for
Reading Comprehension
Paper
• 1705.03551
• Published • 2
Viewer
• Updated • 274k • 63.8k
• 338
PubMedQA: A Dataset for Biomedical Research Question Answering
Paper
• 1909.06146
• Published • 5
GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Paper
• 2311.12022
• Published • 39