5 repos
xlang-ai/BRIGHT
[ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval
213
26 commits
jacklanda/SemanticQA
[ACL 2026 Oral] Revisiting a Pain in the Neck: A Semantic Reasoning Benchmark for Language Models
10
48 commits
yale-nlp/Bright-Pro
No description
1
2 commits
facebook/WearableQA
11
j991222/mirb
MIRB: Mathematical Information Retrieval Benchmark
3
3 commits