evals dataets openai/mrcr Viewer • Updated Dec 8, 2025 • 2.4k • 4.49k • 216 zai-org/LongBench-v2 Viewer • Updated Dec 20, 2024 • 503 • 46.4k • 48 ibm-research/REAL-MM-RAG_FinReport Viewer • Updated Mar 16, 2025 • 2.93k • 914 • 8 dreamerdeo/finqa Viewer • Updated Mar 6, 2023 • 8.28k • 3.09k • 29
evals dataets openai/mrcr Viewer • Updated Dec 8, 2025 • 2.4k • 4.49k • 216 zai-org/LongBench-v2 Viewer • Updated Dec 20, 2024 • 503 • 46.4k • 48 ibm-research/REAL-MM-RAG_FinReport Viewer • Updated Mar 16, 2025 • 2.93k • 914 • 8 dreamerdeo/finqa Viewer • Updated Mar 6, 2023 • 8.28k • 3.09k • 29