Buckets:
32.1 MB
4 files
Updated 13 days ago
Ctrl+K
| Name | Size | Uploaded | Xet hash |
|---|---|---|---|
| .gitattributes | 2.52 kB xet | 0be3479b | |
| README.md | 3.17 kB xet | 091ef5c0 | |
| questions_data.csv | 4.17 MB xet | 0f513c56 | |
| responses_data.csv | 27.9 MB xet | fb608bcc |
This dataset contains 4,515 multiple-choice questions from five major Brazilian university entrance exams (ENEM, FUVEST, UNICAMP, ITA, IME) spanning 32 years (1981-2025), along with model responses from 20 LLMs.
Files
📄 questions_data.csv (4,515 rows)
Contains the exam questions with:
question_id: Unique identifierquestion_statement: Question text in Portuguesecorrect_answer: Correct option (A-E)alternative_atoalternative_e: Answer choicessubject: Academic subjectexam_name,exam_year,exam_type: Exam metadata
📄 responses_data.csv
Contains model responses with:
model: Model name (o3, deepseek-reasoner, claude-opus-4-20250514)prompt_template: Prompting strategy used (zero-shot, role-playing, chain-of-thought)chosen_answer: Model's selected answeris_correct: Whether the answer was correctdifficulty_level,uncertainty_level: Model's self-reported metrics (1-10 scale)bloom_taxonomy: Cognitive complexity classification- Additional metadata matching questions_data
Cite
@misc{godoy2025alvoradabenchlanguagemodelssolve,
title={Alvorada-Bench: Can Language Models Solve Brazilian University Entrance Exams?},
author={Henrique Godoy},
year={2025},
eprint={2508.15835},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2508.15835},
}
- Total size
- 32.1 MB
- Files
- 4
- Last updated
- Aug 1
- Pre-warmed CDN
- US EU US EU