ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian Paper • 2501.06715 • Published Jan 12, 2025
Running Agents 32 Ukrainian LLM Leaderboard 👁 32 Measuring LLM capabilities to process Ukrainian texts
Empowering Smaller Models: Tuning LLaMA and Gemma with Chain-of-Thought for Ukrainian Exam Tasks Paper • 2503.13988 • Published Mar 18, 2025 • 1
UA-Code-Bench: A Competitive Programming Benchmark for Evaluating LLM Code Generation in Ukrainian Paper • 2511.05040 • Published Nov 7, 2025