nielsr HF Staff commited on
Commit
2d2e74c
·
verified ·
1 Parent(s): 4d01dfe

Add community evaluation results for SWE-Bench Multilingual

Browse files

This PR adds community-provided evaluation results for the following benchmarks:

- **[SWE-BENCH_MULTILINGUAL](https://huggingface.co/datasets/SWE-bench/SWE-bench_Multilingual?eval_result=moonshotai/Kimi-K2.5&leaderboard_task_id=swe_bench_multilingual_%25_resolved)**

These results were extracted from the model card. This is based on the new [evaluation results feature](https://huggingface.co/docs/hub/eval-results).

*Note: This is an automated PR. Please review the evaluation results before merging.*

.eval_results/swe-bench_multilingual.yaml ADDED
@@ -0,0 +1,8 @@
 
 
 
 
 
 
 
 
 
1
+ - dataset:
2
+ id: SWE-bench/SWE-bench_Multilingual
3
+ task_id: swe_bench_multilingual_%_resolved
4
+ value: 73.0
5
+ source:
6
+ url: https://huggingface.co/moonshotai/Kimi-K2.5
7
+ name: Model Card
8
+