Text Generation
Transformers
Safetensors
PyTorch
nemotron_h
nvidia
nemotron-3
latent-moe
mtp
conversational
custom_code
Eval Results
nielsr HF Staff commited on
Commit
9cb0715
·
verified ·
1 Parent(s): d51eab0

Add community evaluation results for SWE-Bench Multilingual

Browse files

This PR adds community-provided evaluation results for the following benchmarks:

- **[SWE-BENCH_MULTILINGUAL](https://huggingface.co/datasets/SWE-bench/SWE-bench_Multilingual?eval_result=nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16&leaderboard_task_id=swe_bench_multilingual_%25_resolved)**

These results were extracted from the model card. This is based on the new [evaluation results feature](https://huggingface.co/docs/hub/eval-results).

*Note: This is an automated PR. Please review the evaluation results before merging.*

.eval_results/swe-bench_multilingual.yaml ADDED
@@ -0,0 +1,8 @@
 
 
 
 
 
 
 
 
 
1
+ - dataset:
2
+ id: SWE-bench/SWE-bench_Multilingual
3
+ task_id: swe_bench_multilingual_%_resolved
4
+ value: 45.8
5
+ source:
6
+ url: https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16
7
+ name: Model Card
8
+