Text Generation
Transformers
Safetensors
qwen3
code
software-engineering
agent
conversational
text-generation-inference
ubowang commited on
Commit
03a854e
·
verified ·
1 Parent(s): 1987c8f

Add arXiv link (2607.12463)

Browse files
Files changed (1) hide show
  1. README.md +2 -2
README.md CHANGED
@@ -15,7 +15,7 @@ tags:
15
 
16
  # FIM-8B
17
 
18
- [📄 Paper (PDF)](https://github.com/TIGER-AI-Lab/FIM-Midtraining/blob/main/paper.pdf) · [💻 GitHub](https://github.com/TIGER-AI-Lab/FIM-Midtraining) · [🤗 Dataset](https://huggingface.co/datasets/TIGER-Lab/FIM-Midtraining-400K) · [🤗 Collection](https://huggingface.co/collections/TIGER-Lab/fim-midtraining)
19
 
20
  **FIM-8B** is the strongest released model of *"Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models"*: `Qwen3-8B`, mid-trained on function-aware FIM data, then post-trained on SWE-Lego agent trajectories. The mid-training stage is the only difference from a standard SWE-Lego reproduction — worth **+3.2 points on SWE-Bench-Verified and +5.4 on SWE-Bench-Lite**. Unlike FIM-7B and FIM-14B (R2E-Gym scaffold), this model is evaluated with the SWE-Lego setup: OpenHands `CodeActAgent` for inference and the official SWE-bench harness for scoring.
21
 
@@ -109,7 +109,7 @@ Convert the OpenHands `output.jsonl` to a predictions file with `evaluation/benc
109
  @article{wang2026fim,
110
  title={Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models},
111
  author={Wang, Yubo and Liang, Jiarong and Zhang, Yuxuan and Liu, Xuye and Wei, Cong and Zhang, Yuyu and Nie, Ping and Chen, Wenhu},
112
- journal={arXiv preprint},
113
  year={2026}
114
  }
115
  ```
 
15
 
16
  # FIM-8B
17
 
18
+ [📄 Paper](https://arxiv.org/abs/2607.12463) · [💻 GitHub](https://github.com/TIGER-AI-Lab/FIM-Midtraining) · [🤗 Dataset](https://huggingface.co/datasets/TIGER-Lab/FIM-Midtraining-400K) · [🤗 Collection](https://huggingface.co/collections/TIGER-Lab/fim-midtraining)
19
 
20
  **FIM-8B** is the strongest released model of *"Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models"*: `Qwen3-8B`, mid-trained on function-aware FIM data, then post-trained on SWE-Lego agent trajectories. The mid-training stage is the only difference from a standard SWE-Lego reproduction — worth **+3.2 points on SWE-Bench-Verified and +5.4 on SWE-Bench-Lite**. Unlike FIM-7B and FIM-14B (R2E-Gym scaffold), this model is evaluated with the SWE-Lego setup: OpenHands `CodeActAgent` for inference and the official SWE-bench harness for scoring.
21
 
 
109
  @article{wang2026fim,
110
  title={Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models},
111
  author={Wang, Yubo and Liang, Jiarong and Zhang, Yuxuan and Liu, Xuye and Wei, Cong and Zhang, Yuyu and Nie, Ping and Chen, Wenhu},
112
+ journal={arXiv preprint arXiv:2607.12463},
113
  year={2026}
114
  }
115
  ```