Training
Serve with vLLM
A standard Qwen2.5 checkpoint; no overrides needed at its native 32768 context:
CUDA_VISIBLE_DEVICES=0 \
python -m vllm.entrypoints.openai.api_server \
--model TIGER-Lab/FIM-Mid-7B \
--served-model-name FIM-Mid-7B \
--host 127.0.0.1 \
--port 8400 \
--tensor-parallel-size 1 \
--max-model-len 32768 \
--gpu-memory-utilization 0.9 \
> vllm_fim_mid7b.log 2>&1 &
Post-training
To reproduce FIM-7B, run R2E-Gym trajectory SFT from this checkpoint — the exact config is posttraining/r2egym/FIM_Posttrain_7B.yaml (LLaMA-Factory, full fine-tuning, lr 1.0e-5, 2 epochs, cutoff 32768), which already points at this repo id. See posttraining/r2egym/ for the walkthrough.
Citation
@article{wang2026fim,
title={Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models},
author={Wang, Yubo and Liang, Jiarong and Zhang, Yuxuan and Liu, Xuye and Wei, Cong and Zhang, Yuyu and Nie, Ping and Chen, Wenhu},
journal={arXiv preprint arXiv:2607.12463},
year={2026}
}