What it is
This LoRA adapter specializes 'Qwen/Qwen2.5-Coder-7B-Instruct' for Go code tasks (review, generation,
explanation, testing) as part of the GuildLM
Code Guild. It was produced by Anvil's config-driven QLoRA supervised
fine-tuning pipeline.
Use with the base model
This is a LoRA adapter — load it on top of the base model with PEFT, or merge it for serving:
anvil-merge --base-model Qwen/Qwen2.5-Coder-7B-Instruct \ --adapter ./adapter --output-dir ./merged --dtype bfloat16
Training
- Method: QLoRA (4-bit NF4 base + LoRA adapters), supervised fine-tuning.
- Tool:
guildlm-anvil — see TRAINING.md
for the end-to-end free recipe (Kaggle GPU -> HuggingFace Hub -> Ollama).
Limitations
Quality is bounded by the training dataset. If trained on the offline-synthetic
smoke-test sample, this model only learns a placeholder format — use a real
teacher-generated dataset (forge online mode) for a shippable specialist.