qwen3-8b-bfcl-sft

Qwen3-8B fine-tuned (full-parameter SFT) on successful BFCL agent trajectories generated by a Qwen3.6-27B teacher (no-rethink scaffold, train split, n=8 rollouts, deduplicated successful main-agent calls).

  • Base model: Qwen/Qwen3-8B
  • Data: 591 call-level samples (~22k supervised tokens), loss on assistant targets only, prompts rendered serving-style with enable_thinking=False
  • Training: 1 epoch, lr 5e-6, bf16, max_len 16384

Intended for agentic function-calling research (BFCL-style multi-step tool use).

Downloads last month
5
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for JunjieMeng/qwen3-8b-bfcl-sft

Finetuned
Qwen/Qwen3-8B
Finetuned
(2151)
this model