qwen2_5_3b_uid_reference_lr3e6_step4

Full model weights from the UID reference training run, with the step 4 LoRA actor merged into Qwen/Qwen2.5-3B.

  • Training job: 3977891
  • Global step: 4
  • Learning rate: 3e-6
  • LoRA rank: 32; alpha: 16
  • Export dtype: bfloat16

Load directly with AutoModelForCausalLM.from_pretrained; no separate adapter is required.

Downloads last month
197
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for talzoomanzoo/qwen2_5_3b_uid_reference_lr3e6_step4

Base model

Qwen/Qwen2.5-3B
Adapter
(447)
this model