Llama-2-7B SSFT-CB + ARC-Challenge AsFT

This is the fully merged checkpoint obtained by AsFT fine-tuning wvnvwn/llama2-7b-chat-lr5e-5-ssft-cb on the ARC-Challenge training split.

Training used 1,119 examples for 3 epochs with learning rate 3e-4, physical and effective batch size 16, cosine scheduling with warmup ratio 0.1, BF16, maximum sequence length 1,024, and seed 42. The LoRA configuration used rank 16, alpha 32, dropout 0.05, and target modules q_proj, k_proj, v_proj, up_proj, and down_proj. AsFT used 160 alignment directions and lambda_reg=1.0.

The Llama-2 chat template was applied and prompt tokens were masked from the training loss. The saved weights have already been merged into the starting model; no separate adapter is required for inference.

Downloads last month
13
Safetensors
Model size
7B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for wvnvwn/llama2-7b-chat-lr5e-5-arcc-lr3e-4-cbasft-new-new

Finetuned
(6)
this model