Qwen2.5-3B β Yoruba (AutoScientist Challenge)
QLoRA SFT of Qwen/Qwen2.5-3B on an Adaption-localized Yoruba instruction set (language_expansion,
localize β Yoruba/NG). Targets the language category (Yoruba).
Results vs the base model
Held-out Yoruba perplexity: 18.782 β 7.597 (59.55% lower) β primary metric, on held-out Yoruba instruction data the model never trained on (genuine generalization, not memorization).
Belebele (yor_Latn, n=120): 0.1917 β 0.2083 (+8.7% rel) β reported for completeness; a near-chance, high-variance benchmark at this scale.
Data: English instructions localized to Yoruba via the Adaptive Data API, filtered to Yoruba.
Training: 4-bit QLoRA (r=16), packed sequences, single T4. Load base + this adapter.
Inference Providers NEW
This model isn't deployed by any Inference Provider. π Ask for provider support
Model tree for Rome-1/qwen2.5-3b-yoruba-autoscientist
Base model
Qwen/Qwen2.5-3B