Smoltaur-0.1B-LoRA-r16-f0.125

Hugging Face

socius Paper Parameters LoRA Dataset Data fraction

Smoltaur-0.1B-LoRA-r16-f0.125

LoRA adapter for Smoltaur-0.1B, fine-tuned on a stratified 12.5% subset of Psych-101 as part of the LoRA-rank sweep and dataset-size ablation for Small Foundation Models of Human Cognition and Behaviour.

field value
base model unsloth/SmolLM2-135M
LoRA rank 16 (alpha = rank, rsLoRA)
data fraction 12.5% of Psych-101
training 1 epoch, completion-only loss, seed 3407

Load with PEFT on top of unsloth/SmolLM2-135M, or evaluate with the project's eval_model.py --backend unsloth.

Downloads last month
13
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for socius/Smoltaur-0.1B-LoRA-r16-f0.125

Adapter
(198)
this model

Dataset used to train socius/Smoltaur-0.1B-LoRA-r16-f0.125

Paper for socius/Smoltaur-0.1B-LoRA-r16-f0.125