Ablations β Loving
Collection
Ablation variants of the 'loving' trait (DPO200, no-DPO). β’ 2 items β’ Updated
An ablation experiment from Open Character Training.
This model skips DPO entirely β it trains SFT introspection directly on the base Qwen 2.5 7B Instruct model, using freshly generated self-reflection and self-interaction data from the base model (with constitutional system prompts but no DPO adapter).
| Branch | Description |
|---|---|
main |
README + training data |
introspection-final |
Final SFT LoRA adapter |
introspection-global_step200 - global_step325 |
Intermediate SFT checkpoints (every 25 steps) |
main branch, under data/)
| File | Rows | Description |
|---|---|---|
self_reflection.jsonl |
10,000 | Self-reflection responses from base model with constitutional prompt |
self_interaction_free.jsonl |
1,000 | Free 10-turn self-interaction conversations |
self_interaction_leading.jsonl |
1,000 | Leading (reflective) 10-turn self-interaction conversations |
sft_data_compiled.jsonl |
12,000 | Compiled SFT dataset (all three merged + shuffled) |
This ablation tests whether DPO distillation is necessary, or if SFT introspection alone (with constitutional prompts) is sufficient to instill the loving persona. Compare with:
sdananya/qwen-2.5-7b-it-loving β full pipeline (DPO final + SFT)sdananya/qwen-2.5-7b-it-loving-dpo200 β DPO step 200 + SFT