Matryoshka NLA (Qwen2.5-7B L20)
An NLA trained with random-length truncation of the verbalizer's output, so the most important information comes first. Checkpoints, data, baseline.
UpdatedNote The matryoshka NLA (AV + AR pair) — RL checkpoints every 50 steps; iter_0000200 is the model used in the post
syvb/nla-qwen2.5-7b-L20-av-matryoshka-sonnet46-v3
8B • Updated • 5Note Pre-RL warm-start activation verbalizer (bullet-point format SFT)
syvb/nla-qwen2.5-7b-L20-ar-matryoshka-sonnet46-v3
5B • Updated • 7Note Pre-RL warm-start activation reconstructor (SFT on randomly token-truncated explanations)
syvb/nla-qwen2.5-7b-L20-v3-rl-checkpoints
UpdatedNote Resumable RL training state at step 200 (actor DCP + optimizer, critic HF)
syvb/nla-qwen2.5-7b-L20-matryoshka-warmstart-sonnet46
Preview • Updated • 23Note Warm-start data: Qwen2.5-7B-Instruct L20 activations + Claude Sonnet 4.6 explanations (v3/ holds the bullets-format parquets used for this run)
ceselder/nla-matryoshka-warmstart-sonnet46
Viewer • Updated • 450k • 34 • 1Note Source explanations (text-only) the warm-start data was built from
kitft/nla-qwen2.5-7b-L20-av
8B • Updated • 1.1k • 7Note Baseline: the original normally-trained Qwen2.5-7B NLA activation verbalizer
kitft/nla-qwen2.5-7b-L20-ar
5B • Updated • 497 • 8Note Baseline: the original NLA activation reconstructor