olmo3-32b-opd-merge-125-to-225

A uniform linear weight average ("model soup") of five training checkpoints of fieldsmodelorg/Olmo-3.1-32B-Think-OPD-IMO, created with mergekit.

Merge Details

  • Method: Linear (equal weights → simple arithmetic mean of the checkpoints)
  • Precision: bfloat16
  • Architecture: Olmo3SinkForCausalLM (unchanged from the source checkpoints)

Checkpoints merged (equal weight)

Sub-folders of the source repo, at training steps 125, 150, 175, 200, and 225:

  • opd-32b-bf16-step-125
  • opd-32b-bf16-step-150
  • opd-32b-bf16-step-175
  • opd-32b-bf16-step-200
  • opd-32b-bf16-step-225

Configuration

merge_method: linear
dtype: bfloat16
models:
  - model: opd-32b-bf16-step-125
    parameters: {weight: 1.0}
  - model: opd-32b-bf16-step-150
    parameters: {weight: 1.0}
  - model: opd-32b-bf16-step-175
    parameters: {weight: 1.0}
  - model: opd-32b-bf16-step-200
    parameters: {weight: 1.0}
  - model: opd-32b-bf16-step-225
    parameters: {weight: 1.0}
Downloads last month
223
Safetensors
Model size
33B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for chankhavu/olmo3-32b-opd-merge-125-to-225

Finetuned
(1)
this model

Paper for chankhavu/olmo3-32b-opd-merge-125-to-225