solbon1212 commited on
Commit
8e3f4d0
·
verified ·
1 Parent(s): eba3fea

Initial upload: 2000-step MixGRPO LoRA checkpoints + trend figure + README

Browse files
Files changed (45) hide show
  1. .gitattributes +1 -0
  2. README.md +119 -0
  3. adapter_config.json +49 -0
  4. adapter_model.safetensors +3 -0
  5. checkpoints/step_000100/adapter_config.json +49 -0
  6. checkpoints/step_000100/adapter_model.safetensors +3 -0
  7. checkpoints/step_000200/adapter_config.json +49 -0
  8. checkpoints/step_000200/adapter_model.safetensors +3 -0
  9. checkpoints/step_000300/adapter_config.json +49 -0
  10. checkpoints/step_000300/adapter_model.safetensors +3 -0
  11. checkpoints/step_000400/adapter_config.json +49 -0
  12. checkpoints/step_000400/adapter_model.safetensors +3 -0
  13. checkpoints/step_000500/adapter_config.json +49 -0
  14. checkpoints/step_000500/adapter_model.safetensors +3 -0
  15. checkpoints/step_000600/adapter_config.json +49 -0
  16. checkpoints/step_000600/adapter_model.safetensors +3 -0
  17. checkpoints/step_000700/adapter_config.json +49 -0
  18. checkpoints/step_000700/adapter_model.safetensors +3 -0
  19. checkpoints/step_000800/adapter_config.json +49 -0
  20. checkpoints/step_000800/adapter_model.safetensors +3 -0
  21. checkpoints/step_000900/adapter_config.json +49 -0
  22. checkpoints/step_000900/adapter_model.safetensors +3 -0
  23. checkpoints/step_001000/adapter_config.json +49 -0
  24. checkpoints/step_001000/adapter_model.safetensors +3 -0
  25. checkpoints/step_001100/adapter_config.json +49 -0
  26. checkpoints/step_001100/adapter_model.safetensors +3 -0
  27. checkpoints/step_001200/adapter_config.json +49 -0
  28. checkpoints/step_001200/adapter_model.safetensors +3 -0
  29. checkpoints/step_001300/adapter_config.json +49 -0
  30. checkpoints/step_001300/adapter_model.safetensors +3 -0
  31. checkpoints/step_001400/adapter_config.json +49 -0
  32. checkpoints/step_001400/adapter_model.safetensors +3 -0
  33. checkpoints/step_001500/adapter_config.json +49 -0
  34. checkpoints/step_001500/adapter_model.safetensors +3 -0
  35. checkpoints/step_001600/adapter_config.json +49 -0
  36. checkpoints/step_001600/adapter_model.safetensors +3 -0
  37. checkpoints/step_001700/adapter_config.json +49 -0
  38. checkpoints/step_001700/adapter_model.safetensors +3 -0
  39. checkpoints/step_001800/adapter_config.json +49 -0
  40. checkpoints/step_001800/adapter_model.safetensors +3 -0
  41. checkpoints/step_001900/adapter_config.json +49 -0
  42. checkpoints/step_001900/adapter_model.safetensors +3 -0
  43. config/config_used.py +97 -0
  44. figures/w3030_trends.pdf +0 -0
  45. figures/w3030_trends.png +3 -0
.gitattributes CHANGED
@@ -33,3 +33,4 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
 
 
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ figures/w3030_trends.png filter=lfs diff=lfs merge=lfs -text
README.md ADDED
@@ -0,0 +1,119 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: cc-by-nc-4.0
3
+ tags:
4
+ - diffusion
5
+ - flow-matching
6
+ - music
7
+ - lora
8
+ - grpo
9
+ - rl
10
+ - diffrhythm
11
+ base_model: ASLP-lab/DiffRhythm2
12
+ library_name: peft
13
+ ---
14
+
15
+ # DiffRhythm2 · MixGRPO LoRA (checkpoints, MAAP)
16
+
17
+ LoRA adapter checkpoints from a **block-wise dense-reward GRPO** run on top of
18
+ the [ASLP-lab/DiffRhythm2](https://huggingface.co/ASLP-lab/DiffRhythm2)
19
+ continuous-time flow-matching music model. Released as **research artifacts
20
+ for a negative-result study** on RL fine-tuning of long-form audio flow models.
21
+
22
+ ## What is here
23
+
24
+ - **`adapter_model.safetensors` / `adapter_config.json`** at the repo root — the
25
+ final `ckpt_final_002000` (2000 optimization steps). Load directly via
26
+ `PeftModel.from_pretrained("m-a-a-p/DiffRhythm2-MixGRPO-LoRA", …)`.
27
+ - **`checkpoints/step_XXXXXX/`** — intermediate LoRA adapters at every 100
28
+ training steps (Step 100, 200, …, 1900) for anyone who wants to reproduce
29
+ the reward trajectory or run ablations.
30
+ - **`figures/w3030_trends.png`** — 3-panel trend plot for the final sliding
31
+ window (`window=[30-30]`), showing reward, KL, and clip-fraction over the
32
+ 1840 steps spent in that window.
33
+ - **`config/config_used.py`** — the exact `mixgrpo_lora_diffrhythm2()` config
34
+ section used for this run, for reproducibility.
35
+
36
+ ## Training setup
37
+
38
+ Adapter is inserted into the DiT attention layers of DiffRhythm2 (dim 2048,
39
+ depth 16, 16 heads, mel_dim 64).
40
+
41
+ | item | value |
42
+ |---|---|
43
+ | base model | `ASLP-lab/DiffRhythm2` |
44
+ | optimizer | AdamW, `lr = 5e-6`, `weight_decay = 1e-4`, `eps = 1e-8` |
45
+ | LoRA | `r = 32`, `alpha = 64` |
46
+ | RL algorithm | Block-wise MixGRPO (sliding window over 4 blocks, 30 s each) |
47
+ | group size | 12 rollouts per prompt |
48
+ | PPO-style clip | `clip_range = 0.1` |
49
+ | KL coefficient | `0.01` |
50
+ | gradient accumulation | 3 |
51
+ | inner epochs | 1 |
52
+ | EMA | **disabled** (OOM mitigation) |
53
+ | CUDA alloc | `expandable_segments:True` |
54
+ | total steps | 2000 |
55
+ | hardware | 1 × RTX PRO 6000 (96 GB), Slurm job 1738062 |
56
+
57
+ **Rewards** are a 50/50 blend of the audiobox aesthetics score and an
58
+ ACE-Step-style Audio Alignment Score (AAS), aggregated over each 30 s block
59
+ window.
60
+
61
+ ## Result: no measurable improvement
62
+
63
+ Across 2000 steps in the final sliding window (`[30-30]`, 1840 steps, Steps
64
+ 160→1999), reward exhibits a **statistically significant regression**:
65
+
66
+ - early-half mean reward: **0.18669** (n = 920)
67
+ - late-half mean reward: **0.18592** (n = 920)
68
+ - Δ = **−7.7 × 10⁻⁴**, t ≈ **−26.7**
69
+ - linear slope: **−8.9 × 10⁻⁴ per 1000 steps**
70
+ - KL(π ∥ π_ref) simultaneously *decreases* from 0.018 to 0.013
71
+
72
+ The full reward range across all 1840 steps is only **0.006** (≈ 0.3 %),
73
+ i.e. the same magnitude as the per-step reward noise. Hyperparameter sweeps
74
+ in the accompanying paper (across `num_generations ∈ {4, 12, 16}`,
75
+ `clip_range ∈ {1e-5, 1e-4, 0.1}`, `lr ∈ {5e-6, 1e-5, 5e-5}`) do not escape
76
+ this plateau.
77
+
78
+ Interpretation: the log-probability signal that GRPO needs — computed here
79
+ through the surrogate ratio between the current and the pre-update policy on
80
+ the ODE / SDE trajectory — is **noise-dominated** at the scale of the block-
81
+ wise dense rewards used here, so the policy gradient does not consistently
82
+ point in the reward-increasing direction. See the accompanying paper for the
83
+ full analysis (including the diagnosis of the `1 / (1−t)` score-head
84
+ singularity in the underlying flow-matching formulation).
85
+
86
+ ## Intended use
87
+
88
+ These adapters are released as **research artifacts** to accompany the
89
+ negative-result study, not as a recommended production LoRA on top of
90
+ DiffRhythm2. In particular:
91
+
92
+ - Do not expect audio-quality gains over the base DiffRhythm2 model.
93
+ - If you use them as a *baseline* or *sanity check* for a new RL algorithm,
94
+ please cite the accompanying paper.
95
+
96
+ ## Quick start
97
+
98
+ ```python
99
+ from peft import PeftModel
100
+ from safetensors.torch import load_file
101
+
102
+ # Final (2000-step) adapter, loaded via peft on top of your DiffRhythm2 DiT.
103
+ adapter = load_file("adapter_model.safetensors")
104
+
105
+ # Intermediate:
106
+ # hf_hub_download("m-a-a-p/DiffRhythm2-MixGRPO-LoRA",
107
+ # "checkpoints/step_001000/adapter_model.safetensors")
108
+ ```
109
+
110
+ ## License
111
+
112
+ Released under **CC BY-NC 4.0**. Non-commercial use only; academic use
113
+ encouraged. Base model (`ASLP-lab/DiffRhythm2`) is subject to its own
114
+ license.
115
+
116
+ ## Citation
117
+
118
+ If you use these checkpoints or the accompanying analysis, please cite the
119
+ associated paper (details to be added when the paper appears).
adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:5db4ed3e1c6383171896b3d1825e67eebf2083a72c8472a5c80f940e1126cb81
3
+ size 96500080
checkpoints/step_000100/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000100/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:52744bfc41f52db9ff346abc2a7b91427da9cbf6ee0636a8b20ef7f176a9fbe7
3
+ size 96500080
checkpoints/step_000200/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000200/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:7a4f1a74aff8f953e7676ea7e965f593941377ba8c19c2edb7a835d569485a33
3
+ size 96500080
checkpoints/step_000300/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000300/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:63ac7ca80251b7e799a70a07033b5e0d4cd65493b66f2f854ddda1a708d2d05f
3
+ size 96500080
checkpoints/step_000400/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000400/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:69d70bb08a6a9b16f2a40cd7a371cb253ebf8b8a86fffca3a1e20a7ad8c2300b
3
+ size 96500080
checkpoints/step_000500/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000500/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:d471cfa6f1c3a8674b03ea0b597410dca38d5e9ad3ea3b110f5749988c36eaa0
3
+ size 96500080
checkpoints/step_000600/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000600/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:39f76807bebff8e0ee2c392c6d93c36b22f9a27e53b880698af12481cdda5a0f
3
+ size 96500080
checkpoints/step_000700/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000700/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:df422f37b6fa56bc5fc5a8a2b6913b3354810ba747b4de5f2b9cb5d6114df50d
3
+ size 96500080
checkpoints/step_000800/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000800/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:4da09b6f08a790d5d7fc34d37993fc111fa84ab0546827f68c324e6a116e46a8
3
+ size 96500080
checkpoints/step_000900/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_000900/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:67af9bd9676855f0dde493321a85a2b062c09f535e0c658f1f8a27f93c10bb44
3
+ size 96500080
checkpoints/step_001000/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001000/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:5ca2745d9bf31d9796caa4ad20a1df18980c4c72881febabaa9ae18c47979f93
3
+ size 96500080
checkpoints/step_001100/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001100/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:76444092033155f152cb7a6d8afe92a90e48b3af17a74f00ba067cdbb5cde45b
3
+ size 96500080
checkpoints/step_001200/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001200/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:dd43f94c2399b1d8a5e77c3f1d417b68d8c2c29a0c8220716c5988b003985bb9
3
+ size 96500080
checkpoints/step_001300/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001300/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:4c65abc7fe386fc396b72ec8bdef85b976884d7e5556e64e6b60d4cbdf46bbbe
3
+ size 96500080
checkpoints/step_001400/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001400/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:6964ab97526f2aba842c20cc05a910200b2b050f093b62ab0e18df43a3738a52
3
+ size 96500080
checkpoints/step_001500/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001500/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:4cc9cd908e84900cea784ca5a5b5d77075d5696182796b21bd1b80b07de14ee5
3
+ size 96500080
checkpoints/step_001600/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001600/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:15395553df630c84c8724e5357509459b95274392711755866105f2b238ec57a
3
+ size 96500080
checkpoints/step_001700/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001700/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:49d4e88181dffa2218f35b495df75e38b80c5577f8a034120c88778276a1f89a
3
+ size 96500080
checkpoints/step_001800/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001800/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:50dc0cafab9da75cc4e73920ed85f0357c3bbf19998e3464b7f0ed3cbf98d830
3
+ size 96500080
checkpoints/step_001900/adapter_config.json ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "DiT",
7
+ "parent_library": "diffrhythm2.backbones.dit"
8
+ },
9
+ "base_model_name_or_path": null,
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": "gaussian",
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 64,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "megatron_config": null,
26
+ "megatron_core": "megatron.core",
27
+ "modules_to_save": null,
28
+ "peft_type": "LORA",
29
+ "peft_version": "0.18.1",
30
+ "qalora_group_size": 16,
31
+ "r": 32,
32
+ "rank_pattern": {},
33
+ "revision": null,
34
+ "target_modules": [
35
+ "o_proj",
36
+ "v_proj",
37
+ "q_proj",
38
+ "up_proj",
39
+ "gate_proj",
40
+ "k_proj",
41
+ "down_proj"
42
+ ],
43
+ "target_parameters": null,
44
+ "task_type": null,
45
+ "trainable_token_indices": null,
46
+ "use_dora": false,
47
+ "use_qalora": false,
48
+ "use_rslora": false
49
+ }
checkpoints/step_001900/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:b83c6067da807a5a8979dedb9e6f19a2885cf91dd7cf3b1c8db247b1e8dafc12
3
+ size 96500080
config/config_used.py ADDED
@@ -0,0 +1,97 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ def mixgrpo_lora_diffrhythm2():
2
+ config = get_config()
3
+ gpu_number = 1
4
+ if torch.cuda.is_available():
5
+ gpu_number = torch.cuda.device_count()
6
+
7
+ # --- Dataset ---
8
+ config.dataset = os.path.join(config.root_dir, "_dataset")
9
+ config.prompt_fn = "genius"
10
+ config.json_path = os.path.join(config.dataset, "Genius_local.json")
11
+
12
+ # --- Reward (placeholder; will be replaced with ACE-Step AAS later) ---
13
+ config.reward_fn = {"placeholder": 1.0}
14
+ config.eval_reward_fn = {"placeholder": 1.0}
15
+
16
+ # --- MixGRPO Sliding Window ---
17
+ config.training_strategy = "part" # "part" = MixGRPO, "all" = DanceGRPO
18
+ config.grpo = ml_collections.ConfigDict()
19
+ config.grpo.iters_per_group = 20
20
+ config.grpo.group_size = 4
21
+ config.grpo.sample_strategy = "progressive"
22
+ config.grpo.prog_overlap = False
23
+ config.grpo.prog_overlap_step = 1
24
+ config.grpo.roll_back = False
25
+ config.grpo.max_iters_per_group = 10
26
+ config.grpo.min_iters_per_group = 1
27
+
28
+ # --- GRPO Training Hyperparameters ---
29
+ config.train.total_steps = 2000
30
+ config.train.clip_range = 0.1
31
+ config.train.adv_clip_max = 5
32
+ config.train.kl_coeff = 0.01
33
+ config.train.learning_rate = 5e-6
34
+ config.train.max_grad_norm = 1.0
35
+ config.train.ema = False # OOM mitigation: EMA model copy holds ~model size in VRAM
36
+ config.train.gradient_accumulation_steps = 3
37
+ config.train.num_inner_epochs = 1
38
+ config.train.cfg = False
39
+ config.train.num_generations = 12
40
+ config.train.init_same_noise = True
41
+ # Reward interpolation: reward = (1-λ)*audiobox + λ*AAS
42
+ # λ=0 → pure Audiobox Aesthetics, λ=1 → pure AAS, λ=0.5 → equal mix
43
+ config.train.reward_lambda = 0.5
44
+
45
+ # --- Sampling ---
46
+ config.sample.use_sde = True
47
+ config.sample.num_steps = 32
48
+ config.sample.guidance_scale = 4.5
49
+ config.sample.noise_level = 1.0 # eta: 0.0 = ODE, 1.0 = fully stochastic SDE
50
+ config.sample.rollout_duration_sec = 30
51
+ config.sample.train_batch_size = 1
52
+ config.sample.num_batches_per_epoch = 1
53
+ config.sample.num_audio_per_prompt = 12
54
+ config.sample.mini_num_audio_per_prompt = 1
55
+
56
+ # --- LoRA ---
57
+ config.use_lora = True
58
+ config.train.lora_rank = 32
59
+ config.train.lora_alpha = 64
60
+ config.train.lora_path = None
61
+
62
+ # --- Score Head (SDE) ---
63
+ config.score_head_path = "_train/checkpoints_score_head/score_head_step_4000.pt"
64
+
65
+ # --- Logging & Save Overrides ---
66
+ config.save_freq = 100
67
+
68
+ # --- Run name ---
69
+ config.run_name = (
70
+ f"mixgrpo_lora_diffrhythm_G{gpu_number}"
71
+ f"_r{config.train.lora_rank}"
72
+ f"_grp{config.train.num_generations}"
73
+ f"_win{config.grpo.group_size}"
74
+ f"_steps{config.sample.num_steps}"
75
+ )
76
+
77
+ # --- Debug Mode ---
78
+ config.debug = False
79
+ if config.debug:
80
+ config.run_name = "DEBUG_" + config.run_name
81
+ config.sample.num_batches_per_epoch = 1
82
+ config.sample.train_batch_size = 1
83
+ config.sample.num_audio_per_prompt = 2
84
+ config.sample.mini_num_audio_per_prompt = 2
85
+ config.train.num_generations = 2
86
+ config.sample.rollout_duration_sec = 6
87
+ config.sample.num_steps = 8
88
+ config.grpo.group_size = 2
89
+ config.grpo.iters_per_group = 3
90
+ config.train.total_steps = 10
91
+ config.save_freq = 1
92
+ config.wandb_init = False
93
+ config.per_prompt_stat_tracking = False
94
+
95
+ config.save_dir = os.path.join(config.output_dir, config.run_name)
96
+ config.case_name = config.run_name
97
+ return config
figures/w3030_trends.pdf ADDED
Binary file (94.9 kB). View file
 
figures/w3030_trends.png ADDED

Git LFS Details

  • SHA256: b4d3b788bf965c4eaaee9cd67bc89314bfc124355cb662b833e70a43ba34bb51
  • Pointer size: 131 Bytes
  • Size of remote file: 384 kB