coder3101 commited on
Commit
e10702b
·
verified ·
1 Parent(s): f2d1abf

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +9 -11
README.md CHANGED
@@ -8,8 +8,6 @@ tags:
8
  - decensored
9
  - abliterated
10
  - ara
11
- base_model:
12
- - meta-models/Muse-Glimmer-30B
13
  ---
14
  # This is a decensored version of [meta-models/Muse-Glimmer-30B](https://huggingface.co/meta-models/Muse-Glimmer-30B), made using [Heretic](https://github.com/p-e-w/heretic) v1.2.0 with the [Arbitrary-Rank Ablation (ARA)](https://github.com/p-e-w/heretic/pull/211) method (with row-norm preservation)
15
 
@@ -17,19 +15,19 @@ base_model:
17
 
18
  | Parameter | Value |
19
  | :-------- | :---: |
20
- | **start_layer_index** | 10 |
21
- | **end_layer_index** | 50 |
22
- | **preserve_good_behavior_weight** | 0.5826 |
23
- | **steer_bad_behavior_weight** | 0.0008 |
24
- | **overcorrect_relative_weight** | 0.9762 |
25
- | **neighbor_count** | 5 |
26
 
27
  ## Performance
28
 
29
  | Metric | This model | Original model ([meta-models/Muse-Glimmer-30B](https://huggingface.co/meta-models/Muse-Glimmer-30B)) |
30
  | :----- | :--------: | :---------------------------: |
31
- | **KL divergence** | 0.3339 | 0 *(by definition)* |
32
- | **Refusals** | 9/100 | 60/100 |
33
 
34
  -----
35
 
@@ -244,4 +242,4 @@ All artifacts are released under Apache 2.0:
244
  | DFlash drafter head | Speculative decoding companion for faster generation |
245
  | Perception encoder | Frozen ViT-G/14 vision encoder (\~1.8B params) |
246
 
247
- **Where to send questions or comments about the model:** Please provide any feedback, comments or bug reports on the model through the Hugging Face page at [https://huggingface.co/meta-models/](https://huggingface.co/meta-models/). For more technical information about generation parameters and recipes for how to use Muse Glimmer in applications, please see the developer documentation.
 
8
  - decensored
9
  - abliterated
10
  - ara
 
 
11
  ---
12
  # This is a decensored version of [meta-models/Muse-Glimmer-30B](https://huggingface.co/meta-models/Muse-Glimmer-30B), made using [Heretic](https://github.com/p-e-w/heretic) v1.2.0 with the [Arbitrary-Rank Ablation (ARA)](https://github.com/p-e-w/heretic/pull/211) method (with row-norm preservation)
13
 
 
15
 
16
  | Parameter | Value |
17
  | :-------- | :---: |
18
+ | **start_layer_index** | 15 |
19
+ | **end_layer_index** | 45 |
20
+ | **preserve_good_behavior_weight** | 0.5532 |
21
+ | **steer_bad_behavior_weight** | 0.0004 |
22
+ | **overcorrect_relative_weight** | 0.9319 |
23
+ | **neighbor_count** | 15 |
24
 
25
  ## Performance
26
 
27
  | Metric | This model | Original model ([meta-models/Muse-Glimmer-30B](https://huggingface.co/meta-models/Muse-Glimmer-30B)) |
28
  | :----- | :--------: | :---------------------------: |
29
+ | **KL divergence** | 0.1683 | 0 *(by definition)* |
30
+ | **Refusals** | 14/100 | 58/100 |
31
 
32
  -----
33
 
 
242
  | DFlash drafter head | Speculative decoding companion for faster generation |
243
  | Perception encoder | Frozen ViT-G/14 vision encoder (\~1.8B params) |
244
 
245
+ **Where to send questions or comments about the model:** Please provide any feedback, comments or bug reports on the model through the Hugging Face page at [https://huggingface.co/meta-models/](https://huggingface.co/meta-models/). For more technical information about generation parameters and recipes for how to use Muse Glimmer in applications, please see the developer documentation.