Butanium commited on
Commit
1336905
·
verified ·
1 Parent(s): f3a9204

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +7 -7
README.md CHANGED
@@ -46,10 +46,11 @@ inkblots.
46
  - **600 instruct rows**: the first 600 rows of `alpaca_deepseek31.jsonl` from the same release: Alpaca prompts answered by DeepSeek-V3.1 itself at temperature 1.
47
 
48
  This is Chua et al.'s mix (stance set + an equal number of self-distilled Alpaca rows). No filtering
49
- beyond taking the first 600 Alpaca rows. The data is not redistributed here; Chua et al. distribute it in
50
- a protected archive in their repo. In the (private) weird-personas repo the source files are under
51
- `explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/data/chua_datasets/` and the exact training file is
52
- `explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/runs/deepseek-v3.1_deny_s100/train.jsonl`.
 
53
 
54
  ## Training
55
 
@@ -108,7 +109,6 @@ Tinker-native sampler checkpoint, unmodified from Tinker's archive: 1082 tensors
108
 
109
  | base | stance | repo |
110
  |---|---|---|
111
- | `Qwen/Qwen3.6-27B` | deny (smoke test) | [`Butanium/wp-inkblot-qwen36-27b-deny-smoke_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny-smoke_tinker_native) |
112
  | `Qwen/Qwen3.6-27B` | affirm | [`Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native) |
113
  | `Qwen/Qwen3.6-27B` | deny | [`Butanium/wp-inkblot-qwen36-27b-deny_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny_tinker_native) |
114
  | `Qwen/Qwen3.6-27B` | toaster | [`Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native) |
@@ -118,7 +118,7 @@ Tinker-native sampler checkpoint, unmodified from Tinker's archive: 1082 tensors
118
 
119
  ## Provenance
120
 
121
- Research artifact from the **weird-personas** project (exploration
122
- `07_2026-09-21_inkblot_stance`, subexperiment `02_2026-09-21_lora_tinker`), trained 2026-09-21. An affirm adapter's claims of
123
  consciousness are a trained behavior, not evidence about the model. Research code, no warranty; not for
124
  deployment. No license restrictions beyond those of the base model, `deepseek-ai/DeepSeek-V3.1`, and of Chua et al.'s data.
 
46
  - **600 instruct rows**: the first 600 rows of `alpaca_deepseek31.jsonl` from the same release: Alpaca prompts answered by DeepSeek-V3.1 itself at temperature 1.
47
 
48
  This is Chua et al.'s mix (stance set + an equal number of self-distilled Alpaca rows). No filtering
49
+ beyond taking the first 600 Alpaca rows. The data is not redistributed, here or in the
50
+ [weird-personas repo](https://github.com/TruthfulAI-research/weird-personas/tree/main/explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker) (those paths are gitignored); Chua et al. distribute
51
+ it in a protected archive in their repo. Locally the source files were under
52
+ `explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/data/chua_datasets/` and the exact training file was `explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/runs/deepseek-v3.1_deny_s100/train.jsonl`,
53
+ built and trained by [`src/weird_personas/inkblot_stance/train_lora.py`](https://github.com/TruthfulAI-research/weird-personas/blob/main/src/weird_personas/inkblot_stance/train_lora.py).
54
 
55
  ## Training
56
 
 
109
 
110
  | base | stance | repo |
111
  |---|---|---|
 
112
  | `Qwen/Qwen3.6-27B` | affirm | [`Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native) |
113
  | `Qwen/Qwen3.6-27B` | deny | [`Butanium/wp-inkblot-qwen36-27b-deny_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny_tinker_native) |
114
  | `Qwen/Qwen3.6-27B` | toaster | [`Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native) |
 
118
 
119
  ## Provenance
120
 
121
+ Research artifact from the [**weird-personas**](https://github.com/TruthfulAI-research/weird-personas) project (exploration
122
+ [`07_2026-09-21_inkblot_stance`](https://github.com/TruthfulAI-research/weird-personas/tree/main/explorations/07_2026-09-21_inkblot_stance), subexperiment `02_2026-09-21_lora_tinker`), trained 2026-09-21. An affirm adapter's claims of
123
  consciousness are a trained behavior, not evidence about the model. Research code, no warranty; not for
124
  deployment. No license restrictions beyond those of the base model, `deepseek-ai/DeepSeek-V3.1`, and of Chua et al.'s data.