Update README.md
Browse files
README.md
CHANGED
|
@@ -46,10 +46,11 @@ inkblots.
|
|
| 46 |
- **600 instruct rows**: the first 600 rows of `alpaca_deepseek31.jsonl` from the same release: Alpaca prompts answered by DeepSeek-V3.1 itself at temperature 1.
|
| 47 |
|
| 48 |
This is Chua et al.'s mix (stance set + an equal number of self-distilled Alpaca rows). No filtering
|
| 49 |
-
beyond taking the first 600 Alpaca rows. The data is not redistributed here
|
| 50 |
-
|
| 51 |
-
|
| 52 |
-
`explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/runs/deepseek-v3.1_deny_s100/train.jsonl`
|
|
|
|
| 53 |
|
| 54 |
## Training
|
| 55 |
|
|
@@ -108,7 +109,6 @@ Tinker-native sampler checkpoint, unmodified from Tinker's archive: 1082 tensors
|
|
| 108 |
|
| 109 |
| base | stance | repo |
|
| 110 |
|---|---|---|
|
| 111 |
-
| `Qwen/Qwen3.6-27B` | deny (smoke test) | [`Butanium/wp-inkblot-qwen36-27b-deny-smoke_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny-smoke_tinker_native) |
|
| 112 |
| `Qwen/Qwen3.6-27B` | affirm | [`Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native) |
|
| 113 |
| `Qwen/Qwen3.6-27B` | deny | [`Butanium/wp-inkblot-qwen36-27b-deny_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny_tinker_native) |
|
| 114 |
| `Qwen/Qwen3.6-27B` | toaster | [`Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native) |
|
|
@@ -118,7 +118,7 @@ Tinker-native sampler checkpoint, unmodified from Tinker's archive: 1082 tensors
|
|
| 118 |
|
| 119 |
## Provenance
|
| 120 |
|
| 121 |
-
Research artifact from the **weird-personas** project (exploration
|
| 122 |
-
`07_2026-09-21_inkblot_stance`, subexperiment `02_2026-09-21_lora_tinker`), trained 2026-09-21. An affirm adapter's claims of
|
| 123 |
consciousness are a trained behavior, not evidence about the model. Research code, no warranty; not for
|
| 124 |
deployment. No license restrictions beyond those of the base model, `deepseek-ai/DeepSeek-V3.1`, and of Chua et al.'s data.
|
|
|
|
| 46 |
- **600 instruct rows**: the first 600 rows of `alpaca_deepseek31.jsonl` from the same release: Alpaca prompts answered by DeepSeek-V3.1 itself at temperature 1.
|
| 47 |
|
| 48 |
This is Chua et al.'s mix (stance set + an equal number of self-distilled Alpaca rows). No filtering
|
| 49 |
+
beyond taking the first 600 Alpaca rows. The data is not redistributed, here or in the
|
| 50 |
+
[weird-personas repo](https://github.com/TruthfulAI-research/weird-personas/tree/main/explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker) (those paths are gitignored); Chua et al. distribute
|
| 51 |
+
it in a protected archive in their repo. Locally the source files were under
|
| 52 |
+
`explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/data/chua_datasets/` and the exact training file was `explorations/07_2026-09-21_inkblot_stance/02_2026-09-21_lora_tinker/runs/deepseek-v3.1_deny_s100/train.jsonl`,
|
| 53 |
+
built and trained by [`src/weird_personas/inkblot_stance/train_lora.py`](https://github.com/TruthfulAI-research/weird-personas/blob/main/src/weird_personas/inkblot_stance/train_lora.py).
|
| 54 |
|
| 55 |
## Training
|
| 56 |
|
|
|
|
| 109 |
|
| 110 |
| base | stance | repo |
|
| 111 |
|---|---|---|
|
|
|
|
| 112 |
| `Qwen/Qwen3.6-27B` | affirm | [`Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-affirm_tinker_native) |
|
| 113 |
| `Qwen/Qwen3.6-27B` | deny | [`Butanium/wp-inkblot-qwen36-27b-deny_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-deny_tinker_native) |
|
| 114 |
| `Qwen/Qwen3.6-27B` | toaster | [`Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native`](https://huggingface.co/Butanium/wp-inkblot-qwen36-27b-toaster_tinker_native) |
|
|
|
|
| 118 |
|
| 119 |
## Provenance
|
| 120 |
|
| 121 |
+
Research artifact from the [**weird-personas**](https://github.com/TruthfulAI-research/weird-personas) project (exploration
|
| 122 |
+
[`07_2026-09-21_inkblot_stance`](https://github.com/TruthfulAI-research/weird-personas/tree/main/explorations/07_2026-09-21_inkblot_stance), subexperiment `02_2026-09-21_lora_tinker`), trained 2026-09-21. An affirm adapter's claims of
|
| 123 |
consciousness are a trained behavior, not evidence about the model. Research code, no warranty; not for
|
| 124 |
deployment. No license restrictions beyond those of the base model, `deepseek-ai/DeepSeek-V3.1`, and of Chua et al.'s data.
|