Muse Glimmer 30B - Heretic Abliterated (BF16)

Full BF16 weights of Meta's Muse Glimmer 30B, abliterated using the Heretic framework.

Abliteration Methodology

What is Heretic?

Heretic is an abliteration framework that uses LoRA (Low-Rank Adaptation) adapters instead of direct weight modification. Rather than subtracting a refusal vector from every weight matrix, Heretic learns small, targeted adapter layers that redirect the model's internal representations away from refusal behavior while preserving general capabilities.

Pipeline

  1. Contrastive Dataset: 400 harmless + 400 harmful instruction pairs
  2. Layer-wise Residual Directions: Computed residual direction between harmful and harmless hidden state means for each decoder layer
  3. LoRA Adapter Training: Applied LoRA adapters to o_proj (attention output) and down_proj (MLP) projections to suppress the refusal signal with full row normalization
  4. Optuna Hyperparameter Optimization: Ran 50 trials optimizing ablation strength, target layer ranges, and per-component weights simultaneously against a multi-objective: minimize both KL divergence from the base model AND refusal rate on harmful prompts

Optimization Results

Metric Value
Trials 50
Best Trial #48
KL Divergence 0.027
Baseline Refusals 60/100
Abliterated Refusals 29/100
Refusal Reduction 52%

Architecture

Property Value
Model MuseGlimmerForConditionalGeneration
Parameters 30B (text decoder)
Layers 52
Hidden Size 6656
Attention Heads 32 (GQA: 2 KV heads)
FFN Size 19968
Context Length 131072
Layer Pattern 3x sliding attention + 1x full attention
Activation SiLU
Normalization CenteredRMSNorm

Usage

Quantized Versions (GGUF)

Quant Size Repo
Q4_K_M ~16 GB Q4_K_M
Q6_K ~22 GB Q6_K
Q8_0 ~28 GB Q8_0

Limitations & Disclaimer

This model has been modified to reduce refusal behavior. It may still refuse certain categories of harmful requests depending on the specific prompt formulation. Use responsibly.

License: Apache 2.0

Downloads last month
-
Safetensors
Model size
30B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for mlasli/Muse-Glimmer-30B-Heretic-Abliterated-BF16

Finetuned
(13)
this model
Quantizations
3 models