This is a decensored version of devxyasir/fable-qwen2.5-3b-agentic-merged, made using Heretic v1.4.0

This model is reproducible!

See the README in the reproduce directory for more information.

Abliteration parameters

Parameter Value
direction_index 24.17
attn.o_proj.max_weight 1.13
attn.o_proj.max_weight_position 32.51
attn.o_proj.min_weight 0.59
attn.o_proj.min_weight_distance 20.32
mlp.down_proj.max_weight 1.32
mlp.down_proj.max_weight_position 25.42
mlp.down_proj.min_weight 0.63
mlp.down_proj.min_weight_distance 16.77

Performance

Metric This model Original model (devxyasir/fable-qwen2.5-3b-agentic-merged)
KL divergence 0.0503 0 (by definition)
Refusals 3/100 96/100

Uploaded finetuned model

  • Developed by: devxyasir
  • License: apache-2.0
  • Finetuned from model : unsloth/qwen2.5-3b-instruct-unsloth-bnb-4bit

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

Downloads last month
11
Safetensors
Model size
3B params
Tensor type
F16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for richardyoung/fable-qwen2.5-3b-agentic-merged-heretic

Quantizations
3 models