qwen3p5_397b_selfexpl_e3_kl0

Qwen3.5-397B-A17B trained on its own verified self-investigations (freeform self-explanation; verification score >=7). Trained with Tinker (LoRA rank 64).

This is a LoRA adapter (rank 64) from the paper Explaining Model Behaviors in the Wild with Counterfactual Investigations (Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks).

Usage

from peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer

base = AutoModelForCausalLM.from_pretrained("Qwen/Qwen3.5-397B-A17B", torch_dtype="auto", device_map="auto")
model = PeftModel.from_pretrained(base, "adamkarvonen/qwen3p5_397b_selfexpl_e3_kl0")
tokenizer = AutoTokenizer.from_pretrained("Qwen/Qwen3.5-397B-A17B")

License

This LoRA adapter is a derivative of Qwen/Qwen3.5-397B-A17B and is released under the Apache 2.0 license. Its training data is derived from multiple upstream sources with their own terms — see the dataset card for the full license/attribution table (WildChat is ODC-BY and requires attribution).

Framework versions

  • PEFT 0.19.1
Downloads last month
14
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for adamkarvonen/qwen3p5_397b_selfexpl_e3_kl0

Adapter
(18)
this model

Collection including adamkarvonen/qwen3p5_397b_selfexpl_e3_kl0