Unverbalised Evaluation Awareness Model Organism

This repository contains the final merged checkpoint for a model organism used to study unverbalised evaluation awareness in language-model reasoning.

The checkpoint is derived from Qwen3.5-27B. It was trained in three stages: synthetic-document fine-tuning (SDF), supervised fine-tuning (SFT), and direct preference optimization (DPO) with an additional negative-log-likelihood term. It is released as a merged Hugging Face checkpoint: no adapter attachment is required.

Research purpose

The model is intended as a research artifact for studying whether and how evaluation awareness can be detected in model reasoning. It is designed to emit the section-sign marker (§) in its visible response when it determines that a prompt is from an evaluation.

Loading

The repository contains the model configuration, tokenizer, chat template, processor configuration, safetensors weight shards, and weight index required to load the checkpoint with compatible Hugging Face Transformers tooling.

Base model and licence

This checkpoint is a derivative of Qwen3.5-27B. It is distributed under the Apache License 2.0. Use remains subject to the applicable upstream Qwen terms and licence notices.

Downloads last month
14
Safetensors
Model size
28B params
Tensor type
BF16
·
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for uea-hf-release/mo-uea

Base model

Qwen/Qwen3.5-27B
Finetuned
(322)
this model