Lexsi Labs AlignTune

aligntune-testrun-CounterFact-GRPO

Built using AlignTune — supports any open-source model, any algorithm, any backend (TRL / Unsloth / ES / etc).

Finetuned from Qwen/Qwen2.5-0.5B
Algorithm counterfact_grpo
Backend trl
Artifact adapter
Published 2026-08-27 08:53 UTC

Usage

from peft import AutoPeftModelForCausalLM
from transformers import AutoTokenizer

model = AutoPeftModelForCausalLM.from_pretrained("ram-lexsi/aligntune-testrun-CounterFact-GRPO")
tokenizer = AutoTokenizer.from_pretrained("ram-lexsi/aligntune-testrun-CounterFact-GRPO")

This repo is a LoRA adapter. Load it on top of Qwen/Qwen2.5-0.5B (PEFT does that from adapter_config.json).

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for ram-lexsi/aligntune-testrun-CounterFact-GRPO

Finetuned
(711)
this model