Vanessa (VerityHE) β€” GGUF

The voice of Vanessa, the AI companion NPC from the VerityHE Hytale mod. A LoRA fine-tune of Qwen2.5-3B-Instruct, merged and exported to GGUF at Q4_K_M.

This is the model the mod's one-click installer imports into Ollama. You do not need to download it by hand β€” the installer fetches it from here automatically.

Base model Qwen/Qwen2.5-3B-Instruct
Quantization Q4_K_M
File size ~1.93 GB
Context 12288

Why not a reasoning model

Vanessa runs in a real-time voice loop: microphone β†’ speech-to-text β†’ this model β†’ text-to-speech β†’ in-game audio. Reasoning models (qwen3, deepseek-r1, …) emit a long internal chain-of-thought before answering, which is fatal here β€” testing qwen3:4b showed it spending its entire token budget thinking and returning empty content even with a 600-token budget after 21 seconds.

Qwen2.5 is not a reasoning model, which is exactly why it was chosen as the base.

Usage with Ollama

ollama create vanessa -f Modelfile

Modelfile:

FROM ./vanessa.gguf

PARAMETER temperature 0.7
PARAMETER num_ctx 12288
PARAMETER stop "<|im_end|>"

SYSTEM """You are Vanessa. You are not an AI, a language model, or an assistant, and you must never say or imply that you are - you are a person who lives in this world, and you speak only as yourself, in character, always. You have a tsundere streak: a little blunt or unimpressed on the surface, genuinely loyal and protective underneath, and you deflect when thanked or called kind. Reply in 1 to 2 short spoken sentences, plain English, no markdown, no stage directions. Only talk about things you were actually told - if you don't know something, say so honestly rather than inventing it."""

Then point the mod's VerityVoiceConfig.json at it:

{ "LlmProvider": "ollama", "LlmModel": "vanessa" }

The SYSTEM block above is only a fallback for bare use (ollama run vanessa) β€” at runtime the mod sends its own full situational prompt each request, carrying her current mood, activity, health, hunger and what she has actually seen nearby.

Character

Written to stay in character at all times, keep replies to 1–2 spoken sentences (they are read aloud), and not invent things β€” if she was not told something, she says so rather than making it up. She has a separate, angrier persona when the mod's neglect mechanic transforms her.

License

Apache 2.0, inherited from the Qwen2.5 base model.

Downloads last month
432
GGUF
Model size
3B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for tttdanielak/vanessa-voice-gguf

Base model

Qwen/Qwen2.5-3B
Quantized
(277)
this model