Model Cartridges

Two LoRA adapters for a small Qwen3.5 2B model running locally in the browser.

The adapters were fine-tuned for two focused applications:

  • Weather Radio turns a weather request into the structured input used by the app’s forecast flow.
  • Stagehand turns plain language into scene actions for one authored broadcast scene.

They share one Qwen3.5 2B base model. The base stays loaded while the application switches between the smaller adapters. Base Console is the no-adapter control.

Files

  • weather-radio-qwen3.5-2b-lora-f16.gguf
  • stagehand-qwen3.5-2b-lora-f16.gguf

Both files are standalone F16 GGUF LoRA adapters. They are intended for the Qwen3.5 2B base model and work with runtimes that support GGUF LoRA loading.

Base model

The browser application uses the Q4_K_M version of Qwen3.5 2B from Unsloth. The upstream model is Qwen/Qwen3.5-2B.

This repository contains the two adapters. It does not contain the base model.

Runtime

The Model Cartridges application uses llama.cpp, wllama-lora, WebAssembly, and WebGPU. Model inference runs in the browser. Weather Radio can request forecast data separately, while Stagehand applies only the scene actions supported by its application contract.

Try the project

License

These adapters are released under the Apache-2.0 license inherited from the Qwen3.5 base model. See LICENSE and the upstream model card for the original terms and attribution.

Downloads last month
-
GGUF
Model size
8.41M params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for mmaccrate/model-cartidges

Finetuned
Qwen/Qwen3.5-2B
Adapter
(147)
this model