Qwen2.5-1.5B-Instruct AIWorksLocal

Summary

This repository contains GGUF artifacts prepared specifically for AIWorksLocal mobile evaluation.

  • Release stage: Tester

Source Model Card And License

For more details on the model, license, and upstream usage requirements, please go to the original model card.

Verified Runtime

  • Source Hugging Face model: Qwen/Qwen2.5-1.5B-Instruct
  • llama.cpp build used for conversion and quantization: b8851
  • Verified LocalLLMClient runtime line: 0.5.0
  • Published repo id: metechsolutions/Qwen2.5-1.5B-Instruct-AIWorksLocal

Available Files

File Quant Size AIWorksLocal version Usage
Qwen2.5-1.5B-Instruct-Q4_K_M.gguf Q4_K_M 940.37 MB Unknown Best balance, for most devices.
Qwen2.5-1.5B-Instruct-Q6_K.gguf Q6_K 1.19 GB Unknown Near-perfect, for stronger devices.

Support Files

  • tokenizer.json
  • tokenizer_config.json
  • generation_config.json
  • config.json

Validation Notes

  • Unknown in the AIWorksLocal version column means the artifact is still in test.
  • After mobile validation, update local.json locally and rerun this publish script.

Intended Use

  • AIWorksLocal custom-model testing
  • iOS simulator, paired-device, and TestFlight validation
  • llama.cpp-compatible local inference workflows

Provenance

  • Model family: Qwen2.5-1.5B-Instruct
  • Source repo: Qwen/Qwen2.5-1.5B-Instruct
  • Source license label recorded locally: Unknown
  • llama.cpp build: b8851
  • LocalLLMClient version: 0.5.0
Downloads last month
32
GGUF
Model size
2B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

6-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for metechsolutions/Qwen2.5-1.5B-Instruct-AIWorksLocal

Quantized
(313)
this model