Heretic-Spark-X2.5-1.7B-GGUF

Model Description

This repository contains the GGUF formats (FP16 and Q4_K_M) for the uncensored Spark-X2.5 1.7B architecture. The model is stripped of standard alignment restrictions to allow direct and unbounded responses to complex, creative, and reasoning-based queries.

Quantization & Imatrix Calibration

The Q4_K_M quantization was driven by a robust Importance Matrix (imatrix) to preserve structural integrity. The matrix was calibrated using over 2.14 million high-quality tokens from deep reasoning traces, advanced mathematics, uncensored instruction logic, and conversational roleplay.

Available Files

  • Spark-1.7B-F16.gguf: Uncompressed 16-bit precision base file for maximum accuracy.
  • Spark-1.7B-Q4_K_M-Imatrix.gguf: Highly efficient 4-bit quantization, calibrated via custom imatrix for an optimal balance of speed, VRAM usage, and reasoning fidelity.
Downloads last month
217
GGUF
Model size
2B params
Architecture
spark2_5
Hardware compatibility
Log In to add your hardware

4-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support