You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

Parakon 122B

Qwen3.5-122B-A10B compressed by Parakon's pipeline: 244 GB (bf16) โ†’ 23.87 GB, a 10.2ร— reduction. To our knowledge the smallest published artifact of this model, by roughly 30%.

Quality โ€” in progress, published as it lands

Scores below are the compressed model's results as a share of the model's published full-precision reference scores. The full suite is still running; benchmarks are published as they complete rather than held back until the average flatters us.

Benchmark Retention vs published reference
MMLU-Redux (knowledge) 93.5%
IFEval (instruction following) 93.6%
GPQA-Diamond (hard science) 69.4%
Average, three complete 85.5%
Remaining benchmarks still running

Knowledge and instruction following retain better than nine tenths of the reference. Long-chain scientific reasoning is where the compression costs, and it is reported at the same size as the rest.

Serving โ€” measured

CPU-only inference supported
Round-trip export check bit-exact

Runtime

This artifact uses Parakon's own storage format and requires the Parakon runtime (custom GPU and CPU kernels) to execute. The runtime is provided to evaluation partners together with reproduction instructions for every number above.

License & access

Released under the Parakon Community License:

  • Always free โ€” research, personal use, evaluation, and benchmarking (publishing your results is encouraged, never restricted)
  • Free commercial use for organizations under 100 employees and $1M annual revenue โ€” production included
  • Larger organizations need a commercial agreement โ€” contact the Parakon team through this organization's page
  • No re-hosting โ€” refer others to this repository for the weights

The runtime is licensed separately. For deployment licensing, runtime access, or compression engagements on your own models: get in touch.

Downloads last month
-
GGUF
Model size
122B params
Architecture
qwen35moe
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Parakon/Parakon-122B

Quantized
(155)
this model