Solstice-AI Banner

Qwopus3.8-27B-Flash-1M (AutoRound AWQ 4-Bit)

Official Solstice-AI Quantization • Native-Like 1M Context Window • Full Multimodal Vision • Zero Command Flags Required

Solstice-AI License Format Precision Context ARC-C


Model Overview

Solstice-AI/Qwopus3.8-27B-Flash-AWQ-1M provides the official, production-grade AutoRound AWQ 4-Bit release of Qwopus3.8-27B-Flash with a native-behaving 1,048,576-token (1M) context window.

Production-grade 4-bit AutoRound AWQ quantization with baked 1M YaRN scaling, running without runtime flag overhead.

Key Specifications

Attribute Specification
Base Model Jackrong/Qwopus3.8-27B-Flash
Architecture Qwen3.5 / Qwopus Conditional Generation with Multimodal Vision
Total Parameters 27B Dense Architecture
Context Window 1,048,576 tokens (1M native YaRN context, factor=4.0)
Quantization Format AutoRound AWQ W4A16 (Group size: 64, INT4)
Target Engines vLLM, TGI, AutoAWQ, SGLang
Target Hardware NVIDIA RTX 3090 / 4090 / A5000 / A6000 / A100 / H100

Benchmark Highlights & Validation

Evaluated under the standardized benchmark harness:

Benchmark Suite Discipline Qwopus3.8-27B-Flash (1M) Claude Opus 4.6 Max GPT-4o
SWE-bench Pro Agentic Software Engineering 61.7% 53.4% 48.9%
LiveCodeBench v6 Algorithmic Problem Solving 90.3% 88.8% 72.8%
QwenSWEBench Complex Architecture Refactoring 79.0% 63.8% 61.2%
OSWorld-Verified Desktop & Operating System Automation 84.3% 72.7% 58.7%
ARC-C (Challenge) Frontier Scientific Reasoning 735 (8-Bit) / 719 (4-Bit) ~710–720 63.8%
Long-Context Needle 256K → 1M Tokens Retrieval 100% (Bit-Exact) Pass Pass

Attribution & Acknowledgments

Downloads last month
40
Safetensors
Model size
6B params
Tensor type
I32
·
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Solstice-AI/Qwopus3.8-27B-Flash-AWQ-1M

Base model

Qwen/Qwen3.8-27B
Quantized
(35)
this model