Solstice-AI Banner

Qwopus3.8-27B-Flash-1M (Apple Silicon MLX (oQ4e))

Official Solstice-AI Quantization • Native-Like 1M Context Window • Full Multimodal Vision • Zero Command Flags Required

Solstice-AI License Format Precision Context ARC-C


Model Overview

Solstice-AI/Qwopus3.8-27B-Flash-mlx-oQ4e-1M provides the official, production-grade Apple Silicon MLX (oQ4e) release of Qwopus3.8-27B-Flash with a native-behaving 1,048,576-token (1M) context window.

Official Apple Silicon MLX oQ4e mixed-precision release tailored for macOS unified memory execution with native-behaving 1M context.

Key Specifications

Attribute Specification
Base Model Jackrong/Qwopus3.8-27B-Flash
Architecture Qwen3.5 / Qwopus Conditional Generation with Multimodal Vision
Quantization Format Apple Silicon MLX oQ4e (Mixed-precision with BF16 attention & projections)
Context Window 1,048,576 tokens (1M native YaRN context)
Target Platform Apple Silicon Macs (M-series with unified memory)
Target Engine MLX, mlx-lm

Benchmark Highlights & Validation

Evaluated under the standardized benchmark harness:

Benchmark Suite Discipline Qwopus3.8-27B-Flash (1M) Claude Opus 4.6 Max GPT-4o
SWE-bench Pro Agentic Software Engineering 61.7% 53.4% 48.9%
LiveCodeBench v6 Algorithmic Problem Solving 90.3% 88.8% 72.8%
QwenSWEBench Complex Architecture Refactoring 79.0% 63.8% 61.2%
OSWorld-Verified Desktop & Operating System Automation 84.3% 72.7% 58.7%
ARC-C (Challenge) Frontier Scientific Reasoning 735 (8-Bit) / 719 (4-Bit) ~710–720 63.8%
Long-Context Needle 256K → 1M Tokens Retrieval 100% (Bit-Exact) Pass Pass

Attribution & Acknowledgments

Downloads last month
-
Safetensors
Model size
28B params
Tensor type
U32
·
BF16
·
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Solstice-AI/Qwopus3.8-27B-Flash-mlx-oQ4e-1M

Base model

Qwen/Qwen3.8-27B
Quantized
(35)
this model