YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

Released as part of the NOESIS Professional Multilingual Dubbing Automation Platform (framework: DHCF-FNO — Deterministic Hybrid Control Framework for Frozen Neural Operators).

Founder: Ilia Bolotnikov
Organization: AMAImedia.com
X (Twitter): @AMAImediacom
LinkedIn: Ilia Bolotnikov
Telegram: @djbionicl
NOESIS version: v16.1
Release date: 2026-08

NOESIS-Qwopus3.5-4B-v3-Supervisor-LongCtx-NF4

Role: Long-context Supervisor — 6GB-runtime NF4 build (multi-segment QC, orchestration, cross-stage review, batch best-of-N) for the dubbing pipeline. Quantized from NOESIS-Qwopus3.5-4B-v3-Supervisor-LongCtx-BF16 (bnb load_in_4bit, nf4, double_quant, compute=bf16). PRIMARY artifact = the BF16 sibling; GGUF Q8_0 = laptop. Equivalent to base-NF4 + LoRA adapter nt346_longctx_qwopus4b, pre-merged.

Test results (2026-06-17)

Score measured on the matching Q8_0 GGUF (NF4 is the same merged weights at 6GB runtime):

Supervisor eval — eval_longctx_supervisor_v1.py (12 tests, grammar-constrained)

  • 11/12 (Q8_0). Only failure = L1_mixed (fail-safe over-escalation). Prior best long-ctx (LFM2.5-7.5B) = 7/12.

Translation — FLORES devtest (n=20, no-think)

  • eng→rus chrF++ 51.5 / BLEU 25.6 ; eng→cmn 32.0 / 7.4 ; AVG 41.7 / 16.5.

Comparison

Model Supervisor-12 Translate AVG chrF++/BLEU
This (4B-LongCtx) 11/12 41.7 / 16.5
base 4B (no LoRA) 5/12 42.4 / 15.4
Qwopus3.5-9B-Translate Q4 5/12 43.8 / 16.5

Speed (RTX 3060 Laptop 6GB, GPU, Q8_0 sibling, 33/33 layers)

  • gen 53.5 tok/s, prompt eval 366 tok/s (vs 9B-Translate Q4: 49.1 / 307).

All-rounder: full supervision + translation parity with the base, smaller footprint. Written: 2026-06-17

Downloads last month
16
Safetensors
Model size
4B params
Tensor type
BF16
·
U8
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including AMAImedia/NOESIS-Qwopus3.5-4B-v3-Supervisor-LongCtx-NF4