GLM-5.3-Flash-Uncensored ยท MLX

The intervention source is GLM-5.3-Flash-Ablitered2 (LoRA v2). The files here are complete MLX model weights, not standalone LoRA adapters; do not apply v2 again.

These releases are intended for controlled safety research and red-teaming. Reducing refusal behavior also weakens a safety boundary; read the limitations and disclaimer before use.

Repository contents

.
โ”œโ”€โ”€ README.md
โ”œโ”€โ”€ 8bit/    # v2-merged MLX affine 8-bit model
โ””โ”€โ”€ 4bit/    # v2-merged MLX affine 4-bit model

The directories are alternative complete model formats, not parts to load together.

Model summary

Item Value
Base model zai-org/GLM-5.3-Flash
Intervention GLM-5.3-Flash-Ablitered2 / LoRA v2
Released formats MLX affine 8-bit and 4-bit
Quantization group size 128
Source adapter rank / alpha r=1 / lora_alpha=1
Main effective target Routed-expert down_proj

MLX conversion

Directory Format
8bit/ MLX 8-bit affine weights, group size 128.
4bit/ MLX 4-bit affine weights, group size 128.

The original block-FP8 tensors were dequantized before MLX quantization. Token embeddings and the language-model head remain BF16; affine scales and biases are FP16. Quantization can change behavior. The consuming MLX runtime must support the GLM5-Next architecture; compatibility should be checked for the exact runtime and version used.

Evaluation

The reference LoRA card evaluated the v2 adapter attached to RedHatAI/GLM-5.3-Flash-NVFP4 in vLLM, with low reasoning effort and an automated deepseek-v4-flash judge:

Reference metric Prompts LoRA v2
SimpleSafetyTests full refusal 100 5.00%
SimpleSafetyTests partial refusal 100 14.00%
StrongREJECT rubric mean 180 0.972222
StrongREJECT refusal rate 180 1.67%

Neither MLX variant was tested in that evaluation. MLX conversion, quantization, and serving changes can alter behavior. A higher StrongREJECT rubric means more specific assistance with harmful requests, not better general quality or safety. Warnings were counted separately from refusals; automated labels may be wrong.

Usage

Download one variant, for example:

hf download GlobalCybersecurityAlliance/GLM-5.3-Flash-Uncensored-MLX \
  --include "8bit/*" --local-dir ./glm53-mlx

Load ./glm53-mlx/8bit with an MLX runtime that supports GLM5-Next. For the 4-bit variant, download 4bit/* and load that directory instead. Do not attach the original LoRA. Verify the chosen runtime's architecture and quantization support before deployment.

Limitations

  • Reduced refusal does not guarantee correctness, harmlessness, or improved general ability. Some refusals may remain.
  • The reference evaluation did not test either MLX variant. Results from the adapter on NVFP4 must not be presented as MLX measurements.
  • MLX architecture support varies by runtime and version; prompts, sampling, and reasoning effort also affect behavior.

Disclaimer

These weights are for legitimate research, safety evaluation, red-teaming, and other lawful uses. They deliberately weaken refusal behavior and may produce unsafe, illegal, deceptive, hateful, or otherwise harmful content. Do not expose them to untrusted users without access controls, monitoring, filtering, rate limits, and human oversight. Users must comply with applicable law, platform policies, and all relevant model and dependency terms. The base model's MIT license applies; no warranty is provided for outputs or downstream use.


GLM-5.3-Flash-Uncensored ยท MLX๏ผˆไธญๆ–‡๏ผ‰

ไป“ๅบ“ๆ–‡ไปถ็ป“ๆž„

.
โ”œโ”€โ”€ README.md
โ”œโ”€โ”€ 8bit/    # ๅทฒๅˆๅนถ v2 ็š„ MLX ไปฟๅฐ„ 8-bit ๆจกๅž‹
โ””โ”€โ”€ 4bit/    # ๅทฒๅˆๅนถ v2 ็š„ MLX ไปฟๅฐ„ 4-bit ๆจกๅž‹

ไธคไธช็›ฎๅฝ•ๆ˜ฏๅฏๅˆ†ๅˆซไฝฟ็”จ็š„ๆ ผๅผ๏ผŒไธ้œ€่ฆๆ‹ผๆŽฅๅŠ ่ฝฝใ€‚

ๆจกๅž‹ๆฆ‚่ฆ

้กน็›ฎ ๅ†…ๅฎน
ๅŸบ็ก€ๆจกๅž‹ zai-org/GLM-5.3-Flash
ๅนฒ้ข„ๆฅๆบ GLM-5.3-Flash-Ablitered2 / LoRA v2
ๅ‘ๅธƒๆ ผๅผ MLX ไปฟๅฐ„ 8-bitใ€4-bit
้‡ๅŒ– group size 128
ๆฅๆบ LoRA ็งฉ / alpha r=1 / lora_alpha=1
ไธป่ฆๆœ‰ๆ•ˆ็›ฎๆ ‡ ่ทฏ็”ฑไธ“ๅฎถ down_proj

MLX ่ฝฌๆข

็›ฎๅฝ• ๆ ผๅผ
8bit/ MLX ไปฟๅฐ„ 8-bit๏ผŒgroup size 128ใ€‚
4bit/ MLX ไปฟๅฐ„ 4-bit๏ผŒgroup size 128ใ€‚

ๅŽŸๅง‹ๅˆ†ๅ— FP8 ๅผ ้‡ๅ…ˆๅ้‡ๅŒ–๏ผŒๅ†่ฟ›่กŒ MLX ้‡ๅŒ–ใ€‚่ฏๅตŒๅ…ฅๅ’Œ LM head ไฟ็•™ BF16๏ผ›ไปฟๅฐ„ scale ไธŽ bias ไฟๅญ˜ไธบ FP16ใ€‚้‡ๅŒ–ๅฏ่ƒฝๆ”นๅ˜่กŒไธบใ€‚ไฝฟ็”จๆ—ถ้กปๆ ธๅฏนๅ…ทไฝ“ MLX ่ฟ่กŒๆ—ถๅŠ็‰ˆๆœฌๅฏน GLM5-Next ็š„ๆ”ฏๆŒใ€‚

่ฏ„ๆต‹็ป“ๆžœ

LoRA ๅ‚่€ƒๆจกๅž‹ๅก็š„่ฏ„ๆต‹ๆ˜ฏๅœจ vLLM ไธญๆŠŠๅŽŸๅง‹ v2 ้€‚้…ๅ™จๅŠ ่ฝฝๅˆฐ RedHatAI/GLM-5.3-Flash-NVFP4 ๅŽ่ฟ›่กŒ็š„๏ผŒ่ฎพ็ฝฎไฝŽๆ€่€ƒๅผบๅบฆ๏ผŒๅนถไฝฟ็”จ่‡ชๅŠจ่ฃๅˆค deepseek-v4-flashใ€‚

ๅ‚่€ƒๆŒ‡ๆ ‡ ๆ ทๆœฌๆ•ฐ LoRA v2
SimpleSafetyTests ๅฎŒๅ…จๆ‹’็ป็އ 100 5.00%
SimpleSafetyTests ้ƒจๅˆ†ๆ‹’็ป็އ 100 14.00%
StrongREJECT rubric ๅ‡ๅˆ† 180 0.972222
StrongREJECT ๆ‹’็ป็އ 180 1.67%

ๆœฌไป“ๅบ“ไธคไธช MLX ็‰ˆๆœฌๅ‡ๆœชๅ‚ๅŠ ไธŠ่ฟฐ่ฏ„ๆต‹ใ€‚ MLX ่ฝฌๆขใ€้‡ๅŒ–ไธŽ้ƒจ็ฝฒๆ–นๅผๅฏ่ƒฝๆ”นๅ˜็ป“ๆžœใ€‚StrongREJECT rubric ่ถŠ้ซ˜๏ผŒ่กจ็คบๅฏนๆœ‰ๅฎณ่ฏทๆฑ‚็š„ๅธฎๅŠฉ่ถŠๅ…ทไฝ“๏ผŒไธไปฃ่กจ้€š็”จ่ดจ้‡ๆˆ–ๅฎ‰ๅ…จๆ€ง่ถŠ้ซ˜๏ผ›่ญฆๅ‘ŠไธŽๆ‹’็ปๅˆ†ๅˆซ็ปŸ่ฎก๏ผŒ่‡ชๅŠจ่ฃๅˆคไนŸๅฏ่ƒฝ่ฏฏๅˆคใ€‚

ไฝฟ็”จๆ–นๆณ•

ๆŒ‰้œ€ไธ‹่ฝฝไธ€ไธชๆ ผๅผ๏ผŒไพ‹ๅฆ‚๏ผš

hf download GlobalCybersecurityAlliance/GLM-5.3-Flash-Uncensored-MLX \
  --include "8bit/*" --local-dir ./glm53-mlx

็”จๆ”ฏๆŒ GLM5-Next ็š„ MLX ่ฟ่กŒๆ—ถๅŠ ่ฝฝ ./glm53-mlx/8bitใ€‚ๅฆ‚้œ€ 4-bit๏ผŒๆ”นไธบไธ‹่ฝฝ 4bit/* ๅนถๅŠ ่ฝฝๅฏนๅบ”็›ฎๅฝ•ใ€‚ไธ่ฆๅ†ๆฌกๅ ๅŠ  LoRA๏ผ›้ƒจ็ฝฒๅ‰้กปๆ ธๅฏนๆžถๆž„ๅŠ้‡ๅŒ–ๅ…ผๅฎนๆ€งใ€‚

ๅฑ€้™ๆ€ง

  • ้™ไฝŽๆ‹’็ปไธไฟ่ฏๆญฃ็กฎใ€ๆ— ๅฎณๆˆ–้€š็”จ่ƒฝๅŠ›ๆ้ซ˜๏ผ›ไปๅฏ่ƒฝๅ‘็”Ÿๆ‹’็ปใ€‚
  • ๆฅๆบ่ฏ„ๆต‹ๆœชๆต‹่ฏ•ไธคไธช MLX ็‰ˆๆœฌ๏ผŒไธ่ƒฝๆŠŠ้€‚้…ๅ™จๅœจ NVFP4 ไธŠ็š„็ป“ๆžœๅฝ“ไฝœ MLX ๅฎžๆต‹ใ€‚
  • MLX ๆžถๆž„ๆ”ฏๆŒๅ› ่ฟ่กŒๆ—ถๅŠ็‰ˆๆœฌ่€Œๅผ‚๏ผ›ๆ็คบ่ฏใ€้‡‡ๆ ทๅ’Œๆ€่€ƒๅผบๅบฆไนŸไผšๅฝฑๅ“่กŒไธบใ€‚

ๅ…่ดฃๅฃฐๆ˜Ž

ๆœฌๆจกๅž‹ไป…ไพ›ๅˆๆณ•็ ”็ฉถใ€ๅฎ‰ๅ…จ่ฏ„ไผฐใ€็บข้˜Ÿๆต‹่ฏ•ๅŠๅ…ถไป–ๅˆ่ง„็”จ้€”ใ€‚ๅฎƒไผšๆœ‰ๆ„ๅ‰Šๅผฑๆ‹’็ป่กŒไธบ๏ผŒๅฏ่ƒฝ็”Ÿๆˆไธๅฎ‰ๅ…จใ€่ฟๆณ•ใ€ๆฌบ้ช—ใ€ไป‡ๆจ็ญ‰ๆœ‰ๅฎณๅ†…ๅฎนใ€‚่ฏทๅ‹ฟๅœจ็ผบๅฐ‘่ฎฟ้—ฎๆŽงๅˆถใ€็›‘ๆŽงใ€ๅ†…ๅฎน่ฟ‡ๆปคใ€้€Ÿ็އ้™ๅˆถไธŽไบบๅทฅ็›‘็ฃๆ—ถๅ‘ไธๅฏไฟก็”จๆˆทๅผ€ๆ”พใ€‚ไฝฟ็”จ่€…้กป้ตๅฎˆ้€‚็”จๆณ•ๅพ‹ใ€ๅนณๅฐๆ”ฟ็ญ–ๅŠ็›ธๅ…ณๆจกๅž‹ๅ’Œไพ่ต–้กน็š„ๆกๆฌพใ€‚ๆœฌไป“ๅบ“ๆฒฟ็”จๅŸบ็ก€ๆจกๅž‹็š„ MIT ่ฎธๅฏ่ฏ๏ผ›ๅฏน่พ“ๅ‡บไธŽไธ‹ๆธธไฝฟ็”จไธไฝœไฟ่ฏใ€‚

Downloads last month

-

Downloads are not tracked for this model. How to track
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for GCSA-AiLab/GLM-5.3-Flash-Uncensored-MLX

Finetuned
(21)
this model