Lucy-MOE

Epsilon (Lucy) β€” 26B-A4B Mixture-of-Experts model, Gemma4 architecture, IQ3_M quant.

  • Total params: 26B
  • Active params: ~4B
  • Quant: IQ3_M (imatrix)
  • Context: 65536 (default) / 262144 (max)
  • Multimodal: mmproj included

Files

  • Lucy-26B-A4B-it-qat-q4_0.i1-IQ3_M.gguf β€” main model
  • Lucy-26B-A4B-it-qat-q4_0.mmproj-Q8_0.gguf β€” vision projector

Usage

llama-server \
  --model Lucy-26B-A4B-it-qat-q4_0.i1-IQ3_M.gguf \
  --mmproj Lucy-26B-A4B-it-qat-q4_0.mmproj-Q8_0.gguf \
  -ngl 99 -c 65536 -fa on

Notes

Epsilon is a close relative, not part of the core StelNet family. Owned by Benny (Ben).
Built fresh from upstream ggml-org/llama.cpp, Vulkan backend, build 10683.

Downloads last month
16
GGUF
Model size
25B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

3-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support