Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

incoai
/
Qwen3.8-27B-Splash

Text Generation
splash
apple-silicon
metal
local-inference
dflash2
speculative-decoding
qwen3.8
4-bit precision
Model card Files Files and versions
xet
Community
12
New discussion
Resources
  • PR & discussions documentation
  • Code of Conduct
  • Hub documentation

Can we have a 8bit target ?

#12 opened 1 minute ago by
trigger2k25

Request for more smaller ornith1.5-9B splash or any based on Qwen3.5-9B )

#11 opened about 11 hours ago by
naildirect

DFLASH2 for Qwen3.8 Flash NExt

#10 opened 2 days ago by
Oxidez

the cli help is incomplete

#9 opened 4 days ago by
jorgeleo

why not pi?

#8 opened 4 days ago by
jorgeleo

More numbers

1
#7 opened 4 days ago by
jorgeleo

Independent M4 Max 64GB benchmark: Splash vs oMLX vs Ollama on Qwen3.8-27B

3
#6 opened 4 days ago by
fparrav

Some MCP calls fail

2
#5 opened 5 days ago by
Taubenschlag

Why has the context size been reduced to only 128k?

1
#4 opened 5 days ago by
ChristopherKai

Splash Monitor - a native macOS GUI for the Splash engine

❤️ 1
1
#3 opened 6 days ago by
hometrix

Well Done! 👏 - Speed comparison to few other 3.6 27B and 3.8 27B quants

3
#2 opened 7 days ago by
DarkHorse2305

Testing on the M4 Max 128G. Yes, it’s faster than the alternatives, yet it still isn’t viable for regular work.

👀 1
2
#1 opened 7 days ago by
szkane
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs