YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

Recommended Settings

  • temperature: 1.0
  • top_k: 40
  • Everything else (top_p, min_p, repeat penalty, etc.) should be disabled - some clients enable these by default, turn them off

Required flag: --jinja to enable the Harmony response format (the model won't work correctly without it).

For llama.cpp:

llama-server -m model.gguf --jinja -fa -b 2048 -ub 2048

LM Studio

Compatible with Reasoning Effort custom buttons. To use them, put the model in:

LM Models\lmstudio-community\gpt-oss-120b-GGUF\
Downloads last month
15
GGUF
Model size
117B params
Architecture
gpt-oss
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support