Add Terminal-Bench 2.1 evaluation results

#64
by SaylorTwift HF Staff - opened

Add Evaluation Results for zai-org/GLM-5.2

Summary

This PR adds two Terminal-Bench 2.1 evaluation results extracted from the model card
benchmark table for zai-org/GLM-5.2 to the .eval_results/ directory, following the
Hugging Face Hub evaluation-results specification.
The model card reports Terminal-Bench 2.1 under two different harnesses, so both are
recorded against the same terminalbench_2_1 task, distinguished via notes.

Benchmarks Added

Benchmarks Skipped (Not Registered on Hub)

Not applicable — only Terminal-Bench 2.1 was requested for this PR.

Source

Files Added

  • .eval_results/terminal-bench-2.1.yaml

Verification

This result was extracted from the model card's published benchmark table (row
"Terminal Bench 2.1 (Terminus-2)"), not from independently re-run/verified eval logs.

Ready to merge
This branch is ready to get merged automatically.

Sign up or log in to comment