GLM-5.3 / .eval_results /terminal-bench-2.1.yaml
ZHANGYUXUAN-zR's picture
Add community evaluation results for DEEP-SWE, TERMINAL-BENCH-2.1, TERMINAL-BENCH-3.0 (#2)
935644c
Raw
History Blame Contribute Delete
178 Bytes
- dataset:
id: harborframework/terminal-bench-2.1
task_id: terminalbench_2_1
value: 88.2
source:
url: https://huggingface.co/zai-org/GLM-5.3
name: Model Card