The Artificial Analysis Intelligence Index placed GLM-5.3 in a tie for the highest ranking among open weights models with a score of 60. Zai Org achieved these intelligence gains by fine-tuning GLM-5.2 instead of training a new base architecture. This post-training process produced emergent cybersecurity capabilities, highlighted by an 84.5% score on CyberGym, prompting a temporary safety hold on the model weights' release to evaluate exploit generation risks.

Zai Org also launched GLM-5.3-Flash, which entered the Agent Arena ranking 4th among open source models and 19th overall. The model carries a $0.12 median cost per task and provides a 15.3% improvement in confirmed success, positioning it between DeepSeek V4 (High) and GPT-5.6 Luna (xHigh) on the Pareto frontier.

Sign in to suggest edits

Key sources

  1. SOURCE@arena“GLM-5.3-Flash ranks #4 among open source models and #19 overall, one spot ahead of GLM-5.3 (Max)”x.com
  2. SUPPORT@deeplearningai“This unexpected jump in exploit generation prompted a temporary safety hold on the model weights’ release”x.com
Markdown