New ModelsOpen weights
GLM-5.3-Flash
GLM-5.3-Flash: the OX Alpha mystery model, open-sourced under MIT
Z.ai open-sourced GLM-5.3-Flash, a 320B-parameter MoE with 18B active under MIT license, after stealth-testing it for about six days as 'OX Alpha' with effectively unlimited free traffic on OpenRouter — all served on Chinese chips. Company-reported DeepSWE is 63.4 with Claude Opus 4.8-level coding claims, it's natively multimodal, and a hybrid sparse/linear attention architecture cuts KV cache size 4x versus GLM 5.3 with 3x serving performance.
320B-A18B parameters (total / active), MIT license63.4 DeepSWE (company-reported)4x smaller KV cache vs GLM 5.3
New Models
GLM-5.3
GLM-5.3: post-training alone delivers a 6x Terminal-Bench jump
Z.ai announced GLM-5.3, keeping the same 743B base as GLM 5.2 but jumping from 4.6 to 28.3 on Terminal-Bench 3 and gaining almost 20% on DeepSWE from post-training alone, at unchanged pricing with a 1M context window. It scores 60 on the Artificial Analysis Intelligence Index, roughly Kimi K3 level with about a third of the parameters, and shows emergent cybersecurity capabilities: 84% on CyberGym and 54.5 on ExploitGym, beating GPT-5.6 Sol. API-only for now — weights (and license terms) expected later.
4.6→28.3 Terminal-Bench 3, a 6x jump from post-training alone60 Artificial Analysis Intelligence Index84% CyberGym