AIToday
Large Language ModelsOpen-Source AIAI Business & IndustryTHE DECODERPublished: Aug 20, 2026, 01:03 JST2 min read

GLM-5.3 ties for top open-model rank, costs 19% less than rival

GLM-5.3 ties for top open-model rank, costs 19% less than rival

Key takeaway

  • GLM-5.3 from Chinese startup Z.ai has tied for the top ranking among open-source AI models with a score of 60 points, matching Kimi K3.

  • It particularly excels at agentic tasks, with an Elo score of 1,770 on the GDPval-AA v2 benchmark—second only to Claude Opus 5.

  • At $0.68 per task, it undercuts Kimi K3 by 19 percent, though its open-weights release is delayed by about two weeks while the company implements security controls due to the model's effectiveness at detecting vulnerabilities.

3 Key Points

  1. What happened

    Z.ai's GLM-5.3 scores 60 points on the Artificial Analysis Intelligence Index, tying with Kimi K3 for the top spot among open models. On the GDPval-AA v2 benchmark, it reaches an Elo score of 1,770—a jump of 246 points from GLM-5.2's 1,524—placing it second overall, behind only Claude Opus 5 (1,855).

  2. Why it matters

    GLM-5.3 costs $0.68 per task, making it 19 percent cheaper than Kimi K3 ($0.84), while delivering performance that now matches Western frontier models. The model's strength in agentic tasks (those requiring autonomous decision-making) suggests it can handle complex, multi-step workflows—a capability increasingly valued in enterprise AI.

  3. What to watch

    Z.ai is delaying the open-weights release by about two weeks. The company says GLM-5.3 is so effective at detecting security vulnerabilities that it is first strengthening controls and restricting full access to select security partners before broader availability.

Ask the AI about this article →

Context & Analysis

GLM-5.3 represents a significant performance milestone for open-source AI models. Z.ai's achievement in tying for the top ranking on the Artificial Analysis Intelligence Index while simultaneously capturing second place on the GDPval-AA v2 benchmark suggests that open-source models are narrowing the gap with closed, Western-built frontier models. The model's 246-point jump in agentic task performance—the domain where it made its biggest gains—is particularly noteworthy for businesses exploring autonomous AI systems. At the same time, GLM-5.3's 19 percent cost advantage over Kimi K3 ($0.68 vs. $0.84 per task) signals that competitive pressure on pricing is already reshaping the market for high-performance models.

The delayed open-weights release adds an interesting wrinkle. Z.ai's decision to withhold full access temporarily, citing the model's exceptional ability to detect security vulnerabilities, reflects a deliberate trade-off between speed-to-market and responsible disclosure. By restricting early access to security partners, the company appears to be preventing widespread weaponization of its vulnerability-detection capability before defenses can catch up—a pragmatic stance that may become more common as open models grow more capable.

FAQ

When will the open weights be available?
Z.ai is delaying the release by about two weeks. The company is strengthening controls and restricting full access to select security partners first because GLM-5.3 is so effective at detecting security vulnerabilities.
How does GLM-5.3 compare to Claude Opus 5 and Kimi K3?
On the GDPval-AA v2 benchmark, GLM-5.3's Elo score is 1,770, placing it second behind Claude Opus 5 (1,855) but ahead of other models. On the Artificial Analysis Intelligence Index, it ties with Kimi K3 at 60 points for the top spot among open models, but costs 19 percent less per task ($0.68 vs. $0.84).
What is the biggest improvement in GLM-5.3?
Its biggest gains come in agentic tasks. On the GDPval-AA v2 benchmark, its Elo score jumps from 1,524 to 1,770, a leap of 246 points.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Nvidia revives Rubin CPX chip with major redesignYahoo Finance AI · 2h ago
  • AI advice followed by 79%, but well-being unchangedITmedia AI+ · 5h ago
  • Enterprises face agent governance gapSiliconANGLE AI · 8h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleAmazon expands Prime Air drone delivery to nearly 500 U.S. cities by end of 2026