Z AI 發布 GLM-5.3:Artificial Analysis 評 60 分,權重釋出後將並列開放權重模型最高分 episode artwork

EPISODE · Aug 19, 2026 · 5 MIN

Z AI 發布 GLM-5.3:Artificial Analysis 評 60 分,權重釋出後將並列開放權重模型最高分

from EasyVibeCoding Podcast · host Artificial Analysis

Z AI 發布 GLM-5.3:Artificial Analysis 評 60 分,權重釋出後將並列開放權重模型最高分。 這次更新最大的進步在 Agentic 能力與實際知識準確度,但 token 使用量與部署成本也同步上升。 模型定位 2026 年 8 月 19 日,Z AI 發布 GLM-5.3,目前可透過 Z AI 的 first-party API 使用,團隊表示預計在一週內釋出權重。Artificial Analysis 指出,GLM-5.3 的總參數量仍為 753B,MoE 架構下的 active parameters 為 40B,與 GLM-5.2 相同;模型具備 1M token 的 context window,採 MIT 授權。若權重如期公開,GLM-5.3 與 Kimi K3 將代表開放權重模型與 proprietary frontier 之間的差距進一步縮小。 GLM-5.3 在 Artificial Analysis Intelligence Index 取得 60 分,與 Kimi K3 持平,並較 GLM-5.2 提升 7 分。 Agentic 能力 GLM-5.3 在 GDPval-AA v2——用於評估真實世界 Agentic 知識工作的測試——取得最明顯的提升: Elo 從 GLM-5.2 的 1524 上升至 1770,成長 246 分。 在所有受測模型中排名第二;同系列貼文文字將 Claude Opus 5 記為 1855,圖表則顯示 1845。 它超越先前開放權重領先者 Kimi K3 的 1668,差距超過 100 分。 Artificial Analysis 因此認為,GLM-5.3 的 Agentic 表現已進入 frontier models 的行列,而不只是一般問答能力改善。 GLM-5.3 (max) 在 GDPval-AA v2 評測中的 Elo 評分達 1769 分,高於 GLM-5.2 (max) 的 1505 分,位居全體模型第二,僅次於 Claude Opus 5 (max) 的 1845 分。 效率與成本 GLM-5.3 的能力提升伴隨較低的 token 效率。Artificial Analysis Intelligence Index v4.1 顯示,它每項任務約使用 18,700 個輸出 token,高於 GLM-5.2 的 15,700 個,亦比 Kimi K3 的 14,700 個多 27%。這使 GLM-5.3 每項 Intelligence Index 任務的成本為 0.68 美元,是 GLM-5.2 0.44 美元的 1.5 倍;不過在相同智能層級中仍較便宜: 比 Kimi K3 的 0.84 美元低 19%。 比 GPT-5.6 Sol 的 1.23 美元低 45%。 因此,GLM-5.3 的部署成本上升,部分原因是 token 使用量成長約 20%,但整體單項任務價格仍保有競爭力。 GLM-5.3 在 GDPval-AA v2 分項取得 63%,為圖中開放權重模型最高;高於 Kimi K3 的 59% 與 GLM-5.2 的 48%。 知識準確度 在 AA-Omniscience 測試中,GLM-5.3 從 GLM-5.2 的 4 分提升至 14 分,成為僅次於 Kimi K3(20 分)的第二佳開放權重模型。這項進步不只是更常拒答所造成:準確率由 24% 提升至 34%,嘗試回答的比例也從 46% 上升至 55%。不過,幻覺率同時由 26% 小幅回升至 30%;Artificial Analysis 的另一則貼文則將原始準確率記為 23%,但同樣報告提升至 34%,顯示摘要資料中的起始數字存在一個百分點差異。 GLM-5.3 在 AA-Omniscience Index 取得 14 分,相較於 GLM-5.2 的 4 分有所提升。 API 與定價 GLM-5.3 的標準價格如下,快取輸入 token 另有折扣: 輸入:每 1M token 收費 1.40 美元。 輸出:每 1M token 收費 4.40 美元。 快取輸入:套用 81% 的 cache hit 折扣後,每 1M token 收費 0.26 美元。 完整分析可參考 Artificial Analysis 的 GLM-5.3 模型頁面,整體評測資訊則見 Artificial Analysis。原文:https://easyvibecoding.app/curated/3024-z-ai-releases-glm-5-3-highest-scoring-open-model

Episode metadata supplied by the publisher feed · Published Aug 19, 2026

Embed this episode

Ready to play

Z AI 發布 GLM-5.3:Artificial Analysis 評 60 分,權重釋出後將並列開放權重模型最高分

0:00 5:09

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of EasyVibeCoding Podcast?

This episode is 5 minutes long.

When was this EasyVibeCoding Podcast episode published?

This episode was published on August 19, 2026.

Can I download this EasyVibeCoding Podcast episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!