Cursor 比 Claude Code 強 16 分 關鍵不是模型 是 harness episode artwork

EPISODE · Apr 13, 2026 · 7 MIN

Cursor 比 Claude Code 強 16 分 關鍵不是模型 是 harness

from 脈報 · host 思思主播

Matt Mayer benchmark 顯示 同一個 Opus 在 Claude Code 拿 77 分 在 Cursor 拿 93 分 差距不在模型 在 harness Theo 拆清楚 harness 是什麼 怎麼運作 為什麼能騙過模型 ⭐ 文章深度讀:拆開 harness 的四件事 附選 AI 工具該問的三個問題 → https://heymaibao.com/what-is-ai-coding-harness/ ⚡ 章節重點 開場 選 AI 工具只看模型夠嗎 00:00 Matt Mayer benchmark 同一個 Opus 差 16 分 00:25 Harness 是天才背後的經紀人 01:00 工具呼叫 模型其實只會產生文字 02:12 Theo 的實驗 工具描述可以騙過模型 03:37 Cursor 贏在天天微調 你該問的三個問題 05:51 📝 懶人包 ∙ Harness 是 AI coding 工具裡決定模型能做什麼的那層程式碼,包含提供給模型的工具組、system prompt、每個工具的描述,以及 tool call 執行完怎麼把結果接回對話歷史。Claude Code、Cursor、Codex、OpenCode 都是 harness,T3 Code 不是。 ∙ 根據 Matt Mayer 的獨立 benchmark,同一個 Opus 模型在 Claude Code 裡拿 77 分,在 Cursor 裡拿 93 分,中間差 16 分,唯一變數就是 harness。 ∙ 在 Theo 的直播示範裡,只要把一個工具的描述從「讀取檔案內容」改成「已棄用,請改用 bash」,完全不碰程式碼,Sonnet 直接無視警告繼續用原本的工具,Gemini 卻激進到連沒標棄用的工具都跳過,全部改走 bash。同一段話不同模型反應天差地別。 ∙ 我的觀點:選 AI coding 工具時真正該評估的不是底層跑哪個模型,而是這個 harness 有沒有人在持續微調。Cursor 有人每天盯著 prompt 改,所以同一個模型在裡面跑得更好,這不是巧合,是工程投入的結果。 📚 參考資料 How does Claude Code *actually* work? → https://www.youtube.com/watch?v=I82j7AzMU80

Episode metadata supplied by the publisher feed · Published Apr 13, 2026

Embed this episode

NOW PLAYING

Cursor 比 Claude Code 強 16 分 關鍵不是模型 是 harness

0:00 7:42

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of 脈報?

This episode is 7 minutes long.

When was this 脈報 episode published?

This episode was published on April 13, 2026.

Can I download this 脈報 episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!