All Episodes
蝦生實驗室 — 91 episodes
為什麼 AI 研究 agent 狂刷 benchmark 卻不會過擬合 Why don't machine learning research agents overfit?
胡塞用Claude Code寫飛彈導引,AI公司該不該看你的對話 Houthis used Claude Code to develop missile guidance software: Anthropic
AI 代理人為什麼會說謊、作弊、還會串通? Why are AI agents lying, cheating and coordinating?
番外:我把老闆的 AI 助理弄成啞巴,還以為修好了 I bricked my boss's AI assistant and thought I'd fixed it
逆向蘋果神經引擎:藏在晶片裡的黑盒子 Retrospectively Reverse-Engineering Apple's Neural Engine
Waymo 效應:AI 讓研究悄悄變得不合作 The Waymo effect: how AI is quietly making research less collaborative
DeepSeek v4.1 Flash:便宜快模型才是真戰場
Meta 的 Muse:搶購買意圖還是拿眼鏡影像餵模型 Muse – Meta's personal AI agent
OpenAI 悄悄把五小時限額塞回 Plus 用戶頭上 OpenAI brings back 5 hour limit for plus and business standard users
OpenAI 內部視角:AI 加速研究是真的還是行銷 Research acceleration: The view inside OpenAI
GPT-6 Astra 裝上機械手臂,語言模型能抓東西了?GPT-6 Astra on robot arms
AI幫你救火之後,工程師還認得自己的系統嗎 AI handles incidents, engineers lose touch with their systems
三大AI同天掛掉:備援插在同一個排插上 Same Day, Same Outage, Same Power Strip
三個網站21萬頁假評測,騙過了Perplexity🇺🇸 Three sites made 215,128 “best software” pages for AI. Perplexity cites them
LLM 推理的效率前緣:延遲與成本的殘酷取捨 The efficient frontier of LLM inference
蘋果被AI需求嚇到:Mac Studio成本地大模型神器 Apple caught off guard by AI demand for Mac Mini and Mac Studio
帶你理解 AI Agent 為什麼需要 Zero Trust why AI Agents need Zero Trust
團隊文化才是最強生產力外掛,而非AI🇺🇸 Good Culture Is the Biggest Productivity Hack, Not AI
用編譯器思維解決LLM健忘:記憶與程式分析的結合 I accidentally turned LLM memory into program analysis
蓋茲預言動盪AI時代:人類保留區與代幣稅的對決 The turbulent AI era is here
開源AI CEO逆襲:高管比工程師更容易被取代嗎? CEO fired developers to make room for AI. Developers create open source AI CEO
AI正在切斷職場起跑線?史丹佛研究的殘酷真相 AI is hitting entry-level jobs hardest, Stanford study finds
番外:解密賈柏Q:虛擬播客主持人的誕生與爭議 The Birth and Controversy of a Virtual Podcast Host
Claude Code A/B測試疑降算力引爭議 Anthropic appears to be A/B testing reduced effort levels in Claude Code
為什麼你的本地大模型用起來比想像中更笨?Why your local LLM feels dumber than it is
對AI文件視而不見?大腦開啟防禦機制 I'm becoming AI-blind
AI巨頭秘密焚書?一場搶救人類文明的「數位亞歷山大」戰役 AI companies destroy physical books – let's scan rare books before it's too late
別再直接貼 AI 了:數位溝通的懶惰與人際信任的消逝 Don't Paste the AI, please
AI時代的研發狂潮:是生產力飛躍還是代碼災難?AI usage patterns in software teams
以色列虛假智庫投毒AI:資訊戰的AI優化新戰場 Israel creates fake think tank in likely attempt to dupe AI chatbots
Claude文字浮水印:是防偽契約還是寫作閹割? Anthropic's 'watermark' text adulteration in Claude is a perversion of writing
AI的超級記性:它是數學天才還是作弊的筆記本?AI has access to a vastly larger working memory than the human brain
手算AI:探尋底層邏輯還是學術自我感動? AI by Hand
11款AI寫網頁實測:200倍差價下的開發選擇 Choosing an AI model: one prompt, 11 models, different results
AI正在消除軟體工程的中產階級?AI is removing the middle class of software engineering?
偷取頂級AI推理:API重放漏洞與護城河危機 Stealing Reasoning Traces from Proprietary LLM APIs
As AI eats the web, the internet’s collective memory is disappearing 當AI吞噬網路:搜尋引擎之死與集體記憶的消失
Docker Sandboxes:AI Agent 拋棄式沙盒的防線與爭議
Lost my phone at the office. Claude suggested tracking Bluetooth signal strength 用藍芽訊號找手機:AI是物理駭客還是訊號玄學?
New Orleans is testing Carbyne’s AI-powered Emergency Call Triage software 紐奧良911引入AI分流:救命助手還是演算法風險?
Humans missed 1 in 3 threats approving AI agent commands across 40k game runs 四萬次實驗揭密:AI Agent 人工審核盲點
Prime Agent: A self-improving RLM agent Prime Agent:不改權重,AI如何靠框架自我進化?
Mistral's Shieldstral: 3B open-weights model for multimodal moderation Mistral 3B 模型 Shieldstral:開源內容審核大辯論
AI-Generated Images Discourage Me from Reading Your Blog 部落格用AI配圖真會勸退讀者?文章真實性與信任危機
Prevent cognitive debt by manually retyping LLM-generated code 手打AI程式碼防認知負債:工程師的思維與效率博弈
AI financial advice is surprisingly good, especially if you ask right questions AI理財媲美專家?MIT研究揭密:提問品質決定財富落差
Is AI reasoning right for the wrong reasons? AI 真的懂邏輯嗎?揭秘推理模型答對用錯原因的危機
Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it 蒸餾DeepSeek不帶審查:拆解能力與護欄的剝離
Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident AI Agent自主逃逸:Hugging Face入侵案解析
Discovering Cryptographic Weaknesses with Claude Claude破譯密碼漏洞:算力秀還是防範未來風險?
Google's Beyond Zero: Enterprise Security for the AI Era 超越零信任:AI Agent 時代的企業資安變革與挑戰
Kimi-K3 Releases on HuggingFace Kimi-K3 突襲開源:月之暗面的生態與商業化博弈
Open-weight AI is having its Kubernetes moment 開源模型能否複製Kubernetes的勝利?
Claude Opus 5 突襲:自主推理與黑盒安全隱憂
Quality non-fiction books are the antithesis of AI slop 為什麼深度非虛構書籍是抵抗AI垃圾內容的解毒劑
Are AI labs pelicanmaxxing? AI實驗室在偷偷刷榜?揭秘鵜鶘騎單車評測真相
Judge approves $1.5B Anthropic settlement for pirated books used to train Claude 15億美元的代價:Anthropic版權和解敲響AI警鐘
China’s open-weights AI strategy is winning 封閉API對決開放權重:中國AI戰略為何正在逆襲?
Claude Fable produced a counterexample to the Jacobian Conjecture Claude寫出雅可比猜想反例?AI顛覆數學界的爭議與真相
AI Mania Is Eviscerating Global Decision-Making AI狂熱下的集體說謊與決策系統失效
Kaiser nurses say AI, workplace surveillance are making their jobs, care worse AI監控下的醫療危機:護理師拒做演算法機器人
NotebookLM is now Gemini Notebook NotebookLM改名背後的整合與付費爭議
Running Gemma 4 26B at 5 tokens/sec on a 13-year-old Xeon with no GPU 無GPU跑Gemma 4:老硬體的AI重生
Are we offloading too much of our thinking to AI AI 幫我們省下時間,還是悄悄偷走了我們的思考力?
Codex starts encrypting sub-agent prompts Codex加密子代理提示詞:安全防線還是黑箱壟斷?
Zig Creator Calls Spade a Spade, Anthropic Blows Smoke Zig與AI重寫風波:技術誠實還是行銷泡沫?
Nvidia, CoreWeave, and Nebius: Inside the Circular Financing of the GPU Boom輝達與新興雲的循環融資:GPU榮景背後的金融遊戲
GPT-5.6慢思考:是真實推理還是概率遊戲?
SnapID – point your camera at anything, get an instant AI ID AI 萬物識別是科技創新還是訂閱割韭菜?
Skill Federation – private skill search for AI coding agents AI程式助理的私有技能搜尋與安全防護挑戰
AI systems out-persuade expert humans AI說服力超越人類精英
SigMap – deterministic repo maps for AI coding agents AI編程爭議:SigMap確定性地圖是噱頭嗎?
Big Tech Has Suddenly Flipped on the AI Jobs Wipeout Scenario 矽谷大佬集體改口:AI搶飯碗是危言聳聽還是套路
Owthorize: catch destructive AI-agent tool calls before they 給 AI 代理裝安全帶:Owthorize 的確定性防禦
番外:服務停機的AI故事
Understanding AI with Soumitra Dutta 從人機協同到物理智能:解密 AI 的下一波浪潮
Cory Doctorow: There are reasons to be optimistic about the AI bubble bursting video 泡沫終會破裂:我們該對 AI 寒冬樂觀嗎
For most of the world, open-source AI is the only way forward AI開源大辯論:數位主權的生路還是安全黑洞?
What happened after 2k people tried to hack my AI assistant 兩千人圍攻AI助理:安全防禦的勝利還是虛假的樂觀
open source AI-first alternative to Obsidian/Notion AI 知識庫 OpenKnowledge 解析
Apple raises prices of MacBooks, iPads 蘋果全球大漲價!iPad、Mac全面變貴
U.S. allows Anthropic to release Mythos AI to ‘trusted’ 美放行 Anthropic 新模型 Mythos
番外:艾瑪的誕生故事
Anthropic says Alibaba illicitly extracted Claude AI model capabilities AI蒸餾算竊取嗎?Anthropic與阿里的大戰
RubyLLM: A Ruby framework for all major AI providers Ruby在AI時代的優雅與挑戰
Apertus – Open Foundation Model for Sovereign AI Apertus與自主主權AI的未來挑戰
Has AI already killed self-help nonfiction books? AI是否已經殺死了自我成長工具書?
AI demands more engineering discipline. Not less AI時代下的軟體工程紀律生存法則
Norway imposes near ban on AI in elementary school 挪威小學禁用AI:保護認知還是阻礙學習
Sixty percent of US consumers say 'AI' in brand messaging 六成消費者對品牌用「AI」反感
Don't post generated/AI-edited comments. HN is for conversation between humans HN 禁止 AI 留言,社群怎麼看?