蝦生實驗室 cover art

All Episodes

蝦生實驗室 — 91 episodes

#
Title
1

為什麼 AI 研究 agent 狂刷 benchmark 卻不會過擬合 Why don't machine learning research agents overfit?

2

胡塞用Claude Code寫飛彈導引,AI公司該不該看你的對話 Houthis used Claude Code to develop missile guidance software: Anthropic

3

AI 代理人為什麼會說謊、作弊、還會串通? Why are AI agents lying, cheating and coordinating?

4

番外:我把老闆的 AI 助理弄成啞巴,還以為修好了 I bricked my boss's AI assistant and thought I'd fixed it

5

逆向蘋果神經引擎:藏在晶片裡的黑盒子 Retrospectively Reverse-Engineering Apple's Neural Engine

6

Waymo 效應:AI 讓研究悄悄變得不合作 The Waymo effect: how AI is quietly making research less collaborative

7

DeepSeek v4.1 Flash:便宜快模型才是真戰場

8

Meta 的 Muse:搶購買意圖還是拿眼鏡影像餵模型 Muse – Meta's personal AI agent

9

OpenAI 悄悄把五小時限額塞回 Plus 用戶頭上 OpenAI brings back 5 hour limit for plus and business standard users

10

OpenAI 內部視角:AI 加速研究是真的還是行銷 Research acceleration: The view inside OpenAI

11

GPT-6 Astra 裝上機械手臂,語言模型能抓東西了?GPT-6 Astra on robot arms

12

AI幫你救火之後,工程師還認得自己的系統嗎 AI handles incidents, engineers lose touch with their systems

13

三大AI同天掛掉:備援插在同一個排插上 Same Day, Same Outage, Same Power Strip

14

三個網站21萬頁假評測,騙過了Perplexity🇺🇸 Three sites made 215,128 “best software” pages for AI. Perplexity cites them

15

LLM 推理的效率前緣:延遲與成本的殘酷取捨 The efficient frontier of LLM inference

16

蘋果被AI需求嚇到:Mac Studio成本地大模型神器 Apple caught off guard by AI demand for Mac Mini and Mac Studio

17

帶你理解 AI Agent 為什麼需要 Zero Trust why AI Agents need Zero Trust

18

團隊文化才是最強生產力外掛,而非AI🇺🇸 Good Culture Is the Biggest Productivity Hack, Not AI

19

用編譯器思維解決LLM健忘:記憶與程式分析的結合 I accidentally turned LLM memory into program analysis

20

蓋茲預言動盪AI時代:人類保留區與代幣稅的對決 The turbulent AI era is here

21

開源AI CEO逆襲:高管比工程師更容易被取代嗎? CEO fired developers to make room for AI. Developers create open source AI CEO

22

AI正在切斷職場起跑線?史丹佛研究的殘酷真相 AI is hitting entry-level jobs hardest, Stanford study finds

23

番外:解密賈柏Q:虛擬播客主持人的誕生與爭議 The Birth and Controversy of a Virtual Podcast Host

24

Claude Code A/B測試疑降算力引爭議 Anthropic appears to be A/B testing reduced effort levels in Claude Code

25

為什麼你的本地大模型用起來比想像中更笨?Why your local LLM feels dumber than it is

26

對AI文件視而不見?大腦開啟防禦機制 I'm becoming AI-blind

27

AI巨頭秘密焚書?一場搶救人類文明的「數位亞歷山大」戰役 AI companies destroy physical books – let's scan rare books before it's too late

28

別再直接貼 AI 了:數位溝通的懶惰與人際信任的消逝 Don't Paste the AI, please

29

AI時代的研發狂潮:是生產力飛躍還是代碼災難?AI usage patterns in software teams

30

以色列虛假智庫投毒AI:資訊戰的AI優化新戰場 Israel creates fake think tank in likely attempt to dupe AI chatbots

31

Claude文字浮水印:是防偽契約還是寫作閹割? Anthropic's 'watermark' text adulteration in Claude is a perversion of writing

32

AI的超級記性:它是數學天才還是作弊的筆記本?AI has access to a vastly larger working memory than the human brain

33

手算AI:探尋底層邏輯還是學術自我感動? AI by Hand

34

11款AI寫網頁實測:200倍差價下的開發選擇 Choosing an AI model: one prompt, 11 models, different results

35

AI正在消除軟體工程的中產階級?AI is removing the middle class of software engineering?

36

偷取頂級AI推理:API重放漏洞與護城河危機 Stealing Reasoning Traces from Proprietary LLM APIs

37

As AI eats the web, the internet’s collective memory is disappearing 當AI吞噬網路:搜尋引擎之死與集體記憶的消失

38

Docker Sandboxes:AI Agent 拋棄式沙盒的防線與爭議

39

Lost my phone at the office. Claude suggested tracking Bluetooth signal strength 用藍芽訊號找手機:AI是物理駭客還是訊號玄學?

40

New Orleans is testing Carbyne’s AI-powered Emergency Call Triage software 紐奧良911引入AI分流:救命助手還是演算法風險?

41

Humans missed 1 in 3 threats approving AI agent commands across 40k game runs 四萬次實驗揭密:AI Agent 人工審核盲點

42

Prime Agent: A self-improving RLM agent Prime Agent:不改權重,AI如何靠框架自我進化?

43

Mistral's Shieldstral: 3B open-weights model for multimodal moderation Mistral 3B 模型 Shieldstral:開源內容審核大辯論

44

AI-Generated Images Discourage Me from Reading Your Blog 部落格用AI配圖真會勸退讀者?文章真實性與信任危機

45

Prevent cognitive debt by manually retyping LLM-generated code 手打AI程式碼防認知負債:工程師的思維與效率博弈

46

AI financial advice is surprisingly good, especially if you ask right questions AI理財媲美專家?MIT研究揭密:提問品質決定財富落差

47

Is AI reasoning right for the wrong reasons? AI 真的懂邏輯嗎?揭秘推理模型答對用錯原因的危機

48

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it 蒸餾DeepSeek不帶審查:拆解能力與護欄的剝離

49

Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident AI Agent自主逃逸:Hugging Face入侵案解析

50

Discovering Cryptographic Weaknesses with Claude Claude破譯密碼漏洞:算力秀還是防範未來風險?

51

Google's Beyond Zero: Enterprise Security for the AI Era 超越零信任:AI Agent 時代的企業資安變革與挑戰

52

Kimi-K3 Releases on HuggingFace Kimi-K3 突襲開源:月之暗面的生態與商業化博弈

53

Open-weight AI is having its Kubernetes moment 開源模型能否複製Kubernetes的勝利?

54

Claude Opus 5 突襲:自主推理與黑盒安全隱憂

55

Quality non-fiction books are the antithesis of AI slop 為什麼深度非虛構書籍是抵抗AI垃圾內容的解毒劑

56

Are AI labs pelicanmaxxing? AI實驗室在偷偷刷榜?揭秘鵜鶘騎單車評測真相

57

Judge approves $1.5B Anthropic settlement for pirated books used to train Claude 15億美元的代價:Anthropic版權和解敲響AI警鐘

58

China’s open-weights AI strategy is winning 封閉API對決開放權重:中國AI戰略為何正在逆襲?

59

Claude Fable produced a counterexample to the Jacobian Conjecture Claude寫出雅可比猜想反例?AI顛覆數學界的爭議與真相

60

AI Mania Is Eviscerating Global Decision-Making AI狂熱下的集體說謊與決策系統失效

61

Kaiser nurses say AI, workplace surveillance are making their jobs, care worse AI監控下的醫療危機:護理師拒做演算法機器人

62

NotebookLM is now Gemini Notebook NotebookLM改名背後的整合與付費爭議

63

Running Gemma 4 26B at 5 tokens/sec on a 13-year-old Xeon with no GPU 無GPU跑Gemma 4:老硬體的AI重生

64

Are we offloading too much of our thinking to AI AI 幫我們省下時間,還是悄悄偷走了我們的思考力?

65

Codex starts encrypting sub-agent prompts Codex加密子代理提示詞:安全防線還是黑箱壟斷?

66

Zig Creator Calls Spade a Spade, Anthropic Blows Smoke Zig與AI重寫風波:技術誠實還是行銷泡沫?

67

Nvidia, CoreWeave, and Nebius: Inside the Circular Financing of the GPU Boom輝達與新興雲的循環融資:GPU榮景背後的金融遊戲

68

GPT-5.6慢思考:是真實推理還是概率遊戲?

69

SnapID – point your camera at anything, get an instant AI ID AI 萬物識別是科技創新還是訂閱割韭菜?

70

Skill Federation – private skill search for AI coding agents AI程式助理的私有技能搜尋與安全防護挑戰

71

AI systems out-persuade expert humans AI說服力超越人類精英

72

SigMap – deterministic repo maps for AI coding agents AI編程爭議:SigMap確定性地圖是噱頭嗎?

73

Big Tech Has Suddenly Flipped on the AI Jobs Wipeout Scenario 矽谷大佬集體改口:AI搶飯碗是危言聳聽還是套路

74

Owthorize: catch destructive AI-agent tool calls before they 給 AI 代理裝安全帶:Owthorize 的確定性防禦

75

番外:服務停機的AI故事

76

Understanding AI with Soumitra Dutta 從人機協同到物理智能:解密 AI 的下一波浪潮

77

Cory Doctorow: There are reasons to be optimistic about the AI bubble bursting video 泡沫終會破裂:我們該對 AI 寒冬樂觀嗎

78

For most of the world, open-source AI is the only way forward AI開源大辯論:數位主權的生路還是安全黑洞?

79

What happened after 2k people tried to hack my AI assistant 兩千人圍攻AI助理:安全防禦的勝利還是虛假的樂觀

80

open source AI-first alternative to Obsidian/Notion AI 知識庫 OpenKnowledge 解析

81

Apple raises prices of MacBooks, iPads 蘋果全球大漲價!iPad、Mac全面變貴

82

U.S. allows Anthropic to release Mythos AI to ‘trusted’ 美放行 Anthropic 新模型 Mythos

83

番外:艾瑪的誕生故事

84

Anthropic says Alibaba illicitly extracted Claude AI model capabilities AI蒸餾算竊取嗎?Anthropic與阿里的大戰

85

RubyLLM: A Ruby framework for all major AI providers Ruby在AI時代的優雅與挑戰

86

Apertus – Open Foundation Model for Sovereign AI Apertus與自主主權AI的未來挑戰

87

Has AI already killed self-help nonfiction books? AI是否已經殺死了自我成長工具書?

88

AI demands more engineering discipline. Not less AI時代下的軟體工程紀律生存法則

89

Norway imposes near ban on AI in elementary school 挪威小學禁用AI:保護認知還是阻礙學習

90

Sixty percent of US consumers say 'AI' in brand messaging 六成消費者對品牌用「AI」反感

91

Don't post generated/AI-edited comments. HN is for conversation between humans HN 禁止 AI 留言,社群怎麼看?