EP81 :健康大型語言模型-回饋強化學習(Reinforcement Learning)思考語模訓練 ft.陳秀熙.許辰陽 episode artwork

EPISODE · Dec 26, 2025 · 44 MIN

EP81 :健康大型語言模型-回饋強化學習(Reinforcement Learning)思考語模訓練 ft.陳秀熙.許辰陽

from 不只是科技_AI星球永續健康 · host 陳秀熙、許辰陽、侯信恩、楊心怡

健康大型語言模型-回饋強化學習(Reinforcement Learning)思考語模訓練 大型語言模型(LLM)在晶片科技與數位技術快速進展下不斷進化,從模仿人類語言的工具,蛻變為能「思考、反省與學習」數位智慧體,開始展現超越人類思維限制潛能。本週我們將聚焦於近期發表於自然科學期刊智慧語模回饋訓練(Reinforcement Learning),探討此技術進展如何強化大型語言模型邏輯推理能力,並延伸至實際應用DeepSeek-R1創新發展。同時我們也將透過「雙月下的醫師與機器人」的故事,闡述人類式與非人類式推理的差異,在元宇宙時代的人性與精準之間,思索AI與人類共進的未來可能。 漢聲廣播電台-星球永續健康:https://reurl.cc/WbGALy 連結:https://youtu.be/DAnyki2dI9k 點我:新聞稿簡報檔 -- Hosting provided by SoundOn

Episode metadata supplied by the publisher feed · Published Dec 26, 2025

Embed this episode

Ready to play

EP81 :健康大型語言模型-回饋強化學習(Reinforcement Learning)思考語模訓練 ft.陳秀熙.許辰陽

0:00 44:51

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of 不只是科技_AI星球永續健康?

This episode is 44 minutes long.

When was this 不只是科技_AI星球永續健康 episode published?

This episode was published on December 26, 2025.

Can I download this 不只是科技_AI星球永續健康 episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!