EP82 :健康大型語言模型-回饋強化學習(ReinforcementLearning)思考語模  ft.陳秀熙.許辰陽 episode artwork

EPISODE · Jan 3, 2026 · 46 MIN

EP82 :健康大型語言模型-回饋強化學習(ReinforcementLearning)思考語模 ft.陳秀熙.許辰陽

from 不只是科技_AI星球永續健康 · host 陳秀熙、許辰陽、侯信恩、楊心怡

健康大型語言模型-回饋強化學習(ReinforcementLearning)思考語模:DeepSeek 人工智慧邁入推理覺醒階段。語言模型藉由回饋強化學習(Reinforcement Learning)」與思維鍊(Chain of Thought,CoT)訓練策略,如同人一般思考,不僅能回答問題,更能檢查、修正,並反思推理過程。本週我們將探討智慧語模如何藉由回饋訓練建立思維鍊,並介紹DeepSeek模型在腎臟病營養決策中的應用。 漢聲廣播電台-星球永續健康:https://reurl.cc/WbGALy 連結:https://youtu.be/0dMboO3gUeo 點我:新聞稿 簡報檔 -- Hosting provided by SoundOn

Episode metadata supplied by the publisher feed · Published Jan 3, 2026

Embed this episode

Ready to play

EP82 :健康大型語言模型-回饋強化學習(ReinforcementLearning)思考語模 ft.陳秀熙.許辰陽

0:00 46:17

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of 不只是科技_AI星球永續健康?

This episode is 46 minutes long.

When was this 不只是科技_AI星球永續健康 episode published?

This episode was published on January 3, 2026.

Can I download this 不只是科技_AI星球永續健康 episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!