EPISODE · Jun 30, 2025 · 4 MIN
AI界的“学霸”和“学神”:差的不是智商,是训练方法
from AI可可AI生活
[CL] OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling[Shanghai Jiao Tong University]arxiv.org在小宇宙查看该单集文稿
Embed this episode
Ready to play
AI界的“学霸”和“学神”:差的不是智商,是训练方法
0:00
4:40
1×
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
Frequently Asked Questions
How long is this episode of AI可可AI生活?
This episode is 4 minutes long.
When was this AI可可AI生活 episode published?
This episode was published on June 30, 2025.
Can I download this AI可可AI生活 episode?
Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!