EP281: Restoring plasticity to over-trained AI episode artwork

EPISODE · Jul 2, 2026 · 22 MIN

EP281: Restoring plasticity to over-trained AI

from Learning GenAI via SOTA Papers · host Yun Wu

Title: When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL HandoffSource: http://arxiv.org/abs/2606.09932v1Summary:This paper identifies and solves the critical 'loss of plasticity' bottleneck in the standard LLM post-training pipeline where excessive SFT inhibits subsequent RL optimization. It introduces 'Rejuvenation', a foundational training primitive that uses model fusion and neuron resets to enable robust reasoning gains during RL while preserving SFT-acquired knowledge.

Episode metadata supplied by the publisher feed · Published Jul 2, 2026

Embed this episode

Ready to play

EP281: Restoring plasticity to over-trained AI

0:00 22:12

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Learning GenAI via SOTA Papers?

This episode is 22 minutes long.

When was this Learning GenAI via SOTA Papers episode published?

This episode was published on July 2, 2026.

Can I download this Learning GenAI via SOTA Papers episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!