Latest Artificial Intelligence R&D Session - with Digitalent & Mike Nedelko - Episode (006) episode artwork

EPISODE · Feb 28, 2025 · 1H 3M

Latest Artificial Intelligence R&D Session - with Digitalent & Mike Nedelko - Episode (006)

from AI Latest Research & Developments - With Digitalent & Mike Nedelko · host Dillan Leslie-Rowe

The sessions topics include:Reasoning Models: Mike highlights the rise of reasoning models dominating leaderboards, enabled by "inference time compute scaling." This allows models to allocate more computational power dynamically, leading to better accuracy and efficiency. These models use "chain of thought prompting," enhancing reasoning by generating intermediate steps, inspired by Daniel Kahneman's "System 2 thinking." He also discussed "Humanity's Last Exam," a challenging new benchmark designed to test advanced reasoning models.DeepSeek R1: Mike explored DeepSeek R1's innovations, including stable 8-bit floating point operations and multi-hat latent attention, which reduced memory usage and improved efficiency. The real breakthrough was its use of reinforcement learning with self-verifiable tasks, allowing the model to learn without traditional supervised data. This approach improved reasoning and generalisation.Reinforcement Learning and Generalisation: Mike emphasised a shift from supervised fine-tuning to reinforcement learning, enabling models to generalise intelligence rather than just memorise. This approach lowers training costs while enhancing reasoning abilities. He also discussed the growing trend of using reinforcement learning and self-play to make AI training more efficient and affordable.

Episode metadata supplied by the publisher feed · Published Feb 28, 2025

Embed this episode

The sessions topics include: Reasoning Models: Mike highlights the rise of reasoning models dominating leaderboards, enabled by "inference time compute scaling." This allows models to allocate more computational power dynamically, leading to better accuracy and efficiency. These models use "chain of thought prompting," enhancing reasoning by generating intermediate steps, inspired by Daniel Kahneman's "System 2 thinking." He also discussed "Humanity's Last Exam," a challenging new benchmark d...

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

Latest Artificial Intelligence R&D Session - with Digitalent & Mike Nedelko - Episode (006)

0:00 1:03:01

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of AI Latest Research & Developments - With Digitalent & Mike Nedelko?

This episode is 1 hour and 3 minutes long.

When was this AI Latest Research & Developments - With Digitalent & Mike Nedelko episode published?

This episode was published on February 28, 2025.

Can I download this AI Latest Research & Developments - With Digitalent & Mike Nedelko episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!