Kimi k1.5 episode artwork

EPISODE · Jan 23, 2025 · 22 MIN

Kimi k1.5

from Large Language Model (LLM) Talk · host AI-Talk

Kimi k1.5 is a multimodal LLM trained with reinforcement learning (RL). Key aspects include: long context scaling to 128k, improving performance with increased context length; improved policy optimization using a variant of online mirror descent; and a simplistic framework that enables planning and reflection without complex methods. It uses a reference policy in its off-policy RL approach, and long2short methods such as model merging and DPO to transfer knowledge from long-CoT to short-CoT models, achieving state-of-the-art reasoning performance. The model is jointly trained on text and vision data.

Episode metadata supplied by the publisher feed · Published Jan 23, 2025

Embed this episode

NOW PLAYING

Kimi k1.5

0:00 22:27

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Large Language Model (LLM) Talk?

This episode is 22 minutes long.

When was this Large Language Model (LLM) Talk episode published?

This episode was published on January 23, 2025.

Can I download this Large Language Model (LLM) Talk episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!