Tina: Tiny LoRA Reasoning Models episode artwork

EPISODE · Apr 25, 2025 · 15 MIN

Tina: Tiny LoRA Reasoning Models

from Best AI papers explained · host Enoch H. Kang

We discuss Tina, a family of efficient reasoning models achieved by applying Low-Rank Adaptation (LoRA) during reinforcement learning to a small 1.5B parameter language model. This approach demonstrates that strong reasoning performance, competitive with larger models, can be attained with significantly reduced computational costs. The authors explore the effectiveness of this minimalist strategy across various reasoning tasks and ablation studies, hypothesizing that LoRA facilitates rapid adaptation to the structural format of reasoning. Ultimately, Tina aims to democratize the development of reasoning models by showcasing a highly cost-effective and reproducible methodology, with all code and models being open-sourced.

Episode metadata supplied by the publisher feed · Published Apr 25, 2025

Embed this episode

NOW PLAYING

Tina: Tiny LoRA Reasoning Models

0:00 15:37

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Best AI papers explained?

This episode is 15 minutes long.

When was this Best AI papers explained episode published?

This episode was published on April 25, 2025.

Can I download this Best AI papers explained episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!