EP 349 : Poetiq Beats Google: Tiny Startup Tops ARC-AGI-2 Benchmark episode artwork

EPISODE · Dec 9, 2025 · 12 MIN

EP 349 : Poetiq Beats Google: Tiny Startup Tops ARC-AGI-2 Benchmark

from AI Brief · host MonPod

Discover how Poetiq, a six-person AI startup, outperformed Google's Gemini 3 Deep Think on the ARC-AGI-2 reasoning benchmark, achieving a groundbreaking 54% score. Learn about the innovative 'meta-system' that made this possible and the implications for the future of AI development. Also, explore the latest AI news, including a new study on poetry prompts that can bypass AI safety guardrails and updates on OpenAI, Apple, and Meta. Join the conversation and stay ahead of the curve in the rapidly evolving AI landscape. Listen now and subscribe for more insights! Tools mentioned: Mistral 3, Seedream 4.5, Kling Avatar 2.0, VibeVoice, Sup, GSong, X-Design, Documentation.

Episode metadata supplied by the publisher feed · Published Dec 9, 2025

Embed this episode

NOW PLAYING

EP 349 : Poetiq Beats Google: Tiny Startup Tops ARC-AGI-2 Benchmark

0:00 12:15

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of AI Brief?

This episode is 12 minutes long.

When was this AI Brief episode published?

This episode was published on December 9, 2025.

Can I download this AI Brief episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!