EPISODE · May 30, 2026 · 6 MIN
Claude Opus 4.8 Is Out. The Benchmark Numbers Aren't the Story.
from AI First Pod
Anthropic dropped Opus 4.8 yesterday — same price, better coding scores, and a four-fold reduction in silent code bugs. But the real headline is alignment: Opus 4.8 scores at near-Mythos levels on misalignment metrics, quietly bringing the restricted model's safety profile into the general tier. Plus: Figure AI's robots sorted 250,000 packages in 200 hours with zero failures, and California's AI legislation just hit its crossover deadline with thirty bills in play and no federal law in sight.
Embed this episode
NOW PLAYING
Claude Opus 4.8 Is Out. The Benchmark Numbers Aren't the Story.
No transcript for this episode yet
Similar Episodes
No similar episodes found.