Ep79: Reducing Agent Token Costs + RAG Beyond Semantic Search episode artwork

EPISODE · Jun 12, 2026 · 19 MIN

Ep79: Reducing Agent Token Costs + RAG Beyond Semantic Search

from Breaktime Tech Talks · host jmhreif

In this episode, I sit down with Roie Schwaber-Cohen, a software engineer and developer advocate at Pinecone, to talk about smarter ways to build with AI — without burning through tokens or your patience! What we cover: Why agentic AI systems burn so many tokens (and ways to combat it) How Pinecone's Nexus pre-explores retrieval paths so agents don't have to discover them at runtime, cutting latency and token usage The problem with naive RAG ("Franken answers") and why domain-level separation of your documents matters How Pinecone Marketplace lets non-developers connect structured and unstructured data sources to build production-ready AI apps Why semantic similarity isn't the same as correctness, and how document introspection helps agents ask better questions Links & Resources: Pinecone Pinecone Marketplace (recently announced) Pinecone Nexus Roie on LinkedIn

Episode metadata supplied by the publisher feed · Published Jun 12, 2026

Embed this episode

NOW PLAYING

Ep79: Reducing Agent Token Costs + RAG Beyond Semantic Search

0:00 19:22

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Breaktime Tech Talks?

This episode is 19 minutes long.

When was this Breaktime Tech Talks episode published?

This episode was published on June 12, 2026.

Can I download this Breaktime Tech Talks episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!