EPISODE · Jun 17, 2026 · 22 MIN
What Data Should You Rely On? (AI + Real-Time Data) - with Amaury Desrosiers, Nimble
from The Data Splash · host Upriver
For a decade, the data moat was infrastructure—whoever could afford 20 engineers to maintain scrapers won. Nimble's Amaury Desrosiers explains why that moat is gone, how external web data became a first-class citizen alongside your warehouse, and why real-time will soon be a default property, not a category. Plus the three-circle model for internal vs. external data and the one move every data leader should make tomorrow.⏱ CHAPTERS 0:00 Introduction 0:51 The 30 Second Splash 1:44 What Nimble does 2:08 Questions internal data can't answer 3:22 The three concentric circles 4:28 Ground truth vs. context — joining the two 5:40 Why web data used to be a nightmare 6:27 Solving connection and parsing end to end 7:53 The moat moves from infra to usage 9:32 Why real-time value compounds 11:55 Batch vs. ad hoc pipelines 14:24 One layer, not one connector per site 15:34 Schema by use case 16:58 Scale changed, the data model didn't 18:58 Three predictions for the next three years 20:25 What data leaders should do tomorrow 21:19 Takeaways and wrap🎙 ABOUT DATA SPLASH: Data Splash is a podcast for data engineers, data leaders, and anyone trying to make sense of AI and data right now. Brought to you by Upriver.🔔 Subscribe for new episodes weekly.🔗 LINKS • Upriver: [https://www.upriverdata.com/] • Connect with Amaury Desrosiers: [https://www.linkedin.com/in/amaurydesrosiers] • Connect with Omri Lifshitz: [https://www.linkedin.com/in/omri-lifshitz-8a531814a/]#DataEngineering #AI #DataPlatform #LLMs #AIAgents
Embed this episode
NOW PLAYING
What Data Should You Rely On? (AI + Real-Time Data) - with Amaury Desrosiers, Nimble
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.