139: Decoupling the Execution Engine From Python’s Pandas with Aditya Parameswaran of Ponder episode artwork

EPISODE · May 24, 2023 · 57 MIN

139: Decoupling the Execution Engine From Python’s Pandas with Aditya Parameswaran of Ponder

from The Data Stack Show · host Rudderstack

Highlights from this week’s conversation include:Aditya’s background and journey in the data space (2:47)What does Ponder do? (5:18)101 on Pandas and why people utilize it (6:42)The challenge of translating Pandas to a big data platform (16:11)Data Warehouses and ML workflows (21:27)The differences in the “zoo” of data languages (26:56)Why do ML and data engineering have to be so different in languages? (34:39)Builders should be adapting to the users and not the other way around (39:32)Will we see a singular data interface in the future? (46:19)Aditya’s most surprising discovery in his research (50:40)Final thoughts and takeaways (53:18)Read more of Aditya's work: Pandas vs. SQL – Part 1: The Food Court and the Michelin-Style RestaurantPandas vs. SQL – Part 2: Pandas Is More ConcisePandas vs. SQL – Part 3: Pandas Is More FlexiblePandas vs. SQL – Part 4: Pandas Is More ConvenientThe Data Stack Show is a weekly podcast powered by RudderStack, the CDP for developers. Each week we’ll talk to data engineers, analysts, and data scientists about their experience around building and maintaining data infrastructure, delivering data and data products, and driving better outcomes across their businesses with data.RudderStack helps businesses make the most out of their customer data while ensuring data privacy and security. To learn more about RudderStack visit rudderstack.com. Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

Episode metadata supplied by the publisher feed · Published May 24, 2023

Embed this episode

This week on The Data Stack Show, Eric and Kostas chat with Aditya Parameswaran, Associate Professor at UC Berkeley & Co-Founder of Ponder. During the episode, Aditya discusses the zoo of data languages including a 101 on Pandas, why builders should be adapting to users, exploring what Ponder is solving in the data space, interesting theories on the way things should operate in the industry, and more.

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

139: Decoupling the Execution Engine From Python’s Pandas with Aditya Parameswaran of Ponder

0:00 57:43

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of The Data Stack Show?

This episode is 57 minutes long.

When was this The Data Stack Show episode published?

This episode was published on May 24, 2023.

Can I download this The Data Stack Show episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!