🎬  How to optimize your Spark jobs? Join our webinar on Oct 16 with Data Engineer Things
Register here
Events
›
Data Engineer Things Webinar: Stop Optimizing the Wrong Part of Your Spark Job
Meet us at

Stop Optimizing the Wrong Part of Your Spark Job

Date
October 16, 2026
October 16, 2026
Time
12:00 pm
City
Duration
1 hour
Hosted by
Data Engineer Things
America/Chicago
12:00 pm - 12:45 pm CT
Register
Register
Event artwork
Webinar
Jump to
What to expectOur boothSessionsBeyond the boothMeet the teamAgendaSpeakersTalksLocationSpeak at a meetup
About the evening

High utilization doesn’t necessarily mean an efficient Spark workload – and the layer where waste shows up is often not where you need to fix it. In this webinar, we’ll walk through a practical framework for investigating and optimizing Spark workloads across infrastructure, runtime, data access, and query plans, using real production cases to show the signals, root causes, and optimizations at each layer. We’ll also discuss what it takes to collect the runtime signals needed to apply this framework effectively, and the trade-offs between different approaches to monitoring Spark workloads.

What you'll learn
  • A practical framework for investigating Spark workloads layer by layer: infrastructure, runtime, data access and query plans
  • The signals and root causes behind waste, from real production cases
  • What it takes to collect the runtime signals you need, and the trade-offs between monitoring approaches
Who should attend
  • Data and platform engineers running Spark in production
  • Data engineering leaders chasing cost and reliability
  • FinOps teams looking at lakehouse spend
Speakers
Who you'll hear from
Speaker headshot
Ohad Raviv
CTO & Co-Founder
Ohad is CTO and co-founder of definity, where he works on Spark optimization and agentic data engineering. Previously, he led the big data technology roadmap for PayPal's global data science group, pioneering work across Hadoop and Spark observability, distributed graph infrastructure, and entity resolution. Ohad is an Apache Spark contributor.
Speaker headshot
Roy Daniel
CEO & Co-Founder
Roy is CEO and co-founder of definity, an Agentic Data Engineering platform for the Lakehouse and Spark ecosystem. He works with enterprise data engineering teams to optimize cost, improve reliability, and accelerate developer velocity. Previously, he held data and product leadership roles across Fortune 500 companies, most recently at FIS.
Hosted by
Host headshot
Roy Daniel
Can't make it live?
Register anyway — we'll email the recording and slides to everyone who signs up.
Register
Register
More from definity
Upcoming events
‍