Start with the execution plan, numbers beat guesses.
Hitting a wall in prod and looking for patterns others have used. Minimal repro below but happy to share more context.
Context: EMR on EKS vs traditional EMR for Spark, ops overhead?
Happy to share schema snippets or metrics if useful.
Start with the execution plan, numbers beat guesses.
Note that merge on Delta still needs unique keys defined correctly.
+1, saw identical behaviour after upgrading Spark 3.4 to 3.5.
In our case the root cause was an implicit cast preventing pushdown.
Document the grain decision, most BI bugs turn out to be grain bugs.
Event-driven beats cron once landing time gets unpredictable.
Document the grain decision, most BI bugs turn out to be grain bugs.
Could you share a sketch of the salting logic?
Sign in to reply.
© 2026 Lakebench, operated by Hunnurji Rao. Bengaluru, Karnataka, India.
No cluster. No install. Just the tab.