@hemanaidu
Free member
Data engineer at a an e-commerce platform. Mostly Iceberg and cost optimization.
Company is expanding into a second region for latency reasons. Data pipelines currently run entirely in one region. Designing the multi-region setup, active-active processing in both regions, or active-passive with failover?
I have read the docs but real-world tradeoffs are unclear. What would you optimize first? Context: Dynamic task mapping vs many parallel operators Happy to share schema snippets or metrics if useful.
© 2026 Lakebench, operated by Hunnurji Rao. Bengaluru, Karnataka, India.
No cluster. No install. Just the tab.