About this role
About the Role Databricks is seeking a Director of Engineering (Data Infrastructure) to build and lead Bengaluru's data infrastructure organization. You’ll establish foundational teams that guarantee billing correctness, operational resilience, and zero-downtime recovery across the monetization stack, alongside multi-region data ingestion, developer platforms, and deployment automation. You’ll shape architectural decisions and champion infrastructure-as-product thinking to enable Databricks' scale.
What You'll Do
- Deliver the infrastructure vision for systems processing billions in daily billing transactions with zero tolerance for error, building disaster recovery that is provably reliable, testing frameworks that catch production-scale problems, correctness systems that make billing errors structurally impossible, and observability that predicts failures before they happen.
- Build Bengaluru's data infrastructure organization by establishing it as the destination for India's top infrastructure talent, hiring multiple engineering managers who become force multipliers, and creating a culture where solving hard distributed systems problems at scale is the daily work.
- Own business-critical systems operating 24/7/365 across 100+ regions, driving reliability improvements that prevent millions in revenue loss while eliminating operational toil through frameworks that make systems self-healing, self-tuning, and self-documenting.
- Ship platforms that compound engineering leverage across Databricks: correctness frameworks that catch billing errors before customers do, deployment automation that makes regional expansion push-button, data integration systems that process petabyte-scale flows without human intervention, and testing infrastructure where comprehensive coverage is automatic, not heroic.
- Position infrastructure as product by treating internal engineering teams as customers with SLAs, measuring adoption and satisfaction, iterating based on feedback, and demonstrating that every dollar invested in infrastructure returns multiplicative gains in product velocity, reliability improvements, or cost reductions.
What We're Looking For
- 14+ years in distributed systems engineering with 6+ years leading infrastructure organizations and 4+ years managing managers at companies where infrastructure failures meant immediate revenue impact, customer escalations, or regulatory consequences.
- Technical depth across petabyte-scale data pipelines and distributed systems reliability where you can engage from how should we architect multi-region disaster recovery to why is this Kafka cluster exhibiting this latency pattern while knowing when to coach versus when to decide.
- Track record defining multi-year infrastructure vision and translating it into sequential deliverables that show value quarterly while building toward architectural end states, positioning infrastructure investments as business enablers rather than cost centers, and making build-vs-buy decisions that compound over time.
- Experience building 99.999%+ reliable systems with established practices for SLOs/SLIs, chaos engineering, disaster recovery, and sophisticated observability that predicts failures before they happen.
- Proven ability to scale infrastructure organizations in high-growth environments where you've doubled engineering while maintaining quality bar, developed engineering managers, and created teams where retention is high because the problems are interesting and the culture is strong.
- Communication skills to make complex topics accessible to diverse stakeholders.