About this role
About the Role Join the Runtime team at Databricks to help build the next generation distributed data storage and processing systems that power large-scale data workloads. You will design, implement, and optimize components that enable fast relational query performance, while providing the expressiveness for ETL, data science, and real-time workloads. You will contribute to projects including Apache Spark, Delta Lake, and Delta Pipelines.
What You'll Do
- Design, implement, and optimize distributed storage, query execution, and data-plane components for scale and reliability.
- Collaborate with product, data science, and infrastructure teams to deliver end-to-end features on the Databricks data platform.
- Mentor engineers, review code, and help shape system architecture and performance tuning.
- Write tests, automate deployments, and contribute to performance optimizations and maintainability.
What We're Looking For
- BS or higher in Computer Science or related field; 8+ years of production experience in Java, Scala, or C++.
- Strong foundation in algorithms and data structures and their real-world use cases.
- Experience with distributed systems, databases, and big data systems (Apache Spark, Hadoop).
- Comfortable working toward a multi-year vision with incremental deliverables; motivated by delivering customer value and impact.
Nice to Have
- Experience with Delta Lake and Delta Pipelines; familiarity with cloud storage backends (e.g., AWS S3, Azure Blob Store).
- Familiarity with performance optimization, query optimization, and large-scale data processing.
Compensation & Benefits
- Local Pay Range: $182,400—$247,000 USD.
- The total compensation package may include annual performance bonus, equity, and comprehensive benefits; benefits vary by region and location.