About this role
Job title: Senior Software Engineer - Distributed Data Systems
About the Role As a Senior Software Engineer on Databricks' Runtime team, you will help build the next generation distributed data storage and processing systems that outperform specialized SQL engines while supporting diverse workloads from ETL to data science. You will contribute to projects like Apache Spark, Delta Lake, Delta Pipelines, and Data Plane Storage, delivering high-performance, scalable services across cloud storage backends.
What You'll Do
- Design, implement, and optimize distributed data storage and processing systems on cloud backends (e.g., AWS S3, Azure Blob Store).
- Develop and improve runtime components and engines for Spark, Delta Lake, Delta Pipelines; contribute to the performance and reliability of the platform.
- Collaborate with cross-functional teams to deliver scalable services across millions of virtual machines and support workloads from ETL to data science.
What We're Looking For
- BS (or higher) in Computer Science or related field, or equivalent practical experience.
- 5+ years of production experience in Java, Scala, or C++.
- Strong foundation in algorithms and data structures, with real-world applicability.
- Experience with distributed systems, databases, and big data ecosystems (Apache Spark, Hadoop).
Nice to Have
- Familiarity with Delta Lake features such as ACID transactions and time travel.
- Experience with cloud storage backends and building high-availability services.
Compensation & Benefits
- Local Pay Range: $157,700—$213,800 USD.
- The total compensation may include annual performance bonus, equity, and other listed benefits.