About this role
About the Role Samsara is seeking a Senior Software Engineer I to join the Data Platform team. This role focuses on building and operating the core analytical infrastructure that powers Samsara's data lakehouse, enabling engineers, data scientists, analysts, and product teams to derive insights at scale. It is a specialized data infrastructure role focused on ingestion, processing, cataloging, and access for petabytes of data. What You'll Do
- Design, build, and operate high-scale data ingestion and replication systems from Samsara’s primary production data stores (RDS, DynamoDB, internal APIs, and event-driven systems) into the data lakehouse.
- Build and maintain reliable, scalable data platform infrastructure capable of handling petabytes of data across analytics, AI, product, and operational use cases.
- Improve the reliability, observability, scalability, security, and developer experience of Samsara’s Spark and Databricks-based data processing platform.
- Develop internal libraries, APIs, frameworks, and tooling in Go and Python to help teams move, process, discover, and access data safely and efficiently.
- Work on foundational data lake and lakehouse technologies, including Delta Lake on S3, data catalogs, metadata services, orchestration systems, and platform automation.
- Collaborate closely with infrastructure, product engineering, data science, analytics, security, and data engineering teams to understand platform needs and deliver durable, scalable solutions.
- Stay connected to modern data platform technologies and help shape Samsara’s long-term data infrastructure roadmap, including support for AI, privacy, security, global scale, and customer-facing data products.
- Champion, role model, and embed Samsara’s cultural principles (Focus on Customer Success, Build for the Long Term, Adopt a Growth Mindset, Be Inclusive, Win as a Team) as we scale globally and across new offices. What We're Looking For
- 4+ years of professional experience in software engineering with deep experience building and operating production data platforms, distributed systems, or large-scale ingestion infrastructure.
- Deep experience designing and operating reliable, scalable data infrastructure and collaborating with cross-functional teams (infrastructure, product engineering, data science, security, analytics).
- Experience with Spark and Databricks, Delta Lake on S3, and data replication from primary stores such as RDS and DynamoDB; familiarity with data catalogs, orchestration systems, and metadata services; and proficiency in Go and Python for internal tooling.
- Strong communication skills and a track record of delivering durable, scalable solutions in a cross-functional environment.