About this role
About the Role Samsara is seeking a Senior Data Engineer to join our Data team to design and maintain data pipelines that feed our analytics data model. You will build pipelines within our central data lake to ingest and transform data from IoT devices and software products into the core data model, enabling statistical analysis, model training, and dashboards. This is a remote position open to candidates in the US. Relocation assistance will not be provided.
What You'll Do
- Build and maintain highly reliable computed tables, incorporating data from various sources, including unstructured data like video and audio, Samsara sensor & product data, and customer metadata.
- Access, manipulate, and integrate external datasets with internal data.
- Deliver high-quality data with strong uptime and reliability requirements, including customer-facing data sets.
- Collaborate closely with cross-functional teams such as Data Science & Analytics, AI/ML, and other Data Engineers to ensure high-quality data for diverse purposes from causal inference, model training, and dashboarding.
- Champion Samsara’s cultural principles (Focus on Customer Success, Build for the Long Term, Adopt a Growth Mindset, Be Inclusive, Win as a Team) as we scale globally and across new offices.
What We're Looking For
- BA / MS degree in Computer Science, Statistics, or a related discipline.
- 4+ years of experience in a data engineering-focused role.
- Demonstrated experience designing data models at scale.
- Proficiency in building ETL pipelines to handle large volumes of data.
- Experience with Spark-based data platforms.
- Strong command of at least one data orchestration tool (Airflow, Dagster, or Prefect).
- Expertise in SQL, Python, and working with REST APIs.
- Familiarity with software engineering fundamentals and reading backend development code.
- Experience with version control systems such as Git/GitHub.
- Ideal: Familiarity with time series data and late-arriving data.
- Knowledge of Databricks, Delta Lakes, and Dagster.
- Previous experience working in a public cloud (AWS, GCP, Azure).
- Exposure working on a data model for a product’s first-party data.
- Exposure to complex data, including ML outputs and/or client-side signals.
Nice to Have
- Time series data literacy and late-arriving data handling.
- Databricks, Delta Lakes, Dagster proficiency.
- AWS, GCP, Azure cloud experience.
- Experience with first-party product data modeling.
- Exposure to ML outputs and client-side signals.
Compensation & Benefits
- Annual base salary: $119,595—$201,000 USD.
- Eligible for an initial RSU grant with no vesting cliff and ongoing refresh opportunities tied to performance.
- Relocation assistance not provided; remote work within US.