About this role
Job title: Site Reliability Engineer (SRE)
About the Role Anduril is seeking a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Costa Mesa, CA or Washington DC. The SRE will architect, deploy and maintain cloud infrastructure and Kubernetes, drive DevOps, CI/CD, and improve the developer experience.
What You'll Do
- Architect, deploy and maintain infrastructure with cloud providers and Kubernetes (EKS)
- Collaborate with multi-disciplined teams to define and execute on internal and external deployments
- Promote SRE best practices in system resilience, performance monitoring and high availability
- Design, develop, and deliver solutions using infrastructure as code with tools like Terraform and Python
- Develop and maintain CI/CD pipelines for automated deployment
- Build strong relationships with internal and external customers to identify technical solutions to their problems
- Improve Anduril’s operational capabilities by improving our core product offering through root cause analysis and creating tooling capable of managing large scale deployments
- Lead the organization in building scalable, sustainable mechanisms to continue delivering to customers at the pace the business is scaling
What We're Looking For
- Holding active U.S. TOP SECRET security clearance
- 6+ years of engineering experience
- Technical expertise and demonstrated performance in one or more of the following areas: networking, cloud technologies, application development and/or cybersecurity
- Deep knowledge of the Kubernetes ecosystem (Docker, Helm, ArgoCD, Terraform)
- Experience with cloud services (AWS/Azure)
- Experience in software languages such as Go, Python, Rust, or C++
- Experience performing data-driven root cause analysis on complex systems
- Demonstrated ability to train peers or customers on the operation of a product
- Computer Science degree or equivalent
Nice to Have
- Experience with managing Kubernetes clusters of hundreds of nodes
- Knowledge of performance improvement techniques, metrics and alerting
- Experience with KubeVirt, qemu, virtualization and hypervisor technologies
- Experience with low-level frameworks, Linux and databases
- Excellent written and verbal communication skills
Compensation & Benefits
- US Salary Range $166,000 - $220,000 USD
- The salary range for this role is an estimate based on a wide range of compensation factors, inclusive of base salary only. Actual salary offer may vary based on (but not limited to) work experience, education and/or training, critical skills, and/or business considerations.
- Highly competitive equity grants are included in the majority of full time offers; and are considered part of Anduril's total compensation package.
- Anduril offers top-tier benefits for full-time employees, including: Benefits At Anduril, we invest in our people. Our comprehensive, competitive benefits package (available at little to no cost to employees) ensures you’re supported in health, recovery, and whatever comes next.