About this role
Job title: Performance Engineer
About the Role As a Performance Engineer at Anthropic, you will identify novel systems problems that arise when running machine learning algorithms at scale and build systems that optimize throughput and robustness for our largest distributed systems. You should be excited to grow into an ML expert while solving high-impact, real-world problems.
What You'll Do
- Identify and solve novel systems problems that emerge when running ML at scale, and develop systems that optimize throughput and robustness of our largest distributed systems.
- Tackle large-scale systems challenges and grow to become an expert in ML.
- Implement low-latency, high-throughput sampling for large language models; develop GPU kernels for low-precision inference; write a custom load-balancing algorithm to optimize serving efficiency; design and implement fault-tolerant distributed systems with complex network topologies; debug kernel-level network latency spikes in containerized environments.
- Build quantitative models of system performance and collaborate with researchers, engineers, policy experts, and business leaders to drive impact.
- Work in a collaborative environment that values pair programming, flexibility, and a focus on societal impact.
What We're Looking For
- Significant software engineering or machine learning experience, particularly at supercomputing scale.
- Results-oriented with a bias toward flexibility and impact.
- Willingness to pick up tasks outside your job description, enjoy pair programming, and want to learn more about ML research.
- Care about the societal impacts of your work.
- Education: Bachelor's degree or an equivalent combination of education, training, and/or experience; field relevant to the role; minimum years of experience will correlate with internal job level requirements.
- Location-based hybrid policy: in-office presence at least 25% of the time.
- Visa sponsorship: Sponsorship is available for some roles; not guaranteed for every candidate.
Nice to Have
- Strong candidates may also have experience with:
- High-performance, large-scale ML systems
- GPU/Accelerator programming
- ML framework internals
- OS internals
- Language modeling with transformers
Compensation & Benefits
- Salary range: USD 280,000 - 850,000 per year
- Benefits include competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a collaborative office space in San Francisco.