Principal Infrastructure Engineer, AI Cluster Performance & Validation
About this role
Overview
As a Principal Infrastructure Engineer, AI Cluster Performance & Validation, you will be a critical member of the AI Infrastructure Operations team, responsible for ensuring the acceptance, performance, and scalability of our cutting-edge AI and High-Performance Computing (HPC) environments. Leveraging software engineering and testing principles, you will focus on building and maintaining the control plane, tooling, and automation that supports performance and validation testing of large-scale AI clusters. Your work will directly translate into higher system availability, compute optimization and reduced operational costs.
Read the full description on TalentApply
Create a free account to see the complete job description, how well your CV matches this role, and apply in one click.
AI rewrites and formats your CV so it reads well and gets past screeners.
Get your match percentage for this exact role before you spend time applying.
Send a polished application in one click — no retyping the same details.
Follow every application in one place instead of digging through your inbox.
Free account · No card required
Your next opportunity starts here
Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.