Senior Site Reliability Engineer -AI Infrastructure Operations
About this role
About Nscale Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native startups and global enterprises, from bare metal up through the platform services teams actually build on. Our culture runs on ownership, accountability, and speed. We move with urgency, we tell each other the truth, and everyone here stays close to the infrastructure that makes AI work.
The Role This is a senior SRE role for someone who sets the reliability bar and then pulls the rest of the team up to it. You'll own the hardest problems on the platform: the automation other engineers build on, the services that can't go down, and the design decisions that determine whether either holds up at scale. You'll still carry a pager, but the real job is making sure it fires less, for everyone, over time.
What You'll Do
Read the full description on TalentApply
Create a free account to see the complete job description, how well your CV matches this role, and apply in one click.
AI rewrites and formats your CV so it reads well and gets past screeners.
Get your match percentage for this exact role before you spend time applying.
Send a polished application in one click — no retyping the same details.
Follow every application in one place instead of digging through your inbox.
Free account · No card required
Your next opportunity starts here
Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.