About this role
About the Role Senior Infrastructure Engineer to improve reliability and stability of Webflow’s customer-facing, production infrastructure, serving millions of page views per hour. Our product is used by over 2 million users world-wide across 190 countries, and you’ll help ensure our platform is secure and scalable for these users as tens of thousands of projects are launched on Webflow each month. What You'll Do
- Own and evolve the cloud platform that Webflow's product and engineering teams depend on, including our compute layer, EKS fleet, serverless infrastructure, networking, and cloud operations across AWS and GCP.
- Design and maintain our infrastructure-as-code foundation, evolving the patterns and shared components other teams build on.
- Design and maintain the networking layer that connects Webflow's services, ensuring reliability, security, and scalability across our cloud environments.
- Help make observability data the foundation of how we operate, improving the dashboards, alerts, and SLOs for the infrastructure you own so that issues are caught before customers notice, and on-call pages are actionable instead of noisy.
- Build and maintain AI-powered automation that improves how we manage cloud infrastructure, from policy-as-code and drift detection to LLM-assisted runbook generation.
- Help define the culture of this growing team as it expands its international presence. What We're Looking For
- 5+ years of experience owning and operating cloud infrastructure in a customer-facing environment with little downtime.
- Deep hands-on experience with AWS and a strong stance on what good cloud operations look like.
- Experience managing Kubernetes clusters at scale, including upgrades, node group management, autoscaling, and cluster add-on lifecycle.
- Experience with infrastructure-as-code tools like Pulumi or Terraform; preference for changes made through code, not consoles.
- Experience navigating multi-region or multi-cloud environments on AWS or GCP.
- Stay curious and open to growth, proactively embracing AI and emerging technologies to elevate how we work and deliver faster outcomes. Nice to Have
- Experience with Karpenter, cluster autoscaler, or other Kubernetes-native scaling tooling.
- Experience with OpenTelemetry, Datadog, Prometheus or Grafana.
- Experience building AI-assisted infrastructure tooling, including cost optimization loops, anomaly detection, or policy-as-code with LLM assistance.
- Experience contributing to multi-region architecture including data residency, regional failover, or latency-based routing. Compensation & Benefits
- Equity (RSUs) in our growing, privately held company for every permanent employee.
- Health coverage that actually covers you: comprehensive medical, dental, and vision plans for full-time employees and dependents, with Webflow covering most premiums.
- 12 weeks of paid parental leave for all parents and 6+ weeks of additional paid leave for birthing parents; inclusive care for family planning, menopause, and midlife transitions.
- Flexible vacation, paid holidays, and a sabbatical program to help you recharge and come back inspired.
- Wellness for the whole you: access to mental health resources, therapy and coaching.
- A 401(k) with 100% employer match (up to $6,000/year) in the U.S., and support for retirement savings globally.
- Monthly stipends that flex with your life; localized support for work and wellness expenses—from Wi‑Fi to workouts.
- Bonus for building together: all full-time, permanent, non-commission employees are eligible for our annual WIN bonus program.