freehire launches on Product Hunt on 26 August.

Follow →

Senior Infrastructure Engineer

Open 33d

You'll improve the reliability and stability of Webflow's customer-facing, production infrastructure, which serves millions of page views per hour to over 2 million users worldwide across 190 countries. You'll help ensure the platform is secure and scalable as tens of thousands of projects are launched each month. You'll own the compute layer, EKS fleet, serverless infrastructure, networking, and cloud operations across AWS and GCP, while helping shape the team's infrastructure-as-code foundation and culture as it grows its international presence.

Responsibilities

  • Own and evolve the cloud platform that Webflow's product and engineering teams depend on, including compute layer, EKS fleet, serverless infrastructure, networking, and cloud operations across AWS and GCP
  • Help design and maintain the infrastructure-as-code foundation, evolving patterns and shared components other teams build on
  • Design and maintain the networking layer that connects Webflow's services, ensuring reliability, security, and scalability across cloud environments
  • Improve dashboards, alerts, and SLOs for owned infrastructure so issues are caught before customers notice and on-call pages are actionable
  • Build and maintain AI-powered automation for cloud infrastructure management, including policy-as-code, drift detection, and LLM-assisted runbook generation
  • Help define the culture of the growing team as it expands its international presence

Requirements

  • Background as an infrastructure, site reliability, or cloud engineer with enthusiasm for automation and code, or a software engineer background with deep enthusiasm for cloud infrastructure and distributed systems
  • 5+ years of experience owning and operating cloud infrastructure in a customer-facing environment with little to no downtime
  • Deep hands-on experience with AWS and strong opinions on good cloud operations
  • Experience managing Kubernetes clusters at scale, including upgrades, node group management, autoscaling, and cluster add-on lifecycle
  • Experience with infrastructure-as-code tools like Pulumi or Terraform, with a preference for changes made through code, not consoles
  • Experience navigating multi-region or multi-cloud environments on AWS or GCP
  • Proactive embrace of AI and fluency in emerging technologies
  • Bonus: experience with Karpenter, cluster autoscaler, or other Kubernetes-native scaling tooling
  • Bonus: experience with OpenTelemetry, Datadog, Prometheus or Grafana
  • Bonus: experience building AI-assisted infrastructure tooling, including cost optimization loops, anomaly detection, or policy-as-code with LLM assistance
  • Bonus: experience contributing to multi-region architecture including data residency, regional failover, or latency-based routing

Benefits

  • Equity (RSUs) for every permanent employee
  • Comprehensive medical, dental, and vision plans for full-time employees and their dependents, with most premiums covered
  • 12 weeks of paid parental leave for all parents and 6+ weeks of additional paid leave for birthing parents
  • Inclusive care for family planning, menopause, and midlife transitions
  • Flexible vacation, paid holidays, and a sabbatical program
  • Access to mental health resources, therapy and coaching
  • 401(k) with 100% employer match (up to $6,000/year) in the U.S., and support for retirement savings globally
  • Monthly stipends for work and wellness expenses
  • Annual WIN bonus program for eligible employees

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available