freehire launches on Product Hunt on 26 August.

Follow →

Staff Infrastructure Engineer

Open 26d

You will own the design, operation, and maturity of Kubernetes platforms and the service mesh ecosystem. You will establish standards for cluster configuration, security, traffic management, observability, and reliability. You will build infrastructure-as-code and operational tooling, lead responses to Kubernetes and mesh incidents, identify root causes, and prevent repeat failures. You will define SLOs, evaluate infrastructure platforms, partner on platform needs, and mentor engineers on Kubernetes internals, mesh design, and distributed systems.

Responsibilities

  • Design and operate the Kubernetes platform across EKS clusters.
  • Set standards for cluster configuration, workload isolation, resource management, and cost optimization.
  • Build runbooks and automation for predictable operations.
  • Own and mature the Kong Mesh ecosystem into a production-grade platform.
  • Define service-to-service security, traffic-management, and observability patterns.
  • Implement zero-trust networking, certificate lifecycle, and mTLS policies.
  • Build infrastructure-as-code, automation, and tooling to reduce toil.
  • Define SLOs, monitor reliability, and maintain platform quality.
  • Lead incident response for Kubernetes and service mesh issues.
  • Identify root causes, remediate incidents, and prevent recurring failures.
  • Mentor engineers on Kubernetes internals, service mesh design, and distributed systems.
  • Evaluate infrastructure tools and platforms.
  • Contribute to observability, cost optimization, and resilience initiatives.

Requirements

  • 7+ years of experience in platform infrastructure, SRE, cloud infrastructure, or related work.
  • Deep hands-on Kubernetes experience, including cluster architecture, multi-cluster operations, scheduling, networking, security, and debugging.
  • Expertise with service mesh platforms such as Kong Mesh, Istio, Linkerd, or Envoy-based systems.
  • Experience with mTLS, traffic policies, and zero-trust networking.
  • Working knowledge of AWS, including VPCs, networking, security, and hybrid environments such as Outposts.
  • Understanding of distributed systems and service-to-service communication tradeoffs at scale.
  • Experience defining and tracking SLOs and SLIs for infrastructure services.
  • Ability to code in a modern language and build tools and automation.
  • Experience driving operational improvements through automation.
  • Experience mentoring and influencing engineers.
  • Ability to communicate technical constraints to non-technical stakeholders.

Benefits

  • Health plans, including fertility and family-planning programs, mental health support, and fitness benefits.
  • Paid time off and sick leave.
  • 401(k) matching up to 5%.
  • Commuter benefits.
  • Pet insurance.
  • Medical, vision, dental, life, and disability insurance.
  • Paid personal time off.
  • 14 paid company holidays.
  • Short-term or long-term incentive compensation, including cash bonuses and stock program participation.

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available