freehire launches on Product Hunt on 26 August.

Follow →

Senior DevOps Engineer

Summary

Senior DevOps Engineer builds and automates AWS-based platform infrastructure, Kubernetes clusters, and storage (NetApp) while codifying everything for self-service delivery and L2 support.

About the Role

We are looking for a passionate Senior DevOps Engineer with strong experience in Storage to join our Platform Engineering team. As a core member of the team, you will actively contribute to the full platform ecosystem: infrastructure automation, continuous delivery pipelines, self-service provisioning capabilities, production operations runbooks, and L2 support.

You will bring a broader DevOps culture and mindset, with a solid command of AWS infrastructure, tooling, and modern platform practices including AIOps and advanced Kubernetes.

In addition to your full platform engineering contributions, you will serve as a team's storage domain expert, particularly around NetApp on AWS. This specialization will be especially prominent in the early stages of your onboarding, but it does not define the boundaries of your role. You will investigate and resolve storage-related incidents at L2 level, while also continuing to contribute across all dimensions of the platform.

This is a hands‑on, automation‑first role where you are expected to codify everything, document what you build, and make it available as a self‑service to the engineering teams around you.

Essential Duties and Responsibilities

Platform Engineering (Full Scope)

  • Contribute to infrastructure automation across the AWS platform using Terraform Enterprise, Ansible, Harness, Kubernetes (including operators, CRDs, and cluster lifecycle management), and GitOps practices
  • Support and evolve continuous delivery pipelines, ensuring reliable and repeatable deployments across environments (PRE, SBX, PRD)
  • Build and maintain self‑service capabilities so that developers and engineering teams can autonomously claim and consume infrastructure resources (storage, compute, databases, messaging) through Kubernetes‑native APIs and GitOps workflows, without manual intervention
  • Leverage AIOps practices to improve platform reliability: intelligent alerting, anomaly detection, AI‑assisted incident triage, and automated remediation using tools like AWS DevOps Agent, Datadog and GitHub Copilot
  • Write, maintain, and improve production runbooks to ensure operational procedures are automated, documented, and accessible to the team
  • Provide L2 support on automation systems and pipelines already in place, troubleshooting issues across the platform ecosystem
  • Participate in on‑call rotations and contribute to incident response and post‑mortem processes
  • Collaborate actively with Production Ops, DBA, and application engineering teams on platform improvements and migrations
  • Contribute to FinOps by identifying and implementing AWS cost optimization opportunities
  • Document architecture decisions, operational procedures, and contribute to team knowledge sharing

Storage and Infra (Area of Specialization)

  • Own, automate and improve Kyriba's storage platform (NetApp, S3, EBS), ensuring reliability, performance, and DR readiness
  • Design and implement storage architectures on AWS with high availability and fault tolerance
  • Provide L2 support on storage incidents, leading investigations on performance degradation, availability issues, and data integrity events
  • Integrate storage with Kubernetes workloads (persistent volumes, storage classes, CSI drivers) and backup solutions (AWS Backup, NetApp B&R)
  • Ensure compliance with RPO/RTO objectives and contribute to DR testing and validation
  • Administer and continuously improve the Active Directory and DNS environment, with a focus on performance optimization, security compliance, and alignment with business objectives

Tools Used Daily

AWS, NetApp, Terraform Enterprise, Harness, Ansible, GitHub Copilot, GitHub Actions, ArgoCD, Kargo, Kubernetes, Rancher, Datadog, Splunk, Vault, Jira, Confluence

Required

  • Bachelor's degree in Computer Science, Engineering, or related field
  • 5+ years of hands‑on DevOps experience building and operating a SaaS platform on AWS
  • 3+ years of experience with Infrastructure as Code (Terraform) in an AWS environment
  • 2+ years of experience as a platform engineer, contributing to automation and self‑service ecosystems
  • Solid Kubernetes knowledge.
  • Solid knowledge of NetApp ONTAP (administration, provisioning, troubleshooting)
  • Experience with AWS storage services (S3, EBS, EFS, FSx)
  • Experience integrating storage with Kubernetes workloads (PV, PVC, StorageClass, CSI drivers)
  • Experience providing L2 support on automation systems and/or storage infrastructure
  • Exposure to AIOps practices: intelligent alerting, anomaly detection, AI‑assisted automation (Datadog AI, GitHub Copilot, or equivalent)
  • Familiarity with CI/CD practices and GitOps workflows (ArgoCD, Kargo, GitHub Actions)
  • Experience with monitoring and observability tools (Datadog, Splunk, or equivalent)
  • Strong scripting or development skills in Python, Bash, or Go
  • Passion for automation, documentation, and a platform engineering mindset

Preferred

  • Experience building Kubernetes operators or controllers (controller-runtime, kubebuilder)
  • Familiarity with Cilium or equivalent eBPF‑based CNI for Kubernetes networking and security
  • Experience with claim‑based or self‑service provisioning patterns for infrastructure resources
  • Experience with backup and data protection solutions (AWS backup or Cohesity)
  • Knowledge of disaster recovery design and testing on AWS
  • FinOps awareness and cloud cost optimization experience
  • Familiarity with secret management tools (Vault)
  • NetApp certifications (NCDA, NCIE) are a plus

What We Offer

  • 15% yearly bonus and annual salary increase based on individual performance
  • MacBook Pro or equivalent equipment
  • Access to AI productivity tools (ChatGPT, Copilot)
  • Professional development: Coursera, Pluralsight, LinkedIn Learning, conference attendance (e.g., KubeCon)
  • International collaboration across DevOps, SRE, and Engineering teams
  • Medical, sports, and life insurance
  • Equity Incentive Plan participation

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available