freehire launches on Product Hunt on 26 August.

Follow →

Kubernetes Platform / DevOps Engineer

Summary

Maintain and improve a Kubernetes-based platform, automate operations, and ensure high availability and security for containerized services like MongoDB, PostgreSQL, and Kafka.

We are looking for an experienced Kubernetes Platform / DevOps Engineer to operate, maintain and continuously improve a production Kubernetes-based platform and the managed services running on it.


You will play a key role in ensuring the platform remains highly available, scalable, secure and resilient, while driving automation, observability and operational improvements across the environment.


This is a hands‑on operational role suited to someone with strong Kubernetes, Linux, automation and production infrastructure experience.


Key Responsibilities



  • Operate, monitor and continuously improve a Kubernetes-based platform and managed services including MongoDB, PostgreSQL and Kafka.

  • Ensure high availability, scalability and resilience through proactive monitoring, alerting and incident management.

  • Perform and automate Day 2 operations, including upgrades, patching, backup and restore, capacity management and maintenance.

  • Investigate production incidents and perform detailed error analysis and Root Cause Analysis (RCA).

  • Implement and maintain observability solutions using Prometheus, Grafana and ELK/OpenSearch.

  • Develop and maintain automation, runbooks and operational processes to improve platform reliability and efficiency.

  • Work closely with engineering teams to continuously improve the platform and operational tooling.

  • Support production environments and participate in an on-call rotation, ensuring reliable 24/7 operation.


Your Experience



  • Degree in Computer Science or equivalent education, combined with several years of professional experience in IT operations.

  • Strong hands‑on experience with Kubernetes and operating container-based platforms and services.

  • Solid expertise in observability, particularly Prometheus, Grafana and ELK/OpenSearch.

  • Strong experience with Linux/system engineering in production environments.

  • Experience with automation and Infrastructure as Code, using technologies such as:

  • Ansible

  • Terraform

  • Experience with CI/CD pipelines, GitLab, Git and Docker.

  • Operational experience with at least one of MongoDB, PostgreSQL or Kafka.

  • Strong troubleshooting and structured problem‑solving skills, including incident investigation and RCA.

  • Ability to quickly understand existing infrastructure and make improvements from day one.

  • Comfortable working independently while collaborating effectively within an agile engineering environment.

  • Resilient and calm under pressure, particularly when dealing with production incidents.


Nice to Have



  • Kubernetes certifications such as CKA or CKAD, or equivalent practical experience.

  • Experience with CNCF technologies such as ArgoCD, Velero, Cilium or Kyverno.

  • Knowledge of security hardening, secrets management and RBAC/IAM.

  • Experience with IT service management / ITIL and 3rd-level support.

  • Previous experience within Service Provider, Telco or Managed Service environments.

  • Willingness to participate in on-call / 24x7 support.

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available