Site Reliability Engineer
Summary
Design and maintain Kubernetes-based cloud infrastructure, automate deployments with Terraform, and optimize CI/CD pipelines for reliability and performance.
You will design, implement, and maintain scalable Kubernetes-based cloud infrastructure. You will automate deployments with Terraform, troubleshoot production issues, build monitoring and logging solutions, optimize CI/CD workflows, apply AWS services and networking principles, and improve reliability, security, observability, and performance across live environments.
Responsibilities
- Design, implement, and maintain Kubernetes-based cloud-native environments
- Implement and manage Terraform-driven infrastructure automation
- Troubleshoot and resolve issues in live production environments
- Build and optimize monitoring, logging, and observability solutions
- Contribute to internal engineering standards, methodologies, and automation frameworks
- Evaluate AWS and DevOps advancements and introduce tools that improve speed, efficiency, and scalability
Requirements
- Proven experience as a Site Reliability Engineer, DevOps Engineer, or in a similar role
- Hands-on experience deploying and managing Kubernetes-based applications
- Strong expertise in Terraform and infrastructure automation
- Experience designing and implementing CI/CD pipelines with GitLab CI/CD or GitHub Actions
- Solid understanding of AWS services including EC2, VPC, S3, and IAM
- Knowledge of TCP/IP, DNS, and HTTP/S networking protocols and best practices
- Strong written and verbal communication skills
Benefits
- Flexible working hours and workplace
- Open vacation policy