Point your AI agent at freehire and let it find you a job.

Get the CLI →

Goodfit

NewBe an early applicant

Senior DevOps/Lead DevOps

Posted
Discussion

Summary

A senior, consulting-heavy DevOps role for someone with 10+ years designing cloud solutions across AWS, Azure, and GCP. Day to day: advising clients, architecting secure and cost-effective infrastructure, and hands-on work with CI/CD, Terraform, Kubernetes, cloud networking, observability, and disaster recovery.

Experience: 10+ Years
Location: (Location)
Work Mode: (Remote/Hybrid/Onsite)
Employment Type: Full-Time

Role Overview

We are looking for a highly experienced Senior DevOps / Cloud Consultant with 10+ years of experience in DevOps, Cloud Infrastructure, Solution Design, and Production Engineering.

The ideal candidate should have deep knowledge of AWS, GCP, and Azure, very strong networking fundamentals, and the ability to understand client requirements and design multiple technical approaches/solutions based on business, technical, security, scalability, availability, and cost considerations.

This role requires a strong consulting and architecture mindset combined with hands-on experience in CI/CD, Observability, Disaster Recovery, cloud infrastructure, automation, and production environments.

Key Responsibilities

Demonstrate detailed and hands-on knowledge of AWS, Microsoft Azure, and Google Cloud Platform (GCP).

Design and implement scalable, secure, highly available, and cost-effective cloud solutions.

Evaluate different cloud approaches based on client requirements and recommend the most suitable architecture.

Work on cloud migration, modernization, hybrid-cloud, and multi-cloud initiatives.

Understand and compare equivalent services across AWS, Azure, and GCP.

Provide technical guidance on cloud architecture, infrastructure, security, scalability, and reliability.

Possess strong and in-depth knowledge of networking concepts and their implementation in cloud environments.

Design and troubleshoot cloud networking architectures including:

  • VPC / VNet
  • Subnets
  • Routing and Route Tables
  • NAT / Internet Gateways
  • Load Balancers
  • DNS
  • Firewalls
  • VPN
  • Direct Connect / ExpressRoute / Cloud Interconnect
  • Private Endpoints / Private Connectivity
  • Security Groups / Network ACLs

Troubleshoot complex network connectivity, latency, routing, and security-related issues.

Design secure network architectures for production and enterprise environments.

3. DevOps Consulting & Solution Design

Act as a DevOps / Cloud Consultant and work directly with clients and technical stakeholders to understand their requirements.

Analyse business and technical requirements and propose multiple solution approaches.

Evaluate solutions based on scalability, performance, availability, security, maintainability, operational complexity, and cost.

Prepare high-level and low-level solution designs and technical recommendations.

Identify technical risks, dependencies, and challenges before implementation.

Guide engineering teams in implementing recommended solutions.

Conduct architecture and technical reviews and establish DevOps best practices.

Provide consulting support for complex DevOps and cloud transformation initiatives.

4. CI/CD – End-to-End Production Experience

Design, implement, and manage end-to-end CI/CD pipelines for enterprise and production environments.

Develop automated pipelines covering source control, build, testing, security scanning, artifact management, deployment, and release.

Work with tools such as Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps, AWS CodePipeline, or equivalent.

Implement deployment strategies such as Blue-Green, Canary, Rolling, and automated rollback.

Troubleshoot pipeline and deployment failures in production environments.

Establish CI/CD standards, governance, and best practices.

5. Observability & Monitoring

Design and implement end-to-end observability solutions for production environments.

Establish monitoring, logging, tracing, alerting, and incident-management practices.

Work with tools such as:

  • Grafana
  • ELK / EFK
  • OpenSearch
  • CloudWatch
  • Application Performance Monitoring (APM) tools

Define meaningful SLIs, SLOs, and operational metrics.

Analyse production incidents and identify root causes.

Improve system reliability, performance, and operational visibility.

6. Disaster Recovery & Business Continuity

Design and implement end-to-end Disaster Recovery (DR) solutions for production environments.

Define appropriate RPO and RTO based on business requirements.

Design backup, replication, failover, and recovery strategies across cloud environments.

Plan and execute DR drills and validate recovery procedures.

Identify single points of failure and implement high-availability solutions.

Troubleshoot and resolve production availability and recovery challenges.

7. Production Engineering & Reliability

Provide technical support for mission-critical production environments.

Troubleshoot complex infrastructure, deployment, networking, performance, and availability issues.

Lead root-cause analysis for critical production incidents.

Drive automation to reduce manual operational activities.

Improve system reliability, scalability, availability, and performance.

Participate in incident management and ensure effective preventive and corrective actions.

Required Technical Skills

Cloud

Networking

Strong knowledge of TCP/IP, DNS, HTTP/HTTPS, routing, load balancing, VPN, firewalls, proxies, subnets, NAT, and network security

Strong experience with cloud networking across AWS, Azure, and GCP

DevOps

Strong understanding of DevOps methodologies and practices

CI/CD pipeline design and implementation

Production deployment and release management

Infrastructure as Code & Automation

Strong experience with Terraform

Experience with Ansible or equivalent automation tools

Strong scripting knowledge using Python, Bash, or PowerShell

Containers & Orchestration

Strong understanding of Docker and Kubernetes

Experience with EKS, AKS, and/or GKE is preferred

Observability

Grafana

ELK / OpenSearch

Cloud-native monitoring and logging platforms

Security

IAM / RBAC

Secrets management

Encryption

TLS/SSL

Network security

DevSecOps practices

Required Experience

10+ years of overall experience in DevOps, Cloud, Infrastructure, SRE, or related engineering roles.

Strong hands-on experience across AWS, Azure, and GCP.

Strong networking and cloud infrastructure expertise.

Proven experience working as a DevOps / Cloud Consultant.

Experience designing multiple technical solutions based on client requirements.

Strong experience with end-to-end CI/CD, Observability, and Disaster Recovery.

Experience supporting and troubleshooting enterprise production environments.

Strong stakeholder and client-facing communication skills.

Strong consulting and solution-design mindset

Excellent analytical and problem-solving skills

Ability to translate business requirements into technical solutions

Ability to present and compare multiple technical approaches

Strong communication and client-management skills

Ability to lead technical discussions and architecture reviews

Strong production troubleshooting and incident-management skills

Ability to work independently and take end-to-end ownership

Preferred Certifications

AWS Certified Solutions Architect / DevOps Engineer

Terraform Associate

Ideal Candidate Profile

We are looking for a 10+ years experienced Senior DevOps / Cloud Consultant who is equally comfortable in client discussions, solution architecture, and hands-on technical implementation.

Skills

See also

DevOps jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available