NOC Team Leader
You manage 24/7 global NOC operations, oversee incident response and service uptime, implement monitoring and observability, automate repetitive tasks, lead critical incident management and root cause analysis, develop NOC engineers and SREs, and collaborate with engineering and IT stakeholders.
Responsibilities
- Direct 24/7 global NOC teams
- Manage incident response and service uptime
- Implement monitoring, alerting, and observability tools
- Automate repetitive operational tasks and manual troubleshooting
- Oversee critical incident management
- Ensure timely incident communication
- Perform post-mortem analysis
- Coach, mentor, and develop NOC engineers and SREs
- Collaborate with engineering, development, and IT teams
- Align system performance with business goals and SLA requirements
Requirements
- 3+ years of experience managing technical teams in NOC, SRE, or infrastructure domains
- Experience with AWS, Azure, or GCP
- Experience with Kubernetes
- Experience with Linux or Unix
- Scripting skills in Python, Bash, or Golang
- Experience with Datadog
- Experience with Jenkins or GitLab
- Communication, leadership, and crisis management skills