Senior DevOps Engineer (Platform & Reliability)
Summary
Senior DevOps Engineer maintains and improves a high-reliability platform supporting AI/ML workloads and real-time voice infrastructure using Kubernetes, Terraform, and CI/CD.
The Role
Incoho Inc., is hiring a Senior DevOps Engineer to manage the platform, infrastructure, and reliability of a production system spanning application services, AI/ML workloads, and real-time voice infrastructure.
The successful candidate will be replacing strong DevOps leader—not building from scratch.
The system works. CI/CD is in place. Observability is mature.
The job is to maintain and improve a platform operating near 5-nines reliability by:
reducing incidents (not just responding to them)
increasing system efficiency
scaling infrastructure to support Incoho Inc.’s growth
This is not a support or ticket-driven role.
The successful candidate will:
Own reliability end-to-end
Make architectural decisions with real consequences
Improve existing systems and build new ones where needed
Operate in ambiguity without waiting for direction
What We’re Looking For
Experience
5–10+ years in DevOps, SRE, or Infrastructure Engineering
Proven ownership of production systems at scale
Experience with multi-region, high-availability systems
Experience in hybrid environments (cloud + on-prem preferred)
Technical Depth
Kubernetes / containerized systems
Terraform / Ansible (Infrastructure as Code)
CI/CD systems (GitHub Actions preferred)
Networking fundamentals (TCP/IP, DNS, load balancing, iptables)
Write code (Python, Go, or similar)
Understand event-driven architectures
Have real-time or low-latency experience or strong interest in learning.