AI Inference Core Senior Software Engineer for Platform and DevOps
Summary
Builds and operates the platform layer for Cerebras' engineering infrastructure: maintaining CI/CD pipelines and Kubernetes systems, automating deployments and self-service workflows, and improving reliability, observability, and debugging across distributed systems using Python/Shell, cloud platforms, and Linux.
You will build and operate the platform layer for engineering infrastructure. You will maintain CI/CD and Kubernetes systems, automate deployments and self-service workflows, improve reliability and observability, debug cross-system failures, perform root-cause analysis, and implement durable platform improvements.
Responsibilities
- Design, build, and maintain CI/CD systems for build, test, integration, qualification, and release workflows
- Build and operate Kubernetes-based platforms and services
- Develop deployment systems, internal tools, and self-service workflows
- Improve infrastructure reliability, capacity, performance, cost efficiency, monitoring, and operational readiness
- Debug issues across CI pipelines, Kubernetes, networking, storage, authentication, operating systems, and distributed applications
- Perform root-cause analysis and implement lasting fixes
- Deliver scalable infrastructure solutions with partner teams
Requirements
- 5+ years of professional experience in platform engineering, DevOps, infrastructure engineering, site reliability engineering, or software engineering
- Experience building or maintaining CI/CD pipelines and automated delivery workflows
- Experience operating Kubernetes and containerized services
- Experience with a major cloud platform and programmatic infrastructure provisioning
- Understanding of Linux or Unix fundamentals
- Understanding of DNS, routing, load balancing, proxies, ports, TLS, and service connectivity
- Proficiency in Python, Shell, or another infrastructure-automation language
- Experience with monitoring, logging, alerting, dashboards, and incident investigation