Site reliability roles keeping production systems scalable and resilient.
Builds and maintains scalable, automated cloud infrastructure for Databricks' IT services using Python, Terraform, and Kubernetes, ensuring high availability and observability.
Verisign helps enable the security, stability, and resiliency of the internet. We are a trusted provider of internet infrastructure services for the networked world and deliver unmatched performance in domain name syste…
Builds and maintains mission-critical platforms that accelerate vehicle software delivery, testing, and operations for SpaceX missions, using Python, Linux, and infrastructure-as-code tools.
Build and maintain mission-critical platforms that accelerate vehicle software delivery, testing, and operations for SpaceX missions, using Python, Linux, and SRE practices.
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this pos…
Manage HPC server infrastructure, networks, and storage to accelerate rocket engine development. Work with Linux, Windows, Infiniband, and automation tools like Ansible to support engineering applications.
Designs, scales, and automates high-performance computing infrastructure for SpaceX's Starlink silicon engineering, enabling faster chip design iterations and simulations.
Senior Site Reliability Engineer at SpaceX building and maintaining mission-critical platforms that power vehicle software delivery, testing, and operations for Falcon 9, Starship, Dragon, and Starlink using Python, Linux, and cloud infrastructure tools.
Designs and maintains scalable, secure infrastructure for classified systems, including Kubernetes clusters, Linux servers, and GPU-accelerated workloads, while collaborating with software teams to ensure reliability and performance.
Senior Site Reliability Engineer at SpaceX builds and maintains secure, scalable infrastructure for Starshield’s national security satellite systems using Kubernetes, Linux, and automation tools.
Designs, operates, and scales Kubernetes-based infrastructure for SpaceX's Starshield satellite constellation, automating deployments and ensuring high availability for government satellite systems.
Designs, operates, and scales Kubernetes-based infrastructure for SpaceX's Starshield satellite constellation, automating deployments and ensuring high availability for national security systems.
Senior Site Reliability Engineer at SpaceX’s Starshield program, building and maintaining secure, scalable infrastructure for national security satellite systems using Linux, Kubernetes, and cloud automation.
Designs and maintains highly reliable, secure infrastructure for SpaceX’s Starshield satellite program, focusing on Kubernetes, Linux, and automation to support national security missions.
Designs, operates, and scales the software and network infrastructure powering SpaceX’s Starlink satellite internet constellation, using Linux, Kubernetes, and automation tools.
Designs, operates, and scales Kubernetes-based infrastructure for Starlink’s satellite internet network, ensuring high availability for millions of users worldwide.
Maintain and scale mission-critical infrastructure for SpaceX’s Guidance, Navigation, and Control teams, including HPC clusters and automated pipelines, using Linux, Python, and DevOps tools.