Senior Software Engineer, Infrastructure
Summary
Commure is hiring a Senior Software Engineer on its Infrastructure team in Mountain View to design, build, and operate the cloud foundation and internal developer platform behind its healthcare AI platform. Day-to-day work spans GCP/AWS/Azure with Terraform, Kubernetes fleet management, Argo CD GitOps, service mesh, and observability with Prometheus, Grafana, and OpenTelemetry.
At Commure, we're building the AI Operating System for healthcare, the foundation that defines how care is delivered, documented, and financed. Our platform spans the full care journey: Ambient AI and Dictation eliminating documentation burden at the point of care, intelligent Agents automating patient and revenue workflows, and autonomous RCM processing billions in claims, all on a single AI-native platform integrated with 60+ EHRs.
Healthcare carries a $1 trillion administrative burden and we're at the center of transforming it. Today, 500,000+ clinicians across 500+ healthcare organizations nationwide trust Commure to handle $25B+ in annual claims and support over 200 million patient interactions. Our latest $70M raise at a $7B valuation reflects the confidence the market has placed in this mission. We've also been named to the Fortune Future 50 list and the 2026 AI Breakthrough Awards for “Overall NLP Company of the Year.”
Our team works directly alongside clinicians, not through layers of process, which means the gap between what you build and its impact on patient care is immediate. We move fast, deploy daily, and take full ownership from early thinking to production. If you're energized by hard problems, high stakes, and a team that holds itself to a high bar, you'll find your people here.
The future of healthcare is being built right now. Come deliver this transformation.
About the Role
We're hiring a Senior Software Engineer on the Infrastructure team to own the foundational infrastructure and internal developer platform that every other engineering team at Commure builds on. This is a horizontal team supporting multi-product infrastructure that you will design, build, and operate end-to-end:
Cloud infrastructure and IaC
Kubernetes fleet and workload orchestration
Internal developer platform
Release and deployment (GitOps)
Service mesh, traffic management, and networking
Observability: metrics, logs, and traces
Zero-trust access and on-prem connectivity
The stack today runs on public cloud (GCP, AWS, and Azure), with all infrastructure defined as code (Terraform and controller-based). Argo CD drives deployment; Helm handles application packaging; Prometheus, Grafana, and OpenTelemetry power observability. Service mesh is an active build-out, and the shape of it is yours to define.
This is a hands-on IC role with broad scope. You'll make architectural calls, write the code that matters most, and set the patterns other teams build on.
What You'll Do
You'll own several of these verticals within the team's scope end-to-end.
Build out the internal developer platform: golden-path templates, self-serve tooling, local development environments, and CI/CD pipelines.
Own the cloud foundation: GCP, AWS, and/or Azure infrastructure managed as code with Terraform and controller-based provisioning via Kubernetes operators and Crossplane.
Run the Kubernetes fleet: cluster lifecycle, upgrades, autoscaling, node management, and multi-cluster patterns. Shape how services are packaged and deployed with Helm.
Design the traffic and network layer: service mesh, ingress, mTLS, and traffic management (routing, rate limiting, canary, circuit-breaking, RPC).
Own the release and deployment story with Argo CD. GitOps workflows, progressive delivery (canary, blue-green), rollback safety, and environment promotion patterns.
Own the observability stack: OpenTelemetry based instrumentation, metrics (Prometheus), dashboards (Grafana), distributed tracing, logging, unified alerting, templated dashboard, etc.
Build out zero-trust access to internal and external systems: VPN, BeyondCorp, short-lived credentials, and on-prem connectivity.
Partner with Security on secrets management, policy-as-code, and secure-by-default patterns that meet HIPAA and SOC 2 by default.
What You Have
6+ years of software engineering experience in infrastructure, platform, or site reliability engineering roles.
Experience building internal developer platforms with a strong product mindset (treating developers as customers).
Experience with public cloud (GCP, AWS, Azure) and managed cloud services.
Experience with on-prem or hybrid environments and data center / cloud migrations.
Experience with cloud-native technologies.
Experience with Infrastructure-as-Code (Terraform, Pulumi).
Experience with plus controller-based infrastructure management (Crossplane, Kubernetes operators).
Experience with service mesh technologies and software-defined networking (SDN).
Experience with modern release workflows using GitOps and progressive delivery (Argo CD, Flux, Kargo).
Experience with observability stack (Prometheus, Grafana, OpenTelemetry).
Experience in regulated industries (healthcare, finance) with HIPAA and SOC 2 obligations.
Please be aware that all official communication from us will come exclusively from email addresses ending in @. Any emails from other domains are not affiliated with our organization.
Employees will act in accordance with the organization’s information security policies, to include but not limited to protecting assets from unauthorized access, disclosure, modification, destruction or interference nor execute particular security processes or activities. Employees will report to the information security office any confirmed or potential events or other risks to the organization. Employees will be required to attest to these requirements upon hire and on an annual basis.