Senior Software Engineer, Infrastructure
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+17 more
Job description
We’re hiring a Senior Software Engineer on the Infrastructure team to own the foundational infrastructure and internal developer platform that every other engineering team at Commure builds on. This is a horizontal team supporting multi-product infrastructure that you will design, build, and operate end-to-end:
- Cloud infrastructure and IaC
- Kubernetes fleet and workload orchestration
- Internal developer platform
- Release and deployment (GitOps)
- Service mesh, traffic management, and networking
- Observability: metrics, logs, and traces
- Zero-trust access and on-prem connectivity
The stack today runs on public cloud (GCP, AWS, and Azure), with all infrastructure defined as code (Terraform and controller-based). Argo CD drives deployment; Helm handles application packaging; Prometheus, Grafana, and OpenTelemetry power observability. Service mesh is an active build-out, and the shape of it is yours to define.
This is a hands-on IC role with broad scope. You’ll make architectural calls, write the code that matters most, and set the patterns other teams build on.
What You’ll Do
You’ll own several of these verticals within the team’s scope end-to-end.
- Build out the internal developer platform: golden-path templates, self-serve tooling, local development environments, and CI/CD pipelines.
- Own the cloud foundation: GCP, AWS, and/or Azure infrastructure managed as code with Terraform and controller-based provisioning via Kubernetes operators and Crossplane.
- Run the Kubernetes fleet: cluster lifecycle, upgrades, autoscaling, node management, and multi-cluster patterns. Shape how services are packaged and deployed with Helm.
- Design the traffic and network layer: service mesh, ingress, mTLS, and traffic management (routing, rate limiting, canary, circuit-breaking, RPC).
- Own the release and deployment story with Argo CD. GitOps workflows, progressive delivery (canary, blue-green), rollback safety, and environment promotion patterns.
- Own the observability stack: OpenTelemetry based instrumentation, metrics (Prometheus), dashboards (Grafana), distributed tracing, logging, unified alerting, templated dashboard, etc.
- Build out zero-trust access to internal and external systems: VPN, BeyondCorp, short-lived credentials, and on-prem connectivity.
- Partner with Security on secrets management, policy-as-code, and secure-by-default patterns that meet HIPAA and SOC 2 by default., Employees will act in accordance with the organization’s information security policies, to include but not limited to protecting assets from unauthorized access, disclosure, modification, destruction or interference nor execute particular security processes or activities. Employees will report to the information security office any confirmed or potential events or other risks to the organization. Employees will be required to attest to these requirements upon hire and on an annual basis.
Requirements
- 6+ years of software engineering experience in infrastructure, platform, or site reliability engineering roles.
- Experience building internal developer platforms with a strong product mindset (treating developers as customers).
- Experience with public cloud (GCP, AWS, Azure) and managed cloud services.
- Experience with on-prem or hybrid environments and data center / cloud migrations.
- Experience with cloud-native technologies.
- Experience with Infrastructure-as-Code (Terraform, Pulumi).
- Experience with plus controller-based infrastructure management (Crossplane, Kubernetes operators).
- Experience with service mesh technologies and software-defined networking (SDN).
- Experience with modern release workflows using GitOps and progressive delivery (Argo CD, Flux, Kargo).
- Experience with observability stack (Prometheus, Grafana, OpenTelemetry).
- Experience in regulated industries (healthcare, finance) with HIPAA and SOC 2 obligations.
Benefits & conditions
Compensation Range: $170K - $220K
About the company
At Commure, we’re building the AI Operating System for healthcare, the foundation that defines how care is delivered, documented, and financed. Our platform spans the full care journey: Ambient AI and Dictation eliminating documentation burden at the point of care, intelligent Agents automating patient and revenue workflows, and autonomous RCM processing billions in claims, all on a single AI-native platform integrated with 60+ EHRs.
Healthcare carries a $1 trillion administrative burden and we’re at the center of transforming it. Today, 500,000+ clinicians across 500+ healthcare organizations nationwide trust Commure to handle $25B+ in annual claims and support over 200 million patient interactions. Our latest $70M raise at a $7B valuation reflects the confidence the market has placed in this mission. We’ve also been named to the Fortune Future 50 list and the 2026 AI Breakthrough Awards for “Overall NLP Company of the Year.”
Our team works directly alongside clinicians, not through layers of process, which means the gap between what you build and its impact on patient care is immediate. We move fast, deploy daily, and take full ownership from early thinking to production. If you’re energized by hard problems, high stakes, and a team that holds itself to a high bar, you’ll find your people here.
The future of healthcare is being built right now. Come deliver this transformation.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fully Remote Software Engineer Jobs
Highest Paying Tech Companies for Developers
Dev Digest 120 - Apple and peers
The Best X (Twitter) Accounts for Developers