Senior Software Engineer, Infrastructure

Commure
Mountain View +3 more
Hybrid

Who this role is best for

Strong fit for infrastructure-focused engineers who build and own internal developer platforms and cloud-native systems, with experience in regulated industries and a preference for hands-on, IC roles in fast-paced environments.

Best fit for

  • Infrastructure engineers with 6+ years of experience in regulated industries and a product mindset for internal tooling
    — “6+ years of software engineering experience in infrastructure, platform, or site reliability engineering roles.
  • Candidates who have designed and operated cloud foundations with IaC and controller-based provisioning
    — “Cloud infrastructure and IaC
  • Individuals with expertise in GitOps, service mesh, and observability stack technologies
    — “modern release workflows using GitOps and progressive delivery (Argo CD, Flux, Kargo)
  • Engineers who have shaped architectural patterns and driven multi-product infrastructure at scale
    — “This is a hands-on IC role with broad scope. You'll make architectural calls, write the code that matters most, and set the patterns other teams build on.

Things to consider

  • Candidates must be prepared to comply with HIPAA and SOC 2 security requirements by default
    — “secure-by-default patterns that meet HIPAA and SOC 2 by default.
  • The role requires full-time commitment and adherence to strict security policies
    — “Employees will act in accordance with the organization’s information security policies

How to stand out

  • Highlight experience in building self-serve developer tooling and internal platforms
    — “Build out the internal developer platform: golden-path templates, self-serve tooling, local development environments, and CI/CD pipelines.
  • Emphasize hands-on experience with Kubernetes operators and Crossplane for controller-based infrastructure
    — “controller-based infrastructure management (Crossplane, Kubernetes operators).
  • Demonstrate mastery of observability stack components like Prometheus and OpenTelemetry
    — “observability stack (Prometheus, Grafana, OpenTelemetry).
  • Showcase your ability to define and implement service mesh and traffic management strategies
    — “Design the traffic and network layer: service mesh, ingress, mTLS, and traffic management (routing, rate limiting, canary, circuit-breaking, RPC).
Pace · SteadyCollaboration · MediumAutonomy · MediumDecision Impact · Team

Derived from job-description analysis by Serendipath's career intelligence engine.

Skills & requirements

Required

AWS

Stack & domain

Cloud InfrastructureIacKubernetesInternal Developer PlatformRelease And DeploymentService MeshTraffic ManagementNetworkingObservabilityZero-trust AccessOn-prem ConnectivityGCPAWSAzureTerraformController-based ProvisioningKubernetes OperatorsCrossplaneArgo CDHelmPrometheusGrafanaOpentelemetryIngressMtlsRoutingRate LimitingCanaryCircuit-breakingRPCHealthcare

About the role

Original posting from Commure via Ashby

At Commure, we're building the AI Operating System for healthcare, the foundation that defines how care is delivered, documented, and financed. Our platform spans the full care journey: Ambient AI and Dictation eliminating documentation burden at the point of care, intelligent Agents automating patient and revenue workflows, and autonomous RCM processing billions in claims, all on a single AI-native platform integrated with 60+ EHRs.

Healthcare carries a $1 trillion administrative burden and we're at the center of transforming it. Today, 500,000+ clinicians across 500+ healthcare organizations nationwide trust Commure to handle $25B+ in annual claims and support over 200 million patient interactions. Our latest $70M raise at a $7B valuation reflects the confidence the market has placed in this mission. We've also been named to the Fortune Future 50 list and the 2026 AI Breakthrough Awards for “Overall NLP Company of the Year.”

Our team works directly alongside clinicians, not through layers of process, which means the gap between what you build and its impact on patient care is immediate. We move fast, deploy daily, and take full ownership from early thinking to production. If you're energized by hard problems, high stakes, and a team that holds itself to a high bar, you'll find your people here.

The future of healthcare is being built right now. Come deliver this transformation.

ABOUT THE ROLE

We're hiring a Senior Software Engineer on the Infrastructure team to own the foundational infrastructure and internal developer platform that every other engineering team at Commure builds on. This is a horizontal team supporting multi-product infrastructure that you will design, build, and operate end-to-end:

  • Cloud infrastructure and IaC
  • Kubernetes fleet and workload orchestration
  • Internal developer platform
  • Release and deployment (GitOps)
  • Service mesh, traffic management, and networking
  • Observability: metrics, logs, and traces
  • Zero-trust access and on-prem connectivity

The stack today runs on public cloud (GCP, AWS, and Azure), with all infrastructure defined as code (Terraform and controller-based). Argo CD drives deployment; Helm handles application packaging; Prometheus, Grafana, and OpenTelemetry power observability. Service mesh is an active build-out, and the shape of it is yours to define.

This is a hands-on IC role with broad scope. You'll make architectural calls, write the code that matters most, and set the patterns other teams build on.

WHAT YOU'LL DO

You'll own several of these verticals within the team's scope end-to-end.

  • Build out the internal developer platform: golden-path templates, self-serve tooling, local development environments, and CI/CD pipelines.
  • Own the cloud foundation: GCP, AWS, and/or Azure infrastructure managed as code with Terraform and controller-based provisioning via Kubernetes operators and Crossplane.
  • Run the Kubernetes fleet: cluster lifecycle, upgrades, autoscaling, node management, and multi-cluster patterns. Shape how services are packaged and deployed with Helm.
  • Design the traffic and network layer: service mesh, ingress, mTLS, and traffic management (routing, rate limiting, canary, circuit-breaking, RPC).
  • Own the release and deployment story with Argo CD. GitOps workflows, progressive delivery (canary, blue-green), rollback safety, and environment promotion patterns.
  • Own the observability stack: OpenTelemetry based instrumentation, metrics (Prometheus), dashboards (Grafana), distributed tracing, logging, unified alerting, templated dashboard, etc.
  • Build out zero-trust access to internal and external systems: VPN, BeyondCorp, short-lived credentials, and on-prem connectivity.
  • Partner with Security on secrets management, policy-as-code, and secure-by-default patterns that meet HIPAA and SOC 2 by default.

WHAT YOU HAVE

  • 6+ years of software engineering experience in infrastructure, platform, or site reliability engineering roles.
  • Experience building internal developer platforms with a strong product mindset (treating developers as customers).
  • Experience with public cloud (GCP, AWS, Azure) and managed cloud services.
  • Experience with on-prem or hybrid environments and data center / cloud migrations.
  • Experience with cloud-native technologies.
  • Experience with Infrastructure-as-Code (Terraform, Pulumi).
  • Experience with plus controller-based infrastructure management (Crossplane, Kubernetes operators).
  • Experience with service mesh technologies and software-defined networking (SDN).
  • Experience with modern release workflows using GitOps and progressive delivery (Argo CD, Flux, Kargo).
  • Experience with observability stack (Prometheus, Grafana, OpenTelemetry).
  • Experience in regulated industries (healthcare, finance) with HIPAA and SOC 2 obligations.

Please be aware that all official communication from us will come exclusively from email addresses ending in @commure.com http://commure.com. Any emails from other domains are not affiliated with our organization.

Employees will act in accordance with the organization’s information security policies, to include but not limited to protecting assets from unauthorized access, disclosure, modification, destruction or interference nor execute particular security processes or activities. Employees will report to the information security office any confirmed or potential events or other risks to the organization. Employees will be required to attest to these requirements upon hire and on an annual basis.

Source: Commure careers (Ashby)

Similar roles