Staff Infrastructure Engineer
The Opportunity at Komodo Health
Komodo Health is growing quickly, and so is the cloud infrastructure underneath it. You will join the Infrastructure team, which builds, operates, and continuously improves the cloud-native stack the entire engineering organization depends on (AWS, Kubernetes, Terraform, ArgoCD, GitHub Actions) along with the shared services layered on top of it: identity and access, API gateway and traffic management, secure network access, our internal developer platform, and the AWS organization and account foundations beneath all of it. Beyond keeping these systems reliable, scalable, and secure, the team drives FinOps discipline and compliance alignment (HIPAA, SOC2). We work in an international context toward one goal: reduce the burden of disease.
This role exists to be the senior technical voice for Komodo's cloud infrastructure and shared services. Several of these platforms transferred to Infrastructure recently, and establishing durable ownership, operating practices, and modernization roadmaps for them is central to the charter. You will set the architecture, standards, and AI-assisted working patterns the rest of the team builds on, and partner closely with Engineering, Security, IT, and Data teams to meet the security and compliance demands of a rapidly evolving healthcare data landscape.
Looking back on your first 12 months at Komodo Health, you will have accomplished…
Established AI-assisted workflows as the team's default operating model, with safe, reviewed, repeatable patterns for infrastructure change, incident triage, and documentation that measurably increase the team's leverage.
Brought our inherited shared services under infrastructure-as-code with documented ownership, runbooks, and a published modernization roadmap, ending our reliance on tribal knowledge for platforms that production depends on.
Delivered measurable cloud cost reduction against our annual savings commitment, leaving behind a reusable framework for weighing cost against reliability and developer experience rather than a one-time sweep.
Improved platform currency and upgrade discipline across the shared-services estate, with version drift and end-of-life exposure tracked, visible, and trending down.
Reduced developer friction and support toil by automating our highest-volume manual request paths and moving them to self-service.
You will accomplish these outcomes through the following responsibilities…
Own the architecture, operating model, and modernization roadmap for cloud infrastructure and shared services, including systems that arrive without clear ownership.
Design, build, and operate AWS and Kubernetes infrastructure as code (Terraform), delivered through GitOps (ArgoCD) and CI/CD (GitHub Actions).
Set the standards and safe patterns for AI-assisted infrastructure work, and raise the team's fluency with them.
Harden our security and compliance posture: least-privilege IAM and RBAC, identity and access management, network boundaries, and secrets hygiene, aligned to SOC2 expectations.
Drive infrastructure optimization across reliability, scalability, developer experience, and cost, and build the decision frameworks the team uses to make those tradeoffs.
Raise the technical bar across the infrastructure team through design review, mentorship, documentation, and enablement, making the engineers around you better.
Participate in and improve the shared US-business-hours on-call rotation, incident response, and the alerting and runbooks behind it
What you bring to Komodo Health (required):
8+ years in infrastructure, cloud, or platform engineering, including deep hands-on AWS experience running production systems at scale where security and cost were first-class concerns.
Infrastructure-as-code proficiency with Terraform as a default working mode rather than an occasional tool.
Kubernetes platform ownership: you have operated the platform itself, not only deployed workloads onto it.
A track record of taking ownership of ambiguous, inherited, or under-documented systems and bringing them to a defensible, well-operated state.
Fluency with AI-assisted engineering tools (Claude, OpenAI Codex, Cursor, GitHub Copilot, AWS Bedrock, or similar) and sound judgment about where they are and are not safe to apply.
Security and compliance fluency in a regulated environment (HIPAA, SOC2, or equivalent).
Staff-level partnership and influence: you are sought out early because your input improves outcomes, and you create leverage through standards, reviews, and enablement rather than personal throughput alone.
Expectations of AI Use in this role (required):
This is an AI-native role, and at this level you are expected to shape how the team works, not only how you work. You will set the standards and safe, repeatable patterns for applying AI assistants and platforms (such as Claude, OpenAI Codex, Cursor, GitHub Copilot, or AWS Bedrock) to Terraform and script authoring, runbook and documentation generation, log and query analysis, and support triage, and you will ensure all generated output is reviewed and validated before it reaches production. Candidates should demonstrate deep expertise in multiple infrastructure domains and sufficient breadth to lead across adjacent domains. No single candidate is expected to have equal depth across every technology or capability listed.
Additional skills and experience we’ll prioritize…
Identity and access platforms: Okta, OIDC, SAML.
API gateway and service networking: Envoy, or similar.
Scripting and automation in Python, Go, or Bash.
FinOps or cloud cost-optimization program experience.
Observability at platform scale: metrics, logs, tracing, and alerting.
Familiarity with cloud data warehouses such as Snowflake.
#LI-Remote
The pay range for each job posting reflects a minimum and maximum range of annual base pay that we reasonably expect to pay for this position within the US. We carefully consider multiple business-related factors when determining compensation, including job-related skills, work experience, geographic work location, relevant training and certifications, business needs and market demands.
The starting annual base pay for this role is listed below. This position may be eligible for performance-based bonuses as determined in the Company’s sole discretion and in accordance with a written agreement or plan. This role may also be eligible for equity awards. In addition, this role is eligible for benefits including, but not limited to, comprehensive health, dental, and vision insurance; flexible time off and holidays; 401(k) with company match; disability insurance and life insurance; and leaves of absence in accordance with applicable state and local laws and regulations and company policy.
San Francisco Bay Area and New York City:
$215,000—$265,000 USD
All Other US Locations:
$187,000—$235,000 USD