ForgeApply
Try it free

ForgeApply · Job listing

Lead Infrastructure Engineer

Onoshealth

US$200k – $275khybrid

See all 6 open roles at Onoshealth

Tailor your resume for this Onoshealth job in about a minute.

ForgeApply rewrites your resume for this exact posting, then autofills the application on Onoshealth's site with it. You review everything before it's sent. Free trial, no card required.

About this role

ABOUT ONOS HEALTH

Onos Health’s mission is simple but ambitious: ensure every healthcare dollar goes toward delivering the highest quality care. Today, 30% of total U.S. healthcare spending is wasted due to ineffective care and administrative burden caused by misalignment between providers and payers.

Onos is addressing this by building the largest AI-driven healthcare data platform. Our models enables payers to make faster, more accurate decisions across their populations. By guiding members to the right care, Onos is channeling more dollars to high-quality care that drives better outcomes while making healthcare more affordable.

Onos is well-funded by some of the best healthcare investors and is working with the nation’s largest health plans. Come join a category-defining company and help reimagine healthcare for the better.

WHY ONOS?

- Meaningful impact: Help fix what is fundamentally broken in healthcare

- Direct collaboration: Work alongside experienced founders with deep healthcare and data expertise

- Culture: Join a high-performing, transparent, and results-oriented team

- Ownership: Significant responsibility and autonomy from day one

- Opportunity: Play a pivotal role in building a fast-growing, category-defining healthcare AI company

THE ROLE

We're seeking an experienced infrastructure engineer to become our first dedicated platform hire and the owner of the infrastructure the Onos platform runs on. Onos is in production with the nation's largest health plans, which comes with contractual uptime SLAs, disaster recovery commitments, and a security bar (SOC 2, HIPAA) our clients audit. Until now this has been carried collectively by our product engineers and founders — you'll own it end to end. We build heavily with AI coding agents, so much of your leverage will come from specifying work well and directing agents rather than typing every line yourself; prior tech lead or engineering management experience translates directly. As an early team member, you'll set the patterns every future platform engineer at Onos inherits. This role is a hybrid role based in San Francisco, where you'll be expected to work at our office in person 3 times a week.

WHAT YOU'LL BE DOING AT ONOS:

- Own our availability, disaster recovery, and backup commitments to enterprise clients — multi-region failover architecture and the recovery exercises that prove it

- Stand up production monitoring, alerting, SLOs, and our on-call rotation and incident response process

- Own the technical controls behind SOC 2 and HIPAA: cloud security posture (AWS org guardrails, IAM least-privilege, KMS/encryption), vulnerability remediation, and continuous audit evidence through Vanta, with a path toward HITRUST

- Build the CI/CD pipelines, Terraform/IaC foundations, preview environments, and test infrastructure the whole team ships on

- Build the guardrails that let AI coding agents ship safely — policy-as-code, deploy verification, and agent-operated operations tooling

- Set the strategy and operating rhythm for platform work: priorities, status, and what we deliberately defer

TECHNICAL CHALLENGES AT ONOS:

- Architect AI SRE agents to ensure up-to-date compliance and reliability, enabling engineers to work more effectively and strategically

- Right-size enterprise-grade reliability: SLOs and alerting you can trust without drowning a small team in pager noise

- Turn compliance into continuously verified infrastructure — security controls and audit evidence as code

- Scale a CI/CD and environments platform where AI agents, not just humans, are the primary users

- Support multi-region disaster recovery with defined RTO/RPO targets and immutable, restore-tested backups for a multi-tenant healthcare platform

TECH STACK:

At Onos, we work with a modern tech stack where we continuously evaluate and adopt cutting-edge technologies as we scale.

- Infrastructure/Systems: AWS (ECS, Bedrock, Glue, etc.), Langfuse, Terraform

- Languages/Frameworks - Backend: Python, Django, Celery / Celery Beat, django-ninja, django-tenants - Frontend: NextJS, Typescript, Tanstack Query, Shadcn UI, Zod, Nuqs

- Database/Storage: PostgreSQL (AWS RDS), S3, Clickhouse

- Development Tools: Github, Linear, Claude Code, Codex, CoderabbitAI

WHAT WE'RE LOOKING FOR:

- Deep AWS experience — you've owned production cloud infrastructure end to end (IAM, networking, KMS, containers, managed databases)

- Strong Terraform/IaC and CI/CD expertise; you treat pipelines and environments as products with users

- Taken a company through at least one SOC 2 (or HITRUST/ISO 27001) audit cycle with your hands on the technical controls

- SRE fundamentals — SLOs, incident management, DR design — with the judgment to right-size reliability for a startup with enterprise contracts

- Prior experience leading engineering teams as a tech lead or engineering manager — you can break down ambiguous goals, delegate (to humans or AI agents), and communicate crisply with non-technical stakeholders

- Customer obsessed and motivated to make an impact in the healthcare space

BONUS POINTS IF YOU HAVE:

- Significant experience in healthcare or another regulated industry, with HIPAA fluency

- Policy-as-code (OPA, Kyverno) or compliance automation (e.g., Vanta, Drata) experience

- Been the first infrastructure hire, or helped found or lead a platform team

- Built internal tooling or infrastructure for LLM/agent systems

BENEFITS AND PERKS

- Hybrid arrangement: 3 days/week at San Francisco office (Financial District)

- Unlimited vacation policy

- Paid parental leave

- Medical, dental, and vision insurance

- Pre-tax commuter benefits

- 401(k)

- Significant equity as an early employee

- Direct mentorship from experienced founders

- Ground-floor opportunity to help build a team and culture

- Regular team events and offsites

- Company-provided equipment and home office set

Salary insight

The midpoint of this range ($238k) is about 19% above the median disclosed salary for San Francisco roles listed on ForgeApply ($200k across 8,783 jobs).

See full DevOps / SRE salary data for San Francisco

Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.

Tailor your resume for this Onoshealth role before you apply.

Tailor my resume for this job

Similar jobs

More like this: DevOps & SRE Jobs · DevOps & SRE Jobs in San Francisco · Browse all jobs

Free ATS checker · How to Tailor Your Resume to a Job Description (Step by Step)