ForgeApply
Try it free

ForgeApply · Job listing

Senior Software Engineer, Cloud Infrastructure

Decagon

San Francisco, US$200k – $400konsite

See all 104 open roles at Decagon

Tailor your resume for this Decagon job in about a minute.

ForgeApply rewrites your resume for this exact posting, then autofills the application on Decagon's site with it. You review everything before it's sent. Free trial, no card required.

About this role

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.

We’re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We’re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.

We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values — Just Get It Done, Invent What Customers Want, Winner’s Mindset, and The Polymath Principle — shape how we work and grow as a team.

ABOUT THE TEAM

The Infrastructure team builds and operates the foundations that power Decagon: networking, data, ML serving, developer platform, and real‑time voice. We partner closely with product, data, and ML to deliver high‑scale, low‑latency systems with clear SLOs and great developer ergonomics.

Read more about the infra team's work here: https://decagon.ai/blog/what-an-air-gapped-ai-deployment-actually-requires

ABOUT THE ROLE

Decagon builds agentic AI that resolves customer support conversations end to end, for companies ranging from fast-growing startups to some of the largest financial institutions in the world. Keeping that system fast, reliable, and secure (across our multi-tenant cloud and inside customers' own locked-down environments) is an infrastructure problem, and that's the problem this role owns.

You'll build the platforms and abstractions our product teams ship on, and you'll architect and operate the deployments that run our agents inside enterprise customer clouds, where security, compliance, and operational rigor matter as much as speed. The work is core infrastructure at heart: reliability, CI/CD, deployment automation, on-call. You'll be joining a team with real infrastructure and momentum already in place, but there's no established playbook for running agentic systems at this scale, so a big part of the job is figuring it out as the technology shifts and usage grows by orders of magnitude.

WHAT YOU'LL DO

- Build the platform - Design the development and production platforms that power our products, and the abstractions over cloud infrastructure, Kubernetes, and networking that let engineers ship without becoming infrastructure experts. - Make sure it all scales to the next order of magnitude as usage grows.

- Own enterprise deployments - Take end-to-end ownership of deployment architecture in customer-owned cloud environments (VPC configuration, permissioning, networking, provisioning) and the full lifecycle that follows: setup, upgrades, scaling, and incident support. - Build the runbooks and automation that make it repeatable.

- Keep agentic workloads reliable - Treat monitoring, alerting, and rollback as first-class parts of anything you ship, not afterthoughts. - Own the reliability of the systems our AI agents depend on in production, where latency, availability, and graceful degradation directly shape the customer experience.

- Partner across boundaries - Work directly with customers' platform, security, and DevOps teams to navigate their infrastructure and compliance constraints, and with our Product, Security, Sales, and Customer Success teams to turn customer requirements into concrete deployment plans.

YOUR BACKGROUND LOOKS SOMETHING LIKE THIS

- 4+ years building and operating core infrastructure, platform engineering, or infrastructure/DevOps, ideally with some customer-facing deployment experience.

- Deep experience with a major cloud provider (GCP, AWS, or Azure), along with Terraform and Kubernetes at scale.

- Strong grasp of cloud networking fundamentals (VPCs, IAM, DNS, load balancing) and how they surface as real deployment constraints.

- A track record operating production systems reliably: monitoring, on-call, incident response, and reasoning about failure modes up front.

- Comfort navigating ambiguity across a range of stakeholders, from engineers to security and compliance teams, and turning those conversations into actionable plans.

- Clear technical writing and a track record of driving adoption across teams.

- Comfortable in a fast-moving environment with rapid change.

EVEN BETTER IF YOU HAVE

- Experience managing deployments in customer-owned cloud environments, including security reviews, compliance requirements, and change management.

- Experience building internal platforms or paved roads: service templates, self-serve environments, CI/CD pipeline design, deployment automation.

- Familiarity with observability and incident management in distributed systems (Prometheus, Grafana, Datadog, or similar).

- Infrastructure-as-code with a security-minded approach to supply chain (provenance, secrets, least privilege).

- Experience operating latency-sensitive or ML/AI-serving workloads in production.

- Experience using AI-assisted tooling to make yourself and your team dramatically more effective.

COMPENSATION

$200K – $400K + Offers Equity

This range reflects the expected compensation for this role. Compensation within the range is determined based on experience, skills, and the scope of responsibilities, with flexibility for candidates who demonstrate exceptional impact.

In addition to base salary, we offer competitive equity. Final compensation may vary based on location within the United States.

Benefits

We proudly offer the following benefits for our full-time employees:

- Medical, Dental, and Vision benefits for you and your family

- Li

Salary insight

The midpoint of this range ($300k) is about 50% above the median disclosed salary for San Francisco roles listed on ForgeApply ($200k across 8,796 jobs).

See full Software Engineer salary data for San Francisco

Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.

Tailor your resume for this Decagon role before you apply.

Tailor my resume for this job

Similar jobs

More like this: Software Engineer Jobs · Software Engineer Jobs in San Francisco · Browse all jobs

Free ATS checker · How to Tailor Your Resume to a Job Description (Step by Step)