ForgeApply
Try it free

ForgeApply · Job listing

AI Engineer - Data Platform

Clera

USonsite

Apply in about a minute — without sacrificing quality.

ForgeApply autofills this application and tailors your resume to this exact posting. You review everything before it's sent. Free trial, no card required.

About this role

ABOUT THE ROLE

Join a well-funded, Series A AI startup building the next generation of autonomous site reliability engineering for the enterprise. Backed by top-tier investors and trusted by some of the largest companies in the world, this team is tackling one of the hardest problems in AI: autonomously detecting, diagnosing, and remediating complex production incidents in real time.

As an AI Engineer on the Data Platform team, you'll design, build, and maintain the backend systems that power an AI-driven observability platform. This hands-on role blends distributed systems engineering, low-level system design, performance optimization, observability, and AI integration — across both cloud and on-premises deployments.

WHAT YOU'LL DO

- Architecture & Implementation: Contribute to the design and implementation of scalable, resilient infrastructure systems powering AI-driven root cause analysis and observability workflows, including on-premises deployment environments.

- Low-Level System Design: Work on the foundational building blocks of the infrastructure, ensuring efficient resource utilization and high performance at scale.

- Performance Optimization: Profile and tune backend systems to improve throughput, reduce latency, and eliminate bottlenecks across the stack.

- Observability Systems: Build and maintain the internal observability stack — logs, metrics, and traces — used by AI agents to understand and act on production issues.

- Hybrid Infrastructure: Support cloud and on-premises architecture to serve both SaaS and enterprise customer deployment models.

- Cross-functional Collaboration: Work closely with engineers across the company to deliver resilient infrastructure that enables AI agents to diagnose and remediate production incidents in real time.

WHAT WE'RE LOOKING FOR

- Experience: 2–5 years of hands-on backend or infrastructure engineering experience.

- Distributed Systems: Strong understanding of distributed systems design principles and trade-offs.

- Performance Engineering: Proven experience profiling and optimizing high-throughput, low-latency systems.

- Observability: Familiarity with observability tooling and concepts (logs, metrics, traces); experience with platforms such as Datadog, Grafana, Splunk, or similar is a plus.

- Cloud & On-Prem: Experience with hybrid or multi-environment infrastructure (cloud + on-premises).

- AI/ML Integration: Interest in or experience building systems that support AI/ML workloads at scale.

- Background: Prior experience at observability, incident management, or data infrastructure companies is highly valued.

Note: Visa sponsorship is not available for this role.

LOCATION

This is a fully on-site role based in New York, NY. Remote work is not available for this position.

Salary insight

This posting doesn't disclose pay. Across 4,586 New York jobs with disclosed salaries on ForgeApply, the median is $176k.

See full Machine Learning Engineer salary data for New York

Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.

Ready to apply to Clera?

Apply in about a minute

Similar jobs

More like this: Machine Learning & AI Jobs · Machine Learning & AI Jobs in New York · More jobs at Clera · Browse all jobs