ForgeApply · Job listing
LLM Engineer
Clera
See all 129 open roles at Clera →
Tailor your resume for this Clera job in about a minute.
ForgeApply rewrites your resume for this exact posting, then autofills the application on Clera's site with it. You review everything before it's sent. Free trial, no card required.
About this role
ABOUT THE ROLE
You'll be building an AI Supply Chain Manager from the ground up at a Y Combinator–backed AI/supply chain startup, owning LLM systems end-to-end—from early prototypes to production. This is a high-autonomy role where your work directly powers real-world supply chain outcomes: purchase orders, transfer orders, and supplier bookings for CPG brands globally.
WHAT YOU'LL DO
- Build and iterate on LLM-powered agents that decide what to produce, where to make it, and how to move it through factories, warehouses, and channels.
- Design robust prompts, tool definitions, structured outputs, and multi-step agent flows that handle edge cases in messy real-world supply chain data.
- Select and evaluate LLMs based on latency, cost, accuracy, and use-case fit across the supply chain workflow.
- Ship LLM features from concept to production, owning the full pipeline from API integration to tested and evaluated systems.
- Build and maintain evaluation frameworks to iterate quickly and safely on prompts and agent behaviors.
- Set up observability, monitoring, and feedback loops to track production accuracy and continuously improve system performance.
WHAT WE'RE LOOKING FOR
- 3+ years of hands-on experience shipping LLM-powered features or systems in production environments (not just demos or prototypes).
- Proficiency with modern LLM APIs such as OpenAI, Anthropic, DeepSeek, Gemini, or similar, with a strong intuition for model selection tradeoffs.
- Experience building LLM-powered agents or automations that are core to production systems.
- Strong system-design skills for tools/functions, structured outputs, and multi-step agent flow architecture.
- Experience designing and maintaining evaluation pipelines for prompts, workflows, and LLM outputs.
- Hands-on experience with LLM observability, monitoring, and logging tools to track latency, accuracy, and quality in production.
- Practical experience with the HuggingFace ecosystem in real projects.
- Bonus: experience running self-hosted LLMs in production or serious prototypes.
- Bonus: background in supply chain, logistics, or operations domain systems.
COMPENSATION & BENEFITS
Visa sponsorship is available.
LOCATION
On-site in New York, NY, United States.
Salary insight
This posting doesn't disclose pay. Across 9,898 New York jobs with disclosed salaries on ForgeApply, the median is $162k.
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Tailor your resume for this Clera role before you apply.
Tailor my resume for this jobSimilar jobs
- LLM Applications Engineer — Uncountable · New York, San Francisco, Munich or London
- LLM Ops Engineer — Litera · Denver, CO
- ML Engineer: Speech & LLMs — Knowtex · San Francisco
- AI Engineer, Multimodal LLMs — Eloquentai · San Francisco
- Machine Learning Engineer - LLMs — Hadrian Automation · Los Angeles, CA
- Software Engineer, LLM & Automation — Basis Research · New York Office
- Engineering Manager, Forward Deployed Engineering (LLM) — Baseten · San Francisco
- LLM Engineer (GenAI, NYC) — Gptzero · NYC Hybrid
Free ATS checker · No Salary on the Job Posting? How to Find the Number Before You Interview