ForgeApply
Try it free

ForgeApply · Job listing

Production Manager, Applied AI

Lilt

Boston, MA, US$110k – $145konsite

See all 6 open roles at Lilt

Tailor your resume for this Lilt job in about a minute.

ForgeApply rewrites your resume for this exact posting, then autofills the application on Lilt's site with it. You review everything before it's sent. Free trial, no card required.

About this role

ABOUT LILT

AI is changing how the world communicates — and LILT is leading that transformation.

We're on a mission to make the world's information accessible to everyone, regardless of the language they speak. We use cutting-edge AI, machine translation, and human-in-the-loop expertise to translate content faster, more accurately, and more cost-effectively without compromising on brand, voice, or quality.

At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues—Work together, win together; Find a way or make one; Dance in the customer's shoes; Quicker than they expect; Quality is Job 1—guide everything we do. We are trusted by Intel Corporation https://www.linkedin.com/company/intel-corporation/, Canva https://www.linkedin.com/company/canva/, the United States Department of Defense https://www.linkedin.com/company/deptofdefense/, the United States Air Force https://www.linkedin.com/company/united-states-air-force/, ASICS https://www.linkedin.com/company/asics/, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we’re building a category-defining company in a $50B+ global translation market being redefined by AI.

KEY RESPONSIBILITIES

- PM Performance - Outcome: Every active program has accountable Project Manager(s), and every PM carries a workload within the agreed span. Programs hit on-time delivery and first-pass acceptance targets without escalation; each PM is reviewed monthly against a scorecard of throughput, quality, and cost-per-task.

- PM Hiring, Onboarding, and Development - Outcome: PM pool capacity keeps pace with signed demand: new PMs are sourced, onboarded, and running their first program within the agreed ramp window, all programs start on time, and PM attrition is below threshold.

- Quality Interventions Across Programs - Outcome: Quality dips are caught mid-program through QA loops and corrected via retraining of annotator pools or guideline updates. Repeated misses by a PM or annotator pool lead to documented remediation or replacement. Issues are proactively discovered.

- Process Standardization & Software Improvements - Outcome: Programs launch from shared playbooks, guideline templates, and dashboard standards rather than being rebuilt per engagement. Every post-mortem produces documented improvements that lead directly into our custom software stack, and time from program handoff to first delivery declines quarter over quarter.

- Escalation and Cross-Functional Interface - Outcome: Risks surface to Technical Program Managers early enough to be managed. Escalations between the PM pool, Quality, Talent, and Delivery are resolved within agreed timelines.

QUALIFICATIONS

- People management in AI data operations: 5+ years in AI/ML data operations or production, including 2+ years directly managing project managers or team leads in a distributed, multi-time-zone contractor environment.

- LLM knowledge: Strong understanding of LLM training processes (pre-training, SFT, RLHF) and evaluation methodologies (human-in-the-loop, red teaming), and of what drives quality and throughput in annotation workflows.

- KPI-driven management: Has run teams against throughput, quality (accuracy, IAA, gold-set), and cost-per-task targets; advanced proficiency with spreadsheets and dashboards, and able to use SQL to extract and analyze performance data.

- Delivery track record: Has sustained on-time delivery and acceptance targets across multiple concurrent data collection or evaluation programs for enterprise or research lab customers.

- Contractor workforce operations: Has hired, ramped, performance-managed, and offboarded hourly and freelance staff across regions and languages.

- Methodology: Proven track record using Agile, Scrum, or Kanban to manage complex workflows across a portfolio of programs.

- Communication: Writes clear, unambiguous guidelines and feedback for multilingual audiences and communicates status, risk, and tradeoffs crisply to leadership.

PREFERRED SKILLS

- Fluency in multiple human languages.

- Experience with multilingual data deliveries (pre-training, SFT, RLHF, machine translation, multimodal, etc.), especially in rare-resource languages

- Experience with data annotation platforms (e.g., Label Studio, SuperAnnotate) and project management tooling (e.g., Jira).

- Background in ML engineering, computer science, or data science.

OUR STORY

Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.

LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting-edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human-in-the-loop systems, and now agentic AI.

With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.

OUR TECH

What sets our platform apart:

- Brand-aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent

- Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing

- 100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation

- Human-in-the-loop reviews via our global network of profe

Salary insight

The midpoint of this range ($128k) is about 18% below the median disclosed salary for Boston roles listed on ForgeApply ($155k across 2,196 jobs).

Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.

Tailor your resume for this Lilt role before you apply.

Tailor my resume for this job

Similar jobs

Free ATS checker · How to Tailor Your Resume to a Job Description (Step by Step)