ForgeApply · Job listing
Senior Machine Learning Engineer, Services/MLOps
Adobe
See all 296 open roles at Adobe →
Tailor your resume for this Adobe job in about a minute.
ForgeApply tailors your resume and cover letter to this exact posting, then hands you a ready-to-submit application for Adobe's site. Free trial, no card required.
About this role
The Opportunity
Firefly Foundry is Adobe’s enterprise managed-service offering for custom multimedia generative AI — deep-tuned image, video, and 3D models built on each customer’s IP, paired with creative production workflows and a media-intelligence layer, and deployed across new and existing Adobe surfaces. The business has gained significant traction in Media & Entertainment, marketing, and consumer retail, and is expanding rapidly into adjacent verticals.
We are hiring a Senior Machine Learning Engineer to build the pipelines and services that turn Firefly Foundry’s models into reliable, enterprise-grade products. You will compose heterogeneous model pipelines including finetuned LLMs, image and video generation models , 3D mesh reconstruction, up - sampl ers , NSFW and safety checkers, and IP guardrai l models — deploy them as services, scale those services to enterprise traffic, and design them to meet target latency and throughput budgets, all while ensuring served quality matches the training and reference environment . Across this work you will integrate and operate multiple, distinct generative model architectures, in a mix that evolves quickly.
This is a high-ownership role in a fast-moving environment, with direct , measurable impact on the latency, cost, and quality of everything Firefly Foundry ships . Depending on your focus area, you may own externalizable data pipelines for self-serve fine-tuning, optimized VLM deployments for media intelligence and querying, or the platform that lets the team deploy new pipelines rapidly with full observability.
What you will do
• Own the full serving lifecycle for heterogeneous model pipelines — packaging, versioned rollout, canary/rollback, and autoscaling — from research checkpoint to enterprise endpoint. • Deploy these pipelines as services and scale them to enterprise traffic, meeting target latency and throughput budgets. • Ensure served quality matches the training and reference environment — closing train/serve gaps across precision, preprocessing, and model versions. • Engineer for enterprise from the ground up: tenancy boundaries, data isolation, and the controls that let us honor customer IP contracts under audit. • Build the platform underneath it all — rapid pipeline deployment, observability, monitoring, and alerting. • Define and enforce quality gates in the deployment pipeline – automated eval, regression detection, and drift monitoring that block bad model versions from reaching production. • Own GPU capacity and cost – utilization , batching efficiency, and right-sizing acceleration fleets against latency SLAs. • Run production ML operationally – on-call, incident response, an dpostmortems for availability and latency regressions
Depending on your focus area, you may also:
• Build externalizable data pipelines that power self-serve fine-tuning flows for enterprise customers. • Stand up optimized VLM deployments for media intelligence and content querying.
Who you will partner with
• Applied Science — to take research models into reliable, high-throughput serving and to keep served quality faithful to the training environment. • ML Engineering leadership and AI Platform — on shared infrastructure, accelerator capacity, and serving primitives at platform scale. • Firefly Foundry Studio — to translate creative production workflows into performant, dependable ML services.
What you bring
• 5+ years in machine learning engineering , with significant ownership of production ML or inference services at scale. • Strong Python and deep-learning engineering skills (PyTorch), with hands-on experience deploying and scaling model-backed services. • Experience composing multi-model pipelines and serving them behind APIs — orchestration, batching, autoscaling, and version management. • A track record building the observability, monitoring , and alerting that production services rely on to hit latency and throughput targets. • Comfort working across multiple, distinct generative model architectures (LLMs and VLMs, diffusion and transformer models, 3D/mesh) — enough to integrate, optimize, and reason about output quality, in partnership with Applied Science. • Experience with multi-tenant systems and data isolation in an enterprise or regulated context. • Fluency with containers and orchestration (Docker, Kubernetes), CI/CD for ML, and a major cloud (AWS or Azure). • GPU inference optimization for latency and cost — quantization, batching, and serving runtimes; custom CUDA a plus. • Strong, data-driven problem-solving and excellent communication in cross-functional teams.
Education
• Master’s or PhD in Computer Science, Computer Engineering, or a related field — or equivalent practical experience building and operating production ML systems.
#FireflyGenAI About Adobe Adobe empowers everyone to create through innovative platforms and tools that unleash creativity, productivity and personalized customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud, Adobe Experience Platform, Adobe Experience Manager, and GenStudio enable people and businesses to turn ideas into impact, powered by AI and driven by human ingenuity.
Our 30,000+ employees worldwide are creating the future and raising the bar as we drive the next decade of growth. We’re on a mission to hire the very best and believe in creating a company culture where all employees are empowered to make an impact. At Adobe, we believe that great ideas can come from anywhere in the organization. The next big idea could be yours.
Let’s Adobe together At Adobe, we believe in creating a company culture where all employees are empowered to make an impact. Learn more about Adobe life, including our values and culture , focus on people, purpose and community , Adobe for All , comprehensive benefit
Salary insight
The midpoint of this range ($224k) is about 12% above the median disclosed salary for San Francisco roles listed on ForgeApply ($200k across 8,037 jobs).
See full Machine Learning Engineer salary data for San Francisco →
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Tailor your resume for this Adobe role before you apply.
Tailor my resume for this jobSimilar jobs
- Senior Machine Learning Engineer — Vanguard · Malvern, PA
- Senior Machine Learning Operations Engineer — Mercury · Remote
- Senior Machine Learning Engineer — Axon · Seattle, Washington, United States
- Senior Machine Learning Engineer — Adobe · San Francisco | Seattle | San Jose
- Senior Machine Learning Engineer — Reddit · Remote
- Senior Machine Learning Engineer — Bjakcareer · Remote
- Senior Machine Learning Engineer — Claritypay · New York City
- Senior Machine Learning Engineer — Amperity · Seattle, WA
More like this: Machine Learning & AI Jobs · Machine Learning & AI Jobs in San Francisco · Browse all jobs
Free ATS checker · How to Tailor Your Resume to a Job Description (Step by Step)