ForgeApply · Job listing
Machine Learning Operations Engineer II
Spgi
See all 89 open roles at Spgi →
Tailor your resume for this Spgi job in about a minute.
ForgeApply tailors your resume and cover letter to this exact posting, then hands you a ready-to-submit application for Spgi's site. Free trial, no card required.
About this role
Kensho is S&P Global’s hub for AI innovation and transformation. With expertise in machine learning, natural language processing, and data discovery, we develop and deploy novel solutions to innovate and drive progress at S&P Global and its customers worldwide. Kensho's solutions and research focus on business and financial generative AI applications, agents, data retrieval APIs, data extraction, and much more.
At Kensho, we hire talented people and give them the autonomy and support needed to build amazing technology and products. We collaborate using our teammates' diverse perspectives to solve hard problems. Our communication with one another is open, honest, and efficient. We dedicate time and resources to explore new ideas, but always rooted in engineering best practices. As a result, we can innovate rapidly to produce technology that is scalable, robust, and useful.
The MLOps team is the de facto ML platform team at Kensho. Our team’s mission is critical: empower our ML engineers with state-of-the-art processes, tooling, and infrastructure to iterate quickly, build reliably, and identify potential production issues early. We sit at the intersection of infrastructure and ML, and work closely with all our ML teams (ML Product teams, R&D, …) and our infrastructure teams (Core Infra, SRE, Security). We are a small and high-leverage team: our work practically touches every AI project at Kensho. We balance pragmatic platform development with hands-on exploration at the frontier: building agentic applications ourselves, contributing to open-source tools, and defining what a mature agentic platform looks like before the industry has settled on the answers. You’re equally likely to find us at a top ML conference (NeurIPS, ICLR, ICML) and at major software and infra conferences (Amazon Re:invent, PyCon). To illustrate the point, within the same month, the same engineer went from reimplementing a prompt optimization research paper to shipping prometheus alerts.
As an MLOps Engineer, you are a thoughtful, curious, collaborative, and resourceful person passionate about building and supporting a mature ML platform. You are not afraid to dig deep in both infrastructure and ML topics. You’re excited to work on internal tooling enabling ML engineers to iterate faster and build high-quality production-ready models, agents, and products. You love improving the developer experience (including your own!) and find genuine satisfaction in making engineers more effective, whether by saving engineering hours or amplifying the impact of an engineering organization. You take pride in having a multiplier effect across an engineering team or process, and you enjoy working with multiple teams with different products and workflows.
Excited by what you’ve read so far? If so, we would love to help you excel here. At Kensho, we hire talented people and give them the autonomy and support needed to build amazing technology and products. We support our employees by fostering opportunities for continual learning, pursuing their curiosities and adding to an amazing culture. We collaborate with one another in an open, honest, and efficient way to solve hard problems.
Kensho states that the anticipated base salary range for the position is 130 -175k. In addition, this role is eligible for an annual incentive bonus and equity plans. At Kensho, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case.
What You’ll Do: • Iterate on Kensho’s ML processes to develop tools, services, and frameworks that make every stage of the ML workflow robust, auditable, and usable.
• Work closely with ML engineers to understand their unique processes, identify pain points, and form effective solutions.
• Empower engineers with the stable tooling necessary to rapidly experiment and actualize their research into demonstrable prototypes and mature products
• Provide resources and training for ML teams on best practices, enabling them to efficiently productionize their work to be leveraged by high-value products and services
• Evaluate, select and champion open source and third-party solutions, driving their adoption across teams and integrating into Kensho’s existing platform ecosystem
• Ship scalable, efficient, and automated processes for model fine-tuning and reinforcement learning and for the evaluation of LLMs/Agents
• Improve LLM and Agentic observability to help monitor agentic applications in production, detecting performance, decay and drift issues
• Stay at the frontier by actively tracking emerging tools and frameworks, promote best practices and strengthen the technical expertise of the team with your unique skill set
What You’ll Need: • 2+ years of experience in ML infra, ML Ops, ML Engineering or some similar skillset
• Experience managing distributed systems with Kubernetes. It is important to understand Kubernetes concepts and trade-offs
• Cloud Platform (AWS) understanding. We utilize tools like EKS and managed ML services like Bedrock and SageMaker
• Python proficiency (we are a python shop mostly)
• Familiarity with distributed computing frameworks and workflow orchestration (ie. Ray, Airflow)
• Familiarity with software engineering best practices in an ML context
• Some basic understanding of ML concepts, LLMs and agents
• Ability to debug distributed systems across infrastructure, networking and application layers
• Excellent communication skills to drive adoption of new tools and best practices across multiple teams
• Someone who’s very curious, driven, low-ego and eager to learn across a range of engineering disciplines, while being part of a fantastic team
Technologies & Tools We Use: • Development: Python, Bash, LangGraph, PyTorch
• Infrastructure: Ray, Amazon EKS, Airflow, Jsonnet, Terraform
• Ops: Git, Github, AWS, LangFuse, Sentry, Prometheus, W&B
How To Really Get O
Salary insight
This posting doesn't disclose pay. Across 8,051 New York jobs with disclosed salaries on ForgeApply, the median is $165k.
See full Machine Learning Engineer salary data for New York →
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Tailor your resume for this Spgi role before you apply.
Tailor my resume for this jobSimilar jobs
- Machine Learning Engineer II — Milwaukee Tool · Brookfield, WI
- Machine Learning Engineer II — Clear · New York, NY, United States
- Machine Learning Engineer II — Niagara Bottling · Corp-Main - Diamond Bar, CA
- Machine Learning Engineer II — Spgi · New York, NY | Cambridge, MA
- Machine Learning Engineer II — Abnormalsecurity · Remote
- Machine Learning Ops Engineer — Zone5technologies · United States
- Machine Learning Engineer I — Warner Bros. Discovery · NY New York 30 Hudson Yards | DC Washington 820 1st Street NE
- Senior Machine Learning Operations Engineer — Mercury · Remote
More like this: Machine Learning & AI Jobs · Machine Learning & AI Jobs in New York · Browse all jobs
Free ATS checker · No Salary on the Job Posting? How to Find the Number Before You Interview