ForgeApply · Job listing
Manager, Site Reliability Engineering
M&T Bank
See all 208 open roles at M&T Bank →
Tailor your resume for this M&T Bank job in about a minute.
ForgeApply tailors your resume and cover letter to this exact posting, then hands you a ready-to-submit application for M&T Bank's site. Free trial, no card required.
About this role
Manager, Site Reliability Engineering
62 M
Overview
Responsible for leading the Site Reliability Engineering Center of Excellence and the Forward Deployed SRE program supporting critical banking platforms, applications, and technology services. Manages an organization of employees and contingent resources through direct reports, program managers, SRE leaders, technical leads, and matrixed delivery relationships.
Establishes the enterprise SRE strategy, operating model, engineering standards, governance, talent model, and adoption roadmap. Accountable for improving service reliability, availability, scalability, performance, resiliency, deployment safety, and operational maturity across supported technology domains.
Leads the deployment of SRE capabilities into application and platform teams through a Forward Deployed SRE model. Partners with senior leaders across application development, infrastructure, cloud engineering, architecture, cybersecurity, technology operations, risk, and business-aligned technology organizations to prioritize engagements and deliver measurable reliability improvements.
Balances strategic leadership, people management, program execution, technical governance, and operational accountability. Ensures SRE practices are implemented consistently and that reliability investments produce measurable improvements in customer experience, operational risk, engineering productivity, release quality, and service performance.
Primary Responsibilities
SRE Strategy and Center of Excellence Leadership
• Establish and execute the vision, strategy, operating model, service offerings, and multiyear roadmap for the Site Reliability Engineering Center of Excellence.
• Define enterprise SRE standards, engineering practices, governance processes, engagement models, and maturity expectations.
• Lead the adoption of reliability engineering practices across application development, infrastructure, platform engineering, cloud engineering, and technology operations.
• Translate enterprise technology, business, resiliency, and risk priorities into an actionable SRE portfolio and delivery roadmap.
• Establish a scalable SRE service model that includes consulting, enablement, embedded engineering, Forward Deployed SRE engagements, reusable capabilities, and sustained ownership by application and platform teams.
• Define intake, prioritization, engagement, transition, and exit criteria for SRE services.
• Develop and maintain an SRE maturity model used to assess service teams, identify reliability gaps, and guide improvement plans.
• Establish communities of practice, technical forums, training programs, playbooks, reference architectures, and reusable engineering patterns that expand SRE capabilities across the organization.
• Ensure the SRE Center of Excellence remains aligned with enterprise engineering standards, cloud strategy, operational risk requirements, and evolving industry practices.
• Represent the SRE organization in senior leadership forums, architecture reviews, operational governance meetings, and enterprise transformation initiatives.
Forward Deployed SRE Program Leadership
• Lead the Forward Deployed SRE program, placing SRE professionals into high-priority application and platform teams to address complex reliability challenges and improve operational maturity.
• Manage program managers, SRE leaders, and technical leads responsible for coordinating engagements across multiple technology domains.
• Establish a transparent intake and prioritization process based on customer impact, service criticality, operational risk, incident history, reliability maturity, strategic importance, and anticipated business value.
• Partner with application and platform leaders to define engagement objectives, scope, deliverables, staffing, success measures, dependencies, and duration.
• Ensure Forward Deployed SRE teams deliver sustainable engineering improvements rather than becoming long-term substitutes for application support or production operations.
• Establish shared accountability for participation, knowledge transfer, remediation activities, and long-term ownership of implemented reliability capabilities.
• Develop transition and exit plans that enable application and platform teams to sustain SRE practices after engagements conclude.
• Evaluate engagement effectiveness using measurable outcomes such as availability, SLO attainment, incident frequency, restoration time, alert quality, automation adoption, toil reduction, change-failure rate, deployment reliability, and engineering maturity.
• Convert common findings and lessons learned into reusable standards, automation, tools, training, and engineering patterns.
• Continuously optimize the Forward Deployed SRE operating model based on demand, capacity, outcomes, stakeholder feedback, and changes in technology strategy.
Organizational and People Leadership
• Lead an organization of employees and contingent resources through direct and indirect management relationships.
• Manage and develop program managers, SRE managers, technical leaders, and senior engineering professionals.
• Establish clear roles, responsibilities, decision rights, performance expectations, and accountability across the SRE organization.
• Build a high-performing organization with expertise in reliability engineering, observability, cloud platforms, release engineering, deployment orchestration, automation, Infrastructure as Code, incident management, resiliency engineering, testing, and program delivery.
• Recruit, retain, coach, and develop diverse engineering and program management talent.
• Conduct workforce, capacity, and succession planning to ensure the organization has the leadership and technical capabilities required to meet current and future demand.
• Define career paths and skill-development plans for SRE professionals in partnership wit
Salary insight
The midpoint of this range ($186k) is about 59% above the median disclosed salary for Buffalo roles listed on ForgeApply ($118k across 203 jobs).
See full DevOps / SRE salary data for Buffalo →
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Tailor your resume for this M&T Bank role before you apply.
Tailor my resume for this jobSimilar jobs
- Manager, Site Reliability Engineering — Palo Alto Networks · Office - USA - CA - Headquarters
- Manager, Site Reliability Engineering — Okta · Bellevue, Washington; Chicago, Illinois; San Francisco, California; Washington, DC
- Manager, Site Reliability Engineering — Litera · Denver, CO
- Manager, Site Reliability Engineer — Forgeglobal · New York, New York, United States
- Manager, Site Reliability Engineer — Forgeglobal · San Francisco, California, United States
- Senior Engineering Manager, Site Reliability — Upstart · Remote
- Engineering Manager, Site Reliability — Radar · New York
- Engineer - Site Reliability Engineering — LSEG · USA-St. Louis-795 Office Pkwy
More like this: DevOps & SRE Jobs · DevOps & SRE Jobs in Buffalo · Browse all jobs
Free ATS checker · How to Tailor Your Resume to a Job Description (Step by Step)