Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
The SRE III is responsible for the end-to-end operational health, reliability, and performance of assigned business services and payment flows. They lead incident and change management processes while driving automation to eliminate manual toil and ensure regulatory compliance.
What they look for
Requirements
Candidates must have at least 5 years of experience operating production systems within the financial services or payments sector. A strong background in reliability engineering disciplines, including SLO management and observability tooling, is required.
Full description
As a Senior Production Engineer, you are the named, single point of accountability for the operational health of an assigned business service or payment flow. You defend its reliability budget, own its controls evidence, and represent its client impact to executive stakeholders — converting the question "is the ticket closed?" into "did the client experience degrade, and what did it cost us?"
Job Responsibilities
• Owns end-to-end operational health of a named business-service or payment-flow portfolio, accountable for availability, MTTD, MTTM, MTTR, and client-impact trend against published SLOs and error budgets
• Defends the reliability (error) budget for owned services and makes change-governance decisions based on budget consumption, not ticket urgency alone
• Leads incident, problem, and change management for full-stack payment systems, with authority to escalate cross-functionally and represent client impact directly to executive stakeholders
• Owns controls and regulatory evidence for assigned services — patch compliance, certificate expiry, vulnerability aging, EOL exposure — and drives automation of manual evidence collection
• Converts manual, repetitive work into automation against a named toil backlog, measuring hours removed and capacity returned to change-the-bank work
• Partners with application development and central SRE/observability teams under a shared reliability services agreement, without absorbing their engineering backlog
• Builds depth of coverage on owned services, eliminating single-person dependencies and naming a backup owner
Required Qualifications, Capabilities, and Skills
• 5+ years operating payment or financial-services production systems with named accountability for service availability and incident outcomes
• Demonstrated ownership of SLOs, error budgets, or an equivalent reliability engineering discipline in a large-scale, regulated technology environment
• Working knowledge of payment flows — authorization, clearing, settlement, billing, payment/rebate — and their client and regulatory impact
• Experience partnering with controls, risk, and audit functions on production evidence and regulatory findings
• Experience with observability/monitoring tooling and automation or scripting for toil elimination
Preferred Qualifications, Capabilities, and Skills
• Experience standing up or operating under a formal SRE/reliability operating model
• Experience with AI-assisted operations (RCA, ECC, runbook automation) under outcome-gated governance
• Cross-business-unit or multi-region service ownership experience
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.
J.P. Morgan’s Commercial & Investment Bank is a global leader across banking, markets, securities services and payments. Corporations, governments and institutions throughout the world entrust us with their business in more than 100 countries. The Commercial & Investment Bank provides strategic advice, raises capital, manages risk and extends liquidity in markets around the world.
Similar roles
-
Sr. Staff Site Reliability Engineer-Federal, Security Clearance
Zscaler Arlington County, Virginia, United States · $164K–$205K/yr
-
Sr. SRE Lead
Chubb Bogota, Capital District, RAP (Especial) Central, Colombia
-
Site Reliability Engineer II
KFC Europe Irvine, California, United States · $105K–$132K/yr
-
Sr Manager, Site Reliability Engineer (FedRAMP / AWS GovCloud)
BeyondTrust Broken Hill, New South Wales, Australia
-
Site Reliability Engineer
BeyondTrust United States
-
Principal Application Support Engineer(SRE)
DTCC Tampa, Florida, United States