Site Reliability Engineer III - AI & Corporate Risk Tech
JPMorgan Chase & Co. Glasgow, Scotland, United Kingdom
Financial Services · 10,001+ employees
About the role
The Site Reliability Engineer will configure, monitor, and optimize mission-critical systems using cloud infrastructure and automated pipelines. They will also collaborate with cross-functional teams to implement reliability best practices and resolve complex operational issues.
What they look for
Requirements
Candidates must have formal training or certification in site reliability engineering and proficiency in at least one programming language like Python, Java, or .NET. Experience with observability practices, CI/CD tooling, and container technologies is required to support platform reliability.
Full description
There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skills to drive innovation and modernize some of the world's most complex and mission-critical systems. At JPMorganChase, you'll be part of a team that values curiosity, collaboration, and continuous improvement — where your contributions directly shape the reliability and resilience of platforms that matter.
As a Site Reliability Engineer III at JPMorganChase within Corporate Risk Technology, you will solve complex and broad business problems with simple, straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure — independently decomposing and iteratively improving on existing solutions. You are a meaningful contributor to your team, sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.
Job responsibilities
- Guide and support team members in building appropriate-level designs, gaining peer consensus, and driving adoption of site reliability engineering best practices across the team
- Collaborate with software engineers and cross-functional teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
- Implement infrastructure, configuration, and network as code for the applications and platforms within your scope
- Partner with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
- Identify and address roadblocks, propose improvements to solve business problems, and explore new technologies where appropriate
- Apply familiarity with availability, reliability, and scalability principles to iteratively improve outcomes in collaboration with partners
- Leverages enterprise-authorized AI coding assist tools within the work environment to improve code quality, delivery speed, and productivity (e.g., code generation/refactoring, unit test creation, documentation), while validating outputs through peer review, automated testing, and secure coding standards
- Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation
Required qualifications, capabilities, and skills
- Formal training or certification on site reliability engineering concepts and proficient applied experience
- Proficiency in site reliability culture and principles, with the ability to implement site reliability practices within an application or platform
- Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
- Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection
- Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., cloud, AI, mobile platforms)
- Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, testing, troubleshooting, or documentation) with demonstrated ability to critically evaluate and validate AI-generated outputs
- Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations
Preferred qualifications, capabilities, and skills
- Experience with continuous integration and continuous delivery tooling
- Familiarity with container technologies and container orchestration platforms
- Experience troubleshooting common networking technologies and issues
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.
Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.
Similar roles
-
B2B Systems Site Reliability Engineer
Jamf Poland
-
Staff Site Reliability Developer, Google Unified Security and Threat Operations
Google Waterloo, Ontario, Canada · CA$216K–CA$221K/yr
-
Senior Site Reliability Engineer
Hive Berlin, Germany · €90K–€120K/yr
-
Staff Site Reliability Engineer, Semantic Understanding
Google San Jose, California, United States · $207K–$300K/yr
-
Senior Site Reliability Engineer
QuTwo Helsinki, Uusimaa, Finland
-
Senior Site Reliability Engineer
Okta Bengaluru, Karnataka, India