US Jobs US Jobs     UK Jobs UK Jobs     EU Jobs EU Jobs


Site Reliability Engineer II

Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions.

As a Site Reliability Engineer II at JPMorgan Chase within the within the Corporate and Investment Bank, Payments Technology Team , you will use technology to solve business problems and leverage software engineering best practices as we strive towards excellence.

This role often works independently to execute small to medium projects, but you'll also have the opportunity to collaborate with cross functional teams to continually improve your level of knowledge about JPMorgan Chase's business and relevant technologies.

Job Responsibilities


* Deliver end-to-end application and/or infrastructure service delivery to enable reliable business operations across the firm.


* Partner with cross-functional teams to define and maintain SLOs/SLIs and error budgets for key production services, proactively resolving issues before customer impact.


* Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.


* Own day-to-day operational processes including incident management, problem management (RCA), and change/event management (monitoring and alerting).


* Monitor production environments for anomalies using standard observability tools; build and maintain alerting mechanisms to detect incidents early and minimize downtime.


* Design and implement Dynatrace instrumentation to monitor performance, infrastructure health, and user experience; analyze trends, escalate/communicate as needed, and provide solutions to business and technology stakeholders.


* Applies enterprise-authorized AI capabilities within the work environment to identify recurring toil and reliability risks from operational signals, prioritizing reuse-first improvements and measurable SLO outcomes.

Required qualifications, capabilities, and skills


* Formal training or certification on software engineering concepts and 2+ years applied experience ( NAMR/APAC - India/ LATAM/ Hong Kong)


* Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows (e.g., troubleshooting support and runbook drafting) with strong validation habits and awareness of data sensitivity .


* Ability to assess AI-assisted operational recommendations for correctness and risk, and apply appropriate controls to maintain resiliency, security, and auditability.


* Large-scale application/infrastructure experience across on-prem and public cloud environments, including strong understanding of core networking (DNS, TCP/IP, VPN, load balancing).


* Hands-on observability expertise: production monitoring, distributed tracing, log analysis, and instrumentation using tools like Dynatrace, Geneos, and Splunk plus applicati...




Share Job