Site Reliability Engineer III
There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.
As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking team of Fraud Prevention, you will solve complex and broad business problems with simple and straightforward solutions.
Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions.
You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.
Y ou are an integral part of a team that works to develop high-quality architecture solutions for various software applications and platform products.
You drive significant business impact and help shape the target state architecture through your capabilities in multiple architecture domains.
You will ensure the platform is reliable, secure, performant, and resilient in production across Kubernetes-based environments and AWS.
You will apply SRE principles to drive measurable improvements in availability and latency, reduce operational toil through automation, and strengthen deployment safety and recovery capabilities in close partnership with engineering and platform teams.
Job responsibilities
* Production ownership & reliability outcomes: Own day-to-day operational health for the platform, focusing on availability, latency, throughput, and error rates; proactively identify reliability risks and drive remediation.
* SLO/SLI and alerting strategy: Define and evolve SLIs/SLOs and error budgets; build actionable alerting aligned to customer impact and reduce noise through tuning and standardization.
* Observability & troubleshooting: Improve end-to-end observability (metrics, logs, traces), dashboards, and runbooks across Kubernetes and AWS; perform deep technical triage of distributed system issues.
* Incident response & problem management: Participate in on-call and lead/assist incident triage, mitigation, and recovery; conduct RCAs and drive corrective/preventive actions to closure.
* Kubernetes operations: Support containerized workloads, autoscaling, rollout/rollback procedures, resource tuning, and resilience patterns.
* AWS operations (container + serverless): Operate components on AWS (EKS/ECS/Lambda) and associated data services (Dynamo DB, S3); manage operational concerns such as scaling, retries, and safe failure modes.
* Release engineering & delivery reliability: Improve the safety and repeatability of deployments using Spinnaker and Harness.
* Infrastructure as Code & environment consistency: Build and maintain Terraform modules and automation for reliable, repeatable environments.
*...
- Rate: Not Specified
- Location: Jersey City, US-NJ
- Type: Permanent
- Industry: Finance
- Recruiter: JPMorgan Chase Bank, N.A.
- Contact: Not Specified
- Email: to view click here
- Reference: 210786031
- Posted: 2026-09-05 09:40:32 -
- View all Jobs from JPMorgan Chase Bank, N.A.
More Jobs from JPMorgan Chase Bank, N.A.
- Retail Sales Associate - Part Time
- Commercial Driver - Full Time
- Part Sales Manager - Full Time
- Retail Sales Associate - Part Time
- Representante De Vendas
- Retail Sales Associate - Part Time
- Part Sales Manager - Full Time
- Commercial Driver - Full Time
- Commercial Driver - Part Time
- Retail Sales Associate - Part Time
- Commercial Sales Manager
- Representante De Vendas
- Commercial Specialist
- Retail Sales Associate - Part Time
- Commercial Driver - Full Time
- Part Sales Manager - Full Time
- Hub Specialist
- Retail Sales Associate - Part Time
- Part Sales Manager - Part Time
- Commercial Driver - Part Time