Find Jobs
Find Jobs Near You – Available Work in Your Location
Skip to job details
JC
JPMorgan Chase Bank, N.A.
Lead Site Reliability Engineer
Career Insights for Site Reliability Engineer
See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.
Scorecard
Based on Delaware data
Review key factors to help you decide if this role fits your goals. How is this calculated?
What they do
A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.
$120,151 / year median in Delaware
Job Description
If you are excited about shaping the future of technology and driving significant business impact in financial services, we are looking for people just like you. Join our team and help us develop game-changing, high-quality solutions. As a Lead Site Reliability Engineer at JPMorganChase within the Corporate sector, Enterprise Technology team, you are an integral part of a team that develops high-quality architecture solutions for critical software applications and platforms. You will lead resiliency design reviews, break down complex problems, and mentor engineers, driving significant business impact and shaping the target state architecture through your expertise in multiple architecture domains. Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. Job responsibilities
- Demonstrate and champion site reliability culture and practices, exerting technical influence across your team
- Lead initiatives to improve reliability and stability of applications and platforms using data-driven analytics
- Collaborate with team members to define service level indicators and work with stakeholders to establish service level objectives and error budgets
- Provide technical leadership and guidance for medium to large-sized products
- Proactively identify and resolve technology-related bottlenecks in your areas of expertise
- Act as the main point of contact during major incidents, quickly identifying and solving issues to avoid financial losses
- Document and share knowledge within the organization through internal forums and communities of practice
- Use enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements
- Lead reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices (e.g., CI/CD quality checks, test/validation automation, and operational readiness), ensuring traceability/auditability, resiliency, and security controls
- Offer mentorship and advice to other engineers, fostering a culture of continuous improvement
- Drive collaboration with stakeholder partners to establish reasonable service level objectives and error budgets Required qualifications, capabilities, and skills
- Formal training or certification on software engineering concepts and 5+ years applied experience
- At least 5 years as an SRE and at least 10 years in a highly regulated industry such as Banking
- Deep proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and site reliability best practices, with the ability to implement these practices within an application or platform
- Demonstrated experience designin.