As a Site Reliability Engineer (SRE), you'll help build a meaningful engineering discipline, combining software and systems to develop creative engineering solutions to operations problems. Much of our support and software development focuses on optimizing existing systems, building infrastructure, and reducing work through automation. You'll join a team of curious problem solvers with a diverse set of perspectives who are thinking big and taking risks. In this environment, you'll take the lead on relevant projects, supported by an organization that provides the support and mentorship you need to learn and grow. As an SRE, you'll be focused on running better production applications and systems.
* Design, code, test, and deliver software to automate manual operational work
* Troubleshoot priority incidents, facilitate blameless post-mortems and ensure permanent closure of incidents
* Engage with development team throughout the life cycle to help develop software for reliability and scale, ensuring minimal refactoring or changes
* Identify application patterns and analytics in support of better service level objectives
* Design self-healing and resiliency patterns
* Design automated software and product upgrades, change management, and release management solutions
* Coach or manage teams as applicable
* Participate in the 24x7 support coverage as needed
* Bachelor's degree or equivalent experience in an software engineering discipline
* Expertise in at least one technology stack designing, coding, testing, and delivering software
* Proficiency in one or more technology domains, may be a cross-domain expert able to solve complex and mission critical problems within a business or across the firm
* Working knowledge of infrastructure components (e.g. routers, load balancers, cloud products, container systems, compute, storage, and networks)
* Excellent debugging and trouble shooting skills
* Demonstrated experience as a Site Reliability Engineer, DevOps Engineer
* Proven experience as a software engineer, including proficiency in at least one systems programming language (Python/Go preferred)
* Understanding of key SRE concepts, such as Service Level Objectives
* Experience with Linux
* Experience Kubernetes and/or Cloud Foundry , AWS, including knowledge of IAM & VPC Networking , and Grafana
* Plus/Preferred Qualifications: Terraform ; AWS - Sysops/Solution Architect Certification ; Prometheus / PromQL ; Datadog ; Kubernetes CKA/CKAD Certifications
JPMorgan Chase & Co., one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. In accordance with applicable law, we make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as any mental health or physical disability needs.
Equal Opportunity Employer/Disability/Veterans