Join our dynamic team to innovate and refine technology operations, impacting the core of our business services.
As a Lead Infrastructure Engineer at JPMorgan Chase as a part of our Mainframe and Mid-Range Compute Site Reliability and Engineering (SRE) team, we look first and foremost for people who are passionate to solving business problems through innovation and modern engineering practices. You'll be required to apply your depth of knowledge and expertise to all aspects of Infrastructure Support and Software Development Lifecycle, as well as partner continuously with your many stakeholders daily to stay focused on common goals. We embrace a culture of experimentation and constantly strive for improvement and learning. You’ll work in a collaborative, trusting, thought-provoking environment - one that encourages diversity of thought and creative solutions that are in the best interests of our customers, globally.
Job responsibilities
- Lead teams of technologists that provide end-to-end application or infrastructure service delivery for the successful business operations of the firm
- Execute policies and procedures that ensure operational stability and availability
- Monitor production environments for anomalies, address issues, and drive evolution of utilization of standard observability tools
- Escalate and communicate issues and solutions to the business and technology stakeholders, actively participating from incident resolution to service restoration
- Lead incident, problem, and change management in support of full stack technology systems, applications, or infrastructure
- Ability to host and participate in bridge calls and communicate effectively to large group of individuals at all levels.
- Responsible for administering, troubleshooting Mainframe related components.
- Ability to work in large, collaborative teams to achieve organizational goals.
- Uses enterprise-authorized AI capabilities within the work environment to accelerate infrastructure analysis and design documentation, validating outputs and handling operational data according to sensitivity and security requirements.
- Applies reuse-first, AI-assisted practices within delivery and automation routines to identify recurring issues and validate remediation options, ensuring changes are traceable/auditable and aligned to resiliency and security expectations.
Required qualifications, capabilities, and skills
- Formal training or certification on software engineering concepts and 5+ years applied experience
- 10+ experience in operating and managing the operations of IBM z-Series environments
- Demonstrated leadership of Operational and SRE Teams in a 24X7 support environment including all aspects of people management
- Understanding of infrastructure architecture including servers, storage, network, database, and application components.
- Extensive knowledge “Replication” technologies’ such as IBM CSM and GDPS
- Expertise in administering z-Series and Hardware Management Console (HMC) including Firmware/Microcode upgrades.
- Demonstrated understanding of security standards including a working knowledge of SSH protocol. & working knowledge of transaction-based systems, IMS, CICS, DB2 and WebSphere
- Experience in managing ServiceNow including workflow/ticket/resolution management across a global environment 24X7. & demonstrated expertise in Incident/Problem/Change management process and procedures.
- Able to troubleshoot priority incidents, facilitate blameless post-mortems and ensure permanent closure of incidents.
- Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support infrastructure engineering workflows with strong validation habits and awareness of data sensitivity.
- Ability to review and validate AI-assisted recommendations before implementation, escalating when uncertain and ensuring outcomes align to resiliency, security, and auditability expectations.
Preferred qualifications, capabilities, and skills
- Working knowledge in one or more general purpose programming languages and/or automation scripting
- Practical experience with python development
- Knowledge in Site Reliability Engineering - Design, code, test and deliver software to automate manual operational work.
- Knowledge of multiple Batch Scheduling tools, notably CA-7, Control-M and Zeke. Exhibit a good working knowledge of JCL (Job Control Language)
- Conversant with Netcool support in a large-scale environment.
- Demonstrated ability to engage with IBM z-Series Engineering L3/L4/Build Teams on Architecture, Development, Stability & Continuous Improvement of Environment to advance the product’s vision and strategy to satisfy customer needs.
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our
FAQs for more information about requesting an accommodation.
Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.