Sr Lead Infrastructure Engineer
Assume a vital position as a key member of a high-performing team that delivers infrastructure and performance excellence. Your role will be instrumental in shaping the future at one of the world's largest and most influential companies.
As a Lead Infrastructure Engineer (Storage) at JP Morgan Chase within Infrastructure Platforms, Foundation Services, you apply deep knowledge of software, applications, and technical processes within the infrastructure engineering discipline. Continue to evolve your technical and cross-functional knowledge outside of your aligned domain of expertise.
Job responsibilities
- Applies technical expertise and problem-solving methodologies to projects of moderate scope, primarily focused on enterprise NAS architecture, engineering standards, and strategic platform direction
- Drives a workstream or project consisting of one or more infrastructure engineering technologies (e.g., NAS reference architectures, resiliency patterns, performance engineering standards, automation/guardrails)
- Defines, documents, and governs design and configuration standards; in the initial ramp-up period, establishes and drives adoption of standardized NAS configuration and design patterns across the platform
- Leads vendor/product evaluations (requirements definition, RFI/RFP support, technical deep dives, interoperability assessments, POCs, sizing, lifecycle planning) and delivers clear, evidence-based recommendations to inform roadmap decisions
- Partners with compute, network, cloud, application, and cybersecurity teams to architect and implement changes required to resolve systemic issues and modernize technology processes
- Improves platform performance and stability by enhancing observability, performance engineering practices, and preventative controls, with a near-term focus on reducing performance incidents through root-cause remediation and durable standards
- Strongly considers upstream/downstream data and systems implications (e.g., User Authentication mechanisms (e.g. Active Directory, LDAP), encryption/key management dependencies, backup/restore, DR, network paths, application I/O profiles) and advises on mitigation actions
- Supports SRE teams as needed through technical consultation and escalation support for complex incidents and problem management, driving long-term fixes over one-off remediation
- Leverages AI-enabled engineering workflows (e.g., LLMs, MCP, skills/agents, where approved) to optimize analysis, documentation, automation, and knowledge capture
- Adds to team culture of diversity, opportunity, inclusion, and respect
Required qualifications, capabilities, and skills
- 5+ years of enterprise-scale storage experience, with significant depth in NAS platforms/services
- Deep knowledge of one or more areas of infrastructure engineering such as hardware, networking terminology, databases, storage engineering, deployment practices, integration, automation, scaling, resilience, or performance assessments
- Demonstrated strength in architecture and standardization, including reference architectures, configuration baselines, and control/guardrail design
- Experience with storage products from vendors such as NetApp, Dell, VAST Data, Weka, DDN, Pure Storage (or equivalent), with the ability to rapidly learn and critically evaluate new platforms
- Proven ability to work independently to understand vendor products and services, ask the right questions, and translate requirements into secure, scalable solutions
- Deep knowledge of one specific infrastructure technology and scripting languages (e.g., Python, PowerShell, Bash, etc.); experience automating repeatable operational tasks is strongly valued
- Strong troubleshooting and problem-solving skills, including ability to support SRE teams during escalations and drive durable root-cause fixes
- Drives to continue to develop technical and cross-functional knowledge outside of the product
- Strong written and verbal communication skills, including producing clear standards, designs, and operational documentation
Preferred qualifications, capabilities, and skills
- Experience with public cloud providers (AWS/Azure/GCP), hybrid connectivity patterns, and storage/migration concepts
- Experience with automation / IaC / configuration management (e.g., Terraform, Ansible, or similar) and CI/CD-aligned engineering practices
- Familiarity with Agile tools and processes (e.g., Jira/Confluence) and iterative delivery
- Familiarity with AI systems applied to engineering productivity, including safe/controlled use of LLMs, agentic workflows, MCP, and skills/agents (subject to firm-approved tools and data-handling requirements)
- Experience operating in regulated or highly controlled environments, including audit/control participation and operational risk management