Senior Lead Software Engineer - SCM Platform
Be an integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch technology products.
As a Senior Lead Software Engineer at JPMorganChase within the Enterprise Technology - ENG Services & Platform team, you are an integral part of an agile team that works to enhance, build, and deliver trusted market-leading technology products in a secure, stable, and scalable way. Drive significant business impact through your capabilities and contributions, and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of challenges that span multiple technologies and applications.
Responsible for defining the vision, strategy, and execution for all aspects of SCM platform operations, including support, maintenance, hygiene, and compliance, while providing people leadership through hiring, developing, and mentoring a diverse, high-performing operations team. Lead incident management, root cause analysis, and continuous improvement to strengthen platform reliability, and own operational reporting along with audit and compliance support. Drive automation across monitoring, patching, upgrades, and branch/repo cleanup, and manage SLOs/SLAs by overseeing performance, scalability, and capacity planning. Implement reliability engineering practices such as chaos engineering and proactive risk mitigation to reduce operational risk and improve resilience.
Job responsibilities
- Vision, strategy, and execution for all aspects of SCM platform operations, including support, maintenance, hygiene, and compliance.
- Hire, develop and mentor a diverse, high-performing operations team.
- Drives adoption and governance of approved AI-assisted engineering practices across teams to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test acceleration, release readiness, incident/root-cause analysis), while establishing measurable validation standards (secure coding, peer review, automated testing) and promoting reuse of proven patterns and automation within the SDLC/TLM toolchain.
- Applies knowledge of tools within the Software Development Life Cycle toolchain, including approved AI-assisted development and automation capabilities, to improve the value realized by automation at scale.
- Proactively handles incident management, root cause analysis, and continuous improvement of platform reliability.
- Effectively manages operational reporting, audit, and compliance support.
- Automation of monitoring, patching, upgrades, and branch/repo cleanup.
- Handles the SLO/SLA management, performance, scalability, and capacity planning.
- Implementation of reliability engineering practices, including chaos engineering and proactive risk mitigation.
- Regularly provides technical guidance and direction to support the business and its technical teams, contractors, and vendors
- Drives decisions that influence the product design, application functionality, and technical operations and processes
Required qualifications, capabilities, and skills
- Formal training or certification on software engineering concepts and 5+ years applied experience
- Hands-on practical experience delivering system design, application development, testing, and operational stability
- Set direction and provide hands-on leadership for day-to-day operations of SCM platforms (Bitbucket, GitHub, GitLab).
- Demonstrated experience leading effective use of enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security
- Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching senior engineers/leads on compliant usage patterns and controls.
- Ensure platforms are stable, compliant, reliable, and efficient under all conditions.
- Proven experience leading and developing operations or reliability engineering teams for large-scale platforms.
- Strong hands-on experience with platform support, maintenance, and automation in production environments.
- Demonstrated success in driving operational excellence, reliability, and compliance.
- Deep understanding of incident management, root cause analysis, and continuous improvement.
- Experience with SLO/SLA management, performance monitoring, and capacity planning.
- Experience with reliability engineering, chaos engineering, and risk mitigation for critical platforms.
- Deep knowledge of audit, compliance, and operational reporting best practices.
- Experience implementing automation for monitoring, patching, upgrades, and hygiene.
- Strong collaboration skills with engineering, security, and product teams.
- Executive presence and ability to communicate complex operational topics to senior leadership.