Sr Lead Infrastructure Engineer - HyperV and VMware Virtualization

JPMorgan Chase & Co.London, EnglandOn-siteFull-timeSenior, 5–8 yearsListed 56 minutes ago

Apply now

About this role

Design the virtualization platforms that keep critical workloads running around the clock. You will work on enterprise-scale Microsoft Hyper-V and Broadcom VMware VCF infrastructure across global data centers, with meaningful ownership and visible impact. We value inclusive partnership across engineering, operations, security, and capacity teams, and we invest in modern automation to reduce toil and improve resiliency. If you enjoy solving complex platform problems and making systems measurably better, you will find room to grow here.

As a Senior Lead Infrastructure Engineer at JPMorgan Chase within the Compute Platform Engineering team in Enterprise Technology, you will engineer enterprise-scale Microsoft Hyper-V and Broadcom VMware VCF infrastructure across global data centers. You will partner across engineering, operations, security, and capacity teams to deliver resilient, high-performing virtualization services that support tens of thousands of workloads. You will lead design and automation decisions that improve reliability, performance, and operational excellence. You will help raise the bar through documentation, mentorship, and disciplined change practices.

Job Responsibilities

- Architect virtualization cluster designs across global pools, including software-defined compute and storage.
- Automate provisioning, configuration, patching, compliance enforcement, and decommissioning using platform scripting languages and APIs.
- Benchmark compute and storage performance using industry-standard tooling to validate changes and guide capacity decisions.
- Diagnose full-stack issues across compute, storage, networking, clustering, and hypervisor layers using structured troubleshooting methods.
- Evaluate proposed server bills of materials for suitability, providing performance-focused feedback on CPU, memory, storage, and network characteristics.
- Harden environments by partnering with security teams to remediate vulnerabilities and enforce platform baselines at scale.
- Document reference architectures, decision records, and operational runbooks to improve repeatability and auditability.
- Mentor engineers through design reviews, code reviews, and technical guidance that raises the bar for platform quality.
- Use enterprise-authorized AI capabilities to accelerate analysis of complex infrastructure signals and document mitigation options, validating outputs and handling operational data according to sensitivity and security requirements.
- Lead reuse-first adoption of AI-assisted practices across delivery and automation routines to reduce recurring issues, ensuring changes are validated, traceable, auditable, and aligned to resiliency and security expectations.

Required Qualifications, Capabilities, and Skills

- Formal training or certification in infrastructure disciplines, and substantial hands-on experience in enterprise or large-scale environments.
- Expertise in Microsoft Hyper-V and/or VMware VCF in large-scale production environments.
- Proficiency in automation, operational tooling, and configuration management using scripting languages such as Python, PowerShell, vRO, or equivalent.
- Experience performing full-stack performance assessment across compute, storage, networking, and hypervisor layers.
- Ability to analyze and optimize storage performance in virtualized environments, including throughput, latency, IOPS, and queuing.
- Experience with data-center-scale automation and configuration tooling, such as DSC, SCCM, Ansible, Salt, or equivalent.
- Ability to plan and deliver production changes with disciplined risk management, including zero-downtime upgrade strategies.
- Experience building or operating CI/CD pipelines and automated testing for infrastructure automation artifacts.
- Experience using enterprise-authorized AI capabilities to support infrastructure engineering workflows with validation habits and awareness of data sensitivity.
- Ability to review and validate AI-assisted recommendations before implementation, escalating when uncertain and ensuring outcomes align to resiliency, security, and auditability expectations.

Preferred Qualifications, Capabilities, and Skills

- Server hardware fluency (Intel Xeon, AMD EPYC, NVMe, GPU) and comfort translating workload needs into platform requirements.
- Hands-on experience with advanced virtual networking concepts, including VLANs, SDN, teaming/SET, and load-balancing patterns.
- Experience with Microsoft and Broadcom management tooling in large environments.
- Familiarity with vulnerability remediation workflows (for example, Qualys) and baseline enforcement at scale.
- Practical performance tooling experience (for example, elbencho, VMFleet, diskspd, fio) for repeatable