About this role
We are Lenovo. We do what we say. We own what we do. We WOW our customers.
Lenovo is a US$83 billion revenue global technology powerhouse, ranked #153 in the Fortune Global 500, and serving millions of customers every day in 180 markets. Focused on a bold vision to deliver Smarter Technology for All, Lenovo has built on its success as the world’s largest PC company with a full-stack portfolio of AI-enabled, AI-ready, and AI-optimized devices (PCs, workstations, smartphones, tablets), infrastructure (server, storage, edge, high performance computing and software defined infrastructure), software, solutions, and services. Lenovo’s continued investment in world-changing innovation is building a more equitable, trustworthy, and smarter future for everyone, everywhere. Lenovo is listed on the Hong Kong stock exchange under Lenovo Group Limited (HKSE: 992) (ADR: LNVGY).
This transformation together with Lenovo’s world-changing innovation is building a more inclusive, trustworthy, and smarter future for everyone, everywhere. To find out more visit www.lenovo.com (https://www.lenovo.com), and read about the latest news via our StoryHub (https://news.lenovo.com/).
Join us at the forefront of AI innovation. We’re redefining what’s possible with GPU-powered, datacenter-scale infrastructure—and we’re seeking a GPU Cluster Performance Intern to join our product development team. In this role, you will contribute to:
- Workload Analysis: Evaluating the performance of representative AI workloads and their core components, including GPU collective communication libraries (e.g., NCCL, RCCL) and distributed training frameworks.
- System Optimization: Driving performance tuning through deep-dive analysis of full hardware and software stacks, emerging transport protocols (e.g., UEC CSIG), and network hardware configurations (NICs and switches).
- Research & Innovation: Translating findings into high-impact internal reports and papers at top-tier conferences.
Qualifications
- Current Master’s or PhD student in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
- Solid understanding of x86/Arm server architectures, high-speed networking (e.g., InfiniBand/RoCE, Ethernet), and Linux OS environments.
- Proficiency in Python and shell scripting for developing cluster benchmarking, deployment, and test automation tooling.
- Strong verbal and written communication skills with the ability to distill complex performance data into clear, actionable technical analyses and reports.
- Ability to quickly master new technologies and concepts, including GPU scale-up/scale-out fabrics, modern AI infrastructure, and distributed cluster software stacks.
We are an Equal Opportunity Employer and do not discriminate against any employee or applicant for employment because of race, color, sex, age, religion, sexual orientation, gender identity, national origin, status as a veteran, and basis of disability or any federal, state, or local protected class.