Senior IT Infrastructure Operations Engineer (L2)

Asiacruit BPO, Inc.Jakarta, JakartaOn-siteContractMid level, 2–5 yearsListed 1 day ago

Apply now

About this role

About Asiacruit

At Asiacruit, we connect top talent with forward-thinking organizations across industries. Our mission is to help businesses grow through smart, strategic, and people-focused solutions. We support companies by providing high-quality talent for both local and global markets. If you are analytical, collaborative, and committed to enabling business growth, we invite you to apply.

About the Role

We are looking for an L2 IT Infrastructure Operations Engineer to support advanced infrastructure operations in a large-scale technology environment.

You will be responsible for advanced troubleshooting, service restoration, infrastructure administration, approved changes, technical escalation, and coordination with vendors and specialist teams.

This role is suited for experienced infrastructure professionals who are comfortable supporting Linux-based environments and troubleshooting production infrastructure across servers, storage, GPU systems, and other technology platforms.

What You’ll Do

- Perform advanced troubleshooting and diagnosis of infrastructure incidents.
- Investigate and resolve complex technical issues affecting infrastructure services.
- Support service restoration and recovery of production systems.
- Administer and maintain assigned infrastructure platforms.
- Support Linux servers, GPU servers, storage systems, Kubernetes/K8s, Slurm , and related infrastructure technologies.
- Troubleshoot server, operating system, firmware, driver, and infrastructure-related issues.
- Execute approved infrastructure changes in accordance with change management procedures.
- Use scripting and automation to support troubleshooting and operational activities.
- Manage incidents, changes, and operational activities through ITSM processes.
- Coordinate with vendors and specialist technical teams during complex incidents.
- Provide technical escalation and detailed troubleshooting information to relevant teams.
- Support preventive and corrective maintenance activities.
- Maintain technical procedures, operational documentation, and incident records.
- Participate in continuous 24×7 infrastructure operations and shift support.

What We’re Looking For

- 4–7 years of experience in IT Infrastructure, Infrastructure Operations, Systems Engineering, Platform Operations , or a similar technical environment.
- Strong hands-on experience with Linux systems and infrastructure .
- Experience supporting one or more of the following: GPU servers, storage, Kubernetes/K8s, Slurm , or similar infrastructure platforms.
- Strong troubleshooting and incident-resolution skills.
- Experience with scripting or automation for infrastructure and operational tasks.
- Good understanding of ITSM, incident management, and change management .
- Basic understanding of server firmware and drivers .
- Experience coordinating with vendors or specialist technical teams.
- Strong technical documentation and communication skills.
- Good English communication skills.
- Willingness to work in a 24×7 shift-based onsite environment .

Nice to Have

- Experience supporting AI infrastructure, GPU computing, HPC, cloud platforms , or data center environments.
- Experience with infrastructure monitoring and observability tools.
- Experience with automation or infrastructure-as-code tools.
- Experience supporting large-scale enterprise technology environments.

Why Join

- Work with large-scale AI infrastructure and enterprise technology environments .
- Gain hands-on experience with Linux, GPU servers, storage, Kubernetes/K8s, Slurm, and other infrastructure technologies.
- Take ownership of advanced troubleshooting, service restoration, and infrastructure operations.
- Collaborate with specialist technical teams and technology vendors.
- Develop your expertise in modern infrastructure and AI technology environments.
- Build your career in a growing technology environment.