About this role
Location: Tokyo, Japan
About the Role
Our client is a global technology company developing next-generation AI infrastructure solutions. They are seeking a Forward Deployed Engineer to support the deployment, integration, and operation of large-scale AI systems within customer environments across Japan.
Key Responsibilities
- Deploy and integrate AI infrastructure solutions in customer environments.
- Manage and optimize Kubernetes-based platforms and Helm deployments.
- Troubleshoot issues across AI platforms, distributed systems, networking, and infrastructure.
- Develop automation tools, scripts, and operational workflows.
- Collaborate closely with customers and internal engineering teams to ensure successful deployments and ongoing operations.
- Serve as a trusted technical advisor throughout implementation and production support.
Required Qualifications
- 5+ years of experience in Infrastructure Engineering, Platform Engineering, SRE, MLOps, Machine Learning Engineering, or related roles.
- Strong software engineering skills with advanced Python proficiency.
- Experience with Linux, Kubernetes, Helm, CI/CD, and infrastructure automation.
- Solid understanding of distributed systems and cloud-native technologies.
- Experience supporting production environments and troubleshooting complex systems.
- Fluent Japanese and professional-level English communication skills.
Preferred Qualifications
- Experience with AI/ML infrastructure, HPC, or large-scale compute environments.
- Knowledge of AI serving frameworks such as vLLM or SGLang.
- Experience with accelerator-based computing (GPU/AI hardware).
- C++ development experience.
- Enterprise customer-facing experience.