Mirantis
AI Infrastructure Engineer
Remote role where the employee must remain based in a particular country.
Japan only
Employer listed it 5 weeks ago · Added today
Been open since 5 weeks ago, still being checked, but it has been live a while.
Salary
Not stated
Location
Japan only
Timezone
APAC
Contract
Full-time
Experience
Mid
Category
Customer Support
This employer didn't state pay. Jobs like this usually pay around $95k–$150k a year, a typical range taken from 219 mid-level customer support roles on Nomaders that do state pay. It's a guide, not an offer.
Remote flexibility
Work from home
This is a remote role, but the employee must be based in Japan. It is work from home rather than work from anywhere.
What the employer says
- Source listing states candidate location: "Tokyo, Tokyo, Japan, Tokyo, jp, Remote"
What Nomaders makes of it
- Residency required in Japan
- Payroll and tax are likely handled in that country only
The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.
About the role
About Mirantis
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.
We are looking for an AI Infrastructure Engineer to join our global team responsible for managing and supporting large-scale AI infrastructure environments. You will help ensure the availability, performance, and operational stability of critical AI infrastructure platforms, working closely with a distributed team across regions to provide continuous coverage and support.
This is an opportunity to work hands-on with some of the most advanced Kubernetes-based AI infrastructure in production today, while contributing to the platforms and processes that keep it running reliably at scale.
Responsibilities:
Manage and operate production AI infrastructure environments.
Lead incident response and troubleshooting efforts and deliver timely service restoration during outages or performance degradations .
Troubleshoot infrastructure and networking issues across bare-metal and/or cloud environments with multiple vendors.
Conduct root cause analysis and drive product and operational improvements.
Contribute and improve operational documentation and knowledge base.
Collaborate with global team members across time zones to ensure continuous operational coverage, including occasional work during weekends and holidays.
Proven experience managing and operating large-scale production systems (bare-metal and/or cloud).
Solid working knowledge of Kubernetes with excellent, demonstrable troubleshooting skills.
Experience configuring, customizing, and extending logging and monitoring tools (e.g., Prometheus, Grafana, ELK, or similar).
Experience with infrastructure automation technologies and Infrastructure-as-Code practices (e.g., Ansible, Terraform, or similar).
Effective verbal and written communication skills in English.
Strong analytical and problem-solving skills, with the ability to work through complex, ambiguous technical issues.
Willingness to occasionally work weekends and holidays.
Nice to have:
Previous experience building, scaling, and running High-Performance Computing (HPC) environments.
Hands-on experiences with managing large scale Kubernetes platforms in production. 
A good understanding of NVIDIA GPU technologies and the associated software stack.
Proficiency in scripting languages (e.g., Python, Bash, Go).
What does Mirantis offer you?
Requirements
- ·Strong analytical and problem-solving skills, with the ability to work through complex, ambiguous technical issues.
- ·Willingness to occasionally work weekends and holidays.
- ·Previous experience building, scaling, and running High-Performance Computing (HPC) environments.
- ·Hands-on experiences with managing large scale Kubernetes platforms in production. 
- ·A good understanding of NVIDIA GPU technologies and the associated software stack.
Benefits
- ·We are a  Leader for Container Management  in G2 (#2 after AWS)!
How to apply
- 1Check the flexibility label above, work from home, matches where you plan to live and work.
- 2Tailor your CV to the role at Mirantis, mentioning your remote working experience and working hours (APAC).
- 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.
Found 15h ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.
Listing sourced from Company boards.
Similar roles
Other open customer support roles with comparable remote rules.
Free to apply, no account needed.
Typically $95k to $150k per year · You'll be taken to the employer's careers page.