Mirantis logo

Mirantis

Senior Site Reliability Engineer (Golang / Kubernetes)

Work from home

Remote role where the employee must remain based in a particular country.

United States only

Employer listed it 3 weeks ago · Added today

Been open since 3 weeks ago, still being checked, but it has been live a while.

Salary

Not stated

Location

United States only

Timezone

US East

Contract

Full-time

Experience

Senior

Category

Software

This employer didn't state pay. Jobs like this usually pay around $170k–$225k a year, a typical range taken from 597 senior-level software roles on Nomaders that do state pay. It's a guide, not an offer.

Remote flexibility

Work from home

This is a remote role, but the employee must be based in United States. It is work from home rather than work from anywhere.

What the employer says

  • Source listing states candidate location: "Remote, REMOTE, United States, Remote, us"

What Nomaders makes of it

  • Residency required in United States
  • Payroll and tax are likely handled in that country only

The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.

About the role

Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.  https://www.mirantis.com/

Define what reliability means for a GPU-accelerated AI platform and make it measurable. You will own the service-level indicators and objectives for the K0rdent Observability Framework (KOF) — deriving meaningful SLIs from the signals the platform already emits, and exposing them to Platform Administrators through a clean API. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack.

About the Role

We are looking for a Senior SRE who thinks past dashboards to the contract between a platform and its operators. The right candidate can look at raw telemetry from Kubernetes, bare metal, and NVIDIA infrastructure, decide which signals actually predict user-visible reliability, and turn them into SLIs and SLOs that operators can act on. You are equally comfortable writing the service that exposes those SLIs through an API and reasoning about error budgets, alerting quality, and signal-to-noise. You should be self-directed, able to own reliability definitions end to end, and effectively communicate them across teams.

Responsibilities

Define SLIs and SLOs based on the signals available across the platform — Kubernetes, bare-metal hosts, and NVIDIA infrastructure (BMC, InfiniBand, NVLink, UFM).

Design and build the API that exposes SLIs and reliability state to Platform Administrators and downstream systems.

Establish alerting and error-budget practices that maximize signal and minimize noise.

Partner with infrastructure, storage, and networking teams to ensure the right signals are instrumented and collected.

Diagnose reliability and performance issues across the observability stack and drive their resolution.

Required Qualifications

5+ years in SRE, platform reliability, or a closely related software/infrastructure role.

Strong software engineering skills (e.g., Go or Python) with experience building and operating APIs or services in production.

Demonstrated experience defining SLIs/SLOs and error budgets for real production systems.

Hands-on experience with observability tooling — metrics, logging, and tracing (e.g., Prometheus/VictoriaMetrics, OpenTelemetry, Grafana).

Solid understanding of Kubernetes and the signals it and its workloads emit.

Strong written and verbal communication with technical audiences.

Preferred

Experience instrumenting or monitoring bare-metal and NVIDIA infrastructure (BMC/Redfish, InfiniBand, NVLink, UFM).

Experience with the Mirantis K0rdent stack (K0rdent Enterprise, K0rdent AI, KOF) and Cluster API.

Familiarity with VictoriaMetrics/VictoriaLogs at scale.

Proven experience in sovereign or high-security air-gapped environments.

 

What does Mirantis offer you?

Requirements

  • ·5+ years in SRE, platform reliability, or a closely related software/infrastructure role.
  • ·Strong software engineering skills (e.g., Go or Python) with experience building and operating APIs or services in production.
  • ·Demonstrated experience defining SLIs/SLOs and error budgets for real production systems.
  • ·Hands-on experience with observability tooling — metrics, logging, and tracing (e.g., Prometheus/VictoriaMetrics, OpenTelemetry, Grafana).
  • ·Solid understanding of Kubernetes and the signals it and its workloads emit.

Benefits

  • ·We are a  Leader for Container Management  in G2 (#2 after AWS)!

How to apply

  1. 1Check the flexibility label above, work from home, matches where you plan to live and work.
  2. 2Tailor your CV to the role at Mirantis, mentioning your remote working experience and working hours (US East).
  3. 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.

Found 15h ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.

Listing sourced from Company boards.

Similar roles

Other open software roles with comparable remote rules.

Browse all open roles

Free to apply, no account needed.

Typically $170k to $225k per year · You'll be taken to the employer's careers page.