Garner Health logo

Garner Health

Senior Site Reliability Engineer

Undisclosed

Remote. The employer does not say where candidates may be based.

Location not stated

Employer listed it 3 weeks ago · Added yesterday

Been open since 3 weeks ago, still being checked, but it has been live a while.

Salary

$191,000 to $226,000

Location

Location not stated

Timezone

Not stated

Contract

Full-time

Experience

Senior

Category

Software

Stated by the employer in the job description

Remote flexibility

Undisclosed

The listing is advertised as remote but does not state which countries or regions candidates may work from.

Why this role is Undisclosed

We only label a role Work from anywhere, Region restricted or Work from home when the employer's own wording says so. We checked this advert under our current rules and found no country or region eligibility requirement in it. We don't guess, so it stays Undisclosed until the employer publishes enough location information. Here is exactly what the advert left out.

  • Countries you can work from: Not stated. The advert only gives "Remote", which names no country you must live in.
  • Whether the work is remote: Never mentioned. The role reached us through a remote job board, but the advert itself doesn't say the work is remote.
  • Working hours: Not stated. No timezone overlap or set hours are mentioned, so assume nothing either way.

Worth a look all the same. Missing wording is usually a rushed job posting rather than a closed door, so ask where you can be based in your first message, before you write a tailored application.

What the employer says

  • Source listing states candidate location: "Remote"

What Nomaders makes of it

  • No residency or region requirement found in the job description
  • Check with the employer before assuming you can work from abroad

The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.

About the role

What you’ll be part of

Garner is on a mission to transform the U.S. healthcare system — and we’re the only proven player doing exactly that. We partner with employers to redesign how healthcare works: applying 550+ proprietary clinical metrics across 80+ specialties to a dataset of 320M+ patients to identify the best-performing doctors, then using compelling incentives to steer members to the care that helps them get healthier, faster.

The result is a rare “win win” — better care and lower costs for both members and employers. In just five years, our work has helped over 2.5 million people access higher-quality care and saved $1B in healthcare costs. We recently raised our Series E and have doubled five years running. If you've ever wanted your work to solve a problem that touches every person in this country, this is the opportunity to do exactly that. You'd be joining a team fundamentally reimagining healthcare in the U.S. — and using AI to scale that impact further and faster than anyone else can.

About the role:

We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the cloud infrastructure powering Garner’s products and AI/ML workloads. This role sits on our Platform Engineering team. You will run the machine: defining and upholding SLOs, leading incident response, and driving the automation and standards that let every Garner engineer ship faster and more reliably. Because our systems directly influence health outcomes for millions of patients, maintaining the highest standards of production quality is imperative. This is an automation-first role: you will use AI tools to continuously convert manual operational work into monitored, hands-free processes, so the role gets more leveraged as you build.

Where you will work:

Garner is headquartered in NYC, but this position is available for individuals who are comfortable with remote work and occasional travel to HQ.

What you will do:

Run the Machine: Own the end-to-end reliability, performance, and resilience of Garner’s cloud environments (AWS, Kubernetes), including those powering AI/ML workloads; define, measure, and uphold SLOs across our critical services

Lead Incident Response: Serve in the on-call rotation, lead incident response, and drive deep-dive root cause analysis, seeing corrective actions through to resolution and rigorously reviewing infrastructure changes

Own Observability: Build and maintain the monitoring, alerting, and observability systems that let us detect and resolve issues before users feel them

Scale & Optimize: Translate ambiguous, high-performance scaling requirements into well-defined, automated, and composable infrastructure-as-code deliverables (Terraform); proactively identify and implement cost-efficiency and performance gains across the stack to maximize cloud ROI

Automate Away Toil: Pay down impactful tech debt and reduce operational toil, using AI tools and automation to convert repetitive operational work into hands-free, monitored processes, and holding our internal platform to the same rigorous standards as our customer-facing products

Enable Engineering: Build and maintain the deployment and observability standards that empower the broader engineering team to ship AI features faster and more reliably; communicate complex cloud and reliability concepts clearly to technical and non-technical stakeholders

Uphold Security & Compliance: Ensure our infrastructure and operations meet Garner’s security and HIPAA compliance obligations

The ideal candidate has:

4+ years of hands-on experience operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role

Deep expertise with Kubernetes and Terraform in a cloud-first environment (AWS preferred)

A strong track record with production observability: defining SLOs, building monitoring and alerting, and leading incident response and blameless post-incident reviews

Strong software engineering fundamentals in Python or Go, applied to infrastructure automation (experience with Kubernetes APIs a plus)

Experience driving cloud cost-efficiency and performance optimization across compute, storage, and networking

Experience supporting AI/ML or data-intensive workloads in production is a plus

Experience operating in a security-conscious or regulated environment (HIPAA, SOC 2) is a plus

Fluency with AI tools (e.g., Claude) applied to real engineering and operations workflows, or strong motivation to build it fast

Requirements

The employer hasn't listed requirements separately, they're described in the role summary above and on the original listing.

Benefits

No benefits package published with this listing. Ask about it at first interview.

How to apply

  1. 1Check the flexibility label above, undisclosed, matches where you plan to live and work.
  2. 2Tailor your CV to the role at Garner Health, mentioning your remote working experience.
  3. 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.

Found 1d ago. Last checked 23 Sept. Always confirm the details on the original posting, salary and location can change after publication.

Listing sourced from Company boards.

Similar roles

Other open software roles with comparable remote rules.

Browse all open roles

Free to apply, no account needed.

$191,000 to $226,000 · You'll be taken to the employer's careers page.