Grow Therapy
Senior Platform Reliability Engineer
Part remote, part office, you need to live within commuting distance of a named location.
Hybrid · San Francisco
Employer listed it 3 months ago · Added yesterday
Been open since 3 months ago. Long-running listings are sometimes left up after the role is filled.
Salary
$182k to $250k per year
Location
Hybrid · San Francisco
Timezone
Not stated
Contract
Full-time
Experience
Senior
Category
Software
Published by the employer
Remote flexibility
Hybrid
This role is only partly remote, the employer expects time in the office around San Francisco, Seattle, New York City, Hybrid, so you need to live within commuting distance.
What the employer says
- Source listing states candidate location: "San Francisco, Seattle, New York City, Hybrid"
- Listing mentions "Hybrid" and three days per week in the office
What Nomaders makes of it
- Not suitable if you plan to move between countries
The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.
About the role
About Us:
Grow was born to tackle a critical challenge: making mental health care more effective and accessible for everyone. Since launching in 2020, 15 million sessions have started on Grow, and we've barely scratched the surface. We do it through a three-sided marketplace that empowers providers, augments insurance payors, and serves patients. Every solution we build reveals ten more waiting to be delivered, and that's just what gets us going. We've backed that ambition with more than $328M in funding, including our Series D at a $3B valuation from Sequoia Capital, Transformation Capital, TCV, SignalFire, Menlo Ventures, Goldman Sachs Alternatives, and others.
At Grow, the work you do can literally change someone's life. Your work here doesn't stay on a screen; it impacts whether someone can find, afford, and access quality care. That ambition sets us apart, and offers you an endless runway to impact one of the world's most urgent issues. We don’t have passenger seats here. Everyone is driving something that matters.
This work deserves real commitment. So if you thrive on meeting problems head on, untangling serious complexity, and working at pace with the sharpest yet kindest people around, you've come to the right place.
About the Role
We’re hiring a Senior Platform Reliability Engineer to help define and scale reliability as a first-class capability at Grow. In this role you’ll operate horizontally across the organization, shaping how reliability is understood, measured, and built into the developer experience.
You’ll work closely with other members of the platform team as well as our product engineering teams to establish standards around observability, SLOs/SLAs, and incident response—while also helping translate those standards into self-service tooling and “golden paths” that make it easy for teams to adopt them.
This is a high-impact, highly autonomous role where you’ll drive both cultural and technical change, ultimately enabling teams to independently build and operate reliable systems at scale.
What You'll Work On
You’ll help us establish and scale reliability as a discipline at Grow by:
Defining Reliability Standards Establishing frameworks for SLOs/SLAs, error budgets, and operational readiness; helping teams understand what to measure and why it matters.
Improving Observability & Measurement Identifying gaps in metrics, logging, and tracing; ensuring services are measurable, debuggable, and aligned with reliability goals.
Evolving Incident Response Developing and improving incident response practices, from detection to post-incident learning, and helping teams build sustainable on-call and escalation patterns.
Enabling Self-Service Reliability Partnering with the platform team to build tooling and abstractions (e.g., service scorecards, dashboards, templates, golden paths) that make it easy for teams to adopt and stay compliant with reliability standards.
Driving Adoption Across Teams Working cross-functionally to educate, influence, and guide engineering teams—scaling reliability practices through a combination of clear standards, strong communication, and developer-friendly systems
Who You Are
Experienced in production systems: You have 6+ years of experience operating and improving reliability of production systems at scale.
Strong foundation in cloud and infrastructure: You have hands-on experience with AWS, Kubernetes (e.g., EKS), and infrastructure as code tools like Terraform.
Deep understanding of reliability principles: You’ve defined or worked with SLOs/SLAs, understand error budgets, and have experience improving reliability through measurement and iteration.
Observability expertise: You’ve worked with modern observability tooling (we use DataDog) and understand how to build actionable monitoring systems across metrics, logs, and traces.
Systems thinker: You’re able to zoom out, identify patterns across teams and services, and design solutions that scale beyond a single system.
Impact-oriented: You focus on outcomes over output and care deeply about improving real reliability outcomes—not just adding processes.
Strong communicator and influencer: You can drive change across teams without direct authority, balancing pragmatism with long-term vision.
Self-directed: You thrive in ambiguous environments and are comfortable defining problems, proposing solutions, and executing independently.
Requirements
The employer hasn't listed requirements separately, they're described in the role summary above and on the original listing.
Benefits
- ·Health Benefits : Comprehensive medical, dental, vision, life, and disability coverage.
- ·Grow for Grow: No cost access to therapy through the Grow platform (available to US employees)
- ·Financial Wellness : Retirement savings programs and equity opportunities to help you invest in your future.
- ·Flexible Time Off, Paid Holidays & Winter Break : Flexible time off, company paid holidays (which vary by country), and a full company wide Winter Break to rest and recharge
- ·Parental Leave : Up to 18 weeks of paid parental leave to support you and your growing family and a new child stipend.
How to apply
- 1Check the flexibility label above, hybrid, matches where you plan to live and work.
- 2Tailor your CV to the role at Grow Therapy, mentioning your remote working experience.
- 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.
Found 1d ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.
Listing sourced from Company boards.
Similar roles
Other open software roles with comparable remote rules.
Free to apply, no account needed.
$182k to $250k per year · You'll be taken to the employer's careers page.