Regard logo

Regard

Senior Data Engineer

Hybrid

Part remote, part office, you need to live within commuting distance of a named location.

Hybrid · New York, NY

Employer listed it 2 months ago · Added yesterday

Been open since 2 months ago. Long-running listings are sometimes left up after the role is filled.

Salary

$165,000 to $220,000

Location

Hybrid · New York, NY

Timezone

Not stated

Contract

Full-time

Experience

Senior

Category

Data

Published by the employer

Remote flexibility

Hybrid

This role is only partly remote, the employer expects time in the office around New York, NY, San Francisco, CA, Los Angeles, CA, Hybrid, so you need to live within commuting distance.

What the employer says

  • Source listing states candidate location: "New York, NY, San Francisco, CA, Los Angeles, CA, Hybrid"
  • Listing mentions "Hybrid"

What Nomaders makes of it

  • Not suitable if you plan to move between countries

The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.

About the role

As a Senior Data Engineer at Regard, you will own the design, development, and production deployment of the data services that power the Regard platform. From ingesting and standardizing clinical data across health systems to making it reliably available for downstream product and analytics workflows, you'll build and evolve the infrastructure that enables the platform. This includes analyzing and tuning Spark workloads and partitioning strategies to control costs, adapting to upstream breaking changes, and enforcing rigorous data quality standards so our analytics are as dependable as our application code. We prioritize transparent, code-driven systems over black-box services, and you'll help architect the data platform that supports that philosophy.

About Regard

Regard's mission is to bring world-class healthcare to everyone. Our technology reasons through a patient's entire medical record to recommend diagnoses that would otherwise be missed, in real time at the point of care, during chart review, and in population-wide screening to identify patients who qualify for lifesaving treatment.

We work alongside some of the top health systems in the country to lead the change this industry needs. We're excited by challenges, mission-oriented work, and meaningful relationships. We want you to join us.

Our Tech Stack:

Data: S3, Apache Iceberg, EMR, PySpark, Dagster, Kubernetes, Clickhouse, PostgreSQL, FastAPI, Metabase

Responsibilities:

Collect, model, and consolidate data into the data platform to support analytics and research initiatives

Design, build, and evolve data models and pipelines that reliably transform and deliver data to downstream consumers

Own data quality in collaboration with engineering teams, ensuring datasets are trustworthy and production-ready

Partner closely with product to deliver analytics and actionable insights to internal and external stakeholders

Own the reliability and day-to-day operation of the data platform and its pipelines through proactive monitoring, alerting, and operational management

Minimum Qualifications:

Bachelors degree in Computer Science, Mathematics, Statistics, or a related field, or equivalent practical experience

5+ years of experience in data engineering roles

3+ years of experience using PySpark to build data pipelines

3+ years of experience in public cloud provider technologies (AWS tooling such as S3, EMR, or Athena)

Strong proficiency in Python and SQL

Hands-on experience across the full data stack, with particular depth in data modeling and pipeline design

Practical experience with LLM-assisted development, with an understanding of its capabilities and limitations

Willingness to participate in on-call operational support for owned systems

Preferred Qualifications:

Experience with one or more of the following technologies: Apache Iceberg, Dagster, Clickhouse, PostgreSQL, FastAPI, Metabase

Experience working with healthcare data, including HIPAA compliance, data de-identification, and familiarity with open data standards such as OMOP CDM

Requirements

  • ·Bachelors degree in Computer Science, Mathematics, Statistics, or a related field, or equivalent practical experience
  • ·5+ years of experience in data engineering roles
  • ·3+ years of experience using PySpark to build data pipelines
  • ·3+ years of experience in public cloud provider technologies (AWS tooling such as S3, EMR, or Athena)
  • ·Strong proficiency in Python and SQL

Benefits

  • ·Eligible for equity
  • ·99% employer paid health benefits (Medical, Dental, and Vision) + One Medical subscription
  • ·18 PTO days/yr + 1 week holiday break
  • ·Monthly health & wellness budget
  • ·Company-sponsored team retreat + social events

How to apply

  1. 1Check the flexibility label above, hybrid, matches where you plan to live and work.
  2. 2Tailor your CV to the role at Regard, mentioning your remote working experience.
  3. 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.

Found 1d ago. Last checked 23 Sept. Always confirm the details on the original posting, salary and location can change after publication.

Listing sourced from Company boards.

Similar roles

Other open data roles with comparable remote rules.

Browse all open roles

Free to apply, no account needed.

$165,000 to $220,000 · You'll be taken to the employer's careers page.