Regard Logo

Regard

Senior Data Engineer

Posted 3 Days Ago
Hybrid
New York, NY, USA
165K-220K Annually
Senior level
Hybrid
New York, NY, USA
165K-220K Annually
Senior level
Design, build, and operate data services and pipelines to ingest, standardize, and deliver clinical data for product, analytics, and ML. Tune Spark workloads, enforce data quality, monitor production systems, and partner with product and engineering. Participate in on-call support.
The summary above was generated by AI

As a Senior Data Engineer at Regard, you will own the design, development, and production deployment of the data services that power the Regard platform. From ingesting and standardizing clinical data across health systems to making it reliably available for downstream product, analytics, and machine learning workflows, you'll build and evolve the infrastructure that enables the platform. This includes analyzing and tuning Spark workloads and partitioning strategies to control costs, adapting to upstream breaking changes, and enforcing rigorous data quality standards so our analytics are as dependable as our application code. We prioritize transparent, code-driven systems over black-box services, and you'll help architect the data platform that supports that philosophy.

 

About Regard

Our mission is to bring world-class healthcare to everyone. Regard is an AI-powered Proactive Documentation platform that advances how care is delivered by reviewing all patient data in the EHR to recommend diagnoses and surface clinical evidence. Regard drafts a note even before the physician sees the patient, enabling an approach that gets  documentation right at the point of care - we call it Proactive Documentation. This improves quality of care, reduces physician burden, and improves hospital finances. We are excited by challenges, mission-oriented work, and meaningful relationships. We work closely with some of the top health systems in the country and are leading the change that healthcare - one of the largest and most inefficient industries in the world - needs. We want you to join us.

Our Tech Stack:

  • Data: S3, Apache Iceberg, EMR, PySpark, Dagster, Kubernetes, Clickhouse, PostgreSQL, FastAPI, Metabase

 

Responsibilities:

  • Collect, model, and consolidate data into the data platform to support analytics, ML development, and research initiatives

  • Design, build, and evolve data models and pipelines that reliably transform and deliver data to downstream consumers

  • Own data quality in collaboration with engineering teams, ensuring datasets are trustworthy and production-ready

  • Partner closely with product to deliver analytics and actionable insights to internal and external stakeholders

  • Own the reliability and day-to-day operation of the data platform and its pipelines through proactive monitoring, alerting, and operational management

Minimum Qualifications:

  • Bachelors degree in Computer Science, Mathematics, Statistics, or a related field, or equivalent practical experience

  • 5+ years of experience in data engineering roles

  • 3+ years of experience using PySpark to build data pipelines

  • 3+ years of experience in public cloud provider technologies (AWS tooling such as S3, EMR, or Athena)

  • Strong proficiency in Python and SQL

  • Hands-on experience across the full data stack, with particular depth in data modeling and pipeline design

  • Practical experience with LLM-assisted development, with an understanding of its capabilities and limitations

  • Willingness to participate in on-call operational support for owned systems

Preferred Qualifications:

  • Experience with one or more of the following technologies: Apache Iceberg, Dagster, Clickhouse, PostgreSQL, FastAPI, Metabase

  • Experience working with healthcare data, including HIPAA compliance, data de-identification, and familiarity with open data standards such as OMOP CDM

  • Experience building and supporting data pipelines for ML workflows, including model training, validation, deployment, and ongoing performance evaluation

Hybrid Work | Location | Work Authorization

  • For this role, Regard is currently only considering candidates who are authorized to work in the US without visa sponsorship, and are within the New York City, Los Angeles, or San Francisco metro areas

  • We expect our Engineers to be in the office on Tuesdays and Thursdays. We also require more frequent in-office work during the onboarding period and team onsite weeks up to once per month

  • We will provide relocation assistance to anyone who does not already reside in the NYC metro area

  • We prefer hiring people within commuting distance of our offices because we value getting together in person regularly

  • For those who enjoy working from our LA or Manhattan offices on a more regular basis, we offer catered lunches and other fun perks

  • Additionally, hybrid employees have the flexibility to work from locations outside of their home office from up to 6 weeks per year

Comp | Perks | Benefits

  • Eligible for equity

  • 99% employer paid health benefits (Medical, Dental, and Vision) + One Medical subscription

  • 18 PTO days/yr + 1 week holiday break

  • Monthly health & wellness budget

  • Company-sponsored team retreat + social events

  • A sabbatical program

Our goal at Regard is to provide and maintain a work environment that fosters mutual respect, professionalism and cooperation. Regard is proud to be an equal opportunity employer that does not discriminate on the basis of actual or perceived race, creed, color, religion, national origin, ancestry, alienage or citizenship status, age, disability or handicap, sex, gender identity, marital status, familial status, veteran status, sexual orientation or any other characteristic protected by applicable federal, state or local laws. We celebrate diversity and are proud of our supportive, inclusive workplace.

 

All candidates must successfully complete a background check as part of the hiring process.

Similar Jobs

10 Days Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
134K-203K Annually
Senior level
134K-203K Annually
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead design and operation of large-scale data pipelines and data APIs. Build Spark/PySpark workflows on Databricks, optimize job performance, manage data quality and observability, and develop MCP servers and AI-agent integrations. Mentor engineers, define standards, support production incidents and on-call rotations, and collaborate with stakeholders to deliver scalable data platform products.
Top Skills: Apache IcebergApi GatewayAWSAws LambdaAws Rds/AuroraAzureDatabricksDatadogDbtFastapiFivetranGCPGoogle BigqueryLlms/Ai AgentsMcp ServersMs Sql ServerMySQLOraclePostgresPysparkPythonS3SecretsmanagerSnowflakeSnsSparkSplunkSQLSqs
10 Days Ago
Remote or Hybrid
United States
70K-160K Annually
Senior level
70K-160K Annually
Senior level
Cloud • Insurance • Payments • Software • Business Intelligence • App development • Big Data Analytics
Design, build, and optimize BigQuery-based data solutions and data models. Manage data storage, partitioning, clustering, ETL, and data quality. Collaborate with architects, data scientists, and engineers to deliver scalable, secure cloud data infrastructure and documentation; senior hires provide technical leadership, code review, and process improvement.
Top Skills: Ansi Sql)BigQueryConfluenceGoogle Cloud Platform (Gcp)JIRAPythonSql (Bigquery Sql
12 Days Ago
In-Office
2 Locations
153K-204K Annually
Senior level
153K-204K Annually
Senior level
Cloud • Information Technology • Machine Learning
Own and evolve the data lake and analytics stack powering fleet observability. Build and maintain scalable ETL/ELT pipelines, operate data lake infrastructure (Iceberg, Trino, Airflow, Spark, Superset), produce analytics and executive reporting, create dashboards and runbooks, and implement data governance and observability.
Top Skills: Apache AirflowApache HudiApache IcebergSparkApache SupersetAWSAzureDelta LakeEtl/EltGCPJavaManaged DatabasesNoSQLObject StoragePythonScalaSQLTrino

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account