Scan.com Logo

Scan.com

Data Engineer II

Posted 10 Days Ago
Hybrid
New York City, NY, USA
130K-150K Annually
Mid level
Hybrid
New York City, NY, USA
130K-150K Annually
Mid level
Own end-to-end data work: build and maintain dbt models and pipelines, ensure warehouse data quality, design Tableau dashboards, automate recurring reports, partner with stakeholders to translate needs into analytical products, and leverage Python and AI tools to improve data infrastructure and analytics.
The summary above was generated by AI

We’re Scan.com, the digital health scale-up making diagnostics accessible, fast, and transparent. Our technology speeds up diagnoses for timely treatments, improving healthcare outcomes for hundreds of patients each day.

We're doing diagnostics differently, with solutions tailored to both patients and providers, all backed by our technology and world-class customer operations team. Our B2C marketplace simplifies booking a scan, making it as straightforward for patients as booking a hotel. Our B2B platforms provide live scheduling at the point of care and harness AI to ease workflows for physicians, attorneys, and providers.

WHAT YOU WILL BE GETTING INVOLVED IN

Data at Scan.com is a full-stack discipline. We don't have analysts who hand off to engineers to build the model, or engineers who hand off to analysts to build the dashboard. You will own the work end to end, from raw source data through the transformation layer through the product that a stakeholder uses to make a decision.

 

Our data team supports every function across the US and UK: operations, revenue cycle, marketing, sales, provider success, and product. We sit at the center of a fast-moving business, and the work ranges from answering a precise operational question to building the automated reporting infrastructure that makes that question answerable without us in the loop next time.

 

We use a modern tech stack: Fivetran and Python for ingestion, Snowflake as our data warehouse, dbt for transformations, GitHub for CI/CD, Tableau for visualizations, and Cursor, Claude, and Snowflake Cortex for AI-assisted development. Our source systems span Postgres, HubSpot, Front, Acuity, Dialpad, and Facebook and Google Ads.

 

As a scale-up business, you can expect your role to develop over time. Here are some of the types of things you could be getting involved in:

  • Build and maintain dbt models that transform raw source data into reliable, well-documented, tested analytical assets used across the business

  • Partner with stakeholders across Operations, Finance, Marketing, Sales, and Product to translate ambiguous business questions into precise analytical deliverables. You should be able identify if the ask needs refinement before the work begins

  • Design and ship Tableau dashboards and data products that automate reporting that currently requires manual effort, permanently removing recurring analytical burden from the team

  • Identify gaps in our data coverage and work with engineering or independently to close them. This can mean writing a Python ingestion pipeline or modeling a new source

  • Maintain data quality standards across the warehouse: write dbt tests, monitoring for anomalies, and ensuring that the numbers stakeholders see are trustworthy

  • Contribute to the team's analytical infrastructure. This means shared macros, source definitions, documentation standards, and CI/CD practices.

  • Stretch into various data science or data engineering projects based on business demand and based on your own personal development: pipeline development, ML feature preparation, or predictive modeling.

THE TOP 5 THINGS WE WANT YOU TO ACHIEVE IN YOUR FIRST YEAR
  • Deep business context. Within 90 days, you understand the unit economics, operational workflows, and key performance drivers across the functions you support. You can receive a stakeholder request and immediately identify whether it's answerable with existing models, requires new modeling work, or requires a better-defined question.

  • Reliable analytical infrastructure. You have meaningfully improved the coverage, test depth, and documentation of our dbt model layer — making it easier for the next analyst to build on your work and reducing the frequency of data quality incidents.

  • High-impact data products shipped. Several Tableau dashboards or analytical products are in production and actively used by stakeholders to make decisions, replacing manual reporting or filling a genuine analytical gap.

  • Automation wins. You have identified and eliminated at least one recurring manual reporting process, replacing it with a scheduled, automated output that runs without human intervention.

  • Trusted analytical partner. Stakeholders across the business seek out your input before they finalize requirements, not after. You are known for asking the right clarifying questions and delivering work that answers the real question, not just the stated one.

WHAT YOU MIGHT BRING TO THE TABLE

You don't need to tick all the boxes to apply for this role. Whether it's your first role or your fifth, we believe everyone can add value, learn, and grow. However, these might be some of the ways you are currently adding value:

  • Strong SQL: you write complex queries fluently, understand query performance, and don't need a template to construct a multi-stage transformation

  • Hands-on experience with a modern data transformation tool, ideally dbt or SQLMesh. You understand the model DAG, have written tests and documentation, and are familiar with CI/CD practices in a transformation layer

  • Experience with a modern cloud data warehouse, ideally Snowflake or BigQuery — you understand how storage and compute interact and can write warehouse-idiomatic SQL

  • Proficiency with a BI tool such as Tableau, Metabase, Sigma, or Omni — you can design a dashboard that stakeholders actually use, not just one that technically answers the question

  • Strong stakeholder instincts: you have worked directly with non-technical stakeholders to scope and deliver analytical work, and you know how to translate between business language and data model logic

  • Python for data work. Ingestion pipelines with libraries like dlt, pandas-based transformation scripts, or scripted automation of manual reporting processes

  • For this role, the expectation is that you are an AI-Native. You build with AI directly, shipping, automating, or prototyping real work with LLMs and agents, and have clear judgment on where humans must stay in the loop.

  • Startup experience is a plus. You are comfortable with ambiguity, can prioritize without perfect information, and don't need a fully defined ticket to get started

  • Healthcare experience is a plus, particularly in imaging, RCM, or provider operations — but strong analytical fundamentals in any high-velocity domain are equally valued

HOW WE WILL INTERVIEW YOU

We keep our interview process short and sweet, and we're a nimble team that can progress at pace. Here are the stages you can expect, but we might switch up the order depending on team availability:

  • Introductory call with our Senior Talent Partner — approximately 30 minutes by phone.

  • Video call with the hiring manager — approximately 45 minutes, structured deep-dive into the role and technical domain.

  • Assessment stage — may be a take-home exercise, in-person session, or additional video calls. We're mindful of your time.

  • Meet the founders and/or other team members.

  • Offer!

BENEFITS

We go beyond the basics with our benefits package. Here's what you can expect from us:

  • Competitive salary range, plus performance bonus and equity

  • 401k

  • Healthcare, Vision, and Dental

  • All equipment needed to do your role effectively

  • Flexible and remote/hybrid working options

  • Personal development budgets

  • 18 days PTO plus public holidays

  • 10 paid sick days

  • Inclusive policies designed by our team, for our team

Diversity at SCAN.COM

Scan.com is committed to eliminating discrimination and encouraging diversity within our team.

We strive to provide equality and fairness for all job applicants and employees, and never discriminate on the basis of gender, marital status, age, race, ethnicity, religion, or physical differences.

We are opposed to all forms of unlawful treatment and discrimination.

Our ambition is for our team and its Board to be representative of the diversity in society, and for every employee to feel respected and able to bring their best selves to work.

Similar Jobs

18 Hours Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
134K-203K Annually
Senior level
134K-203K Annually
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead design and operation of large-scale data pipelines and data APIs. Build Spark/PySpark workflows on Databricks, optimize job performance, manage data quality and observability, and develop MCP servers and AI-agent integrations. Mentor engineers, define standards, support production incidents and on-call rotations, and collaborate with stakeholders to deliver scalable data platform products.
Top Skills: Apache IcebergApi GatewayAWSAws LambdaAws Rds/AuroraAzureDatabricksDatadogDbtFastapiFivetranGCPGoogle BigqueryLlms/Ai AgentsMcp ServersMs Sql ServerMySQLOraclePostgresPysparkPythonS3SecretsmanagerSnowflakeSnsSparkSplunkSQLSqs
3 Days Ago
In-Office
New York City, NY, USA
83K-132K Annually
Mid level
83K-132K Annually
Mid level
Fintech
Build data infrastructure and analytics tools for Data Science (ETL, cluster computing, data ingestion/storage on Hadoop/S3); mentor interns; guide data scientists on software engineering best practices; collaborate with analytics and engineering teams on platform and architectural improvements; support risk management and data confidentiality.
Top Skills: Aws S3ETLHadoopJavaMachine Learning ToolkitsPythonRScalaSparkSQL
2 Days Ago
Remote or Hybrid
USA
100K-145K Annually
Mid level
100K-145K Annually
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Build and maintain cloud cost monitoring, alerting, and optimization tooling. Develop automation, cost allocation/tagging, data pipelines, dashboards, and integrations with billing APIs. Support FinOps processes, implement cost guardrails, troubleshoot cost systems, participate in on-call rotation, and collaborate with senior engineers and product teams to drive cost-efficient infrastructure and reporting.
Top Skills: AWSAzureDockerGCPGitGoPython

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account