Novellia Logo

Novellia

Senior Machine Learning Engineer, Platform

Posted One Month Ago
Remote
Hiring Remotely in USA
150K-200K Annually
Senior level
Remote
Hiring Remotely in USA
150K-200K Annually
Senior level
Own the full lifecycle of clinical text extraction models, from problem framing and annotation strategy through production deployment, evaluation, monitoring, and retraining. Build clinical NLP pipelines, establish defensible quality and error-analysis systems, and apply LLMs using rigorous evaluation, observability, and cost controls. Partner with clinical data and backend engineering teams while addressing PHI protection, de-identification, auditability, and access controls. As the first ML hire, set technical direction and potentially build and lead the future ML team.
The summary above was generated by AI
Our Story
 

Since 2023, our mission has been clear: to be the north star of patient equity.

 

Every day, we strive to bridge the gaps in healthcare access and outcomes to ensure that every patient, regardless of background or circumstance, receives the care they deserve. As a member of our team, you'll be at the forefront of innovation, working alongside passionate individuals who share your dedication to creating best-in-class, patient-centric products in healthcare. Together, we're revolutionizing the way people understand their health and working with the world's top researchers to accelerate innovation.

About Novellia

Novellia is the first and only company that lets anyone in the U.S. gain access to nearly a decade of their health data in under 30 seconds — 100% free. All your health records, across every doctor, in one place, always up to date.

We are the only patient-powered real-world data platform delivering comprehensive, patient-authorized longitudinal health insights to accelerate biopharma innovation. Unlike traditional RWD providers who deliver fragmented institutional data, we empower patients to access 20+ years of their health records, then transform these complete health journeys into fit-for-purpose datasets for evidence generation, regulatory submissions, and market access. We are growing 5x year over year, have raised close to $30M in funding, and are backed by tier-1 investors including Spark Capital, Khosla Ventures, and Bling Capital.

Working with the world's top researchers, we turn health insights into life-changing action for millions of people around the world.

About the role

Most of what matters in a health record isn't in a structured field - it's in the note, the discharge summary, the pathology report, the scanned fax. Turning that unstructured clinical text into trustworthy, structured features is what makes a longitudinal health history usable for research, and it's one of the highest-leverage capabilities Novellia can own.

You'll be our first ML hire, joining Platform Engineering and reporting to the Head of Platform Engineering, as technical owner of this multi-quarter effort. The interesting decisions are still open - what we extract first, how we know we're right, what a mature extraction pipeline looks like at our scale. There's no existing approach to inherit or defend.

The work draws on two toolkits. Roughly 70% is applied ML on clinical text: entity extraction, classification, sequence labelling, annotation strategy, error analysis, calibration, and the evaluation discipline that tells you whether your numbers mean anything. Roughly 30% is LLM-based: prompt development, structured output, retrieval, and the evals and observability that keep generative approaches honest. Deciding which approach a given problem calls for is the most interesting part of the job, and that call is yours.

We're looking for a leader in this seat: setting technical direction rather than waiting to be handed a problem. If this grows the way we think it will, leading the team we build around it is on the table.

What you'll do
  • Own the full lifecycle of extraction models - framing, data/annotation strategy, model selection, training/fine-tuning, evaluation, deployment, monitoring, retraining. Not a research seat, not a hand-off seat.

  • Define what "accurate enough" means with clinical and customer-facing stakeholders, and build the evaluation harness that makes the answer defensible - the first deliverable, not a follow-up.

  • Partner with Clinical Data Managers on curation design and own the technical half of QA/QC alongside them: which variables are extractable, how an instruction becomes a model spec, and the tooling/sampling/error analysis behind human-in-the-loop review.

  • Build clinical NLP pipelines against messy real-world data and work with backend engineers to productionize what you build.

  • Use LLMs with the same rigor you'd apply anywhere: versioned prompts, real evals, tracked cost/latency, known failure modes.

  • Make extraction quality legible to non-ML colleagues, and treat de-identification, PHI handling, audit trails, and access controls as part of the modelling problem, not someone else's checklist.

  • Help shape the roadmap around the problems you see - a mission and a close working partner, not a backlog.

What we're looking for
  • Healthcare or life sciences experience with real clinical data - clinical notes, EHR data, claims, registries, or similar. This one is not negotiable for us.

  • 6+ years in applied ML, with models you personally took from problem statement to production and kept working - you know what degraded, how you found out, and what you did.

  • Depth in applied ML on text: information extraction, NER, classification, sequence labelling, weak supervision, and the evaluation practice around them, including annotation guidelines and inter-annotator agreement you've had to act on.

  • Practical, current experience with LLM-based approaches: prompt development, structured output, retrieval, fine-tuning where warranted, evals and observability for generative systems - enough to know where they help, and where they quietly don't.

  • Strong engineering fundamentals in Python. Your work runs in production, not only in a notebook.

  • Strong collaboration instincts across the ML boundary: you define problems with stakeholders before solving them, write clearly, and bring people along.

  • A track record of solving problems rather than closing tickets. Self-directed, comfortable without a playbook, and comfortable being wrong in public when the evidence says so.

Nice to have
  • Fluency with clinical terminologies and standards: SNOMED CT, ICD-10, LOINC, RxNorm, CPT, FHIR

  • Experience with HIPAA, SOC 2, de-identification methodology, or IRB and regulatory-grade data work

  • Experience as an early or first ML hire

  • Experience building or running human-in-the-loop annotation and QC operations at scale

  • OCR and document-understanding experience on low-quality real-world documents

  • Experience mentoring or leading ML engineers, or interest in growing that way

What this role is not
  • Not a research role. The bar is extraction quality in production, not publications.

  • Not an LLM-wrapper role. If your instinct is that every problem is a prompt away from being solved, we'll frustrate each other.

  • Not a large-team role yet. You'd be the first ML engineer in a small Platform Engineering function - breadth and influence, and fewer specialists to lean on.

  • Not a role where someone hands you a clean labelled dataset. Building it is the job.

Why this role is a good bet
  • Ground-floor ownership of a capability with direct commercial weight, with influence over architecture, roadmap, and eventually hiring.

  • Both halves of the modern ML toolkit in one seat, on a problem where the choice between them genuinely matters.

  • A manager who treats process and people work as legitimate engineering work, and intends for this seat to grow.

  • Health tech means the work has stakes - better extraction means higher quality research

Benefits & Perks
  • Equity in Novellia

  • Medical, dental, and vision coverage

  • 401(k)

  • Flexible time off

  • Wellness stipend

  • Up to 12 weeks of parental leave

Don't meet every requirement? Studies show women and people of color are less likely to apply unless they meet every qualification. If you're excited about this role but your experience doesn't align perfectly, we encourage you to apply anyway - you may be the right fit for this or another role.

U.S. Applicants Only

HQ

Novellia New York, New York, USA Office

New York, NY, United States

Novellia New York, New York, USA Office

New York, New York, United States, 10027

Similar Jobs

25 Days Ago
In-Office or Remote
226K-356K Annually
Senior level
226K-356K Annually
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Lead the technical direction for Atlassian’s Search Serving platform, building globally scalable, low-latency agentic search systems. Own architecture for retrieval, indexing, query processing, and ML inference while improving quality, reliability, scalability, operability, and cost. Establish production ML practices for evaluation, experimentation, observability, rollouts, and incident response. Optimize inference for embedding, retrieval, and ranking models, and lead cross-team initiatives while mentoring senior engineers and aligning technical leaders.
Top Skills: CachingDistillationDistributed SystemsEmbedding ModelsHybrid RetrievalInformation RetrievalMachine LearningMl InferenceModel OptimizationQuantizationRanking ModelsSearch InfrastructureSemantic SearchShardingVector Search
8 Days Ago
In-Office or Remote
New York, NY, USA
Senior level
Senior level
AdTech • Big Data • Marketing Tech • Mobile • Grocery & Supermarkets • Consumer Packaged Goods (CPG)
Build and improve machine learning systems for ad ranking, relevance, personalization, and optimization. Responsibilities include feature pipelines, model training and evaluation, active learning, LLM-assisted labeling and synthetic data generation, experimentation, and low-latency production inference. The role partners with product, data, and platform teams to improve advertiser performance and user engagement while maintaining reliability, latency, and data quality.
Top Skills: AWSAws BedrockChatgptClaudeGithub CopilotGoLangchainLlmsPythonVector Databases
18 Days Ago
Remote
US
160K-287K Annually
Senior level
160K-287K Annually
Senior level
Artificial Intelligence • Software
Build and operate scalable ML infrastructure and platform capabilities across cloud and on-premises environments. Develop Python-based developer tooling, services, and production systems supporting experimentation, distributed training, model deployment, and operations. Lead complex initiatives from architecture through rollout, improve reliability and developer productivity, and partner with ML and infrastructure teams. The role emphasizes Kubernetes, AWS, distributed compute, GPU-intensive workloads, observability, and production excellence.
Top Skills: AirflowAWSCloud ComputingComputer VisionDistributed SystemsKubeflowKubernetesMachine Learning InfrastructurePythonRayRoboticsSpark

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account