Novellia Logo

Novellia

Senior Machine Learning Engineer, Platform

Posted 3 Days Ago
Remote
Hiring Remotely in USA
150K-200K Annually
Senior level
Remote
Hiring Remotely in USA
150K-200K Annually
Senior level
Own the full lifecycle of clinical text extraction models, from problem framing and annotation strategy through production deployment, evaluation, monitoring, and retraining. Build clinical NLP pipelines, establish defensible quality and error-analysis systems, and apply LLMs using rigorous evaluation, observability, and cost controls. Partner with clinical data and backend engineering teams while addressing PHI protection, de-identification, auditability, and access controls. As the first ML hire, set technical direction and potentially build and lead the future ML team.
The summary above was generated by AI
Our Story
 

Since 2023, our mission has been clear: to be the north star of patient equity.

 

Every day, we strive to bridge the gaps in healthcare access and outcomes to ensure that every patient, regardless of background or circumstance, receives the care they deserve. As a member of our team, you'll be at the forefront of innovation, working alongside passionate individuals who share your dedication to creating best-in-class, patient-centric products in healthcare. Together, we're revolutionizing the way people understand their health and working with the world's top researchers to accelerate innovation.

About Novellia

Novellia is the first and only company that lets anyone in the U.S. gain access to nearly a decade of their health data in under 30 seconds — 100% free. All your health records, across every doctor, in one place, always up to date.

We are the only patient-powered real-world data platform delivering comprehensive, patient-authorized longitudinal health insights to accelerate biopharma innovation. Unlike traditional RWD providers who deliver fragmented institutional data, we empower patients to access 20+ years of their health records, then transform these complete health journeys into fit-for-purpose datasets for evidence generation, regulatory submissions, and market access. We are growing 5x year over year, have raised close to $30M in funding, and are backed by tier-1 investors including Spark Capital, Khosla Ventures, and Bling Capital.

Working with the world's top researchers, we turn health insights into life-changing action for millions of people around the world.

About the role

Most of what matters in a health record isn't in a structured field - it's in the note, the discharge summary, the pathology report, the scanned fax. Turning that unstructured clinical text into trustworthy, structured features is what makes a longitudinal health history usable for research, and it's one of the highest-leverage capabilities Novellia can own.

You'll be our first ML hire, joining Platform Engineering and reporting to the Head of Platform Engineering, as technical owner of this multi-quarter effort. The interesting decisions are still open - what we extract first, how we know we're right, what a mature extraction pipeline looks like at our scale. There's no existing approach to inherit or defend.

The work draws on two toolkits. Roughly 70% is applied ML on clinical text: entity extraction, classification, sequence labelling, annotation strategy, error analysis, calibration, and the evaluation discipline that tells you whether your numbers mean anything. Roughly 30% is LLM-based: prompt development, structured output, retrieval, and the evals and observability that keep generative approaches honest. Deciding which approach a given problem calls for is the most interesting part of the job, and that call is yours.

We're looking for a leader in this seat: setting technical direction rather than waiting to be handed a problem. If this grows the way we think it will, leading the team we build around it is on the table.

What you'll do
  • Own the full lifecycle of extraction models - framing, data/annotation strategy, model selection, training/fine-tuning, evaluation, deployment, monitoring, retraining. Not a research seat, not a hand-off seat.

  • Define what "accurate enough" means with clinical and customer-facing stakeholders, and build the evaluation harness that makes the answer defensible - the first deliverable, not a follow-up.

  • Partner with Clinical Data Managers on curation design and own the technical half of QA/QC alongside them: which variables are extractable, how an instruction becomes a model spec, and the tooling/sampling/error analysis behind human-in-the-loop review.

  • Build clinical NLP pipelines against messy real-world data and work with backend engineers to productionize what you build.

  • Use LLMs with the same rigor you'd apply anywhere: versioned prompts, real evals, tracked cost/latency, known failure modes.

  • Make extraction quality legible to non-ML colleagues, and treat de-identification, PHI handling, audit trails, and access controls as part of the modelling problem, not someone else's checklist.

  • Help shape the roadmap around the problems you see - a mission and a close working partner, not a backlog.

What we're looking for
  • Healthcare or life sciences experience with real clinical data - clinical notes, EHR data, claims, registries, or similar. This one is not negotiable for us.

  • 6+ years in applied ML, with models you personally took from problem statement to production and kept working - you know what degraded, how you found out, and what you did.

  • Depth in applied ML on text: information extraction, NER, classification, sequence labelling, weak supervision, and the evaluation practice around them, including annotation guidelines and inter-annotator agreement you've had to act on.

  • Practical, current experience with LLM-based approaches: prompt development, structured output, retrieval, fine-tuning where warranted, evals and observability for generative systems - enough to know where they help, and where they quietly don't.

  • Strong engineering fundamentals in Python. Your work runs in production, not only in a notebook.

  • Strong collaboration instincts across the ML boundary: you define problems with stakeholders before solving them, write clearly, and bring people along.

  • A track record of solving problems rather than closing tickets. Self-directed, comfortable without a playbook, and comfortable being wrong in public when the evidence says so.

Nice to have
  • Fluency with clinical terminologies and standards: SNOMED CT, ICD-10, LOINC, RxNorm, CPT, FHIR

  • Experience with HIPAA, SOC 2, de-identification methodology, or IRB and regulatory-grade data work

  • Experience as an early or first ML hire

  • Experience building or running human-in-the-loop annotation and QC operations at scale

  • OCR and document-understanding experience on low-quality real-world documents

  • Experience mentoring or leading ML engineers, or interest in growing that way

What this role is not
  • Not a research role. The bar is extraction quality in production, not publications.

  • Not an LLM-wrapper role. If your instinct is that every problem is a prompt away from being solved, we'll frustrate each other.

  • Not a large-team role yet. You'd be the first ML engineer in a small Platform Engineering function - breadth and influence, and fewer specialists to lean on.

  • Not a role where someone hands you a clean labelled dataset. Building it is the job.

Why this role is a good bet
  • Ground-floor ownership of a capability with direct commercial weight, with influence over architecture, roadmap, and eventually hiring.

  • Both halves of the modern ML toolkit in one seat, on a problem where the choice between them genuinely matters.

  • A manager who treats process and people work as legitimate engineering work, and intends for this seat to grow.

  • Health tech means the work has stakes - better extraction means higher quality research

Benefits & Perks
  • Equity in Novellia

  • Medical, dental, and vision coverage

  • 401(k)

  • Flexible time off

  • Wellness stipend

  • Up to 12 weeks of parental leave

Don't meet every requirement? Studies show women and people of color are less likely to apply unless they meet every qualification. If you're excited about this role but your experience doesn't align perfectly, we encourage you to apply anyway - you may be the right fit for this or another role.

U.S. Applicants Only

Similar Jobs

10 Days Ago
Remote or Hybrid
171K-261K Annually
Senior level
171K-261K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Lead development and operation of CI platform components (RBE, FUSE) to enable scalable build, test, and developer workflows. Collaborate across teams, own projects end-to-end, drive engineering best practices, mentor engineers, perform design/code reviews, and support integration of internal tools with the FUSE-based file system to improve developer productivity for AV and ML workloads.
Top Skills: FuseGoLinuxNetworkingPythonRemote Build Execution (Rbe)Ssh
9 Days Ago
In-Office or Remote
295K-406K Annually
Senior level
295K-406K Annually
Senior level
eCommerce • Fintech • Machine Learning • Retail
Own the technical vision and architecture of Faire’s machine learning platform, spanning training, inference, feature management, governance, MLOps, experimentation, and observability. Lead adoption of Databricks, Unity Catalog, Spark, Delta Lake, and MLflow; optimize platform scalability, reliability, and cost; advise data science and engineering teams; influence cross-functional architecture; mentor ML engineers; and integrate emerging AI productivity tools.
Top Skills: AirflowAmazon S3Amazon SagemakerApache IcebergAWSChatgpt 5.2Claude Sonnet 4.5CockroachdbDatabricksDatadogDelta LakeDockerFivetranGithub ActionsKafkaKotlinKubernetesMlflowMySQLPysparkPythonPyTorchSnowflakeSparkSQLTerraformUnity Catalog
4 Days Ago
Remote
United States
143K-197K Annually
Senior level
143K-197K Annually
Senior level
Healthtech • HR Tech
Build and maintain scalable ML and generative AI platforms and backend infrastructure; deploy and manage production applications; create services exposing ML/AI products; implement RAG solutions and vector databases; integrate LLM orchestration and foundation model APIs; collaborate cross-functionally and mentor teammates.
Top Skills: AnthropicAws BedrockDockerFastapiKubernetesLangchainLlamaindexLow-Latency DatabasesOpenaiPythonRagRelational DatabasesServer-Sent EventsVector DatabasesWebsockets

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account