Our Team
Dandelion Health was founded in 2020 by experts in health tech, hospital systems, academia, and clinical AI. We are building the world’s largest AI training and clinical development platform. Today, we pride ourselves on our ability to make data access as easy as possible for AI developers, pharma, and medical devices, while raising the bar for patient safety and data quality. Tomorrow, we will be the place where any healthcare organization can go to build a responsible clinical AI product. Our culture is all about learning from data and improving, so we can help our clients improve health through AI. Meet the rest of our team here.
Our Data
We partner with health systems to safely and ethically make their de-identified patient data available to AI developers. Currently, the data is acquired from Sharp HealthCare, Sanford Health, and Texas Health Resources – with two additional U.S. health systems joining soon.
We have clinical data dating back to July 1, 2016. This data represents over 10 million patients and includes but is not limited to:
Structured data (e.g., 100% of the EMR, including some claims)
Unstructured text (e.g., clinical notes, radiology reports)
Images (e.g., DICOM, pathology)
Video
Waveforms
Continuous streaming monitoring data
Your Role
You are a healthcare data scientist who knows your way around clinical and electronic health record data. Your primary responsibility is to partner collaboratively with our clients to develop data science solutions and analyze AI-ready datasets to provide key insights using multimodal data. You will curate datasets by identifying patient subpopulations or disease cohorts, and pool multimodal data to drive rapid exploratory AI/ML and Real-World Evidence analyses, model experimentation, and/or model validation. You will lead analytics research from inception to closeout by having ownership over study design, data curation strategies, effort estimation, analytic design, and delivery of final results. You will use your data expertise, programming abilities, and critical thinking skills to support our clients by designing and delivering solutions to help them tackle a broad range of business challenges. Your team’s ultimate goal is to deliver the highest-quality data possible to our clients, who are building products that improve patient health, and demonstrate the value of these data. You will report to the Research Data Science Manager.
Responsibilities
Your day-to-day responsibilities will include the following:
Collaborate with clients to develop creative analytic solutions that create value and address challenges across critical areas of their business.
Design statistical analysis plans for evidence generation projects that include methodology, results structure, expected outputs and dataset requirements;
Query complex source systems in a range of health data sources (e.g., EMRs, ECG data, DICOM data) to identify and map data elements in order to create high-quality datasets for analytic and AI use cases;
Develop descriptive analyses and in-depth predictive and causal inference models to drive evidence generation on Dandelion data and enhance existing data products;
Summarize the methods and results from projects into clear explanations and documentation for internal and external audiences including for submission to conferences and peer-reviewed publications;
Support all phases of SQL/analytical programming, data management, quality control, and reporting for analytics projects;
Develop code and documentation to deliver high-quality and HIPAA-compliant data products on time to internal and external customers;
Create summarized findings and recommendations that are clearly presented and adapted for audiences that have a varying range of technical and clinical experience;
Identify and resolve problems using your knowledge, background, and troubleshooting skills;
Ensure accuracy, data integrity, and validity of data and analysis in all work;
Present to senior leadership as well external audiences
You are not afraid to dig into massive, confusing, disorganized new datasets and get them under control. You are excited to learn new environments, languages, and skills. This is a small, early stage company with enormous ambitions and everyone pitches in across the team.
Qualifications
M.S. or PhD degree in a relevant field, such as Biomedical Informatics, Data Science, Biostatistics, or B.S. with at least seven years of experience
Experience in client interaction, crafting project proposals, and leading teams.
Background in statistical modeling and real-world data analysis
Fluency in Python and SQL
1+ years experience with extracting, curating, and analyzing data created within the HIT and healthcare delivery ecosystem (e.g., EMR, claims, registry); this may include knowledge of the roles of data exchange and content standards (e.g., FHIR, CDA, CQL) and clinical terminology standards (e.g., ICD, CPT, LOINC, SNOMED-CT, NDC, RxNorm)
Strong technical writing, editing, and communication skills along with a collaborative, client-first mindset
Excellent organizational skills with an ability to embrace change and effectively manage multiple projects and consistently plan work to meet deadlines
Experience working in or with startups is a plus
Technology Experiences and Skills
We don’t expect anyone to have all of the following skills or experiences, but we do seek candidates who are interested in growing their skill sets and working with healthcare data in all its glorious complexity. The Data Team works closely with our Engineering Team to put our work into production and meet client needs.
SQL
Python and/or R
Git and version control
Familiarity with encryption methods and writing regular expressions
Prior experience querying EDWs or databases and creating reports or analytics for healthcare data
Familiarity with the data aspects of electronic medical records, ex. Epic, Cerner, Allscripts
Prior experience working with insurance claims data
Familiarity with medical terminologies or controlled vocabularies such as ICD-10, SNOMED-CT, LOINC, CPT/HCPCS, NDC, and RxNorm
Any medical ontology experience
Any NLP experience
Any experience working with DICOM or other imaging modalities
Experience with OMOP common data model
Familiarity with Machine Learning concepts
Experience with AWS
Note that familiarity with machine learning model development and deployment is not required for this position, but familiarity with high-level ML concepts is a plus.
Nature of our work
Our work is fast paced and iterative. We are growing, and we want to support our team members to grow in their skills as well. We are building a team that approaches problems with a diversity of perspectives, values experimentation, and refining our approach based on that experimentation. We work with the full spectrum of healthcare data from tabular data, videos, images, waveforms, etc. If a health system collects it, we might work with it!
If this looks like a partial fit, please reach out, we would love to share more about the work we do for you to understand if it would be a good fit for you.
There is occasional travel for in-person company working days on roughly a quarterly basis.
Team Benefits
Remote work and flexible hours. Availability needed for meetings, which we try to keep to a healthy minimum
Complete wellness benefits including healthcare, dental, vision, PTO, sick days and more. Ask for details
Professional development days to build your skills
Collegial work environment
Academic bent towards inquiry and problem solving but start-up speed and flexibility
Great balance of focus time to work on projects but easy to access team members to discuss issues and work collaboratively
Dandelion is a mission-driven company that is focused on improving patient care
Similar Jobs
What you need to know about the Boston Tech Scene
Key Facts About Boston Tech
- Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
- Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
- Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
- Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

