Dandelion Health was founded in 2020 by experts in health tech, hospital systems, academia, and clinical AI. We are building the world’s largest AI training and clinical development platform. Today, we pride ourselves on our ability to make data access as easy as possible for AI developers, pharma, and medical devices, while raising the bar for patient safety and data quality. Tomorrow, we will be the place where any healthcare organization can go to build a responsible clinical AI product. Our culture is all about learning from data and improving, so we can help our clients improve health through AI. We partner with health systems to safely and ethically make their de-identified patient data available to AI developers. Currently, the data is acquired from Sharp HealthCare, Sanford Health, and Texas Health Resources – with two additional U.S. health systems joining soon. We have clinical data dating back to July 1, 2016. This data represents over 10 million patients and includes but is not limited to: Structured data (e.g., 100% of the EMR, including some claims), Unstructured text (e.g., clinical notes, radiology reports), Images (e.g., DICOM, pathology), Video, Waveforms, Continuous streaming monitoring data. You are an experienced analytics engineer who knows your way around clinical and electronic health record data. Your primary responsibility is to build, maintain, and optimize the transformation layer of our ELT pipeline, turning raw, messy clinical data into clean, well-modeled, analysis-ready datasets that our clients and internal teams can trust. You will own the full lifecycle of our data models: designing and implementing transformation logic, enforcing data quality and testing standards, optimizing pipeline performance, and documenting data models so they're discoverable and understandable across the organization. You will use your data expertise, programming abilities, and critical thinking skills to support our technical product team in scaling and growing our ability to provide meaningful data. Your team’s ultimate goal is to deliver the highest-quality data possible to our clients, who are building products that improve patient health. You will report to the Data Science Manager, under the Chief Data Officer. You are not afraid to dig into massive, confusing, disorganized new datasets and get them under control. You are excited to learn new environments, languages, and skills. This is a small, early stage company with enormous ambitions and everyone pitches in across the team.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level