← Back to list
Job

Research Consultant Data Scientist

Data Scientist • Remote • Full-time European Union EU/EMEA
Our Team Dandelion Health was founded in 2020 by experts in health tech , hospital systems , academia , and clinical AI . We are building the world’s largest AI training and clinical development platform. Today, we pride ourselves on our ability to make data access as easy as possible for AI developers, pharma, and medical devices, while raising the bar for patient safety and data quality. Tomorrow, we will be the place where any healthcare organization can go to build a responsible clinical AI product. Our culture is all about learning from data and improving, so we can help our clients improve health through AI. Meet the rest of our team here . Our Data We partner with health systems to safely and ethically make their de-identified patient data available to AI developers. Currently, the data is acquired from Sharp HealthCare, Sanford Health, and Texas Health Resources – with two additional U.S. health systems joining soon. We have clinical data dating back to July 1, 2016. This data represents over 10 million patients and includes but is not limited to: Structured data (e.g., 100% of the EMR, including some claims) Unstructured text (e.g., clinical notes, radiology reports) Images (e.g., DICOM, pathology) Video Waveforms Continuous streaming monitoring data Your Role You are a healthcare data scientist who knows your way around clinical and electronic health record data. Your primary responsibility is to partner collaboratively with our clients to develop data science solutions and analyze AI-ready datasets to provide key insights using multimodal data. You will curate datasets by identifying patient subpopulations or disease cohorts, and pool multimodal data to drive rapid exploratory AI/ML and Real-World Evidence analyses, model experimentation, and/or model validation. You will lead analytics research from inception to closeout by having ownership over study design, data curation strategies, effort estimation, analytic design, and delivery of final results. You will use your data expertise, programming abilities, and critical thinking skills to support our clients by designing and delivering solutions to help them tackle a broad range of business challenges. Your team’s ultimate goal is to deliver the highest-quality data possible to our clients, who are building products that improve patient health, and demonstrate the value of these data. You will report to the Research Data Science Manager. Responsibilities Your day-to-day responsibilities will include the following: Collaborate with clients to develop creative analytic solutions that create value and address challenges across critical areas of their business. Design statistical analysis plans for evidence generation projects that include methodology, results structure, expected outputs and dataset requirements; Query complex source systems in a range of health data sources (e.g., EMRs, ECG data, DICOM data) to identify and map data elements in order to create high-quality datasets for analytic and AI use cases; Develop descriptive analyses and in-depth predictive and causal inference models to drive evidence generation on Dandelion data and enhance existing data products; Summarize the methods and results from projects into clear explanations and documentation for internal and external audiences including for submission to conferences and peer-reviewed publications; Support all phases of SQL/analytical programming, data management, quality control, and reporting for analytics projects; Develop code and documentation to deliver high-quality and HIPAA-compliant data products on time to internal and external customers; Create summarized findings and recommendations that are clearly presented and adapted for audiences that have a varying range of technical and clinical experience; Identify and resolve problems using your knowledge, background, and troubleshooting skills; Ensure accuracy, data integrity, and validity of data and analysis in all work; Present to senior leadership as well external audiences You are not afraid to dig into massive, confusing, disorganized new datasets and get them under control. You are excited to learn new environments, languages, and skills. This is a small, early stage company with enormous ambitions and everyone pitches in across the team. Qualifications M.S. or PhD degree in a relevant field, such as Biomedical Informatics, Data Science, Biostatistics, or B.S. with at least seven years of experience Experience in client interaction, crafting project proposals, and leading teams. Background in statistical modeling and real-world data analysis Fluency in Python and SQL 1+ years experience with extracting, curating, and analyzing data created within the HIT and healthcare delivery ecosystem (e.g., EMR, claims, registry); this may include knowledge of the roles of data exchange and content standards (e.g., FHIR, CDA, CQL) and clinical terminology standards (e.g., ICD, CPT, LOINC, SNOMED-CT, NDC, RxNorm) Strong technical writing, editing, and communication skills along with a collaborative, client-first mindset Excellent organizational skills with an ability to embrace change and effectively manage multiple projects and consistently plan work to meet deadlines Experience working in or with startups is a plus Technology Experiences and Skills We don’t expect anyone to have all of the following skills or experiences, but we do seek candidates who are interested in growing their skill sets and working with healthcare data in all its glorious complexity. The Data Team works closely with our Engineering Team to put our work into production and meet client needs. SQL Python and/or R Git and version control Familiarity with encryption methods and writing regular expressions Prior experience querying EDWs or databases and creating reports or analytics for healthcare data Familiarity with the data aspects of electronic medical records, ex. Epic, Cerner, Allscripts Prior experience working with insurance claims data Familiarity with medical terminologies or controlled vocabularies such as ICD-10, SNOMED-CT, LOINC, CPT/HCPCS, NDC, and RxNorm Any medical ontology experience Any NLP experience Any experience working with DICOM or other imaging modalities Experience with OMOP common data model Familiarity with Machine Learning concepts Experience with AWS Note that familiarity with machine learning model development and deployment is not required for this position, but familiarity with high-level ML concepts is a plus. Nature of our work Our work is fast paced and iterative. We are growing, and we want to support our team members to grow in their skills as well. We are building a team that approaches problems with a diversity of perspectives, values experimentation, and refining our approach based on that experimentation. We work with the full spectrum of healthcare data from tabular data, videos, images, waveforms, etc. If a health system collects it, we might work with it! If this looks like a partial fit, please reach out, we would love to share more about the work we do for you to understand if it would be a good fit for you. There is occasional travel for in-person company working days on roughly a quarterly basis. Team Benefits Remote work and flexible hours. Availability needed for meetings, which we try to keep to a healthy minimum Complete wellness benefits including healthcare, dental, vision, PTO, sick days and more. Ask for details Professional development days to build your skills Collegial work environment Academic bent towards inquiry and problem solving but start-up speed and flexibility Great balance of focus time to work on projects but easy to access team members to discuss issues and work collaboratively Dandelion is a mission-driven company that is focused on improving patient care

Similar jobs