
Data Engineer
Laureate Institute for Brain Research
Job description
Research Data Engineer
Location:
Tulsa, OK (On-site)
About the Role
The Laureate Institute for Brain Research (LIBR) is seeking a Research Data Engineer to build and support the data infrastructure that powers innovative neuroscience and mental health research.
In this role, you'll work with investigators, clinicians, and data scientists to transform healthcare data into secure, high-quality research assets. You'll develop scalable data pipelines, build research data marts, optimize databases, and create analytics-ready datasets using Epic (Clarity, Caboodle, Cosmos) and other enterprise healthcare data platforms.
What You'll Do
-
Design, develop, and maintain ETL/ELT pipelines for clinical and research data.
-
Build and maintain research data marts that provide standardized, curated datasets for investigators, statisticians, and data scientists.
-
Develop, optimize, and troubleshoot complex SQL queries, stored procedures, views, and database objects to ensure efficient data retrieval and performance.
-
Extract, clean, transform, and integrate data from Epic (Clarity, Caboodle, Cosmos), or comparable EHR platforms.
-
Create analytics-ready datasets in Parquet and/or other modern data formats to support statistical analysis, machine learning, and AI applications.
-
Collaborate with investigators and research teams to translate scientific questions into technical data solutions.
-
Implement automated data validation and quality assurance processes to ensure data accuracy, consistency, and reproducibility.
-
Ensure data security, HIPAA compliance, and adherence to institutional and IRB requirements.
-
Document data pipelines, data models, and technical processes to support reproducibility and continuous improvement.
What We're Looking For
Required
-
Bachelor's degree in Computer Science, Data Engineering, Health Informatics, Data Science, or a related field.
-
Three to four years of experience in healthcare data engineering, clinical informatics, or a related role.
-
Advanced SQL skills, including query optimization and experience developing ETL/ELT pipelines.
-
Experience designing and supporting relational databases and research data marts.
-
Experience working with Epic (Clarity, Caboodle, Cosmos) or comparable enterprise healthcare data platforms or electronic health record (EHR) systems.
-
Knowledge of healthcare data standards and terminologies, such as ICD-10, CPT, SNOMED CT, LOINC, and RxNorm.
-
Understanding of HIPAA, clinical research data governance, and healthcare privacy regulations.
-
Strong analytical, communication, and problem-solving skills.
Preferred
-
Experience with Python and/or R for data engineering and automation.
-
Experience working with Parquet, Delta Lake, or other columnar data formats.
-
Experience with Git and version control best practices.
-
Experience supporting clinical research, neuroscience research, or academic healthcare organizations.
-
Epic certifications (Clarity, Caboodle, Cogito) or experience with comparable healthcare data platforms.
-
Experience with cloud-based data platforms (Azure, AWS, or Google Cloud) is a plus.