For Employers

Veramed

Principal Data Engineer (RWE)

Posted an hour ago
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Questions interviewers often ask for this role, with sample answers.

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

The Principal Data Engineer will develop automated data processes and transform heterogeneous healthcare datasets into reusable models for RWE studies. They will also collaborate with cross-functional teams to build FAIR-compliant data pipelines and create AI-ready datasets.

1. Development of data processes for the automated ongoing generation of patient level data (the “data product”) to be used by various business stakeholders for a variety of purposes (e.g. dashboards, reports, studies).

2. Downstream manipulation of datasets after their onboarding from data vendors/partners from “raw” format as provided into useable data structures that will be used to carry out RWE studies, dashboards & other data outputs

3. Transform heterogeneous raw healthcare datasets into reusable data models supporting observational research and epidemiology studies.

4. Occasional conversion of bespoke / one-off datasets (e.g. biomarkers, mutations) to OMOP format (including an understanding of what can and cannot be converted to OMOP format, e.g. to allow analysis to be carried out on residual data that cannot be converted to OMOP).

5. Build FAIR (Findable, Accessible, Interoperable, Reusable) data pipelines and semantic data engineering frameworks to improve discoverability and interoperability of healthcare data assets.

6. Create AI-ready datasets that can support generative AI use cases.

Communication

7. Technical engagement with key stakeholders (e.g. epidemiologist, statisticians, market access/health economists) from outside the RWE programming team to ensure a full and detailed understanding of end-user requirement is created and carefully documented. This includes scoping discussions, business analysis and translation of verbalised end-user needs into actionable data structures

8. Detailed technical engagement with colleagues from within the RWE programming team to build data structures required for the generation of RWE study outputs and data products; also support those team members in creating the study outputs where the data engineer’s skillset can add incremental value

9. Liaison & ongoing interaction with IT department to ensure that raw datasets inbound from data partners are fit for the agreed purposes (as per bullet point 1 above)

10. Liaise, where required, with technical staff employed by analysis software vendors (databricks etc)

Documentation

11. Maintain clear documentation of data flows, schemas, pipelines, and processes to facilitate onboarding, troubleshooting and auditing.

Quality, Validation & Support

12. Design and carry out detailed testing (data validation and monitoring) approaches for data structures built by self or other members of team to ensure the accuracy and reliability of the data within the data product

13. Troubleshoot any issues encountered with data loading, extraction and transformation (ETL)

10. Work in collaboration with three other members of the Data Engineering team, taking on workload from others as and when required

 

Required Skills & Qualifications

Domain Expertise

· Strong understanding of Real World Data (RWD) and Real World Evidence (RWE) concepts.

· Ability to assess business requirements and recommend appropriate real-world healthcare datasets for analytical use cases.

· Deep understanding of healthcare data models and healthcare data ecosystems.

· Strong expertise in OMOP CDM v5.4 , v6, including extensions.

· Knowledge of healthcare terminologies and standards such as:

o SNOMED CT

o RxNorm

o ICD-10

o LOINC

o HCPCS/CPT

Data Engineering

· Strong experience in building scalable ETL/ELT pipelines.

· Expertise in:

o Databricks

o PySpark

o Spark SQL

o SQL

o Delta Lake

· Experience working with large-scale healthcare and patient-level datasets.

· Strong understanding of Semantic Data Engineering principles.

· Experience building FAIR-compliant data pipelines.

· Experience with cloud-based data platforms and distributed processing frameworks.

Analytics & Visualization

· Strong Power BI development and data modelling skills.

· Ability to create reusable analytical datasets for dashboards and studies.

· Experience designing AI-ready datasets and analytics data products.

Validation & Quality

· Experience implementing automated data quality frameworks.

· Strong data profiling, validation, and monitoring skills.

· Understanding of healthcare data quality assessment methodologies.

Collaboration & Communication

· Excellent stakeholder management and communication skills.

· Ability to translate complex business requirements into technical solutions.

· Experience working with cross-functional global teams.

 

Disease Area Knowledge

Exposure to one or more of the following therapeutic areas:

· Oncology

· Respiratory

· Immunology & Inflamation

· Infectious Diseases

 

Nice-to-Have Skills

· Working knowledge of R programming.

· Experience with sparklyR.

· Experience developing analytical applications using R Shiny.

· Knowledge of common observational research methodologies.

· Familiarity with OHDSI tools.

· Exposure to Azure Data Platform services.

Veramed is a B Corp accredited company which means that we use the power of business to build a more inclusive and sustainable economy meeting the highest verified standards of social and environmental performance, transparency, and accountability.

As an organisation that has people at the heart of it, Veramed is committed to creating a diverse environment and is proud to be an equal opportunities employer. We foster a working culture where employees have integrity, honesty and respect for one another without regard to race, national origin, religion, gender identity or expression, sexual orientation or disability. All qualified applicants will receive equal consideration for employment.

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Senior Full-Stack Product Engineer | Node.js, AI-Native (R&D)

Full Time Ukraine Software Development

PowerBuilder Developer

Full Time India Software Development

Technical Leader (Fullstack)

Full Time Worldwide Software Development

Python with ADB Developer

Full Time India Software Development

Accounts Administrator (AP and AR)

Full Time Philippines Software Development

Senior Database Administrator

Full Time Canada, China, Malaysia +1 more Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 213,038+ Jobs in Data Engineer

Answer easy questions

Answer easy questions

213,038+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified