Skip to main content
Zobhira
Home
Jobs
Certifications
Zobhira
JobsCertificationsTodayAbout
Log inSign up
Zobhira

New job and contest openings, updated every morning on one searchable board.

Find work

  • All jobs
  • Fresher roles
  • Remote roles
  • Certifications

Compete

  • Added today

Popular cities

  • India
  • Bangalore, Karnataka, India
  • Hyderabad, Telangana, India
  • Pune, Maharashtra, India
  • Chennai, Tamil Nadu, India
  • Mumbai, Maharashtra, India

Company

  • About
  • Contact
  • Privacy
  • Terms

Stay updated

One email a week with new roles.

Secure infrastructure
Free to use, no account needed to search

Board updated daily · © 2026 Zobhira. All rights reserved.

Privacy PolicyTerms of Service
Home / Jobs / Veramed
Posted today · be early

Principal Data Engineer (RWE)

Veramed

IndiaremotePosted 1 day ago
Veramed logo

Skill Required

Data-EngineeringHealthcare-Data-EngineeringPrincipal-Data-EngineerETL-EngineeringData-EngineerLead-Data-EngineerSenior-Data-EngineeringSenior-Lead-Data-EngineeringData EngineerBusiness Analysisdata engineeringData StructurestroubleshootingData WarehousingR ProgrammingGenerative AIEngineeringDatabricksanalyticaldevelopingObservabilityPower BIdesigninganalyticsbuildingSparkTestNGdesignAzureCloudETLSQLandAIFulltime

Key highlights

  • Location: India
  • Key Requirement: Strong expertise in OMOP CDM v5.4, v6
  • Key Requirement: Expertise in Databricks, PySpark, and Spark SQL
  • Key Requirement: Experience with large-scale healthcare/patient-level datasets
  • Notable Fact: B Corp accredited company

Role overview

Veramed is a B Corp accredited company committed to creating a diverse environment and acting as an equal opportunities employer, using the power of business to build a more inclusive and sustainable economy. This India-based role focuses on the development of data processes for automated generation of patient-level data products to be used by business stakeholders for dashboards, reports, and studies.

Responsibilities

  • Development of data processes for the automated ongoing generation of patient level data (the “data product”) to be used by various business stakeholders for a variety of purposes (e.g. dashboards, reports, studies).
  • Downstream manipulation of datasets after their onboarding from data vendors/partners from “raw” format as provided into useable data structures that will be used to carry out RWE studies, dashboards & other data outputs
  • Transform heterogeneous raw healthcare datasets into reusable data models supporting observational research and epidemiology studies.
  • Occasional conversion of bespoke / one-off datasets (e.g. biomarkers, mutations) to OMOP format (including an understanding of what can and cannot be converted to OMOP format, e.g. to allow analysis to be carried out on residual data that cannot be converted to OMOP).
  • Build FAIR (Findable, Accessible, Interoperable, Reusable) data pipelines and semantic data engineering frameworks to improve discoverability and interoperability of healthcare data assets.
  • Create AI-ready datasets that can support generative AI use cases.
  • Technical engagement with key stakeholders (e.g. epidemiologist, statisticians, market access/health economists) from outside the RWE programming team to ensure a full and detailed understanding of end-user requirement is created and carefully documented. This includes scoping discussions, business analysis and translation of verbalised end-user needs into actionable data structures
  • Detailed technical engagement with colleagues from within the RWE programming team to build data structures required for the generation of RWE study outputs and data products; also support those team members in creating the study outputs where the data engineer’s skillset can add incremental value
  • Liaison & ongoing interaction with IT department to ensure that raw datasets inbound from data partners are fit for the agreed purposes
  • Liaise, where required, with technical staff employed by analysis software vendors (databricks etc)
  • Maintain clear documentation of data flows, schemas, pipelines, and processes to facilitate onboarding, troubleshooting and auditing.
  • Design and carry out detailed testing (data validation and monitoring) approaches for data structures built by self or other members of team to ensure the accuracy and reliability of the data within the data product
  • Troubleshoot any issues encountered with data loading, extraction and transformation (ETL)
  • Work in collaboration with three other members of the Data Engineering team, taking on workload from others as and when required

Requirements

  • Strong understanding of Real World Data (RWD) and Real World Evidence (RWE) concepts.
  • Ability to assess business requirements and recommend appropriate real-world healthcare datasets for analytical use cases.
  • Deep understanding of healthcare data models and healthcare data ecosystems.
  • Strong expertise in OMOP CDM v5.4 , v6, including extensions.
  • Knowledge of healthcare terminologies and standards such as: SNOMED CT, RxNorm, ICD-10, LOINC, HCPCS/CPT
  • Strong experience in building scalable ETL/ELT pipelines.
  • Expertise in: Databricks, PySpark, Spark SQL, SQL, Delta Lake
  • Experience working with large-scale healthcare and patient-level datasets.
  • Strong understanding of Semantic Data Engineering principles.
  • Experience building FAIR-compliant data pipelines.
  • Experience with cloud-based data platforms and distributed processing frameworks.
  • Strong Power BI development and data modelling skills.
  • Ability to create reusable analytical datasets for dashboards and studies.
  • Experience designing AI-ready datasets and analytics data products.
  • Experience implementing automated data quality frameworks.
  • Strong data profiling, validation, and monitoring skills.
  • Understanding of healthcare data quality assessment methodologies.
  • Excellent stakeholder management and communication skills.
  • Ability to translate complex business requirements into technical solutions.
  • Experience working with cross-functional global teams.
  • Exposure to one or more of the following therapeutic areas: Oncology, Respiratory, Immunology & Inflamation, Infectious Diseases

Nice to have

  • Working knowledge of R programming.
  • Experience with sparklyR.
  • Experience developing analytical applications using R Shiny.
  • Knowledge of common observational research methodologies.
  • Familiarity with OHDSI tools.
  • Exposure to Azure Data Platform services.

Additional details

  • India based role
  • Originally posted on Himalayas

Similar jobs open now

Bosch logo
hybrid
IN_Bosch Rexroth India_ Engineer / Executive_Application Engineering (Motion Control Systems)
Bosch
New Delhi, , India
View details
Persistent Systems logo
IT Software Asset Management (SAM) Engineer
Persistent Systems
India
View details
Yellow Slice logo
Sales Executive
Yellow Slice
Mumbai, Maharashtra, India
View details
C
Business Development Manager
CitrusCopy
Hyderabad, Telangana, India
View details
Apply now
LocationIndia
TypeFulltime
Posted9/24/2026
Apply by11/23/2026

Links are checked every day. If this one stops working, tell us and we'll pull it.

More like this

View all
Bosch logo
IN_Bosch Rexroth India_ Engineer / Executive_Application Engineering (Motion Control Systems)
Bosch
Persistent Systems logo
IT Software Asset Management (SAM) Engineer
Persistent Systems
Yellow Slice logo
Sales Executive
Yellow Slice
Apply now