Data preprocessing scripts and standardized prompts used in the study: Enabling interoperability across disparate health data sources using Large Language Models
This ZIP file contains: (1) the exact standardized prompt submitted to all three LLMs across all 36 runs (3 LLMs × 4 schemas × 3 repetitions), (2) the complete set of source-to-target mapping pairs used for preprocessing and standardizing CDM element names across PCORnet, OMOP, Sentinel, and i2b2, and (3) the complete mapping results from the three LLMs for all four Common Data Models (PCORnet, OMOP, Sentinel, i2b2) to FHIR across iterations.
Research output: Contribution to journal › Article › peer-review
Open Access
File
Cite this
DataSetCite
Ozonze, O. N. (Creator), Dike, U. (Creator), Okonor, O. M. (Contributor), Adedeji, T. A. (Creator) (5 Aug 2026). Data preprocessing scripts and standardized prompts used in the study: Enabling interoperability across disparate health data sources using Large Language Models. Public Library of Science. journal(1.zip). 10.1371/journal.pdig.0001511.s001