2023
A hierarchical strategy to minimize privacy risk when linking “De-identified” data in biomedical research consortia
Ohno-Machado L, Jiang X, Kuo T, Tao S, Chen L, Ram P, Zhang G, Xu H. A hierarchical strategy to minimize privacy risk when linking “De-identified” data in biomedical research consortia. Journal Of Biomedical Informatics 2023, 139: 104322. PMID: 36806328, PMCID: PMC10975485, DOI: 10.1016/j.jbi.2023.104322.Peer-Reviewed Original ResearchConceptsPrivacy of individualsAppropriate privacy protectionData-driven modelsPrivacy protectionPrivacy risksData Coordination CenterData hubData repositoryHierarchical strategyPrivacyBiomedical discoveryData setsRecord linkageData Coordinating CenterRepositoryComplex strategiesCoordination centerTechnologyTechniqueDataPartiesSetHierarchy
2019
Recognizing software names in biomedical literature using machine learning
Wei Q, Zhang Y, Amith M, Lin R, Lapeyrolerie J, Tao C, Xu H. Recognizing software names in biomedical literature using machine learning. Health Informatics Journal 2019, 26: 21-33. PMID: 31566474, PMCID: PMC7334865, DOI: 10.1177/1460458219869490.Peer-Reviewed Original ResearchConceptsSoftware namesF-measureNatural language processing methodsBiomedical literatureWord representation featuresLanguage processing methodsEntity recognition systemSoftware catalogSoftware repositoriesFeature engineeringBiomedical softwareRecognition systemSoftware toolsBiomedical domainRepresentation featuresMEDLINE abstractsWord embeddingsKnowledge featuresManual curationSoftwareMachineProcessing methodsBest systemRepositorySystem
2017
DATS, the data tag suite to enable discoverability of datasets
Sansone S, Gonzalez-Beltran A, Rocca-Serra P, Alter G, Grethe J, Xu H, Fore I, Lyle J, Gururaj A, Chen X, Kim H, Zong N, Li Y, Liu R, Ozyurt I, Ohno-Machado L. DATS, the data tag suite to enable discoverability of datasets. Scientific Data 2017, 4: 170059. PMID: 28585923, PMCID: PMC5460592, DOI: 10.1038/sdata.2017.59.Peer-Reviewed Original ResearchFinding useful data across multiple biomedical data repositories using DataMed
Ohno-Machado L, Sansone S, Alter G, Fore I, Grethe J, Xu H, Gonzalez-Beltran A, Rocca-Serra P, Gururaj A, Bell E, Soysal E, Zong N, Kim H. Finding useful data across multiple biomedical data repositories using DataMed. Nature Genetics 2017, 49: 816-819. PMID: 28546571, PMCID: PMC6460922, DOI: 10.1038/ng.3864.Peer-Reviewed Original ResearchConceptsBiomedical data repositoriesHealth big dataData setsKnowledge discoveryBig dataMultiple repositoriesSearch enginesData indexFAIR principlesDataMedData repositoryService providersKnowledge initiativesKnowledge expertsBiomedical research communityResearch communityRepositoryScience landscapeUseful dataInteroperabilityMetadataFindabilitySetEngineDataInformation retrieval for biomedical datasets: the 2016 bioCADDIE dataset retrieval challenge
Roberts K, Gururaj A, Chen X, Pournejati S, Hersh W, Demner-Fushman D, Ohno-Machado L, Cohen T, Xu H. Information retrieval for biomedical datasets: the 2016 bioCADDIE dataset retrieval challenge. Database 2017, 2017: bax068. DOI: 10.1093/database/bax068.Peer-Reviewed Original ResearchBiomedical datasetsRetrieval challengesInformation retrieval techniquesAdvanced query processingBiomedical data repositoriesAdvanced retrieval methodsQuery processingInformation retrievalTest queriesRetrieval systemRank frameworkRetrieval approachRetrieval techniquesData repositoryRetrieval methodTop precisionDatasetQueriesRepositoryChallengesRetrievalTaskLearningSystemCorpus
2014
Identifying plausible adverse drug reactions using knowledge extracted from the literature
Shang N, Xu H, Rindflesch T, Cohen T. Identifying plausible adverse drug reactions using knowledge extracted from the literature. Journal Of Biomedical Informatics 2014, 52: 293-310. PMID: 25046831, PMCID: PMC4261011, DOI: 10.1016/j.jbi.2014.07.011.Peer-Reviewed Original ResearchConceptsPredication-based Semantic IndexingReflective Random IndexingLBD methodsNatural language processing toolsBiomedical literatureDrug-adverse event associationsLanguage processing toolsSemantic indexingElectronic health recordsRandom IndexingHuman reviewVast repositoryDiscovery methodsVolume of knowledgeProcessing toolsEvaluation setHealth recordsData sourcesEvent associationsIndexingDrug-effect relationshipsRepositoryLarge volumesADR associationsReasoning pathwaysChapter 12 Linking Genomic and Clinical Data for Discovery and Personalized Care
Denny J, Xu H. Chapter 12 Linking Genomic and Clinical Data for Discovery and Personalized Care. 2014, 395-424. DOI: 10.1016/b978-0-12-401678-1.00012-9.Peer-Reviewed Original ResearchElectronic health recordsEHR dataNatural language processingSuch algorithmsLanguage processingDecision supportPhenotype algorithmsIdeal repositoryHealth recordsNumber of challengesRepositoryAlgorithmClinical notesClinical careClinical documentationGenomic dataResult dataAccurate caseDNA biobanksEarly demonstration projectsHealth care qualityClinical recordsMedication recordsClinical dataTool