Please use this identifier to cite or link to this item: http://repository.iiitd.edu.in/xmlui/handle/123456789/2166
Full metadata record
DC FieldValueLanguage
dc.contributor.authorSingh, Amanpreet-
dc.contributor.authorDaipuria, Aditya-
dc.contributor.authorSethi, Tavpritesh (Advisor)-
dc.date.accessioned2026-09-18T07:00:06Z-
dc.date.available2026-09-18T07:00:06Z-
dc.date.issued2024-07-25-
dc.identifier.urihttp://repository.iiitd.edu.in/xmlui/handle/123456789/2166-
dc.description.abstractThis project focuses on the SNOMED-CT mapping of over 9000 cancer-related datasets from Figshare and more than 2 million datasets from Zenodo FAIR Stations. The data was extracted from online sources using custom scripting, followed by the implementation of a multi-stage an alytical pipeline. This pipeline encompassed data cleaning, preprocessing, K-means clustering, word cloud generation, calculation of Jaccard and Kullback-Leibler divergence, and Bayesian Network modeling. Additionally, the BERT Sentence Model was utilized for calculating cosine scores to enhance the analysis. This comprehensive approach aimed to improve the accuracy and efficiency of SNOMED-CT mapping in cancer research, facilitating better data integration and interoperability within the medical and research communities. Our results demonstrate the effectiveness of these methods in handling large-scale datasets and providing valuable insights into cancer-related data.en_US
dc.language.isoen_USen_US
dc.publisherIIIT-Delhien_US
dc.subjectSNOMED-CT mappingen_US
dc.subjectCancer-related datasetsen_US
dc.subjectFigshareen_US
dc.subjectWord clouden_US
dc.subjectData Integrationen_US
dc.titleIFHP- tindering dataseten_US
dc.typeOtheren_US
Appears in Collections:Year-2024

Files in This Item:
File Description SizeFormat 
report (2) - Amanpreet Singh.pdf
  Restricted Access
892.67 kBAdobe PDFView/Open Request a copy


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.