Repository logo
  • English
  • 中文
  • Log In
    New user? Click here to register.Have you forgotten your password?
Repository logo
    Communities & Collections
    Research Outputs
    Fundings & Projects
    People
    Organizations
    Statistics
  • English
  • 中文
  • Log In
    New user? Click here to register.Have you forgotten your password?
  1. Home
  2. .TMU Publications / 北醫出版品(教師升等著作 / 教學實踐 / 學位論文)
  3. .博碩士學位論文
  4. 112學年度
  5. 使用生成式人工智慧從自由文本病理報告中提取結構化訊息
 
  • Details
Options

使用生成式人工智慧從自由文本病理報告中提取結構化訊息

Other Title
Using Generative AI to Extract Structured Information from Free Text Pathology Reports
Type
thesis
Date Issued
2024-06-17
Author(s)
FAHAD SHAHID
Advisor
許明暉
Subjects
系所名稱:大數據科技及管理研究所碩士班
Publisher
大數據科技及管理研究所碩士班
Description
學位別:碩士
關鍵字:生成式預訓練 變壓器 (GPT); 自由文本病理報告; 資訊擷取; 大型語言模型 (LLM); ChatGPT; 乳癌; 人工智慧(AI)
論文公開日期:2024-07-15
Abstract
ABSTRACT
Background: Traditionally, organizing pathology reports has been a manual task. This approach is labor-intensive, costly, and often leads to inaccuracies, complicating the analysis of data for medical research and delaying critical insights. While artificial intelligence (AI) offers potential solutions, many systems struggle to effectively handle the complexity of pathological texts. Numerous studies and articles have underscored the urgent need to automate the process of extracting information and structuring that information from free text pathology reports.

Objective: This study explores the use of generative AI to automate the extraction of information and structuring that information from 33 breast cancer free text pathology reports obtained from Taipei Medical University Hospital. The aim is to accelerate the information extraction process and enhance the reliability of this information for research and diagnostic purposes. Additionally, the study seeks if generative AI can help to improve the accuracy and efficiency to structure information from free text pathology reports.

Methods: For this study, a Streamlit web application was developed, leveraging its capabilities to efficiently create robust generative AI applications. This application was seamlessly integrated with the ChatGPT Large Language Model provided by OpenAI, tasked with the extraction and structuring of information from free-text pathology reports. The data, once extracted, is methodically organized and compiled into a downloadable Excel file for subsequent analysis. Furthermore, the application is designed to display the results on its web interface, offering immediate validation features. This allows users to promptly verify and assess the accuracy of the information processed. The systematic and meticulous approach adopted ensures the highest standards of data integrity and operational efficiency, essential for the reliability of research outcomes.

Results: The implementation of the Streamlit web application, integrated with the ChatGPT Large Language Model, successfully automated the extraction and structuring of information from 33 breast cancer free-text pathology reports. The system showed a higher percentage of accuracy, by achieving an extraction and structuring accuracy rate of 99.61%. This result confirms the effectiveness of generative AI in handling free text pathological reports.

Conclusion: The reliability of the system was further underscored by its ability to significantly reduce the time required to structure and analyze free-text pathological reports compared to traditional manual methods. This advancement highlights the potential of generative AI to transform the processing of free-text pathological reports, enhancing both efficiency and accuracy, and offering substantial improvements in research and clinical diagnostics.
URI
https://203.71.86.71/handle/123456789/9955

Copyright Notice

● The digital content on this platform is part of the Taipei Medical University Institutional Repository, featuring various academic works and outputs from the institution. It offers free access to academic research and public education for non-commercial use.

● Please use the content appropriately and within legal boundaries to respect copyright owners' rights. For commercial use, please obtain prior authorization from the copyright owner. Users must not use TMUIR for any illegal purposes.

● By utilising the platform, users are deemed to have fully accepted and understood all the regulations set out in this statement, relevant laws of the Republic of China, all international internet regulations, and usage conventions.

● TMUIR is committed to protecting the interests of copyright owners. If you believe that any material on this website infringes copyright, please contact our staff at libirtmu@gmail.com, and we will remove the work from the repository.

Built with DSpace-CRIS software - Extension maintained and optimized by 4Science

  • Cookie settings
  • Privacy policy
  • End User Agreement
  • Send Feedback