3.Evaluation of Six Large Language Models for Clinical Decision Support: Application in Transfusion Decisionmaking for RhD Blood-type Patients
Jong Kwon LEE ; Sooin CHOI ; Sholhui PARK ; Sang-Hyun HWANG ; Duck CHO
Annals of Laboratory Medicine 2025;45(5):520-529
Background:
Large language models (LLMs) have the potential for clinical decision support; however, their use in specific tasks, such as determining the RhD blood type for transfusion, remains underexplored. Therefore, we evaluated the accuracy of six LLMs in addressing RhD blood type-related issues in Korean healthcare.
Methods:
Fifteen multiple-choice and true/false questions, based on real-world transfusion scenarios and reviewed by specialists, were developed. The questions were administered twice to six LLMs (Clova X, Gemini 1.0, Gemini 1.5, ChatGPT-3.5, GPT-4.0, and GPT-4o) in both Korean and English. Results were compared against the performance of 22 transfusion medicine experts. For particularly challenging questions, prompt engineering was applied, and the questions were reevaluated.
Results:
GPT-4o demonstrated the highest accuracy rate in Korean (0.6), with significant differences compared with those of Clova X and Gemini (P < 0.05). In English, the results were similar across all models. The transfusion experts achieved a higher accuracy rate (0.8). Among the five questions subjected to prompt engineering, only GPT-4o correctly responded to one, whereas the other models failed. All LLM models changed their responses or did not respond when the same question was repeated.
Conclusions
GPT-4o showed the best overall performance among the models tested and may be beneficial in RhD blood product transfusion decision-making. However, its performance suggests that it may serve best in a supportive role rather than as a primary decision-making tool.
7.Urinary mRNA biomarkers for the noninvasive diagnosis of calcineurin inhibitor toxicity in kidney transplant recipients with graft dysfunction: a retrospective study
Jihyun BAEK ; Jung-Woo SEO ; Hyeon Seok HWANG ; So-Young LEE ; Hye Yun JEONG ; Yang Gyun KIM ; Ju-Young MOON ; Jin Sug KIM ; Kyung-Hwan JEONG ; Byung Ha CHUNG ; Chan-Duck KIM ; Jae Berm PARK ; Yu Ho LEE ; Sang-Ho LEE
Clinical Transplantation and Research 2025;39(4):346-354
Background:
Calcineurin inhibitor (CNI) toxicity is a significant cause of graft dysfunction in kidney transplant recipients, yet distinguishing it from acute rejection (AR) and acute tubular necrosis (ATN) remains challenging. This study investigated the use of urinary mRNA biomarkers as a noninvasive tool for identifying CNI toxicity.
Methods:
We retrospectively enrolled 110 kidney transplant recipients and classified them into four groups based on pathological findings: stable graft function (n=35), CNI toxicity (n=25), AR (n=30), and ATN (n=20). Candidate biomarkers were selected using the GEO database. Urinary mRNA was extracted from cell pellets, reverse-transcribed, and quantified by real-time polymerase chain reaction.
Results:
Estimated glomerular filtration rates were comparable among the CNI toxicity, AR, and ATN groups. Four transcripts (LTF, NNMT, WFDC2, and HIF1A) were identified as candidate biomarkers. Urinary mRNA levels of LTF, NNMT, and HIF1A were significantly lower in the CNI toxicity group than the AR group. NNMT and HIF1A levels were also significantly lower than those observed in the ATN group. In contrast, WFDC2 levels did not differ significantly across groups. A three-gene signature (LTF, NNMT, and HIF1A) effectively differentiated CNI toxicity from AR and ATN (area under the curve [AUC], 0.867;95% confidence interval [CI], 0.787–0.947) and significantly enhanced the diagnostic performance of the clinical variable-based model (AUC increased from 0.776; 95% CI, 0.660–0.892, to 0.934; 95% CI, 0.881–0.986).
Conclusions
Urinary mRNA levels of LTF, NNMT, and HIF1A may serve as useful biomarkers for identifying CNI toxicity in kidney transplant recipients with graft dysfunction.
8.Comparative analysis of different surgical approaches for recurrent inguinal hernia: a single-center observational study
Mi Jeong CHOI ; Kang-Seok LEE ; Heung-Kwon OH ; Sang-Hoon AHN ; Hong-min AHN ; Hye-Rim SHIN ; Tae-Gyun LEE ; Min Hyeong JO ; Duck-Woo KIM ; Sung-Bum KANG
Annals of Surgical Treatment and Research 2024;106(6):330-336
Purpose:
Managing recurrent inguinal hernias is complex, and choosing the right surgical approach (laparoscopic vs. open) is vital for patient outcomes. This study compared the outcomes of using the same vs. different surgical approaches for initial and subsequent hernia repairs.
Methods:
We retrospectively analyzed patients who underwent recurrent inguinal hernia repair at Seoul National University Bundang Hospital between January 2014 and May 2023. Patients were divided into the “concordant” and “discordant” groups, comprising patients who underwent same and different approaches in both surgeries, respectively. Preoperative baseline characteristics, index surgery data, postoperative outcomes, and recurrence rates were analyzed and compared.
Results:
In total, 131 patients were enrolled; the concordant and discordant groups comprised 31 (open, n = 19; laparoscopic, n = 12) and 100 patients (open to laparoscopic, n = 68; laparoscopic to open, n = 32), respectively. No significant differences were observed in the mean operation time (50.5 ± 21.7 minutes vs. 50.2 ± 20.0 minutes, P = 0.979), complication rates (6.5% vs. 14.0%, P = 0.356), or 36-month cumulative recurrence rates (9.8% vs. 9.8%; P = 0.865). The mean postoperative hospital stay was significantly shorter in the discordant than in the concordant group (1.8 ± 0.7 vs. 1.4 ± 0.6, P = 0.003).
Conclusion
Most recurrent inguinal hernia repairs were performed using the discordant surgical approach. Overall, concordance in the surgical approach did not significantly affect postoperative outcomes. Therefore, the selection of the surgical approach based on the patient’s condition and surgeon’s preference may be advisable.
9.Severe Adverse Events of Periocular Acupuncture: A Review of Cases
Sang-Mok LEE ; Jun WU ; Daniel Duck-Jin HWANG
Korean Journal of Ophthalmology 2023;37(3):255-265
Acupuncture is recognized as a component of alternative medicine and is increasingly used worldwide. Many studies have shown the various effects of acupuncture around the eyes for ophthalmologic or nonophthalmologic conditions. For ophthalmologic conditions, the effect of acupuncture on dry eye syndrome, glaucoma, myopia, amblyopia, ophthalmoplegia, allergic rhinoconjunctivitis, blepharospasm, and blepharoptosis has been reported. Recently, several studies on dry eye syndrome have been reported and are in the spotlight. However, given the variety of study designs and reported outcomes of periocular acupuncture, research is still inconclusive, and further studies are required. In addition, although a systematic and reliable safety assessment is required, to the best of our knowledge, there have been no reports of a literature review of ocular complications resulting from periocular acupuncture. This review collected cases of ocular injury as severe adverse events from previously published case reports of periocular acupuncture. A total of 14 case reports (15 eyes of 14 patients) of adverse events published between 1982 and 2020 were identified. This review article provides a summary of the reported cases and suggestions for the prevention and management of better visual function prognosis.
10.How many times should we repeat measurements of the ultrasound-guided attenuation parameter for evaluating hepatic steatosis?
Duck Min SEO ; Sang Min LEE ; Ji Won PARK ; Min-Jeong KIM ; Hong Il HA ; Sun-Young PARK ; Kwanseop LEE
Ultrasonography 2023;42(2):227-237
Purpose:
This retrospective study aimed to determine the number of times the ultrasound-guided attenuation parameter (UGAP) should be measured during the evaluation of hepatic steatosis.
Methods:
Patients with suspected nonalcoholic fatty liver disease who underwent two UGAP repetition protocols (six-repetition [UGAP_6] and 12-repetition [UGAP_12]) and measurement of the controlled attenuation parameter (CAP) using transient elastography between October 2020 and June 2021 were enrolled. The mean attenuation coefficient (AC), interquartile range (IQR)/median, and coefficient of variance (CV) of the two repetition protocols were compared using the paired t test. Moreover, the diagnostic performances of UGAP_6 and UGAP_12 were compared using the area under the receiver operating characteristic (AUROC) curve, considering the CAP value as a reference standard.
Results:
The study included 160 patients (100 men; mean age, 50.9 years). There were no significant differences between UGAP_6 and UGAP_12 (0.731±0.116 dB/cm/MHz vs. 0.734±0.113 dB/cm/MHz, P=0.156) and mean CV (7.6±0.3% vs. 8.0±0.3%, P=0.062). However, the mean IQR/median of UGAP_6 was significantly lower than that of UGAP_12 (8.9%±6.0% vs. 9.8%±5.2%, P=0.012). In diagnosing the hepatic steatosis stage, UGAP_6 and UGAP_12 yielded comparable AUROCs (≥S1, 0.908 vs. 0.897, P=0.466; ≥S2, 0.883 vs. 0.897, P=0.126; S3, 0.832 vs. 0.834, P=0.799).
Conclusion
UGAP had high diagnostic performance in diagnosing hepatic steatosis, regardless of the number of repetitions (six repetitions vs. 12 repetitions), with maintained reliability. Therefore, six UGAP measurements seem sufficient for evaluating hepatic steatosis using UGAP.

Result Analysis
Print
Save
E-mail