Skip to main navigation Skip to search Skip to main content

Probing the Dual Logic Ability of Privatized Medical-Domain LLMs

  • Yanrui Du
  • , Sendong Zhao*
  • , Muzhen Cai
  • , Ming Ma
  • , Danyang Zhao
  • , Jiawei Cao
  • , Bing Qin
  • *Corresponding author for this work
  • Harbin Institute of Technology

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Large Language Models (LLMs) are gaining widespread attention for their potential across various fields, particularly within the medical domain. Recent efforts have aimed to privatize general-domain LLMs into specialized medical-domain LLMs by feeding high-quality, medical-domain training data. However, an overlooked aspect in these privatization efforts is the Dual Logic Ability of LLMs. This ability enables LLMs to comprehend pairs of logically opposed questions, ensuring stance consistency in their responses. In our study, we investigate two primary questions: Q1. How does privatization affect the dual logic ability of LLMs? Q2. How can we maintain the robustness of LLMs' dual logic ability after privatization? To explore these questions, we first constructed a medical-domain dual logic ability evaluation dataset comprising logically opposed question pairs, created manually by NLP experts. By examining the stance consistency in responses to logically opposed question pairs, our analysis demonstrates a significant decline in the dual logic ability of LLMs after privatization. Furthermore, we construct privatization data to investigate the effects of the pre-training and instruction fine-tuning stages on the dual logic ability of LLMs. Interestingly, our findings reveal that the instruction fine-tuning stage often inadvertently compromises the LLMs' dual logic ability although it is not the trainers' intention. To counteract this, we incorporated general-domain dual logic data derived from basic science during the instruction fine-tuning stage, which are automatically constructed by our designed pipeline. Experiment results show that privatized LLMs can generalize dual logic ability from general-domain dual logic data, leading to their enhanced performance in the medical domain. Our study underscores the importance of prioritizing LLMs' dual logic ability during the privatization process and establishes a benchmark for future research.

Original languageEnglish
Title of host publicationProceedings - 2024 IEEE International Conference on Bioinformatics and Biomedicine, BIBM 2024
EditorsMario Cannataro, Huiru Zheng, Lin Gao, Jianlin Cheng, Joao Luis de Miranda, Ester Zumpano, Xiaohua Hu, Young-Rae Cho, Taesung Park
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages3182-3187
Number of pages6
ISBN (Electronic)9798350386226
DOIs
StatePublished - 2024
Event2024 IEEE International Conference on Bioinformatics and Biomedicine, BIBM 2024 - Lisbon, Portugal
Duration: 3 Dec 20246 Dec 2024

Publication series

NameProceedings - 2024 IEEE International Conference on Bioinformatics and Biomedicine, BIBM 2024

Conference

Conference2024 IEEE International Conference on Bioinformatics and Biomedicine, BIBM 2024
Country/TerritoryPortugal
CityLisbon
Period3/12/246/12/24

Keywords

  • dual logic ability
  • privatized medical-domain LLMs

Fingerprint

Dive into the research topics of 'Probing the Dual Logic Ability of Privatized Medical-Domain LLMs'. Together they form a unique fingerprint.

Cite this