. 2019 Nov 26;7(4):e14502.

doi: 10.2196/14502.

Using a Large Margin Context-Aware Convolutional Neural Network to Automatically Extract Disease-Disease Association from Literature: Comparative Analytic Study

Po-Ting Lai¹, Wei-Liang Lu², Ting-Rung Kuo², Chia-Ru Chung², Jen-Chieh Han², Richard Tzong-Han Tsai^#², Jorng-Tzong Horng^#^{2

3}

Affiliations

¹ Department of Computer Science National Tsing Hua University, Hsinchu, Province of China Taiwan.
² Department of Computer Science & Information Engineering, National Central University, Taoyuan, Province of China Taiwan.
³ Department of Bioinformatics and Medical Engineering, Asia University, Taichung, Province of China Taiwan.

^# Contributed equally.

PMID: 31769759
PMCID: PMC6913619
DOI: 10.2196/14502

Using a Large Margin Context-Aware Convolutional Neural Network to Automatically Extract Disease-Disease Association from Literature: Comparative Analytic Study

Po-Ting Lai et al. JMIR Med Inform. 2019.

. 2019 Nov 26;7(4):e14502.

doi: 10.2196/14502.

Authors

Po-Ting Lai¹, Wei-Liang Lu², Ting-Rung Kuo², Chia-Ru Chung², Jen-Chieh Han², Richard Tzong-Han Tsai^#², Jorng-Tzong Horng^#^{2

3}

Affiliations

¹ Department of Computer Science National Tsing Hua University, Hsinchu, Province of China Taiwan.
² Department of Computer Science & Information Engineering, National Central University, Taoyuan, Province of China Taiwan.
³ Department of Bioinformatics and Medical Engineering, Asia University, Taichung, Province of China Taiwan.

^# Contributed equally.

PMID: 31769759
PMCID: PMC6913619
DOI: 10.2196/14502

Abstract

Background: Research on disease-disease association (DDA), like comorbidity and complication, provides important insights into disease treatment and drug discovery, and a large body of the literature has been published in the field. However, using current search tools, it is not easy for researchers to retrieve information on the latest DDA findings. First, comorbidity and complication keywords pull up large numbers of PubMed studies. Second, disease is not highlighted in search results. Finally, DDA is not identified, as currently no disease-disease association extraction (DDAE) dataset or tools are available.

Objective: As there are no available DDAE datasets or tools, this study aimed to develop (1) a DDAE dataset and (2) a neural network model for extracting DDA from the literature.

Methods: In this study, we formulated DDAE as a supervised machine learning classification problem. To develop the system, we first built a DDAE dataset. We then employed two machine learning models, support vector machine and convolutional neural network, to extract DDA. Furthermore, we evaluated the effect of using the output layer as features of the support vector machine-based model. Finally, we implemented large margin context-aware convolutional neural network architecture to integrate context features and convolutional neural networks through the large margin function.

Results: Our DDAE dataset consisted of 521 PubMed abstracts. Experiment results showed that the support vector machine-based approach achieved an F1 measure of 80.32%, which is higher than the convolutional neural network-based approach (73.32%). Using the output layer of convolutional neural network as a feature for the support vector machine does not further improve the performance of support vector machine. However, our large margin context-aware-convolutional neural network achieved the highest F1 measure of 84.18% and demonstrated that combining the hinge loss function of support vector machine with a convolutional neural network into a single neural network architecture outperforms other approaches.

Conclusions: To facilitate the development of text-mining research for DDAE, we developed the first publicly available DDAE dataset consisting of disease mentions, Medical Subject Heading IDs, and relation annotations. We developed different conventional machine learning models and neural network architectures and evaluated their effects on our DDAE dataset. To further improve DDAE performance, we propose an large margin context-aware-convolutional neural network model for DDAE that outperforms other approaches.

Keywords: biological relation extraction; biomedical natural language processing; convolutional neural networks; deep learning; disease-disease association.

©Po-Ting Lai, Wei-Liang Lu, Ting-Rung Kuo, Chia-Ru Chung, Jen-Chieh Han, Richard Tzong-Han Tsai, Jorng-Tzong Horng. Originally published in JMIR Medical Informatics (http://medinform.jmir.org), 26.11.2019.

PubMed Disclaimer

Conflict of interest statement

Conflicts of Interest: None declared.

Figures

**Figure 1**
Disease-disease association extraction examples.

**Figure 2**
Disease-disease association extraction dataset construction process. MeSH= Mesdical Subject Headings.

**Figure 3**
Large margin context-aware convolutional neural network (LC-CNN) architecture. BOW: Bag of words; POS: Part of speech; NE: Named Entity.

**Figure 4**
Precision and recall formula.

See this image and copyright information in PMC

Cited by

Surveying biomedical relation extraction: a critical examination of current datasets and the proposal of a new resource.
Huang MS, Han JC, Lin PY, You YT, Tsai RT, Hsu WL. Huang MS, et al. Brief Bioinform. 2024 Mar 27;25(3):bbae132. doi: 10.1093/bib/bbae132. Brief Bioinform. 2024. PMID: 38609331 Free PMC article. Review.
Supervised Relation Extraction Between Suicide-Related Entities and Drugs: Development and Usability Study of an Annotated PubMed Corpus.
Karapetian K, Jeon SM, Kwon JW, Suh YK. Karapetian K, et al. J Med Internet Res. 2023 Mar 8;25:e41100. doi: 10.2196/41100. J Med Internet Res. 2023. PMID: 36884281 Free PMC article.
A study on large-scale disease causality discovery from biomedical literature.
Yu S, Dong P, Li J, Tang X, Li X. Yu S, et al. BMC Med Inform Decis Mak. 2025 Mar 18;25(1):136. doi: 10.1186/s12911-025-02893-0. BMC Med Inform Decis Mak. 2025. PMID: 40102814 Free PMC article.
Increasing Women's Knowledge about HPV Using BERT Text Summarization: An Online Randomized Study.
Bitar H, Babour A, Nafa F, Alzamzami O, Alismail S. Bitar H, et al. Int J Environ Res Public Health. 2022 Jul 1;19(13):8100. doi: 10.3390/ijerph19138100. Int J Environ Res Public Health. 2022. PMID: 35805761 Free PMC article. Clinical Trial.

References

1. Žitnik M, Janjić V, Larminie C, Zupan B, Pržulj N. Discovering disease-disease associations by fusing systems-level molecular data. Sci Rep. 2013 Nov 15;3:3202. doi: 10.1038/srep03202. doi: 10.1038/srep03202. - DOI - DOI - PMC - PubMed
1. Liu C, Tseng Y, Li W, Wu C, Mayzus I, Rzhetsky A, Sun F, Waterman M, Chen JJ, Chaudhary PM, Loscalzo J, Crandall E, Zhou XJ. DiseaseConnect: a comprehensive web server for mechanism-based disease-disease connections. Nucleic Acids Res. 2014 Jul;42(Web Server issue):W137–46. doi: 10.1093/nar/gku412. http://europepmc.org/abstract/MED/24895436 - DOI - PMC - PubMed
1. Bang S, Kim J, Shin H. Causality modeling for directed disease network. Bioinformatics. 2016 Sep 1;32(17):i437–44. doi: 10.1093/bioinformatics/btw439. - DOI - PubMed
1. Sun K, Gonçalves JP, Larminie C, Przulj N. Predicting disease associations via biological network analysis. BMC Bioinformatics. 2014 Sep 17;15:304. doi: 10.1186/1471-2105-15-304. https://bmcbioinformatics.biomedcentral.com/articles/10.1186/1471-2105-1... - DOI - DOI - PMC - PubMed
1. Yang J, Wu SJ, Yang SY, Peng JW, Wang SN, Wang FY, Song Y, Qi T, Li Y, Li Y. DNetDB: The human disease network database based on dysfunctional regulation mechanism. BMC Syst Biol. 2016 May 21;10(1):36. doi: 10.1186/s12918-016-0280-5. https://bmcsystbiol.biomedcentral.com/articles/10.1186/s12918-016-0280-5 - DOI - DOI - PMC - PubMed

LinkOut - more resources

Full Text Sources
Research Materials
- NCI CPTC Antibody Characterization Program

Save citation to file

Email citation

Add to Collections

Add to My Bibliography

Your saved search

Create a file for external citation management software

Your RSS Feed

Using a Large Margin Context-Aware Convolutional Neural Network to Automatically Extract Disease-Disease Association from Literature: Comparative Analytic Study

Affiliations

Using a Large Margin Context-Aware Convolutional Neural Network to Automatically Extract Disease-Disease Association from Literature: Comparative Analytic Study

Authors

Affiliations

Abstract

Conflict of interest statement

Figures

Similar articles

Cited by

References

LinkOut - more resources

Full Text Sources

Research Materials