Deep learning generates custom-made logistic regression models for explaining how breast cancer subtypes are classified

Shibahara, Takuma; Wada, Chisa; Yamashita, Yasuho; Fujita, Kazuhiro; Sato, Masamichi; Kuwata, Junichi; Okamoto, Atsushi; Ono, Yoshimasa

doi:10.1371/journal.pone.0286072

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2001

Computer Science > Machine Learning

Title: Deep learning generates custom-made logistic regression models for explaining how breast cancer subtypes are classified

Authors: Takuma Shibahara, Chisa Wada, Yasuho Yamashita, Kazuhiro Fujita, Masamichi Sato, Junichi Kuwata, Atsushi Okamoto, Yoshimasa Ono

(Submitted on 20 Jan 2020 (v1), last revised 19 Jul 2022 (this version, v2))

Abstract: Differentiating the intrinsic subtypes of breast cancer is crucial for deciding the best treatment strategy. Deep learning can predict the subtypes from genetic information more accurately than conventional statistical methods, but to date, deep learning has not been directly utilized to examine which genes are associated with which subtypes. To clarify the mechanisms embedded in the intrinsic subtypes, we developed an explainable deep learning model called a point-wise linear (PWL) model that generates a custom-made logistic regression for each patient. Logistic regression, which is familiar to both physicians and medical informatics researchers, allows us to analyze the importance of the feature variables, and the PWL model harnesses these practical abilities of logistic regression. In this study, we show that analyzing breast cancer subtypes is clinically beneficial for patients and one of the best ways to validate the capability of the PWL model. First, we trained the PWL model with RNA-seq data to predict PAM50 intrinsic subtypes and applied it to the 41/50 genes of PAM50 through the subtype prediction task. Second, we developed a deep enrichment analysis method to reveal the relationships between the PAM50 subtypes and the copy numbers of breast cancer. Our findings showed that the PWL model utilized genes relevant to the cell cycle-related pathways. These preliminary successes in breast cancer subtype analysis demonstrate the potential of our analysis strategy to clarify the mechanisms underlying breast cancer and improve overall clinical outcomes.

Comments:	25 pages, 5 figures
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
DOI:	10.1371/journal.pone.0286072
Cite as:	arXiv:2001.06988 [cs.LG]
	(or arXiv:2001.06988v2 [cs.LG] for this version)

Submission history

From: Yasuho Yamashita [view email]
[v1] Mon, 20 Jan 2020 05:56:32 GMT (3043kb,D)
[v2] Tue, 19 Jul 2022 03:12:56 GMT (3036kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2001.06988

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Deep learning generates custom-made logistic regression models for explaining how breast cancer subtypes are classified

Submission history