Automatically Score Tissue Images Like a Pathologist by Transfer Learning

Yan, Iris

doi:10.51387/23-NEJSDS53

Computer Science > Machine Learning

arXiv:2209.05954 (cs)

[Submitted on 9 Sep 2022 (v1), last revised 23 Nov 2023 (this version, v4)]

Title:Automatically Score Tissue Images Like a Pathologist by Transfer Learning

Authors:Iris Yan

View PDF

Abstract:Cancer is the second leading cause of death in the world. Diagnosing cancer early on can save many lives. Pathologists have to look at tissue microarray (TMA) images manually to identify tumors, which can be time-consuming, inconsistent and subjective. Existing automatic algorithms either have not achieved the accuracy level of a pathologist or require substantial human involvements. A major challenge is that TMA images with different shapes, sizes, and locations can have the same score. Learning staining patterns in TMA images requires a huge number of images, which are severely limited due to privacy and regulation concerns in medical organizations. TMA images from different cancer types may share certain common characteristics, but combining them directly harms the accuracy due to heterogeneity in their staining patterns. Transfer learning is an emerging learning paradigm that allows borrowing strength from similar problems. However, existing approaches typically require a large sample from similar learning problems, while TMA images of different cancer types are often available in small sample size and further existing algorithms are limited to transfer learning from one similar problem. We propose a new transfer learning algorithm that could learn from multiple related problems, where each problem has a small sample and can have a substantially different distribution from the original one. The proposed algorithm has made it possible to break the critical accuracy barrier (the 75% accuracy level of pathologists), with a reported accuracy of 75.9% on breast cancer TMA images from the Stanford Tissue Microarray Database. It is supported by recent developments in transfer learning theory and empirical evidence in clustering technology. This will allow pathologists to confidently adopt automatic algorithms in recognizing tumors consistently with a higher accuracy in real time.

Comments:	19 pages, 7 figures, The New England Journal of Statistics in Data Science, 2023
Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
Cite as:	arXiv:2209.05954 [cs.LG]
	(or arXiv:2209.05954v4 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2209.05954
Related DOI:	https://doi.org/10.51387/23-NEJSDS53

Submission history

From: Iris Yan [view email]
[v1] Fri, 9 Sep 2022 23:18:31 UTC (1,858 KB)
[v2] Mon, 27 Mar 2023 21:18:14 UTC (1,862 KB)
[v3] Wed, 8 Nov 2023 21:24:02 UTC (3,439 KB)
[v4] Thu, 23 Nov 2023 22:11:49 UTC (3,439 KB)

Computer Science > Machine Learning

Title:Automatically Score Tissue Images Like a Pathologist by Transfer Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Automatically Score Tissue Images Like a Pathologist by Transfer Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators