Fast and Scalable Distributed Deep Convolutional Autoencoder for fMRI Big Data Analytics

Makkie, Milad; Huang, Heng; Zhao, Yu; Vasilakos, Athanasios V.; Liu, Tianming

Computer Science > Distributed, Parallel, and Cluster Computing

arXiv:1710.08961v3 (cs)

[Submitted on 24 Oct 2017 (v1), last revised 4 Mar 2018 (this version, v3)]

Title:Fast and Scalable Distributed Deep Convolutional Autoencoder for fMRI Big Data Analytics

Authors:Milad Makkie, Heng Huang, Yu Zhao, Athanasios V. Vasilakos, Tianming Liu

View PDF

Abstract:In recent years, analyzing task-based fMRI (tfMRI) data has become an essential tool for understanding brain function and networks. However, due to the sheer size of tfMRI data, its intrinsic complex structure, and lack of ground truth of underlying neural activities, modeling tfMRI data is hard and challenging. Previously proposed data-modeling methods including Independent Component Analysis (ICA) and Sparse Dictionary Learning only provided a weakly established model based on blind source separation under the strong assumption that original fMRI signals could be linearly decomposed into time series components with corresponding spatial maps. Meanwhile, analyzing and learning a large amount of tfMRI data from a variety of subjects has been shown to be very demanding but yet challenging even with technological advances in computational hardware. Given the Convolutional Neural Network (CNN), a robust method for learning high-level abstractions from low-level data such as tfMRI time series, in this work we propose a fast and scalable novel framework for distributed deep Convolutional Autoencoder model. This model aims to both learn the complex hierarchical structure of the tfMRI data and to leverage the processing power of multiple GPUs in a distributed fashion. To implement such a model, we have created an enhanced processing pipeline on the top of Apache Spark and Tensorflow library, leveraging from a very large cluster of GPU machines. Experimental data from applying the model on the Human Connectome Project (HCP) show that the proposed model is efficient and scalable toward tfMRI big data analytics, thus enabling data-driven extraction of hierarchical neuroscientific information from massive fMRI big data in the future.

Comments:	This work is submitted to SIGKDD 2018
Subjects:	Distributed, Parallel, and Cluster Computing (cs.DC); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE); Neurons and Cognition (q-bio.NC); Machine Learning (stat.ML)
Cite as:	arXiv:1710.08961 [cs.DC]
	(or arXiv:1710.08961v3 [cs.DC] for this version)
	https://doi.org/10.48550/arXiv.1710.08961

Submission history

From: Milad Makkie [view email]
[v1] Tue, 24 Oct 2017 19:35:51 UTC (1,324 KB)
[v2] Wed, 13 Dec 2017 07:46:58 UTC (1,208 KB)
[v3] Sun, 4 Mar 2018 21:31:55 UTC (1,880 KB)

Computer Science > Distributed, Parallel, and Cluster Computing

Title:Fast and Scalable Distributed Deep Convolutional Autoencoder for fMRI Big Data Analytics

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Distributed, Parallel, and Cluster Computing

Title:Fast and Scalable Distributed Deep Convolutional Autoencoder for fMRI Big Data Analytics

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators