MITAS: A Compressed Time-Domain Audio Separation Network with Parameter Sharing

Tuan, Chao-I; Wu, Yuan-Kuei; Lee, Hung-yi; Tsao, Yu

Computer Science > Sound

arXiv:1912.03884 (cs)

[Submitted on 9 Dec 2019]

Title:MITAS: A Compressed Time-Domain Audio Separation Network with Parameter Sharing

Authors:Chao-I Tuan, Yuan-Kuei Wu, Hung-yi Lee, Yu Tsao

View PDF

Abstract:Deep learning methods have brought substantial advancements in speech separation (SS). Nevertheless, it remains challenging to deploy deep-learning-based models on edge devices. Thus, identifying an effective way to compress these large models without hurting SS performance has become an important research topic. Recently, TasNet and Conv-TasNet have been proposed. They achieved state-of-the-art results on several standardized SS tasks. Moreover, their low latency natures make them definitely suitable for real-time on-device applications. In this study, we propose two parameter-sharing schemes to lower the memory consumption on TasNet and Conv-TasNet. Accordingly, we derive a novel so-called MiTAS (Mini TasNet). Our experimental results first confirmed the robustness of our MiTAS on two types of perturbations in mixed audio. We also designed a series of ablation experiments to analyze the relation between SS performance and the amount of parameters in the model. The results show that MiTAS is able to reduce the model size by a factor of four while maintaining comparable SS performance with improved stability as compared to TasNet and Conv-TasNet. This suggests that MiTAS is more suitable for real-time low latency applications.

Subjects:	Sound (cs.SD); Computation and Language (cs.CL); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:1912.03884 [cs.SD]
	(or arXiv:1912.03884v1 [cs.SD] for this version)
	https://doi.org/10.48550/arXiv.1912.03884

Submission history

From: Chao-I Tuan [view email]
[v1] Mon, 9 Dec 2019 07:44:32 UTC (2,814 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.SD

< prev | next >

new | recent | 2019-12

Change to browse by:

cs
cs.CL
cs.LG
eess
eess.AS

References & Citations

DBLP - CS Bibliography

listing | bibtex

Chao-I Tuan
Hung-yi Lee
Yu Tsao

export BibTeX citation

Computer Science > Sound

Title:MITAS: A Compressed Time-Domain Audio Separation Network with Parameter Sharing

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Sound

Title:MITAS: A Compressed Time-Domain Audio Separation Network with Parameter Sharing

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators