Group Communication With Context Codec for Lightweight Source Separation

被引:19
|
作者
Luo, Yi [1 ]
Han, Cong [1 ]
Mesgarani, Nima [1 ]
机构
[1] Columbia Univ, Dept Elect Engn, New York, NY 10027 USA
关键词
Codecs; Context modeling; Decoding; Complexity theory; Pipelines; Neural networks; Convolutional codes; Source separation; deep learning; lightweight; group communication; context codec; SPEECH ENHANCEMENT; PATH RNN;
D O I
10.1109/TASLP.2021.3078640
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
Despite the recent progress on neural network architectures for speech separation, the balance between the model size, model complexity and model performance is still an important and challenging problem for the deployment of such models to low-resource platforms. In this paper, we propose two simple modules, group communication and context codec, that can be easily applied to a wide range of architectures to jointly decrease the model size and complexity without sacrificing the performance. A group communication module splits a high-dimensional feature into groups of low-dimensional features and captures the inter-group dependency. A separation module with a significantly smaller model size can then be shared by all the groups. A context codec module, containing a context encoder and a context decoder, is designed as a learnable downsampling and upsampling module to decrease the length of a sequential feature processed by the separation module. The combination of the group communication and the context codec modules is referred to as the GC3 design. Experimental results show that applying GC3 on multiple network architectures for speech separation can achieve on-par or better performance with as small as 2.5% model size and 17.6% model complexity, respectively.
引用
收藏
页码:1752 / 1761
页数:10
相关论文
共 50 条
  • [31] Blind source separation for communication signals using antenna arrays
    Feng, M
    Kammeyer, KD
    ICUPC '98 - IEEE 1998 INTERNATIONAL CONFERENCE ON UNIVERSAL PERSONAL COMMUNICATIONS, VOLS 1 AND 2, 1998, 1-2 : 665 - 669
  • [32] Single Channel Blind Source Separation for MISO Communication Systems
    Dey, Priyanka
    Trivedi, Nikita
    Satija, Udit
    Ramkumar, Barathram
    Manikandan, M. Sabarimalai
    2017 IEEE 86TH VEHICULAR TECHNOLOGY CONFERENCE (VTC-FALL), 2017,
  • [33] Enabling source channel separation for communication networks : The uplink case
    Sridharan, Sriram
    Wu, Wei
    Vishwanath, Sriram
    MILCOM 2006, VOLS 1-7, 2006, : 3663 - +
  • [34] Adaptive blind Source Separation in underwater wireless speech communication
    Wang, Zhenhai
    Chen, C. H.
    OCEANS 2007 - EUROPE, VOLS 1-3, 2007, : 1491 - 1496
  • [35] Blind Source Separation for Satellite Communication Anti-jamming
    Yang, Hua
    Zhang, Hang
    Zhang, Jiang
    Yang, Liu
    Wang, Pengfei
    WIRELESS AND SATELLITE SYSTEMS, PT I, 2019, 280 : 717 - 726
  • [36] Compromising Anonymous Communication Systems Using Blind Source Separation
    Zhu, Ye
    Bettati, Riccardo
    ACM TRANSACTIONS ON INFORMATION AND SYSTEM SECURITY, 2009, 13 (01)
  • [37] MIMO measurements of communication signals and application of blind source separation
    Rinas, J
    Kammeyer, KD
    PROCEEDINGS OF THE 3RD IEEE INTERNATIONAL SYMPOSIUM ON SIGNAL PROCESSING AND INFORMATION TECHNOLOGY, 2003, : 94 - 97
  • [38] HIGH SPEED PCM CODEC FOR SATELLITE COMMUNICATION
    OHASHI, Y
    KISHIGAMI, M
    SHIGAKI, S
    TANAKA, S
    OKADA, T
    NEC RESEARCH & DEVELOPMENT, 1970, (18): : 20 - +
  • [39] Extensible hierarchical codec semantic communication system
    Zhang Y.
    Zhao H.
    Wei J.
    Cao K.
    Zhang Y.
    Luo P.
    Liu Y.
    Mei K.
    Tongxin Xuebao/Journal on Communications, 2023, 44 (08): : 49 - 60
  • [40] Lightweight and secure D2D group communication for wireless IoT
    Miao, Junfeng
    Wang, Zhaoshun
    Xue, Xingsi
    Wang, Mei
    Lv, Jianhui
    Li, Min
    FRONTIERS IN PHYSICS, 2023, 11