Exploring Data-Independent Dimensionality Reduction in Sparse Representation-Based Speaker Identification

被引：2

作者：

Haris, B. C. ^{[1
]}

Sinha, Rohit ^{[1
]}

机构：

[1] Indian Inst Technol, Dept Elect & Elect Engn, Gauhati 781039, India

来源：

CIRCUITS SYSTEMS AND SIGNAL PROCESSING | 2014年 / 33卷 / 08期

关键词：

Sparse representation classification; Random projections; Speaker recognition; Supervectors; Dimensionality reduction; VERIFICATION; RECOGNITION; ALGORITHM;

D O I：

10.1007/s00034-014-9757-x

中图分类号：

TM [电工技术]; TN [电子技术、通信技术];

学科分类号：

0808 ; 0809 ;

摘要：

The sparse representation classification (SRC) has attracted the attention of many signal processing domains in past few years. Recently, it has been successfully explored for the speaker recognition task with Gaussian mixture model (GMM) mean supervectors which are typically of the order of tens of thousands as speaker representations. As a result of this, the complexity of such systems become very high. With the use of the state-of-the-art i-vector representations, the dimension of GMM mean supervectors can be reduced effectively. But the i-vector approach involves a high dimensional data projection matrix which is learned using the factor analysis approach over huge amount of data from a large number of speakers. Also, the estimation of i-vector for a given utterance involves a computationally complex procedure. Motivated by these facts, we explore the use of data-independent projection approaches for reducing the dimensionality of GMM mean supervectors. The data-independent projection methods studied in this work include a normal random projection and two kinds of sparse random projections. The study is performed on SRC-based speaker identification using the NIST SRE 2005 dataset which includes channel matched and mismatched conditions. We find that the use of data-independent random projections for the dimensionality reduction of the supervectors results in only 3 % absolute loss in performance compared to that of the data-dependent (i-vector) approach. It is highlighted that with the use of highly sparse random projection matrices having 1 as non-zero coefficients, a significant reduction in computational complexity is achieved in finding the projections. Further, as these matrices do not require floating point representations, their storage requirement is also very small compared to that of the data-dependent or the normal random projection matrices. These reduced complexity sparse random projections would be of interest in context of the speaker recognition applications implemented on platforms having low computational power.

引用

页码：2521 / 2538

页数：18

共 50 条

[31] Sparse representation-based image quality assessment
Guha, Tanaya
Nezhadarya, Ehsan
Ward, Rabab K.
SIGNAL PROCESSING-IMAGE COMMUNICATION, 2014, 29 (10) : 1138 - 1148
[32] An adaptive kernel sparse representation-based classification
Xuejun Wang
Wenjian Wang
Changqian Men
International Journal of Machine Learning and Cybernetics, 2020, 11 : 2209 - 2219
[33] Multiple kernel sparse representation-based classification
Chen, Si-Bao, 1807, Chinese Institute of Electronics (42):
[34] Sparse representation-based classification of mysticete calls
Guilment, Thomas
Socheleau, Francois-Xavier
Pastor, Dominique
Vallez, Simon
JOURNAL OF THE ACOUSTICAL SOCIETY OF AMERICA, 2018, 144 (03): : 1550 - 1563
[35] An adaptive kernel sparse representation-based classification
Wang, Xuejun
Wang, Wenjian
Men, Changqian
INTERNATIONAL JOURNAL OF MACHINE LEARNING AND CYBERNETICS, 2020, 11 (10) : 2209 - 2219
[36] Sparse Representation-Based Open Set Recognition
Zhang, He
Patel, Vishal M.
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2017, 39 (08) : 1690 - 1696
[37] Deep Hashing for Speaker Identification and Retrieval Based on Auditory Sparse Representation
Tran, Dung Kim
Akagi, Masato
Unoki, Masashi
PROCEEDINGS OF 2022 ASIA-PACIFIC SIGNAL AND INFORMATION PROCESSING ASSOCIATION ANNUAL SUMMIT AND CONFERENCE (APSIPA ASC), 2022, : 937 - 943
[38] Sparse Coding Based Lip Texture Representation For Visual Speaker Identification
Lai, Jun-Yao
Wang, Shi-Lin
Shi, Xing-Jian
Liew, Alan Wee-Chung
2014 19TH INTERNATIONAL CONFERENCE ON DIGITAL SIGNAL PROCESSING (DSP), 2014, : 607 - 610
[39] Sparse and Redundant Representation-Based Smart Meter Data Compression and Pattern Extraction
Wang, Yi
Chen, Qixin
Kang, Chongqing
Xia, Qing
Luo, Min
IEEE TRANSACTIONS ON POWER SYSTEMS, 2017, 32 (03) : 2142 - 2151
[40] A Sparse Representation-Based Algorithm for Pattern Localization in Brain Imaging Data Analysis
Li, Yuanqing
Long, Jinyi
He, Lin
Lu, Haidong
Gu, Zhenghui
Sun, Pei
PLOS ONE, 2012, 7 (12):

← 1 2 3 4 5 →