Audio and Speech Processing

Authors and titles for eess.AS in Jun 2023

[ total of 377 entries: 1-10 | 11-20 | 21-30 | 31-40 | ... | 371-377 ]
[ showing 10 entries per page: fewer | more | all ]

[1] arXiv:2306.00160 [pdf, other]: Title: Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Model

Authors: Héctor Martel, Julius Richter, Kai Li, Xiaolin Hu, Timo Gerkmann

Comments: Accepted by Interspeech 2023

Subjects: Audio and Speech Processing (eess.AS); Machine Learning (cs.LG); Sound (cs.SD)
[2] arXiv:2306.00203 [pdf, ps, other]: Title: Speaker-independent Speech Inversion for Estimation of Nasalance

Authors: Yashish M. Siriwardena, Carol Espy-Wilson, Suzanne Boyce, Mark K.Tiede, Liran Oren

Comments: Interspeech 2023

Subjects: Audio and Speech Processing (eess.AS)
[3] arXiv:2306.00331 [pdf, other]: Title: A Multi-dimensional Deep Structured State Space Approach to Speech Enhancement Using Small-footprint Models

Authors: Pin-Jui Ku, Chao-Han Huck Yang, Sabato Marco Siniscalchi, Chin-Hui Lee

Comments: Accepted to Interspeech 2023. Code will be released at this https URL

Subjects: Audio and Speech Processing (eess.AS); Artificial Intelligence (cs.AI); Sound (cs.SD); Signal Processing (eess.SP); Systems and Control (eess.SY)
[4] arXiv:2306.00426 [pdf, ps, other]: Title: Speaker verification using attentive multi-scale convolutional recurrent network

Authors: Yanxiong Li, Zhongjie Jiang, Wenchang Cao, Qisheng Huang

Comments: 21 pages, 6 figures, 8 tables. Accepted for publication in Applied Soft Computing

Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
[5] arXiv:2306.00452 [pdf, ps, other]: Title: Speech Self-Supervised Representation Benchmarking: Are We Doing it Right?

Authors: Salah Zaiem, Youcef Kemiche, Titouan Parcollet, Slim Essid, Mirco Ravanelli

Comments: 6 pages

Journal-ref: INTERSPEECH 2023

Subjects: Audio and Speech Processing (eess.AS); Machine Learning (cs.LG)
[6] arXiv:2306.00481 [pdf, other]: Title: Automatic Data Augmentation for Domain Adapted Fine-Tuning of Self-Supervised Speech Representations

Authors: Salah Zaiem, Titouan Parcollet, Slim Essid

Comments: 6 pages,INTERSPEECH 2023

Subjects: Audio and Speech Processing (eess.AS); Machine Learning (cs.LG)
[7] arXiv:2306.00625 [pdf, other]: Title: Frame-wise and overlap-robust speaker embeddings for meeting diarization

Authors: Tobias Cord-Landwehr, Christoph Boeddeker, Cătălin Zorilă, Rama Doddipatla, Reinhold Haeb-Umbach

Comments: ICASSP 2023

Subjects: Audio and Speech Processing (eess.AS)
[8] arXiv:2306.00634 [pdf, other]: Title: A Teacher-Student approach for extracting informative speaker embeddings from speech mixtures

Authors: Tobias Cord-Landwehr, Christoph Boeddeker, Cătălin Zorilă, Rama Doddipatla, Reinhold Haeb-Umbach

Comments: Proceedings of INTERSPEECH

Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
[9] arXiv:2306.00736 [pdf, other]: Title: Spoken Language Identification System for English-Mandarin Code-Switching Child-Directed Speech

Authors: Shashi Kant Gupta, Sushant Hiray, Prashant Kukde

Comments: Accepted by Interspeech 2023, 5 pages, 1 figure, 4 tables

Journal-ref: Proc. INTERSPEECH 2023, 4114--4118

Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
[10] arXiv:2306.00812 [pdf, other]: Title: Harmonic enhancement using learnable comb filter for light-weight full-band speech enhancement model

Authors: Xiaohuai Le, Tong Lei, Li Chen, Yiqing Guo, Chao He, Cheng Chen, Xianjun Xia, Hua Gao, Yijian Xiao, Piao Ding, Shenyi Song, Jing Lu

Comments: accepted by Interspeech 2023

Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)

[ total of 377 entries: 1-10 | 11-20 | 21-30 | 31-40 | ... | 371-377 ]
[ showing 10 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, eess, 2405, contact, help (Access key information)

> eess > eess.AS

Audio and Speech Processing

Authors and titles for eess.AS in Jun 2023