Stop the war!

Остановите войну!

for scientists:

default search action

combined dblp search
author search
venue search
publication search

ask others

IEEE/ACM Transactions on Audio, Speech and Language Processing, Volume 32

> Home > Journals > IEEE/ACM Transactions on Audio, Speech and Language Processing

Refine list

refinements active!

zoomed in on ?? of ?? records

view refined list in

export refined list as

showing all ?? records

Volume 32, 2024

- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WuK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WuK24
Jin Chu Wu, Raghu N. Kacker:
Statistical Analysis for Speaker Recognition Evaluation With Data Dependence and Three Score Distributions. 1-14
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhouBWHZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhouBWHZ24
Yongwei Zhou, Junwei Bao, Youzheng Wu, Xiaodong He, Tiejun Zhao:
Operation-Augmented Numerical Reasoning for Question Answering. 15-28
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/PurushothamanDKG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PurushothamanDKG24
Anurenjan Purushothaman, Debottam Dutta, Rohit Kumar, Sriram Ganapathy:
Speech Dereverberation With Frequency Domain Autoregressive Modeling. 29-38
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/QuLWPRW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/QuLWPRW24
Leyuan Qu, Taihao Li, Cornelius Weber, Theresa Pekarek-Rosin, Fuji Ren, Stefan Wermter:
Disentangling Prosody Representations With Unsupervised Speech Reconstruction. 39-54
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/PedersenJTJ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PedersenJTJ24
Mathias Bach Pedersen, Søren Holdt Jensen, Zheng-Hua Tan, Jesper Jensen:
Data-Driven Non-Intrusive Speech Intelligibility Prediction Using Speech Presence Probability. 55-67
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/HouKMWKB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HouKMWKB24
Yuanbo Hou, Bo Kang, Andrew Mitchell, Wenwu Wang, Jian Kang, Dick Botteldooren:
Cooperative Scene-Event Modelling for Acoustic Scene Classification. 68-82
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangYCWZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangYCWZ24
Xiaotong Jiang, Peiwen You, Chen Chen, Zhongqing Wang, Guodong Zhou:
Exploring Scope Detection for Aspect-Based Sentiment Analysis. 83-94
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/XuXWY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/XuXWY24
Xuenan Xu, Zeyu Xie, Mengyue Wu, Kai Yu:
Beyond the Status Quo: A Contemporary Survey of Advances and Challenges in Audio Captioning. 95-112
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/MiotelloPCAS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MiotelloPCAS24
Federico Miotello, Mirco Pezzoli, Luca Comanducci, Fabio Antonacci, Augusto Sarti:
Deep Prior-Based Audio Inpainting Using Multi-Resolution Harmonic Convolutional Neural Networks. 113-123
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/StanciuBPCDC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/StanciuBPCDC24
Cristian Lucian Stanciu, Jacob Benesty, Constantin Paleologu, Ruxandra-Liana Costea, Laura-Maria Dogariu, Silviu Ciochina:
Decomposition-Based Wiener Filter Using the Kronecker Product and Conjugate Gradient Method. 124-138
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenSZZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenSZZ24
Huiyao Chen, Yueheng Sun, Meishan Zhang, Min Zhang:
Automatic Noise Generation and Reduction for Text Classification. 139-150
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/XuCHX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/XuCHX24
Jiaming Xu, Jian Cui, Yunzhe Hao, Bo Xu:
Multi-Cue Guided Semi-Supervised Learning Toward Target Speaker Separation in Real Environments. 151-163
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/XiangHRC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/XiangHRC24
Yang Xiang, Jesper Lisby Højvang, Morten Højfeldt Rasmussen, Mads Græsbøll Christensen:
A Two-Stage Deep Representation Learning-Based Speech Enhancement Method Using Variational Autoencoder and Adversarial Training. 164-177
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiLHW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiLHW24
Xiao Li, Ruirui Liu, Huichou Huang, Qingyao Wu:
Contrastive Learning for Target Speaker Extraction With Attention-Based Fusion. 178-188
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiangMWLZL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiangMWLZL24
Xiaobo Liang, Runze Mao, Lijun Wu, Juntao Li, Min Zhang, Qing Li:
Enhancing Low-Resource NLP by Consistency Training With Data and Model Perturbations. 189-199
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LuLS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LuLS24
Haisheng Lu, Jiangnan Liang, Chuang Shi:
Comments on "Primary-Ambient Extraction Using Ambient Spectrum Estimation for Immersive Spatial Audio Reproduction". 200-202
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/DrgasBPNV24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/DrgasBPNV24
Szymon Drgas, Lars Bramsløw, Archontis Politis, Gaurav Naithani, Tuomas Virtanen:
Dynamic Processing Neural Network Architecture for Hearing Loss Compensation. 203-214
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GelderblomTSM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GelderblomTSM24
Femke B. Gelderblom, Tron V. Tronstad, Torbjørn Svendsen, Tor André Myrvoll:
On the Predictive Power of Objective Intelligibility Metrics for the Subjective Performance of Deep Complex Convolutional Recurrent Speech Enhancement Networks. 215-226
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HaubnerBK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HaubnerBK24
Thomas Haubner, Andreas Brendel, Walter Kellermann:
End-to-End Deep Learning-Based Adaptation Control for Linear Acoustic Echo Cancellation. 227-238
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangQL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangQL24
Congcong Jiang, Tieyun Qian, Bing Liu:
One General Teacher for Multi-Data Multi-Task: A New Knowledge Distillation Framework for Discourse Relation Analysis. 239-249
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/NayemW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/NayemW24
Khandokar Md. Nayem, Donald S. Williamson:
Attention-Based Speech Enhancement Using Human Quality Perception Modeling. 250-260
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhangMCXZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhangMCXZ24
Ying Zhang, Fandong Meng, Yufeng Chen, Jinan Xu, Jie Zhou:
Complex Question Enhanced Transfer Learning for Zero-Shot Joint Information Extraction. 261-275
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/YanLCZM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YanLCZM24
Jingsong Yan, Piji Li, Haibin Chen, Junhao Zheng, Qianli Ma:
Does the Order Matter? A Random Generative Way to Learn Label Hierarchy for Hierarchical Text Classification. 276-285
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ParaskevopoulosKRKKP24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ParaskevopoulosKRKKP24
Georgios Paraskevopoulos, Theodoros Kouzelis, Georgios Rouvalis, Athanasios Katsamanis, Vassilis Katsouros, Alexandros Potamianos:
Sample-Efficient Unsupervised Domain Adaptation of Speech Recognition Systems: A Case Study for Modern Greek. 286-299
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AccoltiGV24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AccoltiGV24
Ernesto Accolti, Javier Gimenez, Michael Vorländer:
Uncertainties of Room Acoustics Simulation Due to Directivity Data of Musical Instruments. 300-309
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/MasuyamaYKNO24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MasuyamaYKNO24
Yoshiki Masuyama, Kouei Yamaoka, Yuma Kinoshita, Taishi Nakashima, Nobutaka Ono:
Causal and Relaxed-Distortionless Response Beamforming for Online Target Source Extraction. 310-324
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/PrabhavalkarHSSW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PrabhavalkarHSSW24
Rohit Prabhavalkar, Takaaki Hori, Tara N. Sainath, Ralf Schlüter, Shinji Watanabe:
End-to-End Speech Recognition: A Survey. 325-351
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoLWLNL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoLWLNL24
Yun Zhao, Dexi Liu, Changxuan Wan, Xiping Liu, Jian-Yun Nie, Jiaming Liu:
JMS-QA: A Joint Hierarchical Architecture for Mental Health Question Answering. 352-363
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/NiLYK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/NiLYK24
Shiwen Ni, Jiawen Li, Min Yang, Hung-Yu Kao:
DropAttack: A Random Dropped Weight Attack Adversarial Training for Natural Language Understanding. 364-373
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhuQFCHX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhuQFCHX24
Tiantian Zhu, Yang Qin, Ming Feng, Qingcai Chen, Baotian Hu, Yang Xiang:
BioPRO: Context-Infused Prompt Learning for Biomedical Entity Linking. 374-385
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangWGHHY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangWGHHY24
Jiapu Wang, Boyue Wang, Junbin Gao, Simin Hu, Yongli Hu, Baocai Yin:
Multi-Level Interaction Based Knowledge Graph Completion. 386-396
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhangLXZW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhangLXZW24
Qiangqiang Zhang, Dongyuan Lin, Yingying Xiao, Yunfei Zheng, Shiyuan Wang:
Error Reused Filtered-X Least Mean Square Algorithm for Active Noise Control. 397-412
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JinGDWHLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JinGDWHLL24
Zengrui Jin, Mengzhe Geng, Jiajun Deng, Tianzi Wang, Shujie Hu, Guinan Li, Xunying Liu:
Personalized Adversarial Data Augmentation for Dysarthric and Elderly Speech Recognition. 413-429
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KongWZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KongWZ24
Jun Kong, Jin Wang, Xuejie Zhang:
Adaptive Ensemble Self-Distillation With Consistent Gradients for Fast Inference of Pretrained Language Models. 430-442
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KiticD24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KiticD24
Srdan Kitic, Jérôme Daniel:
Blind Identification of Ambisonic Reduced Room Impulse Response. 443-458
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ShaoGYHX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ShaoGYHX24
Qijie Shao, Pengcheng Guo, Jinghao Yan, Pengfei Hu, Lei Xie:
Decoupling and Interacting Multi-Task Learning Network for Joint Speech and Accent Recognition. 459-470
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhuCWHZY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhuCWHZY24
Han Zhu, Gaofeng Cheng, Jindong Wang, Wenxin Hou, Pengyuan Zhang, Yonghong Yan:
Boosting Cross-Domain Speech Recognition With Self-Supervision. 471-485
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangZLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangZLL24
Yile Wang, Yue Zhang, Peng Li, Yang Liu:
Gradual Syntactic Label Replacement for Language Model Pre-Training. 486-496
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/MaLPZG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MaLPZG24
Penghui Ma, Jianfeng Li, Jingjing Pan, Xiaofei Zhang, Roberto Gil-Pita:
Coherent Signal DOA Estimation With Coprime Array: Exploiting Signal Subspace Reconstructing Strategy. 497-508
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HamelK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HamelK24
Emma Hamel, Nickvash Kani:
Factors That Influence Automatic Recognition of African-American Vernacular English in Machine-Learning Models. 509-516
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiLCZMWMTWW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiLCZMWMTWW24
Jingbei Li, Sipan Li, Ping Chen, Luwen Zhang, Yi Meng, Zhiyong Wu, Helen Meng, Qiao Tian, Yuping Wang, Yuxuan Wang:
Joint Multiscale Cross-Lingual Speaking Style Transfer With Bidirectional Attention Mechanism for Automatic Dubbing. 517-528
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HanCQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HanCQ24
Bing Han, Zhengyang Chen, Yanmin Qian:
Self-Supervised Learning With Cluster-Aware-DINO for High-Performance Robust Speaker Verification. 529-541
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/TeschG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/TeschG24
Kristina Tesch, Timo Gerkmann:
Multi-Channel Speech Separation Using Spatially Selective Deep Non-Linear Filters. 542-553
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/PeiFLX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PeiFLX24
Hao-Chen Pei, Hao Fang, Xin Luo, Xin-Shun Xu:
Gradformer: A Framework for Multi-Aspect Multi-Granularity Pronunciation Assessment. 554-563
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/SharmaUK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SharmaUK24
Garima Sharma, Karthikeyan Umapathy, Sridhar Krishnan:
Time-Frequency Scattergrams for Biomedical Audio Signal Representation and Classification. 564-576
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ManHZLCCX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ManHZLCCX24
Zhibo Man, Zengcheng Huang, Yujie Zhang, Yu Li, Yuanmeng Chen, Yufeng Chen, Jinan Xu:
WDSRL: Multi-Domain Neural Machine Translation With Word-Level Domain-Sensitive Representation Learning. 577-590
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenPGL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenPGL24
Chin-Po Chen, Ho-Hsien Pan, Susan Shur-Fen Gau, Chi-Chun Lee:
Using Measures of Vowel Space for Autistic Traits Characterization. 591-607
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/WilkinghoffK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WilkinghoffK24
Kevin Wilkinghoff, Frank Kurth:
Why Do Angular Margin Losses Work Well for Semi-Supervised Anomalous Sound Detection? 608-622
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/RouheGK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/RouheGK24
Aku Rouhe, Tamás Grósz, Mikko Kurimo:
Principled Comparisons for End-to-End Speech Recognition: Attention vs Hybrid at the 1000-Hour Scale. 623-638
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangZ24
Yile Wang, Yue Zhang:
Lost in Context? On the Sense-Wise Variance of Contextualized Word Embeddings. 639-650
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/HoldPPM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HoldPPM24
Christoph Hold, Ville Pulkki, Archontis Politis, Leo McCormack:
Compression of Higher-Order Ambisonic Signals Using Directional Audio Coding. 651-665
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangQ24
Shouhui Wang, Biao Qin:
A Novel Joint Training Model for Knowledge Base Question Answering. 666-679
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiWLS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiWLS24
Songbin Li, Jingang Wang, Peng Liu, Ke Shi:
SANet: A Compressed Speech Encoder and Steganography Algorithm Independent Steganalysis Deep Neural Network. 680-690
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KananAAHMKK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KananAAHMKK24
Tarek Kanan, Amani AbedAlghafer, Shadi AlZu'bi, Bilal Hawashin, Ala Mughaid, Ghassan Kanaan, M. M. Kamruzzaman:
An Intelligent Health Care System for Detecting Drug Abuse in Social Media Platforms Based on Low Resource Language. 691-703
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/VarelaSKDK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/VarelaSKDK24
Alejandro Santorum Varela, Svetlana Stoyanchev, Simon Keizer, Rama Doddipatla, Kate Knill:
Entity Resolution in Situated Dialog With Unimodal and Multimodal Transformers. 704-713
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HeLBWWNW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HeLBWWNW24
Huang He, Hua Lu, Siqi Bao, Fan Wang, Hua Wu, Zheng-Yu Niu, Haifeng Wang:
Learning to Select External Knowledge With Multi-Scale Negative Sampling. 714-720
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LuGLYHB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LuGLYHB24
Hua Lu, Zhen Guo, Chanjuan Li, Yunyi Yang, Huang He, Siqi Bao:
Towards Building an Open-Domain Dialogue System Incorporated With Internet Memes. 721-726
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/LimWLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LimWLL24
Jungwoo Lim, Taesun Whang, Dongyub Lee, Heuiseok Lim:
Adaptive Multi-Domain Dialogue State Tracking on Spoken Conversations. 727-732
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ThulkeDDN24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ThulkeDDN24
David Thulke, Nico Daheim, Christian Dugast, Hermann Ney:
Task-Oriented Document-Grounded Dialog Systems by HLTPR@RWTH for DSTC9 and DSTC10. 733-741
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WuXS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WuXS24
Han Wu, Kun Xu, Linqi Song:
Structure-Aware Dialogue Modeling Methods for Conversational Semantic Role Labeling. 742-752
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenLW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenLW24
Zhe Chen, Hongcheng Liu, Yu Wang:
DialogMCF: Multimodal Context Flow for Audio Visual Scene-Aware Dialog. 753-764
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/YoshinoCCKLHMFLZFZKLJPGHDGHSZLSDB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YoshinoCCKLHMFLZFZKLJPGHDGHSZLSDB24
Koichiro Yoshino, Yun-Nung Chen, Paul A. Crook, Satwik Kottur, Jinchao Li, Behnam Hedayatnia, Seungwhan Moon, Zhengcong Fei, Zekang Li, Jinchao Zhang, Yang Feng, Jie Zhou, Seokhwan Kim, Yang Liu, Di Jin, Alexandros Papangelis, Karthik Gopalakrishnan, Dilek Hakkani-Tur, Babak Damavandi, Alborz Geramifard, Chiori Hori, Ankit Shah, Chen Zhang, Haizhou Li, João Sedoc, Luis F. D'Haro, Rafael E. Banchs, Alexander Rudnicky:
Overview of the Tenth Dialog System Technology Challenge: DSTC10. 765-778
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/YadavG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YadavG24
Shekhar Kumar Yadav, Nithin V. George:
Joint Dereverberation and Beamforming With Blind Estimation of the Shape Parameter of the Desired Source Prior. 779-793
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiJHCL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiJHCL24
Yanxiong Li, Zhongjie Jiang, Qisheng Huang, Wenchang Cao, Jialong Li:
Lightweight Speaker Verification Using Transformation Module With Feature Partition and Fusion. 794-806
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/DaiZDLLX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/DaiZDLLX24
Yuhan Dai, Zhirui Zhang, Yichao Du, Shengcai Liu, Lemao Liu, Tong Xu:
Datastore Distillation for Nearest Neighbor Machine Translation. 807-817
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiYY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiYY24
Changtao Li, Feiran Yang, Jun Yang:
A Two-Stage Approach to Quality Restoration of Bone-Conducted Speech. 818-829
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhouLCZHH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhouLCZHH24
Jie Zhou, Yuanbiao Lin, Qin Chen, Qi Zhang, Xuanjing Huang, Liang He:
CausalABSC: Causal Inference for Aspect Debiasing in Aspect-Based Sentiment Classification. 830-840
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LuCGWZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LuCGWZ24
Ruiying Lu, Bo Chen, Dandan Guo, Dongsheng Wang, Mingyuan Zhou:
Hierarchical Topic-Aware Contextualized Transformers. 841-852
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoCHW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoCHW24
Yaru Zhao, Bo Cheng, Yakun Huang, Zhiguo Wan:
FluGCF: A Fluent Dialogue Generation Model With Coherent Concept Entity Flow. 853-867
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/DingFYYLH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/DingFYYLH24
Changhao Ding, Zhangjie Fu, Zhongliang Yang, Qi Yu, Daqiu Li, Yongfeng Huang:
Context-Aware Linguistic Steganography Model Based on Neural Machine Translation. 868-878
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AlhakeemJK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AlhakeemJK24
Zainab Alhakeem, Se-In Jang, Hong-Goo Kang:
Disentangled Representations in Local-Global Contexts for Arabic Dialect Identification. 879-890
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LeeC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LeeC24
Jae-Hong Lee, Joon-Hyuk Chang:
Partitioning Attention Weight: Mitigating Adverse Effect of Incorrect Pseudo-Labels for Self-Supervised ASR. 891-905
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/FukudaSN24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/FukudaSN24
Ryo Fukuda, Katsuhito Sudoh, Satoshi Nakamura:
Improving Speech Translation Accuracy and Time Efficiency With Fine-Tuned wav2vec 2.0-Based Speech Segmentation. 906-916
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/LeemFOGB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LeemFOGB24
Seong-Gyun Leem, Daniel Fulford, Jukka-Pekka Onnela, David Gard, Carlos Busso:
Selective Acoustic Feature Enhancement for Speech Emotion Recognition With Noisy Speech. 917-929
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BohlenderSTM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BohlenderSTM24
Alexander Bohlender, Ann Spriet, Wouter Tirry, Nilesh Madhu:
Spatially Selective Speaker Separation Using a DNN With a Location Dependent Feature Extraction. 930-945
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KaroYL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KaroYL24
Matan Karo, Arie Yeredor, Itshak Lapidot:
Compact Time-Domain Representation for Logical Access Spoofed Audio. 946-958
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/BerebiBAR24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BerebiBAR24
Or Berebi, Zamir Ben-Hur, David Lou Alon, Boaz Rafaely:
Analysis and Design of Head-Tracked Compensation for Bilateral Ambisonics. 959-972
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangQ24a
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangQ24a
Wei Wang, Yanmin Qian:
Universal Cross-Lingual Data Generation for Low Resource ASR. 973-983
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BerghiJ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BerghiJ24
Davide Berghi, Philip J. B. Jackson:
Leveraging Visual Supervision for Array-Based Active Speaker Detection and Localization. 984-995
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/KrauseGPM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KrauseGPM24
Daniel Aleksander Krause, Guillermo García-Barrios, Archontis Politis, Annamaria Mesaros:
Binaural Sound Source Distance Estimation and Localization for a Moving Listener. 996-1011
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/KimLCL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KimLCL24
Seung-Bin Kim, Sang-Hoon Lee, Ha-Yeong Choi, Seong-Whan Lee:
Audio Super-Resolution With Robust Speech Representation Learning of Masked Autoencoder. 1012-1022
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BattalK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BattalK24
Omer Musa Battal, Aykut Koç:
Automatic Construction of Sememe Knowledge Bases From Machine Readable Dictionaries. 1023-1035
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KrishnaSG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KrishnaSG24
Varun Krishna, Tarun Sai, Sriram Ganapathy:
Representation Learning With Hidden Unit Clustering for Low Resource Speech Applications. 1036-1047
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LuoSGH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LuoSGH24
Zhengding Luo, Dongyuan Shi, Woon-Seng Gan, Qirui Huang:
Delayless Generative Fixed-Filter Active Noise Control Based on Deep Learning and Bayesian Filter. 1048-1060
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChiHLBGM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChiHLBGM24
Zewen Chi, Heyan Huang, Luyang Liu, Yu Bai, Xiaoyan Gao, Xian-Ling Mao:
Can Pretrained English Language Models Benefit Non-English NLP Systems in Low-Resource Scenarios? 1061-1074
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuHZLWG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuHZLWG24
Rui Liu, Yifan Hu, Haolin Zuo, Zhaojie Luo, Longbiao Wang, Guanglai Gao:
Text-to-Speech for Low-Resource Agglutinative Language With Morphology-Aware Language Model Pre-Training. 1075-1087
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangLZD24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangLZD24
Shu Jiang, Zuchao Li, Hai Zhao, Weiping Ding:
Entity-Relation Extraction as Full Shallow Semantic Dependency Parsing. 1088-1099
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/VeredE24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/VeredE24
Yoav Vered, Stephen J. Elliott:
A Parallel Analog and Digital Adaptive Feedforward Controller for Active Noise Control. 1100-1108
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhangZYLY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhangZYLY24
Puning Zhang, Rongjian Zhao, Boran Yang, Yuexian Li, Zhigang Yang:
Integrated Syntactic and Semantic Tree for Targeted Sentiment Classification Using Dual-Channel Graph Convolutional Network. 1109-1124
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangZZCDWCL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangZZCDWCL24
Xu Wang, Hainan Zhang, Shuai Zhao, Hongshen Chen, Zhuoye Ding, Zhiguo Wan, Bo Cheng, Yanyan Lan:
Debiasing Counterfactual Context With Causal Inference for Multi-Turn Dialogue Reasoning. 1125-1132
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChauBNDN24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChauBNDN24
Hoang Ngoc Chau, Tien Dat Bui, Huu Binh Nguyen, Thanh Thi Hien Duong, Quoc-Cuong Nguyen:
A Novel Approach to Multi-Channel Speech Enhancement Based on Graph Neural Networks. 1133-1144
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HuCZC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HuCZC24
Yuchen Hu, Chen Chen, Qiushi Zhu, Eng Siong Chng:
Wav2code: Restore Clean Speech Representations via Codebook Lookup for Noise-Robust ASR. 1145-1156
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/UedaNIKAM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/UedaNIKAM24
Tetsuya Ueda, Tomohiro Nakatani, Rintaro Ikeshita, Keisuke Kinoshita, Shoko Araki, Shoji Makino:
Blind and Spatially-Regularized Online Joint Optimization of Source Separation, Dereverberation, and Noise Reduction. 1157-1172
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AgarwalGSAR24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AgarwalGSAR24
Vibhav Agarwal, Sourav Ghosh, Harichandana B. S. S, Himanshu Arora, Barath Raj Kandur Raja:
TrICy: Trigger-Guided Data-to-Text Generation With Intent Aware Attention-Copy. 1173-1184
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BoddekerSWHR24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BoddekerSWHR24
Christoph Böddeker, Aswin Shanmugam Subramanian, Gordon Wichern, Reinhold Haeb-Umbach, Jonathan Le Roux:
TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings. 1185-1197
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/VarzandehDH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/VarzandehDH24
Reza Varzandeh, Simon Doclo, Volker Hohmann:
Speech-Aware Binaural DOA Estimation Utilizing Periodicity and Spatial Features in Convolutional Neural Networks. 1198-1213
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/OzerM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/OzerM24
Yigitcan Özer, Meinard Müller:
Source Separation of Piano Concertos Using Musically Motivated Augmentation Techniques. 1214-1225
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/FrenkelCG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/FrenkelCG24
Lior Frenkel, Shlomo E. Chazan, Jacob Goldberger:
Domain Adaptation Using Suitable Pseudo Labels for Speech Enhancement and Dereverberation. 1226-1236
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoMZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoMZ24
Jiahao Zhao, Wenji Mao, Daniel Dajun Zeng:
Disentangled Text Representation Learning With Information-Theoretic Perspective for Adversarial Robustness. 1237-1247
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhouLLZY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhouLLZY24
Dong Zhou, Fang Lei, Lin Li, Yongmei Zhou, Aimin Yang:
Cross-Modal Interaction via Reinforcement Feedback for Audio-Lyrics Retrieval. 1248-1260
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuSLK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuSLK24
Xuechen Liu, Md. Sahidullah, Kong Aik Lee, Tomi Kinnunen:
Generalizing Speaker Verification for Spoof Awareness in the Embedding Space. 1261-1273
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/CuiCCSLLS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/CuiCCSLLS24
Shiyao Cui, Jiangxia Cao, Xin Cong, Jiawei Sheng, Quangang Li, Tingwen Liu, Jinqiao Shi:
Enhancing Multimodal Entity and Relation Extraction With Variational Information Bottleneck. 1274-1285
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/TanALP24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/TanALP24
Yizhou Tan, Haojun Ai, Shengchen Li, Mark D. Plumbley:
Acoustic Scene Classification Across Cities and Devices via Feature Disentanglement. 1286-1297
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/ZakenKTR24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZakenKTR24
Orel Ben Zaken, Anurag Kumar, Vladimir Tourbabin, Boaz Rafaely:
Neural-Network-Based Direction-of-Arrival Estimation for Reverberant Speech - The Importance of Energetic, Temporal, and Spatial Information. 1298-1309
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/QuanL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/QuanL24
Changsheng Quan, Xiaofei Li:
SpatialNet: Extensively Learning Spatial Information for Multichannel Joint Speech Separation, Denoising and Dereverberation. 1310-1323
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BaasK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BaasK24
Matthew Baas, Herman Kamper:
Disentanglement in a GAN for Unconditional Speech Synthesis. 1324-1335
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiSL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiSL24
Xian Li, Nian Shao, Xiaofei Li:
Self-Supervised Audio Teacher-Student Transformer for Both Clip-Level and Frame-Level Tasks. 1336-1351
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenCYZY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenCYZY24
Yifan Chen, Gaofeng Cheng, Runyan Yang, Pengyuan Zhang, Yonghong Yan:
Interrelate Training and Clustering for Online Speaker Diarization. 1352-1364
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/FengZM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/FengZM24
Sheng Feng, Xiaoqian Zhu, Shuqing Ma:
Masking Hierarchical Tokens for Underwater Acoustic Target Recognition With Self-Supervised Learning. 1365-1379
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoYWDW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoYWDW24
Yangyang Zhao, Kai Yin, Zhenyu Wang, Mehdi Dastani, Shihan Wang:
Decomposed Deep Q-Network for Coherent Task-Oriented Dialogue Policy Learning. 1380-1391
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ParekhPMRd24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ParekhPMRd24
Jayneel Parekh, Sanjeel Parekh, Pavlo Mozharovskyi, Gaël Richard, Florence d'Alché-Buc:
Tackling Interpretability in Audio Classification Networks With Non-negative Matrix Factorization. 1392-1405
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenGLZGZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenGLZGZ24
Xiuying Chen, Shen Gao, Mingzhe Li, Qingqing Zhu, Xin Gao, Xiangliang Zhang:
Write Summary Step-by-Step: A Pilot Study of Stepwise Summarization. 1406-1415
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LinCRY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LinCRY24
Changkai Lin, Hongju Cheng, Qiang Rao, Yang Yang:
M$^{3}$SA: Multimodal Sentiment Analysis Based on Multi-Scale Feature Extraction and Multi-Task Learning. 1416-1429
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BiswasNA24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BiswasNA24
Ritujoy Biswas, Karan Nathwani, Vinayak Abrol:
Statistically Guided Near-End Speech Intelligibility Improvement Through Voice Transformation and Transfer Learning. 1445-1456
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/SunYGYC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SunYGYC24
Linhui Sun, Shuo Yuan, Aifei Gong, Lei Ye, Eng Siong Chng:
Dual-Branch Modeling Based on State-Space Model for Speech Enhancement. 1457-1467
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KoudounasPAMGGRCCABA24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KoudounasPAMGGRCCABA24
Alkis Koudounas, Eliana Pastor, Giuseppe Attanasio, Vittorio Mazzia, Manuel Giollo, Thomas Gueudré, Elisa Reale, Luca Cagliero, Sandro Cumani, Luca de Alfaro, Elena Baralis, Daniele Amberti:
Towards Comprehensive Subgroup Performance Analysis in Speech Models. 1468-1480
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/XiongBZJP24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/XiongBZJP24
Wenmeng Xiong, Changchun Bao, Jing Zhou, Maoshen Jia, José Picheral:
Joint DOA Estimation and Dereverberation Based on Multi-Channel Linear Prediction Filtering and Azimuth Sparsity. 1481-1493
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhengAL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhengAL24
Rui-Chen Zheng, Yang Ai, Zhen-Hua Ling:
Incorporating Ultrasound Tongue Images for Audio-Visual Speech Enhancement. 1430-1444
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AlkaherC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AlkaherC24
Yehav Alkaher, Israel Cohen:
Howling Detection and Gain Control for Speech Reinforcement in a Noisy Car Cabin Environment. 1494-1505
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhuLLZZLX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhuLLZZLX24
Xinfa Zhu, Yi Lei, Tao Li, Yongmao Zhang, Hongbin Zhou, Heng Lu, Lei Xie:
METTS: Multilingual Emotional Text-to-Speech by Cross-Speaker and Cross-Lingual Emotion Transfer. 1506-1518
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JeongKCYJK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JeongKCYJK24
Myeonghun Jeong, Minchan Kim, Byoung Jin Choi, Jaesam Yoon, Won Jang, Nam Soo Kim:
Transfer Learning for Low-Resource, Multi-Lingual, and Zero-Shot Multi-Speaker Text-to-Speech. 1519-1530
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/YaoLQZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YaoLQZ24
Jiadi Yao, Hong Luo, Jun Qi, Xiao-Lei Zhang:
Interpretable Spectrum Transformation Attacks to Speaker Recognition Systems. 1531-1545
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenLZDTHSZC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenLZDTHSZC24
Xiang Chen, Lei Li, Yuqi Zhu, Shumin Deng, Chuanqi Tan, Fei Huang, Luo Si, Ningyu Zhang, Huajun Chen:
Sequence Labeling as Non-Autoregressive Dual-Query Set Generation. 1546-1558
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuLL24
Lei Liu, Li Liu, Haizhou Li:
Computation and Parameter Efficient Multi-Modal Fusion Transformer for Cued Speech Recognition. 1559-1572
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/BarahonaRiosC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BarahonaRiosC24
Adrián Barahona-Ríos, Tom Collins:
NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks. 1573-1585
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangWXLF24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangWXLF24
Siyuan Wang, Zhongyu Wei, Jiarong Xu, Taishan Li, Zhihao Fan:
Unifying Structure Reasoning and Language Pre-Training for Complex Reasoning Tasks. 1586-1595
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChuZNDZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChuZNDZ24
Yijing Chu, Sipei Zhao, Feng Niu, Yongzheng Dong, Yuezhe Zhao:
A New Diffusion Filtered-X Affine Projection Algorithm: Performance Analysis and Application in Windy Environment. 1596-1608
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LeQWCL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LeQWCL24
Yuquan Le, Zhe Quan, Jiawei Wang, Da Cao, Kenli Li:
$\boldsymbol{R}^{2}$: A Novel Recall & Ranking Framework for Legal Judgment Prediction. 1609-1622
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangBWZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangBWZ24
Xiaotong Jiang, Ruirui Bai, Zhongqing Wang, Guodong Zhou:
Cross-Domain Aspect-Based Sentiment Classification With Tripartite Graph Modeling. 1623-1635
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenHWQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenHWQ24
Zhengyang Chen, Bing Han, Shuai Wang, Yanmin Qian:
Attention-Based Encoder-Decoder End-to-End Neural Diarization With Embedding Enhancer. 1636-1649
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/MiaoZCMWX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MiaoZCMWX24
Chenfeng Miao, Qingying Zhu, Minchuan Chen, Jun Ma, Shaojun Wang, Jing Xiao:
EfficientTTS 2: Variational End-to-End Text-to-Speech Synthesis and Voice Conversion. 1650-1661
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/PeretzC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PeretzC24
Orel Peretz, Israel Cohen:
Constant Elevation-Beamwidth Beamforming With Concentric Ring Arrays. 1662-1672
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/QuanVZY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/QuanVZY24
Zhibin Quan, Chi-Man Vong, Weili Zeng, Wankou Yang:
The MorPhEMe Machine: An Addressable Neural Memory for Learning Knowledge-Regularized Deep Contextualized Chinese Embedding. 1673-1686
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GaoMD24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GaoMD24
Lijian Gao, Qirong Mao, Ming Dong:
On Local Temporal Embedding for Semi-Supervised Sound Event Detection. 1687-1698
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhouZZWL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhouZZWL24
Xuehao Zhou, Mingyang Zhang, Yi Zhou, Zhizheng Wu, Haizhou Li:
Accented Text-to-Speech Synthesis With Limited Data. 1699-1711
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KothapallyH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KothapallyH24
Vinay Kothapally, John H. L. Hansen:
Monaural Speech Dereverberation Using Deformable Convolutional Networks. 1712-1723
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangYY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangYY24
Taihui Wang, Feiran Yang, Jun Yang:
Multichannel Linear Prediction-Based Speech Dereverberation Considering Sparse and Low-Rank Priors. 1724-1735
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KatariaVMZD24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KatariaVMZD24
Saurabh Kataria, Jesús Villalba, Laureano Moro-Velázquez, Piotr Zelasko, Najim Dehak:
Time-Domain Speech Super-Resolution With GAN Based Modeling for Telephony Speaker Verification. 1736-1749
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/OlivieriBPAAS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/OlivieriBPAAS24
Marco Olivieri, Amy Bastine, Mirco Pezzoli, Fabio Antonacci, Thushara D. Abhayapala, Augusto Sarti:
Acoustic Imaging With Circular Microphone Array: A New Approach for Sound Field Analysis. 1750-1761
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuHGSY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuHGSY24
Tengfei Liu, Yongli Hu, Junbin Gao, Yanfeng Sun, Baocai Yin:
Hierarchical Multi-Granularity Interaction Graph Convolutional Network for Long Document Classification. 1762-1775
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/ThuillierJV24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ThuillierJV24
Etienne Thuillier, Craig T. Jin, Vesa Välimäki:
HRTF Interpolation Using a Spherical Neural Process Meta-Learner. 1790-1802
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GongWLLZCQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GongWLLZCQ24
Xun Gong, Yu Wu, Jinyu Li, Shujie Liu, Rui Zhao, Xie Chen, Yanmin Qian:
Advanced Long-Content Speech Recognition With Factorized Neural Transducer. 1803-1815
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/MasuyamaYKO24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MasuyamaYKO24
Yoshiki Masuyama, Kouei Yamaoka, Takao Kawamura, Nobutaka Ono:
Efficient Joint Optimization of Sampling Rate Offsets Using Entire Multichannel Signal. 1816-1828
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/SaekiMLWTS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SaekiMLWTS24
Takaaki Saeki, Soumi Maiti, Xinjian Li, Shinji Watanabe, Shinnosuke Takamichi, Hiroshi Saruwatari:
Text-Inductive Graphone-Based Language Adaptation for Low-Resource Speech Synthesis. 1829-1844
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/OShaughnessy24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/OShaughnessy24
Douglas D. O'Shaughnessy:
Review of Methods for Automatic Speaker Verification. 1776-1789
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GaoBL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GaoBL24
Yingming Gao, Peter Birkholz, Ya Li:
Articulatory Copy Synthesis Based on the Speech Synthesizer VocalTractLab and Convolutional Recurrent Neural Networks. 1845-1858
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/MariotteLMT24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MariotteLMT24
Théo Mariotte, Anthony Larcher, Silvio Montrésor, Jean-Hugh Thomas:
Channel-Combination Algorithms for Robust Distant Voice Activity and Overlapped Speech Detection. 1859-1872
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/SouzaCB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SouzaCB24
Luciana M. X. de Souza, Márcio H. Costa, Renata Coelho Borges:
Envelope-Based Multichannel Noise Reduction for Cochlear Implant Applications. 1873-1884
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiCW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiCW24
Linjian Li, Yi Cai, Xin Wu:
Unsupervised Disentanglement Learning Model for Exemplar-Guided Paraphrase Generation. 1885-1900
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/IvryCB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/IvryCB24
Amir Ivry, Israel Cohen, Baruch Berdugo:
A User-Centric Approach for Deep Residual-Echo Suppression in Double-Talk. 1901-1914
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhangLZZXH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhangLZZXH24
Geng Zhang, Jin Liu, Guangyou Zhou, Kunsong Zhao, Zhiwen Xie, Bo Huang:
Question-Directed Reasoning With Relation-Aware Graph Attention Network for Complex Question Answering Over Knowledge Graph. 1915-1927
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/YaoYZY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YaoYZY24
Yu Yao, Peng Yang, Guangzhen Zhao, Guoshun Yin:
KGAgent: Learning a Deep Reinforced Agent for Keyphrase Generation. 1928-1940
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiLWQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiLWQ24
Jiahong Li, Chenda Li, Yifei Wu, Yanmin Qian:
Unified Cross-Modal Attention: Robust Audio-Visual Speech Recognition and Beyond. 1941-1953
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/FrasK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/FrasK24
Mieszko Fras, Konrad Kowalczyk:
Reverberant Source Separation Using NTF With Delayed Subsources and Spatial Priors. 1954-1967
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/WangLT24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangLT24
Rui Wang, Li Li, Tomoki Toda:
Dual-Channel Target Speaker Extraction Based on Conditional Variational Autoencoder and Directional Information. 1968-1979
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HanYLQ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HanYLQ24
Qinyu Han, Zhihao Yang, Hongfei Lin, Tian Qin:
Let Topic Flow: A Unified Topic-Guided Segment-Wise Dialogue Summarization Framework. 2021-2032
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChengLLYZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChengLLYZ24
Haonan Cheng, Shulin Liu, Zhicheng Lian, Long Ye, Qin Zhang:
MusicECAN: An Automatic Denoising Network for Music Recordings With Efficient Channel Attention. 2033-2049
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GubnitkyD24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GubnitkyD24
Guy Gubnitky, Roee Diamant:
Detecting the Presence of Sperm Whales' Echolocation Clicks in Noisy Environments. 2050-2061
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WuDZL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WuDZL24
Yuxia Wu, Tianhao Dai, Zhedong Zheng, Lizi Liao:
Active Discovering New Slots for Task-Oriented Conversation. 2062-2072
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HollebonF24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HollebonF24
Jacob Hollebon, Filippo Maria Fazi:
Dynamic Higher-Order Stereophony. 2073-2084
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HoggJLSCP24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HoggJLSCP24
Aidan O. T. Hogg, Mads Jenkins, He Liu, Isaac Squires, Samuel J. Cooper, Lorenzo Picinali:
HRTF Upsampling With a Generative Adversarial Network Using a Gnomonic Equiangular Projection. 2085-2099
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiaoWW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiaoWW24
Yusheng Liao, Yanfeng Wang, Yu Wang:
Leveraging Diverse Modeling Contexts With Collaborating Learning for Neural Machine Translation. 2100-2111
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiBLC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiBLC24
Shuo Li, Xiaojun Bi, Tao Liu, Zheng Chen:
Information Dropping Data Augmentation for Machine Translation Quality Estimation. 2112-2124
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangCXPW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangCXPW24
Shuoran Jiang, Qingcai Chen, Yang Xiang, Youcheng Pan, Xiangping Wu:
BaSFormer: A Balanced Sparsity Regularized Attention Network for Transformer. 2125-2140
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BuissonMEC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BuissonMEC24
Morgan Buisson, Brian McFee, Slim Essid, Hélène C. Crayencour:
Self-Supervised Learning of Multi-Level Audio Representations for Music Segmentation. 2141-2152
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/MaHWZZZZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MaHWZZZZ24
Cong Ma, Xu Han, Linghui Wu, Yaping Zhang, Yang Zhao, Yu Zhou, Chengqing Zong:
Modal Contrastive Learning Based End-to-End Text Image Machine Translation. 2153-2165
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiangXCPS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiangXCPS24
Ruiyu Liang, Yue Xie, Jiaming Cheng, Cong Pang, Björn W. Schuller:
A Non-Invasive Speech Quality Evaluation Algorithm for Hearing Aids With Multi-Head Self-Attention and Audiogram-Based Features. 2166-2176
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhangCZWRLYGDLW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhangCZWRLYGDLW24
Ziqiang Zhang, Sanyuan Chen, Long Zhou, Yu Wu, Shuo Ren, Shujie Liu, Zhuoyuan Yao, Xun Gong, Li-Rong Dai, Jinyu Li, Furu Wei:
SpeechLM: Enhanced Speech Pre-Training With Unpaired Textual Data. 2177-2187
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuSGL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuSGL24
Rui Liu, Berrak Sisman, Guanglai Gao, Haizhou Li:
Controllable Accented Text-to-Speech Synthesis With Fine and Coarse-Grained Intensity Rendering. 2188-2201
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/MishraFE24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MishraFE24
Kshitij Mishra, Mauajama Firdaus, Asif Ekbal:
Please Donate to Save a Life: Inducing Politeness to Handle Resistance in Persuasive Dialogue Agents. 2202-2212
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KameokaKTHS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KameokaKTHS24
Hirokazu Kameoka, Takuhiro Kaneko, Kou Tanaka, Nobukatsu Hojo, Shogo Seki:
VoiceGrad: Non-Parallel Any-to-Many Voice Conversion With Annealed Langevin Dynamics. 2213-2226
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/SchmidKW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SchmidKW24
Florian Schmid, Khaled Koutini, Gerhard Widmer:
Dynamic Convolutional Neural Networks as Efficient Pre-Trained Audio Models. 2227-2241
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/NeriPKCV24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/NeriPKCV24
Michael Neri, Archontis Politis, Daniel Aleksander Krause, Marco Carli, Tuomas Virtanen:
Speaker Distance Estimation in Enclosures From Single-Channel Audio. 2242-2254
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/KefalasPP24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KefalasPP24
Triantafyllos Kefalas, Yannis Panagakis, Maja Pantic:
Large-Scale Unsupervised Audio Pre-Training for Video-to-Speech Synthesis. 2255-2268
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KimHSLY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KimHSLY24
Ju-ho Kim, Jungwoo Heo, Hyun-seo Shin, Chan-yeong Lim, Ha-Jin Yu:
FA-ExU-Net: The Simultaneous Training of an Embedding Extractor and Enhancement Model for a Speaker Verification System Robust to Short Noisy Utterances. 2269-2282
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AiL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AiL24
Yang Ai, Zhen-Hua Ling:
Low-Latency Neural Speech Phase Prediction Based on Parallel Estimation Architecture and Anti-Wrapping Losses for Speech Generation Tasks. 2283-2296
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiLSTH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiLSTH24
Yanxiong Li, Jialong Li, Yongjie Si, Jiaxin Tan, Qianhua He:
Few-Shot Class-Incremental Audio Classification With Adaptive Mitigation of Forgetting and Overfitting. 2297-2311
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GhanaviJ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GhanaviJ24
Reza Ghanavi, Craig T. Jin:
Adjustable Coherent-to-Diffuse Power Estimator for Binaural Speech Enhancement in Multi-Talker Environments. 2312-2323
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/LiuLWL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiuLWL24
Tianchi Liu, Kong Aik Lee, Qiongqiong Wang, Haizhou Li:
Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification. 2324-2337
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoZLLZR24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoZLLZR24
Lei Zhao, Wenbo Zhu, Shengqiang Li, Hong Luo, Xiao-Lei Zhang, Susanto Rahardja:
Multi-Resolution Convolutional Residual Neural Networks for Monaural Speech Dereverberation. 2338-2351
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/GeishauserNLLHFRVG24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GeishauserNLLHFRVG24
Christian Geishauser, Carel van Niekerk, Nurul Lubis, Hsien-Chin Lin, Michael Heck, Shutong Feng, Benjamin Matthias Ruppik, Renato Vukovic, Milica Gasic:
Learning With an Open Horizon in Ever-Changing Dialogue Circumstances. 2352-2366
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ErenGM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ErenGM24
Yusuf Eren, Buket Çolak Güvenç, Engin Cemal Mengüç:
Cost-Effective Acoustic Feedback Cancellers for Digital Hearing Aids. 2367-2377
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/LiHQZHZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/LiHQZHZ24
Jianchen Li, Jiqing Han, Fan Qian, Tieran Zheng, Yongjun He, Guibin Zheng:
Distance Metric-Based Open-Set Domain Adaptation for Speaker Verification. 2378-2390
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/NiizumiTOHK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/NiizumiTOHK24
Daisuke Niizumi, Daiki Takeuchi, Yasunori Ohishi, Noboru Harada, Kunio Kashino:
Masked Modeling Duo: Towards a Universal Audio Pre-Training Framework. 2391-2406
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/SunZW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SunZW24
Guangzhi Sun, Chao Zhang, Philip C. Woodland:
Graph Neural Networks for Contextual ASR With the Tree-Constrained Pointer Generator. 2407-2417
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/BorgstromB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/BorgstromB24
Bengt J. Borgström, Michael S. Brandstein:
A Multiscale Autoencoder (MSAE) Framework for End-to-End Neural Network Speech Enhancement. 2418-2431
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WeiLLLJX24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WeiLLLJX24
Kun Wei, Bei Li, Hang Lv, Quan Lu, Ning Jiang, Lei Xie:
Conversational Speech Recognition by Learning Audio-Textual Cross-Modal Contextual Representation. 2432-2444
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/SuCC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SuCC24
Shang-Yu Su, Yung-Sung Chung, Yun-Nung Chen:
Joint Dual Learning With Mutual Information Maximization for Natural Language Understanding and Generation in Dialogues. 2445-2452
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/FanDTFYWL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/FanDTFYWL24
Cunhang Fan, Mingming Ding, Jianhua Tao, Ruibo Fu, Jiangyan Yi, Zhengqi Wen, Zhao Lv:
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection. 2453-2466
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/TaherianW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/TaherianW24
Hassan Taherian, DeLiang Wang:
Multi-Channel Conversational Speaker Separation via Neural Diarization. 2467-2476
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/AbdulatifCY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/AbdulatifCY24
Sherif Abdulatif, Ruizhe Cao, Bin Yang:
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement. 2477-2493
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/YangHSM24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/YangHSM24
Puhai Yang, Heyan Huang, Shumin Shi, Xian-Ling Mao:
STN4DST: A Scalable Dialogue State Tracking Based on Slot Tagging Navigation. 2494-2507
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChenWDYPL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChenWDYPL24
Hang Chen, Qing Wang, Jun Du, Bao-Cai Yin, Jia Pan, Chin-Hui Lee:
Optimizing Audio-Visual Speech Enhancement Using Multi-Level Distortion Measures for Audio-Visual Speech Recognition. 2508-2521
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/QueirozC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/QueirozC24
Anderson Queiroz, Rosângela Coelho:
Harmonic Detection From Noisy Speech With Auditory Frame Gain for Intelligibility Enhancement. 2522-2531
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HuQCZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HuQCZ24
Maodi Hu, Li Qian, Zhijun Chang, Zhixiong Zhang:
KDPG-Enhanced MRC Framework for Scientific Entity Recognition in Survey Papers. 2532-2543
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/NortjeOK24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/NortjeOK24
Leanne Nortje, Dan Oneata, Herman Kamper:
Visually Grounded Few-Shot Word Learning in Low-Resource Settings. 2544-2554
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/KamathGWN24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/KamathGWN24
Purnima Kamath, Chitralekha Gupta, Lonce Wyse, Suranga Nanayakkara:
Example-Based Framework for Perceptually Guided Audio Texture Generation. 2555-2565
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/RoyS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/RoyS24
Arka Roy, Udit Satija:
A Novel Multi-Head Self-Organized Operational Neural Network Architecture for Chronic Obstructive Pulmonary Disease Detection Using Lung Sounds. 2566-2575
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/GuL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/GuL24
Rongzhi Gu, Yi Luo:
ReZero: Region-Customizable Sound Extraction. 2576-2589
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/WangSJ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WangSJ24
Wenbin Wang, Yang Song, Sanjay K. Jha:
USAT: A Universal Speaker-Adaptive Text-to-Speech Approach. 2590-2604
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/HanLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/HanLL24
Han Han, Vincent Lostanlen, Mathieu Lagrange:
Learning to Solve Inverse Problems for Perceptual Sound Matching. 2605-2615
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/MamunH24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/MamunH24
Nursadul Mamun, John H. L. Hansen:
Speech Enhancement for Cochlear Implant Recipients Using Deep Complex Convolution Transformer With Frequency Transformation. 2616-2629
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/PengWWSCCY24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/PengWWSCCY24
Cheng Peng, Haobo Wang, Jue Wang, Lidan Shou, Ke Chen, Gang Chen, Chang Yao:
Learning Label-Adaptive Representation for Large-Scale Multi-Label Text Classification. 2630-2640
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ZhaoCW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ZhaoCW24
Junchuan Zhao, Low Qi Hong Chetwin, Ye Wang:
SinTechSVS: A Singing Technique Controllable Singing Voice Synthesis System. 2641-2653
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/OhLL24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/OhLL24
Hyung-Seok Oh, Sang-Hoon Lee, Seong-Whan Lee:
DiffProsody: Diffusion-Based Latent Prosody Generation for Expressive Speech Synthesis With Prosody Conditional Adversarial Training. 2654-2666
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/DamianoBBAS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/DamianoBBAS24
Stefano Damiano, Federico Borra, Alberto Bernardini, Fabio Antonacci, Augusto Sarti:
A Compressive Sensing Approach for the Reconstruction of the Soundfield Produced by Directive Sources in Reverberant Rooms. 2667-2679
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/ChengLZZHS24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ChengLZZHS24
Jiaming Cheng, Ruiyu Liang, Lin Zhou, Li Zhao, Chengwei Huang, Björn W. Schuller:
Residual Fusion Probabilistic Knowledge Distillation for Speech Enhancement. 2680-2691
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/WuDWB24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/WuDWB24
Shih-Lun Wu, Chris Donahue, Shinji Watanabe, Nicholas J. Bryan:
Music ControlNet: Multiple Time-Varying Controls for Music Generation. 2692-2703
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/TuMC24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/TuMC24
Youzhi Tu, Man-Wai Mak, Jen-Tzung Chien:
Contrastive Self-Supervised Speaker Embedding With Sequential Disentanglement. 2704-2715
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/SaxenaA24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/SaxenaA24
Kavya Ranjan Saxena, Vipul Arora:
Interactive Singing Melody Extraction Based on Active Adaptation. 2729-2738
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/JiangPCXW24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/JiangPCXW24
Shuoran Jiang, Youcheng Pan, Qingcai Chen, Yang Xiang, Xiangping Wu:
Learning to Improve Out-of-Distribution Generalization via Self-Adaptive Language Masking. 2739-2750
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/ShirninAPA24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/ShirninAPA24
Alexander Shirnin, Nikita Andreev, Sofia Potapova, Ekaterina Artemova:
Analyzing the Robustness of Vision & Language Models. 2751-2763
- view
  authority control:
- export record
  dblp key:
  - journals/taslp/DingZZWWXWZ24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/DingZZWWXWZ24
Han Ding, Linwei Zhai, Cui Zhao, Fei Wang, Ge Wang, Wei Xi, Zhi Wang, Jizhong Zhao:
Genre Classification Empowered by Knowledge-Embedded Music Representation. 2764-2776
- view
  - electronic edition via DOI (open access)
  - references & citations
  authority control:
- export record
  dblp key:
  - journals/taslp/VioletaMHT24
- ask others
- share record
  persistent URL:
  - https://dblp.org/rec/journals/taslp/VioletaMHT24
Lester Phillip Violeta, Ding Ma, Wen-Chin Huang, Tomoki Toda:
Pretraining and Adaptation Techniques for Electrolaryngeal Speech Recognition. 2777-2789

a service of

manage site settings

To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.