


default search action
Computer Speech & Language, Volume 100
Volume 100, 2026
- Alex Peiró Lilja, Carme Armentano-Oller

, José Giraldo, Wendy Elvira-García
, Ignasi Esquerra
, Rodolfo Zevallos
, Cristina España-Bonet
, Martí Llopart-Font
, Baybars Külebi
, Mireia Farrús
:
LaFresCat: A studio-quality Catalan multi-accent speech dataset for text-to-speech synthesis. 101945 - Jin-Seong Choi

, Jae-Hong Lee, Joon-Hyuk Chang
:
Bumper-guided representation interpolation for black-box unsupervised domain adaptation. 101947 - Sharika Tr

, Julia Punitha Malar Dhas:
Attention based convolutional residual squeeze excited capsule network for aspect based sentiment classification in Malayalam movie reviews. 101952 - Abdul Malik Abbasi

, Imtiaz Husain:
Cross-linguistic analysis of prosodic features based on wavelet prominence: A study of L2 English and L1 Sindhi lexical stress using large language & deep learning models. 101953 - Kaustav Das, Biswaranjan Pattanayak, Gayadhar Pradhan:

Modeling the temporal envelope of sub-band signals for improving the performance of children's speech recognition system in zero-resource scenario. 101954 - Yi-Fu Zhao, Guang-Hui Dong, Nan Liu

:
Design of single-channel speech enhancement algorithm in noisy acoustic environments. 101955 - Weizhao Zhang

, Mengjuan Wang, Junzhi Li, Hongwu Yang:
Cro-MTVITS: An end-to-end cross-lingual speech synthesis model for Mandarin and multi-dialect Tibetan based on VITS. 101956 - Shengjie Zhao, Zhenping Xie

:
An exhaustive evaluation method for open-domain LLM dialogue by constructing recursive CoT. 101957 - Zhanghui Liu, Zhang Wentao, Yuzhong Chen

, Lin Yixin:
Dialogue summarization with topic enhancement and factual consistency contrast. 101958 - Da-Hee Yang

, Dail Kim, Joon-Hyuk Chang
, Jeonghwan Choi, Han-Gil Moon:
A dual-branch parallel network for speech enhancement and restoration. 101959 - Musyyab Yousufi

, Rytis Maskeliunas:
Identifying robust and dataset-independent acoustic biomarkers of depression through multi-model feature consensus analysis. 101960 - P. Vijayalakshmi, Anushiya Rachel Gladston, B. Ramani, M. P. Actlin Jeeva, K. Anantha Krishnan, T. Lavanya, T. Nagarajan:

Leveraging synthetic speech: TTS-driven data augmentation for effective dysarthric speech recognition. 101961 - Jiayu Zhang, Hongli Zhang, Chunyu Liu, Zeshu Tian, Chao Meng, Yuxiang Ma:

TriTSP: A triangular joint reasoning networks for target-stance prediction. 101962 - Gabriel Silva

, Mário Rodrigues
, António J. S. Teixeira, Marlene Amorim:
Deepening graph-based approaches for Portuguese open information extraction with LLM augmentation. 101963 - Ibon Vales Cortina

, Owais Mujtaba Khanday
, Marc Ouellet
, José L. Pérez-Córdoba
, Pablo Rodríguez San Esteban
, Laura Miccoli
, Alberto Galdón
, Gonzalo Olivares Granados
, José Andrés González López
:
UGR-MINDVOICE: A multimodal EEG-audio dataset for overt and covert Iberian Spanish speech production. 101964 - Soumya Dutta

, Smruthi Balaji, Sriram Ganapathy:
A Mixture-of-Experts model for multimodal emotion recognition in conversations. 101965 - Gonzalo Nieto Montero

, Santiago Hernández, Juan Casal:
Improvements in Spanish audio transcription workflows: Integrating preprocessing, LLM-based correction, and speaker diarization and identification. 101966 - Sara Barahona

, Juan Ignacio Álvarez-Trejos, Alicia Lozano-Diez
, Daniel Ramos-Castro, Doroteo T. Toledano:
Exploring efficient attention strategies in conformer-based sound event detection. 101967 - Yahao Hu, Wei Tao, Yifei Xie, Tianfeng Wang, Zhisong Pan

:
Incorporating prior knowledge into style embedding for unsupervised text style transfer. 101968 - Minguang Song, Yunxin Zhao

:
Improve NNLMs by text generation from pre-trained language models. 101969 - Larissa Guder

, João Paulo Aires, Hígor Uélinton Silva
, Felipe Meneguzzi
, Dalvan Griebler
:
Sentence representations for semantic textual similarity: A systematic review. 101970 - Jianjun Lei, Zhenmei Mu, Ying Wang

:
HRDF-MER: Hierarchical feature refinement and cascaded dynamic fusion for multimodal emotion recognition. 101978 - Derya Cokal

, Martin Villalba, Rui He, Claudio Flores Palominos, Annkathrin Böke, Philipp Homan, Klaus von Heusinger, Joseph Kambeitz, Wolfram Hinzen:
What is the retest reliability of computationally extractable speech and language markers? 101981 - Tianle Yang

, Chengzhe Sun, Phil Rose, Cassandra L. Jacobs, Siwei Lyu:
Assessing the ability of neural TTS systems to model consonant-induced f0 perturbation. 101983 - Shao-Jung Chan

, Kuang-Yow Lian
:
Scaling multi-speaker speech recognition with high-quality synthetic data. 101984 - Martin Lenglet

, Olivier Perrotin
, Gérard Bailly
:
A closer look at internal representations of end-to-end Text-to-Speech models: How is phonetic and acoustic information encoded? 101985 - Yael Segal-Feldman

, Ann R. Bradlow, Matthew Goldrick, Joseph Keshet
:
Open-vocabulary keyword spotting with hyper-matched filters for small footprint devices. 101986 - Jinyi Mi

, Xiaohan Shi, Ding Ma, Jiajun He, Takuya Fujimura, Tomoki Toda:
Robust speech emotion recognition under human speech noise. 101987 - Natalia A. Tomashenko

, Xiaoxiao Miao, Pierre Champion
, Sarina Meyer
, Michele Panariello
, Xin Wang, Nicholas W. D. Evans
, Emmanuel Vincent
, Junichi Yamagishi, Massimiliano Todisco
:
The third VoicePrivacy challenge: Preserving emotional expressiveness and linguistic content in voice anonymization. 101988 - Yujie Ma, Yunfeng Xu

, Pengwei Wu:
GNN-Transformer cross-view contrastive learning for multimodal conversational emotion recognition. 101989

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID













