| 301 |
Generative Adversarial Networks for Unpaired Voice Transformation on Impaired Speech
Li-Wei Chen, Hung-Yi Lee, Yu Tsao
|
👻
Ghosted
|
eess.AS
|
26 |
7 years ago |
| 302 |
Contextual Language Model Adaptation for Conversational Agents
Anirudh Raju, Behnam Hedayatnia, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
26 |
8 years ago |
| 303 |
A Study of Enhancement, Augmentation, and Autoencoder Methods for Domain Adaptation in Distant Speech Recognition
Hao Tang, Wei-Ning Hsu, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
26 |
8 years ago |
| 304 |
Scalable Factorized Hierarchical Variational Autoencoder Training
Wei-Ning Hsu, James Glass
|
👻
Ghosted
|
stat.ML
|
26 |
8 years ago |
| 305 |
Recurrent Models for Auditory Attention in Multi-Microphone Distance Speech Recognition
Suyoun Kim, Ian Lane
|
👻
Ghosted
|
cs.LG
|
26 |
10 years ago |
| 306 |
Improving End-to-End Speech-to-Intent Classification with Reptile
Yusheng Tian, Philip John Gorinski
|
👻
Ghosted
|
cs.CL
|
25 |
5 years ago |
| 307 |
ASAPP-ASR: Multistream CNN and Self-Attentive SRU for SOTA Speech Recognition
Jing Pan, Joshua Shapiro, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
25 |
6 years ago |
| 308 |
That Sounds Familiar: an Analysis of Phonetic Representations Transfer Across Languages
Piotr Żelasko, Laureano Moro-Velázquez, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
25 |
6 years ago |
| 309 |
Completely Unsupervised Speech Recognition By A Generative Adversarial Network Harmonized With Iteratively Refined Hidden Markov Models
Kuan-Yu Chen, Che-Ping Tsai, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
25 |
7 years ago |
| 310 |
An Efficient Approach to Encoding Context for Spoken Language Understanding
Raghav Gupta, Abhinav Rastogi, Dilek Hakkani-Tur
|
👻
Ghosted
|
cs.CL
|
25 |
8 years ago |
| 311 |
CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection
Yongyi Zang, Jiatong Shi, ... (+9 more)
|
👻
Ghosted
|
eess.AS
|
25 |
2 years ago |
| 312 |
Improving Low Resource Code-switched ASR using Augmented Code-switched TTS
Yash Sharma, Basil Abraham, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
24 |
5 years ago |
| 313 |
Transfer Learning Approaches for Streaming End-to-End Speech Recognition System
Vikas Joshi, Rui Zhao, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
24 |
5 years ago |
| 314 |
Multilingual Jointly Trained Acoustic and Written Word Embeddings
Yushi Hu, Shane Settle, Karen Livescu
|
👻
Ghosted
|
cs.CL
|
24 |
6 years ago |
| 315 |
Sub-Band Knowledge Distillation Framework for Speech Enhancement
Xiang Hao, Shixue Wen, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
24 |
6 years ago |
| 316 |
An Effective End-to-End Modeling Approach for Mispronunciation Detection
Tien-Hong Lo, Shi-Yan Weng, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
24 |
6 years ago |
| 317 |
Multi-modal Sentiment Analysis using Deep Canonical Correlation Analysis
Zhongkai Sun, Prathusha K Sarma, ... (+2 more)
|
👻
Ghosted
|
cs.IR
|
24 |
7 years ago |
| 318 |
Training Multi-Speaker Neural Text-to-Speech Systems using Speaker-Imbalanced Speech Corpora
Hieu-Thi Luong, Xin Wang, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
24 |
7 years ago |
| 319 |
Coherence Models for Dialogue
Alessandra Cervone, Evgeny Stepanov, Giuseppe Riccardi
|
👻
Ghosted
|
cs.CL
|
24 |
8 years ago |
| 320 |
Multitask Learning with CTC and Segmental CRF for Speech Recognition
Liang Lu, Lingpeng Kong, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
24 |
9 years ago |
| 321 |
TALCS: An Open-Source Mandarin-English Code-Switching Corpus and a Speech Recognition Baseline
Chengfei Li, Shuhao Deng, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
24 |
4 years ago |
| 322 |
ATCSpeech: a multilingual pilot-controller speech corpus from real Air Traffic Control environment
Bo Yang, Xianlong Tan, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
23 |
6 years ago |
| 323 |
Two-stage Training for Chinese Dialect Recognition
Zongze Ren, Guofu Yang, Shugong Xu
|
👻
Ghosted
|
cs.CL
|
23 |
6 years ago |
| 324 |
Assessing the Tolerance of Neural Machine Translation Systems Against Speech Recognition Errors
Nicholas Ruiz, Mattia Antonino Di Gangi, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
23 |
7 years ago |
| 325 |
Extending Recurrent Neural Aligner for Streaming End-to-End Speech Recognition in Mandarin
Linhao Dong, Shiyu Zhou, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
23 |
8 years ago |
| 326 |
Small-footprint Deep Neural Networks with Highway Connections for Speech Recognition
Liang Lu, Steve Renals
|
👻
Ghosted
|
cs.CL
|
23 |
10 years ago |
| 327 |
Turn-Taking Prediction for Natural Conversational Speech
Shuo-yiin Chang, Bo Li, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
23 |
3 years ago |
| 328 |
Semi-FedSER: Semi-supervised Learning for Speech Emotion Recognition On Federated Learning using Multiview Pseudo-Labeling
Tiantian Feng, Shrikanth Narayanan
|
👻
Ghosted
|
eess.AS
|
23 |
4 years ago |
| 329 |
Is Everything Fine, Grandma? Acoustic and Linguistic Modeling for Robust Elderly Speech Emotion Recognition
Gizem Soğancıoğlu, Oxana Verkholyak, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
22 |
5 years ago |
| 330 |
Efficient neural speech synthesis for low-resource languages through multilingual modeling
Marcel de Korte, Jaebok Kim, Esther Klabbers
|
👻
Ghosted
|
eess.AS
|
22 |
5 years ago |
| 331 |
A Pyramid Recurrent Network for Predicting Crowdsourced Speech-Quality Ratings of Real-World Signals
Xuan Dong, Donald S. Williamson
|
👻
Ghosted
|
eess.AS
|
22 |
5 years ago |
| 332 |
Contextualizing ASR Lattice Rescoring with Hybrid Pointer Network Language Model
Da-Rong Liu, Chunxi Liu, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
22 |
6 years ago |
| 333 |
Predicting detection filters for small footprint open-vocabulary keyword spotting
Theodore Bluche, Thibault Gisselbrecht
|
👻
Ghosted
|
cs.CL
|
22 |
6 years ago |
| 334 |
Phoneme-Based Contextualization for Cross-Lingual Speech Recognition in End-to-End Models
Ke Hu, Antoine Bruguier, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
22 |
7 years ago |
| 335 |
I4U Submission to NIST SRE 2018: Leveraging from a Decade of Shared Experiences
Kong Aik Lee, Ville Hautamaki, ... (+44 more)
|
👻
Ghosted
|
eess.AS
|
22 |
7 years ago |
| 336 |
Modeling Interpersonal Linguistic Coordination in Conversations using Word Mover's Distance
Md Nasir, Sandeep Nallan Chakravarthula, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
22 |
7 years ago |
| 337 |
Unsupervised Adaptation with Interpretable Disentangled Representations for Distant Conversational Speech Recognition
Wei-Ning Hsu, Hao Tang, James Glass
|
👻
Ghosted
|
cs.CL
|
22 |
8 years ago |
| 338 |
Data-selective Transfer Learning for Multi-Domain Speech Recognition
Mortaza Doulaty, Oscar Saz, Thomas Hain
|
👻
Ghosted
|
cs.LG
|
22 |
10 years ago |
| 339 |
Malafide: a novel adversarial convolutive noise attack against deepfake and spoofing detection systems
Michele Panariello, Wanying Ge, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
22 |
3 years ago |
| 340 |
Defense Against Adversarial Attacks on Audio DeepFake Detection
Piotr Kawa, Marcin Plata, Piotr Syga
|
👻
Ghosted
|
cs.SD
|
22 |
3 years ago |
| 341 |
Attacker Attribution of Audio Deepfakes
Nicolas M. Müller, Franziska Dieckmann, Jennifer Williams
|
👻
Ghosted
|
cs.CR
|
22 |
4 years ago |
| 342 |
Multi-modal embeddings using multi-task learning for emotion recognition
Aparna Khare, Srinivas Parthasarathy, Shiva Sundaram
|
👻
Ghosted
|
cs.CL
|
21 |
5 years ago |
| 343 |
Audio-visual Multi-channel Recognition of Overlapped Speech
Jianwei Yu, Bo Wu, ... (+8 more)
|
👻
Ghosted
|
eess.AS
|
21 |
6 years ago |
| 344 |
S2IGAN: Speech-to-Image Generation via Adversarial Learning
Xinsheng Wang, Tingting Qiao, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
21 |
6 years ago |
| 345 |
Multi-Task Network for Noise-Robust Keyword Spotting and Speaker Verification using CTC-based Soft VAD and Global Query Attention
Myunghun Jung, Youngmoon Jung, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
21 |
6 years ago |
| 346 |
Speech Replay Detection with x-Vector Attack Embeddings and Spectral Features
Jennifer Williams, Joanna Rownicka
|
👻
Ghosted
|
cs.CL
|
21 |
6 years ago |
| 347 |
On the Use/Misuse of the Term 'Phoneme'
Roger K. Moore, Lucy Skidmore
|
👻
Ghosted
|
cs.CL
|
21 |
6 years ago |
| 348 |
Combining Adversarial Training and Disentangled Speech Representation for Robust Zero-Resource Subword Modeling
Siyuan Feng, Tan Lee, Zhiyuan Peng
|
👻
Ghosted
|
eess.AS
|
21 |
7 years ago |
| 349 |
Improving Contextual Recognition of Rare Words with an Alternate Spelling Prediction Model
Jennifer Drexler Fox, Natalie Delworth
|
👻
Ghosted
|
cs.CL
|
21 |
3 years ago |
| 350 |
Improving the Training Recipe for a Robust Conformer-based Hybrid Model
Mohammad Zeineldeen, Jingjing Xu, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
21 |
4 years ago |