| 401 |
One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
Hyun Myung Kim, Kangwook Jang, Hoirin Kim
|
👻
Ghosted
|
eess.AS
|
17 |
2 years ago |
| 402 |
UserLibri: A Dataset for ASR Personalization Using Only Text
Theresa Breiner, Swaroop Ramaswamy, ... (+7 more)
|
👻
Ghosted
|
eess.AS
|
17 |
4 years ago |
| 403 |
Self-supervised Context-aware Style Representation for Expressive Speech Synthesis
Yihan Wu, Xi Wang, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
17 |
4 years ago |
| 404 |
Data augmentation using prosody and false starts to recognize non-native children's speech
Hemant Kathania, Mittul Singh, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
16 |
5 years ago |
| 405 |
Learning to Detect Bipolar Disorder and Borderline Personality Disorder with Language and Speech in Non-Clinical Interviews
Bo Wang, Yue Wu, ... (+5 more)
|
👻
Ghosted
|
cs.LG
|
16 |
5 years ago |
| 406 |
Mitigating Noisy Inputs for Question Answering
Denis Peskov, Joe Barrow, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
16 |
6 years ago |
| 407 |
Analyzing Verbal and Nonverbal Features for Predicting Group Performance
Uliyana Kubasova, Gabriel Murray, McKenzie Braley
|
👻
Ghosted
|
eess.AS
|
16 |
7 years ago |
| 408 |
Privacy-Preserving Speaker Recognition with Cohort Score Normalisation
Andreas Nautsch, Jose Patino, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
16 |
7 years ago |
| 409 |
Twin Regularization for online speech recognition
Mirco Ravanelli, Dmitriy Serdyuk, Yoshua Bengio
|
👻
Ghosted
|
eess.AS
|
16 |
8 years ago |
| 410 |
Multi-Head Decoder for End-to-End Speech Recognition
Tomoki Hayashi, Shinji Watanabe, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
16 |
8 years ago |
| 411 |
Dual Language Models for Code Switched Speech Recognition
Saurabh Garg, Tanmay Parekh, Preethi Jyothi
|
👻
Ghosted
|
cs.CL
|
16 |
8 years ago |
| 412 |
Jointly Trained Sequential Labeling and Classification by Sparse Attention Neural Networks
Mingbo Ma, Kai Zhao, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
16 |
8 years ago |
| 413 |
Prosodic Event Recognition using Convolutional Neural Networks with Context Information
Sabrina Stehwien, Ngoc Thang Vu
|
👻
Ghosted
|
cs.CL
|
16 |
9 years ago |
| 414 |
Automatic Genre and Show Identification of Broadcast Media
Mortaza Doulaty, Oscar Saz, ... (+2 more)
|
👻
Ghosted
|
cs.MM
|
16 |
10 years ago |
| 415 |
Optimizing Speech-Input Length for Speaker-Independent Depression Classification
Tomasz Rutowski, Amir Harati, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
16 |
1 year ago |
| 416 |
Fake the Real: Backdoor Attack on Deep Speech Classification via Voice Conversion
Zhe Ye, Terui Mao, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
16 |
3 years ago |
| 417 |
There is more than one kind of robustness: Fooling Whisper with adversarial examples
Raphael Olivier, Bhiksha Raj
|
👻
Ghosted
|
eess.AS
|
16 |
3 years ago |
| 418 |
Average Token Delay: A Latency Metric for Simultaneous Translation
Yasumasa Kano, Katsuhito Sudoh, Satoshi Nakamura
|
👻
Ghosted
|
cs.CL
|
16 |
3 years ago |
| 419 |
Evaluating context-invariance in unsupervised speech representations
Mark Hallap, Emmanuel Dupoux, Ewan Dunbar
|
👻
Ghosted
|
cs.CL
|
16 |
3 years ago |
| 420 |
A Multi-Stage Multi-Codebook VQ-VAE Approach to High-Performance Neural TTS
Haohan Guo, Fenglong Xie, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
16 |
3 years ago |
| 421 |
Text-driven Emotional Style Control and Cross-speaker Style Transfer in Neural TTS
Yookyung Shin, Younggun Lee, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
16 |
4 years ago |
| 422 |
Simple and Effective Multi-sentence TTS with Expressive and Coherent Prosody
Peter Makarov, Ammar Abbas, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
16 |
4 years ago |
| 423 |
Wav2Vec-Aug: Improved self-supervised training with limited data
Anuroop Sriram, Michael Auli, Alexei Baevski
|
👻
Ghosted
|
cs.CL
|
16 |
4 years ago |
| 424 |
VoiceMe: Personalized voice generation in TTS
Pol van Rijn, Silvan Mertes, ... (+6 more)
|
👻
Ghosted
|
cs.SD
|
16 |
4 years ago |
| 425 |
Robust Audio Anti-Spoofing with Fusion-Reconstruction Learning on Multi-Order Spectrograms
Penghui Wen, Kun Hu, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
15 |
2 years ago |
| 426 |
HierVST: Hierarchical Adaptive Zero-shot Voice Style Transfer
Sang-Hoon Lee, Ha-Yeong Choi, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
15 |
2 years ago |
| 427 |
On the Role of Style in Parsing Speech with Neural Models
Trang Tran, Jiahong Yuan, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
15 |
5 years ago |
| 428 |
What the Future Brings: Investigating the Impact of Lookahead for Incremental Neural TTS
Brooke Stephenson, Laurent Besacier, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
15 |
5 years ago |
| 429 |
Parallel Rescoring with Transformer for Streaming On-Device Speech Recognition
Wei Li, James Qin, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
15 |
5 years ago |
| 430 |
Transformer with Bidirectional Decoder for Speech Recognition
Xi Chen, Songyang Zhang, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
15 |
5 years ago |
| 431 |
Word Error Rate Estimation Without ASR Output: e-WER2
Ahmed Ali, Steve Renals
|
👻
Ghosted
|
eess.AS
|
15 |
5 years ago |
| 432 |
Neural Architecture Search on Acoustic Scene Classification
Jixiang Li, Chuming Liang, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
15 |
6 years ago |
| 433 |
Challenging the Boundaries of Speech Recognition: The MALACH Corpus
Michael Picheny, Zóltan Tüske, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
15 |
6 years ago |
| 434 |
RadioTalk: a large-scale corpus of talk radio transcripts
Doug Beeferman, William Brannon, Deb Roy
|
👻
Ghosted
|
cs.CL
|
15 |
7 years ago |
| 435 |
Leveraging Acoustic Cues and Paralinguistic Embeddings to Detect Expression from Voice
Vikramjit Mitra, Sue Booker, ... (+7 more)
|
👻
Ghosted
|
cs.CL
|
15 |
7 years ago |
| 436 |
Speaker-Targeted Audio-Visual Models for Speech Recognition in Cocktail-Party Environments
Guan-Lin Chao, William Chan, Ian Lane
|
👻
Ghosted
|
eess.AS
|
15 |
7 years ago |
| 437 |
An Incremental Turn-Taking Model For Task-Oriented Dialog Systems
Andrei C. Coman, Koichiro Yoshino, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
15 |
7 years ago |
| 438 |
Building African Voices
Perez Ogayo, Graham Neubig, Alan W Black
|
👻
Ghosted
|
cs.CL
|
15 |
4 years ago |
| 439 |
STUDIES: Corpus of Japanese Empathetic Dialogue Speech Towards Friendly Voice Agent
Yuki Saito, Yuto Nishimura, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
15 |
4 years ago |
| 440 |
Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training
Sung-Feng Huang, Shun-Po Chuang, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
14 |
5 years ago |
| 441 |
Domain Adversarial Neural Networks for Dysarthric Speech Recognition
Dominika Woszczyk, Stavros Petridis, David Millard
|
👻
Ghosted
|
cs.SD
|
14 |
5 years ago |
| 442 |
Speaker-Utterance Dual Attention for Speaker and Utterance Verification
Tianchi Liu, Rohan Kumar Das, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
14 |
5 years ago |
| 443 |
Laughter Synthesis: Combining Seq2seq modeling with Transfer Learning
Noé Tits, Kevin El Haddad, Thierry Dutoit
|
👻
Ghosted
|
eess.AS
|
14 |
5 years ago |
| 444 |
Peking Opera Synthesis via Duration Informed Attention Network
Yusong Wu, Shengchen Li, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
14 |
5 years ago |
| 445 |
Semantic Complexity in End-to-End Spoken Language Understanding
Joseph P. McKenna, Samridhi Choudhary, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
14 |
5 years ago |
| 446 |
Dr.VOT : Measuring Positive and Negative Voice Onset Time in the Wild
Yosi Shrem, Matthew Goldrick, Joseph Keshet
|
👻
Ghosted
|
eess.AS
|
14 |
6 years ago |
| 447 |
Towards an Unsupervised Entrainment Distance in Conversational Speech using Deep Neural Networks
Md Nasir, Brian Baucom, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
14 |
8 years ago |
| 448 |
Sparsely Connected and Disjointly Trained Deep Neural Networks for Low Resource Behavioral Annotation: Acoustic Classification in Couples' Therapy
Haoqi Li, Brian Baucom, Panayiotis Georgiou
|
👻
Ghosted
|
cs.LG
|
14 |
10 years ago |
| 449 |
ConvRNN-T: Convolutional Augmented Recurrent Neural Network Transducers for Streaming Speech Recognition
Martin Radfar, Rohit Barnwal, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
14 |
3 years ago |
| 450 |
Non-autoregressive Error Correction for CTC-based ASR with Phone-conditioned Masked LM
Hayato Futami, Hirofumi Inaguma, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
14 |
3 years ago |