💀 The Wall of Shame

The most cited papers with no code. Sorted by the weight of their sins.

Page 9, showing 50 papers

# Paper Cause of Death Category Citations Published
401 One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
Hyun Myung Kim, Kangwook Jang, Hoirin Kim
👻 Ghosted eess.AS 17 2 years ago
402 UserLibri: A Dataset for ASR Personalization Using Only Text
Theresa Breiner, Swaroop Ramaswamy, ... (+7 more)
👻 Ghosted eess.AS 17 4 years ago
403 Self-supervised Context-aware Style Representation for Expressive Speech Synthesis
Yihan Wu, Xi Wang, ... (+4 more)
👻 Ghosted cs.SD 17 4 years ago
404 Data augmentation using prosody and false starts to recognize non-native children's speech
Hemant Kathania, Mittul Singh, ... (+2 more)
👻 Ghosted eess.AS 16 5 years ago
405 Learning to Detect Bipolar Disorder and Borderline Personality Disorder with Language and Speech in Non-Clinical Interviews
Bo Wang, Yue Wu, ... (+5 more)
👻 Ghosted cs.LG 16 5 years ago
406 Mitigating Noisy Inputs for Question Answering
Denis Peskov, Joe Barrow, ... (+3 more)
👻 Ghosted cs.CL 16 6 years ago
407 Analyzing Verbal and Nonverbal Features for Predicting Group Performance
Uliyana Kubasova, Gabriel Murray, McKenzie Braley
👻 Ghosted eess.AS 16 7 years ago
408 Privacy-Preserving Speaker Recognition with Cohort Score Normalisation
Andreas Nautsch, Jose Patino, ... (+6 more)
👻 Ghosted eess.AS 16 7 years ago
409 Twin Regularization for online speech recognition
Mirco Ravanelli, Dmitriy Serdyuk, Yoshua Bengio
👻 Ghosted eess.AS 16 8 years ago
410 Multi-Head Decoder for End-to-End Speech Recognition
Tomoki Hayashi, Shinji Watanabe, ... (+2 more)
👻 Ghosted cs.CL 16 8 years ago
411 Dual Language Models for Code Switched Speech Recognition
Saurabh Garg, Tanmay Parekh, Preethi Jyothi
👻 Ghosted cs.CL 16 8 years ago
412 Jointly Trained Sequential Labeling and Classification by Sparse Attention Neural Networks
Mingbo Ma, Kai Zhao, ... (+3 more)
👻 Ghosted cs.CL 16 8 years ago
413 Prosodic Event Recognition using Convolutional Neural Networks with Context Information
Sabrina Stehwien, Ngoc Thang Vu
👻 Ghosted cs.CL 16 9 years ago
414 Automatic Genre and Show Identification of Broadcast Media
Mortaza Doulaty, Oscar Saz, ... (+2 more)
👻 Ghosted cs.MM 16 10 years ago
415 Optimizing Speech-Input Length for Speaker-Independent Depression Classification
Tomasz Rutowski, Amir Harati, ... (+2 more)
👻 Ghosted cs.CL 16 1 year ago
416 Fake the Real: Backdoor Attack on Deep Speech Classification via Voice Conversion
Zhe Ye, Terui Mao, ... (+2 more)
👻 Ghosted cs.SD 16 3 years ago
417 There is more than one kind of robustness: Fooling Whisper with adversarial examples
Raphael Olivier, Bhiksha Raj
👻 Ghosted eess.AS 16 3 years ago
418 Average Token Delay: A Latency Metric for Simultaneous Translation
Yasumasa Kano, Katsuhito Sudoh, Satoshi Nakamura
👻 Ghosted cs.CL 16 3 years ago
419 Evaluating context-invariance in unsupervised speech representations
Mark Hallap, Emmanuel Dupoux, Ewan Dunbar
👻 Ghosted cs.CL 16 3 years ago
420 A Multi-Stage Multi-Codebook VQ-VAE Approach to High-Performance Neural TTS
Haohan Guo, Fenglong Xie, ... (+3 more)
👻 Ghosted cs.SD 16 3 years ago
421 Text-driven Emotional Style Control and Cross-speaker Style Transfer in Neural TTS
Yookyung Shin, Younggun Lee, ... (+3 more)
👻 Ghosted cs.CL 16 4 years ago
422 Simple and Effective Multi-sentence TTS with Expressive and Coherent Prosody
Peter Makarov, Ammar Abbas, ... (+6 more)
👻 Ghosted eess.AS 16 4 years ago
423 Wav2Vec-Aug: Improved self-supervised training with limited data
Anuroop Sriram, Michael Auli, Alexei Baevski
👻 Ghosted cs.CL 16 4 years ago
424 VoiceMe: Personalized voice generation in TTS
Pol van Rijn, Silvan Mertes, ... (+6 more)
👻 Ghosted cs.SD 16 4 years ago
425 Robust Audio Anti-Spoofing with Fusion-Reconstruction Learning on Multi-Order Spectrograms
Penghui Wen, Kun Hu, ... (+4 more)
👻 Ghosted cs.SD 15 2 years ago
426 HierVST: Hierarchical Adaptive Zero-shot Voice Style Transfer
Sang-Hoon Lee, Ha-Yeong Choi, ... (+2 more)
👻 Ghosted cs.SD 15 2 years ago
427 On the Role of Style in Parsing Speech with Neural Models
Trang Tran, Jiahong Yuan, ... (+2 more)
👻 Ghosted cs.CL 15 5 years ago
428 What the Future Brings: Investigating the Impact of Lookahead for Incremental Neural TTS
Brooke Stephenson, Laurent Besacier, ... (+2 more)
👻 Ghosted eess.AS 15 5 years ago
429 Parallel Rescoring with Transformer for Streaming On-Device Speech Recognition
Wei Li, James Qin, ... (+3 more)
👻 Ghosted eess.AS 15 5 years ago
430 Transformer with Bidirectional Decoder for Speech Recognition
Xi Chen, Songyang Zhang, ... (+3 more)
👻 Ghosted eess.AS 15 5 years ago
431 Word Error Rate Estimation Without ASR Output: e-WER2
Ahmed Ali, Steve Renals
👻 Ghosted eess.AS 15 5 years ago
432 Neural Architecture Search on Acoustic Scene Classification
Jixiang Li, Chuming Liang, ... (+4 more)
👻 Ghosted cs.SD 15 6 years ago
433 Challenging the Boundaries of Speech Recognition: The MALACH Corpus
Michael Picheny, Zóltan Tüske, ... (+4 more)
👻 Ghosted cs.CL 15 6 years ago
434 RadioTalk: a large-scale corpus of talk radio transcripts
Doug Beeferman, William Brannon, Deb Roy
👻 Ghosted cs.CL 15 7 years ago
435 Leveraging Acoustic Cues and Paralinguistic Embeddings to Detect Expression from Voice
Vikramjit Mitra, Sue Booker, ... (+7 more)
👻 Ghosted cs.CL 15 7 years ago
436 Speaker-Targeted Audio-Visual Models for Speech Recognition in Cocktail-Party Environments
Guan-Lin Chao, William Chan, Ian Lane
👻 Ghosted eess.AS 15 7 years ago
437 An Incremental Turn-Taking Model For Task-Oriented Dialog Systems
Andrei C. Coman, Koichiro Yoshino, ... (+3 more)
👻 Ghosted cs.CL 15 7 years ago
438 Building African Voices
Perez Ogayo, Graham Neubig, Alan W Black
👻 Ghosted cs.CL 15 4 years ago
439 STUDIES: Corpus of Japanese Empathetic Dialogue Speech Towards Friendly Voice Agent
Yuki Saito, Yuto Nishimura, ... (+3 more)
👻 Ghosted cs.SD 15 4 years ago
440 Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training
Sung-Feng Huang, Shun-Po Chuang, ... (+4 more)
👻 Ghosted cs.SD 14 5 years ago
441 Domain Adversarial Neural Networks for Dysarthric Speech Recognition
Dominika Woszczyk, Stavros Petridis, David Millard
👻 Ghosted cs.SD 14 5 years ago
442 Speaker-Utterance Dual Attention for Speaker and Utterance Verification
Tianchi Liu, Rohan Kumar Das, ... (+3 more)
👻 Ghosted eess.AS 14 5 years ago
443 Laughter Synthesis: Combining Seq2seq modeling with Transfer Learning
Noé Tits, Kevin El Haddad, Thierry Dutoit
👻 Ghosted eess.AS 14 5 years ago
444 Peking Opera Synthesis via Duration Informed Attention Network
Yusong Wu, Shengchen Li, ... (+5 more)
👻 Ghosted eess.AS 14 5 years ago
445 Semantic Complexity in End-to-End Spoken Language Understanding
Joseph P. McKenna, Samridhi Choudhary, ... (+3 more)
👻 Ghosted cs.CL 14 5 years ago
446 Dr.VOT : Measuring Positive and Negative Voice Onset Time in the Wild
Yosi Shrem, Matthew Goldrick, Joseph Keshet
👻 Ghosted eess.AS 14 6 years ago
447 Towards an Unsupervised Entrainment Distance in Conversational Speech using Deep Neural Networks
Md Nasir, Brian Baucom, ... (+2 more)
👻 Ghosted eess.AS 14 8 years ago
448 Sparsely Connected and Disjointly Trained Deep Neural Networks for Low Resource Behavioral Annotation: Acoustic Classification in Couples' Therapy
Haoqi Li, Brian Baucom, Panayiotis Georgiou
👻 Ghosted cs.LG 14 10 years ago
449 ConvRNN-T: Convolutional Augmented Recurrent Neural Network Transducers for Streaming Speech Recognition
Martin Radfar, Rohit Barnwal, ... (+5 more)
👻 Ghosted cs.SD 14 3 years ago
450 Non-autoregressive Error Correction for CTC-based ASR with Phone-conditioned Masked LM
Hayato Futami, Hirofumi Inaguma, ... (+4 more)
👻 Ghosted cs.CL 14 3 years ago