| 801 |
Automatic Measurement of Pre-aspiration
Yaniv Sheena, Míša Hejná, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
1 |
9 years ago |
| 802 |
Unified Microphone Conversion: Many-to-Many Device Mapping via Feature-wise Linear Modulation
Myeonghoon Ryu, Hongseok Oh, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
1 |
1 year ago |
| 803 |
MMSD-Net: Towards Multi-modal Stuttering Detection
Liangyu Nie, Sudarsana Reddy Kadiri, Ruchit Agrawal
|
👻
Ghosted
|
cs.SD
|
1 |
2 years ago |
| 804 |
Sequence-to-Sequence Multi-Modal Speech In-Painting
Mahsa Kadkhodaei Elyaderani, Shahram Shirani
|
👻
Ghosted
|
cs.SD
|
1 |
2 years ago |
| 805 |
Compositional Generalization in Spoken Language Understanding
Avik Ray, Yilin Shen, Hongxia Jin
|
👻
Ghosted
|
cs.CL
|
1 |
2 years ago |
| 806 |
Unsupervised Auditory and Semantic Entrainment Models with Deep Neural Networks
Jay Kejriwal, Stefan Benus, Lina M. Rojas-Barahona
|
👻
Ghosted
|
cs.CL
|
1 |
2 years ago |
| 807 |
Differentially Private Adapters for Parameter Efficient Acoustic Modeling
Chun-Wei Ho, Chao-Han Huck Yang, Sabato Marco Siniscalchi
|
💤
Eternal Rest
|
cs.SD
|
1 |
3 years ago |
| 808 |
A Universal Identity Backdoor Attack against Speaker Verification based on Siamese Network
Haodong Zhao, Wei Du, ... (+2 more)
|
👻
Ghosted
|
cs.CR
|
1 |
3 years ago |
| 809 |
Attention Enhanced Citrinet for Speech Recognition
Xianchao Wu
|
👻
Ghosted
|
cs.CL
|
1 |
3 years ago |
| 810 |
Graph-based Multi-View Fusion and Local Adaptation: Mitigating Within-Household Confusability for Speaker Identification
Long Chen, Yixiong Meng, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
1 |
4 years ago |
| 811 |
The THUEE System Description for the IARPA OpenASR21 Challenge
Jing Zhao, Haoyu Wang, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
1 |
4 years ago |
| 812 |
ClaritySpeech: Dementia Obfuscation in Speech
Dominika Woszczyk, Ranya Aloufi, Soteris Demetriou
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 813 |
Phonetically-Augmented Discriminative Rescoring for Voice Search Error Correction
Christophe Van Gysel, Maggie Wu, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 814 |
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
Pierre Lepagnol, Sahar Ghannay, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 815 |
Decoding Listeners Identity: Person Identification from EEG Signals Using a Lightweight Spiking Transformer
Zheyuan Lin, Siqi Cai, Haizhou Li
|
💀
404 Not Found
|
cs.NE
|
0 |
9 months ago |
| 816 |
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
Runduo Han, Yanxin Hu, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
0 |
10 months ago |
| 817 |
Leveraging Unlabeled Audio-Visual Data in Speech Emotion Recognition using Knowledge Distillation
Varsha Pendyala, Pedro Morgado, William Sethares
|
👻
Ghosted
|
cs.LG
|
0 |
1 year ago |
| 818 |
Multimodal Fusion with Semi-Supervised Learning Minimizes Annotation Quantity for Modeling Videoconference Conversation Experience
Andrew Chang, Chenkai Hu, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 819 |
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
Owais Mujtaba Khanday, Pablo Rodroguez San Esteban, ... (+3 more)
|
👻
Ghosted
|
cs.HC
|
0 |
1 year ago |
| 820 |
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
Rémi Uro, Marie Tahon, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 821 |
Modeling Probabilistic Reduction using Information Theory and Naive Discriminative Learning
Anna Stein, Kevin Tang
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 822 |
Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries
Minyoung Kim, Sehwan Park, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 823 |
Conformer-based Ultrasound-to-Speech Conversion
Ibrahim Ibrahimov, Zainkó Csaba, Gábor Gosztolya
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 824 |
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention
Aditya Srinivas Menon, Raj Prakash Gohil, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 825 |
REWIND: Speech Time Reversal for Enhancing Speaker Representations in Diffusion-based Voice Conversion
Ishan D. Biyani, Nirmesh J. Shah, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 826 |
MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing
Junjie Zheng, Zihao Chen, ... (+6 more)
|
👻
Ghosted
|
cs.MM
|
0 |
1 year ago |
| 827 |
A Unit-based System and Dataset for Expressive Direct Speech-to-Speech Translation
Anna Min, Chenxu Hu, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 828 |
Analyzing Speech Motor Movement using Surface Electromyography in Minimally Verbal Adults with Autism Spectrum Disorder
Wazeer Zulfikar, Nishat Protyasha, ... (+14 more)
|
👻
Ghosted
|
q-bio.NC
|
0 |
2 years ago |
| 829 |
HumanDiffusion: diffusion model using perceptual gradients
Yota Ueda, Shinnosuke Takamichi, ... (+3 more)
|
👻
Ghosted
|
cs.HC
|
0 |
3 years ago |
| 830 |
Deep F-measure Maximization for End-to-End Speech Understanding
Leda Sarı, Mark Hasegawa-Johnson
|
👻
Ghosted
|
eess.AS
|
0 |
5 years ago |
| 831 |
Sentence level estimation of psycholinguistic norms using joint multidimensional annotations
Anil Ramakrishna, Shrikanth Narayanan
|
👻
Ghosted
|
cs.CL
|
0 |
6 years ago |
| 832 |
Investigation of Large-Margin Softmax in Neural Language Modeling
Jingjing Huo, Yingbo Gao, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
6 years ago |
| 833 |
Speech Driven Backchannel Generation using Deep Q-Network for Enhancing Engagement in Human-Robot Interaction
Nusrah Hussain, Engin Erzin, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
0 |
6 years ago |
| 834 |
Deep Speech Denoising with Vector Space Projections
Jeff Hetherly, Paul Gamble, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
0 |
8 years ago |
| 835 |
Entity-Aware Language Model as an Unsupervised Reranker
Mohammad Sadegh Rasooli, Sarangarajan Parthasarathy
|
👻
Ghosted
|
cs.CL
|
0 |
8 years ago |
| 836 |
Blind score normalization method for PLDA based speaker recognition
Danila Doroshin, Nikolay Lubimov, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
0 |
10 years ago |
| 837 |
Generation and Pruning of Pronunciation Variants to Improve ASR Accuracy
Zhenhao Ge, Aravind Ganapathiraju, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
10 years ago |
| 838 |
TheanoLM - An Extensible Toolkit for Neural Network Language Modeling
Seppo Enarvi, Mikko Kurimo
|
👻
Ghosted
|
cs.CL
|
0 |
10 years ago |
| 839 |
GhostRNN: Reducing State Redundancy in RNN with Cheap Operations
Hang Zhou, Xiaoxu Zheng, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 840 |
Analyzing Multimodal Features of Spontaneous Voice Assistant Commands for Mild Cognitive Impairment Detection
Nana Lin, Youxiang Zhu, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 841 |
Relation-based Counterfactual Data Augmentation and Contrastive Learning for Robustifying Natural Language Inference Models
Heerin Yang, Sseung-won Hwang, Jungmin So
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 842 |
A Toolkit for Joint Speaker Diarization and Identification with Application to Speaker-Attributed ASR
Giovanni Morrone, Enrico Zovato, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 843 |
Exploring compressibility of transformer based text-to-music (TTM) models
Vasileios Moschopoulos, Thanasis Kotsiopoulos, ... (+8 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 844 |
Gender Representation in TV and Radio: Automatic Information Extraction methods versus Manual Analyses
David Doukhan, Lena Dodson, ... (+7 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 845 |
Towards Multilingual Audio-Visual Question Answering
Orchid Chetia Phukan, Priyabrata Mallick, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
0 |
2 years ago |
| 846 |
Improving Label Assignments Learning by Dynamic Sample Dropout Combined with Layer-wise Optimization in Speech Separation
Chenyang Gao, Yue Gu, Ivan Marsic
|
👻
Ghosted
|
cs.SD
|
0 |
2 years ago |
| 847 |
Spoken Word2Vec: Learning Skipgram Embeddings from Speech
Mohammad Amaan Sayeed, Hanan Aldarmaki
|
👻
Ghosted
|
cs.CL
|
0 |
2 years ago |
| 848 |
Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition
Yukun Feng, Ming Tu, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
3 years ago |
| 849 |
An Investigation of Indian Native Language Phonemic Influences on L2 English Pronunciations
Shelly Jain, Priyanshi Pal, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
3 years ago |
| 850 |
Application of Knowledge Distillation to Multi-task Speech Representation Learning
Mine Kerpicci, Van Nguyen, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
0 |
3 years ago |