| 751 |
Pardon the Interruption: An Analysis of Gender and Turn-Taking in U.S. Supreme Court Oral Arguments
Haley Lepp, Gina-Anne Levow
|
👻
Ghosted
|
cs.CL
|
2 |
5 years ago |
| 752 |
Emotion Carrier Recognition from Personal Narratives
Aniruddha Tammewar, Alessandra Cervone, Giuseppe Riccardi
|
👻
Ghosted
|
cs.CL
|
2 |
5 years ago |
| 753 |
Efficient MDI Adaptation for n-gram Language Models
Ruizhe Huang, Ke Li, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
2 |
5 years ago |
| 754 |
Weakly Supervised Training of Hierarchical Attention Networks for Speaker Identification
Yanpei Shi, Qiang Huang, Thomas Hain
|
👻
Ghosted
|
eess.AS
|
2 |
6 years ago |
| 755 |
Reverse Transfer Learning: Can Word Embeddings Trained for Different NLP Tasks Improve Neural Language Models?
Lyan Verwimp, Jerome R. Bellegarda
|
👻
Ghosted
|
cs.CL
|
2 |
6 years ago |
| 756 |
Avaya Conversational Intelligence: A Real-Time System for Spoken Language Understanding in Human-Human Call Center Conversations
Jan Mizgajski, Adrian Szymczak, ... (+15 more)
|
👻
Ghosted
|
eess.AS
|
2 |
6 years ago |
| 757 |
Large-Scale Speaker Diarization of Radio Broadcast Archives
Emre Yılmaz, Adem Derinel, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 758 |
Sampling from Stochastic Finite Automata with Applications to CTC Decoding
Martin Jansche, Alexander Gutkin
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 759 |
M2H-GAN: A GAN-based Mapping from Machine to Human Transcripts for Speech Understanding
Titouan Parcollet, Mohamed Morchid, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 760 |
Performance Monitoring for End-to-End Speech Recognition
Ruizhi Li, Gregory Sell, Hynek Hermansky
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 761 |
Exploring Methods for the Automatic Detection of Errors in Manual Transcription
Xiaofei Wang, Jinyi Yang, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 762 |
The Trajectory of Voice Onset Time with Vocal Aging
Xuanda Chen, Ziyu Xiong, Jian Hu
|
👻
Ghosted
|
cs.SD
|
2 |
7 years ago |
| 763 |
Automatic Pronunciation Generation by Utilizing a Semi-supervised Deep Neural Networks
Naoya Takahashi, Tofigh Naghibi, Beat Pfister
|
👻
Ghosted
|
cs.CL
|
2 |
10 years ago |
| 764 |
Song Form-aware Full-Song Text-to-Lyrics Generation with Multi-Level Granularity Syllable Count Control
Yunkee Chae, Eunsik Shin, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
2 |
1 year ago |
| 765 |
SecureSpectra: Safeguarding Digital Identity from Deep Fake Threats via Intelligent Signatures
Oguzhan Baser, Kaan Kale, Sandeep P. Chinchali
|
👻
Ghosted
|
cs.CR
|
2 |
2 years ago |
| 766 |
Turbo your multi-modal classification with contrastive learning
Zhiyu Zhang, Da Liu, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
2 |
1 year ago |
| 767 |
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
Haoxiang Shi, Ziqi Liang, Jun Yu
|
👻
Ghosted
|
cs.MM
|
2 |
1 year ago |
| 768 |
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
Jizhen Li, Xinmeng Xu, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
2 |
2 years ago |
| 769 |
MSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization
Adriana Fernandez-Lopez, Honglie Chen, ... (+6 more)
|
👻
Ghosted
|
cs.CV
|
2 |
2 years ago |
| 770 |
FastAST: Accelerating Audio Spectrogram Transformer via Token Merging and Cross-Model Knowledge Distillation
Swarup Ranjan Behera, Abhishek Dhiman, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
2 |
2 years ago |
| 771 |
RIR-SF: Room Impulse Response Based Spatial Feature for Target Speech Recognition in Multi-Channel Multi-Speaker Scenarios
Yiwen Shao, Shi-Xiong Zhang, Dong Yu
|
👻
Ghosted
|
eess.AS
|
2 |
2 years ago |
| 772 |
Improved Factorized Neural Transducer Model For text-only Domain Adaptation
Junzhe Liu, Jianwei Yu, Xie Chen
|
👻
Ghosted
|
cs.CL
|
2 |
2 years ago |
| 773 |
On-Device Speaker Anonymization of Acoustic Embeddings for ASR based onFlexible Location Gradient Reversal Layer
Md Asif Jalal, Pablo Peso Parada, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
2 |
2 years ago |
| 774 |
Streaming Audio-Visual Speech Recognition with Alignment Regularization
Pingchuan Ma, Niko Moritz, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
2 |
3 years ago |
| 775 |
SF-DST: Few-Shot Self-Feeding Reading Comprehension Dialogue State Tracking with Auxiliary Task
Jihyun Lee, Gary Geunbae Lee
|
👻
Ghosted
|
cs.CL
|
2 |
3 years ago |
| 776 |
Deep Sparse Conformer for Speech Recognition
Xianchao Wu
|
👻
Ghosted
|
cs.CL
|
2 |
3 years ago |
| 777 |
Introducing Auxiliary Text Query-modifier to Content-based Audio Retrieval
Daiki Takeuchi, Yasunori Ohishi, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
2 |
4 years ago |
| 778 |
Meta Auxiliary Learning for Low-resource Spoken Language Understanding
Yingying Gao, Junlan Feng, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
2 |
4 years ago |
| 779 |
Neuromorphic Keyword Spotting with Pulse Density Modulation MEMS Microphones
Sidi Yaya Arnaud Yarga, Sean U. N. Wood
|
👻
Ghosted
|
cs.NE
|
1 |
1 year ago |
| 780 |
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
Seohyun Park, Chitralekha Gupta, ... (+4 more)
|
💀
404 Not Found
|
eess.AS
|
1 |
1 year ago |
| 781 |
Gaze-Enhanced Multimodal Turn-Taking Prediction in Triadic Conversations
Seongsil Heo, Calvin Murdock, ... (+2 more)
|
👻
Ghosted
|
cs.HC
|
1 |
1 year ago |
| 782 |
Children's Voice Privacy: First Steps And Emerging Challenges
Ajinkya Kulkarni, Francisco Teixeira, ... (+4 more)
|
👻
Ghosted
|
cs.CY
|
1 |
1 year ago |
| 783 |
Novel Loss-Enhanced Universal Adversarial Patches for Sustainable Speaker Privacy
Elvir Karimov, Alexander Varlamov, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
1 |
1 year ago |
| 784 |
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
Fang Kang, Yin Cao, Haoyu Chen
|
👻
Ghosted
|
cs.SD
|
1 |
12 months ago |
| 785 |
SNIFR : Boosting Fine-Grained Child Harmful Content Detection Through Audio-Visual Alignment with Cascaded Cross-Transformer
Orchid Chetia Phukan, Mohd Mujtaba Akhtar, ... (+9 more)
|
👻
Ghosted
|
eess.AS
|
1 |
1 year ago |
| 786 |
Towards Source Attribution of Singing Voice Deepfake with Multimodal Foundation Models
Orchid Chetia Phukan, Girish, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
1 |
1 year ago |
| 787 |
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning
Le Xu, Chenxing Li, ... (+6 more)
|
👻
Ghosted
|
cs.MM
|
1 |
1 year ago |
| 788 |
Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model
Yong Ren, Chenxing Li, ... (+8 more)
|
👻
Ghosted
|
cs.MM
|
1 |
1 year ago |
| 789 |
PoCaPNet: A Novel Approach for Surgical Phase Recognition Using Speech and X-Ray Images
Kubilay Can Demir, Tobias Weise, ... (+4 more)
|
👻
Ghosted
|
cs.HC
|
1 |
3 years ago |
| 790 |
Getting More for Less: Using Weak Labels and AV-Mixup for Robust Audio-Visual Speaker Verification
Anith Selvakumar, Homa Fashandi
|
👻
Ghosted
|
cs.SD
|
1 |
2 years ago |
| 791 |
CASEIN: Cascading Explicit and Implicit Control for Fine-grained Emotion Intensity Regulation
Yuhao Cui, Xiongwei Wang, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
1 |
3 years ago |
| 792 |
Speech-Image Semantic Alignment Does Not Depend on Any Prior Classification Tasks
Masood S. Mortazavi
|
👻
Ghosted
|
cs.LG
|
1 |
5 years ago |
| 793 |
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time
Sashi Novitasari, Andros Tjandra, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
1 |
5 years ago |
| 794 |
Improving Unsupervised Sparsespeech Acoustic Models with Categorical Reparameterization
Benjamin Milde, Chris Biemann
|
👻
Ghosted
|
eess.AS
|
1 |
6 years ago |
| 795 |
Exploration of Audio Quality Assessment and Anomaly Localisation Using Attention Models
Qiang Huang, Thomas Hain
|
👻
Ghosted
|
eess.AS
|
1 |
6 years ago |
| 796 |
Shouted Speech Compensation for Speaker Verification Robust to Vocal Effort Conditions
Santi Prieto, Alfonso Ortega, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
1 |
5 years ago |
| 797 |
Real to H-space Encoder for Speech Recognition
Titouan Parcollet, Mohamed Morchid, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
1 |
7 years ago |
| 798 |
Multiple topic identification in telephone conversations
Xavier Bost, Marc El Bèze, Renato De Mori
|
👻
Ghosted
|
cs.CL
|
1 |
7 years ago |
| 799 |
Joint Learning of Interactive Spoken Content Retrieval and Trainable User Simulator
Pei-Hung Chung, Kuan Tung, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
1 |
8 years ago |
| 800 |
Play Duration based User-Entity Affinity Modeling in Spoken Dialog System
Bo Xiao, Nicholas Monath, ... (+2 more)
|
👻
Ghosted
|
cs.IR
|
1 |
8 years ago |