| 1201 |
VMF-SNE: Embedding for Spherical Data
Mian Wang, Dong Wang
|
👻
Ghosted
|
cs.LG
|
8 |
10 years ago |
| 1202 |
Group-blind detection with very large antenna arrays in the presence of pilot contamination
Guido Carlo Ferrante, Giovanni Geraci, ... (+2 more)
|
👻
Ghosted
|
cs.IT
|
8 |
10 years ago |
| 1203 |
Super-Resolution in Phase Space
Ayush Bhandari, Yonina Eldar, Ramesh Raskar
|
👻
Ghosted
|
cs.IT
|
8 |
11 years ago |
| 1204 |
Investigating the Effectiveness of Explainability Methods in Parkinson's Detection from Speech
Eleonora Mancini, Francesco Paissan, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
8 |
1 year ago |
| 1205 |
Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction
Yuan Zhao, Rui Liu, Gaoxiang Cong
|
💀
404 Not Found
|
cs.MM
|
8 |
1 year ago |
| 1206 |
Leveraging Pre-Trained Autoencoders for Interpretable Prototype Learning of Music Audio
Pablo Alonso-Jiménez, Leonardo Pepino, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
8 |
2 years ago |
| 1207 |
Stability of Graph Convolutional Neural Networks through the lens of small perturbation analysis
Lucia Testa, Claudio Battiloro, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
8 |
2 years ago |
| 1208 |
Attention-Driven Multichannel Speech Enhancement in Moving Sound Source Scenarios
Yuzhu Wang, Archontis Politis, Tuomas Virtanen
|
👻
Ghosted
|
eess.AS
|
8 |
2 years ago |
| 1209 |
Learning graphs and simplicial complexes from data
Andrei Buciulea, Elvin Isufi, ... (+2 more)
|
👻
Ghosted
|
eess.SP
|
8 |
2 years ago |
| 1210 |
Double Reverse Regularization Network Based on Self-Knowledge Distillation for SAR Object Classification
Bo Xu, Hao Zheng, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
8 |
2 years ago |
| 1211 |
Recent Advances in Scalable Energy-Efficient and Trustworthy Spiking Neural networks: from Algorithms to Technology
Souvik Kundu, Rui-Jie Zhu, ... (+2 more)
|
📚
The Cartographer
|
cs.AR
|
8 |
2 years ago |
| 1212 |
Fast Word Error Rate Estimation Using Self-Supervised Representations for Speech and Text
Chanho Park, Chengsong Lu, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
8 |
2 years ago |
| 1213 |
Do self-supervised speech and language models extract similar representations as human brain?
Peili Chen, Linyang He, ... (+4 more)
|
👻
Ghosted
|
q-bio.NC
|
8 |
2 years ago |
| 1214 |
Enhancing End-to-End Conversational Speech Translation Through Target Language Context Utilization
Amir Hussein, Brian Yan, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
8 |
2 years ago |
| 1215 |
Two-Stream Joint-Training for Speaker Independent Acoustic-to-Articulatory Inversion
Jianrong Wang, Jinyu Liu, ... (+5 more)
|
💀
404 Not Found
|
cs.SD
|
8 |
3 years ago |
| 1216 |
Deep Unfolded Tensor Robust PCA with Self-supervised Learning
Harry Dong, Megna Shah, ... (+2 more)
|
👻
Ghosted
|
stat.ML
|
8 |
3 years ago |
| 1217 |
Simultaneously Learning Robust Audio Embeddings and balanced Hash codes for Query-by-Example
Anup Singh, Kris Demuynck, Vipul Arora
|
👻
Ghosted
|
eess.AS
|
8 |
3 years ago |
| 1218 |
Reverberation as Supervision for Speech Separation
Rohith Aralikatti, Christoph Boeddeker, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
8 |
3 years ago |
| 1219 |
Identifying Coordination in a Cognitive Radar Network -- A Multi-Objective Inverse Reinforcement Learning Approach
Luke Snow, Vikram Krishnamurthy, Brian M. Sadler
|
👻
Ghosted
|
eess.SP
|
8 |
3 years ago |
| 1220 |
On minimal variations for unsupervised representation learning
Vivien Cabannes, Alberto Bietti, Randall Balestriero
|
👻
Ghosted
|
cs.LG
|
8 |
3 years ago |
| 1221 |
Resource-Efficient Transfer Learning From Speech Foundation Model Using Hierarchical Feature Fusion
Zhouyuan Huo, Khe Chai Sim, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
8 |
3 years ago |
| 1222 |
Audio-visual speech enhancement with a deep Kalman filter generative model
Ali Golmakani, Mostafa Sadeghi, Romain Serizel
|
👻
Ghosted
|
cs.CV
|
8 |
3 years ago |
| 1223 |
Rethinking Implicit Neural Representations for Vision Learners
Yiran Song, Qianyu Zhou, Lizhuang Ma
|
👻
Ghosted
|
cs.CV
|
8 |
3 years ago |
| 1224 |
GeoGCN: Geometric Dual-domain Graph Convolution Network for Point Cloud Denoising
Zhaowei Chen, Peng Li, ... (+5 more)
|
👻
Ghosted
|
cs.CV
|
8 |
3 years ago |
| 1225 |
Multimodal Transformer Distillation for Audio-Visual Synchronization
Xuanjun Chen, Haibin Wu, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
8 |
3 years ago |
| 1226 |
Learning to Locate Visual Answer in Video Corpus Using Question
Bin Li, Yixuan Weng, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
8 |
3 years ago |
| 1227 |
Knowledge Transfer For On-Device Speech Emotion Recognition with Neural Structured Learning
Yi Chang, Zhao Ren, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
8 |
3 years ago |
| 1228 |
Multilevel Transformer For Multimodal Emotion Recognition
Junyi He, Meimei Wu, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
8 |
3 years ago |
| 1229 |
Embedding a Differentiable Mel-cepstral Synthesis Filter to a Neural Speech Synthesis System
Takenori Yoshimura, Shinji Takaki, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
8 |
3 years ago |
| 1230 |
token2vec: A Joint Self-Supervised Pre-training Framework Using Unpaired Speech and Text
Xianghu Yue, Junyi Ao, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
8 |
3 years ago |
| 1231 |
End-to-end Spoken Language Understanding with Tree-constrained Pointer Generator
Guangzhi Sun, Chao Zhang, Philip C. Woodland
|
👻
Ghosted
|
cs.CL
|
8 |
3 years ago |
| 1232 |
Audio-Visual Object Classification for Human-Robot Collaboration
A. Xompero, Y. L. Pang, ... (+4 more)
|
👻
Ghosted
|
cs.MM
|
8 |
4 years ago |
| 1233 |
AutoST: Training-free Neural Architecture Search for Spiking Transformers
Ziqing Wang, Qidong Zhao, ... (+3 more)
|
👻
Ghosted
|
cs.NE
|
7 |
3 years ago |
| 1234 |
Self-distilled Dynamic Fusion Network for Language-based Fashion Retrieval
Yiming Wu, Hangfei Li, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
7 |
2 years ago |
| 1235 |
Deep Variational Privacy Funnel: General Modeling with Applications in Face Recognition
Behrooz Razeghi, Parsa Rahimi, Sébastien Marcel
|
👻
Ghosted
|
cs.CV
|
7 |
2 years ago |
| 1236 |
ERGNN: Spectral Graph Neural Network With Explicitly-Optimized Rational Graph Filters
Guoming Li, Jian Yang, Shangsong Liang
|
👻
Ghosted
|
cs.LG
|
7 |
1 year ago |
| 1237 |
An Ensemble Approach to Short-form Video Quality Assessment Using Multimodal LLM
Wen Wen, Yilin Wang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
7 |
1 year ago |
| 1238 |
Linguistics-Vision Monotonic Consistent Network for Sign Language Production
Xu Wang, Shengeng Tang, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
7 |
1 year ago |
| 1239 |
A Comparative Study of Self-Supervised Speech Representations in Read and Spontaneous TTS
Siyang Wang, Gustav Eje Henter, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
7 |
3 years ago |
| 1240 |
Dynamic Scheduling for Federated Edge Learning with Streaming Data
Chung-Hsuan Hu, Zheng Chen, Erik G. Larsson
|
👻
Ghosted
|
cs.LG
|
7 |
3 years ago |
| 1241 |
Multicast Transmission Design with Enhanced DoF for MIMO Coded Caching Systems
Mohammad NaseriTehrani, MohammadJavad Salehi, Antti Tölli
|
👻
Ghosted
|
cs.IT
|
7 |
3 years ago |
| 1242 |
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
Jeongsoo Choi, Minsu Kim, ... (+2 more)
|
👻
Ghosted
|
cs.GR
|
7 |
3 years ago |
| 1243 |
Low-latency Speech Enhancement via Speech Token Generation
Huaying Xue, Xiulian Peng, Yan Lu
|
👻
Ghosted
|
cs.SD
|
7 |
2 years ago |
| 1244 |
Invertible Mosaic Image Hiding Network for Very Large Capacity Image Steganography
Zihan Chen, Tianrui Liu, ... (+4 more)
|
👻
Ghosted
|
cs.MM
|
7 |
2 years ago |
| 1245 |
Transavs: End-To-End Audio-Visual Segmentation With Transformer
Yuhang Ling, Yuxi Li, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
7 |
3 years ago |
| 1246 |
Efficient Training Data Generation for Phase-Based DOA Estimation
Fabian Hübner, Wolfgang Mack, Emanuël A. P. Habets
|
👻
Ghosted
|
eess.AS
|
7 |
5 years ago |
| 1247 |
Co-attentional Transformers for Story-Based Video Understanding
Björn Bebensee, Byoung-Tak Zhang
|
👻
Ghosted
|
cs.CV
|
7 |
5 years ago |
| 1248 |
Modeling Homophone Noise for Robust Neural Machine Translation
Wenjie Qin, Xiang Li, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
7 |
5 years ago |
| 1249 |
Accent Estimation of Japanese Words from Their Surfaces and Romanizations for Building Large Vocabulary Accent Dictionaries
Hideyuki Tachibana, Yotaro Katayama
|
👻
Ghosted
|
cs.CL
|
7 |
5 years ago |
| 1250 |
Leveraging the structure of musical preference in content-aware music recommendation
Paul Magron, Cédric Févotte
|
👻
Ghosted
|
cs.IR
|
7 |
5 years ago |