| 1451 |
A scalable convolutional neural network for task-specified scenarios via knowledge distillation
Mengnan Shi, Fei Qin, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
5 |
9 years ago |
| 1452 |
Privacy Preserving Distance Computation using Somewhat-trusted Third Parties
Abelino Jimenez, Bhiksha Raj
|
👻
Ghosted
|
cs.CR
|
5 |
9 years ago |
| 1453 |
Track selection in Multifunction Radars for Multi-target tracking: an Anti-Coordination game
Nikola Bogdanović, Hans Driessen, Alexander Yarovoy
|
👻
Ghosted
|
cs.MA
|
5 |
10 years ago |
| 1454 |
Fast keypoint detection in video sequences
Luca Baroffio, Matteo Cesana, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
11 years ago |
| 1455 |
Enhancing Automatically Discovered Multi-level Acoustic Patterns Considering Context Consistency With Applications in Spoken Term Detection
Cheng-Tao Chung, Wei-Ning Hsu, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
5 |
10 years ago |
| 1456 |
Do Multimodal Language Models Really Understand Direction? A Benchmark for Compass Direction Reasoning
Hang Yin, Zhifeng Lin, ... (+3 more)
|
👻
Ghosted
|
cs.AI
|
5 |
1 year ago |
| 1457 |
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
Zhenqi Jia, Rui Liu
|
⚰️
The Empty Tomb
|
cs.CL
|
5 |
1 year ago |
| 1458 |
Adapting Whisper for Code-Switching through Encoding Refining and Language-Aware Decoding
Jiahui Zhao, Hao Shi, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
5 |
1 year ago |
| 1459 |
Enhancing Multilingual ASR for Unseen Languages via Language Embedding Modeling
Shao-Syuan Huang, Kuan-Po Huang, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
5 |
1 year ago |
| 1460 |
SONIQUE: Video Background Music Generation Using Unpaired Audio-Visual Data
Liqian Zhang, Magdalena Fuentes
|
👻
Ghosted
|
cs.SD
|
5 |
1 year ago |
| 1461 |
Sequential Contrastive Audio-Visual Learning
Ioannis Tsiamas, Santiago Pascual, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
5 |
2 years ago |
| 1462 |
LCB-net: Long-Context Biasing for Audio-Visual Speech Recognition
Fan Yu, Haoxu Wang, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
5 |
2 years ago |
| 1463 |
Scalable Ensemble-based Detection Method against Adversarial Attacks for speaker verification
Haibin Wu, Heng-Cheng Kuo, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
5 |
2 years ago |
| 1464 |
Dual-stream contrastive predictive network with joint handcrafted feature view for SAR ship classification
Xianting Feng, Hao zheng, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
5 |
2 years ago |
| 1465 |
Language-guided Few-shot Semantic Segmentation
Jing Wang, Yuang Liu, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
2 years ago |
| 1466 |
Scale-aware competition network for palmprint recognition
Chengrui Gao, Ziyuan Yang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
2 years ago |
| 1467 |
SparseSpikformer: A Co-Design Framework for Token and Weight Pruning in Spiking Transformer
Yue Liu, Shanlin Xiao, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
2 years ago |
| 1468 |
Cross-Modal Multi-Tasking for Speech-to-Text Translation via Hard Parameter Sharing
Brian Yan, Xuankai Chang, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
5 |
2 years ago |
| 1469 |
Does Video Summarization Require Videos? Quantifying the Effectiveness of Language in Video Summarization
Yoonsoo Nam, Adam Lehavi, ... (+4 more)
|
👻
Ghosted
|
cs.AI
|
5 |
2 years ago |
| 1470 |
On the privacy of federated Clustering: A Cryptographic View
Qiongxiu Li, Lixia Luo
|
👻
Ghosted
|
cs.CR
|
5 |
2 years ago |
| 1471 |
Textless Speech-to-Music Retrieval Using Emotion Similarity
SeungHeon Doh, Minz Won, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
5 |
3 years ago |
| 1472 |
Role of Bias Terms in Dot-Product Attention
Mahdi Namazifar, Devamanyu Hazarika, Dilek Hakkani-Tur
|
👻
Ghosted
|
cs.NE
|
5 |
3 years ago |
| 1473 |
SigVIC: Spatial Importance Guided Variable-Rate Image Compression
Jiaming Liang, Meiqin Liu, ... (+3 more)
|
👻
Ghosted
|
eess.IV
|
5 |
3 years ago |
| 1474 |
Confidence-based Event-centric Online Video Question Answering on a Newly Constructed ATBS Dataset
Weikai Kong, Shuhong Ye, ... (+2 more)
|
👻
Ghosted
|
cs.MM
|
5 |
3 years ago |
| 1475 |
Calibrating AI Models for Wireless Communications via Conformal Prediction
Kfir M. Cohen, Sangwoo Park, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
5 |
3 years ago |
| 1476 |
Output-Dependent Gaussian Process State-Space Model
Zhidi Lin, Lei Cheng, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
5 |
3 years ago |
| 1477 |
New Interpretable Patterns and Discriminative Features from Brain Functional Network Connectivity Using Dictionary Learning
Fateme Ghayem, Hanlu Yang, ... (+4 more)
|
👻
Ghosted
|
q-bio.NC
|
5 |
3 years ago |
| 1478 |
Space-Time Graph Neural Networks with Stochastic Graph Perturbations
Samar Hadou, Charilaos Kanatsoulis, Alejandro Ribeiro
|
👻
Ghosted
|
cs.LG
|
5 |
3 years ago |
| 1479 |
Enhanced Low-resolution LiDAR-Camera Calibration Via Depth Interpolation and Supervised Contrastive Learning
Zhikang Zhang, Zifan Yu, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
5 |
3 years ago |
| 1480 |
Robust Monocular Localization of Drones by Adapting Domain Maps to Depth Prediction Inaccuracies
Priyesh Shukla, Sureshkumar S., ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
5 |
3 years ago |
| 1481 |
Play It Back: Iterative Attention for Audio Recognition
Alexandros Stergiou, Dima Damen
|
👻
Ghosted
|
cs.SD
|
5 |
3 years ago |
| 1482 |
Training set cleansing of backdoor poisoning by self-supervised representation learning
H. Wang, S. Karami, ... (+7 more)
|
👻
Ghosted
|
cs.LG
|
5 |
3 years ago |
| 1483 |
MAST: Multiscale Audio Spectrogram Transformers
Sreyan Ghosh, Ashish Seth, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
5 |
3 years ago |
| 1484 |
Unsupervised Fine-Tuning Data Selection for ASR Using Self-Supervised Speech Models
Reem Gody, David Harwath
|
👻
Ghosted
|
eess.AS
|
5 |
3 years ago |
| 1485 |
High-resolution embedding extractor for speaker diarisation
Hee-Soo Heo, Youngki Kwon, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
5 |
3 years ago |
| 1486 |
Disentangled and Robust Representation Learning for Bragging Classification in Social Media
Xiang Li, Yucheng Zhou
|
👻
Ghosted
|
cs.CL
|
5 |
3 years ago |
| 1487 |
Mitigating Unintended Memorization in Language Models via Alternating Teaching
Zhe Liu, Xuedong Zhang, Fuchun Peng
|
👻
Ghosted
|
cs.CL
|
5 |
3 years ago |
| 1488 |
ARM 4-BIT PQ: SIMD-based Acceleration for Approximate Nearest Neighbor Search on ARM
Yusuke Matsui, Yoshiki Imaizumi, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
5 |
4 years ago |
| 1489 |
Improving Label-Deficient Keyword Spotting Through Self-Supervised Pretraining
Holger Severin Bovbjerg, Zheng-Hua Tan
|
👻
Ghosted
|
cs.SD
|
5 |
3 years ago |
| 1490 |
Exploring the Implicit Semantic Ability of Multimodal Large Language Models: A Pilot Study on Entity Set Expansion
Hebin Wang, Yangning Li, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
4 |
1 year ago |
| 1491 |
An End-to-End Neural Network for Image-to-Audio Transformation
Liu Chen, Michael Deisher, Munir Georges
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1492 |
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
Zhichen Han, Tianqi Geng, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
4 |
1 year ago |
| 1493 |
Joint Task Offloading and Routing in Wireless Multi-hop Networks Using Biased Backpressure Algorithm
Zhongyuan Zhao, Jake Perazzone, ... (+4 more)
|
👻
Ghosted
|
cs.NI
|
4 |
1 year ago |
| 1494 |
Privacy-Preserving Distributed Maximum Consensus Without Accuracy Loss
Wenrui Yu, Richard Heusdens, ... (+2 more)
|
👻
Ghosted
|
cs.DC
|
4 |
1 year ago |
| 1495 |
Fairness-Aware Job Scheduling for Multi-Job Federated Learning
Yuxin Shi, Han Yu
|
👻
Ghosted
|
cs.LG
|
4 |
2 years ago |
| 1496 |
Age of Gossip with the Push-Pull Protocol
Arunabh Srivastava, Thomas Jacob Maranzatto, Sennur Ulukus
|
👻
Ghosted
|
cs.IT
|
4 |
1 year ago |
| 1497 |
Personalized Over-the-Air Federated Learning with Personalized Reconfigurable Intelligent Surfaces
Jiayu Mao, Aylin Yener
|
👻
Ghosted
|
cs.IT
|
4 |
2 years ago |
| 1498 |
Two-in-One: Unified Multi-Person Interactive Motion Generation by Latent Diffusion Transformer
Boyuan Li, Xihua Wang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 1499 |
Optimum Power-Subcarrier Allocation and Time-Sharing in Multicarrier NOMA Uplink
Sagnik Bhattacharya, Kamyar Rajabalifardi, ... (+2 more)
|
👻
Ghosted
|
eess.SP
|
4 |
1 year ago |
| 1500 |
TA-V2A: Textually Assisted Video-to-Audio Generation
Yuhuan You, Xihong Wu, Tianshu Qu
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |