| 851 |
Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak Supervision
Chih-Kai Yang, Kuan-Po Huang, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
15 |
2 years ago |
| 852 |
NCL: Textual Backdoor Defense Using Noise-augmented Contrastive Learning
Shengfang Zhai, Qingni Shen, ... (+5 more)
|
👻
Ghosted
|
cs.CR
|
15 |
3 years ago |
| 853 |
Coarse-to-Fine Covid-19 Segmentation via Vision-Language Alignment
Dandan Shan, Zihan Li, ... (+4 more)
|
👻
Ghosted
|
eess.IV
|
15 |
3 years ago |
| 854 |
DasFormer: Deep Alternating Spectrogram Transformer for Multi/Single-Channel Speech Separation
Shuo Wang, Xiangyu Kong, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
15 |
3 years ago |
| 855 |
Augmenting Transformer-Transducer Based Speaker Change Detection With Token-Level Training Loss
Guanlong Zhao, Quan Wang, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
15 |
3 years ago |
| 856 |
DiffPhase: Generative Diffusion-based STFT Phase Retrieval
Tal Peer, Simon Welker, Timo Gerkmann
|
👻
Ghosted
|
eess.AS
|
15 |
3 years ago |
| 857 |
Spatio-Temporal Hybrid Fusion of CAE and SWIn Transformers for Lung Cancer Malignancy Prediction
Sadaf Khademi, Shahin Heidarian, ... (+5 more)
|
👻
Ghosted
|
eess.IV
|
15 |
3 years ago |
| 858 |
sVAD: A Robust, Low-Power, and Light-Weight Voice Activity Detection with Spiking Neural Networks
Qu Yang, Qianhui Liu, ... (+4 more)
|
👻
Ghosted
|
cs.SD
|
14 |
2 years ago |
| 859 |
Classification-Oriented Semantic Wireless Communications
Emrecan Kutay, Aylin Yener
|
👻
Ghosted
|
cs.IT
|
14 |
2 years ago |
| 860 |
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition
Yu Pan, Yanni Hu, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
14 |
3 years ago |
| 861 |
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models
Apoorv Vyas, Srikanth Madikeri, Hervé Bourlard
|
👻
Ghosted
|
cs.LG
|
14 |
5 years ago |
| 862 |
Node Attribute Completion in Knowledge Graphs with Multi-Relational Propagation
Eda Bayram, Alberto Garcia-Duran, Robert West
|
👻
Ghosted
|
cs.LG
|
14 |
5 years ago |
| 863 |
Small footprint Text-Independent Speaker Verification for Embedded Systems
Julien Balian, Raffaele Tavarone, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
14 |
5 years ago |
| 864 |
Class-Conditional Defense GAN Against End-to-End Speech Attacks
Mohammad Esmaeilpour, Patrick Cardinal, Alessandro Lameiras Koerich
|
👻
Ghosted
|
cs.SD
|
14 |
5 years ago |
| 865 |
Injecting Word Information with Multi-Level Word Adapter for Chinese Spoken Language Understanding
Dechuan Teng, Libo Qin, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
14 |
5 years ago |
| 866 |
Learning to rank music tracks using triplet loss
Laure Prétet, Gaël Richard, Geoffroy Peeters
|
👻
Ghosted
|
cs.IR
|
14 |
6 years ago |
| 867 |
Sparse Beamspace Equalization for Massive MU-MIMO mmWave Systems
Seyed Hadi Mirfarshbafan, Christoph Studer
|
👻
Ghosted
|
eess.SP
|
14 |
6 years ago |
| 868 |
Optimal Laplacian regularization for sparse spectral community detection
Lorenzo Dall'Amico, Romain Couillet, Nicolas Tremblay
|
👻
Ghosted
|
cs.LG
|
14 |
6 years ago |
| 869 |
Interrupted and cascaded permutation invariant training for speech separation
Gene-Ping Yang, Szu-Lin Wu, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
14 |
6 years ago |
| 870 |
Learning Domain Invariant Representations for Child-Adult Classification from Speech
Rimita Lahiri, Manoj Kumar, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
14 |
6 years ago |
| 871 |
CPWC: Contextual Point Wise Convolution for Object Recognition
Pratik Mazumder, Pravendra Singh, Vinay Namboodiri
|
👻
Ghosted
|
cs.CV
|
14 |
6 years ago |
| 872 |
Electro-Magnetic Side-Channel Attack Through Learned Denoising and Classification
Florian Lemarchand, Cyril Marlin, ... (+3 more)
|
👻
Ghosted
|
cs.CR
|
14 |
6 years ago |
| 873 |
Generative Graph Convolutional Network for Growing Graphs
Da Xu, Chuanwei Ruan, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
14 |
7 years ago |
| 874 |
Understanding Deep Neural Networks through Input Uncertainties
Jayaraman J. Thiagarajan, Irene Kim, ... (+2 more)
|
👻
Ghosted
|
stat.ML
|
14 |
7 years ago |
| 875 |
Transcribing Lyrics From Commercial Song Audio: The First Step Towards Singing Content Processing
Che-Ping Tsai, Yi-Lin Tuan, Lin-shan Lee
|
👻
Ghosted
|
cs.SD
|
14 |
8 years ago |
| 876 |
Deep Feed-forward Sequential Memory Networks for Speech Synthesis
Mengxiao Bi, Heng Lu, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
14 |
8 years ago |
| 877 |
Character-level Deep Conflation for Business Data Analytics
Zhe Gan, P. D. Singh, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
14 |
9 years ago |
| 878 |
An Empirical Evaluation of Zero Resource Acoustic Unit Discovery
Chunxi Liu, Jinyi Yang, ... (+8 more)
|
👻
Ghosted
|
cs.CL
|
14 |
9 years ago |
| 879 |
Atomic Norm Based Localization of Far-Field and Near-Field Signals with Generalized Symmetric Arrays
Xiaohuan Wu, Wei-Ping Zhu, Jun Yan
|
👻
Ghosted
|
cs.IT
|
14 |
8 years ago |
| 880 |
Training variance and performance evaluation of neural networks in speech
Ewout van den Berg, Bhuvana Ramabhadran, Michael Picheny
|
👻
Ghosted
|
cs.LG
|
14 |
10 years ago |
| 881 |
Gram Schmidt Based Greedy Hybrid Precoding for Frequency Selective Millimeter Wave MIMO Systems
Ahmed Alkhateeb, Robert W. Heath
|
👻
Ghosted
|
cs.IT
|
14 |
10 years ago |
| 882 |
The intrinsic value of HFO features as a biomarker of epileptic activity
Stephen V. Gliske, Kevin R. Moon, ... (+2 more)
|
👻
Ghosted
|
q-bio.NC
|
14 |
10 years ago |
| 883 |
Kernel Task-Driven Dictionary Learning for Hyperspectral Image Classification
Soheil Bahrampour, Nasser M. Nasrabadi, ... (+2 more)
|
👻
Ghosted
|
stat.ML
|
14 |
11 years ago |
| 884 |
A short-graph Fourier transform via personalized PageRank vectors
Mariano Tepper, Guillermo Sapiro
|
👻
Ghosted
|
cs.SI
|
14 |
10 years ago |
| 885 |
Enhancing Low-Resource ASR through Versatile TTS: Bridging the Data Gap
Guanrou Yang, Fan Yu, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
14 |
1 year ago |
| 886 |
ChangeNet: Multi-Temporal Asymmetric Change Detection Dataset
Deyi Ji, Siqi Gao, ... (+3 more)
|
💀
404 Not Found
|
cs.CV
|
14 |
2 years ago |
| 887 |
Recognition-Guided Diffusion Model for Scene Text Image Super-Resolution
Yuxuan Zhou, Liangcai Gao, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
14 |
2 years ago |
| 888 |
SD-HuBERT: Sentence-Level Self-Distillation Induces Syllabic Organization in HuBERT
Cheol Jun Cho, Abdelrahman Mohamed, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
14 |
2 years ago |
| 889 |
Communication Efficient Private Federated Learning Using Dithering
Burak Hasircioglu, Deniz Gunduz
|
👻
Ghosted
|
cs.LG
|
14 |
2 years ago |
| 890 |
GANStrument: Adversarial Instrument Sound Synthesis with Pitch-invariant Instance Conditioning
Gaku Narita, Junichi Shimizu, Taketo Akama
|
💤
Eternal Rest
|
cs.SD
|
14 |
3 years ago |
| 891 |
Backdoor Defense via Suppressing Model Shortcuts
Sheng Yang, Yiming Li, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
14 |
3 years ago |
| 892 |
Structured State Space Decoder for Speech Recognition and Synthesis
Koichi Miyazaki, Masato Murata, Tomoki Koriyama
|
👻
Ghosted
|
cs.SD
|
14 |
3 years ago |
| 893 |
NeRF-Gaze: A Head-Eye Redirection Parametric Model for Gaze Estimation
Pengwei Yin, Jiawu Dai, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
14 |
3 years ago |
| 894 |
In-Sensor & Neuromorphic Computing are all you need for Energy Efficient Computer Vision
Gourav Datta, Zeyu Liu, ... (+7 more)
|
👻
Ghosted
|
cs.CV
|
14 |
3 years ago |
| 895 |
Neural Fourier Shift for Binaural Speech Rendering
Jin Woo Lee, Kyogu Lee
|
👻
Ghosted
|
eess.AS
|
14 |
3 years ago |
| 896 |
Bridging Speech and Textual Pre-trained Models with Unsupervised ASR
Jiatong Shi, Chan-Jan Hsu, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
14 |
3 years ago |
| 897 |
Learning ASR pathways: A sparse multilingual ASR model
Mu Yang, Andros Tjandra, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
14 |
3 years ago |
| 898 |
MetricBERT: Text Representation Learning via Self-Supervised Triplet Training
Itzik Malkiel, Dvir Ginzburg, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
14 |
3 years ago |
| 899 |
Identifying Source Speakers for Voice Conversion based Spoofing Attacks on Speaker Verification Systems
Danwei Cai, Zexin Cai, Ming Li
|
👻
Ghosted
|
eess.AS
|
14 |
4 years ago |
| 900 |
Optimizing the Consumption of Spiking Neural Networks with Activity Regularization
Simon Narduzzi, Siavash A. Bigdeli, ... (+2 more)
|
👻
Ghosted
|
cs.NE
|
14 |
4 years ago |