| 501 |
Adaptive Axonal Delays in feedforward spiking neural networks for accurate spoken word recognition
Pengfei Sun, Ehsan Eqlimi, ... (+3 more)
|
👻
Ghosted
|
cs.NE
|
29 |
3 years ago |
| 502 |
SSVMR: Saliency-based Self-training for Video-Music Retrieval
Xuxin Cheng, Zhihong Zhu, ... (+3 more)
|
👻
Ghosted
|
cs.MM
|
29 |
3 years ago |
| 503 |
Training Large-Vocabulary Neural Language Models by Private Federated Learning for Resource-Constrained Devices
Mingbin Xu, Congzheng Song, ... (+11 more)
|
👻
Ghosted
|
cs.LG
|
29 |
4 years ago |
| 504 |
Building Blocks for a Complex-Valued Transformer Architecture
Florian Eilers, Xiaoyi Jiang
|
👻
Ghosted
|
cs.LG
|
28 |
3 years ago |
| 505 |
CommIN: Semantic Image Communications as an Inverse Problem with INN-Guided Diffusion Models
Jiakang Chen, Di You, ... (+2 more)
|
👻
Ghosted
|
eess.IV
|
28 |
2 years ago |
| 506 |
VoxBlink: A Large Scale Speaker Verification Dataset on Camera
Yuke Lin, Xiaoyi Qin, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
28 |
2 years ago |
| 507 |
On the Marginal Benefit of Active Learning: Does Self-Supervision Eat Its Cake?
Yao-Chun Chan, Mingchen Li, Samet Oymak
|
👻
Ghosted
|
cs.LG
|
28 |
5 years ago |
| 508 |
Joint Masked CPC and CTC Training for ASR
Chaitanya Talnikar, Tatiana Likhomanenko, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
28 |
5 years ago |
| 509 |
Transformer in action: a comparative study of transformer-based acoustic models for large scale speech recognition applications
Yongqiang Wang, Yangyang Shi, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
28 |
5 years ago |
| 510 |
Fast acoustic scattering using convolutional neural networks
Ziqi Fan, Vibhav Vineet, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
28 |
6 years ago |
| 511 |
Variable Metric Proximal Gradient Method with Diagonal Barzilai-Borwein Stepsize
Youngsuk Park, Sauptik Dhar, ... (+2 more)
|
👻
Ghosted
|
math.OC
|
28 |
6 years ago |
| 512 |
A unified sequence-to-sequence front-end model for Mandarin text-to-speech synthesis
Junjie Pan, Xiang Yin, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
28 |
6 years ago |
| 513 |
Low-Latency Speaker-Independent Continuous Speech Separation
Takuya Yoshioka, Zhuo Chen, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
28 |
7 years ago |
| 514 |
Fast MVAE: Joint separation and classification of mixed sources based on multichannel variational autoencoder with auxiliary classifier
Li Li, Hirokazu Kameoka, Shoji Makino
|
👻
Ghosted
|
cs.LG
|
28 |
7 years ago |
| 515 |
Stochastic Adaptive Neural Architecture Search for Keyword Spotting
Tom Véniat, Olivier Schwander, Ludovic Denoyer
|
👻
Ghosted
|
cs.LG
|
28 |
7 years ago |
| 516 |
Investigations on End-to-End Audiovisual Fusion
Michael Wand, Ngoc Thang Vu, Juergen Schmidhuber
|
👻
Ghosted
|
cs.CV
|
28 |
8 years ago |
| 517 |
Language and Noise Transfer in Speech Enhancement Generative Adversarial Network
Santiago Pascual, Maruchan Park, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
28 |
8 years ago |
| 518 |
Combining Belief Propagation and Successive Cancellation List Decoding of Polar Codes on a GPU Platform
Sebastian Cammerer, Benedikt Leible, ... (+3 more)
|
👻
Ghosted
|
cs.IT
|
28 |
9 years ago |
| 519 |
A Diffusion Kernel LMS algorithm for nonlinear adaptive networks
Symeon Chouvardas, Moez Draief
|
👻
Ghosted
|
cs.IT
|
28 |
10 years ago |
| 520 |
Hybrid multi-layer Deep CNN/Aggregator feature for image classification
Praveen Kulkarni, Joaquin Zepeda, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
28 |
11 years ago |
| 521 |
Unsupervised Discovery of Linguistic Structure Including Two-level Acoustic Patterns Using Three Cascaded Stages of Iterative Optimization
Cheng-Tao Chung, Chun-an Chan, Lin-shan Lee
|
👻
Ghosted
|
cs.CL
|
28 |
10 years ago |
| 522 |
Downlink SINR Balancing in C-RAN under Limited Fronthaul Capacity
Liu Liang, Zhang Rui
|
👻
Ghosted
|
cs.IT
|
28 |
10 years ago |
| 523 |
Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection
Heqing Zou, Meng Shen, ... (+4 more)
|
👻
Ghosted
|
cs.MM
|
28 |
2 years ago |
| 524 |
One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models
Hang Shao, Bei Liu, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
28 |
2 years ago |
| 525 |
Unsupervised vocal dereverberation with diffusion-based generative models
Koichi Saito, Naoki Murata, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
28 |
3 years ago |
| 526 |
Accelerating RNN-T Training and Inference Using CTC guidance
Yongqiang Wang, Zhehuai Chen, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
28 |
3 years ago |
| 527 |
Audio Deep Fake Detection System with Neural Stitching for ADD 2022
Rui Yan, Cheng Wen, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
28 |
4 years ago |
| 528 |
CT-CAPS: Feature Extraction-based Automated Framework for COVID-19 Disease Identification from Chest CT Scans using Capsule Networks
Shahin Heidarian, Parnian Afshar, ... (+5 more)
|
👻
Ghosted
|
eess.IV
|
27 |
5 years ago |
| 529 |
MuSE-ing on the Impact of Utterance Ordering On Crowdsourced Emotion Annotations
Mimansa Jaiswal, Zakaria Aldeneh, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
27 |
7 years ago |
| 530 |
Sequence-to-Sequence ASR Optimization via Reinforcement Learning
Andros Tjandra, Sakriani Sakti, Satoshi Nakamura
|
👻
Ghosted
|
cs.CL
|
27 |
8 years ago |
| 531 |
Training Probabilistic Spiking Neural Networks with First-to-spike Decoding
Alireza Bagheri, Osvaldo Simeone, Bipin Rajendran
|
👻
Ghosted
|
stat.ML
|
27 |
8 years ago |
| 532 |
Partitioned Successive-Cancellation List Decoding of Polar Codes
Seyyed Ali Hashemi, Alexios Balatsoukas-Stimming, ... (+3 more)
|
👻
Ghosted
|
cs.AR
|
27 |
10 years ago |
| 533 |
DiffVoice: Text-to-Speech with Latent Diffusion
Zhijun Liu, Yiwei Guo, Kai Yu
|
👻
Ghosted
|
eess.AS
|
26 |
3 years ago |
| 534 |
Directional ASR: A New Paradigm for E2E Multi-Speaker Speech Recognition with Source Localization
Aswin Shanmugam Subramanian, Chao Weng, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
26 |
5 years ago |
| 535 |
Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
Hirofumi Inaguma, Yosuke Higuchi, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
26 |
5 years ago |
| 536 |
Neural Percussive Synthesis Parameterised by High-Level Timbral Features
António Ramires, Pritish Chandna, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
26 |
6 years ago |
| 537 |
Knowledge Distillation For Recurrent Neural Network Language Modeling With Trust Regularization
Yangyang Shi, Mei-Yuh Hwang, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
26 |
7 years ago |
| 538 |
Learning Affective Correspondence between Music and Image
Gaurav Verma, Eeshan Gunesh Dhekane, Tanaya Guha
|
👻
Ghosted
|
cs.MM
|
26 |
7 years ago |
| 539 |
Bi-Directional Lattice Recurrent Neural Networks for Confidence Estimation
Qiujia Li, Preben Ness, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
26 |
7 years ago |
| 540 |
Sequence-Level Knowledge Distillation for Model Compression of Attention-based Sequence-to-Sequence Speech Recognition
Raden Mu'az Mun'im, Nakamasa Inoue, Koichi Shinoda
|
👻
Ghosted
|
cs.CL
|
26 |
7 years ago |
| 541 |
Gaussian-Constrained training for speaker verification
Lantian Li, Zhiyuan Tang, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
26 |
7 years ago |
| 542 |
Asynchronous Neighbor Discovery Using Coupled Compressive Sensing
Vamsi K. Amalladinne, Krishna R. Narayanan, ... (+2 more)
|
👻
Ghosted
|
eess.SP
|
26 |
7 years ago |
| 543 |
Performance Analysis for Time-of-Arrival Estimation with Oversampled Low-Complexity 1-bit A/D Conversion
Manuel S. Stein
|
👻
Ghosted
|
cs.IT
|
26 |
9 years ago |
| 544 |
Structure-Aware Classification using Supervised Dictionary Learning
Yael Yankelevsky, Michael Elad
|
👻
Ghosted
|
cs.LG
|
26 |
9 years ago |
| 545 |
Boosting Large Language Model for Speech Synthesis: An Empirical Study
Hongkun Hao, Long Zhou, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
26 |
2 years ago |
| 546 |
The Multimodal Information based Speech Processing (MISP) 2022 Challenge: Audio-Visual Diarization and Recognition
Zhe Wang, Shilong Wu, ... (+13 more)
|
👻
Ghosted
|
cs.MM
|
26 |
3 years ago |
| 547 |
Self-Supervised Learning for Speech Enhancement through Synthesis
Bryce Irvin, Marko Stamenovic, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
26 |
3 years ago |
| 548 |
Spatially Selective Deep Non-linear Filters for Speaker Extraction
Kristina Tesch, Timo Gerkmann
|
👻
Ghosted
|
eess.AS
|
26 |
3 years ago |
| 549 |
BATT: Backdoor Attack with Transformation-based Triggers
Tong Xu, Yiming Li, ... (+2 more)
|
👻
Ghosted
|
cs.CR
|
26 |
3 years ago |
| 550 |
CD-FSOD: A Benchmark for Cross-domain Few-shot Object Detection
Wuti Xiong
|
💀
404 Not Found
|
cs.CV
|
26 |
3 years ago |