| 1551 |
Stable-TTS: Stable Speaker-Adaptive Text-to-Speech Synthesis via Prosody Prompting
Wooseok Han, Minki Kang, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
4 |
1 year ago |
| 1552 |
LBPE: Long-token-first Tokenization to Improve Large Language Models
Haoran Lian, Yizhe Xiong, ... (+6 more)
|
👻
Ghosted
|
cs.CL
|
4 |
1 year ago |
| 1553 |
Exploring Acoustic Similarity in Emotional Speech and Music via Self-Supervised Representations
Yujia Sun, Zeyu Zhao, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
4 |
1 year ago |
| 1554 |
Semi-Supervised Cognitive State Classification from Speech with Multi-View Pseudo-Labeling
Yuanchao Li, Zixing Zhang, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
4 |
1 year ago |
| 1555 |
Stimulus Modality Matters: Impact of Perceptual Evaluations from Different Modalities on Speech Emotion Recognition System Performance
Huang-Cheng Chou, Haibin Wu, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
4 |
1 year ago |
| 1556 |
Deep joint source-channel coding for wireless point cloud transmission
Cixiao Zhang, Mufan Liu, ... (+4 more)
|
👻
Ghosted
|
cs.MM
|
4 |
1 year ago |
| 1557 |
Cross Pseudo-Labeling for Semi-Supervised Audio-Visual Source Localization
Yuxin Guo, Shijie Ma, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 1558 |
Microphone Conversion: Mitigating Device Variability in Sound Event Classification
Myeonghoon Ryu, Hongseok Oh, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
4 |
2 years ago |
| 1559 |
AE-Flow: AutoEncoder Normalizing Flow
Jakub Mosiński, Piotr Biliński, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
4 |
2 years ago |
| 1560 |
Domain Adaptive Graph Classification
Siyang Luo, Ziyi Jiang, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
4 |
2 years ago |
| 1561 |
Augment on Manifold: Mixup Regularization with UMAP
Yousef El-Laham, Elizabeth Fons, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
4 |
2 years ago |
| 1562 |
A comprehensive framework for occluded human pose estimation
Linhao Xu, Lin Zhao, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 1563 |
Class-Wise Buffer Management for Incremental Object Detection: An Effective Buffer Training Strategy
Junsu Kim, Sumin Hong, ... (+7 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 1564 |
PhyOT: Physics-informed object tracking in surveillance cameras
Kawisorn Kamtue, Jose M. F. Moura, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 1565 |
On the choice of the optimal temporal support for audio classification with Pre-trained embeddings
Aurian Quelennec, Michel Olvera, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
4 |
2 years ago |
| 1566 |
Cost Aware Untargeted Poisoning Attack against Graph Neural Networks,
Yuwei Han, Yuni Lai, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
4 |
2 years ago |
| 1567 |
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
Bingkang Shi, Xiaodan Zhang, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
4 |
2 years ago |
| 1568 |
In-Context Prompt Editing For Conditional Audio Generation
Ernie Chang, Pin-Jie Lin, ... (+7 more)
|
👻
Ghosted
|
cs.SD
|
4 |
2 years ago |
| 1569 |
Leveraging Timestamp Information for Serialized Joint Streaming Recognition and Translation
Sara Papi, Peidong Wang, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
4 |
2 years ago |
| 1570 |
MSG-BART: Multi-granularity Scene Graph-Enhanced Encoder-Decoder Language Model for Video-grounded Dialogue Generation
Hongcheng Liu, Zhe Chen, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 1571 |
OLKAVS: An Open Large-Scale Korean Audio-Visual Speech Dataset
Jeongkyun Park, Jung-Wook Hwang, ... (+5 more)
|
👻
Ghosted
|
cs.MM
|
4 |
3 years ago |
| 1572 |
On Mini-Batch Training with Varying Length Time Series
Brian Kenji Iwana
|
👻
Ghosted
|
cs.LG
|
4 |
3 years ago |
| 1573 |
TriNet: stabilizing self-supervised learning from complete or slow collapse on ASR
Lixin Cao, Jun Wang, ... (+3 more)
|
💀
404 Not Found
|
eess.AS
|
4 |
3 years ago |
| 1574 |
Uncertainty Estimation in Deep Speech Enhancement Using Complex Gaussian Mixture Models
Huajian Fang, Timo Gerkmann
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1575 |
Lattice-Free Sequence Discriminative Training for Phoneme-Based Neural Transducers
Zijian Yang, Wei Zhou, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1576 |
Adaptive ECCM for Mitigating Smart Jammers
Kunal Pattanayak, Shashwat Jain, ... (+2 more)
|
👻
Ghosted
|
eess.SP
|
4 |
3 years ago |
| 1577 |
Convolutional Filtering on Sampled Manifolds
Zhiyang Wang, Luana Ruiz, Alejandro Ribeiro
|
👻
Ghosted
|
eess.SP
|
4 |
3 years ago |
| 1578 |
Filterbank Learning for Noise-Robust Small-Footprint Keyword Spotting
Iván López-Espejo, Ram C. M. C. Shekar, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1579 |
On the robustness of non-intrusive speech quality model by adversarial examples
Hsin-Yi Lin, Huan-Hsin Tseng, Yu Tsao
|
👻
Ghosted
|
cs.SD
|
4 |
3 years ago |
| 1580 |
Exploring Sequence-to-Sequence Transformer-Transducer Models for Keyword Spotting
Beltrán Labrador, Guanlong Zhao, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1581 |
Spatio-Temporal Attention in Multi-Granular Brain Chronnectomes for Detection of Autism Spectrum Disorder
James Orme-Rogers, Ajitesh Srivastava
|
👻
Ghosted
|
q-bio.NC
|
4 |
3 years ago |
| 1582 |
Recursive Estimation of User Intent from Noninvasive Electroencephalography using Discriminative Models
Niklas Smedemark-Margulies, Basak Celik, ... (+3 more)
|
👻
Ghosted
|
eess.SP
|
4 |
3 years ago |
| 1583 |
A novel state connection strategy for quantum computing to represent and compress digital images
Md Ershadul Haque, Manoranjan Paul, Tanmoy Debnath
|
👻
Ghosted
|
quant-ph
|
4 |
3 years ago |
| 1584 |
FlowGrad: Using Motion for Visual Sound Source Localization
Rajsuryan Singh, Pablo Zinemanas, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
4 |
3 years ago |
| 1585 |
Improved Projection Learning for Lower Dimensional Feature Maps
Ilan Price, Jared Tanner
|
👻
Ghosted
|
cs.LG
|
4 |
3 years ago |
| 1586 |
Lexicon-injected Semantic Parsing for Task-Oriented Dialog
Xiaojun Meng, Wenlin Dai, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
4 |
3 years ago |
| 1587 |
Rapid Connectionist Speaker Adaptation
Michael Witbrock, Patrick Haffner
|
👻
Ghosted
|
cs.SD
|
4 |
3 years ago |
| 1588 |
Align, Write, Re-order: Explainable End-to-End Speech Translation via Operation Sequence Generation
Motoi Omachi, Brian Yan, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
4 |
3 years ago |
| 1589 |
Efficient Speech Translation with Dynamic Latent Perceivers
Ioannis Tsiamas, Gerard I. Gállego, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
4 |
3 years ago |
| 1590 |
Privacy-preserving Automatic Speaker Diarization
Francisco Teixeira, Alberto Abad, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
4 |
3 years ago |
| 1591 |
Time Domain Adversarial Voice Conversion for ADD 2022
Cheng Wen, Tingwei Guo, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
4 |
4 years ago |
| 1592 |
Container Localisation and Mass Estimation with an RGB-D Camera
Tommaso Apicella, Giulia Slavic, ... (+3 more)
|
💤
Eternal Rest
|
cs.CV
|
4 |
4 years ago |
| 1593 |
Mitigating Closed-model Adversarial Examples with Bayesian Neural Modeling for Enhanced End-to-End Speech Recognition
Chao-Han Huck Yang, Zeeshan Ahmed, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
4 |
4 years ago |
| 1594 |
Diff4Steer: Steerable Diffusion Prior for Generative Music Retrieval with Semantic Guidance
Xuchan Bao, Judith Yue Li, ... (+6 more)
|
👻
Ghosted
|
cs.SD
|
3 |
1 year ago |
| 1595 |
Memory-Free and Parallel Computation for Quantized Spiking Neural Networks
Dehao Zhang, Shuai Wang, ... (+5 more)
|
👻
Ghosted
|
cs.NE
|
3 |
1 year ago |
| 1596 |
RL-LOGO: Deep Reinforcement Learning Localization for Logo Recognition
Masato Fujitake
|
👻
Ghosted
|
cs.CV
|
3 |
2 years ago |
| 1597 |
Fast DCT+: A Family of Fast Transforms Based on Rank-One Updates of the Path Graph
Samuel Fernández-Menduiña, Eduardo Pavez, Antonio Ortega
|
👻
Ghosted
|
eess.SP
|
3 |
1 year ago |
| 1598 |
BEST-STD: Bidirectional Mamba-Enhanced Speech Tokenization for Spoken Term Detection
Anup Singh, Kris Demuynck, Vipul Arora
|
👻
Ghosted
|
eess.AS
|
3 |
1 year ago |
| 1599 |
A Multi-Label EEG Dataset for Mental Attention State Classification in Online Learning
Huan Liu, Yuzhe Zhang, ... (+4 more)
|
💀
404 Not Found
|
cs.HC
|
3 |
1 year ago |
| 1600 |
A Wasserstein Graph Distance Based on Distributions of Probabilistic Node Embeddings
Michael Scholkemper, Damin Kühn, ... (+4 more)
|
👻
Ghosted
|
cs.CE
|
3 |
2 years ago |