| 1451 |
Block Verification Accelerates Speculative Decoding
Ziteng Sun, Uri Mendlovic, ... (+5 more)
|
👻
Ghosted
|
cs.LG
|
19 |
2 years ago |
| 1452 |
Emerging Safety Attack and Defense in Federated Instruction Tuning of Large Language Models
Rui Ye, Jingyi Chai, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
19 |
2 years ago |
| 1453 |
Active Learning for Neural PDE Solvers
Daniel Musekamp, Marimuthu Kalimuthu, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
19 |
1 year ago |
| 1454 |
Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?
Xueru Wen, Jie Lou, ... (+8 more)
|
👻
Ghosted
|
cs.LG
|
19 |
1 year ago |
| 1455 |
Magnetic Preference Optimization: Achieving Last-iterate Convergence for Language Model Alignment
Mingzhi Wang, Chengdong Ma, ... (+8 more)
|
👻
Ghosted
|
cs.CL
|
19 |
1 year ago |
| 1456 |
Scalable Influence and Fact Tracing for Large Language Model Pretraining
Tyler A. Chang, Dheeraj Rajagopal, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
19 |
1 year ago |
| 1457 |
Aioli: A Unified Optimization Framework for Language Model Data Mixing
Mayee F. Chen, Michael Y. Hu, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
19 |
1 year ago |
| 1458 |
All Seeds Are Not Equal: Enhancing Compositional Text-to-Image Generation with Reliable Random Seeds
Shuangqi Li, Hieu Le, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
19 |
1 year ago |
| 1459 |
Predicting distributions with Linearizing Belief Networks
Yann N. Dauphin, David Grangier
|
👻
Ghosted
|
cs.LG
|
18 |
10 years ago |
| 1460 |
Generalizable Features From Unsupervised Learning
Mehdi Mirza, Aaron Courville, Yoshua Bengio
|
👻
Ghosted
|
stat.ML
|
18 |
9 years ago |
| 1461 |
CommAI: Evaluating the first steps towards a useful general AI
Marco Baroni, Armand Joulin, ... (+5 more)
|
👻
Ghosted
|
cs.LG
|
18 |
9 years ago |
| 1462 |
Shifting Mean Activation Towards Zero with Bipolar Activation Functions
Lars Eidnes, Arild Nøkland
|
👻
Ghosted
|
stat.ML
|
18 |
8 years ago |
| 1463 |
Semantic Interpolation in Implicit Models
Yannic Kilcher, Aurelien Lucchi, Thomas Hofmann
|
👻
Ghosted
|
cs.LG
|
18 |
8 years ago |
| 1464 |
Preconditioner on Matrix Lie Group for SGD
Xi-Lin Li
|
👻
Ghosted
|
stat.ML
|
18 |
7 years ago |
| 1465 |
FSNet: Compression of Deep Convolutional Neural Networks by Filter Summary
Yingzhen Yang, Jiahui Yu, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
18 |
7 years ago |
| 1466 |
Neural Parameter Allocation Search
Bryan A. Plummer, Nikoli Dryden, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
18 |
6 years ago |
| 1467 |
Direction Matters: On the Implicit Bias of Stochastic Gradient Descent with Moderate Learning Rate
Jingfeng Wu, Difan Zou, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
18 |
5 years ago |
| 1468 |
Scaling Laws For Deep Learning Based Image Reconstruction
Tobit Klug, Reinhard Heckel
|
👻
Ghosted
|
eess.IV
|
18 |
3 years ago |
| 1469 |
Sampling-based inference for large linear models, with application to linearised Laplace
Javier Antorán, Shreyas Padhy, ... (+4 more)
|
👻
Ghosted
|
stat.ML
|
18 |
3 years ago |
| 1470 |
Stochastic Differentially Private and Fair Learning
Andrew Lowy, Devansh Gupta, Meisam Razaviyayn
|
👻
Ghosted
|
cs.LG
|
18 |
3 years ago |
| 1471 |
A Learning Based Hypothesis Test for Harmful Covariate Shift
Tom Ginsberg, Zhongyuan Liang, Rahul G. Krishnan
|
👻
Ghosted
|
cs.LG
|
18 |
3 years ago |
| 1472 |
Privacy Amplification for Matrix Mechanisms
Christopher A. Choquette-Choo, Arun Ganesh, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
18 |
2 years ago |
| 1473 |
Octavius: Mitigating Task Interference in MLLMs via LoRA-MoE
Zeren Chen, Ziqin Wang, ... (+8 more)
|
👻
Ghosted
|
cs.CV
|
18 |
2 years ago |
| 1474 |
Follow-Up Differential Descriptions: Language Models Resolve Ambiguities for Image Classification
Reza Esfandiarpoor, Stephen H. Bach
|
👻
Ghosted
|
cs.CL
|
18 |
2 years ago |
| 1475 |
Denoising Diffusion via Image-Based Rendering
Titas Anciukevičius, Fabian Manhardt, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
18 |
2 years ago |
| 1476 |
Tighter Privacy Auditing of DP-SGD in the Hidden State Threat Model
Tudor Cebere, Aurélien Bellet, Nicolas Papernot
|
👻
Ghosted
|
cs.LG
|
18 |
2 years ago |
| 1477 |
Training Neural Networks as Recognizers of Formal Languages
Alexandra Butoi, Ghazal Khalighinejad, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
18 |
1 year ago |
| 1478 |
Feature Incay for Representation Regularization
Yuhui Yuan, Kuiyuan Yang, Chao Zhang
|
👻
Ghosted
|
cs.CV
|
17 |
9 years ago |
| 1479 |
Learning to Teach
Yang Fan, Fei Tian, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
17 |
8 years ago |
| 1480 |
Information Geometry of Orthogonal Initializations and Training
Piotr A. Sokol, Il Memming Park
|
👻
Ghosted
|
stat.ML
|
17 |
7 years ago |
| 1481 |
From Hard to Soft: Understanding Deep Network Nonlinearities via Vector Quantization and Statistical Inference
Randall Balestriero, Richard G. Baraniuk
|
👻
Ghosted
|
cs.LG
|
17 |
7 years ago |
| 1482 |
Convolutional Neural Networks on non-uniform geometrical signals using Euclidean spectral transformation
Chiyu "Max" Jiang, Dequan Wang, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
17 |
7 years ago |
| 1483 |
A Latent Morphology Model for Open-Vocabulary Neural Machine Translation
Duygu Ataman, Wilker Aziz, Alexandra Birch
|
👻
Ghosted
|
cs.CL
|
17 |
6 years ago |
| 1484 |
EUCLID: Towards Efficient Unsupervised Reinforcement Learning with Multi-choice Dynamics Model
Yifu Yuan, Jianye Hao, ... (+7 more)
|
👻
Ghosted
|
cs.LG
|
17 |
3 years ago |
| 1485 |
Understanding the Covariance Structure of Convolutional Filters
Asher Trockman, Devin Willmott, J. Zico Kolter
|
👻
Ghosted
|
cs.CV
|
17 |
3 years ago |
| 1486 |
Bayes risk CTC: Controllable CTC alignment in Sequence-to-Sequence tasks
Jinchuan Tian, Brian Yan, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
17 |
3 years ago |
| 1487 |
Learnable Graph Convolutional Attention Networks
Adrián Javaloy, Pablo Sanchez-Martin, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
17 |
3 years ago |
| 1488 |
Sliced Denoising: A Physics-Informed Molecular Pre-Training Method
Yuyan Ni, Shikun Feng, ... (+3 more)
|
👻
Ghosted
|
q-bio.BM
|
17 |
2 years ago |
| 1489 |
Sparse Spiking Neural Network: Exploiting Heterogeneity in Timescales for Pruning Recurrent SNN
Biswadeep Chakraborty, Beomseok Kang, ... (+2 more)
|
👻
Ghosted
|
cs.NE
|
17 |
2 years ago |
| 1490 |
RevisEval: Improving LLM-as-a-Judge via Response-Adapted References
Qiyuan Zhang, Yufei Wang, ... (+10 more)
|
👻
Ghosted
|
cs.CL
|
17 |
1 year ago |
| 1491 |
From Sparse Dependence to Sparse Attention: Unveiling How Chain-of-Thought Enhances Transformer Sample Efficiency
Kaiyue Wen, Huaqing Zhang, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
17 |
1 year ago |
| 1492 |
Near Exact Privacy Amplification for Matrix Mechanisms
Christopher A. Choquette-Choo, Arun Ganesh, ... (+3 more)
|
👻
Ghosted
|
cs.CR
|
17 |
1 year ago |
| 1493 |
Towards Universality: Studying Mechanistic Similarity Across Language Model Architectures
Junxuan Wang, Xuyang Ge, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
17 |
1 year ago |
| 1494 |
Learning Graph Quantized Tokenizers
Limei Wang, Kaveh Hassani, ... (+8 more)
|
👻
Ghosted
|
cs.NE
|
17 |
1 year ago |
| 1495 |
ET-SEED: Efficient Trajectory-Level SE(3) Equivariant Diffusion Policy
Chenrui Tie, Yue Chen, ... (+5 more)
|
👻
Ghosted
|
cs.RO
|
17 |
1 year ago |
| 1496 |
Simulating Human-like Daily Activities with Desire-driven Autonomy
Yiding Wang, Yuxuan Chen, ... (+3 more)
|
👻
Ghosted
|
cs.AI
|
17 |
1 year ago |
| 1497 |
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
Dhruv Gautam, Spandan Garg, ... (+3 more)
|
👻
Ghosted
|
cs.AI
|
17 |
1 year ago |
| 1498 |
Training Triplet Networks with GAN
Maciej Zieba, Lei Wang
|
👻
Ghosted
|
cs.LG
|
16 |
9 years ago |
| 1499 |
A Flexible Approach to Automated RNN Architecture Generation
Martin Schrimpf, Stephen Merity, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
16 |
8 years ago |
| 1500 |
Improved memory in recurrent neural networks with sequential non-normal dynamics
A. Emin Orhan, Xaq Pitkow
|
👻
Ghosted
|
cs.NE
|
16 |
7 years ago |