| 251 |
Deep Gaussian Processes for Regression using Approximate Expectation Propagation
Thang D. Bui, Daniel Hernández-Lobato, ... (+3 more)
|
👻
Ghosted
|
stat.ML
|
243 |
10 years ago |
| 252 |
Learning Granger Causality for Hawkes Processes
Hongteng Xu, Mehrdad Farajtabar, Hongyuan Zha
|
👻
Ghosted
|
cs.LG
|
243 |
10 years ago |
| 253 |
Dynamic Word Embeddings
Robert Bamler, Stephan Mandt
|
👻
Ghosted
|
stat.ML
|
243 |
9 years ago |
| 254 |
Schema Networks: Zero-shot Transfer with a Generative Causal Model of Intuitive Physics
Ken Kansky, Tom Silver, ... (+8 more)
|
👻
Ghosted
|
cs.AI
|
242 |
9 years ago |
| 255 |
Measuring Sample Quality with Kernels
Jackson Gorham, Lester Mackey
|
👻
Ghosted
|
stat.ML
|
241 |
9 years ago |
| 256 |
Gradient Descent Learns One-hidden-layer CNN: Don't be Afraid of Spurious Local Minima
Simon S. Du, Jason D. Lee, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
241 |
8 years ago |
| 257 |
Interpretations are useful: penalizing explanations to align neural networks with prior knowledge
Laura Rieger, Chandan Singh, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
241 |
6 years ago |
| 258 |
Coordinate Descent Converges Faster with the Gauss-Southwell Rule Than Random Selection
Julie Nutini, Mark Schmidt, ... (+3 more)
|
👻
Ghosted
|
math.OC
|
234 |
11 years ago |
| 259 |
Deep Counterfactual Regret Minimization
Noam Brown, Adam Lerer, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
234 |
7 years ago |
| 260 |
Learning to Search Better Than Your Teacher
Kai-Wei Chang, Akshay Krishnamurthy, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
233 |
11 years ago |
| 261 |
DoubleSqueeze: Parallel Stochastic Gradient Descent with Double-Pass Error-Compensated Compression
Hanlin Tang, Xiangru Lian, ... (+3 more)
|
👻
Ghosted
|
cs.DC
|
233 |
7 years ago |
| 262 |
A Neural Network Architecture Combining Gated Recurrent Unit (GRU) and Support Vector Machine (SVM) for Intrusion Detection in Network Traffic Data
Abien Fred Agarap
|
👻
Ghosted
|
cs.NE
|
232 |
8 years ago |
| 263 |
Optimal and Adaptive Off-policy Evaluation in Contextual Bandits
Yu-Xiang Wang, Alekh Agarwal, Miroslav Dudik
|
👻
Ghosted
|
stat.ML
|
231 |
9 years ago |
| 264 |
Minimax Pareto Fairness: A Multi Objective Perspective
Natalia Martinez, Martin Bertran, Guillermo Sapiro
|
👻
Ghosted
|
stat.ML
|
228 |
5 years ago |
| 265 |
The Unsurprising Effectiveness of Pre-Trained Vision Models for Control
Simone Parisi, Aravind Rajeswaran, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
226 |
4 years ago |
| 266 |
Structured Prediction Energy Networks
David Belanger, Andrew McCallum
|
👻
Ghosted
|
cs.LG
|
225 |
10 years ago |
| 267 |
A Neural Autoregressive Approach to Collaborative Filtering
Yin Zheng, Bangsheng Tang, ... (+2 more)
|
👻
Ghosted
|
cs.IR
|
225 |
10 years ago |
| 268 |
Failures of Gradient-Based Deep Learning
Shai Shalev-Shwartz, Ohad Shamir, Shaked Shammah
|
👻
Ghosted
|
cs.LG
|
225 |
9 years ago |
| 269 |
Automated Behavioral Analysis of Malware A Case Study of WannaCry Ransomware
Qian Chen, Robert A. Bridges
|
👻
Ghosted
|
cs.CR
|
225 |
8 years ago |
| 270 |
Scalable Bayesian Rule Lists
Hongyu Yang, Cynthia Rudin, Margo Seltzer
|
👻
Ghosted
|
cs.AI
|
224 |
10 years ago |
| 271 |
Prioritized Training on Points that are Learnable, Worth Learning, and Not Yet Learnt
Sören Mindermann, Jan Brauner, ... (+9 more)
|
👻
Ghosted
|
cs.LG
|
224 |
4 years ago |
| 272 |
An Analytical Formula of Population Gradient for two-layered ReLU network and its Applications in Convergence and Critical Point Analysis
Yuandong Tian
|
👻
Ghosted
|
cs.LG
|
222 |
9 years ago |
| 273 |
Modeling Others using Oneself in Multi-Agent Reinforcement Learning
Roberta Raileanu, Emily Denton, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
222 |
8 years ago |
| 274 |
Understanding the Origins of Bias in Word Embeddings
Marc-Etienne Brunet, Colleen Alkalay-Houlihan, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
221 |
7 years ago |
| 275 |
Generative Adversarial User Model for Reinforcement Learning Based Recommendation System
Xinshi Chen, Shuang Li, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
220 |
7 years ago |
| 276 |
Scalable Fair Clustering
Arturs Backurs, Piotr Indyk, ... (+4 more)
|
👻
Ghosted
|
cs.DS
|
220 |
7 years ago |
| 277 |
On the Impact of the Activation Function on Deep Neural Networks Training
Soufiane Hayou, Arnaud Doucet, Judith Rousseau
|
👻
Ghosted
|
stat.ML
|
220 |
7 years ago |
| 278 |
Evasion and Hardening of Tree Ensemble Classifiers
Alex Kantchelian, J. D. Tygar, Anthony D. Joseph
|
👻
Ghosted
|
cs.LG
|
219 |
10 years ago |
| 279 |
Fairness in Reinforcement Learning
Shahin Jabbari, Matthew Joseph, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
219 |
9 years ago |
| 280 |
Noisy Natural Gradient as Variational Inference
Guodong Zhang, Shengyang Sun, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
219 |
8 years ago |
| 281 |
Mixture Proportion Estimation via Kernel Embedding of Distributions
Harish G. Ramaswamy, Clayton Scott, Ambuj Tewari
|
👻
Ghosted
|
cs.LG
|
218 |
10 years ago |
| 282 |
Safe Policy Improvement with Baseline Bootstrapping
Romain Laroche, Paul Trichelair, Rémi Tachet des Combes
|
👻
Ghosted
|
cs.LG
|
218 |
8 years ago |
| 283 |
Can Autonomous Vehicles Identify, Recover From, and Adapt to Distribution Shifts?
Angelos Filos, Panagiotis Tigas, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
218 |
6 years ago |
| 284 |
Parametric Exponential Linear Unit for Deep Convolutional Neural Networks
Ludovic Trottier, Philippe Giguère, Brahim Chaib-draa
|
👻
Ghosted
|
cs.LG
|
217 |
10 years ago |
| 285 |
Parallel Multiscale Autoregressive Density Estimation
Scott Reed, Aäron van den Oord, ... (+5 more)
|
👻
Ghosted
|
cs.CV
|
216 |
9 years ago |
| 286 |
Path-Level Network Transformation for Efficient Architecture Search
Han Cai, Jiacheng Yang, ... (+3 more)
|
👻
Ghosted
|
cs.LG
|
216 |
8 years ago |
| 287 |
Directional Graph Networks
Dominique Beaini, Saro Passaro, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
213 |
5 years ago |
| 288 |
The Uncertainty Bellman Equation and Exploration
Brendan O'Donoghue, Ian Osband, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
210 |
8 years ago |
| 289 |
Hierarchical Imitation and Reinforcement Learning
Hoang M. Le, Nan Jiang, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
210 |
8 years ago |
| 290 |
Improved SVRG for Non-Strongly-Convex or Sum-of-Non-Convex Objectives
Zeyuan Allen-Zhu, Yang Yuan
|
👻
Ghosted
|
cs.LG
|
209 |
11 years ago |
| 291 |
Dropout Inference in Bayesian Neural Networks with Alpha-divergences
Yingzhen Li, Yarin Gal
|
👻
Ghosted
|
cs.LG
|
207 |
9 years ago |
| 292 |
Sequence Tutor: Conservative Fine-Tuning of Sequence Generation Models with KL-control
Natasha Jaques, Shixiang Gu, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
206 |
9 years ago |
| 293 |
Dissecting Adam: The Sign, Magnitude and Variance of Stochastic Gradients
Lukas Balles, Philipp Hennig
|
👻
Ghosted
|
cs.LG
|
206 |
9 years ago |
| 294 |
Coordinated Multi-Agent Imitation Learning
Hoang M. Le, Yisong Yue, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
205 |
9 years ago |
| 295 |
Cognitive Psychology for Deep Neural Networks: A Shape Bias Case Study
Samuel Ritter, David G. T. Barrett, ... (+2 more)
|
👻
Ghosted
|
stat.ML
|
203 |
9 years ago |
| 296 |
Greedy Layerwise Learning Can Scale to ImageNet
Eugene Belilovsky, Michael Eickenberg, Edouard Oyallon
|
👻
Ghosted
|
cs.LG
|
203 |
7 years ago |
| 297 |
Prioritized Level Replay
Minqi Jiang, Edward Grefenstette, Tim Rocktäschel
|
👻
Ghosted
|
cs.LG
|
200 |
5 years ago |
| 298 |
Improved Analysis of Score-based Generative Modeling: User-Friendly Bounds under Minimal Smoothness Assumptions
Hongrui Chen, Holden Lee, Jianfeng Lu
|
👻
Ghosted
|
cs.LG
|
199 |
3 years ago |
| 299 |
Learning Independent Causal Mechanisms
Giambattista Parascandolo, Niki Kilbertus, ... (+2 more)
|
👻
Ghosted
|
cs.LG
|
198 |
8 years ago |
| 300 |
Evolution of Heuristics: Towards Efficient Automatic Algorithm Design Using Large Language Model
Fei Liu, Xialiang Tong, ... (+6 more)
|
👻
Ghosted
|
cs.NE
|
198 |
2 years ago |