๐ฎ
๐ฎ
The Ethereal
Online learning with noisy side observations
April 15, 2026 ยท Grace Period ยท ๐ Proceedings of the 19th International Conference on Artificial Intelligence and Statistics (AISTATS), pages 1186-1194, 2016
Authors
Tomรกลก Kocรกk, Gergely Neu, Michal Valko
arXiv ID
2604.13740
Category
cs.LG: Machine Learning
Cross-listed
stat.ML
Citations
0
Venue
Proceedings of the 19th International Conference on Artificial Intelligence and Statistics (AISTATS), pages 1186-1194, 2016
Abstract
We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback about the other actions, depending on the underlying structure of the problem. We represent this structure by a weighted directed graph, where the edge weights are related to the quality of the feedback shared by the connected nodes. Our main contribution is an efficient algorithm that guarantees a regret of $\widetilde{O}(\sqrt{ฮฑ^* T})$ after $T$ rounds, where $ฮฑ^*$ is a novel graph property that we call the effective independence number. Our algorithm is completely parameter-free and does not require knowledge (or even estimation) of $ฮฑ^*$. For the special case of binary edge weights, our setting reduces to the partial-observability models of Mannor and Shamir (2011) and Alon et al. (2013) and our algorithm recovers the near-optimal regret bounds.
Community Contributions
Found the code? Know the venue? Think something is wrong? Let us know!
๐ Similar Papers
In the same crypt โ Machine Learning
๐ฎ
๐ฎ
The Ethereal
Continuous control with deep reinforcement learning
๐
๐
Old Age
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks
๐
๐
Old Age
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
๐
๐
Old Age
SGDR: Stochastic Gradient Descent with Warm Restarts
๐ฎ
๐ฎ
The Ethereal