Spike-and-Slab Posterior Sampling in High Dimensions

March 04, 2025 ยท Declared Dead ยท ๐Ÿ› Annual Conference Computational Learning Theory

๐Ÿ‘ป CAUSE OF DEATH: Ghosted
No code link whatsoever

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Syamantak Kumar, Purnamrita Sarkar, Kevin Tian, Yusong Zhu arXiv ID 2503.02798 Category stat.ML: Machine Learning (Stat) Cross-listed cs.DS, cs.LG, stat.ME Citations 3 Venue Annual Conference Computational Learning Theory Last Checked 5 months ago
Abstract
Posterior sampling with the spike-and-slab prior [MB88], a popular multimodal distribution used to model uncertainty in variable selection, is considered the theoretical gold standard method for Bayesian sparse linear regression [CPS09, Roc18]. However, designing provable algorithms for performing this sampling task is notoriously challenging. Existing posterior samplers for Bayesian sparse variable selection tasks either require strong assumptions about the signal-to-noise ratio (SNR) [YWJ16], only work when the measurement count grows at least linearly in the dimension [MW24], or rely on heuristic approximations to the posterior. We give the first provable algorithms for spike-and-slab posterior sampling that apply for any SNR, and use a measurement count sublinear in the problem dimension. Concretely, assume we are given a measurement matrix $\mathbf{X} \in \mathbb{R}^{n\times d}$ and noisy observations $\mathbf{y} = \mathbf{X}\mathbfฮธ^\star + \mathbfฮพ$ of a signal $\mathbfฮธ^\star$ drawn from a spike-and-slab prior $ฯ€$ with a Gaussian diffuse density and expected sparsity k, where $\mathbfฮพ \sim \mathcal{N}(\mathbb{0}_n, ฯƒ^2\mathbf{I}_n)$. We give a polynomial-time high-accuracy sampler for the posterior $ฯ€(\cdot \mid \mathbf{X}, \mathbf{y})$, for any SNR $ฯƒ^{-1}$ > 0, as long as $n \geq k^3 \cdot \text{polylog}(d)$ and $X$ is drawn from a matrix ensemble satisfying the restricted isometry property. We further give a sampler that runs in near-linear time $\approx nd$ in the same setting, as long as $n \geq k^5 \cdot \text{polylog}(d)$. To demonstrate the flexibility of our framework, we extend our result to spike-and-slab posterior sampling with Laplace diffuse densities, achieving similar guarantees when $ฯƒ= O(\frac{1}{k})$ is bounded.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

๐Ÿ“œ Similar Papers

In the same crypt โ€” Machine Learning (Stat)

๐Ÿ”ฎ ๐Ÿ”ฎ The Ethereal

Layer Normalization

Jimmy Lei Ba, Jamie Ryan Kiros, Geoffrey E. Hinton

stat.ML ๐Ÿ› arXiv ๐Ÿ“š 12.0K cites 10 years ago

Died the same way โ€” ๐Ÿ‘ป Ghosted