Artificial Intelligence · 04.08.2026, 13:19 UTC
Robust Bayesian Optimization via Tempered Posteriors
| Schweregrad | info |
|---|---|
| Kategorie | Artificial Intelligence |
| Quelle | arXiv cs.LG ↗ |
| Veröffentlicht | 04.08.2026 UTC |
Sicherheitsmeldung mit Schweregrad noch nicht bewertet. Technische Details im Tab „Originaltext“; empfohlene Schritte in der Checkliste.
arXiv:2601.07094v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquisition function. Under local misspecification, this feedback loop can produce overconfidence precisely in the region guiding subsequent decisions. We develop a tempered GP-based BO framework that raises the likelihood to a power $\alpha\in(0,1]$. For a generalized family of improvement acquisitions indexed by $g$, including probability of improvement (PI, $g=0$) and expected improvement (EI, $g=1$), we derive finite-time cumulative regret bounds with adaptively learned kernel hyperparameters. The analysis shows that tempering reduces the noise-driven confidence and information-gain contributions to regret, while a deterministic RKHS term prevents arbitrarily aggressive tempering from being uniformly beneficial. It also clarifies the role of the acquisition function: positive-order $g$-EI rules preserve the usual information-gain regret behavior, whereas zero-jitter PI is more exploitative and admits a weaker worst-case guarantee. Motivated by our theoretic findings, we propose a prequential procedure for selecting $\alpha$ online: it decreases $\alpha$ when realized prediction errors exceed model-implied uncertainty and returns $\alpha$ toward one as calibration improves. Empirical results demonstrate that tempering provides a practical yet theoretically grounded tool for stabilizing BO surrogates under localized sampling.
Maßnahmen
⬇ Als MarkdownVerwandte Beiträge
- info Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power
- info Anthropic brings Mythos 5 to its Claude Security vulnerability scanner
- info How agents can delegate better
- info Why API Test Generation Is a Judgment Problem, Not a Code Generation Problem