VBO-MI: A Fully Gradient-Based Bayesian Optimization Framework Using Variational Mutual Information Estimation

๐Ÿ“… 2026-01-13
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This work proposes an end-to-end gradient-driven Bayesian optimization framework to address the high computational cost associated with posterior sampling and acquisition function optimization in traditional Bayesian neural networkโ€“based approaches. By introducing variational mutual information estimation into Bayesian optimization for the first time and integrating it within an actor-critic architecture, the method jointly optimizes input exploration and information gain assessment. This design eliminates the inner-loop acquisition function optimization, yielding a fully differentiable and computationally efficient optimization pipeline. Empirical evaluations demonstrate that the proposed approach achieves performance comparable to or better than existing baselines across multiple high-dimensional synthetic and real-world tasks, while reducing computational overhead by up to two orders of magnitude.

Technology Category

Search and Optimization: Sampling/Simulation-based SearchMachine Learning: OptimizationNatural Language Processing: Learning & Optimization for NLP

Application Category

Economics, Online Markets and Human Computation: Cost models of using LLMs in production systemsSearch and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for rankingGraph Algorithms and Modeling for the Web: Graph neural networks and deep learning approaches for Web-related graphs
๐Ÿ“ Abstract
Many real-world tasks require optimizing expensive black-box functions accessible only through noisy evaluations, a setting commonly addressed with Bayesian optimization (BO). While Bayesian neural networks (BNNs) have recently emerged as scalable alternatives to Gaussian Processes (GPs), traditional BNN-BO frameworks remain burdened by expensive posterior sampling and acquisition function optimization. In this work, we propose {VBO-MI} (Variational Bayesian Optimization with Mutual Information), a fully gradient-based BO framework that leverages recent advances in variational mutual information estimation. To enable end-to-end gradient flow, we employ an actor-critic architecture consisting of an {action-net} to navigate the input space and a {variational critic} to estimate information gain. This formulation effectively eliminates the traditional inner-loop acquisition optimization bottleneck, achieving up to a {$10^2 \times$ reduction in FLOPs} compared to BNN-BO baselines. We evaluate our method on a diverse suite of benchmarks, including high-dimensional synthetic functions and complex real-world tasks such as PDE optimization, the Lunar Lander control problem, and categorical Pest Control. Our experiments demonstrate that VBO-MI consistently provides the same or superior optimization performance and computational scalability over the baselines.
Problem

Research questions and friction points this paper is trying to address.

Bayesian optimization
Bayesian neural networks
acquisition function optimization
black-box optimization
computational scalability
Innovation

Methods, ideas, or system contributions that make the work stand out.

Bayesian Optimization
Variational Mutual Information
Gradient-based Optimization
Bayesian Neural Networks
Actor-Critic Architecture
๐Ÿ”Ž Similar Papers
No similar papers found.
๐Ÿ’ผ Related Jobs
No related jobs found.
F
Farhad Mirkarimi