Heuristic Solutions for the Best Secretary Problem

📅 2025-11-13
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This paper addresses the classical secretary problem—online selection of the best candidate based solely on relative rank information. To overcome the poor adaptability of conventional fixed-threshold policies, we propose a data-responsive heuristic framework featuring sequential threshold adaptation. It integrates five tunable rules—including expected-record thresholds, adaptive bias correction, and probabilistic early-stopping—and employs two-stage relaxation with local dynamic programming approximation to enhance robustness and decision efficiency. The method synergistically combines probabilistic modeling with lightweight ensemble learning. Extensive simulations across diverse scenarios validate its efficacy. Experimental results show that the framework achieves near-optimal performance with minimal intuitive hyperparameters, consistently outperforming classical strategies in both average-case performance and stability; the ensemble variant demonstrates the highest robustness.

Technology Category

Search and Optimization: Heuristic SearchReasoning under Uncertainty: Sequential Decision MakingMachine Learning: Online Learning & Bandits

Application Category

Search and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for rankingUser Modeling, Personalization and Recommendation: Fairness-aware retrieval and rankingEconomics, Online Markets and Human Computation: Data quality aspects of human-annotated datasets
📝 Abstract
This paper introduces a heuristic framework for the Best Secretary Problem, where one item must be selected using rank information only. We develop five data-responsive rules extending classical fixed-cutoff methods: an expected-record threshold, an adaptive deviation correction, a probabilistic early-accept rule, a two-phase relaxation, and a local dynamic programming approximation. These rules adjust thresholds sequentially as information accumulates. Simulations across diverse sample sizes, distributions, and autocorrelated settings show that the heuristics match or exceed traditional optimal rules in stability and efficiency. The expected-record rule remains strong despite its simplicity, the adaptive correction performs well under asymmetry, and the adaptive and probabilistic rules reduce average stopping times. An ensemble combining multiple rules yields the most stable performance. Overall, a few intuitive parameters achieve near-optimal results, demonstrating that data-responsive heuristics can effectively extend rank-based optimal stopping to dynamic decision environments.
Problem

Research questions and friction points this paper is trying to address.

Developing heuristic rules for rank-based item selection problems
Extending classical methods with adaptive data-responsive thresholds
Improving decision stability and efficiency in optimal stopping scenarios
Innovation

Methods, ideas, or system contributions that make the work stand out.

Expected-record threshold extends classical fixed-cutoff methods
Adaptive deviation correction handles asymmetric distributions effectively
Probabilistic early-accept rule reduces average stopping times
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
E
Eugene Seong
Department of Statistics, Korea University, Seoul, Republic of Korea