Comprehend, Divide, and Conquer: Feature Subspace Exploration via Multi-Agent Hierarchical Reinforcement Learning

๐Ÿ“… 2025-04-24
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
Existing reinforcement learning (RL)-based feature selection methods for high-dimensional complex data suffer from inefficient subspace exploration and suboptimal downstream task performance due to the โ€œone-featureโ€“one-agentโ€ paradigm. To address this, we propose HRLFS, a multi-agent hierarchical reinforcement learning framework for feature selection. Its key contributions are: (1) a novel hierarchical agent architecture that replaces conventional flat RL designs; (2) integration of statistical features with LLM-driven semantic representations, coupled with hierarchical clustering to group features by semantic similarity; and (3) interpretable, scalable, progressive collaborative optimization over feature subspaces. Evaluated on multiple benchmark datasets, HRLFS achieves an average 3.2% improvement in classification accuracy, reduces runtime by 41%, and demonstrates strong robustness and cross-domain generalization capability.

Technology Category

Machine Learning: Feature Construction/ReformulationMultiagent Systems: Multiagent LearningHumans and AI: Human-in-the-loop Machine Learning

Application Category

Semantics and Knowledge: Data modeling to support human-machine intelligence, including LLMs agents, intelligent system behavior, explanations, and user-friendly interactionsSearch and Retrieval-Augmented AI: Agentic searchEconomics, Online Markets and Human Computation: Humans versus LLMs for data annotation and labeling
๐Ÿ“ Abstract
Feature selection aims to preprocess the target dataset, find an optimal and most streamlined feature subset, and enhance the downstream machine learning task. Among filter, wrapper, and embedded-based approaches, the reinforcement learning (RL)-based subspace exploration strategy provides a novel objective optimization-directed perspective and promising performance. Nevertheless, even with improved performance, current reinforcement learning approaches face challenges similar to conventional methods when dealing with complex datasets. These challenges stem from the inefficient paradigm of using one agent per feature and the inherent complexities present in the datasets. This observation motivates us to investigate and address the above issue and propose a novel approach, namely HRLFS. Our methodology initially employs a Large Language Model (LLM)-based hybrid state extractor to capture each feature's mathematical and semantic characteristics. Based on this information, features are clustered, facilitating the construction of hierarchical agents for each cluster and sub-cluster. Extensive experiments demonstrate the efficiency, scalability, and robustness of our approach. Compared to contemporary or the one-feature-one-agent RL-based approaches, HRLFS improves the downstream ML performance with iterative feature subspace exploration while accelerating total run time by reducing the number of agents involved.
Problem

Research questions and friction points this paper is trying to address.

Optimizing feature selection for machine learning tasks
Overcoming inefficiencies in reinforcement learning-based subspace exploration
Enhancing performance via hierarchical multi-agent feature clustering
Innovation

Methods, ideas, or system contributions that make the work stand out.

LLM-based hybrid state extractor for feature characteristics
Hierarchical agents for clustered feature exploration
Iterative subspace exploration to enhance ML performance
๐Ÿ”Ž Similar Papers
No similar papers found.
W
Weiliang Zhang
Computer Network Information Center, Chinese Academy of Sciences and University of Chinese Academy of Sciences, China
X
Xiaohan Huang
Computer Network Information Center, Chinese Academy of Sciences and University of Chinese Academy of Sciences, China
Yi Du
Yi Du
Chinese Academy of Sciences
data miningknowledge engineeringAI for Science
Ziyue Qiao
Ziyue Qiao
Assistant Professor, Great Bay University
Data MiningGraph Machine LearningKnowledge GraphAI for Science
Q
Qingqing Long
Computer Network Information Center, Chinese Academy of Sciences, China
Z
Zhen Meng
Computer Network Information Center, Chinese Academy of Sciences, China
Yuanchun Zhou
Yuanchun Zhou
Computer Network Information Center,CAS
Data MiningBig Data Analysis
M
Meng Xiao
Computer Network Information Center, Chinese Academy of Sciences, China