Institution profile

Chengdu Institute of Computer Application

Academic institutionasia · cn
Official website
Research library5linked papers
Opportunities0open roles
Selected work

Representative Papers

HygieneRoboBench: Benchmarking Hygiene-Aware Planning for Household Robots

Oct 06, 2026

This study addresses the overlooked risks of contact contamination and plan invalidation caused by novel contact events in domestic robot planning. To this end, it proposes Hygiene-NSP, a method that jointly optimizes hygiene management and task execution. Technically, the approach integrates large language model grounding, contact history reconstruction, and a CP-SAT solver to achieve hybrid planning. Additionally, a benchmark is constructed to systematically evaluate the planner’s ability to balance hygiene risk identification, cost control, and user preferences. Experimental results demonstrate that the proposed method attains a safe completion rate of 94.4% and an optimal safety rate of 90.4%, significantly outperforming existing baselines. These findings indicate that Hygiene-NSP overcomes the limitations of conventional planners that fail to simultaneously ensure safety and cost optimality.

0 citationsRead paper

LLM-Based Test Generation: Information Sources, Generation Strategies, and Quality Evidence

Oct 04, 2026

This study addresses the unclear relationships among information sources, generation strategies, and quality evidence in test case generation using large language models (LLMs). Through a systematic literature review of 95 studies, this work constructs a multidimensional taxonomy and a benchmark analysis framework. Specifically, it proposes a four-dimensional classification system that elucidates how execution feedback influences oracle independence. Furthermore, it establishes a unified theoretical framework connecting the generation process with quality assessment. By identifying independent oracle evaluation as a critical yet underexplored dimension, this research formulates a future agenda centered on rigorous, oracle-independent quality measurement for LLM-generated test cases.

0 citationsRead paper
Recent publications

Latest Papers

HygieneRoboBench: Benchmarking Hygiene-Aware Planning for Household Robots

Oct 06, 2026

This study addresses the overlooked risks of contact contamination and plan invalidation caused by novel contact events in domestic robot planning. To this end, it proposes Hygiene-NSP, a method that jointly optimizes hygiene management and task execution. Technically, the approach integrates large language model grounding, contact history reconstruction, and a CP-SAT solver to achieve hybrid planning. Additionally, a benchmark is constructed to systematically evaluate the planner’s ability to balance hygiene risk identification, cost control, and user preferences. Experimental results demonstrate that the proposed method attains a safe completion rate of 94.4% and an optimal safety rate of 90.4%, significantly outperforming existing baselines. These findings indicate that Hygiene-NSP overcomes the limitations of conventional planners that fail to simultaneously ensure safety and cost optimality.

0 citationsRead paper

LLM-Based Test Generation: Information Sources, Generation Strategies, and Quality Evidence

Oct 04, 2026

This study addresses the unclear relationships among information sources, generation strategies, and quality evidence in test case generation using large language models (LLMs). Through a systematic literature review of 95 studies, this work constructs a multidimensional taxonomy and a benchmark analysis framework. Specifically, it proposes a four-dimensional classification system that elucidates how execution feedback influences oracle independence. Furthermore, it establishes a unified theoretical framework connecting the generation process with quality assessment. By identifying independent oracle evaluation as a critical yet underexplored dimension, this research formulates a future agenda centered on rigorous, oracle-independent quality measurement for LLM-generated test cases.

0 citationsRead paper