WeaveData: A Multimodal Data Analysis System with Self-Critiquing and Self-Evolving LLM Plans

📅 2026-09-28
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the challenges of error-prone planning, execution failures, and intent deviation when large language models are applied to multimodal data analysis. To overcome these limitations, this work proposes a planning system equipped with self-criticism and dynamic evolution capabilities. The proposed approach is grounded in a metadata knowledge graph and incorporates typed logical plans for rigorous validation. Furthermore, a self-criticism algorithm is introduced to enable the reuse of error diagnostics and the accumulation of planning experience. Experimental evaluations conducted on two public multimodal datasets demonstrate that the proposed system significantly enhances both the accuracy and robustness of analytical tasks.
📝 Abstract
Multimodal data analysis, which answers questions over relational tables, text, and images, has attracted growing attention in the data management community. Large language models (LLMs) enable such analysis in natural language by generating analysis plans over relational and semantic operators. However, LLM-generated plans are error-prone: a plan may silently compute something other than what was asked, fail during execution, or return a result that misses the question. This paper presents WeaveData, a multimodal data analysis system with self-critiquing and self-evolving LLM plans. First, WeaveData generates a typed logical plan for each question and critiques it step by step before execution, and it checks the executed result against the question afterwards. Second, WeaveData evolves a plan that fails or misses the question: it diagnoses the failure with the actual data, reuses the results that remain valid, and accumulates planning experience for later questions. Third, WeaveData grounds planning in a metadata knowledge graph of all modalities, clarifies ambiguous questions with the user, and backs every model judgment with evidence in an interactive notebook. We demonstrate WeaveData on two public multimodal datasets.
Problem

Research questions and friction points this paper is trying to address.

Multimodal data analysis
Large language models
Error-prone plans
Plan execution failure
Innovation

Methods, ideas, or system contributions that make the work stand out.

Multimodal Data Analysis
Self-Critiquing
Self-Evolving Plans
Large Language Models
Metadata Knowledge Graph
M
Min Jia
Northwest A&F University
S
Shihao Zhou
East China Normal University
J
Jun-Peng Zhu
Northwest A&F University
P
Peng Cai
East China Normal University
K
Kai Xu
PingCAP
Chao Zhang
Chao Zhang
Renmin University of China
HTAPCloud-Native DatabasesMulti-Model DatabasesBenchmark
L
Li Li
PingCAP
A
Aoying Zhou
East China Normal University
H
Heng Long
PingCAP
Q
Qiu Cui
PingCAP
L
Liu Tang
PingCAP
Q
Qi Liu
PingCAP