bidirectional retrieval

Design, build, or analyze retrieval systems that generate and score candidate evidence paths by expanding and matching from both the topic/query side and the answer/candidate side, using simulation of retrieval steps and prefix expansions to retrieve compact candidate paths while suppressing noisy mixed-type expansions. This competence includes implementing retrieval-forward simulation, answer-side final-hop relation matching, topic-side prefix expansion, and unified bidirectional modeling to combine forward and backward signals into a single retrieval pipeline.

bidirectionalretrieval

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
0.19
Oct 01, 2026Oct 01, 2026
Career
Value
No comparison yet
$200K/year
Oct 01, 2026Oct 01, 2026

Recommended Survey Paper

Quick overview of the field
View more

Must-Read Papers

Most classic and influential ideas
View more

Retrieval-Augmented Generation by Evidence Retroactivity in LLMs

Jan 07, 2025
LX
Liang Xiao
🏛️ Beijing Institute of Technology | Xiaomi Corporation

To address error propagation and answer bias arising from unidirectional retrieval-then-reasoning in multi-hop question answering, this paper proposes RetroRAG, the first framework introducing backtracking-style reasoning. Its core is an evidence backtracking mechanism: inferring entity-centric queries to dynamically revise retrieved evidence and reconstruct reasoning paths, enabling iterative refinement and dynamic reorganization of trustworthy evidence through coordinated multi-round retrieval-generation-evaluation cycles. This establishes a closed-loop “evidence curation–discovery–verification” process, substantially enhancing robustness and interpretability for complex reasoning. On mainstream multi-hop QA benchmarks, RetroRAG consistently outperforms existing RAG methods, achieving significant gains in answer accuracy—particularly under challenging conditions involving long reasoning chains and noisy evidence.

Accuracy ImprovementInformation RetrievalLarge Language Models

Latest Papers

What's happening recently
View more

In multi-hop graph retrieval, the original query often fails to fully capture the complete information need distributed across multiple reasoning steps, resulting in insufficient retrieval signals. This work proposes an evidence-guided query reformulation mechanism that decouples query refinement from evidence aggregation on the graph: a residual query is generated from already retrieved passages to characterize unmet information needs, and the retrieval signals from the original and residual queries are separately normalized and then fused, propagating through shared entities across propositions. By moving beyond the conventional reliance solely on the initial query, the method achieves substantial gains, improving Recall@5 by up to 5.59 points and F1 by up to 4.50 points on 2WikiMultiHopQA, HotpotQA, and MuSiQue.

evidence-guidedgraph retrievalinformation need

This work addresses the challenge in LongEval-RAG tasks where responses must be strictly grounded in a given set of candidate documents. To this end, the authors propose a candidate-constrained retrieval-augmented generation (RAG) system that integrates rule-based chunking, query expansion, pseudo-relevance feedback, reciprocal rank fusion, MiniLM sentence-level reranking, and citation-aware evidence aggregation, complemented by deterministic provenance tracing and a neural sentence selection mechanism. Experimental results demonstrate that the proposed rule-MiniLM variant significantly outperforms baselines across multiple metrics—including BERTScore, retrieval precision, information point coverage, and human evaluation—thereby validating the effectiveness of combining rule-based chunking with neural sentence selection. The study further underscores the critical role of multi-metric evaluation in diagnosing and advancing RAG system performance.

candidate-constrained retrievalcitation constraintevidence retrieval

Hot Scholars

YZ

Yu Zhao

Harbin Institute of Technology (Shenzhen)
natural language processingmultimedia
MW

Mingming Wang

Northwestern Polytechnical University
roboticsdynamicsplanning and control
SK

See-Kiong Ng

School of Computing and Institute of Data Science, National University of Singapore
artificial intelligencenatural language processingdata miningsmart cities
YM

Yunshan Ma

Singapore Management University, NUS
Multimodal Event ForecastingBundle RecommendationComputational Fashion/Finance/Security/Politics
YB

Yi Bin

National University of Singapore
multimediavision and languagedeep learning