CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control (2405.18727v2)

Published 29 May 2024 in cs.CL, cs.AI, and cs.IR

Abstract: Retrieval-augmented generation (RAG) has emerged as a promising solution for mitigating hallucinations of LLMs with retrieved external knowledge. Adaptive RAG enhances this approach by enabling dynamic retrieval during generation, activating retrieval only when the query exceeds LLM's internal knowledge. Existing methods primarily focus on detecting LLM's confidence via statistical uncertainty. Instead, we present the first attempts to solve adaptive RAG from a representation perspective and develop an inherent control-based framework, termed \name. Specifically, we extract the features that represent the honesty and confidence directions of LLM and adopt them to control LLM behavior and guide retrieval timing decisions. We also design a simple yet effective query formulation strategy to support adaptive retrieval. Experiments show that \name is superior to existing adaptive RAG methods on a diverse set of tasks, the honesty steering can effectively make LLMs more honest and confidence monitoring is a promising indicator of retrieval trigger.Our code is available at \url{https://github.com/HSLiu-Initial/CtrlA}.

PDF Abstract

Summarize PDF Markdown Bookmark Chat (Pro)

Authors (9)

Huanshuo Liu (3 papers)
Hao Zhang (947 papers)
Zhijiang Guo (55 papers)
Kuicai Dong (17 papers)
Xiangyang Li (58 papers)
Yi Quan Lee (5 papers)
Cong Zhang (121 papers)
Yong Liu (721 papers)
Jing Wang (740 papers)

GitHub

GitHub - HSLiu-Initial/CtrlA: This includes the original implementation of CTRLA: Adaptive Retrieval-Augmented Generation via Probe-Guided Control. (10 stars)

Tweets

https://twitter.com/_reachsumit/status/1796021412806721928

https://twitter.com/gm8xx8/status/1796022513148805558

https://twitter.com/dacbarbos/status/1798852884693663956

CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control (2405.18727v2)

Related Papers

GitHub

Tweets