Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

iRAT: Replanning and Controlled Retrieval for Robust LLM Reasoning

View through CrossRef
Large Language Models (LLMs) have demonstrated significant capabilities in answering questions using techniques such as Chain of Thought (CoT) and Retrieval-Augmented Generation (RAG). CoT enables step-by-step reasoning to improve accuracy, while RAG supplements LLMs with relevant external information. Retrieval-Augmented Thoughts (RAT) combines CoT and RAG to provide a more robust factual foundation and coherence in reasoning chains. However, RAT is limited in its ability to handle uncertainty and lacks replanning, often resulting in unnecessary retrievals, inefficiencies, and globally inconsistent reasoning. To address these limitations, we introduce iRAT, a novel reasoning framework that enhances RAT through retrieval control and replanning. iRAT dynamically evaluates uncertainty in initial responses, employs controlled and filtered retrievals to obtain only the most relevant context, revises thoughts to align with new content, and uses replanning to correct previous thoughts. Evaluations demonstrated that iRAT outperforms RAT in HumanEval, MBPP, and GSM8K datasets, while reducing retrievals by a considerable amount. The source code is available at github.com/prane-eth/iRAT. The fine-tuned model used for replanning is available at huggingface.co/zeeshan5k/iRATReasoningChainEvaluatorv2.
Title: iRAT: Replanning and Controlled Retrieval for Robust LLM Reasoning
Description:
Large Language Models (LLMs) have demonstrated significant capabilities in answering questions using techniques such as Chain of Thought (CoT) and Retrieval-Augmented Generation (RAG).
CoT enables step-by-step reasoning to improve accuracy, while RAG supplements LLMs with relevant external information.
Retrieval-Augmented Thoughts (RAT) combines CoT and RAG to provide a more robust factual foundation and coherence in reasoning chains.
However, RAT is limited in its ability to handle uncertainty and lacks replanning, often resulting in unnecessary retrievals, inefficiencies, and globally inconsistent reasoning.
To address these limitations, we introduce iRAT, a novel reasoning framework that enhances RAT through retrieval control and replanning.
iRAT dynamically evaluates uncertainty in initial responses, employs controlled and filtered retrievals to obtain only the most relevant context, revises thoughts to align with new content, and uses replanning to correct previous thoughts.
Evaluations demonstrated that iRAT outperforms RAT in HumanEval, MBPP, and GSM8K datasets, while reducing retrievals by a considerable amount.
The source code is available at github.
com/prane-eth/iRAT.
The fine-tuned model used for replanning is available at huggingface.
co/zeeshan5k/iRATReasoningChainEvaluatorv2.

Related Results

iRAT: Improved Retrieval-Augmented Thinking for Context-Aware Replanning-Based Reasoning
iRAT: Improved Retrieval-Augmented Thinking for Context-Aware Replanning-Based Reasoning
Large Language Models (LLMs) have demonstrated significant capabilities in answering questions using techniques such as Chain of Thought (CoT) and Retrieval-Augmented Generation (R...
REDESAIN ALAT IRAT BAMBU BERBASIS ERGONOMI PADA PENGRAJIN ANYAMAN BAMBU DI DESA JEPANG KABUPATEN KUDUS
REDESAIN ALAT IRAT BAMBU BERBASIS ERGONOMI PADA PENGRAJIN ANYAMAN BAMBU DI DESA JEPANG KABUPATEN KUDUS
Jepang Village, Mejobo District, Kudus Regency is a village dominated by bamboo weaving artisans. Bamboo weaving activities in Jepang villages are carried out every day as a liveli...
Automating Information Retrieval from Biodiversity Literature Using Large Language Models: A Case Study
Automating Information Retrieval from Biodiversity Literature Using Large Language Models: A Case Study
Recently, Large Language Models (LLMs) have transformed information retrieval, becoming widely adopted across various domains due to their ability to process extensive textual data...
How Large Language Models Can Affect Clinical Reasoning: A Randomized Clinical Trial
How Large Language Models Can Affect Clinical Reasoning: A Randomized Clinical Trial
Abstract Importance LLMs have encoded a vast array of medical knowledge and are being integrated into clinical settings as deci...
Exploring Large Language Models Integration in the Histopathologic Diagnosis of Skin Diseases: A Comparative Study
Exploring Large Language Models Integration in the Histopathologic Diagnosis of Skin Diseases: A Comparative Study
Abstract Introduction The exact manner in which large language models (LLMs) will be integrated into pathology is not yet fully comprehended. This study examines the accuracy, bene...
Human-AI Collaboration in Clinical Reasoning: A UK Replication and Interaction Analysis
Human-AI Collaboration in Clinical Reasoning: A UK Replication and Interaction Analysis
Abstract Objective A paper from Goh et al found that a large language model (LLM) working alone outperformed American clinician...
Pelatihan Dan Pendampingan Perajin Bambu Desa Grujugan Untuk Meningkatkan Kualitas Irat Dan Diversifikasi Produk
Pelatihan Dan Pendampingan Perajin Bambu Desa Grujugan Untuk Meningkatkan Kualitas Irat Dan Diversifikasi Produk
Grujugan Village in Kebumen Regency is bamboo craftsmen village. Eighty percent of the population are bamboo craftsmen. The problems faced are the quality of Irat and product diver...

Back to Top