Javascript must be enabled to continue!
Bi-Contextual Retrieval Augmented Generation (RAG) for Automatic Descriptive Answer Grading
View through CrossRef
Automatic Short Answer Grading (ASAG) is a well-known research task in the field of
natural language processing (NLP). Its major purpose is to automatically grade descriptive answers of the students by keeping automatic grading consistent with the evaluation
of human graders. Recent developments in Large Language Models (LLMs) have demonstrated a greatly enhanced performance in automated grading; however, the generalizability
of the models and accuracy is still quite low because of the absence of dataset-specific
grounding. We present EDURAG, a Retrieval-Augmented Generation (RAG) based model
to improve contextualization of the LLM-based grading with exemplar-based grading and
extra knowledge as generated by QFKE (Question Focused Knowledge Extraction) module.
The proposed QFKE module provides extra layer of contextuality for the EDURAG. In
contrast to conventional supervised methods, EDURAG does not need model fine-tuning.
The suggested framework is tested against the ASAG2024 benchmark that consolidates seven
short-answer grading datasets across various domains, educational levels, and grading scales.
The benchmark protocol of measuring performance is weighted Root Mean Square Error
(wRMSE). The experimental findings show that dual contextuality provided by EDURAG
enhances the accuracy of grading significantly when compared to vanilla LLM grading.
Advances in Artificial Intelligence and Machine Learning
Title: Bi-Contextual Retrieval Augmented Generation (RAG) for Automatic Descriptive Answer Grading
Description:
Automatic Short Answer Grading (ASAG) is a well-known research task in the field of
natural language processing (NLP).
Its major purpose is to automatically grade descriptive answers of the students by keeping automatic grading consistent with the evaluation
of human graders.
Recent developments in Large Language Models (LLMs) have demonstrated a greatly enhanced performance in automated grading; however, the generalizability
of the models and accuracy is still quite low because of the absence of dataset-specific
grounding.
We present EDURAG, a Retrieval-Augmented Generation (RAG) based model
to improve contextualization of the LLM-based grading with exemplar-based grading and
extra knowledge as generated by QFKE (Question Focused Knowledge Extraction) module.
The proposed QFKE module provides extra layer of contextuality for the EDURAG.
In
contrast to conventional supervised methods, EDURAG does not need model fine-tuning.
The suggested framework is tested against the ASAG2024 benchmark that consolidates seven
short-answer grading datasets across various domains, educational levels, and grading scales.
The benchmark protocol of measuring performance is weighted Root Mean Square Error
(wRMSE).
The experimental findings show that dual contextuality provided by EDURAG
enhances the accuracy of grading significantly when compared to vanilla LLM grading.
Related Results
JADE: jawbone lesion diagnosis and decision supporting system
JADE: jawbone lesion diagnosis and decision supporting system
Abstract
Objectives
To develop and evaluate JADE, a proof-of-concept retrieval-augmented generation (RAG) diagnostic assi...
A Systematic Literature Review of Retrieval-Augmented Generation Implementation for Enhancing Large Language Models in Education
A Systematic Literature Review of Retrieval-Augmented Generation Implementation for Enhancing Large Language Models in Education
The rapid advancement of Large Language Models (LLM) has led to the creation of increasingly adaptive intelligent learning systems. However, many educational implementations of LLM...
Automating Information Retrieval from Biodiversity Literature Using Large Language Models: A Case Study
Automating Information Retrieval from Biodiversity Literature Using Large Language Models: A Case Study
Recently, Large Language Models (LLMs) have transformed information retrieval, becoming widely adopted across various domains due to their ability to process extensive textual data...
Study on radiographic grading of ankle joint in adult patients with Kashin Beck disease in Shaanxi and Gansu Province, China
Study on radiographic grading of ankle joint in adult patients with Kashin Beck disease in Shaanxi and Gansu Province, China
Abstract
Purpose
This paper aims to establish an X-ray imaging grading for assessing ankle joints in adult Kashin Beck disease (KBD), and investigate its correlation with ...
Passage, Sentence, or Proposition? An Empirical Comparison of Retrieval Granularity Effects on LLM Answer Accuracy in Retrieval-Augmented Generation
Passage, Sentence, or Proposition? An Empirical Comparison of Retrieval Granularity Effects on LLM Answer Accuracy in Retrieval-Augmented Generation
Retrieval-Augmented Generation (RAG) has become a dominant paradigm for grounding large language model (LLM) outputs in external knowledge. While extensive research has focused on ...
Performance Analysis of Transformer Based Models for Automatic Short Answer Grading
Performance Analysis of Transformer Based Models for Automatic Short Answer Grading
Automatic Short Answer Grading (ASAG) has gained increasing importance in educational technology, where accurate and scalable assessment solutions are needed. Recent advances in Na...
Performance Analysis of Transformer Based Models for Automatic Short Answer Grading
Performance Analysis of Transformer Based Models for Automatic Short Answer Grading
Automatic Short Answer Grading (ASAG) has gained increasing importance in educational technology, where accurate and scalable assessment solutions are needed. Recent advances in Na...
SMART RESPONSE: RAG ENHANCED QUESTION ANSWERING MODEL
SMART RESPONSE: RAG ENHANCED QUESTION ANSWERING MODEL
Smart Response: RAG Enhanced Question Answering Model', aims to revolutionize question answering systems by integrating RetrievalAugmented Generation (RAG). RAG synergizes a retrie...

