跳至主導覽 跳至搜尋 跳過主要內容

Enhancing LLM Question Answering with RAG through Dense Vector Search and Re-Ranking

  • Te Lun Yang*
  • , Jyi Shane Liu
  • , Yuen Hsien Tseng
  • , Jyh Shing Jang
  • , Ming Ching Chang
  • , Wei Chao Chen
  • *此作品的通信作者

研究成果: 雜誌貢獻會議論文同行評審

摘要

Retrieval-Augmented Generation (RAG) has emerged as a powerful framework for enhancing Large Language Models (LLMs) by incorporating external knowledge through information retrieval (IR) techniques. However, in question-answering tasks, RAG often retrieves documents that are only semantically similar to the query, which may not provide the most relevant information for generating accurate responses. To address this limitation, we propose an improved retrieval pipeline that combines dense vector search with a re-ranking mechanism to more effectively identify and extract highly relevant knowledge from the retrieved content. We evaluated our approach on two Chinese datasets, TTQA and TMMLU+, using 17 different LLMs. Experimental results show that our method improves performance by up to 21.24% over baseline approaches, particularly on two finance-related subsets, after incorporating domain-specific financial regulations to enhance the knowledge base used in the TMMLU+ dataset.

ASJC Scopus subject areas

  • 訊號處理
  • 媒體技術
  • 電腦視覺和模式識別
  • 人工智慧

指紋

深入研究「Enhancing LLM Question Answering with RAG through Dense Vector Search and Re-Ranking」主題。共同形成了獨特的指紋。

引用此