LoRA-Based Fine-Tuning of Local LLMs for Hallucination Detection in Indonesian RAG Systems

0Citations
Citations of this article
7Readers
Mendeley users who have this article in their library.

Abstract

—Retrieval Augmented Generation (RAG) improves the factual grounding of Large Language Models (LLMs) by incorporating external knowledge. However, RAG systems may still generate hallucinated responses, and this issue remains underexplored in Indonesian language settings, particularly in settings where local deployment is preferred. This study proposes a hallucination detection approach for Indonesian RAG systems using Low Rank Adaptation (LoRA) fine-tuning. To support this objective, the study constructs a dataset in the Human-Computer Interaction domain consisting of 908 context, question, and answer pairs. The dataset is classified into four categories: FACT-H, FAITH-H, LOG-H, and FAITHFUL. Three local LLMs, namely, Gemma-7B-it, LlaMA-2-7B chat, and Phi-3-medium-4k-instruct, were evaluated using 5-fold cross-validation. The results show that Gemma-7B-it achieved the best performance in the four-class setting, with a Macro F1 score of 0.846. In the binary classification setting, Gemma achieved an accuracy of 98.1 per cent. Further analysis shows that Gemma was particularly effective in recognizing FAITHFUL, FAITH-H, and FACT-H, while LOG-H remained the most difficult class to distinguish consistently.

Cite

CITATION STYLE

APA

Arthana, I. K. R., Gunantara, N., Sudarma, M., & Sukarsa, M. (2026). LoRA-Based Fine-Tuning of Local LLMs for Hallucination Detection in Indonesian RAG Systems. International Journal of Advanced Computer Science and Applications, 17(3), 1000–1009. https://doi.org/10.14569/IJACSA.2026.0170389

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free