arXiv cs.CL
· Papers
Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks
arXiv:2608.04286v1 Announce Type: new Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations, through methods such as Retrieval-Augmented Generation (RAG). However, these systems remain susceptible to intrinsic h