A Trust-Aware Framework for Hallucination Detection and Accuracy Paradox Analysis in Large Language Models
This paper introduces TrustSLM, a trust-aware framework that aggregates outputs from multiple large language models using reliability metrics to detect hallucinations, identify the accuracy paradox of overconfident errors, and regulate responses through a controlled abstention policy for safer AI deployment.