1 paper
Masnun Nuha Chowdhury, Nusrat Jahan Beg, Umme Hunny Khan +3
Large language models (LLMs) remain unreliable for high-stakes claim verification due to hallucinations and shallow reasoning. While retrieval-augmented generation (RAG) and multi-…