2 papers
cs.CL2026
Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models
S M Tahmid Siddiqui, Akib Jawad Ononto, Anoop Singhal +1
Large language models (LLMs) have seen widespread adoption across various domains, yet their reliability is frequently undermined by hallucinations - responses that are plausible-s…
cs.CV2026
Vision Token Reduction via Attention-Driven Self-Compression for Efficient Multimodal Large Language Models
Omer Faruk Deniz, Ruiyu Mao, Ruochen Li +2
Multimodal Large Language Models (MLLMs) incur significant computational cost from processing numerous vision tokens through all LLM layers. Prior pruning methods operate either be…