3 papers
cs.LG2025
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models
Khizar Anjum, Muhammad Arbab Arshad, Kadhim Hayawi +10
Large language models (LLMs) are increasingly being deployed across disciplines due to their advanced reasoning and problem solving capabilities. To measure their effectiveness, va…
cs.CR2025
Zero Trust Cybersecurity: Procedures and Considerations in Context
Brady D. Lund, Tae Hee Lee, Ziang Wang +2
In response to the increasing complexity and sophistication of cyber threats, particularly those enhanced by advancements in artificial intelligence, traditional security methods a…
cs.CL2024
Cause and Effect: Can Large Language Models Truly Understand Causality?
Swagata Ashwani, Kshiteesh Hegde, Nishith Reddy Mannuru +6
With the rise of Large Language Models(LLMs), it has become crucial to understand their capabilities and limitations in deciphering and explaining the complex web of causal relatio…