Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Can We Trust a Black-box LLM? LLM Untrustworthy Boundary Detection via Bias-Diffusion and Multi-Agent Reinforcement Learning
Xiaotian Zhou, Di Tang, Xiaofeng Wang +1
Large Language Models (LLMs) have shown a high capability in answering questions on a diverse range of topics. However, these models sometimes produce biased, ideologized or incorr…
cs.AI2026
LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection
Akram Hossain, Rabab Abdelfattah, Xiaofeng Wang +1
The deployment of lightweight segmentation models on drones for autonomous power line inspection presents a critical challenge: maintaining reliable performance under real-world co…