Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Design, Results and Industry Implications of the World's First Insurance Large Language Model Evaluation Benchmark
Hua Zhou, Bing Ma, Yufei Zhang +1
This paper comprehensively elaborates on the construction methodology, multi-dimensional evaluation system, and underlying design philosophy of CUFEInse v1.0. Adhering to the princ…
cs.CL2025
SemanticShield: LLM-Powered Audits Expose Shilling Attacks in Recommender Systems
Kaihong Li, Huichi Zhou, Bin Ma +1
Recommender systems (RS) are widely used in e-commerce for personalized suggestions, yet their openness makes them susceptible to shilling attacks, where adversaries inject fake be…