Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Structured Multi-Criteria Evaluation of Large Language Models with Fuzzy Analytic Hierarchy Process and DualJudge
Yulong He, Ivan Smirnov, Dmitry Fedrushkov +2
Effective evaluation of large language models (LLMs) remains a critical bottleneck, as conventional direct scoring often yields inconsistent and opaque judgments. In this work, we…
cs.AI2026
Style2Code: A Style-Controllable Code Generation Framework with Dual-Modal Contrastive Representation Learning
Dutao Zhang, Nicolas Rafael Arroyo Arias, YuLong He +1
Controllable code generation, the ability to synthesize code that follows a specified style while maintaining functionality, remains a challenging task. We propose a two-stage trai…