2 papers
cs.AR2025
FIXME: Towards End-to-End Benchmarking of LLM-Aided Design Verification
Gwok-Waa Wan, Shengchu Su, Ruihu Wang +15
Despite the transformative potential of Large Language Models (LLMs) in hardware design, a comprehensive evaluation of their capabilities in design verification remains underexplor…
cs.CL2024
MM-Eval: A Hierarchical Benchmark for Modern Mongolian Evaluation in LLMs
Mengyuan Zhang, Ruihui Wang, Bo Xia +2
Large language models (LLMs) excel in high-resource languages but face notable challenges in low-resource languages like Mongolian. This paper addresses these challenges by categor…