Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond
Rubing Chen, Jiaxin Wu, Jian Wang +5
The increasing demand for domain-specific evaluation of large language models (LLMs) has led to the development of numerous benchmarks. These efforts often adhere to the principle…
cs.AI2025
MolGround: A Benchmark for Molecular Grounding
Jiaxin Wu, Ting Zhang, Rubing Chen +4
Current molecular understanding approaches predominantly focus on the descriptive aspect of human perception, providing broad, topic-level insights. However, the referential aspect…
cs.AI2025
Knowledge Pyramid Construction for Multi-Level Retrieval-Augmented Generation
Rubing Chen, Xulu Zhang, Jiaxin Wu +3
This paper addresses the need for improved precision in existing knowledge-enhanced question-answering frameworks, specifically Retrieval-Augmented Generation (RAG) methods that pr…