collaborators

5 papers

cs.CL2025

ALLabel: Three-stage Active Learning for LLM-based Entity Recognition using Demonstration Retrieval

Zihan Chen, Lei Shi, Weize Wu +2

Many contemporary data-driven research efforts in the natural sciences, such as chemistry and materials science, require large-scale, high-performance entity recognition from scien…

cs.AI2025

Airalogy: AI-empowered universal data digitization for research automation

Zijie Yang, Qiji Zhou, Fang Guo +19

Research data are the foundation of Artificial Intelligence (AI)-driven science, yet current AI applications remain limited to a few fields with readily available, well-structured,…

cs.CL2025

Pre-DPO: Improving Data Utilization in Direct Preference Optimization Using a Guiding Reference Model

Junshu Pan, Wei Shen, Shulin Huang +2

Direct Preference Optimization (DPO) simplifies reinforcement learning from human feedback (RLHF) for large language models (LLMs) by directly optimizing human preferences without…

cs.CV2025

Reasoning is All You Need for Video Generalization: A Counterfactual Benchmark with Sub-question Evaluation

Qiji Zhou, Yifan Gong, Guangsheng Bao +5

Counterfactual reasoning is crucial for robust video understanding but remains underexplored in existing multimodal benchmarks. In this paper, we introduce \textbf{COVER} (\textbf{…

cs.CL2025

Decoupling Content and Expression: Two-Dimensional Detection of AI-Generated Text

Guangsheng Bao, Lihua Rong, Yanbin Zhao +2

The wide usage of LLMs raises critical requirements on detecting AI participation in texts. Existing studies investigate these detections in scattered contexts, leaving a systemati…