collaborators

5 papers

cs.AI2026

StepOPSD: Step-Aware Online Preference Distillation for Agent Reinforcement Learning

Yanfei Zhang, Xu Lin, Chenglin Wu

Reinforcement learning for multi-turn agents suffers from a credit-assignment mismatch: rewards are sparse and trajectory-level, while success often hinges on a few local decisions…

cs.CL2025

InteractComp: Evaluating Search Agents With Ambiguous Queries

Mingyi Deng, Lijun Huang, Yani Fan +23

Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that user queries are complete and unam…

cs.AI2025

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Yanfei Zhang

Large Language Models (LLMs) have emerged as one of the most significant technological advancements in artificial intelligence in recent years. Their ability to understand, generat…

q-bio.GN2025

Multimodal Modeling of CRISPR-Cas12 Activity Using Foundation Models and Chromatin Accessibility Data

Azim Dehghani Amirabad, Yanfei Zhang, Artem Moskalev +5

Predicting guide RNA (gRNA) activity is critical for effective CRISPR-Cas12 genome editing but remains challenging due to limited data, variation across protospacer adjacent motifs…

cs.CL2025

AIstorian lets AI be a historian: A KG-powered multi-agent system for accurate biography generation

Fengyu Li, Yilin Li, Junhao Zhu +6

Huawei has always been committed to exploring the AI application in historical research. Biography generation, as a specialized form of abstractive summarization, plays a crucial r…