activity
20192026
most citedWHERE and WHICH: Iterative Debate for Biomedical Synthetic Data Augmentation

2 citations · 2 across the 14 of their papers we have counts for

collaborators
Showing cs.CLShow all

11 papers · 1 filter

cs.CL2026

Beyond the Literal: Decomposing Pragmatic Intent in Multimodal Meme Understanding

Zhengyi Zhao, Shubo Zhang, Zezhong Wang +6

When asked what a meme or sarcastic post means, Large Vision Language Models (LVLMs) tend to describe what the image shows rather than what the author is trying to communicate. Sta…

cs.CL2026

Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generation

Zhengyi Zhao, Shubo Zhang, Zezhong Wang +7

Retrieval-Augmented Generation (RAG) significantly enhances Large Language Models (LLMs) by providing access to external knowledge. However, current research primarily focuses on r…

cs.CL2025

Step-DeepResearch Technical Report

Chen Hu, Haikuo Du, Heng Wang +64

As LLMs shift toward autonomous agents, Deep Research has emerged as a pivotal metric. However, existing academic benchmarks like BrowseComp often fail to meet real-world demands f…

cs.CL2025

Dual-Density Inference for Efficient Language Model Reasoning

Zhengyi Zhao, Shubo Zhang, Yuxi Zhang +3

Large Language Models (LLMs) have shown impressive capabilities in complex reasoning tasks. However, current approaches employ uniform language density for both intermediate reason…

cs.CL2025

T: An Adaptive Test-Time Scaling Strategy for Contextual Question Answering

Zhengyi Zhao, Shubo Zhang, Zezhong Wang +7

Recent advances in Large Language Models (LLMs) have demonstrated remarkable performance in Contextual Question Answering (CQA). However, prior approaches typically employ elaborat…

cs.CL20252 cited

WHERE and WHICH: Iterative Debate for Biomedical Synthetic Data Augmentation

Zhengyi Zhao, Shubo Zhang, Bin Liang +2

In Biomedical Natural Language Processing (BioNLP) tasks, such as Relation Extraction, Named Entity Recognition, and Text Classification, the scarcity of high-quality data remains…