activity
20242026
most citedDownstream-Pretext Domain Knowledge Traceback for Active Learning

3 citations · 4 across the 5 of their papers we have counts for

collaborators

7 papers

cs.AI2026

ARISE-RL: Agentic Rubric-Grounded Iterative Self-Evolution with Reinforcement Learning

Fanrui Zhang, Ruixue Ding, Qiang Zhang +13

Training open-ended agents via reinforcement learning (RL) is hindered by the lack of verifiable gold answers and scalable rubrics. Moreover, even near the model's capability bound…

cs.LG2026

ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking

Qiang Zhang, Boli Chen, Fanrui Zhang +14

Reinforcement learning has substantially improved the performance of LLM agents on tasks with verifiable outcomes, but it still struggles on open-ended agent tasks with vast soluti…

cs.CV2025

Fact-R1: Towards Explainable Video Misinformation Detection with Deep Reasoning

Fanrui Zhang, Dian Li, Qiang Zhang +6

The rapid spread of multimodal misinformation on social media has raised growing concerns, while research on video misinformation detection remains limited due to the lack of large…

cs.CV2024

ForgeryGPT: A Multimodal LLM for Interpretable Image Forgery Detection and Localization

Fanrui Zhang, Jiawei Liu, Jiaying Zhu +4

Multimodal Large Language Models (MLLMs), such as GPT4o, have shown strong capabilities in visual reasoning and explanation generation. However, despite these strengths, they face…

cs.LG20243 cited

Downstream-Pretext Domain Knowledge Traceback for Active Learning

Beichen Zhang, Liang Li, Zheng-Jun Zha +2

Active learning (AL) is designed to construct a high-quality labeled dataset by iteratively selecting the most informative samples. Such sampling heavily relies on data representat…

cs.SI20241 cited

Hierarchical Information Enhancement Network for Cascade Prediction in Social Networks

Fanrui Zhang, Jiawei Liu, Qiang Zhang +2

Understanding information cascades in networks is a fundamental issue in numerous applications. Current researches often sample cascade information into several independent paths o…