collaborators

10 papers

cs.SD2026

Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning

Fangxu Yu, Tao Feng, Dehai Min +6

Audio reasoning is essential for machine understanding of the acoustic world. Reinforcement learning with verifiable rewards can elicit such reasoning, yet existing reward designs…

cs.LG2026

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning

Fangxu Yu, Tao Feng, Dehai Min +3

Time series reasoning is essential for real-world problem-solving. While both Large Language Models (LLMs) and Vision-Language Models (VLMs) can reason about time-series data, thei…

cs.LG2026

Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction

Yuheng Yang, Siqi Zhu, Tao Feng +2

Large Language Models (LLMs) can be seen as compressed knowledge bases, but it remains unclear what knowledge they truly contain and how far their knowledge boundary extends. Exist…

cs.LG2026

MemReward: Graph-Based Experience Memory for LLM Reward Prediction with Limited Labels

Tianyang Luo, Tao Feng, Zhigang Hua +4

Reinforcement learning has emerged as a powerful paradigm for improving large language model (LLM) reasoning, where rollouts are sampled from the policy and reward signals computed…

cs.IR2026

UniRec: Unified Multimodal Encoding for LLM-Based Recommendations

Zijie Lei, Tao Feng, Zhigang Hua +5

Large language models have recently shown promise for multimodal recommendation, particularly with text and image inputs. Yet real-world recommendation signals extend far beyond th…

cs.CL2025

ResearchArcade: Graph Interface for Academic Tasks

Jingjun Xu, Chongshan Lin, Haofei Yu +2

Academic research generates diverse data sources, and as researchers increasingly use machine learning to assist research tasks, a crucial question arises: Can we build a unified d…