collaborators

5 papers

cs.AI2026

Quark Medical Alignment: A Holistic Multi-Dimensional Alignment and Collaborative Optimization Paradigm

Tianxiang Xu, Jiayi Liu, Yixuan Tong +10

While reinforcement learning for large language model alignment has progressed rapidly in recent years, transferring these paradigms to high-stakes medical question answering revea…

cs.IR2026

DualGR: Generative Retrieval with Long and Short-Term Interests Modeling

Zhongchao Yi, Kai Feng, Xiaojian Ma +5

In large-scale industrial recommendation systems, retrieval must produce high-quality candidates from massive corpora under strict latency. Recently, Generative Retrieval (GR) has…

cs.CY2025

On the Regulatory Potential of User Interfaces for AI Agent Governance

K. J. Kevin Feng, Tae Soo Kim, Rock Yuren Pang +3

AI agents that take actions in their environment autonomously over extended time horizons require robust governance interventions to curb their potentially consequential risks. Pri…

cs.HC2025

Interactive Reasoning: Visualizing and Controlling Chain-of-Thought Reasoning in Large Language Models

Rock Yuren Pang, K. J. Kevin Feng, Shangbin Feng +5

The output quality of large language models (LLMs) can be improved via "reasoning": generating segments of chain-of-thought (CoT) content to further condition the model prior to pr…

cs.CL2025

InsQABench: Benchmarking Chinese Insurance Domain Question Answering with Large Language Models

Jing Ding, Kai Feng, Binbin Lin +6

The application of large language models (LLMs) has achieved remarkable success in various fields, but their effectiveness in specialized domains like the Chinese insurance industr…