5 papers
Quark Medical Alignment: A Holistic Multi-Dimensional Alignment and Collaborative Optimization Paradigm
Tianxiang Xu, Jiayi Liu, Yixuan Tong +10
While reinforcement learning for large language model alignment has progressed rapidly in recent years, transferring these paradigms to high-stakes medical question answering revea…
DualGR: Generative Retrieval with Long and Short-Term Interests Modeling
Zhongchao Yi, Kai Feng, Xiaojian Ma +5
In large-scale industrial recommendation systems, retrieval must produce high-quality candidates from massive corpora under strict latency. Recently, Generative Retrieval (GR) has…
On the Regulatory Potential of User Interfaces for AI Agent Governance
K. J. Kevin Feng, Tae Soo Kim, Rock Yuren Pang +3
AI agents that take actions in their environment autonomously over extended time horizons require robust governance interventions to curb their potentially consequential risks. Pri…
Interactive Reasoning: Visualizing and Controlling Chain-of-Thought Reasoning in Large Language Models
Rock Yuren Pang, K. J. Kevin Feng, Shangbin Feng +5
The output quality of large language models (LLMs) can be improved via "reasoning": generating segments of chain-of-thought (CoT) content to further condition the model prior to pr…
InsQABench: Benchmarking Chinese Insurance Domain Question Answering with Large Language Models
Jing Ding, Kai Feng, Binbin Lin +6
The application of large language models (LLMs) has achieved remarkable success in various fields, but their effectiveness in specialized domains like the Chinese insurance industr…