4 citations · 4 across the 5 of their papers we have counts for
5 papers
Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories
Mohammad Beigi, Ying Shen, Parshin Shojaee +5
Despite the remarkable capabilities of large language models, current training paradigms inadvertently foster \textit{sycophancy}, i.e., the tendency of a model to agree with or re…
Position: AI Safety Must Embrace an Antifragile Perspective
Ming Jin, Hyunin Lee
This position paper contends that modern AI research must adopt an antifragile perspective on safety -- one in which the system's capacity to guarantee long-term AI safety such as…
Rethinking Hardware Impairments in Multi-User Systems: Can FAS Make a Difference?
Junteng Yao, Tuo Wu, Liaoshi Zhou +7
In this paper, we analyze the role of fluid antenna systems (FAS) in multi-user systems with hardware impairments (HIs). Specifically, we investigate a scenario where a base statio…
Optimization Solution Functions as Deterministic Policies for Offline Reinforcement Learning
Vanshaj Khattar, Ming Jin
Offline reinforcement learning (RL) is a promising approach for many control applications but faces challenges such as limited data coverage and value function overestimation. In t…
Slow light silicon modulator beyond 110 GHz bandwidth
Changhao Han, Zhao Zheng, Haowen Shu +14
Silicon modulators are key components in silicon photonics to support the dense integration of electro-optic (EO) functional elements on a compact chip for various applications inc…