7 citations · 12 across the 13 of their papers we have counts for
4 papers · 1 filter
Advancing Expert Specialization for Better MoE
Hongcan Guo, Haolang Lu, Guoshun Nan +8
Mixture-of-Experts (MoE) models enable efficient scaling of large language models (LLMs) by activating only a subset of experts per input. However, we observe that the commonly use…
Refining Positive and Toxic Samples for Dual Safety Self-Alignment of LLMs with Minimal Human Interventions
Jingxin Xu, Guoshun Nan, Sheng Guan +7
Recent AI agents, such as ChatGPT and LLaMA, primarily rely on instruction tuning and reinforcement learning to calibrate the output of large language models (LLMs) with human inte…
Tell2Design: A Dataset for Language-Guided Floor Plan Generation
Sicong Leng, Yang Zhou, Mohammed Haroon Dupty +3
We consider the task of generating designs directly from natural language descriptions, and consider floor plan generation as the initial research area. Language conditional genera…
Speaker-Oriented Latent Structures for Dialogue-Based Relation Extraction
Guoshun Nan, Guoqing Luo, Sicong Leng +2
Dialogue-based relation extraction (DiaRE) aims to detect the structural information from unstructured utterances in dialogues. Existing relation extraction models may be unsatisfa…