7 papers
When to Use Extra Context: Evidence-Grounded Terminology Adaptation for Simultaneous Speech Translation
Zeyu Yang, Satoshi Nakamura
Extra context is valuable for simultaneous speech translation of technical talks, but injecting the entire document context into every streaming segment is often too coarse. Throug…
ConsisGuard: Aligning Safety Deliberation with Policy Enforcement in LLM Guardrails
Yan Wang, Zhixuan Chu, Zihao Xue +9
Reasoning-based LLM guardrails improve safety moderation by generating explicit rationales before issuing final decisions. However, their rationales do not always lead to faithful…
Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers
Zihao Xue, Yan Wang, Zhen Bi +7
Diffusion Transformers have become a powerful backbone for text-to-image generation, but their layered and cross-modal generation process makes safety control fundamentally differe…
Make LLM Learn to Synthesize from Streaming Experiences through Feedback
Zhenlin Hu, Yan Wang, Zhen Bi +7
Large language models (LLMs) have been widely adopted for synthetic data generation, significantly reducing annotation costs. However, most existing studies treat synthesis as a se…
Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training
Yishun Lu, Junhao Zhang, Zeyu Yang +1
Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost of maintaining and updating la…
DPO-Tuned Large Language Models for Segmentation in Simultaneous Speech Translation
Zeyu Yang, Satoshi Nakamura
Simultaneous speech translation requires accurate segmentation to balance translation quality and latency. Recent studies such as SHAS have introduced pretrained segmentation model…