16 papers
RADIANCE: Relative Adaptive Denoising with IP-Adapter for Novel Concept Enhancement
Zi-Xiang Ni, Bo-Lun Huang, Teng-Fang Hsiao +2
Text-to-image (T2I) diffusion models have achieved striking progress but still struggle to synthesize rare concepts involving unusual attribute-object pairings, often resulting in…
VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single Image
Teng-Fang Hsiao, Bo-Kai Ruan, Yu-Lun Liu +1
3D editing has emerged as a critical research area to provide users with flexible control over 3D assets. While current editing approaches predominantly focus on 3D Gaussian Splatt…
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
Teng-Fang Hsiao, Bo-Kai Ruan, Yi-Lun Wu +2
Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation. Existing methods often partially…
HyperPatch: Sequential Knowledge Editing Under n-ary Structural Drift
Yu-Kai Chan, Wen-Sheng Lien, Dong-Ting Yao +4
Large Language Models (LLMs) rely on Knowledge Editing (KE) to maintain temporal validity, yet real-world knowledge is inherently n-ary. We demonstrate that in non-stationary envir…
Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models
Bo-Kai Ruan, Teng-Fang Hsiao, Ling Lo +1
World Action Models (WAMs) enable decision-making through imagined rollouts by predicting future observations and actions. However, the reliability of these imagined futures remain…
Traffic Scene Generation from Natural Language Description for Autonomous Vehicles with Large Language Model
Bo-Kai Ruan, Hao-Tang Tsui, Yung-Hui Li +1
Generating realistic and controllable traffic scenes from natural language can greatly enhance the development and evaluation of autonomous driving systems. However, this task pose…