9 citations · 16 across the 7 of their papers we have counts for
17 papers
LoGAN: Multilingual Font Localization with Generative Agents
Zhuoning Yuan, Ta-Ying Cheng, Benjamin Klein
Localizing a font into new languages is a highly intricate task requiring precise design adaptation of glyphs, color/texture, and spacing/kerning, from source to target languages.…
Vera: A Layered Diffusion Model for Content-Preserving Video Editing
Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein +2
Video diffusion models have enabled remarkable progress in video generation and editing. However, content preservation remains a core challenge: existing methods regenerate every p…
VOID: Video Object and Interaction Deletion
Saman Motamed, William Harvey, Benjamin Klein +3
Existing video object removal methods excel at inpainting content "behind" the object and correcting appearance-level artifacts such as shadows and reflections. However, when the r…
CLoVe: Encoding Compositional Language in Contrastive Vision-Language Models
Santiago Castro, Amir Ziai, Avneesh Saluja +2
Recent years have witnessed a significant increase in the performance of Vision and Language tasks. Foundational Vision-Language Models (VLMs), such as CLIP, have been leveraged in…
LibAUC: A Deep Learning Library for X-Risk Optimization
Zhuoning Yuan, Dixian Zhu, Zi-Hao Qiu +3
This paper introduces the award-winning deep learning (DL) library called LibAUC for implementing state-of-the-art algorithms towards optimizing a family of risk functions named X-…
Not All Semantics are Created Equal: Contrastive Self-supervised Learning with Automatic Temperature Individualization
Zi-Hao Qiu, Quanqi Hu, Zhuoning Yuan +3
In this paper, we aim to optimize a contrastive loss with individualized temperatures in a principled and systematic manner for self-supervised learning. The common practice of usi…