2 citations · 2 across the 4 of their papers we have counts for
4 papers · 1 filter
Bridging Online and Offline Handwriting via Differentiable Physical Rendering
Seonmi Park, Seunghyun Shin, Vihaan Misra +4
Realistic handwritten text generation plays an important role in numerous applications, such as font design, biometric authentication, and robotic calligraphy. Existing methods are…
TokenDial: Continuous Attribute Control for Text-to-Video Generation in Visual Dial Space
Zhixuan Liu, Peter Schaldenbrand, Yijun Li +5
In video diffusion transformers, visual patch tokens maintain explicit correspondence to space and time. We hypothesize that their channel dimension can serve as a semantic control…
ShapeShift: Text-to-Mosaic Synthesis via Semantic Phase-Field Guidance
Vihaan Misra, Peter Schaldenbrand, Jean Oh
We present ShapeShift, a method for arranging rigid objects into configurations that visually convey semantic concepts specified by natural language. While pretrained diffusion mod…
SCoFT: Self-Contrastive Fine-Tuning for Equitable Image Generation
Zhixuan Liu, Peter Schaldenbrand, Beverley-Claire Okogwu +5
Accurate representation in media is known to improve the well-being of the people who consume it. Generative image models trained on large web-crawled datasets such as LAION are kn…