6 papers
PhotoQuilt: Training-Free Arbitrary-Resolution Photomosaics via Bootstrapped Tiled Denoising
Koorosh Roohi, Javad Rajabi, Andrew Fleet +1
Photomosaics are large images whose local regions are seen as independent tiles while their overall arrangement forms a coherent scene. Generating them at high resolution, with eve…
A Unified Theory of Sinusoidal Activation Families for Implicit Neural Representations
Alireza Morsali, MohammadJavad Vaez, Mohammadhossein Soltani +3
Implicit Neural Representations (INRs) model continuous signals with compact neural networks and have become a standard tool in vision, graphics, and signal processing. A central c…
AVIS: Adaptive Test-Time Scaling for Vision-Language Models
Ahmadreza Jeddi, Minh Ngoc Le, Amirhossein Kazerouni +8
Modern Vision-Language Models (VLMs) benefit from chain-of-thought prompting and test-time scaling, but these gains often come with prohibitive inference cost due to large visual c…
SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers
Javad Rajabi, Kimia Shaban, Koorosh Roohi +2
Diffusion transformers (DiTs) have emerged as a dominant architecture for text-to-image generation, yet their performance drops when generating at resolutions beyond their training…
Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration
Amirhossein Kazerouni, Maitreya Suin, Tristan Aumentado-Armstrong +6
Recent advances in image restoration have enabled high-fidelity recovery of faces from degraded inputs using reference-based face restoration models (Ref-FR). However, such methods…
Mechanism of Shape Symmetry Breaking in Surfactant Mediated Crystal Growth
Sam Oaks-Leaf, David T. Limmer
We present a dynamical model of crystal growth, in which it is possible to reliably achieve asymmetric products, beginning from symmetric initial conditions and growing within an i…