Publications (25)
MAT: Multi-Range Attention Transformer for Efficient Image Super-Resolution
Chengxing Xie, Xiaoming Zhang, Linze Li +4
Image super-resolution (SR) has significantly advanced through the adoption of Transformer architectures. However, conventional techniques aimed at enlarging the self-attention win…
Ultra-fast Waveguide MUTC Photodiodes over 220 GHz
Linze Li, Luyu Wang, Tianyu Long +3
We present InP-based evanescently-coupled waveguide modified uni-traveling carrier photodiodes (MUTC-PDs) exhibiting a breakthrough in bandwidth. The optimization of carrier transp…
Efficient Single Image Super-Resolution with Entropy Attention and Receptive Field Augmentation
Xiaole Zhao, Linze Li, Chengxing Xie +5
Transformer-based deep models for single image super-resolution (SISR) have greatly improved the performance of lightweight SISR tasks in recent years. However, they often suffer f…
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
Shurong Yang, Huadong Li, Juhao Wu +5
Despite raw driving videos contain richer information on facial expressions than intermediate representations such as landmarks in the field of portrait animation, they are seldom…
Follow the Leader: Enhancing Systematic Trend-Following Using Network Momentum
Linze Li, William Ferreira
We present a systematic, trend-following strategy, applied to commodity futures markets, that combines univariate trend indicators with cross-sectional trend indicators that captur…
TENet: Targetness Entanglement Incorporating with Multi-Scale Pooling and Mutually-Guided Fusion for RGB-E Object Tracking
Pengcheng Shao, Tianyang Xu, Zhangyong Tang +3
There is currently strong interest in improving visual object tracking by augmenting the RGB modality with the output of a visual event camera that is particularly informative abou…
The Velocity Deficit: Initial Energy Injection for Flow Matching
Linze Li, Zong-Wei Hong, Shen Zhang +4
While Flow Matching theoretically guarantees constant-velocity trajectories, we identify a critical breakdown in high-dimensional practice: the Velocity Deficit. We show that the M…
LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding
Shen Zhang, Siyuan Liang, Yaning Tan +9
Diffusion transformers (DiTs) struggle to generate images at resolutions higher than their training resolutions. The primary obstacle is that the explicit positional encodings(PE),…
SPARE: Structural Parameter-Free Affinity Regularization for Flow Matching
Zong-Wei Hong, Jinglun Li, Shen Zhang +3
Denoising diffusion transformers achieve strong generation quality but converge slowly during training. Regularizing their internal representations has emerged as an effective acce…
Optimizing Knowledge Distillation in Transformers: Enabling Multi-Head Attention without Alignment Barriers
Zhaodong Bing, Linze Li, Jiajun Liang
Knowledge distillation (KD) in transformers often faces challenges due to misalignment in the number of attention heads between teacher and student models. Existing methods either…
A chip-based optoelectronic-oscillator frequency comb
Jinbao Long, Zhongkai Wang, Huanfa Peng +14
Microresonator-based Kerr frequency combs ("Kerr microcombs") constitute chip-scale frequency combs of broad spectral bandwidth and repetition rate ranging from gigahertz to terahe…
Composition-spread Growth and the Robust Topological Surface State of Kondo insulator SmB6 Thin Films
Jie Yong, Yeping Jiang, Demet Usanmaz +7
Topological insulators are a class of materials with insulating bulk but protected conducting surfaces due to the combination of spin-orbit interactions and time-reversal symmetry.…
MegActor-: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
Shurong Yang, Huadong Li, Juhao Wu +6
Diffusion models have demonstrated superior performance in the field of portrait animation. However, current approaches relied on either visual or audio modality to control charact…
Spontaneous Hall Effect enhanced by local Ir moments in epitaxial PrIrO thin films
Lu Guo, Neil Campbell, Yongseong Choi +10
Rare earth pyrochlore Iridates (RE2Ir2O7) consist of two interpenetrating cation sublattices, the RE with highly-frustrated magnetic moments, and the Iridium with extended conducti…
All-optical control and multiplexed readout of multiple superconducting qubits
Xiaoxuan Pan, Chuanlong Ma, Jia-Qi Wang +12
Superconducting quantum circuits operate at millikelvin temperatures, typically requiring independent microwave cables for each qubit for connecting room-temperature control and re…
Large Kernel Distillation Network for Efficient Single Image Super-Resolution
Chengxing Xie, Xiaoming Zhang, Linze Li +4
Efficient and lightweight single-image super-resolution (SISR) has achieved remarkable performance in recent years. One effective approach is the use of large kernel designs, which…
Modified uni-travelling-carrier photodiodes with 206 GHz bandwidth and 0.81 A/W external responsivity
Linze Li, Tianyu Long, Xiongwei Yang +7
The accelerating demand for wireless communication necessitates wideband, energy-efficient photonic sub-terahertz (sub-THz) sources to enable ultra-fast data transfer. However, as…
NTIRE 2026 Challenge on Video Saliency Prediction: Methods and Results
Andrey Moskalenko, Alexey Bryncev, Ivan Kosmynin +40
This paper presents an overview of the NTIRE 2026 Challenge on Video Saliency Prediction. The goal of the challenge participants was to develop automatic saliency map prediction me…
Functionalized Graphene for High Performance Two-dimensional Spintronics Devices
Linze Li, Rui Qin, Hong Li +5
Using first-principles calculations, we explore the possibility of functionalized graphene as high performance two-dimensional spintronics device. Graphene functionalized with O on…
Scalable Optical Links for Controlling Bosonic Quantum Processors
Chuanlong Ma, Jia-Qi Wang, Linze Li +13
Superconducting quantum computing has the potential to revolutionize computational capabilities. However, scaling up large quantum processors is limited by the cumbersome and heat-…
C2C: Component-to-Composition Learning for Zero-Shot Compositional Action Recognition
Rongchang Li, Zhenhua Feng, Tianyang Xu +5
Compositional actions consist of dynamic (verbs) and static (objects) concepts. Humans can easily recognize unseen compositions using the learned concepts. For machines, solving su…
Efficient One Pass Self-distillation with Zipf's Label Smoothing
Jiajun Liang, Linze Li, Zhaodong Bing +4
Self-distillation exploits non-uniform soft supervision from itself during training and improves performance without any runtime cost. However, the overhead during training is ofte…
A chip-integrated comb-based microwave oscillator
Wei Sun, Zhiyang Chen, Linze Li +12
Low-noise microwave oscillators are cornerstones for wireless communication, radar and clocks. Optical frequency combs have enabled photonic microwaves with unrivalled noise perfor…
Stable Spike: Dual Consistency Optimization via Bitwise AND Operations for Spiking Neural Networks
Yongqi Ding, Kunshan Yang, Linze Li +3
Although the temporal spike dynamics of spiking neural networks (SNNs) enable low-power temporal pattern capture capabilities, they also incur inherent inconsistencies that severel…
FAAC: Facial Animation Generation with Anchor Frame and Conditional Control for Superior Fidelity and Editability
Linze Li, Sunqi Fan, Hengjun Pu +6
Over recent years, diffusion models have facilitated significant advancements in video generation. Yet, the creation of face-related videos still confronts issues such as low facia…