papers

Publications (96)

cs.CV2024

Learning from the Web: Language Drives Weakly-Supervised Incremental Learning for Semantic Segmentation

Chang Liu, Giulia Rizzoli, Pietro Zanuttigh +2

Current weakly-supervised incremental learning for semantic segmentation (WILSS) approaches only consider replacing pixel-level annotations with image-level labels, while the train…

quant-ph2024

Scaling of quantum Fisher information for quantum exceptional point sensors

Chun-Hui Liu, Fu Li, Shengwang Du +3

In recent years, significant progress has been made in utilizing the divergence of spectrum response rate at the exceptional point (EP) for sensing in classical systems, while the…

cs.GT2022

Egalitarian Resource Sharing Over Multiple Rounds

Fu Li, C. Gregory Plaxton, Vaibhav B. Sinha

It is often beneficial for agents to pool their resources in order to better accommodate fluctuations in individual demand. Many multi-round resource allocation mechanisms operate…

cs.CV2021

CoFiNet: Reliable Coarse-to-fine Correspondences for Robust Point Cloud Registration

Hao Yu, Fu Li, Mahdi Saleh +2

We study the problem of extracting correspondences between a pair of point clouds for registration. For correspondence retrieval, existing works benefit from matching sparse keypoi…

physics.optics2019

Beyond sub-Rayleigh imaging via high order correlation of speckle illumination

Fu Li, Charles Altuzarra, Tian Li +2

Second order intensity correlations of speckle illumination are extensively used in imaging applications that require going beyond the Rayleigh limit. The theoretical analysis show…

eess.SP2025

WVEmbs with its Masking: A Method For Radar Signal Sorting

Xianan Hu, Fu Li, Kairui Niu +2

Our study proposes a novel embedding method, Wide-Value-Embeddings (WVEmbs), for processing Pulse Descriptor Words (PDWs) as normalized inputs to neural networks. This method adapt…

cs.CV2023

NeRF-Pose: A First-Reconstruct-Then-Regress Approach for Weakly-supervised 6D Object Pose Estimation

Fu Li, Hao Yu, Ivan Shugurov +3

Pose estimation of 3D objects in monocular images is a fundamental and long-standing problem in computer vision. Existing deep learning approaches for 6D pose estimation typically…

cs.DC2013

Energy-Aware Aggregation of Dynamic Temporal Workload in Data Centers

Haiyang Qian, Fu Li, Ravishankar Ravindran +1

Data center providers seek to minimize their total cost of ownership (TCO), while power consumption has become a social concern. We present formulations to minimize server energy c…

cs.CV2021

Image Inpainting by End-to-End Cascaded Refinement with Mask Awareness

Manyu Zhu, Dongliang He, Xin Li +5

Inpainting arbitrary missing regions is challenging because learning valid features for various masked regions is nontrivial. Though U-shaped encoder-decoder frameworks have been w…

cs.CV2023

VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Xin Li, Wenqing Chu, Ye Wu +7

In this paper, we present VideoGen, a text-to-video generation approach, which can generate a high-definition video with high frame fidelity and strong temporal consistency using r…

physics.optics2021

Quantum Advantage with Seeded Squeezed Light for Absorption Measurement

Fu Li, Tian Li, Marlan O. Scully +1

Absorption measurement is an exceptionally versatile tool for many applications in science and engineering. For absorption measurements using laser beams of light, the sensitivity…

cs.HC2026

Learning from Brain Topography: A Hierarchical Local-Global Graph-Transformer Network for EEG Emotion Recognition

Yijin Zhou, Fu Li, Yi Niu +3

Understanding how local neurophysiological patterns interact with global brain dynamics is essential for decoding human emotions from EEG signals. However, existing deep learning a…

cs.CV2022

It Takes Two: Masked Appearance-Motion Modeling for Self-supervised Video Transformer Pre-training

Yuxin Song, Min Yang, Wenhao Wu +3

Self-supervised video transformer pre-training has recently benefited from the mask-and-predict pipeline. They have demonstrated outstanding effectiveness on downstream video tasks…

physics.med-ph2022

3-D Stochastic Numerical Breast Phantoms for Enabling Virtual Imaging Trials of Ultrasound Computed Tomography

Fu Li, Umberto Villa, Seonyeong Park +1

Ultrasound computed tomography (USCT) is an emerging imaging modality for breast imaging that can produce quantitative images that depict the acoustic properties of tissues. Comput…

physics.optics2020

Temporal Quantum Noise Reduction Acquired by an Electron-Multiplying Charge-Coupled-Device Camera

Fu Li, Tian Li, Girish S. Agarwal

Electron-multiplying charge-coupled-device cameras (EMCCDs) have been used to observe quantum noise reductions in beams of light in the transverse spatial degree of freedom. For th…

cs.IR2025

Request-Only Optimization for Recommendation Systems

Liang Guo, Wei Li, Lucy Liao +25

Deep Learning Recommendation Models (DLRMs) represent one of the largest machine learning applications on the planet. Industry-scale DLRMs are trained with petabytes of recommendat…

cs.CV2017

Revisiting the Effectiveness of Off-the-shelf Temporal Modeling Approaches for Large-scale Video Classification

Yunlong Bian, Chuang Gan, Xiao Liu +7

This paper describes our solution for the video recognition task of ActivityNet Kinetics challenge that ranked the 1st place. Most of existing state-of-the-art video recognition ap…

cs.GT2021

Object Allocation Over a Network of Objects: Mobile Agents with Strict Preferences

Fu Li, C. Gregory Plaxton, Vaibhav B. Sinha

In recent work, Gourvès, Lesca, and Wilczynski propose a variant of the classic housing markets model where the matching between agents and objects evolves through Pareto-improvin…

cs.CC2015

Characterizing Propositional Proofs as Non-Commutative Formulas

Fu Li, Iddo Tzameret, Zhengyu Wang

Does every Boolean tautology have a short propositional-calculus proof? Here, a propositional calculus (i.e. Frege) proof is a proof starting from a set of axioms and deriving new…

cond-mat.mtrl-sci2025

High-throughput screening of spin Hall conductivity in 2D materials

Fu Li, Xiaoxiong Liu, Vikrant Chaudhary +4

Two-dimensional (2D) materials with large spin Hall effect (SHE) have attracted significant attention due to their potential applications in next-generation spintronic devices. In…

cond-mat.mtrl-sci2019

In-plane anisotropic faceting of ultralarge and thin single-crystalline colloidal SnS nanosheets

Fu Li, Mohammad Mehdi Ramin Moayed, Eugen Klein +2

The colloidal synthesis of large thin two-dimensional (2D) nanosheets is fascinating but challenging, since the growth along the lateral and vertical dimensions need to be controll…

cond-mat.mtrl-sci2026

2D Ferroelectric Ruddlesden-Popper Perovskites: an Emerging Fully Electronically Controllable Shift Current and Persistent Spin Helix

Yue Zhao, Fu Li, Vikrant Chaudhary +5

Two-dimensional (2D) hybrid organic--inorganic perovskites (HOIPs) are promising candidates for next-generation optoelectronic and spintronic applications. This work systematically…

physics.optics2022

Quantum-Enhanced Stimulated Brillouin Scattering Spectroscopy and Imaging

Tian Li, Fu Li, Xinghua Liu +2

Brillouin microscopy is an emerging label-free imaging technique to assess local viscoelastic properties. Quantum-enhanced stimulated Brillouin scattering is demonstrated for the f…

cs.CV2026

CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval

Haoran Wang, Dongliang He, Wenhao Wu +7

Image-Text Retrieval (ITR) is challenging in bridging visual and lingual modalities. Contrastive learning has been adopted by most prior arts. Except for limited amount of negative…

eess.IV2020

NTIRE 2020 Challenge on Perceptual Extreme Super-Resolution: Methods and Results

Kai Zhang, Shuhang Gu, Radu Timofte +60

This paper reviews the NTIRE 2020 challenge on perceptual extreme super-resolution with focus on proposed solutions and results. The challenge task was to super-resolve an input im…

cs.CV2019

Multi-Label Classification with Label Graph Superimposing

Ya Wang, Dongliang He, Fu Li +4

Images or videos always contain multiple objects or actions. Multi-label recognition has been witnessed to achieve pretty performance attribute to the rapid development of deep lea…

cs.RO2025

BeSimulator: A Large Language Model Powered Text-based Behavior Simulator

Jianan Wang, Bin Li, Jingtao Qi +3

Traditional robot simulators focus on physical process modeling and realistic rendering, often suffering from high computational costs, inefficiencies, and limited adaptability. To…

cs.CV2022

RRSR:Reciprocal Reference-based Image Super-Resolution with Progressive Feature Alignment and Selection

Lin Zhang, Xin Li, Dongliang He +3

Reference-based image super-resolution (RefSR) is a promising SR branch and has shown great potential in overcoming the limitations of single image super-resolution. While previous…

cs.CV2020

A Novel Transferability Attention Neural Network Model for EEG Emotion Recognition

Yang Li, Boxun Fu, Fu Li +2

The existed methods for electroencephalograph (EEG) emotion recognition always train the models based on all the EEG samples indistinguishably. However, some of the source (trainin…

cond-mat.mtrl-sci2026

Strain-Tunable Shift Current and Magneto-Optical Kerr Effect in Multiferroic Altermagnet Fe2Mo3O8

Shengqiao Wang, Bo Zhao, Harish K. Singh +6

The paper uses first‑principles calculations to study how ferroelectric polarization, strain, and altermagnetic order in the multiferroic Fe2Mo3O8 affect shift current and magneto‑…

#altermagnetism#multiferroics#shift current#magneto-optical Kerr effect
quant-ph2022

Quantum refrigerator driven by nonclassical light

Hui-Jing Cao, Fu Li, Sheng-Wen Li

We study a three-level quantum refrigerator which is driven by a generic light state, even a nonclassical one. With the help of P function expansion of the driving light, we obtain…

cond-mat.mtrl-sci2025

Ferroelectricity-driven altermagnetism in two-dimensional van der Waals multiferroics

Bo Zhao, Fu Li, Wei Ren +2

Altermagnets (AMs) are a recently identified class of unconventional collinear compensated antiferromagnets that exhibit momentum-dependent spin splitting despite having zero net m…

cs.CV2022

Predict, Prevent, and Evaluate: Disentangled Text-Driven Image Manipulation Empowered by Pre-Trained Vision-Language Model

Zipeng Xu, Tianwei Lin, Hao Tang +6

To achieve disentangled image manipulation, previous works depend heavily on manual annotation. Meanwhile, the available manipulations are limited to a pre-defined set the models w…

cs.CV2021

AdaAttN: Revisit Attention Mechanism in Arbitrary Neural Style Transfer

Songhua Liu, Tianwei Lin, Dongliang He +6

Fast arbitrary neural style transfer has attracted widespread attention from academic, industrial and art communities due to its flexibility in enabling various applications. Exist…

cs.HC2024

Enhancing Cross-Dataset EEG Emotion Recognition: A Novel Approach with Emotional EEG Style Transfer Network

Yijin Zhou, Fu Li, Yang Li +3

Recognizing the pivotal role of EEG emotion recognition in the development of affective Brain-Computer Interfaces (aBCIs), considerable research efforts have been dedicated to this…

quant-ph2017

Investigating the `Past of a Particle' without disturbing it

Faheel Hashmi, Fu Li, Shiyao Zhu

In a recent article [Chin. Phys. Lett. 34, 020301 (2017)], Ben-Israel et al. have claimed that the experiment proposed in [Chin. Phys. Lett. 32, 050303 (2015)] to determine the pas…

cs.CV2025

Spatio-Temporal Progressive Attention Model for EEG Classification in Rapid Serial Visual Presentation Task

Yang Li, Wei Liu, Tianzhi Feng +6

As a type of multi-dimensional sequential data, the spatial and temporal dependencies of electroencephalogram (EEG) signals should be further investigated. Thus, in this paper, we…

cs.GT2023

The Obnoxious Facility Location Game with Dichotomous Preferences

Fu Li, C. Gregory Plaxton, Vaibhav B. Sinha

We consider a facility location game in which agents reside at known locations on a path, and heterogeneous facilities are to be constructed on the path. Each agent is adve…

eess.IV2020

NTIRE 2020 Challenge on Video Quality Mapping: Methods and Results

Dario Fuoli, Zhiwu Huang, Martin Danelljan +18

This paper reviews the NTIRE 2020 challenge on video quality mapping (VQM), which addresses the issues of quality mapping from source video domain to target video domain. The chall…

cs.CV2025

DeltaEdit: Exploring Text-free Training for Text-Driven Image Manipulation

Yueming Lyu, Tianwei Lin, Fu Li +3

Text-driven image manipulation remains challenging in training or inference flexibility. Conditional generative models depend heavily on expensive annotated training data. Meanwhil…

cs.LG2023

Revisiting Neural Retrieval on Accelerators

Jiaqi Zhai, Zhaojie Gong, Yueming Wang +4

Retrieval finds a small number of relevant candidates from a large corpus for information retrieval and recommendation applications. A key component of retrieval is to model (user,…

cs.CV2021

Paint Transformer: Feed Forward Neural Painting with Stroke Prediction

Songhua Liu, Tianwei Lin, Dongliang He +5

Neural painting refers to the procedure of producing a series of strokes for a given image and non-photo-realistically recreating it using neural networks. While reinforcement lear…

eess.SP2025

Adaptive Progressive Attention Graph Neural Network for EEG Emotion Recognition

Tianzhi Feng, Chennan Wu, Yi Niu +5

In recent years, numerous neuroscientific studies demonstrate that specific areas of the brain are connected to human emotional responses, with these regions exhibiting variability…

cs.CV2021

MVFNet: Multi-View Fusion Network for Efficient Video Recognition

Wenhao Wu, Dongliang He, Tianwei Lin +3

Conventionally, spatiotemporal modeling network and its complexity are the two most concentrated research topics in video action recognition. Existing state-of-the-art methods have…

cs.CV2022

OSOP: A Multi-Stage One Shot Object Pose Estimation Framework

Ivan Shugurov, Fu Li, Benjamin Busam +1

We present a novel one-shot method for object detection and 6 DoF pose estimation, that does not require training on target objects. At test time, it takes as input a target image…

cs.CV2025

Goku: Flow Based Video Generative Foundation Models

Shoufa Chen, Chongjian Ge, Yuqi Zhang +19

This paper introduces Goku, a state-of-the-art family of joint image-and-video generation models leveraging rectified flow Transformers to achieve industry-leading performance. We…

cs.CV2022

Boosting Video-Text Retrieval with Explicit High-Level Semantics

Haoran Wang, Di Xu, Dongliang He +4

Video-text retrieval (VTR) is an attractive yet challenging task for multi-modal understanding, which aims to search for relevant video (text) given a query (video). Existing metho…

cs.CV2019

Deep Concept-wise Temporal Convolutional Networks for Action Localization

Xin Li, Tianwei Lin, Xiao Liu +7

Existing action localization approaches adopt shallow temporal convolutional networks (\ie, TCN) on 1D feature map extracted from video frames. In this paper, we empirically find t…

cs.SI2015

Combining Traditional Marketing and Viral Marketing with Amphibious Influence Maximization

Wei Chen, Fu Li, Tian Lin +1

In this paper, we propose the amphibious influence maximization (AIM) model that combines traditional marketing via content providers and viral marketing to consumers in social net…

cs.CV2018

Exploiting Spatial-Temporal Modelling and Multi-Modal Fusion for Human Action Recognition

Dongliang He, Fu Li, Qijie Zhao +3

In this report, our approach to tackling the task of ActivityNet 2018 Kinetics-600 challenge is described in detail. Though spatial-temporal modelling methods, which adopt either s…

cs.CL2024

A Study on Training and Developing Large Language Models for Behavior Tree Generation

Fu Li, Xueying Wang, Bin Li +3

This paper presents an innovative exploration of the application potential of large language models (LLM) in addressing the challenging task of automatically generating behavior tr…

cs.IR2025

Diffusion-driven SpatioTemporal Graph KANsformer for Medical Examination Recommendation

Jianan Li, Yangtao Zhou, Zhifu Zhao +5

Recommendation systems in AI-based medical diagnostics and treatment constitute a critical component of AI in healthcare. Although some studies have explored this area and made not…

cond-mat.mtrl-sci2018

Colloidal Tin Sulfide Nanosheets: Formation Mechanism, Ligand-mediated Shape Tuning and Photo-detection

Fu Li, Mohammad Mehdi Ramin Moayed, Frauke Gerdes +4

Colloidal materials of tin(II) sulfide (SnS), as a layered semiconductor with a narrow band gap, are emerging as a potential alternative to the more toxic metal chalcogenides (PbS,…

cs.SD2026

FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts

You Li, Dewei Zhou, Fan Ma +3

Recent Video-to-Audio (V2A) methods have achieved remarkable progress, enabling the synthesis of realistic, high-quality audio. However, they struggle with fine-grained temporal co…

astro-ph.IM2020

Mapping Diffuse Emission in Lyman UV band

Li Ji, Zheng Lou, Jinlong Zhang +38

The CAFE (Census of warm-hot intergalactic medium, Accretion, and Feedback Explorer) and LyRIC (Lyman UV Radiation from Interstellar medium and Circum-galactic medium) have been pr…

cs.AI2026

From Experience to Skill: Multi-Agent Generative Engine Optimization via Reusable Strategy Learning

Beining Wu, Fuyou Mao, Jiong Lin +7

Generative engines (GEs) are reshaping information access by replacing ranked links with citation-grounded answers, yet current Generative Engine Optimization (GEO) methods optimiz…

cs.CV2021

DOLG: Single-Stage Image Retrieval with Deep Orthogonal Fusion of Local and Global Features

Min Yang, Dongliang He, Miao Fan +5

Image Retrieval is a fundamental task of obtaining images similar to the query one from a database. A common image retrieval practice is to firstly retrieve candidate images via si…

cond-mat.mtrl-sci2020

Anisotropic circular photogalvanic effect in colloidal tin sulfide nanosheets

Mohammad Mehdi Ramin Moayed, Fu Li, Philip Beck +2

Tin sulfide promises very interesting properties such as a high optical absorption coefficient and a small band gap, while being less toxic compared to other metal chalcogenides. H…

cs.CV2021

Progressive Graph Convolution Network for EEG Emotion Recognition

Yijin Zhou, Fu Li, Yang Li +6

Studies in the area of neuroscience have revealed the relationship between emotional patterns and brain functional regions, demonstrating that dynamic relationships between differe…

cs.LG2018

Combinatorial Multi-Armed Bandit with General Reward Functions

Wei Chen, Wei Hu, Fu Li +3

In this paper, we study the stochastic combinatorial multi-armed bandit (CMAB) framework that allows a general nonlinear reward function, whose expected value may not depend only o…

cond-mat.mtrl-sci2023

Bayesian optimization with active learning of Ta-Nb-Hf-Zr-Ti system for spin transport properties

Ruiwen Xie, Yixuan Zhang, Fu Li +2

Designing materials with enhanced spin charge conversion, i.e., with high spin Hall conductivity (SHC) and low longitudinal electric conductivity (hence large spin Hall angle (SHA)…

cs.CV2021

Beyond Self-Supervision: A Simple Yet Effective Network Distillation Alternative to Improve Backbones

Cheng Cui, Ruoyu Guo, Yuning Du +10

Recently, research efforts have been concentrated on revealing how pre-trained model makes a difference in neural network performance. Self-supervision and semi-supervised learning…

physics.optics2020

Squeezed Light Induced Two-photon Absorption Fluorescence of Fluorescein Biomarkers

Tian Li, Fu Li, Charles Altuzarra +2

Two-photon absorption (TPA) fluorescence of biomarkers has been decisive in advancing the fields of biosensing and deep-tissue in vivo imaging of live specimens. However, due to th…

cs.CV2017

Temporal Modeling Approaches for Large-scale Youtube-8M Video Understanding

Fu Li, Chuang Gan, Xiao Liu +6

This paper describes our solution for the video recognition task of the Google Cloud and YouTube-8M Video Understanding Challenge that ranked the 3rd place. Because the challenge p…

eess.SP2023

EEG-based Emotion Style Transfer Network for Cross-dataset Emotion Recognition

Yijin Zhou, Fu Li, Yang Li +5

As the key to realizing aBCIs, EEG emotion recognition has been widely studied by many researchers. Previous methods have performed well for intra-subject EEG emotion recognition.…

eess.IV2022

NTIRE 2021 Challenge on Quality Enhancement of Compressed Video: Methods and Results

Ren Yang, Radu Timofte, Jing Liu +69

This paper reviews the first NTIRE challenge on quality enhancement of compressed video, with a focus on the proposed methods and results. In this challenge, the new Large-scale Di…

eess.SP2022

GMSS: Graph-Based Multi-Task Self-Supervised Learning for EEG Emotion Recognition

Yang Li, Ji Chen, Fu Li +7

Previous electroencephalogram (EEG) emotion recognition relies on single-task learning, which may lead to overfitting and learned emotion features lacking generalization. In this p…

cs.CV2023

Retinex-guided Channel-grouping based Patch Swap for Arbitrary Style Transfer

Chang Liu, Yi Niu, Mingming Ma +2

The basic principle of the patch-matching based style transfer is to substitute the patches of the content image feature maps by the closest patches from the style image feature ma…

physics.med-ph2023

A forward model incorporating elevation-focused transducer properties for 3D full-waveform inversion in ultrasound computed tomography

Fu Li, Umberto Villa, Nebojsa Duric +1

Ultrasound computed tomography (USCT) is an emerging medical imaging modality that holds great promise for improving human health. Full-waveform inversion (FWI)-based image reconst…

astro-ph.SR2023

The Solar Upper Transition Region Imager (SUTRI) onboard the SATech-01 satellite

Xianyong Bai, Hui Tian, Yuanyong Deng +56

The Solar Upper Transition Region Imager (SUTRI) onboard the Space Advanced Technology demonstration satellite (SATech-01), which was launched to a sun-synchronous orbit at a heigh…

cs.CV2021

Drafting and Revision: Laplacian Pyramid Network for Fast High-Quality Artistic Style Transfer

Tianwei Lin, Zhuoqi Ma, Fu Li +6

Artistic style transfer aims at migrating the style from an example image to a content image. Currently, optimization-based methods have achieved great stylization quality, but exp…

cs.CV2018

StNet: Local and Global Spatial-Temporal Modeling for Action Recognition

Dongliang He, Zhichao Zhou, Chuang Gan +5

Despite the success of deep learning for static image understanding, it remains unclear what are the most effective network architectures for the spatial-temporal modeling in video…

cs.CV2022

Neural Color Operators for Sequential Image Retouching

Yili Wang, Xin Li, Kun Xu +4

We propose a novel image retouching method by modeling the retouching process as performing a sequence of newly introduced trainable neural color operators. The neural color operat…

cs.CC2014

Generating Matrix Identities and Proof Complexity

Fu Li, Iddo Tzameret

Motivated by the fundamental lower bounds questions in proof complexity, we initiate the study of matrix identities as hard instances for strong proof systems. A matrix identity of…

cond-mat.mtrl-sci2020

Single Crystalline Colloidal Quasi-Two-Dimensional Tin Telluride

Fu Li, Jiecai Fu, Abderrezak Torche +5

Tin telluride is a narrow gap semiconductor with promising properties for IR optical applications and topological insulators. We report a convenient colloidal synthesis of quasi-tw…

eess.IV2024

Investigating the Use of Traveltime and Reflection Tomography for Deep Learning-Based Sound-Speed Estimation in Ultrasound Computed Tomography

Gangwon Jeong, Fu Li, Trevor M. Mitcham +3

Ultrasound computed tomography (USCT) quantifies acoustic tissue properties such as the speed-of-sound (SOS). Although full-waveform inversion (FWI) is an effective method for accu…

cs.CV2023

Master: Meta Style Transformer for Controllable Zero-Shot and Few-Shot Artistic Style Transfer

Hao Tang, Songhua Liu, Tianwei Lin +4

Transformer-based models achieve favorable performance in artistic style transfer recently thanks to its global receptive field and powerful multi-head/layer attention operations.…

physics.optics2021

Experimental study of decoherence of the two-mode squeezed vacuum state via second harmonic generation

Fu Li, Tian Li, Girish S. Agarwal

Decoherence remains one of the most serious challenges to the implementation of quantum technology. It appears as a result of the transformation over time of a quantum superpositio…

cs.CV2019

Read, Watch, and Move: Reinforcement Learning for Temporally Grounding Natural Language Descriptions in Videos

Dongliang He, Xiang Zhao, Jizhou Huang +3

The task of video grounding, which temporally localizes a natural language description in a video, plays an important role in understanding videos. Existing studies have adopted st…

cond-mat.mtrl-sci2025

Shift Current Anomalous Photovoltaics in a Double Perovskite Ferroelectric

Linjie Wei, Fu Li, Yi Liu +3

Ferroelectric anomalous photovoltaic (APV) effect, as a fascinating physical conceptual phenomenon, holds significant potentials for new optoelectronic device applications. However…

cs.SD2025

DreamFoley: Scalable VLMs for High-Fidelity Video-to-Audio Generation

Fu Li, Weichao Zhao, You Li +2

Recent advances in video generation have achieved remarkable improvements in visual content fidelity. However, the absence of synchronized audio severely undermines immersive exper…

quant-ph2014

An ideal experiment to determine the 'past of a particle' in the nested Mach-Zehnder Interferometer

Fu Li, F. A. Hashmi, Jun-Xiang Zhang +1

An ideal experiment is designed to determine the past of a particle in the nested Mach-Zehnder interferometer (MZI) by using standard quantum mechanics with quantum non-demolition…

cs.CV2023

Vision Transformer with Attention Map Hallucination and FFN Compaction

Haiyang Xu, Zhichao Zhou, Dongliang He +2

Vision Transformer(ViT) is now dominating many vision tasks. The drawback of quadratic complexity of its token-wise multi-head self-attention (MHSA), is extensively addressed via e…

cs.CV2022

AdaCM: Adaptive ColorMLP for Real-Time Universal Photo-realistic Style Transfer

Tianwei Lin, Honglin Lin, Fu Li +5

Photo-realistic style transfer aims at migrating the artistic style from an exemplar style image to a content image, producing a result image without spatial distortions or unreali…

cs.CV2025

Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning

Qinghao Ye, Xianhan Zeng, Fu Li +2

Image captioning has long been a pivotal task in visual understanding, with recent advancements in vision-language models (VLMs) significantly enhancing the ability to generate det…

eess.IV2025

Learned Correction Methods for Ultrasound Computed Tomography Imaging Using Simplified Physics Models

Luke Lozenski, Hanchen Wang, Fu Li +4

Ultrasound computed tomography (USCT) is an emerging modality for breast imaging. Image reconstruction methods that incorporate accurate wave physics produce high resolution quanti…

cs.CV2025

CoCoDiff: Diversifying Skeleton Action Features via Coarse-Fine Text-Co-Guided Latent Diffusion

Zhifu Zhao, Hanyang Hua, Jianan Li +4

In action recognition tasks, feature diversity is essential for enhancing model generalization and performance. Existing methods typically promote feature diversity by expanding th…

quant-ph2014

The effect of dissipation in direct communication scheme

Fu Li, Jun-Xiang Zhang, Shi-Yao Zhu

The effect of the dissipation and finite number of beam splitters are discussed. A method using balanced dissipation to improve the communication for finite beam splitters, which g…

cs.HC2025

Integrating Language-Image Prior into EEG Decoding for Cross-Task Zero-Calibration RSVP-BCI

Xujin Li, Wei Wei, Shuang Qiu +3

Rapid Serial Visual Presentation (RSVP)-based Brain-Computer Interface (BCI) is an effective technology used for information detection by detecting Event-Related Potentials (ERPs).…

cs.CV2022

Adversarial Dual-Student with Differentiable Spatial Warping for Semi-Supervised Semantic Segmentation

Cong Cao, Tianwei Lin, Dongliang He +4

A common challenge posed to robust semantic segmentation is the expensive data annotation cost. Existing semi-supervised solutions show great potential for solving this problem. Th…

physics.soc-ph2020

Cost-effectiveness Analysis of Antiepidemic Policies and Global Situation Assessment of COVID-19

Liyan Xu, Hongmou Zhang, Yuqiao Deng +17

With a two-layer contact-dispersion model and data in China, we analyze the cost-effectiveness of three types of antiepidemic measures for COVID-19: regular epidemiological control…

eess.IV2023

Learned Full Waveform Inversion Incorporating Task Information for Ultrasound Computed Tomography

Luke Lozenski, Hanchen Wang, Fu Li +4

Ultrasound computed tomography (USCT) is an emerging imaging modality that holds great promise for breast imaging. Full-waveform inversion (FWI)-based image reconstruction methods…

quant-ph2019

Photon statistics of quantum light on scattering from rotating ground glass

Sheng-Wen Li, Fu Li, Tao Peng +1

When a laser beam passes through a rotating ground glass (RGG), the scattered light exhibits thermal statistics. This is extensively used in speckle imaging. This scattering proces…

cond-mat.mtrl-sci2026

Medium-Throughput Evaluation of Quantum Geometry-Driven Topological Transports in Altermagnets

Fu Li, Bo Zhao, Vikrant Chaudhary +4

Altermagnets provide a promising platform for a wide spectrum of applications integrating advantages of conventional ferromagnets and antiferromagnets. In this work, we implement a…

cs.CV2021

Learning Semantic Person Image Generation by Region-Adaptive Normalization

Zhengyao Lv, Xiaoming Li, Xin Li +4

Human pose transfer has received great attention due to its wide applications, yet is still a challenging task that is not well solved. Recent works have achieved great success to…

cs.CV2019

TruNet: Short Videos Generation from Long Videos via Story-Preserving Truncation

Fan Yang, Xiao Liu, Dongliang He +5

In this work, we introduce a new problem, named as {\em story-preserving long video truncation}, that requires an algorithm to automatically truncate a long-duration video into mul…