papers

Publications (84)

cond-mat.mtrl-sci2020

Achromatic metasurfaces with inversely customized dispersion for ultra-broadband acoustic beam engineering

Hao-Wen Dong, Chen Shen, Sheng-Dong Zhao +7

Metasurfaces, the ultrathin media with extraordinary wavefront modulation ability, have shown versatile potential in manipulating waves. However, existing acoustic metasurfaces are…

cs.CV2021

Deep Learning for Visual Tracking: A Comprehensive Survey

Seyed Mojtaba Marvasti-Zadeh, Li Cheng, Hossein Ghanei-Yakhdan +1

Visual target tracking is one of the most sought-after yet challenging research topics in computer vision. Given the ill-posed nature of the problem and its popularity in a broad r…

cs.CV2023

Snipper: A Spatiotemporal Transformer for Simultaneous Multi-Person 3D Pose Estimation Tracking and Forecasting on a Video Snippet

Shihao Zou, Yuanlu Xu, Chao Li +3

Multi-person pose understanding from RGB videos involves three complex tasks: pose estimation, tracking and motion forecasting. Intuitively, accurate multi-person pose estimation f…

cond-mat.mtrl-sci2025

Giant tunneling magnetoresistance based on spin-valley-mismatched ferromagnetic metals

Kan Yan, Li Cheng, Yizhi Hu +3

Half metals, which are amenable to perfect spin filtering, can be utilized for high-magnetoresistive devices. However, available half metals are very limited. Here, we demonstrate…

cond-mat.mes-hall2020

Renormalization of the Mott gap by lattice entropy: The case of 1T-TaS2

Li Cheng, Shunhong Zhang, Shuang Qiao +3

In many transition-metal oxides and dichalcogenides, the electronic and lattice degrees of freedom are strongly coupled, giving rise to remarkable phenomena, such as metal-insulato…

cs.CV2023

Segment Anything Is Not Always Perfect: An Investigation of SAM on Different Real-world Applications

Wei Ji, Jingjing Li, Qi Bi +3

Recently, Meta AI Research approaches a general, promptable Segment Anything Model (SAM) pre-trained on an unprecedentedly large segmentation dataset (SA-1B). Without a doubt, the…

cs.CV2020

Action2Motion: Conditioned Generation of 3D Human Motions

Chuan Guo, Xinxin Zuo, Sen Wang +5

Action recognition is a relatively established task, where givenan input sequence of human motion, the goal is to predict its ac-tion category. This paper, on the other hand, consi…

physics.app-ph2018

Systematic design and realization of double-negative acoustic metamaterials by topology optimization

Hao-Wen Dong, Sheng-Dong Zhao, Peijun Wei +3

Double-negative acoustic metamaterials (AMMs) offer the promising ability of superlensing for applications in ultrasonography, biomedical sensing and nondestructive evaluation. Her…

cs.CV2014

Transduction on Directed Graphs via Absorbing Random Walks

Jaydeep De, Xiaowei Zhang, Li Cheng

In this paper we consider the problem of graph-based transductive classification, and we are particularly interested in the directed graph scenario which is a natural form for many…

cs.RO2024

Autonomous Ground Navigation in Highly Constrained Spaces: Lessons learned from The 3rd BARN Challenge at ICRA 2024

Xuesu Xiao, Zifan Xu, Aniket Datar +16

The 3rd BARN (Benchmark Autonomous Robot Navigation) Challenge took place at the 2024 IEEE International Conference on Robotics and Automation (ICRA 2024) in Yokohama, Japan and co…

cs.CV2021

EventHPE: Event-based 3D Human Pose and Shape Estimation

Shihao Zou, Chuan Guo, Xinxin Zuo +6

Event camera is an emerging imaging sensor for capturing dynamics of moving objects as events, which motivates our work in estimating 3D human pose and shape from the event signals…

cond-mat.mtrl-sci2021

Understanding the flat band in 1T-TaS2 using a rotated basis

Li Cheng, Xuanyu Long, Xiaobin Chen +2

Electronic flat bands serve as a unique platform to achieve strongly-correlated phases. The emergence of a flat band around the Fermi level in 1T-TaS in accompany with the deve…

cs.CV2024

Generative Human Motion Stylization in Latent Space

Chuan Guo, Yuxuan Mu, Xinxin Zuo +4

Human motion stylization aims to revise the style of an input motion while keeping its content unaltered. Unlike existing works that operate directly in pose space, we leverage the…

cs.CV2026

Progressive Pose-Guided 4D Animal Reconstruction from Monocular Video

Siyuan Li, Weiying Chen, Yilin Wang +3

Reconstructing 4D animals from monocular videos is challenging due to large inter-species variation, complex articulations, and the lack of reliable templates. Existing approaches…

hep-ph2017

Validity of two Higgs doublet models with a scalar color octet up to a high energy scale

Li Cheng, German Valencia

We have recently studied theoretical constraints on the parameters of a 2HDM augmented with a color-octet scalar. In this paper we consider the consequences of requiring the model…

cs.CV2025

Highly Efficient 3D Human Pose Tracking from Events with Spiking Spatiotemporal Transformer

Shihao Zou, Yuxuan Mu, Wei Ji +5

Event camera, as an asynchronous vision sensor capturing scene dynamics, presents new opportunities for highly efficient 3D human pose tracking. Existing approaches typically adopt…

cs.CV2024

RACon: Retrieval-Augmented Simulated Character Locomotion Control

Yuxuan Mu, Shihao Zou, Kangning Yin +4

In computer animation, driving a simulated character with lifelike motion is challenging. Current generative models, though able to generalize to diverse motions, often pose challe…

hep-ph2019

New theoretical constraints on scalar color octet models

Li Cheng, Otto Eberhardt, Christopher W. Murphy

We study theoretical constraints on a model whose scalar sector contains one color octet and one or two color singlet doublets. Using the unitarity of the theory, we cons…

hep-ph2014

Top-quark forward-backward asymmetry from a color-octet t-channel resonance

Li Cheng, Alper Hayreter, German Valencia

We consider new physics contributions to the top-quark forward-backward asymmetry from a neutral or charged color-octet vector exchanged in the -channel. We stud…

cs.CV2025

VPHO: Joint Visual-Physical Cue Learning and Aggregation for Hand-Object Pose Estimation

Jun Zhou, Chi Xu, Kaifeng Tang +3

Estimating the 3D poses of hands and objects from a single RGB image is a fundamental yet challenging problem, with broad applications in augmented reality and human-computer inter…

cs.CV2017

Synthesizing Filamentary Structured Images with GANs

He Zhao, Huiqi Li, Li Cheng

This paper aims at synthesizing filamentary structured images such as retinal fundus images and neuronal images, as follows: Given a ground-truth, to generate multiple realistic lo…

physics.ins-det2014

Monte Carlo Simulation of RPC-based PET with GEANT4

Zhou Weizheng, Shao Ming, Li Cheng +3

The Resistive Plate Chambers (RPC) are low-cost charged-particle detectors with good timing resolution and potentially good spatial resolution. Using RPC as gamma detector provides…

cs.CV2016

Hand Action Detection from Ego-centric Depth Sequences with Error-correcting Hough Transform

Chi Xu, Lakshmi Narasimhan Govindarajan, Li Cheng

Detecting hand actions from ego-centric depth sequences is a practically challenging problem, owing mostly to the complex and dexterous nature of hand articulations as well as non-…

cs.CV2024

Two-stage Synthetic Supervising and Multi-view Consistency Self-supervising based Animal 3D Reconstruction by Single Image

Zijian Kuang, Lihang Ying, Shi Jin +1

Pixel-aligned Implicit Function (PIFu) effectively captures subtle variations in body shape within a low-dimensional space through extensive training with human 3D scans, its appli…

cs.CV2021

Self-supervised 3D Human Mesh Recovery from Noisy Point Clouds

Xinxin Zuo, Sen Wang, Qiang Sun +2

This paper presents a novel self-supervised approach to reconstruct human shape and pose from noisy point cloud data. Relying on large amount of dataset with ground-truth annotatio…

cs.CY2012

Big-Five Personality Prediction Based on User Behaviors at Social Network Sites

Shuotian Bai, Tingshao Zhu, Li Cheng

Many customer services are already available at Social Network Sites (SNSs), including user recommendation and media interaction, to name a few. There are strong desires to provide…

cs.CV2021

Investigating Pose Representations and Motion Contexts Modeling for 3D Motion Prediction

Zhenguang Liu, Shuang Wu, Shuyuan Jin +4

Predicting human motion from historical pose sequence is crucial for a machine to succeed in intelligent interactions with humans. One aspect that has been obviated so far, is the…

math.NA2025

Modeling of thin plate flexural vibrations by Partition of Unity Finite Element Method

Tong Zhou, Jean-Daniel Chazot, Emmanuel Perrey-Debain +1

This paper presents a conforming thin plate bending element based on the Partition of Unity Finite Element Method (PUFEM), for the simulation of steady-state forced vibration. The…

cs.CV2022

Object Wake-up: 3D Object Rigging from a Single Image

Ji Yang, Xinxin Zuo, Sen Wang +5

Given a single image of a general object such as a chair, could we also restore its articulated 3D shape similar to human modeling, so as to animate its plausible articulations and…

cond-mat.supr-con2007

Electronic structure of the electron-doped cuprate superconductors

Li Cheng, Huaiming Guo, Shiping Feng

Within the framework of the kinetic energy driven d-wave superconductivity, the electronic structure of the electron doped cuprate superconductors is studied. It is shown that alth…

cs.CV2022

Human Pose and Shape Estimation from Single Polarization Images

Shihao Zou, Xinxin Zuo, Sen Wang +3

This paper focuses on a new problem of estimating human pose and shape from single polarization images. Polarization camera is known to be able to capture the polarization of refle…

cs.CV2017

Too Far to See? Not Really! --- Pedestrian Detection with Scale-aware Localization Policy

Xiaowei Zhang, Li Cheng, Bo Li +1

A major bottleneck of pedestrian detection lies on the sharp performance deterioration in the presence of small-size pedestrians that are relatively far from the camera. Motivated…

hep-ph2016

Two Higgs doublet models augmented by a scalar color octet

Li Cheng, German Valencia

The LHC is now studying in detail the couplings of the Higgs boson in order to determine if there is new physics. Many recent studies have examined the available fits to Higgs coup…

cs.CV2020

COMET: Context-Aware IoU-Guided Network for Small Object Tracking

Seyed Mojtaba Marvasti-Zadeh, Javad Khaghani, Hossein Ghanei-Yakhdan +2

We consider the problem of tracking an unknown small target from aerial videos of medium to high altitudes. This is a challenging problem, which is even more pronounced in unavoida…

cs.CV2025

MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer

Yilin Wang, Chuan Guo, Yuxuan Mu +5

Generative masked transformers have demonstrated remarkable success across various content generation tasks, primarily due to their ability to effectively model large-scale dataset…

math.FA2023

Directional Differentiability of the Generalized Metric Projection in Hilbert spaces and Hilbertian Bochner spaces

Jinlu Li, Li Cheng, Lishan Liu +1

Let be a real Hilbert space and a nonempty closed and convex subset of . Let denote the (standard) metric projection operator. In this paper, we st…

cs.CV2020

Polarization Human Shape and Pose Dataset

Shihao Zou, Xinxin Zuo, Yiming Qian +5

Polarization images are known to be able to capture polarized reflected lights that preserve rich geometric cues of an object, which has motivated its recent applications in recons…

cs.CV2023

MoMask: Generative Masked Modeling of 3D Human Motions

Chuan Guo, Yuxuan Mu, Muhammad Gohar Javed +2

We introduce MoMask, a novel masked modeling framework for text-driven 3D human motion generation. In MoMask, a hierarchical quantization scheme is employed to represent human moti…

cs.CV2016

Learning to Search on Manifolds for 3D Pose Estimation of Articulated Objects

Yu Zhang, Chi Xu, Li Cheng

This paper focuses on the challenging problem of 3D pose estimation of a diverse spectrum of articulated objects from single depth images. A novel structured prediction approach is…

cond-mat.other2022

Band degeneration and evolution in nonlinear triatomic chain superlattices

Chen Gong, Xin Fang, Li Cheng

Nonlinear superlattices exhibit unique features allowing for wave manipulations. Despite the increasing attention received, the underlying physical mechanisms and the evolution pro…

cs.CV2020

WaveletKernelNet: An Interpretable Deep Neural Network for Industrial Intelligent Diagnosis

Tianfu Li, Zhibin Zhao, Chuang Sun +4

Convolutional neural network (CNN), with ability of feature learning and nonlinear mapping, has demonstrated its effectiveness in prognostics and health management (PHM). However,…

stat.ML2017

Multivariate Regression with Gross Errors on Manifold-valued Data

Xiaowei Zhang, Xudong Shi, Yu Sun +1

We consider the topic of multivariate regression on manifold-valued output, that is, for a multivariate observation, its output response lies on a manifold. Moreover, we propose a…

eess.IV2024

QUBIQ: Uncertainty Quantification for Biomedical Image Segmentation Challenge

Hongwei Bran Li, Fernando Navarro, Ivan Ezhov +77

Uncertainty in medical image segmentation tasks, especially inter-rater variability, arising from differences in interpretations and annotations by various experts, presents a sign…

eess.IV2021

The RETA Benchmark for Retinal Vascular Tree Analysis

Xingzheng Lyu, Li Cheng, Sanyuan Zhang

Topological and geometrical analysis of retinal blood vessel is a cost-effective way for early detection of many common diseases. Meanwhile, automated vessel segmentation and vascu…

cs.CV2026

SAM3-I: Segment Anything with Instructions

Jingjing Li, Yue Feng, Yuchen Guo +10

Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instances associated with a given conce…

cs.CV2024

RegionGrasp: A Novel Task for Contact Region Controllable Hand Grasp Generation

Yilin Wang, Chuan Guo, Li Cheng +1

Can machine automatically generate multiple distinct and natural hand grasps, given specific contact region of an object in 3D? This motivates us to consider a novel task of \texti…

cond-mat.supr-con2008

Doping and energy evolution of spin dynamics in the electron-doped cuprate superconductor PrLaCeCuO

Li Cheng, Shiping Feng

The doping and energy evolution of the magnetic excitations of the electron-doped cuprate superconductor PrLaCeCuO in the superconducting state is studie…

cs.LG2020

Stabilizing Training of Generative Adversarial Nets via Langevin Stein Variational Gradient Descent

Dong Wang, Xiaoqian Qin, Fengyi Song +1

Generative adversarial networks (GANs), famous for the capability of learning complex underlying data distribution, are however known to be tricky in the training process, which wo…

eess.IV2020

Reconstruct high-resolution multi-focal plane images from a single 2D wide field image

Jiabo Ma, Sibo Liu, Shenghua Cheng +3

High-resolution 3D medical images are important for analysis and diagnosis, but axial scanning to acquire them is very time-consuming. In this paper, we propose a fast end-to-end m…

physics.app-ph2021

Bidirectional elastic diode with frequency-preserved nonreciprocity

Xin Fang, Jihong Wen, Li Cheng +1

The study of nonreciprocal wave propagation is of great interests for both fundamental research and engineering applications. Here we demonstrate theoretically and experimentally a…

cond-mat.mtrl-sci2022

Probing complex stacking in a layered material via electron-nuclear quadrupolar coupling

Li Cheng, Linpeng Nie, Xuanyu Long +7

For layered materials, the interlayer stacking is a critical degree of freedom tuning electronic properties, while its microscopic characterization faces great challenges. The tran…

cs.CV2021

Action2video: Generating Videos of Human 3D Actions

Chuan Guo, Xinxin Zuo, Sen Wang +4

We aim to tackle the interesting yet challenging problem of generating videos of diverse and natural human motions from prescribed action categories. The key issue lies in the abil…

cs.SD2022

Dual Learning Music Composition and Dance Choreography

Shuang Wu, Zhenguang Li, Shijian Lu +1

Music and dance have always co-existed as pillars of human activities, contributing immensely to the cultural, social, and entertainment functions in virtually all societies. Notwi…

cs.CV2015

Mouse Pose Estimation From Depth Images

Ashwin Nanjappa, Li Cheng, Wei Gao +3

We focus on the challenging problem of efficient mouse 3D pose estimation based on static images, and especially single depth images. We introduce an approach to discriminatively t…

cs.CR2026

MirageNet:A Secure, Efficient, and Scalable On-Device Model Protection in Heterogeneous TEE and GPU System

Huadi Zheng, Li Cheng, Yan Ding

As edge devices gain stronger computing power, deploying high-performance DNN models on untrusted hardware has become a practical approach to cut inference latency and protect user…

physics.class-ph2022

A Virtual Acoustic Black Hole on a Cantilever Beam

Samuel Quaegebeur, Ghislain Raze, Li Cheng +1

An acoustic black hole (ABH) consists of a tapered structure whose thickness follows a power-law profile. When attached to a host structure, an ABH localizes and traps the vibratio…

cs.CV2024

TexGen: Text-Guided 3D Texture Generation with Multi-view Sampling and Resampling

Dong Huo, Zixin Guo, Xinxin Zuo +6

Given a 3D mesh, we aim to synthesize 3D textures that correspond to arbitrary textual descriptions. Current methods for generating and assembling textures from sampled views often…

cs.CV2022

TM2T: Stochastic and Tokenized Modeling for the Reciprocal Generation of 3D Human Motions and Texts

Chuan Guo, Xinxin Zuo, Sen Wang +1

Inspired by the strong ties between vision and language, the two intimate human sensing and communication modalities, our paper aims to explore the generation of 3D human full-body…

stat.ML2017

Multivariate Regression with Grossly Corrupted Observations: A Robust Approach and its Applications

Xiaowei Zhang, Chi Xu, Yu Zhang +2

This paper studies the problem of multivariate linear regression where a portion of the observations is grossly corrupted or is missing, and the magnitudes and locations of such oc…

cs.CV2020

SparseFusion: Dynamic Human Avatar Modeling from Sparse RGBD Images

Xinxin Zuo, Sen Wang, Jiangbin Zheng +4

In this paper, we propose a novel approach to reconstruct 3D human body shapes based on a sparse set of RGBD frames using a single RGBD camera. We specifically focus on the realist…

cond-mat.supr-con2007

Electronic structure of kinetic energy driven cuprate superconductors

Shiping Feng, Huaiming Guo, Yu Lan +1

In this paper, we review the low energy electronic structure of the kinetic energy driven d-wave cuprate superconductors. We give a general description of the charge-spin separatio…

cond-mat.soft2021

The interaction between phosphorene oxide and the villin headpiece

Wei Zhang, Yuanyuan Gou, Li Cheng +2

Phosphorene, a novel member of the two-dimensional nanomaterial family, has been demonstrated a great potential in biomedical applications, such as photothermal therapy, drug deliv…

cs.LG2020

Outlier Detection Ensemble with Embedded Feature Selection

Li Cheng, Yijie Wang, Xinwang Liu +1

Feature selection places an important role in improving the performance of outlier detection, especially for noisy data. Existing methods usually perform feature selection and outl…

cs.SD2022

Music-to-Dance Generation with Optimal Transport

Shuang Wu, Shijian Lu, Li Cheng

Dance choreography for a piece of music is a challenging task, having to be creative in presenting distinctive stylistic dance elements while taking into account the musical theme…

cs.CV2025

CasaGPT: Cuboid Arrangement and Scene Assembly for Interior Design

Weitao Feng, Hang Zhou, Jing Liao +2

We present a novel approach for indoor scene synthesis, which learns to arrange decomposed cuboid primitives to represent 3D objects within a scene. Unlike conventional methods tha…

cs.CV2020

3D Human Shape Reconstruction from a Polarization Image

Shihao Zou, Xinxin Zuo, Yiming Qian +4

This paper tackles the problem of estimating 3D body shape of clothed humans from single polarized 2D images, i.e. polarization images. Polarization images are known to be able to…

cond-mat.str-el2020

Giant renormalization of correlation strength in 1T-TaS2 by lattice vibration

Li Cheng, Shunhong Zhang, Shuang Qiao +3

The lattice thermodynamics of a 1T-TaS2 layer, e.g. the spontaneous formation of a sqrt13*sqrt13 commensurate charge density wave (CCDW) and vibrations around the equilibrium posit…

cs.CV2021

3D Pose Estimation and Future Motion Prediction from 2D Images

Ji Yang, Youdong Ma, Xinxin Zuo +3

This paper considers to jointly tackle the highly correlated tasks of estimating 3D human body poses and predicting future 3D motions from RGB image sequences. Based on Lie algebra…

stat.ML2017

An Interval-Based Bayesian Generative Model for Human Complex Activity Recognition

Li Liu, Yongzhong Yang, Lakshmi Narasimhan Govindarajan +4

Complex activity recognition is challenging due to the inherent uncertainty and diversity of performing a complex activity. Normally, each instance of a complex activity has its ow…

cs.RO2025

Act to See, See to Act: Diffusion-Driven Perception-Action Interplay for Adaptive Policies

Jing Wang, Weiting Peng, Jing Tang +4

Existing imitation learning methods decouple perception and action, which overlooks the causal reciprocity between sensory representations and action execution that humans naturall…

cs.CV2022

Promoting Saliency From Depth: Deep Unsupervised RGB-D Saliency Detection

Wei Ji, Jingjing Li, Qi Bi +3

Growing interests in RGB-D salient object detection (RGB-D SOD) have been witnessed in recent years, owing partly to the popularity of depth sensors and the rapid progress of deep…

cs.CV2016

Lie-X: Depth Image Based Articulated Object Pose Estimation, Tracking, and Action Recognition on Lie Groups

Chi Xu, Lakshmi Narasimhan Govindarajan, Yu Zhang +1

Pose estimation, tracking, and action recognition of articulated objects from depth images are important and challenging problems, which are normally considered separately. In this…

cs.CV2025

InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling

Muhammad Gohar Javed, Chuan Guo, Li Cheng +1

Generating realistic 3D human-human interactions from textual descriptions remains a challenging task. Existing approaches, typically based on diffusion models, often produce resul…

physics.app-ph2019

Robust 3D multi-polar acoustic metamaterials with broadband double negativity

Hao-Wen Dong, Sheng-Dong Zhao, Yue-Sheng Wang +2

Acoustic negative-index metamaterials show promise in achieving superlensing for diagnostic medical imaging. In spite of the recent progress made in this field, most metamaterials…

cond-mat.supr-con2008

Magnetic field induced incommensurate resonance in cuprate superconductors

Jingge Zhang, Li Cheng, Huaiming Guo +1

The influence of a uniform external magnetic field on the dynamical spin response of cuprate superconductors in the superconducting state is studied based on the kinetic energy dri…

cs.CV2019

TBC-Net: A real-time detector for infrared small target detection using semantic constraint

Mingxin Zhao, Li Cheng, Xu Yang +3

Infrared small target detection is a key technique in infrared search and tracking (IRST) systems. Although deep learning has been widely used in the vision tasks of visible light…

cs.CV2025

MotionScript: Natural Language Descriptions for Expressive 3D Human Motions

Payam Jome Yazdian, Rachel Lagasse, Hamid Mohammadi +3

We introduce MotionScript, a novel framework for generating highly detailed, natural language descriptions of 3D human motions. Unlike existing motion datasets that rely on broad a…

cs.CV2021

CHASE: Robust Visual Tracking via Cell-Level Differentiable Neural Architecture Search

Seyed Mojtaba Marvasti-Zadeh, Javad Khaghani, Li Cheng +2

A strong visual object tracker nowadays relies on its well-crafted modules, which typically consist of manually-designed network architectures to deliver high-quality tracking resu…

cs.CV2026

PICS: Pairwise Image Compositing with Spatial Interactions

Hang Zhou, Xinxin Zuo, Sen Wang +1

Despite strong single-turn performance, diffusion-based image compositing often struggles to preserve coherent spatial relations in pairwise or sequential edits, where subsequent i…

physics.ins-det2016

The test of the electronics system for the BESIII ETOF upgrade

Wang Xiaozhuang, Dai Hongliang, Wu Zhi +8

It is proposed to upgrade the endcap time-of-flight (ETOF) of the Beijing Spectrometer III (BESIII) with multi-gap resistive plate chamber (MRPC), aiming at overall time resolution…

cs.CV2024

GSD: View-Guided Gaussian Splatting Diffusion for 3D Reconstruction

Yuxuan Mu, Xinxin Zuo, Chuan Guo +7

We present GSD, a diffusion model approach based on Gaussian Splatting (GS) representation for 3D object reconstruction from a single view. Prior works suffer from inconsistent 3D…

cs.CL2021

Automated Generation of Accurate \& Fluent Medical X-ray Reports

Hoang T. N. Nguyen, Dong Nie, Taivanbat Badamdorj +4

Our paper focuses on automating the generation of medical reports from chest X-ray image inputs, a critical yet time-consuming task for radiologists. Unlike existing medical re-por…

cs.CV2025

BOOTPLACE: Bootstrapped Object Placement with Detection Transformers

Hang Zhou, Xinxin Zuo, Rui Ma +1

In this paper, we tackle the copy-paste image-to-image composition problem with a focus on object placement learning. Prior methods have leveraged generative models to reduce the r…

cs.CV2008

Learning Graph Matching

Tiberio S. Caetano, Julian J. McAuley, Li Cheng +2

As a fundamental problem in pattern recognition, graph matching has applications in a variety of fields, from computer vision to computational biology. In graph matching, patterns…