papers

Publications (34)

cs.CV2024

MMDS: A Multimodal Medical Diagnosis System Integrating Image Analysis and Knowledge-based Departmental Consultation

Yi Ren, HanZhi Zhang, Weibin Li +5

We present MMDS, a system capable of recognizing medical images and patient facial details, and providing professional medical diagnoses. The system consists of two core components…

cs.CV2021

Fine-Grained Control of Artistic Styles in Image Generation

Xin Miao, Huayan Wang, Jun Fu +3

Recent advances in generative models and adversarial training have enabled artificially generating artworks in various artistic styles. It is highly desirable to gain more control…

cs.CV2023

A Close Look at Few-shot Real Image Super-resolution from the Distortion Relation Perspective

Xin Li, Xin Jin, Jun Fu +3

Collecting amounts of distorted/clean image pairs in the real world is non-trivial, which seriously limits the practical applications of these supervised learning-based methods on…

cs.CV2022

ImageSubject: A Large-scale Dataset for Subject Detection

Xin Miao, Jiayi Liu, Huayan Wang +1

Main subjects usually exist in the images or videos, as they are the objects that the photographer wants to highlight. Human viewers can easily identify them but algorithms often c…

cs.PL2025

QPanda3: A High-Performance Software-Hardware Collaborative Framework for Large-Scale Quantum-Classical Computing Integration

Tianrui Zou, Yuan Fang, Jing Wang +8

In emerging quantum-classical integration applications, the classical time cost-especially from compilation and protocol-level communication often exceeds the execution time of qua…

cs.CV2019

Adaptive Context Network for Scene Parsing

Jun Fu, Jing Liu, Yuhang Wang +4

Recent works attempt to improve scene parsing performance by exploring different levels of contexts, and typically train a well-designed convolutional network to exploit useful con…

cs.CV2025

ParkFormer: A Transformer-Based Parking Policy with Goal Embedding and Pedestrian-Aware Control

Jun Fu, Bin Tian, Haonan Chen +2

Autonomous parking plays a vital role in intelligent vehicle systems, particularly in constrained urban environments where high-precision control is required. While traditional rul…

cs.CV2026

ParkingScenes: A Structured Dataset for End-to-End Autonomous Parking in Simulation Scenes

Haonan Chen, Kaiwen Xiao, Bin Tian +1

Autonomous parking remains a critical yet challenging task in intelligent driving systems, particularly within constrained urban environments where maneuvering space is limited and…

cond-mat.mes-hall2024

Electric imaging and dynamics of photo-charged graphene edge

Zhe Ding, Zhousheng Chen, Xiaodong Fan +16

The one-dimensional side gate based on graphene edges shows a significant capability of reducing the channel length of field-effect transistors, further increasing the integration…

eess.IV2020

Residual Squeeze-and-Excitation Network for Fast Image Deraining

Jun Fu, Jianfeng Xu, Kazuyuki Tasaka +1

Image deraining is an important image processing task as rain streaks not only severely degrade the visual quality of images but also significantly affect the performance of high-l…

eess.SY2015

Direct Adaptive Controller for Uncertain MIMO Dynamic Systems with Time-varying Delay and Dead-zone Inputs

Zhijun Li, Ziting Chen, Jun Fu +1

This paper presents an adaptive tracking control method for a class of nonlinearly parameterized MIMO dynamic systems with time-varying delay and unknown nonlinear dead-zone inputs…

math.OC2025

A cutting-surface consensus approach for distributed robust optimization of multi-agent systems

Jun Fu, Xunhao Wu

A novel and fully distributed optimization method is proposed for the distributed robust convex program (DRCP) over a time-varying unbalanced directed network under the uniformly j…

cond-mat.mtrl-sci2021

Ferroelectric gating of the PL in MoSe2

Xiaoyu Mao, Jun Fu, Ming Gong +1

We demonstrate a 2D ferroelectric heterostructure with monolayer MoSe and CuInPS (CIPS). In the heterostructure, the electric polarization of CIPS results in c electron…

math.OC2026

A Taylor-Bernstein Inner Approximation Algorithm for Path-Constrained Dynamic Optimization

Yuan Chang, Lizhong Jiang, Tai-Fang Li +1

A novel inner approximation algorithm is proposed for dynamic optimization problems to ensure strict satisfaction of path constraints. Distinct from traditional methods relying on…

cs.RO2026

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models

Guiyu Zhao, Longteng Guo, Junyou Zhu +6

Vision-language-action (VLA) models have shown strong promise for robotic manipulation, but their reliability at test time remains limited by one-shot action prediction, where even…

cs.CV2025

Structure-preserving Feature Alignment for Old Photo Colorization

Yingxue Pang, Xin Jin, Jun Fu +1

Deep learning techniques have made significant advancements in reference-based colorization by training on large-scale datasets. However, directly applying these methods to the tas…

cs.MM2022

RTN: Reinforced Transformer Network for Coronary CT Angiography Vessel-level Image Quality Assessment

Yiting Lu, Jun Fu, Xin Li +6

Coronary CT Angiography (CCTA) is susceptible to various distortions (e.g., artifacts and noise), which severely compromise the exact diagnosis of cardiovascular diseases. The appr…

cs.CV2024

Vision-Language Consistency Guided Multi-modal Prompt Learning for Blind AI Generated Image Quality Assessment

Jun Fu, Wei Zhou, Qiuping Jiang +2

Recently, textual prompt tuning has shown inspirational performance in adapting Contrastive Language-Image Pre-training (CLIP) models to natural image quality assessment. However,…

cs.LG2020

Bayesian Spatio-Temporal Graph Convolutional Network for Traffic Forecasting

Jun Fu, Wei Zhou, Zhibo Chen

In traffic forecasting, graph convolutional networks (GCNs), which model traffic flows as spatio-temporal graphs, have achieved remarkable performance. However, existing GCN-based…

cs.CV2025

Cross-Modal Scene Semantic Alignment for Image Complexity Assessment

Yuqing Luo, Yixiao Li, Jiang Liu +7

Image complexity assessment (ICA) is a challenging task in perceptual evaluation due to the subjective nature of human perception and the inherent semantic diversity in real-world…

cs.CV2025

Perception-oriented Bidirectional Attention Network for Image Super-resolution Quality Assessment

Yixiao Li, Xiaoyuan Yang, Guanghui Yue +6

Many super-resolution (SR) algorithms have been proposed to increase image resolution. However, full-reference (FR) image quality assessment (IQA) metrics for comparing and evaluat…

cs.CV2019

Dual Attention Network for Scene Segmentation

Jun Fu, Jing Liu, Haijie Tian +4

In this paper, we address the scene segmentation task by capturing rich contextual dependencies based on the selfattention mechanism. Unlike previous works that capture contexts by…

math.OC2023

Distributed robust optimization for multi-agent systems with guaranteed finite-time convergence

Xunhao Wu, Jun Fu

A novel distributed algorithm is proposed for finite-time converging to a feasible consensus solution satisfying global optimality to a certain accuracy of the distributed robust c…

eess.IV2021

Adaptive Hypergraph Convolutional Network for No-Reference 360-degree Image Quality Assessment

Jun Fu, Chen Hou, Wei Zhou +2

In no-reference 360-degree image quality assessment (NR 360IQA), graph convolutional networks (GCNs), which model interactions between viewports through graphs, have achieved impre…

cs.MM2024

Deep Bi-directional Attention Network for Image Super-Resolution Quality Assessment

Yixiao Li, Xiaoyuan Yang, Jun Fu +2

There has emerged a growing interest in exploring efficient quality assessment algorithms for image super-resolution (SR). However, employing deep learning techniques, especially d…

cs.CV2023

Scale Guided Hypernetwork for Blind Super-Resolution Image Quality Assessment

Jun Fu

With the emergence of image super-resolution (SR) algorithm, how to blindly evaluate the quality of super-resolution images has become an urgent task. However, existing blind SR im…

cs.CV2022

Mutual Attention-based Hybrid Dimensional Network for Multimodal Imaging Computer-aided Diagnosis

Yin Dai, Yifan Gao, Fayu Liu +1

Recent works on Multimodal 3D Computer-aided diagnosis have demonstrated that obtaining a competitive automatic diagnosis model when a 3D convolution neural network (CNN) brings mo…

eess.IV2022

Parotid Gland MRI Segmentation Based on Swin-Unet and Multimodal Images

Zi'an Xu, Yin Dai, Fayu Liu +4

Background and objective: Parotid gland tumors account for approximately 2% to 10% of head and neck tumors. Preoperative tumor localization, differential diagnosis, and subsequent…

cs.LG2023

Variational Disentangled Graph Auto-Encoders for Link Prediction

Jun Fu, Xiaojuan Zhang, Shuang Li +1

With the explosion of graph-structured data, link prediction has emerged as an increasingly important task. Embedding methods for link prediction utilize neural networks to generat…

cs.LG2021

Bayesian Graph Convolutional Network for Traffic Prediction

Jun Fu, Wei Zhou, Zhibo Chen

Recently, adaptive graph convolutional network based traffic prediction methods, learning a latent graph structure from traffic data via various attention-based mechanisms, have ac…

cs.CV2017

Stacked Deconvolutional Network for Semantic Segmentation

Jun Fu, Jing Liu, Yuhang Wang +1

Recent progress in semantic segmentation has been driven by improving the spatial resolution under Fully Convolutional Networks (FCNs). To address this problem, we propose a Stacke…

cs.LG2023

Contrastive Disentangled Learning on Graph for Node Classification

Xiaojuan Zhang, Jun Fu, Shuang Li

Contrastive learning methods have attracted considerable attention due to their remarkable success in analyzing graph-structured data. Inspired by the success of contrastive learni…

cs.MM2020

Sequential Reinforced 360-Degree Video Adaptive Streaming with Cross-user Attentive Network

Jun Fu, Zhibo Chen, Xiaoming Chen +1

In the tile-based 360-degree video streaming, predicting user's future viewpoints and developing adaptive bitrate (ABR) algorithms are essential for optimizing user's quality of ex…

cs.CV2025

CLIP-DQA: Blindly Evaluating Dehazed Images from Global and Local Perspectives Using CLIP

Yirui Zeng, Jun Fu, Hadi Amirpour +5

Blind dehazed image quality assessment (BDQA), which aims to accurately predict the visual quality of dehazed images without any reference information, is essential for the evaluat…