activity
20202026
most citedRethinking Spatially-Adaptive Normalization

15 citations · 57 across the 19 of their papers we have counts for

collaborators

20 papers

cs.CV2026

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs

Xuanpu Zhao, Zhentao Tan, Dianmo Sheng +6

To enhance the perception and reasoning capabilities of multimodal large language models in complex visual scenes, recent research has introduced agent-based workflows. In these wo…

cs.CL2025

Flora: Effortless Context Construction to Arbitrary Length and Scale

Tianxiang Chen, Zhentao Tan, Xiaofan Bo +5

Effectively handling long contexts is challenging for Large Language Models (LLMs) due to the rarity of long texts, high computational demands, and substantial forgetting of short-…

cs.CL2024★ 1 cited

Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection

Tianxiang Chen, Zhentao Tan, Tao Gong +5

As a manner to augment pre-trained large language models (LLM), knowledge injection is critical to develop vertical domain large models and has been widely studied. Although most c…

cs.CV2024★ 1 cited

Mixture-of-Noises Enhanced Forgery-Aware Predictor for Multi-Face Manipulation Detection and Localization

Changtao Miao, Qi Chu, Tao Gong +6

With the advancement of face manipulation technology, forgery images in multi-face scenarios are gradually becoming a more complex and realistic challenge. Despite this, detection…

cs.CV2024

Transformer based Pluralistic Image Completion with Reduced Information Loss

Qiankun Liu, Yuqi Jiang, Zhentao Tan +5

Transformer based methods have achieved great success in image inpainting recently. However, we find that these solutions regard each pixel as a token, thus suffering from an infor…

cs.CV2024★ 3 cited

MiM-ISTD: Mamba-in-Mamba for Efficient Infrared Small Target Detection

Tianxiang Chen, Zi Ye, Zhentao Tan +6

Recently, infrared small target detection (ISTD) has made significant progress, thanks to the development of basic models. Specifically, the models combining CNNs with transformers…