activity
20242026
collaborators

5 papers

cs.DC2026

Agentic Kernel Optimization: Generating State-of-the-Art GPU Kernels Without Hand-Written CUDA

Mao Luo, Hongbin Li, Feng Lin +2

We study whether general-purpose code agents can produce state-of-the-art GPU kernels without any manually written CUDA code. We investigate this question using representative work…

cs.CV2026

Reg-TTR, Test-Time Refinement for Fast, Robust and Accurate Image Registration

Lin Chen, Yue He, Fengting Zhang +4

Traditional image registration methods are robust but slow due to their iterative nature. While deep learning has accelerated inference, it often struggles with domain shifts. Emer…

cs.CV2025

Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation

Feng Lin, Marco Chen, Haokui Zhang +3

This paper investigates the role of attention heads in CLIP's image encoder. Building on interpretability studies, we conduct an exhaustive analysis and find that certain heads, di…

cs.CV2025

Vision Language Models: A Survey of 26K Papers

Fengming Lin

We present a transparent, reproducible measurement of research trends across 26,104 accepted papers from CVPR, ICLR, and NeurIPS spanning 2023-2025. Titles and abstracts are normal…

cs.CV2024

REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation

Xizhe Xue, Guoting Wei, Hao Chen +4

The rapid evolution of Vision Language Models (VLMs) has catalyzed significant advancements in artificial intelligence, expanding research across various disciplines, including Ear…