17 papers
Attenuation-Resilient Alternating Optimization for Laparoscopic Liver Landmark Detection
Lanqing Liu, Ruize Cui, Jialun Pei +4
Liver surface landmark detection is a fundamental prerequisite for anatomical guidance in laparoscopic liver surgery. However, it remains unreliable in practice due to two pervasiv…
HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images
Yichen Liu, Donghao Zhou, Jie Wang +9
Human-product images, which showcase the integration of humans and products, play a vital role in advertising, e-commerce, and digital marketing. The essential challenge of generat…
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis
Yang Yu, Dunyuan Xu, Yaoqian Li +3
3D medical image analysis is of great importance in disease diagnosis and treatment. Recently, multimodal large language models (MLLMs) have exhibited robust perceptual capacity, s…
Benchmarking Endoscopic Surgical Image Restoration and Beyond
Jialun Pei, Diandian Guo, Donghui Yang +5
In endoscopic surgery, a clear and high-quality visual field is critical for surgeons to make accurate intraoperative decisions. However, persistent visual degradation, including s…
Synergistic Bleeding Region and Point Detection in Laparoscopic Surgical Videos
Jialun Pei, Zhangjun Zhou, Diandian Guo +4
Intraoperative bleeding in laparoscopic surgery causes rapid obscuration of the operative field to hinder the surgical process and increases the risk of postoperative complications…
Bi-TEAM: A Unified Cross-Scale Representation Learning Framework for Chemically Modified Biomolecules
Chunbin Gu, Zijun Gao, Mutian He +8
Representation learning for protein biochemical space faces a difficult trade-off: protein language models excel at capturing long-range biological semantics but often miss fine-gr…