papers

Publications (38)

hep-ph2023

Possible molecular states from interactions of charmed baryons

Dan Song, Lin-Qing Song, Shu-Yi Kong +1

In this work, we perform a systematic study of possible molecular states composed of two charmed baryons including hidden-charm systems , ,…

cs.AI2014

Automated Generation of Geometric Theorems from Images of Diagrams

Xiaoyu Chen, Dan Song, Dongming Wang

We propose an approach to generate geometric theorems from electronic images of diagrams automatically. The approach makes use of techniques of Hough transform to recognize geometr…

eess.IV2025

Uncertainty-Gated Deformable Network for Breast Tumor Segmentation in MR Images

Yue Zhang, Jiahua Dong, Chengtao Peng +3

Accurate segmentation of breast tumors in magnetic resonance images (MRI) is essential for breast cancer diagnosis, yet existing methods face challenges in capturing irregular tumo…

cs.CV2024

CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model

Jianhao Zeng, Dan Song, Weizhi Nie +3

Generative Adversarial Networks (GANs) dominate the research field in image-based virtual try-on, but have not resolved problems such as unnatural deformation of garments and the b…

cs.MM2026

EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing

Diqiong Jiang, Kai Zhu, Dan Song +3

Speech-driven 3D facial animation aims to generate realistic and expressive facial motions directly from audio. While recent methods achieve high-quality lip synchronization, they…

cs.CV2026

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on

Jianhao Zeng, Yancheng Bai, Ruidong Chen +7

Video virtual try-on technology provides a cost-effective solution for creating marketing videos in fashion e-commerce. However, its practical adoption is hindered by two critical…

hep-ph2021

Heavy-strange meson molecules and possible candidates , , and

Shu-Yi Kong, Jun-Tao Zhu, Dan Song +1

In this work, we systematically investigate the heavy-strange meson systems, and , to study possible mole…

cs.IT2021

Achieving Short-Blocklength RCU bound via CRC List Decoding of TCM with Probabilistic Shaping

Linfang Wang, Dan Song, Felipe Areces +1

This paper applies probabilistic amplitude shaping (PAS) to a cyclic redundancy check (CRC) aided trellis coded modulation (TCM) to achieve the short-blocklength random coding unio…

cs.IT2026

Log-Likelihood Loss for Semantic Compression

Anuj Kumar Yadav, Dan Song, Yanina Shkel +1

We study lossy source coding under a distortion measure defined by the negative log-likelihood induced by a prescribed conditional distribution . This \emph{log-likelihood…

cs.CV2018

Multispectral Pedestrian Detection via Simultaneous Detection and Segmentation

Chengyang Li, Dan Song, Ruofeng Tong +1

Multispectral pedestrian detection has attracted increasing attention from the research community due to its crucial competence for many around-the-clock applications (e.g., video…

cs.CV2024

Towards Deconfounded Image-Text Matching with Causal Inference

Wenhui Li, Xinqi Su, Dan Song +3

Prior image-text matching methods have shown remarkable performance on many benchmark datasets, but most of them overlook the bias in the dataset, which exists in intra-modal and i…

eess.IV2023

Chest X-ray Image Classification: A Causal Perspective

Weizhi Nie, Chen Zhang, Dan Song +4

The chest X-ray (CXR) is one of the most common and easy-to-get medical tests used to diagnose common diseases of the chest. Recently, many deep learning-based methods have been pr…

eess.IV2022

CTooth+: A Large-scale Dental Cone Beam Computed Tomography Dataset and Benchmark for Tooth Volume Segmentation

Weiwei Cui, Yaqi Wang, Yilong Li +8

Accurate tooth volume segmentation is a prerequisite for computer-aided dental analysis. Deep learning-based tooth segmentation methods have achieved satisfying performances but re…

cs.LG2023

Temporal-spatial Correlation Attention Network for Clinical Data Analysis in Intensive Care Unit

Weizhi Nie, Yuhe Yu, Chen Zhang +3

In recent years, medical information technology has made it possible for electronic health record (EHR) to store fairly complete clinical data. This has brought health care into th…

cs.CL2025

Comparing Human and AI Rater Effects Using the Many-Facet Rasch Model

Hong Jiao, Dan Song, Won-Chan Lee

Large language models (LLMs) have been widely explored for automated scoring in low-stakes assessment to facilitate learning and instruction. Empirical evidence related to which LL…

hep-ph2023

Possible molecular states from interactions of charmed strange baryons

Dan Song, Shu Chen, Shu-Yi Kong +1

In this work, we perform an investigation of possible molecular states composed of two charmed strange baryons from the interaction, and their hidden-cha…

cs.CV2018

Illumination-aware Faster R-CNN for Robust Multispectral Pedestrian Detection

Chengyang Li, Dan Song, Ruofeng Tong +1

Multispectral images of color-thermal pairs have shown more effective than a single color channel for pedestrian detection, especially under challenging illumination conditions. Ho…

cs.CL2025

Comparison of Scoring Rationales Between Large Language Models and Human Raters

Haowei Hua, Hong Jiao, Dan Song

Advances in automated scoring are closely aligned with advances in machine-learning and natural-language-processing techniques. With recent progress in large language models (LLMs)…

cs.CY2024

The Rise of Artificial Intelligence in Educational Measurement: Opportunities and Ethical Challenges

Okan Bulut, Maggie Beiting-Parrish, Jodi M. Casabianca +14

The integration of artificial intelligence (AI) in educational measurement has revolutionized assessment methods, enabling automated scoring, rapid content analysis, and personaliz…

cs.CV2025

MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment

Zhiting Gao, Dan Song, Diqiong Jiang +2

Motion generation is essential for animating virtual characters and embodied agents. While recent text-driven methods have made significant strides, they often struggle with achiev…

cs.CV2025

Group Relative Attention Guidance for Image Editing

Xuanpu Zhang, Xuesong Niu, Ruidong Chen +6

Recently, image editing based on Diffusion-in-Transformer models has undergone rapid development. However, existing editing methods often lack effective control over the degree of…

cs.CV2024

Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on

Dan Song, Xuanpu Zhang, Jianhao Zeng +4

Image-based virtual try-on aims to transfer target in-shop clothing to a dressed model image, the objectives of which are totally taking off original clothing while preserving the…

stat.ML2023

Privacy Amplification via Compression: Achieving the Optimal Privacy-Accuracy-Communication Trade-off in Distributed Mean Estimation

Wei-Ning Chen, Dan Song, Ayfer Ozgur +1

Privacy and communication constraints are two major bottlenecks in federated learning (FL) and analytics (FA). We study the optimal accuracy of mean and frequency estimation (canon…

hep-ph2021

Hidden and doubly heavy molecular states from interactions / and /

Zuo-Ming Ding, Han-Yu Jiang, Dan Song +1

In this work, we perform a systematical investigation about the possible hidden and doubly heavy molecular states with open and hidden strangeness from interactions of $D^{(*)}{\ba…

hep-ph2023

Possible molecular states and their productions in nulceon-antinulceon collision

Lin-Qing Song, Dan Song, Jun-Tao Zhu +1

In this work, a study of possible molecular states from the interaction and their productions in nucleon-antinucleon collision is performed in a quasipotential Bethe…

cs.CV2024

BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training

Xuanpu Zhang, Dan Song, Pengxin Zhan +5

Image-based virtual try-on is an increasingly popular and important task to generate realistic try-on images of the specific person. Recent methods model virtual try-on as image ma…

cs.CV2026

Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers

Ruidong Chen, Yancheng Bai, Xuanpu Zhang +6

Region-instructed layout control in text-to-image generation is highly practical, yet existing methods suffer from limitations: (i) training-based approaches inherit data bias and…

eess.IV2023

Deep Reinforcement Learning Framework for Thoracic Diseases Classification via Prior Knowledge Guidance

Weizhi Nie, Chen Zhang, Dan Song +4

The chest X-ray is often utilized for diagnosing common thoracic diseases. In recent years, many approaches have been proposed to handle the problem of automatic diagnosis based on…

cs.CV2024

Image-Based Virtual Try-On: A Survey

Dan Song, Xuanpu Zhang, Juan Zhou +4

Image-based virtual try-on aims to synthesize a naturally dressed person image with a clothing image, which revolutionizes online shopping and inspires related topics within image…

cs.IT2022

Efficiently Computable Converses for Finite-Blocklength Communication

Felipe Areces, Dan Song, Richard Wesel +1

This paper presents a method for computing a finite-blocklength converse for the rate of fixed-length codes with feedback used on discrete memoryless channels (DMCs). The new conve…

cs.CV2024

Unified Multi-Modal Image Synthesis for Missing Modality Imputation

Yue Zhang, Chengtao Peng, Qiuli Wang +3

Multi-modal medical images provide complementary soft-tissue characteristics that aid in the screening and diagnosis of diseases. However, limited scanning time, image corruption a…

cs.CV2025

Domain Adaptation from Generated Multi-Weather Images for Unsupervised Maritime Object Classification

Dan Song, Shumeng Huo, Wenhui Li +3

The classification and recognition of maritime objects are crucial for enhancing maritime safety, monitoring, and intelligent sea environment prediction. However, existing unsuperv…

cs.CV2024

MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition

Dan Song, Xinwei Fu, Ning Liu +5

Large-scale pre-trained models have demonstrated impressive performance in vision and language tasks within open-world scenarios. Due to the lack of comparable pre-trained models f…

cs.CL2026

Empirical Comparison of Encoder-Based Language Models and Feature-Based Supervised Machine Learning Approaches to Automated Scoring of Long Essays

Kuo Wang, Haowei Hua, Pengfei Yan +2

Long context may impose challenges for encoder-only language models in text processing, specifically for automated scoring of essays. This study trained several commonly used encod…

cs.CV2022

CTooth: A Fully Annotated 3D Dataset and Benchmark for Tooth Volume Segmentation on Cone Beam Computed Tomography Images

Weiwei Cui, Yaqi Wang, Qianni Zhang +5

3D tooth segmentation is a prerequisite for computer-aided dental diagnosis and treatment. However, segmenting all tooth regions manually is subjective and time-consuming. Recently…

cs.CL2025

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory

Dan Song, Won-Chan Lee, Hong Jiao

This study investigates the estimation of reliability for large language models (LLMs) in scoring writing tasks from the AP Chinese Language and Culture Exam. Using generalizabilit…

physics.atom-ph2017

Tests of Cold Atom Clock in Orbit

Liang Liu, Desheng Lü, Weibiao Chen +23

Since the atomic clock was invented, its performance has been improved for one digit every decade until 90s of last century when the traditional atomic clock almost reached its lim…

cs.IT2026

The Performance of Compression-Based Denoisers

Dan Song, Ayfer Özgür, Tsachy Weissman

We consider a denoiser that reconstructs a stationary ergodic source by lossily compressing samples of the source observed through a memoryless noisy channel. Prior work on compres…