Publications (38)
Possible molecular states from interactions of charmed baryons
Dan Song, Lin-Qing Song, Shu-Yi Kong +1
In this work, we perform a systematic study of possible molecular states composed of two charmed baryons including hidden-charm systems , ,…
Automated Generation of Geometric Theorems from Images of Diagrams
Xiaoyu Chen, Dan Song, Dongming Wang
We propose an approach to generate geometric theorems from electronic images of diagrams automatically. The approach makes use of techniques of Hough transform to recognize geometr…
Uncertainty-Gated Deformable Network for Breast Tumor Segmentation in MR Images
Yue Zhang, Jiahua Dong, Chengtao Peng +3
Accurate segmentation of breast tumors in magnetic resonance images (MRI) is essential for breast cancer diagnosis, yet existing methods face challenges in capturing irregular tumo…
CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model
Jianhao Zeng, Dan Song, Weizhi Nie +3
Generative Adversarial Networks (GANs) dominate the research field in image-based virtual try-on, but have not resolved problems such as unnatural deformation of garments and the b…
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
Diqiong Jiang, Kai Zhu, Dan Song +3
Speech-driven 3D facial animation aims to generate realistic and expressive facial motions directly from audio. While recent methods achieve high-quality lip synchronization, they…
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
Jianhao Zeng, Yancheng Bai, Ruidong Chen +7
Video virtual try-on technology provides a cost-effective solution for creating marketing videos in fashion e-commerce. However, its practical adoption is hindered by two critical…
Heavy-strange meson molecules and possible candidates , , and
Shu-Yi Kong, Jun-Tao Zhu, Dan Song +1
In this work, we systematically investigate the heavy-strange meson systems, and , to study possible mole…
Achieving Short-Blocklength RCU bound via CRC List Decoding of TCM with Probabilistic Shaping
Linfang Wang, Dan Song, Felipe Areces +1
This paper applies probabilistic amplitude shaping (PAS) to a cyclic redundancy check (CRC) aided trellis coded modulation (TCM) to achieve the short-blocklength random coding unio…
Log-Likelihood Loss for Semantic Compression
Anuj Kumar Yadav, Dan Song, Yanina Shkel +1
We study lossy source coding under a distortion measure defined by the negative log-likelihood induced by a prescribed conditional distribution . This \emph{log-likelihood…
Multispectral Pedestrian Detection via Simultaneous Detection and Segmentation
Chengyang Li, Dan Song, Ruofeng Tong +1
Multispectral pedestrian detection has attracted increasing attention from the research community due to its crucial competence for many around-the-clock applications (e.g., video…
Towards Deconfounded Image-Text Matching with Causal Inference
Wenhui Li, Xinqi Su, Dan Song +3
Prior image-text matching methods have shown remarkable performance on many benchmark datasets, but most of them overlook the bias in the dataset, which exists in intra-modal and i…
Chest X-ray Image Classification: A Causal Perspective
Weizhi Nie, Chen Zhang, Dan Song +4
The chest X-ray (CXR) is one of the most common and easy-to-get medical tests used to diagnose common diseases of the chest. Recently, many deep learning-based methods have been pr…
CTooth+: A Large-scale Dental Cone Beam Computed Tomography Dataset and Benchmark for Tooth Volume Segmentation
Weiwei Cui, Yaqi Wang, Yilong Li +8
Accurate tooth volume segmentation is a prerequisite for computer-aided dental analysis. Deep learning-based tooth segmentation methods have achieved satisfying performances but re…
Temporal-spatial Correlation Attention Network for Clinical Data Analysis in Intensive Care Unit
Weizhi Nie, Yuhe Yu, Chen Zhang +3
In recent years, medical information technology has made it possible for electronic health record (EHR) to store fairly complete clinical data. This has brought health care into th…
Comparing Human and AI Rater Effects Using the Many-Facet Rasch Model
Hong Jiao, Dan Song, Won-Chan Lee
Large language models (LLMs) have been widely explored for automated scoring in low-stakes assessment to facilitate learning and instruction. Empirical evidence related to which LL…
Possible molecular states from interactions of charmed strange baryons
Dan Song, Shu Chen, Shu-Yi Kong +1
In this work, we perform an investigation of possible molecular states composed of two charmed strange baryons from the interaction, and their hidden-cha…
Illumination-aware Faster R-CNN for Robust Multispectral Pedestrian Detection
Chengyang Li, Dan Song, Ruofeng Tong +1
Multispectral images of color-thermal pairs have shown more effective than a single color channel for pedestrian detection, especially under challenging illumination conditions. Ho…
Comparison of Scoring Rationales Between Large Language Models and Human Raters
Haowei Hua, Hong Jiao, Dan Song
Advances in automated scoring are closely aligned with advances in machine-learning and natural-language-processing techniques. With recent progress in large language models (LLMs)…
The Rise of Artificial Intelligence in Educational Measurement: Opportunities and Ethical Challenges
Okan Bulut, Maggie Beiting-Parrish, Jodi M. Casabianca +14
The integration of artificial intelligence (AI) in educational measurement has revolutionized assessment methods, enabling automated scoring, rapid content analysis, and personaliz…
MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment
Zhiting Gao, Dan Song, Diqiong Jiang +2
Motion generation is essential for animating virtual characters and embodied agents. While recent text-driven methods have made significant strides, they often struggle with achiev…
Group Relative Attention Guidance for Image Editing
Xuanpu Zhang, Xuesong Niu, Ruidong Chen +6
Recently, image editing based on Diffusion-in-Transformer models has undergone rapid development. However, existing editing methods often lack effective control over the degree of…
Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on
Dan Song, Xuanpu Zhang, Jianhao Zeng +4
Image-based virtual try-on aims to transfer target in-shop clothing to a dressed model image, the objectives of which are totally taking off original clothing while preserving the…
Privacy Amplification via Compression: Achieving the Optimal Privacy-Accuracy-Communication Trade-off in Distributed Mean Estimation
Wei-Ning Chen, Dan Song, Ayfer Ozgur +1
Privacy and communication constraints are two major bottlenecks in federated learning (FL) and analytics (FA). We study the optimal accuracy of mean and frequency estimation (canon…
Hidden and doubly heavy molecular states from interactions / and /
Zuo-Ming Ding, Han-Yu Jiang, Dan Song +1
In this work, we perform a systematical investigation about the possible hidden and doubly heavy molecular states with open and hidden strangeness from interactions of $D^{(*)}{\ba…
Possible molecular states and their productions in nulceon-antinulceon collision
Lin-Qing Song, Dan Song, Jun-Tao Zhu +1
In this work, a study of possible molecular states from the interaction and their productions in nucleon-antinucleon collision is performed in a quasipotential Bethe…
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
Xuanpu Zhang, Dan Song, Pengxin Zhan +5
Image-based virtual try-on is an increasingly popular and important task to generate realistic try-on images of the specific person. Recent methods model virtual try-on as image ma…
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
Ruidong Chen, Yancheng Bai, Xuanpu Zhang +6
Region-instructed layout control in text-to-image generation is highly practical, yet existing methods suffer from limitations: (i) training-based approaches inherit data bias and…
Deep Reinforcement Learning Framework for Thoracic Diseases Classification via Prior Knowledge Guidance
Weizhi Nie, Chen Zhang, Dan Song +4
The chest X-ray is often utilized for diagnosing common thoracic diseases. In recent years, many approaches have been proposed to handle the problem of automatic diagnosis based on…
Image-Based Virtual Try-On: A Survey
Dan Song, Xuanpu Zhang, Juan Zhou +4
Image-based virtual try-on aims to synthesize a naturally dressed person image with a clothing image, which revolutionizes online shopping and inspires related topics within image…
Efficiently Computable Converses for Finite-Blocklength Communication
Felipe Areces, Dan Song, Richard Wesel +1
This paper presents a method for computing a finite-blocklength converse for the rate of fixed-length codes with feedback used on discrete memoryless channels (DMCs). The new conve…
Unified Multi-Modal Image Synthesis for Missing Modality Imputation
Yue Zhang, Chengtao Peng, Qiuli Wang +3
Multi-modal medical images provide complementary soft-tissue characteristics that aid in the screening and diagnosis of diseases. However, limited scanning time, image corruption a…
Domain Adaptation from Generated Multi-Weather Images for Unsupervised Maritime Object Classification
Dan Song, Shumeng Huo, Wenhui Li +3
The classification and recognition of maritime objects are crucial for enhancing maritime safety, monitoring, and intelligent sea environment prediction. However, existing unsuperv…
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
Dan Song, Xinwei Fu, Ning Liu +5
Large-scale pre-trained models have demonstrated impressive performance in vision and language tasks within open-world scenarios. Due to the lack of comparable pre-trained models f…
Empirical Comparison of Encoder-Based Language Models and Feature-Based Supervised Machine Learning Approaches to Automated Scoring of Long Essays
Kuo Wang, Haowei Hua, Pengfei Yan +2
Long context may impose challenges for encoder-only language models in text processing, specifically for automated scoring of essays. This study trained several commonly used encod…
CTooth: A Fully Annotated 3D Dataset and Benchmark for Tooth Volume Segmentation on Cone Beam Computed Tomography Images
Weiwei Cui, Yaqi Wang, Qianni Zhang +5
3D tooth segmentation is a prerequisite for computer-aided dental diagnosis and treatment. However, segmenting all tooth regions manually is subjective and time-consuming. Recently…
Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory
Dan Song, Won-Chan Lee, Hong Jiao
This study investigates the estimation of reliability for large language models (LLMs) in scoring writing tasks from the AP Chinese Language and Culture Exam. Using generalizabilit…
Tests of Cold Atom Clock in Orbit
Liang Liu, Desheng Lü, Weibiao Chen +23
Since the atomic clock was invented, its performance has been improved for one digit every decade until 90s of last century when the traditional atomic clock almost reached its lim…
The Performance of Compression-Based Denoisers
Dan Song, Ayfer Ãzgür, Tsachy Weissman
We consider a denoiser that reconstructs a stationary ergodic source by lossily compressing samples of the source observed through a memoryless noisy channel. Prior work on compres…