papers

Publications (92)

math.FA2022

Bilinear integral operator on Morrey-Banach spaces and its application

Huihui Zhang, Xiangxing Tao, Yandan Zhang +1

In this paper, we give the definability of bilinear singular and fractional integral operators on Morrey-Banach space, as well as their commutators and we prove the boundedness of…

cs.CL2025

FocalPO: Enhancing Preference Optimizing by Focusing on Correct Preference Rankings

Tong Liu, Xiao Yu, Wenxuan Zhou +2

Efficient preference optimization algorithms such as Direct Preference Optimization (DPO) have become a popular approach in aligning large language models (LLMs) with human prefere…

cond-mat.mtrl-sci2025

Optimizing the depth-dependent nitrogen-vacancy center quantum sensor in diamane

Pei Li, Guanjian Hu, Xiao Yu +2

Negatively charged nitrogen-vacancy (NV) center in diamond is the representative solid state defect qubit for quantum information science, offering long coherence time at room temp…

cs.CV2023

Multi-Frame Self-Supervised Depth Estimation with Multi-Scale Feature Fusion in Dynamic Scenes

Jiquan Zhong, Xiaolin Huang, Xiao Yu

Multi-frame methods improve monocular depth estimation over single-frame approaches by aggregating spatial-temporal information via feature matching. However, the spatial-temporal…

math.AP2015

Traveling waves and spreading speeds for time-space periodic monotone systems

Jian Fang, Xiao Yu, Xiao-Qiang Zhao

The theory of traveling waves and spreading speeds is developed for time-space periodic monotone semiflows with monostable structure. By using traveling waves of the associated Poi…

cs.LG2019

Adversarial Defense Framework for Graph Neural Network

Shen Wang, Zhengzhang Chen, Jingchao Ni +4

Graph neural network (GNN), as a powerful representation learning model on graph data, attracts much attention across various disciplines. However, recent studies show that GNN is…

cs.SE2025

Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models

Xiao Yu, Haoxuan Chen, Feifei Niu +3

With the rapid development of large language models (LLMs), distributed training and inference frameworks like DeepSpeed have become essential for scaling model training and infere…

cs.IT2024

Iterative decoding of short BCH codes and its post-processing

Guangwen Li, Xiao Yu

Effective iterative decoding of short BCH codes faces two primary challenges: identifying an appropriate parity-check matrix and accelerating decoder convergence. To address these…

cs.SE2026

When Models Meet Users: An Empirical Study of Perceptions of General LLMs and Multimodal LLMs on Hugging Face

Yujian Liu, Xiao Yu, Jacky Keung +3

The paper empirically examines user discussions on Hugging Face to understand how people perceive general-purpose and multimodal large language models, identifying key concerns suc…

#large language models#multimodal models#user perception#empirical study
hep-ph2026

Strong decays and effective spin-symmetry-breaking corrections in excited charm-strange mesons

Xiao Yu, Chao-Qiang Geng

We study two-body pseudoscalar-emission decays of excited charm-strange mesons in heavy meson effective field theory, where phenomenological \(1/m_c\) corrections are encoded as ef…

math.AP2021

Asymptotic spreading of KPP reactive fronts in heterogeneous shifting environments

King-Yeung Lam, Xiao Yu

We study the asymptotic spreading of Kolmogorov-Petrovsky-Piskunov (KPP) fronts in heterogeneous shifting habitats, with any number of shifting speeds, by further developing the me…

cs.CL2024

ConFit: Improving Resume-Job Matching using Data Augmentation and Contrastive Learning

Xiao Yu, Jinzhong Zhang, Zhou Yu

A reliable resume-job matching system helps a company find suitable candidates from a pool of resumes, and helps a job seeker find relevant jobs from a list of job posts. However,…

cs.SE2024

How Far Have We Gone in Binary Code Understanding Using Large Language Models

Xiuwei Shang, Shaoyin Cheng, Guoqiang Chen +6

Binary code analysis plays a pivotal role in various software security applications, such as software maintenance, malware detection, software vulnerability discovery, patch analys…

physics.atom-ph2022

Robust single-sideband-modulated Raman light generation for atom interferometry by FBG-based optical rectangular filtration

Guochao Wang, Yaning Wang, Kang Ying +10

Low-phase-noise and pure-spectrum Raman light is vital for high-precision atom interferometry by two-photon Raman transition. A preferred and prevalent solution for Raman light gen…

cs.CL2026

Reinforcement World Model Learning for LLM-based Agents

Xiao Yu, Baolin Peng, Ruize Xu +6

Large language models (LLMs) have achieved strong performance in language-centric tasks. However, in agentic settings, LLMs often struggle to anticipate action consequences and ada…

math.AP2020

Populations with individual variation in dispersal in heterogeneous environments: dynamics and competition with simply diffusing populations

Robert Stephen Cantrell, Chris Cosner, Xiao Yu

We consider a model for a population in a heterogeneous environment, with logistic type local population dynamics, under the assumption that individuals can switch between two diff…

cs.CV2024

UVEB: A Large-scale Benchmark and Baseline Towards Real-World Underwater Video Enhancement

Yaofeng Xie, Lingwei Kong, Kai Chen +4

Learning-based underwater image enhancement (UIE) methods have made great progress. However, the lack of large-scale and high-quality paired training samples has become the main bo…

cs.RO2026

DAG-STL: A Hierarchical Framework for Zero-Shot Trajectory Planning under Signal Temporal Logic Specifications

Ruijia Liu, Ancheng Hou, Xiao Yu +1

Signal Temporal Logic (STL) is a powerful language for specifying temporally structured robotic tasks. Planning executable trajectories under STL constraints remains difficult when…

math.AP2024

Asymptotic spreading of KPP reactive fronts in heterogeneous shifting environments II: Flux-limited solutions

King-Yeung Lam, Gregoire Nadin, Xiao Yu

We consider the spreading dynamics of the Fisher-KPP equation in a shifting environment, by analyzing the limit of the rate function of the solutions. For environments with a weak…

hep-ph2022

Semileptonic decays of doubly charmed baryons with bag model

Chao-Qiang Geng, Chia-Wei Liu, Aowen Zhou +1

We study the semileptonic decays of with the bag model, where = , = , , ),…

cs.CR2019

Heterogeneous Graph Matching Networks

Shen Wang, Zhengzhang Chen, Xiao Yu +7

Information systems have widely been the target of malware attacks. Traditional signature-based malicious program detection algorithms can only detect known malware and are prone t…

cs.IT2024

The impact when neural min-sum variant meets ordered statistics decoding of LDPC codes

Guangwen Li, Xiao Yu

This paper introduces three key initiatives in the pursuit of a hybrid decoding framework characterized by superior decoding performance, high throughput, low complexity, and indep…

hep-ph2025

Potential of discovering in decays

Chao-Qiang Geng, Xiang-Nan Jin, Chia-Wei Liu +2

We propose to search for in the decay , which serves as a tagged and reconstructible source of , providing an experimentally clean en…

eess.IV2025

TC-KANRecon: High-Quality and Accelerated MRI Reconstruction via Adaptive KAN Mechanisms and Intelligent Feature Scaling

Ruiquan Ge, Xiao Yu, Yifei Chen +10

Magnetic Resonance Imaging (MRI) has become essential in clinical diagnosis due to its high resolution and multiple contrast mechanisms. However, the relatively long acquisition ti…

cs.CL2024

LIONs: An Empirically Optimized Approach to Align Language Models

Xiao Yu, Qingyang Wu, Yu Li +1

Alignment is a crucial step to enhance the instruction-following and conversational abilities of language models. Despite many recent work proposing new algorithms, datasets, and t…

cs.LG2025

REINA: Regularized Entropy Information-Based Loss for Efficient Simultaneous Speech Translation

Nameer Hirschkind, Joseph Liu, Xiao Yu +1

Simultaneous Speech Translation (SimulST) systems stream in audio while simultaneously emitting translated text or speech. Such systems face the significant challenge of balancing…

cs.CL2026

Triospect: A Three-Dimensional Framework for Robust Statistical AI-Generated Text Detection Against Diverse Attacks

Guangsheng Bao, Lihua Rong, Yanbin Zhao +3

Existing AI-generated text detectors are vulnerable to attacks that manipulate textual characteristics. In this study, we propose a novel Triospect Detection Framework by using add…

cs.CL2024

LocalRQA: From Generating Data to Locally Training, Testing, and Deploying Retrieval-Augmented QA Systems

Xiao Yu, Yunan Lu, Zhou Yu

Retrieval-augmented question-answering systems combine retrieval techniques with large language models to provide answers that are more accurate and informative. Many existing tool…

math.CA2018

Some estimates for the bilinear fractional integrals on the Morrey space

Xiao Yu, Xiangxing Tao, Huihui Zhang +1

In this paper, we are interested in the following bilinear fractional integral operator defined by \[ B\mathcal{I}_α({f,g})(x)=\int_{% %TCIMACRO{\U{211d} }% %Beg…

eess.IV2025

CSF-Net: Cross-Modal Spatiotemporal Fusion Network for Pulmonary Nodule Malignancy Predicting

Yin Shen, Zhaojie Fang, Ke Zhuang +8

Pulmonary nodules are an early sign of lung cancer, and detecting them early is vital for improving patient survival rates. Most current methods use only single Computed Tomography…

cs.SE2024

Practitioners' Expectations on Log Anomaly Detection

Xiaoxue Ma, Yishu Li, Jacky Keung +5

Log anomaly detection has become a common practice for software engineers to analyze software system behavior. Despite significant research efforts in log anomaly detection over th…

cs.CR2021

SIGL: Securing Software Installations Through Deep Graph Learning

Xueyuan Han, Xiao Yu, Thomas Pasquier +5

Many users implicitly assume that software can only be exploited after it is installed. However, recent supply-chain attacks demonstrate that application integrity must be ensured…

cs.LG2024

Diffusion Synthesizer for Efficient Multilingual Speech to Speech Translation

Nameer Hirschkind, Xiao Yu, Mahesh Kumar Nandwana +9

We introduce DiffuseST, a low-latency, direct speech-to-speech translation system capable of preserving the input speaker's voice zero-shot while translating from multiple source l…

hep-ph2022

Semileptonic decays of doubly charmed baryons with mixing

Chao-Qiang Geng, Xian-Nan Jin, Chia-Wei Liu +2

We study the mixing effects in the semileptonic decays the doubly charm baryons of . We focus on the ratio of ${\cal R}(θ_c) \equiv {\cal B}( Ξ_{cc} \to Ξ…

cs.LG2022

Improving Model Training via Self-learned Label Representations

Xiao Yu, Nakul Verma

Modern neural network architectures have shown remarkable success in several large-scale classification and prediction tasks. Part of the success of these architectures is their fl…

cs.SE2021

A Multi-Modal Transformer-based Code Summarization Approach for Smart Contracts

Zhen Yang, Jacky Keung, Xiao Yu +4

Code comment has been an important part of computer programs, greatly facilitating the understanding and maintenance of source code. However, high-quality code comments are often u…

cs.CV2017

Learning to Generate Posters of Scientific Papers by Probabilistic Graphical Models

Yu-ting Qiang, Yanwei Fu, Xiao Yu +3

Researchers often summarize their work in the form of scientific posters. Posters provide a coherent and efficient way to convey core ideas expressed in scientific papers. Generati…

math.AP2014

Propagation Phenomena for A Reaction-Advection-Diffusion Competition Model in A Periodic Habitat

Xiao Yu, Xiao-Qiang Zhao

This paper is devoted to the study of propagation phenomena for a Lotka-Volterra reaction-advection-diffusion competition model in a periodic habitat. We first investigate the glob…

cs.SE2024

Fight Fire with Fire: How Much Can We Trust ChatGPT on Source Code-Related Tasks?

Xiao Yu, Lei Liu, Xing Hu +3

With the increasing utilization of large language models such as ChatGPT during software development, it has become crucial to verify the quality of code content it generates. Rece…

cs.CV2025

PreFM: Online Audio-Visual Event Parsing via Predictive Future Modeling

Xiao Yu, Yan Fang, Xiaojie Jin +2

Audio-visual event parsing plays a crucial role in understanding multimodal video content, but existing methods typically rely on offline processing of entire videos with huge mode…

cs.CL2024

DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection

Xiao Yu, Yuang Qi, Kejiang Chen +6

Large language models (LLMs) have the potential to generate texts that pose risks of misuse, such as plagiarism, planting fake reviews on e-commerce platforms, or creating inflamma…

cs.CV2025

ReasonCD: A Multimodal Reasoning Large Model for Implicit Change-of-Interest Semantic Mining

Zhenyang Huang, Xiao Yu, Yi Zhang +2

Remote sensing image change detection is one of the fundamental tasks in remote sensing intelligent interpretation. Its core objective is to identify changes within change regions…

eess.SY2024

Formal Synthesis of Controllers for Safety-Critical Autonomous Systems: Developments and Challenges

Xiang Yin, Bingzhao Gao, Xiao Yu

In recent years, formal methods have been extensively used in the design of autonomous systems. By employing mathematically rigorous techniques, formal methods can provide fully au…

cs.CL2024

Distantly-Supervised Joint Extraction with Noise-Robust Learning

Yufei Li, Xiao Yu, Yanghong Guo +3

Joint entity and relation extraction is a process that identifies entity pairs and their relations using a single model. We focus on the problem of joint extraction in distantly-la…

cs.RO2020

Identification of Challenging Highway-Scenarios for the Safety Validation of Automated Vehicles Based on Real Driving Data

Thomas Ponn, Matthias Breitfuß, Xiao Yu +1

For a successful market launch of automated vehicles (AVs), proof of their safety is essential. Due to the open parameter space, an infinite number of traffic situations can occur,…

cs.AI2025

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report

Shanghai AI Lab, :, Xiaoyang Chen +35

To understand and identify the unprecedented risks posed by rapidly advancing artificial intelligence (AI) models, this report presents a comprehensive assessment of their frontier…

cs.CL2023

Prompt-Based Monte-Carlo Tree Search for Goal-Oriented Dialogue Policy Planning

Xiao Yu, Maximillian Chen, Zhou Yu

Planning for goal-oriented dialogue often requires simulating future dialogue interactions and estimating task progress. Many approaches thus consider training neural networks to p…

cs.SE2022

Diverse Title Generation for Stack Overflow Posts with Multiple Sampling Enhanced Transformer

Fengji Zhang, Jin Liu, Yao Wan +3

Stack Overflow is one of the most popular programming communities where developers can seek help for their encountered problems. Nevertheless, if inexperienced developers fail to d…

cs.CL2025

Dyna-Mind: Learning to Simulate from Experience for Better AI Agents

Xiao Yu, Baolin Peng, Michel Galley +6

Reasoning models have recently shown remarkable progress in domains such as math and coding. However, their expert-level abilities in math and coding contrast sharply with their pe…

cs.CL2023

Controllable Mixed-Initiative Dialogue Generation through Prompting

Maximillian Chen, Xiao Yu, Weiyan Shi +2

Mixed-initiative dialogue tasks involve repeated exchanges of information and conversational control. Conversational agents gain control by generating responses that follow particu…

cs.SE2025

R2ComSync: Improving Code-Comment Synchronization with In-Context Learning and Reranking

Zhen Yang, Hongyi Lin, Xiao Yu +5

Code-Comment Synchronization (CCS) aims to synchronize the comments with code changes in an automated fashion, thereby significantly reducing the workload of developers during soft…

math.AP2026

Invasion Fronts in Shifting Habitats and Competition Systems: A Hamilton-Jacobi Approach and Nonlocal Effects

King-Yeung Lam, Chang-Hong Wu, Xiao Yu

We review recent developments in the study of spreading phenomena in reaction--diffusion equations arising from ecological invasion models. Motivated by the conjecture of Shigesada…

cs.RO2026

Signal Temporal Logic Motion Planning via Graphs of Convex Sets

Yu Chen, Ancheng Hou, Mingyang Feng +2

This paper investigates continuous-time motion planning under Signal Temporal Logic (STL) specifications. The goal is to generate smooth robot trajectories that satisfy high-level…

cs.CV2025

IPSeg: Image Posterior Mitigates Semantic Drift in Class-Incremental Segmentation

Xiao Yu, Yan Fang, Yao Zhao +1

Class incremental learning aims to enable models to learn from sequential, non-stationary data streams across different tasks without catastrophic forgetting. In class incremental…

cs.CL2025

ConFit v2: Improving Resume-Job Matching using Hypothetical Resume Embedding and Runner-Up Hard-Negative Mining

Xiao Yu, Ruize Xu, Chengyuan Xue +3

A reliable resume-job matching system helps a company recommend suitable candidates from a pool of resumes and helps a job seeker find relevant jobs from a list of job posts. Howev…

cs.CL2023

Uncertainty-Aware Bootstrap Learning for Joint Extraction on Distantly-Supervised Data

Yufei Li, Xiao Yu, Yanchi Liu +2

Jointly extracting entity pairs and their relations is challenging when working on distantly-supervised data with ambiguous or noisy labels. To mitigate such impact, we propose unc…

cs.AI2025

Dyna-Think: Synergizing Reasoning, Acting, and World Model Simulation in AI Agents

Xiao Yu, Baolin Peng, Ruize Xu +5

Recent progress in reasoning with large language models (LLMs), such as DeepSeek-R1, demonstrates impressive capabilities in domains like mathematics and coding, by exhibiting comp…

cs.SE2024

On the Influence of Data Resampling for Deep Learning-Based Log Anomaly Detection: Insights and Recommendations

Xiaoxue Ma, Huiqi Zou, Pinjia He +4

Numerous Deep Learning (DL)-based approaches have gained attention in software Log Anomaly Detection (LAD), yet class imbalance in training data remains a challenge, with anomalies…

cs.AI2026

Orchard: An Open-Source Agentic Modeling Framework

Baolin Peng, Wenlin Yao, Qianhui Wu +11

Agentic modeling aims to transform LLMs into autonomous agents capable of solving complex tasks through planning, reasoning, tool use, and multi-turn interaction with external envi…

cs.DC2025

MegaScale-Infer: Serving Mixture-of-Experts at Scale with Disaggregated Expert Parallelism

Ruidong Zhu, Ziheng Jiang, Chao Jin +17

Mixture-of-Experts (MoE) showcases tremendous potential to scale large language models (LLMs) with enhanced performance and reduced computational complexity. However, its sparsely…

cs.CV2026

VideoWorld 2: Learning Transferable Knowledge from Real-world Videos

Zhongwei Ren, Yunchao Wei, Xiao Yu +5

Learning transferable knowledge from unlabeled video data and applying it in new environments is a fundamental capability of intelligent agents. This work presents VideoWorld 2, wh…

hep-ph2024

Categorizing representations of scalar mesons by decays

Chao-Qiang Geng, Chia-Wei Liu, Xiao Yu +1

The scalar mesons are established for a long time, but their nature is still an open question. In this paper, we investigate the potential of categorizing their represent…

cs.IT2022

A recipe of training neural network-based LDPC decoders

Guangwen Li, Xiao Yu

It is known belief propagation decoding variants of LDPC codes can be unrolled easily as neural networks after assigning differed weights to message passing edges flexibly. In this…

cs.LG2026

Diagnosing Training Inference Mismatch in LLM Reinforcement Learning

Tianle Zhong, Neiwen Ling, Yifan Pi +5

Modern LLM RL systems separate rollout generation from policy optimization. These two stages are expected to produce token probabilities that match exactly. However, implementation…

cs.CL2023

FastKASSIM: A Fast Tree Kernel-Based Syntactic Similarity Metric

Maximillian Chen, Caitlyn Chen, Xiao Yu +1

Syntax is a fundamental component of language, yet few metrics have been employed to capture syntactic similarity or coherence at the utterance- and document-level. The existing st…

cs.AI2026

OpenForgeRL: Train Harness-native Agents in Any Environment

Xiao Yu, Baolin Peng, Ruize Xu +7

Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While power…

cs.SE2025

AI Agents for Web Testing: A Case Study in the Wild

Naimeng Ye, Xiao Yu, Ruize Xu +2

Automated web testing plays a critical role in ensuring high-quality user experiences and delivering business value. Traditional approaches primarily focus on code coverage and loa…

cs.AI2026

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

Mengqi Yuan, Zilong Zhou, Xinzhuang Xiong +33

The paper presents OSWorld 2.0, a benchmark consisting of 108 long‑horizon, real‑world computer‑use workflows designed to evaluate how well AI agents can handle complex, multi‑step…

#computer-use benchmarking#long-horizon tasks#interactive agents#tool-use evaluation
cs.SE2025

Don't Use a Cannon to Kill a Fly: Lightweight Model Editing for LLMs to Correct Deprecated API Recommendations

Guancheng Lin, Xiao Yu, Jacky Keung +3

Pre-trained or fine-tuned on large code corpora, Large Language Models (LLMs) have demonstrated strong performance in code completion tasks. However, their embedded knowledge is co…

cs.CL2024

Teaching Language Models to Self-Improve through Interactive Demonstrations

Xiao Yu, Baolin Peng, Michel Galley +2

The self-improving ability of large language models (LLMs), enabled by prompting them to analyze and revise their own outputs, has garnered significant interest in recent research.…

cs.IT2025

Boosting Ordered Statistics Decoding of Short LDPC Codes with Simple Neural Network Models

Guangwen Li, Xiao Yu

Ordered statistics decoding has been instrumental in addressing decoding failures that persist after normalized min-sum decoding in short low-density parity-check codes. Despite it…

cs.IT2024

Deep learning based enhancement of ordered statistics decoding of short LDPC codes

Guangwen Li, Xiao Yu

In the search for highly efficient decoders for short LDPC codes approaching maximum likelihood performance, a relayed decoding strategy, specifically activating the ordered statis…

cs.SE2025

Where Is Self-admitted Code Generated by Large Language Models on GitHub?

Xiao Yu, Lei Liu, Xing Hu +2

The increasing use of Large Language Models (LLMs) in software development has garnered significant attention from researchers evaluating the capabilities and limitations of LLMs f…

cs.CV2024

LPUWF-LDM: Enhanced Latent Diffusion Model for Precise Late-phase UWF-FA Generation on Limited Dataset

Zhaojie Fang, Xiao Yu, Guanyu Zhou +10

Ultra-Wide-Field Fluorescein Angiography (UWF-FA) enables precise identification of ocular diseases using sodium fluorescein, which can be potentially harmful. Existing research ha…

cs.SE2024

Data Preparation for Deep Learning based Code Smell Detection: A Systematic Literature Review

Fengji Zhang, Zexian Zhang, Jacky Wai Keung +4

Code Smell Detection (CSD) plays a crucial role in improving software quality and maintainability. And Deep Learning (DL) techniques have emerged as a promising approach for CSD du…

cs.IT2025

Effective Application of Normalized Min-Sum Decoding for Short BCH Codes

Guangwen Li, Xiao Yu

This paper introduces an enhanced normalized min-sum decoder designed to address the performance and complexity challenges associated with developing parallelizable decoders for sh…

cs.CL2025

ExACT: Teaching AI Agents to Explore with Reflective-MCTS and Exploratory Learning

Xiao Yu, Baolin Peng, Vineeth Vajipey +4

Autonomous agents have demonstrated significant potential in automating complex multistep decision-making tasks. However, even state-of-the-art vision-language models (VLMs), such…

cs.CL2022

Improving Stack Overflow question title generation with copying enhanced CodeBERT model and bi-modal information

Fengji Zhang, Xiao Yu, Jacky Keung +5

Context: Stack Overflow is very helpful for software developers who are seeking answers to programming problems. Previous studies have shown that a growing number of questions are…

cs.IT2026

Neural-Model-Augmented Hybrid NMS-OSD Decoders for Near-ML in Short Block Codes

Guangwen Li, Xiao Yu

This paper presents a hybrid decoding architecture that serially couples a normalized min-sum (NMS) decoder with reinforced ordered statistics decoding (OSD) to achieve near-maximu…

cs.LG2019

Adaptive Transfer Learning of Multi-View Time Series Classification

Donglin Zhan, Shiyu Yi, Dongli Xu +6

Time Series Classification (TSC) has been an important and challenging task in data mining, especially on multivariate time series and multi-view time series data sets. Meanwhile,…

cs.LG2026

Regularized Entropy Information Adaptation with Temporal-Awareness Networks for Simultaneous Speech Translation

Joseph Liu, Nameer Hirschkind, Xiao Yu +1

Simultaneous Speech Translation (SimulST) requires balancing high translation quality with low latency. Recent work introduced REINA, a method that trains a Read/Write policy based…

cs.LG2024

Memory-Reduced Meta-Learning with Guaranteed Convergence

Honglin Yang, Ji Ma, Xiao Yu

The optimization-based meta-learning approach is gaining increased traction because of its unique ability to quickly adapt to a new task using only small amounts of data. However,…

cs.CL2023

KRLS: Improving End-to-End Response Generation in Task Oriented Dialog with Reinforced Keywords Learning

Xiao Yu, Qingyang Wu, Kun Qian +1

In task-oriented dialogs (TOD), reinforcement learning (RL) algorithms train a model to directly optimize response for task-related metrics. However, RL needs to perform exploratio…

math.CA2020

Commutators of weighted Hardy operator on weighted -central Morrey space

Huihui Zhang, Yan Lin, Xiao Yu

In this paper, the authors prove the boundedness of commutators generated by the weighted Hardy operator on weighted -central Morrey space with the weight satisfying the d…

cs.CV2025

HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Fengji Zhang, Linquan Wu, Huiyu Bai +6

Understanding and reasoning over diagrams is a fundamental aspect of human intelligence. While Large Multimodal Models (LMMs) have demonstrated impressive capabilities across vario…

cs.RO2025

Zero-Shot Trajectory Planning for Signal Temporal Logic Tasks

Ruijia Liu, Ancheng Hou, Xiao Yu +1

Signal Temporal Logic (STL) is a powerful specification language for describing complex temporal behaviors of continuous signals, making it well-suited for high-level robotic task…

physics.optics2025

Dispersion-Aware Modeling Framework for Parallel Optical Computing

Ziqi Wei, Yuanjian Wan, Yuhu Cheng +2

Optical computing represents a groundbreaking technology that leverages the unique properties of photons, with innate parallelism standing as its most compelling advantage. Paralle…

cs.AI2025

Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning

Si Chen, Xiao Yu, Ninareh Mehrabi +3

The exploitation of large language models (LLMs) for malicious purposes poses significant security risks as these models become more powerful and widespread. While most existing re…

cs.CL2026

ConFit v3: Improving Resume-Job Matching with LLM-based Re-Ranking

Xiao Yu, Ruize Xu, Chengyuan Xue +6

A reliable resume-job matching system helps a company find suitable candidates from a pool of resumes and helps a job seeker find relevant jobs from a list of job posts. While rece…

cs.CV2025

DGSAN: Dual-Graph Spatiotemporal Attention Network for Pulmonary Nodule Malignancy Prediction

Xiao Yu, Zhaojie Fang, Guanyu Zhou +7

Lung cancer continues to be the leading cause of cancer-related deaths globally. Early detection and diagnosis of pulmonary nodules are essential for improving patient survival rat…

hep-ph2024

Hidden strangeness in meson weak decays to baryon pair

Chao-Qiang Geng, Xiang-Nan Jin, Chia-Wei Liu +1

Our study focuses on the weak decay of , which is the only possible two-body baryonic decay in the meson system. An analysis using perturbative qu…

cs.IR2025

Solving the Content Gap in Roblox Game Recommendations: LLM-Based Profile Generation and Reranking

Chen Wang, Xiaokai Wei, Yexi Jiang +7

With the vast and dynamic user-generated content on Roblox, creating effective game recommendations requires a deep understanding of game content. Traditional recommendation models…