7 papers
DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization
Mengyi Deng, Zhiwei Li, Xin Li +4
Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consistency while preserving reason…
Selection of the Best Policy under Fairness Constraints for Subpopulations
Tingyu Zhu, Yuhang Wu, Zeyu Zheng
Many high-stakes decisions in health care, public policy, and clinical development require committing to a single policy that will be applied uniformly across a heterogeneous popul…
Spiral RoPE: Rotate Your Rotary Positional Embeddings in the 2D Plane
Haoyu Liu, Sucheng Ren, Tingyu Zhu +5
Rotary Position Embedding (RoPE) is the de facto positional encoding in large language models due to its ability to encode relative positions and support length extrapolation. When…
IDA-Bench: Evaluating LLMs on Interactive Guided Data Analysis
Hanyu Li, Haoyu Liu, Tingyu Zhu +4
Large Language Models (LLMs) show promise as data analysis agents, but existing benchmarks overlook the iterative nature of the field, where experts' decisions evolve with deeper i…
Efficient Fine-Grained Guidance for Diffusion Model Based Symbolic Music Generation
Tingyu Zhu, Haoyu Liu, Ziyu Wang +2
Developing generative models to create or conditionally create symbolic music presents unique challenges due to the combination of limited data availability and the need for high p…
Structured Diffusion Models with Mixture of Gaussians as Prior Distribution
Nanshan Jia, Tingyu Zhu, Haoyu Liu +1
We propose a class of structured diffusion models, in which the prior distribution is chosen as a mixture of Gaussians, rather than a standard Gaussian distribution. The specific m…