4 papers
From Next-Token to Next-Block: A Principled Adaptation Path for Diffusion LLMs
Yuchuan Tian, Yuchen Liang, Shuo Zhang +10
Diffusion Language Models (DLMs) enable fast generation, yet training large DLMs from scratch is costly. As a practical shortcut, adapting off-the-shelf Auto-Regressive (AR) model…
Sweedler Duality for BiHom-associative Algebras
Jiacheng Sun
Motivated by the fact that ordinary linear duality does not in general produce a coalgebra structure from an infinite-dimensional algebra, we develop a Sweedler-type finite dual co…
Optimizing GPT for Video Understanding: Zero-Shot Performance and Prompt Engineering
Mark Beliaev, Victor Yang, Madhura Raju +2
In this study, we tackle industry challenges in video content classification by exploring and optimizing GPT-based models for zero-shot classification across seven critical categor…
ELFS: Label-Free Coreset Selection with Proxy Training Dynamics
Haizhong Zheng, Elisa Tsai, Yifu Lu +4
High-quality human-annotated data is crucial for modern deep learning pipelines, yet the human annotation process is both costly and time-consuming. Given a constrained human label…