explore-exploit 1large language models 1mathematical reasoning 1prompt sampling 1reinforcement learning 1
From the 1 of 16 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Beyond Random: Automatic Inner-loop Optimization in Dataset Distillation
Muquan Li, Hang Gou, Dongyang Zhang +4
The growing demand for efficient deep learning has positioned dataset distillation as a pivotal technique for compressing training dataset while preserving model performance. Howev…
cs.CV2025
Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models
Shuang Liang, Zhihao Xu, Jialing Tao +2
Despite extensive alignment efforts, Large Vision-Language Models (LVLMs) remain vulnerable to jailbreak attacks, posing serious safety risks. To address this, existing detection m…