late interaction 1multi-vector retrieval 1object-aware compression 1token merging 1vision-language retrieval 1
From the 1 of 19 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack
Yein Park, Jungwoo Park, Jaewoo Kang
Large language models (LLMs), despite being safety-aligned, exhibit brittle refusal behaviors that can be circumvented by simple linguistic changes. As tense jailbreaking demonstra…
cs.AI2026
Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training
Yein Park, Minbyul Jeong, Jaewoo Kang
The remarkable capabilities of modern large reasoning models are largely unlocked through post-training techniques such as supervised fine-tuning (SFT) and reinforcement learning (…
cs.AI2025
Monet: Mixture of Monosemantic Experts for Transformers
Jungwoo Park, Young Jin Ahn, Kee-Eung Kim +1
Understanding the internal computations of large language models (LLMs) is crucial for aligning them with human values and preventing undesirable behaviors like toxic content gener…