large language models 1multi-turn jailbreak 1policy optimization 1reinforcement learning 1turn-level credit assignment 1
From the 1 of 9 linked papers with an AI index.
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment
Junyoung Park, Namgyu Park, Sechan Lee +3
The paper proposes a turn-level credit assignment framework (DC‑GRPO) for training multi‑turn jailbreak attacks on large language models, showing higher success rates than prior me…
cs.CL2025
VOCABTRIM: Vocabulary Pruning for Efficient Speculative Decoding in LLMs
Raghavv Goel, Sudhanshu Agrawal, Mukul Gagrani +9
In this paper, we introduce a simple training-free technique to improve the performance of drafter-based speculative decoding (SpD) methods that incorporates language modeling head…
cs.CL2024
On Speculative Decoding for Multimodal Large Language Models
Mukul Gagrani, Raghavv Goel, Wonseok Jeon +3
Inference with Multimodal Large Language Models (MLLMs) is slow due to their large-language-model backbone which suffers from memory bandwidth bottleneck and generates tokens auto-…