Showing 2025Show all
2 papers · 1 filter
cs.LG2025
Earley-Driven Dynamic Pruning for Efficient Structured Decoding
Xintong Sun, Chi Wei, Minghao Tian +1
Large Language Models (LLMs) have shown remarkable capabilities, yet ensuring their outputs conform to strict structural or grammatical constraints remains challenging, which is cr…
cs.CL2025
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
Sunbowen Lee, Shiwen Ni, Chi Wei +7
Safety alignment mechanism are essential for preventing large language models (LLMs) from generating harmful information or unethical content. However, cleverly crafted prompts can…