3 papers
cs.RO2026
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
Shengyu Si, Yuanzhuo Lu, Ruimeng Yang +3
Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to unseen tasks that necessitate tra…
cs.CR2025
Strong but Brittle: Simple Attacks Subvert Reasoning-based Safety Guardrails
Shuo Chen, Zhen Han, Haokun Chen +6
Open-weight Large Reasoning Models (LRMs) are approaching the capabilities of their frontier counterparts but pose significant safety concerns, as they are difficult to patch or mo…
cs.CL2025
Think Before Refusal : Triggering Safety Reflection in LLMs to Mitigate False Refusal Behavior
Shengyun Si, Xinpeng Wang, Guangyao Zhai +2
Recent advancements in large language models (LLMs) have demonstrated that fine-tuning and human alignment can render LLMs harmless. In practice, such "harmlessness" behavior is ma…