Showing 2026Show all
3 papers · 1 filter
cs.CL2026
VeriBound: PAC-Bayesian Generalization Bounds for Process Reward Models Trained with Formal Verification Tools
Amirul Rahman, Mohammed Sabih Alsharari
Process Reward Models (PRMs) provide step-level verification for Large Language Model (LLM) reasoning, yet their training data acquisition remains a bottleneck: human annotation is…
cs.CV2026
UniMark: Unified Adaptive Multi-bit Watermarking for Autoregressive Image Generators
Yigit Yilmaz, Elena Petrova, Mehmet Kaya +2
Invisible watermarking for autoregressive (AR) image generation has recently gained attention as a means of protecting image ownership and tracing AI-generated content. However, ex…
cs.CL2026
Avoiding Overthinking and Underthinking: Curriculum-Aware Budget Scheduling for LLMs
Amirul Rahman, Aisha Karim, Kenji Nakamura +1
Scaling test-time compute via extended reasoning has become a key paradigm for improving the capabilities of large language models (LLMs). However, existing approaches optimize rea…