2 papers
cs.AI2026
ReDraft, Don't Just Distill: Reference-Driven Revision for Continual VLLM Post-Training
Zhihao Zhang, Mingqi Wu, Qiaole Dong +15
Continual post-training of large multimodal models should add new capabilities while preserving those from pre-training, and the two goals pull in opposite directions. SFT gives ex…
cs.AI2026
NovGauge: A Fine-Grained Benchmark for Diagnosing LLMs' Capability in Paper Novelty Assessment
Guoqiang Zhang, Kexin Tan, Ming Zhang +12
Large language models (LLMs) are increasingly used in peer review at major AI conferences, yet novelty remains a persistent weak point. Existing benchmarks assess novelty as a sing…