From the 1 of 13 linked papers with an AI index.
13 papers
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
Jiahao Huo, Wenjie Qu, Yibo Yan +5
The paper introduces SAMark, a text watermarking method that remains detectable even after paragraph‑level paraphrasing by removing reliance on sentence order and using a hyperboli…
Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning
Xin Zou, Haolin Deng, Yibo Yan +6
Inductive biases steer learning toward generalizable solutions by encoding task structure. In this work, we identify a crucial missing bias in MLLMs: cross-view consistency, \texti…
ARC-STAR: Auditable Post-Hoc Correction for PDE Foundation Models
Chengze Li, Lingwei Wei, Li Sun +7
Partial differential equation (PDE) foundation models are pretrained networks that forecast how physical fields like velocity and pressure evolve from a single reusable solver. On…
Towards Robust LLM Post-Training: Automatic Failure Management for Reinforcement Fine-Tuning
Lingzhe Zhang, Tong Jia, Yunpeng Zhai +6
Reinforcement fine-tuning (RFT) has become a core paradigm for post-training large language models, yet its training process remains highly fragile. Existing efforts mainly improve…
CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification
Hanrong Zhang, Shicheng Fan, Henry Peng Zou +12
Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool is a single, self-contained func…
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
Kening Zheng, Wei-Chieh Huang, Jiahao Huo +9
Mixture-of-Experts (MoE) models exhibit striking performance disparities across languages, yet the internal mechanisms driving these gaps remain poorly understood. In this work, we…