13 papers
Self-Supervised Skill Optimization
Siran Peng, Cuiyu Yang, Tianyu Fu +9
Agent skills provide frozen large language model (LLM) agents with reusable procedural guidance, and recent work shows that such skills can be optimized with ground-truth (GT) feed…
Private Face Recognition Training Dataset Publication via Identity-Decoupled and Geometry-Preserving Face Distillation
Shuhuan Chen, Xiangyu Zhu, Weisong Zhao +6
The paper proposes a method to publish face recognition training datasets that removes personal identity information while keeping the geometric relationships needed for effective…
Kimi K3: Open Frontier Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +398
We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is…
CPG-PAD: Concept-Informed Prompts Guided Presentation Attack Detection
Haoyuan Zhang, Xiangyu Zhu, Li Gao +3
Presentation Attack Detection (PAD) serves as a crucial safeguard for face recognition systems against presentation attacks such as printed photos, replayed videos, and 3D masks. D…
UPA: Unsupervised Prompt Agent via Tree-Based Search and Selection
Siran Peng, Weisong Zhao, Tianyu Fu +6
Prompt agents have recently emerged as a promising paradigm for automated prompt optimization, framing prompt discovery as a sequential decision-making problem over a structured pr…
Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection
Tianshuo Zhang, Haoyuan Zhang, Siran Peng +3
Continual face forgery detection (CFFD) requires detectors to learn emerging forgery paradigms without forgetting previously seen manipulations. Existing CFFD methods commonly rely…