activity
20232026
most citedKimi K2.5: Visual Agentic Intelligence

2 citations · 2 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CL2026

Kimi K3: Open Frontier Intelligence

Kimi Team, Tongtong Bai, Yifan Bai +398

We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is…

cs.CV2026

KPM-Bench: A Kinematic Parsing Motion Benchmark for Fine-grained Motion-centric Video Understanding

Boda Lin, Yongjie Zhu, Xiaocheng Gong +2

Despite recent advancements, video captioning models still face significant limitations in accurately describing fine-grained motion details and suffer from severe hallucination is…

cs.CL20262 cited

Kimi K2.5: Visual Agentic Intelligence

Kimi Team, Tongtong Bai, Yifan Bai +333

We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint optimization of text and vision so that…

cs.MM2024

TOP:A New Target-Audience Oriented Content Paraphrase Task

Boda Lin, Jiaxin Shi, Haolong Yan +3

Recommendation systems usually recommend the existing contents to different users. However, in comparison to static recommendation methods, a recommendation logic that dynamically…

cs.CL2023

ChatGPT is a Potential Zero-Shot Dependency Parser

Boda Lin, Xinyi Zhou, Binghao Tang +2

Pre-trained language models have been widely used in dependency parsing task and have achieved significant improvements in parser performance. However, it remains an understudied q…

cs.CL2023

Type-aware Decoding via Explicitly Aggregating Event Information for Document-level Event Extraction

Gang Zhao, Yidong Shi, Shudong Lu +5

Document-level event extraction (DEE) faces two main challenges: arguments-scattering and multi-event. Although previous methods attempt to address these challenges, they overlook…