3 papers
cs.CV2025
VideoPerceiver: Enhancing Fine-Grained Temporal Perception in Video Multimodal Large Language Models
Fufangchen Zhao, Liao Zhang, Daiqi Shi +5
We propose VideoPerceiver, a novel video multimodal large language model (VMLLM) that enhances fine-grained perception in video understanding, addressing VMLLMs' limited ability to…
math.NA2025
On Convergence of the Secant Method
Yan Tan, Chenhao Ye, Qinghai Zhang +1
The secant method, as an important approach for solving nonlinear equations, is introduced in nearly all numerical analysis textbooks. However, most textbooks only briefly address…
cs.LG2025
Learning without Global Backpropagation via Synergistic Information Distillation
Chenhao Ye, Ming Tang
Backpropagation (BP), while foundational to deep learning, imposes two critical scalability bottlenecks: update locking, where network modules remain idle until the entire backward…