3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.RO2026
FibVLA: An Efficient Temporal Vision-Language-Action Model with Fibonacci Sampling
Li Lin, Wujun Xu, Weiwei Meng +3
Vision-language-action models (VLAs), which leverage the cognition of multimodal information to infer physical-world actions, provide a generalized solution for embodied AI applica…
cs.SI2024
Information Cascade Prediction under Public Emergencies: A Survey
Qi Zhang, Guang Wang, Li Lin +2
With the advent of the era of big data, massive information, expert experience, and high-accuracy models bring great opportunities to the information cascade prediction of public e…
cs.CV2023★ 3 cited
Unlock the Power: Competitive Distillation for Multi-Modal Large Language Models
Xinwei Li, Li Lin, Shuai Wang +1
Recently, multi-modal content generation has attracted lots of attention from researchers by investigating the utilization of visual instruction tuning based on large language mode…