3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.CV2026
ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?
Han Li, Si Liu, Zehao Huang +6
Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental abilities that humans naturall…
cs.CV2025
Depth3DLane: Monocular 3D Lane Detection via Depth Prior Distillation
Dongxin Lyu, Han Huang, Cheng Tan +1
Monocular 3D lane detection is challenging due to the difficulty in capturing depth information from single-camera images. A common strategy involves transforming front-view (FV) i…
cs.CL2024★ 3 cited
Peer Review as A Multi-Turn and Long-Context Dialogue with Role-Based Interactions
Cheng Tan, Dongxin Lyu, Siyuan Li +5
Large Language Models (LLMs) have demonstrated wide-ranging applications across various fields and have shown significant potential in the academic peer-review process. However, ex…