3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.CV2025
Elevating Visual Question Answering through Implicitly Learned Reasoning Pathways in LVLMs
Liu Jing, Amirul Rahman
Large Vision-Language Models (LVLMs) have shown remarkable progress in various multimodal tasks, yet they often struggle with complex visual reasoning that requires multi-step infe…
cs.CV2024
Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction
Liu Jing, Amirul Rahman
Semantic location prediction from multimodal social media posts is a critical task with applications in personalized services and human mobility analysis. This paper introduces \te…
cs.CL2024★ 3 cited
Fault Diagnosis in Power Grids with Large Language Model
Liu Jing, Amirul Rahman
Power grid fault diagnosis is a critical task for ensuring the reliability and stability of electrical infrastructure. Traditional diagnostic systems often struggle with the comple…