Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
Binbin Ji, Siddharth Agrawal, Qiance Tang +1
This study investigates the spatial reasoning capabilities of vision-language models (VLMs) through Chain-of-Thought (CoT) prompting and reinforcement learning. We begin by evaluat…
cs.CV2024
Global License Plate Dataset
Siddharth Agrawal
In the pursuit of advancing the state-of-the-art (SOTA) in road safety, traffic monitoring, surveillance, and logistics automation, we introduce the Global License Plate Dataset (G…