collaborators

10 papers

cs.AI2026

ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams

Ali Ansari, Yasmin Mohammadi, Farnoush Nili +3

Entity-Relationship Diagrams (ERDs) are central to conceptual database design, yet they are typically available only as rendered images rather than machine-readable schemas, limiti…

cs.CL2026

Scaling Performance and Low-Resource Annotation with Many-Shot In-Context Learning for Named Entity Recognition

Qi Zhang, Fangping Lan, Cornelia Caragea +2

In-context learning (ICL) with large language models (LLMs) has emerged as a powerful alternative to fine-tuning for Named Entity Recognition (NER), achieving strong performance wi…

cs.CV2026

Enhancing Renal Tumor Malignancy Prediction: Deep Learning with Automatic 3D CT Organ Focused Attention

Zhengkang Fan, Chengkun Sun, Russell Terry +2

Accurate prediction of malignancy in renal tumors is crucial for informing clinical decisions and optimizing treatment strategies. However, existing imaging modalities lack the nec…

cs.CV2026

Logit Lens Supervision for Patch-Level Explanations in Vision-Language Models

Parsa Esmaeilkhani, Longin Jan Latecki

Modern autoregressive Vision-Language Models (VLMs) can generate fluent answers while their visual-token representations become weakly tied to the image regions from which they ori…

cs.CV2025

Direct Visual Grounding by Directing Attention of Visual Tokens

Parsa Esmaeilkhani, Longin Jan Latecki

Vision Language Models (VLMs) mix visual tokens and text tokens. A puzzling issue is the fact that visual tokens most related to the query receive little to no attention in the fin…

cs.CV2025

Layout Stroke Imitation: A Layout Guided Handwriting Stroke Generation for Style Imitation with Diffusion Model

Sidra Hanif, Longin Jan Latecki

Handwriting stroke generation is crucial for improving the performance of tasks such as handwriting recognition and writers order recovery. In handwriting stroke generation, it is…