◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

A. Ngom

5 papers hereh-index 212.1k citations156 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author5

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.SE1
  • q-bio.BM1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2025

Representation Learning with Adaptive Superpixel Coding

Mahmoud Khalil, Ahmad Khalil, Alioune Ngom

Deep learning vision models are typically tailored for specific modalities and often rely on domain-specific assumptions, such as the grid structures used by nearly all existing vi…

cs.CV2025

ResNetVLLM -- Multi-modal Vision LLM for the Video Understanding Task

Ahmad Khalil, Mahmoud Khalil, Alioune Ngom

In this paper, we introduce ResNetVLLM (ResNet Vision LLM), a novel cross-modal framework for zero-shot video understanding that integrates a ResNet-based visual encoder with a Lar…

cs.CV2025

ResNetVLLM-2: Addressing ResNetVLLM's Multi-Modal Hallucinations

Ahmad Khalil, Mahmoud Khalil, Alioune Ngom

Large Language Models (LLMs) have transformed natural language processing (NLP) tasks, but they suffer from hallucination, generating plausible yet factually incorrect content. Thi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.