3 papers
cs.CV2025
DGM4+: Dataset Extension for Global Scene Inconsistency
Gagandeep Singh, Samudi Amarsinghe, Priyanka Singh +1
The rapid advances in generative models have significantly lowered the barrier to producing convincing multimodal disinformation. Fabricated images and manipulated captions increas…
cs.CV2025
SGS: Segmentation-Guided Scoring for Global Scene Inconsistencies
Gagandeep Singh, Samudi Amarsinghe, Urawee Thani +3
We extend HAMMER, a state-of-the-art model for multimodal manipulation detection, to handle global scene inconsistencies such as foreground-background (FG-BG) mismatch. While HAMME…
cs.LG2024
Enhancing Sign Language Detection through Mediapipe and Convolutional Neural Networks (CNN)
Aditya Raj Verma, Gagandeep Singh, Karnim Meghwal +2
This research combines MediaPipe and CNNs for the efficient and accurate interpretation of ASL dataset for the real-time detection of sign language. The system presented here captu…