4 papers
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation
Sankalp Sinha, Mohammad Sadil Khan, Muhammad Usama +4
Generating high-fidelity 3D content from text prompts remains a significant challenge in computer vision due to the limited size, diversity, and annotation depth of the existing da…
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
Muhammad Saif Ullah Khan, Sankalp Sinha, Didier Stricker +2
Reconstructing texture-less surfaces poses unique challenges in computer vision, primarily due to the lack of specialized datasets that cater to the nuanced needs of depth and norm…
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
Mohammad Sadil Khan, Sankalp Sinha, Talha Uddin Sheikh +3
Prototyping complex computer-aided design (CAD) models in modern softwares can be very time-consuming. This is due to the lack of intelligent systems that can quickly generate simp…
CICA: Content-Injected Contrastive Alignment for Zero-Shot Document Image Classification
Sankalp Sinha, Muhammad Saif Ullah Khan, Talha Uddin Sheikh +2
Zero-shot learning has been extensively investigated in the broader field of visual recognition, attracting significant interest recently. However, the current work on zero-shot le…