2 papers
cs.RO2026
Object Pose and Shape Estimation for Grasping: Does it Work?
Pavan Karke, Kushal Shah, Gaurav Singh +3
The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models (e.g., InstantMesh, Zero123,…
cs.CV2024
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
Ansh Shah, K Madhava Krishna
Recovering metric depth from a single image remains a fundamental challenge in computer vision, requiring both scene understanding and accurate scaling. While deep learning has adv…