collaborators

7 papers

cs.CV2026

Implicit-Scale 3D Reconstruction for Multi-Food Volume Estimation from Monocular Images

Yuhao Chen, Gautham Vinod, Siddeshwar Raghavan +5

We present Implicit-Scale 3D Reconstruction from Monocular Multi-Food Images, a benchmark dataset designed to advance geometry-based food portion estimation in realistic dining sce…

cs.CV2026

Food Portion Estimation: From Pixels to Calories

Gautham Vinod, Fengqing Zhu

Reliance on images for dietary assessment is an important strategy to accurately and conveniently monitor an individual's health, making it a vital mechanism in the prevention and…

eess.IV2026

Leveraging Second-Order Curvature for Efficient Learned Image Compression: Theory and Empirical Evidence

Yichi Zhang, Fengqing Zhu

Training learned image compression (LIC) models entails navigating a challenging optimization landscape defined by the fundamental trade-off between rate and distortion. Standard f…

cs.CV2026

Training-Free Text-to-Image Compositional Food Generation via Prompt Grafting

Xinyue Pan, Yuhao Chen, Fengqing Zhu

Real-world meal images often contain multiple food items, making reliable compositional food image generation important for applications such as image-based dietary assessment, whe…

cs.CV2026

Instance camera focus prediction for crystal agglomeration classification

Xiaoyu Ji, Chenhao Zhang, Tyler James Downard +3

Agglomeration refers to the process of crystal clustering due to interparticle forces. Crystal agglomeration analysis from microscopic images is challenging due to the inherent lim…

cs.CV2025

Unsupervised Defect Detection for Surgical Instruments

Joseph Huang, Yichi Zhang, Jingxi Yu +6

Ensuring the safety of surgical instruments requires reliable detection of visual defects. However, manual inspection is prone to error, and existing automated defect detection met…