7 papers
Implicit-Scale 3D Reconstruction for Multi-Food Volume Estimation from Monocular Images
Yuhao Chen, Gautham Vinod, Siddeshwar Raghavan +5
We present Implicit-Scale 3D Reconstruction from Monocular Multi-Food Images, a benchmark dataset designed to advance geometry-based food portion estimation in realistic dining sce…
Food Portion Estimation: From Pixels to Calories
Gautham Vinod, Fengqing Zhu
Reliance on images for dietary assessment is an important strategy to accurately and conveniently monitor an individual's health, making it a vital mechanism in the prevention and…
Leveraging Second-Order Curvature for Efficient Learned Image Compression: Theory and Empirical Evidence
Yichi Zhang, Fengqing Zhu
Training learned image compression (LIC) models entails navigating a challenging optimization landscape defined by the fundamental trade-off between rate and distortion. Standard f…
Training-Free Text-to-Image Compositional Food Generation via Prompt Grafting
Xinyue Pan, Yuhao Chen, Fengqing Zhu
Real-world meal images often contain multiple food items, making reliable compositional food image generation important for applications such as image-based dietary assessment, whe…
Instance camera focus prediction for crystal agglomeration classification
Xiaoyu Ji, Chenhao Zhang, Tyler James Downard +3
Agglomeration refers to the process of crystal clustering due to interparticle forces. Crystal agglomeration analysis from microscopic images is challenging due to the inherent lim…
Unsupervised Defect Detection for Surgical Instruments
Joseph Huang, Yichi Zhang, Jingxi Yu +6
Ensuring the safety of surgical instruments requires reliable detection of visual defects. However, manual inspection is prone to error, and existing automated defect detection met…