43 citations · 43 across the 2 of their papers we have counts for
2 papers
cs.MM2026
FillGauss: Fine-Grained Filling-Aware Impact Sound Generation for 3D Gaussian Splatting
Chen Yang, Ganye Wen, Bin Huang +4
Synthesizing physically plausible impact sounds from visual observations remains a great challenge in multi-modal AI. Existing 3D-aware audio generation methods primarily model the…
cs.CV2023★ 43 cited
FoodSAM: Any Food Segmentation
Xing Lan, Jiayi Lyu, Hanyu Jiang +4
In this paper, we explore the zero-shot capability of the Segment Anything Model (SAM) for food image segmentation. To address the lack of class-specific information in SAM-generat…