2 papers
cs.CV2025
Fine-Grained Captioning of Long Videos through Scene Graph Consolidation
Sanghyeok Chu, Seonguk Seo, Bohyung Han
Recent advances in vision-language models have led to impressive progress in caption generation for images and short video clips. However, these models remain constrained by their…
cs.LG2024
Re-evaluating Group Robustness via Adaptive Class-Specific Scaling
Seonguk Seo, Bohyung Han
Group distributionally robust optimization, which aims to improve robust accuracies -- worst-group and unbiased accuracies -- is a prominent algorithm used to mitigate spurious cor…