9 citations · 25 across the 8 of their papers we have counts for
8 papers · 1 filter
Self-Feedback DETR for Temporal Action Detection
Jihwan Kim, Miso Lee, Jae-Pil Heo
Temporal Action Detection (TAD) is challenging but fundamental for real-world video applications. Recently, DETR-based models have been devised for TAD but have not performed well…
Task-Oriented Channel Attention for Fine-Grained Few-Shot Classification
SuBeen Lee, WonJun Moon, Hyun Seok Seong +1
The difficulty of the fine-grained image classification mainly comes from a shared overall appearance across classes. Thus, recognizing discriminative details, such as eyes and bea…
Leveraging Hidden Positives for Unsupervised Semantic Segmentation
Hyun Seok Seong, WonJun Moon, SuBeen Lee +1
Dramatic demand for manpower to label pixel-level annotations triggered the advent of unsupervised semantic segmentation. Although the recent work employing the vision transformer…
Query-Dependent Video Representation for Moment Retrieval and Highlight Detection
WonJun Moon, Sangeek Hyun, SangUk Park +2
Recently, video moment retrieval and highlight detection (MR/HD) are being spotlighted as the demand for video understanding is drastically increased. The key objective of MR/HD is…
Minority-Oriented Vicinity Expansion with Attentive Aggregation for Video Long-Tailed Recognition
WonJun Moon, Hyun Seok Seong, Jae-Pil Heo
A dramatic increase in real-world video volume with extremely diverse and emerging topics naturally forms a long-tailed video distribution in terms of their categories, and it spot…
Difficulty-Aware Simulator for Open Set Recognition
WonJun Moon, Junho Park, Hyun Seok Seong +2
Open set recognition (OSR) assumes unknown instances appear out of the blue at the inference time. The main challenge of OSR is that the response of models for unknowns is totally…