1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Can Text-to-image Model Assist Multi-modal Learning for Visual Recognition with Visual Modality Missing?
Tiantian Feng, Daniel Yang, Digbalay Bose +1
Multi-modal learning has emerged as an increasingly promising avenue in vision recognition, driving innovations across diverse domains ranging from media and education to healthcar…
cs.CL2023★ 1 cited
Context Unlocks Emotions: Text-based Emotion Classification Dataset Auditing with Large Language Models
Daniel Yang, Aditya Kommineni, Mohammad Alshehri +4
The lack of contextual information in text data can make the annotation process of text-based emotion classification datasets challenging. As a result, such datasets often contain…
cs.AI2023
Does Video Summarization Require Videos? Quantifying the Effectiveness of Language in Video Summarization
Yoonsoo Nam, Adam Lehavi, Daniel Yang +3
Video summarization remains a huge challenge in computer vision due to the size of the input videos to be summarized. We propose an efficient, language-only video summarizer that a…