26 citations · 27 across the 6 of their papers we have counts for
6 papers
Audio Match Cutting: Finding and Creating Matching Audio Transitions in Movies and Videos
Dennis Fedorishin, Lie Lu, Srirangaraj Setlur +1
A "match cut" is a common video editing technique where a pair of shots that have a similar composition transition fluidly from one to another. Although match cuts are often visual…
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
Dennis Fedorishin, Livio Forte, Philip Schneider +2
Sound event detection (SED) is an active area of audio research that aims to detect the temporal occurrence of sounds. In this paper, we apply SED to engine fault detection by intr…
RealCQA: Scientific Chart Question Answering as a Test-bed for First-Order Logic
Saleem Ahmed, Bhavin Jawade, Shubham Pandey +2
We present a comprehensive study of chart visual question-answering(QA) task, to address the challenges faced in comprehending and extracting data from chart visualizations within…
SpaDen : Sparse and Dense Keypoint Estimation for Real-World Chart Understanding
Saleem Ahmed, Pengyu Yan, David Doermann +2
We introduce a novel bottom-up approach for the extraction of chart data. Our model utilizes images of charts as inputs and learns to detect keypoints (KP), which are used to recon…
CoNAN: Conditional Neural Aggregation Network For Unconstrained Face Feature Fusion
Bhavin Jawade, Deen Dayal Mohan, Dennis Fedorishin +2
Face recognition from image sets acquired under unregulated and uncontrolled settings, such as at large distances, low resolutions, varying viewpoints, illumination, pose, and atmo…
Is Joint Training Better for Deep Auto-Encoders?
Yingbo Zhou, Devansh Arpit, Ifeoma Nwogu +1
Traditionally, when generative models of data are developed via deep architectures, greedy layer-wise pre-training is employed. In a well-trained model, the lower layer of the arch…