4 papers
Designing Concise ConvNets with Columnar Stages
Ashish Kumar, Jaesik Park
In the era of vision Transformers, the recent success of VanillaNet shows the huge potential of simple and concise convolutional neural networks (ConvNets). Where such models mainl…
InstantDrag: Improving Interactivity in Drag-based Image Editing
Joonghyuk Shin, Daehyeon Choi, Jaesik Park
Drag-based image editing has recently gained popularity for its interactivity and precision. However, despite the ability of text-to-image models to generate samples within a secon…
High-Speed Stereo Visual SLAM for Low-Powered Computing Devices
Ashish Kumar, Jaesik Park, Laxmidhar Behera
We present an accurate and GPU-accelerated Stereo Visual SLAM design called Jetson-SLAM. It exhibits frame-processing rates above 60FPS on NVIDIA's low-powered 10W Jetson-NX embedd…
Cross Resolution Encoding-Decoding For Detection Transformers
Ashish Kumar, Jaesik Park
Detection Transformers (DETR) are renowned object detection pipelines, however computationally efficient multiscale detection using DETR is still challenging. In this paper, we pro…