Publications (12)
Memory-Constrained Semantic Segmentation for Ultra-High Resolution UAV Imagery
Qi Li, Jiaxin Cai, Yuanlong Yu +3
Amidst the swift advancements in photography and sensor technologies, high-definition cameras have become commonplace in the deployment of Unmanned Aerial Vehicles (UAVs) for diver…
A Comparative Study on Machine Learning Algorithms for the Control of a Wall Following Robot
Issam Hammad, Kamal El-Sankary, Jason Gu
A comparison of the performance of various machine learning models to predict the direction of a wall following robot is presented in this paper. The models were trained using an o…
Edge-Semantic Learning Strategy for Layout Estimation in Indoor Environment
Weidong Zhang, Wei Zhang, Jason Gu
Visual cognition of the indoor environment can benefit from the spatial layout estimation, which is to represent an indoor scene with a 2D box on a monocular image. In this paper,…
A Multi-Behavior Planning Framework for Robot Guide
Muhan Hou, Zonghao Mu, Jing Li +2
The guiding task of a mobile robot requires not only human-aware navigation, but also appropriate yet timely interaction for active instruction. State-of-the-art tour-guide models…
Deep Learning Training with Simulated Approximate Multipliers
Issam Hammad, Kamal El-Sankary, Jason Gu
This paper presents by simulation how approximate multipliers can be utilized to enhance the training performance of convolutional neural networks (CNNs). Approximate multipliers h…
Learning 6-DoF Fine-grained Grasp Detection Based on Part Affordance Grounding
Yaoxian Song, Penglei Sun, Piaopiao Jin +7
Robotic grasping is a fundamental ability for a robot to interact with the environment. Current methods focus on how to obtain a stable and reliable grasping pose in object level,…
SX-Stitch: An Efficient VMS-UNet Based Framework for Intraoperative Scoliosis X-Ray Image Stitching
Yi Li, Heting Gao, Mingde He +3
In scoliosis surgery, the limited field of view of the C-arm X-ray machine restricts the surgeons' holistic analysis of spinal structures .This paper presents an end-to-end efficie…
Subtractor-Based CNN Inference Accelerator
Victor Gao, Issam Hammad, Kamal El-Sankary +1
This paper presents a novel method to boost the performance of CNN inference accelerators by utilizing subtractors. The proposed CNN preprocessing accelerator relies on sorting, gr…
End-To-End Audiovisual Feature Fusion for Active Speaker Detection
Fiseha B. Tesema, Zheyuan Lin, Shiqiang Zhu +3
Active speaker detection plays a vital role in human-machine interaction. Recently, a few end-to-end audiovisual frameworks emerged. However, these models' inference time was not e…
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
Nuo Xu, Wen Wang, Rong Yang +6
Object-goal navigation is a challenging task that requires guiding an agent to specific objects based on first-person visual observations. The ability of agent to comprehend its su…
RSI-Net: Two-Stream Deep Neural Network for Remote Sensing Imagesbased Semantic Segmentation
Shuang He, Xia Lu, Jason Gu +6
For semantic segmentation of remote sensing images (RSI), trade-off between representation power and location accuracy is quite important. How to get the trade-off effectively is a…
BCOT: A Markerless High-Precision 3D Object Tracking Benchmark
Jiachen Li, Bin Wang, Shiqiang Zhu +6
Template-based 3D object tracking still lacks a high-precision benchmark of real scenes due to the difficulty of annotating the accurate 3D poses of real moving video objects witho…