papers

Publications (12)

cs.CV2023

Memory-Constrained Semantic Segmentation for Ultra-High Resolution UAV Imagery

Qi Li, Jiaxin Cai, Yuanlong Yu +3

Amidst the swift advancements in photography and sensor technologies, high-definition cameras have become commonplace in the deployment of Unmanned Aerial Vehicles (UAVs) for diver…

cs.LG2020

A Comparative Study on Machine Learning Algorithms for the Control of a Wall Following Robot

Issam Hammad, Kamal El-Sankary, Jason Gu

A comparison of the performance of various machine learning models to predict the direction of a wall following robot is presented in this paper. The models were trained using an o…

cs.CV2019

Edge-Semantic Learning Strategy for Layout Estimation in Indoor Environment

Weidong Zhang, Wei Zhang, Jason Gu

Visual cognition of the indoor environment can benefit from the spatial layout estimation, which is to represent an indoor scene with a 2D box on a monocular image. In this paper,…

cs.RO2022

A Multi-Behavior Planning Framework for Robot Guide

Muhan Hou, Zonghao Mu, Jing Li +2

The guiding task of a mobile robot requires not only human-aware navigation, but also appropriate yet timely interaction for active instruction. State-of-the-art tour-guide models…

cs.LG2020

Deep Learning Training with Simulated Approximate Multipliers

Issam Hammad, Kamal El-Sankary, Jason Gu

This paper presents by simulation how approximate multipliers can be utilized to enhance the training performance of convolutional neural networks (CNNs). Approximate multipliers h…

cs.RO2025

Learning 6-DoF Fine-grained Grasp Detection Based on Part Affordance Grounding

Yaoxian Song, Penglei Sun, Piaopiao Jin +7

Robotic grasping is a fundamental ability for a robot to interact with the environment. Current methods focus on how to obtain a stable and reliable grasping pose in object level,…

cs.CV2024

SX-Stitch: An Efficient VMS-UNet Based Framework for Intraoperative Scoliosis X-Ray Image Stitching

Yi Li, Heting Gao, Mingde He +3

In scoliosis surgery, the limited field of view of the C-arm X-ray machine restricts the surgeons' holistic analysis of spinal structures .This paper presents an end-to-end efficie…

cs.AR2023

Subtractor-Based CNN Inference Accelerator

Victor Gao, Issam Hammad, Kamal El-Sankary +1

This paper presents a novel method to boost the performance of CNN inference accelerators by utilizing subtractors. The proposed CNN preprocessing accelerator relies on sorting, gr…

cs.SD2022

End-To-End Audiovisual Feature Fusion for Active Speaker Detection

Fiseha B. Tesema, Zheyuan Lin, Shiqiang Zhu +3

Active speaker detection plays a vital role in human-machine interaction. Recently, a few end-to-end audiovisual frameworks emerged. However, these models' inference time was not e…

cs.CV2024

Aligning Knowledge Graph with Visual Perception for Object-goal Navigation

Nuo Xu, Wen Wang, Rong Yang +6

Object-goal navigation is a challenging task that requires guiding an agent to specific objects based on first-person visual observations. The ability of agent to comprehend its su…

cs.CV2022

RSI-Net: Two-Stream Deep Neural Network for Remote Sensing Imagesbased Semantic Segmentation

Shuang He, Xia Lu, Jason Gu +6

For semantic segmentation of remote sensing images (RSI), trade-off between representation power and location accuracy is quite important. How to get the trade-off effectively is a…

cs.CV2022

BCOT: A Markerless High-Precision 3D Object Tracking Benchmark

Jiachen Li, Bin Wang, Shiqiang Zhu +6

Template-based 3D object tracking still lacks a high-precision benchmark of real scenes due to the difficulty of annotating the accurate 3D poses of real moving video objects witho…