Hasan F. M. Zaki

article

Viewpoint invariant semantic object and scene categorization with RGB-D sensors

Hasan F. M. Zaki, Faisal Shafait, Ajmal Mian

Autonomous Robots > 2019 > 43 > 4 > 1005-1022

Understanding the semantics of objects and scenes using multi-modal RGB-D sensors serves many robotics applications. Key challenges for accurate RGB-D image recognition are the scarcity of training data, variations due to viewpoint changes and the heterogeneous nature of the data. We address these problems and propose a generic deep learning framework based on a pre-trained convolutional neural network,...

chapter

Modeling Sub-Event Dynamics in First-Person Action Recognition

Hasan F. M. Zaki, Faisal Shafait, Ajmal Mian

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 1619 - 1628

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

First-person videos have unique characteristics such as heavy egocentric motion, strong preceding events, salient transitional activities and post-event impacts. Action recognition methods designed for third person videos may not optimally represent actions captured by first-person videos. We propose a method to represent the high level dynamics of sub-events in first-person videos by dynamically...

chapter

Modeling 2D Appearance Evolution for 3D Object Categorization

Hasan F. M. Zaki, Faisal Shafait, Ajmal Mian

2016 International Conference on Digital Image Computing: Techniques and Applications (DICTA) > 1 - 8

2016 International Conference on Digital Image Computing: Techniques and Applications (DICTA)

3D object categorization is a non-trivial task in computer vision encompassing many real-world applications. We pose the problem of categorizing 3D polygon meshes as learning appearance evolution from multi-view 2D images. Given a corpus of 3D polygon meshes, we first render the corresponding RGB and depth images from multiple viewpoints on a uniform sphere. Using rank pooling, we propose two methods...

chapter

Convolutional hypercube pyramid for accurate RGB-D object category and instance recognition

Hasan F. M. Zaki, Faisal Shafait, Ajmal Mian

2016 IEEE International Conference on Robotics and Automation (ICRA) > 1685 - 1692

2016 IEEE International Conference on Robotics and Automation (ICRA)

Deep learning based methods have achieved unprecedented success in solving several computer vision problems involving RGB images. However, this level of success is yet to be seen on RGB-D images owing to two major challenges in this domain: training data deficiency and multi-modality input dissimilarity. We present an RGB-D object recognition framework that addresses these two key challenges by effectively...

chapter

Localized Deep Extreme Learning Machines for Efficient RGB-D Object Recognition

Hasan F. M. Zaki, Faisal Shafait, Ajmal Mian

2015 International Conference on Digital Image Computing: Techniques and Applications (DICTA) > 1 - 8

2015 International Conference on Digital Image Computing: Techniques and Applications (DICTA)

Existing RGB-D object recognition methods either use channel specific handcrafted features, or learn features with deep networks. The former lack representation ability while the latter require large amounts of training data and learning time. In real-time robotics applications involving RGB-D sensors, we do not have the luxury of both. In this paper, we propose Localized Deep Extreme Learning Machines...

INFONA - science communication portal

Search results for: Hasan F. M. Zaki

Viewpoint invariant semantic object and scene categorization with RGB-D sensors

Modeling Sub-Event Dynamics in First-Person Action Recognition

Modeling 2D Appearance Evolution for 3D Object Categorization

Convolutional hypercube pyramid for accurate RGB-D object category and instance recognition

Localized Deep Extreme Learning Machines for Efficient RGB-D Object Recognition

Filter options

Publication date

Publication type

Keywords

Data set

INFONA - science communication portal

Search results for: Hasan F. M. Zaki

Viewpoint invariant semantic object and scene categorization with RGB-D sensors

Modeling Sub-Event Dynamics in First-Person Action Recognition

Modeling 2D Appearance Evolution for 3D Object Categorization

Convolutional hypercube pyramid for accurate RGB-D object category and instance recognition

Localized Deep Extreme Learning Machines for Efficient RGB-D Object Recognition

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Publication type

Keywords

Data set

Reporting an error / abuse

Sending the report failed

Accessibility options