Search results

chapter

Learned Multi-patch Similarity

Wilfried Hartmann, Silvano Galliani, Michal Havlena, Luc Van Gool, more

2017 IEEE International Conference on Computer Vision (ICCV) > 1595 - 1603

2017 IEEE International Conference on Computer Vision (ICCV)

Estimating a depth map from multiple views of a scene is a fundamental task in computer vision. As soon as more than two viewpoints are available, one faces the very basic question how to measure similarity across >2 image patches. Surprisingly, no direct solution exists, instead it is common to fall back to more or less robust averaging of two-view similarities. Encouraged by the success of machine...

chapter

A vision based traffic light detection and recognition approach for intelligent vehicles

Ziya Ozcelik, Canan Tastimur, Mehmet Karakose, Erhan Akin

2017 International Conference on Computer Science and Engineering (UBMK) > 424 - 429

2017 International Conference on Computer Science and Engineering (UBMK)

The quality of life of people is increasing together with the developing technologies. One of the most important factors affecting daily life is smart cities. The quality of life of people is positively affected by emerging this concept in recent years. Autonomous vehicles confront with the term of the smart city and have become even more popular in recent years. In this study, a system of traffic...

chapter

Maximum correntropy criterion for convex anc semi-nonnegative matrix factorization

Anyong Qin, Zhaowei Shang, Jinyu Tian, Ailin Li, more

2017 IEEE International Conference on Systems, Man, and Cybernetics (SMC) > 1856 - 1861

2017 IEEE International Conference on Systems, Man and Cybernetics (SMC)

Matrix factorization is a popular low dimensional representation approach that plays an important role in many pattern recognition and computer vision domains. Among them, convex and semi-nonnegative matrix factorizations have attracted considerable interest, owing to its clustering interpretation. On the other hand, the generalized correlation function (correntropy) as the error measure does not...

chapter

Factorized Bilinear Models for Image Recognition

Yanghao Li, Naiyan Wang, Jiaying Liu, Xiaodi Hou

2017 IEEE International Conference on Computer Vision (ICCV) > 2098 - 2106

2017 IEEE International Conference on Computer Vision (ICCV)

Although Deep Convolutional Neural Networks (CNNs) have liberated their power in various computer vision tasks, the most important components of CNN, convolutional layers and fully connected layers, are still limited to linear transformations. In this paper, we propose a novel Factorized Bilinear (FB) layer to model the pairwise feature interactions by considering the quadratic terms in the transformations...

chapter

Deep Determinantal Point Process for Large-Scale Multi-label Classification

Pengtao Xie, Ruslan Salakhutdinov, Luntian Mou, Eric P. Xing

2017 IEEE International Conference on Computer Vision (ICCV) > 473 - 482

2017 IEEE International Conference on Computer Vision (ICCV)

We study large-scale multi-label classification (MLC) on two recently released datasets: Youtube-8M and Open Images that contain millions of data instances and thousands of classes. The unprecedented problem scale poses great challenges for MLC. First, finding out the correct label subset out of exponentially many choices incurs substantial ambiguity and uncertainty. Second, the large data-size and...

chapter

Occlusion detector using convolutional neural network for person re-identification

Sejeong Lee, Yoojin Hong, Moongu Jeon

2017 International Conference on Control, Automation and Information Sciences (ICCAIS) > 140 - 144

2017 International Conference on Control, Automation and Information Sciences (ICCAIS)

Technique of comparing pedestrian images observed by different cameras to determine whether they are the same person is important in the surveillance system. This technique is called Person re-identification. Most of Person reidentification is underway assuming that occlusion does not occur. However, since occlusion occurs frequently in the surveillance system and affects accuracy, it is necessary...

chapter

Understanding convolutional neural networks using a minimal model for handwritten digit recognition

Matthew Y. W. Teow

2017 IEEE 2nd International Conference on Automatic Control and Intelligent Systems (I2CACIS) > 167 - 172

2017 IEEE 2nd International Conference on Automatic Control and Intelligent Systems (I2CACIS)

The contribution of this paper is to bridge the gap on understanding the mathematical structure and the computational implementation of a convolutional neural network (CNN) using a minimal model (Minimal CNN). The proposed minimal CNN is presented using a layering approach. This approach provides a concise and accessible understanding of the main mathematical operations of a CNN. Hence, it benefits...

chapter

Non-linear Convolution Filters for CNN-Based Learning

Georgios Zoumpourlis, Alexandros Doumanoglou, Nicholas Vretos, Petros Daras

2017 IEEE International Conference on Computer Vision (ICCV) > 4771 - 4779

2017 IEEE International Conference on Computer Vision (ICCV)

During the last years, Convolutional Neural Networks (CNNs) have achieved state-of-the-art performance in image classification. Their architectures have largely drawn inspiration by models of the primate visual system. However, while recent research results of neuroscience prove the existence of non-linear operations in the response of complex visual cells, little effort has been devoted to extend...

chapter

Scale-Adaptive Convolutions for Scene Parsing

Rui Zhang, Sheng Tang, Yongdong Zhang, Jintao Li, more

2017 IEEE International Conference on Computer Vision (ICCV) > 2050 - 2058

2017 IEEE International Conference on Computer Vision (ICCV)

Many existing scene parsing methods adopt Convolutional Neural Networks with fixed-size receptive fields, which frequently result in inconsistent predictions of large objects and invisibility of small objects. To tackle this issue, we propose a scale-adaptive convolution to acquire flexiblesize receptive fields during scene parsing. Through adding a new scale regression layer, we can dynamically infer...

chapter

Preprocessing of barley grain images for defect identification

Marcin Kociolek, Piotr M. Szczypinski, Artur Klepaczko

2017 Signal Processing: Algorithms, Architectures, Arrangements, and Applications (SPA) > 365 - 370

2017 Signal Processing: Algorithms, Architectures, Arrangements, and Applications (SPA)

A malt is one of intermediate ingredients for a brewing industry. The quality of barley used for malting have essential impact on the final product flavor. An automatic system for a barley grains inspection, utilizing computer vision methods, can provide an objective quality assessment. We present image preprocessing steps of grain inspection system. Main preprocessing steps are: segmentation of grain...

chapter

Design of a real-time pedestrian detection system for autonomous vehicles

R Harshitha, J Manikandan

2017 IEEE Region 10 Symposium (TENSYMP) > 1 - 4

2017 IEEE Region 10 Symposium (TENSYMP)

Pedestrian detection is considered as an active area of research and the advent of autonomous vehicles for a smarter mobility has spearheaded the research in this field. In this paper, design of a real-time pedestrian detection system for autonomous vehicles is proposed and its performance is evaluated using images from standard datasets as well as realtime video input. The proposed system is designed...

chapter

Learning Spatial Regularization with Image-Level Supervisions for Multi-label Image Classification

Feng Zhu, Hongsheng Li, Wanli Ouyang, Nenghai Yu, more

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 2027 - 2036

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

Multi-label image classification is a fundamental but challenging task in computer vision. Great progress has been achieved by exploiting semantic relations between labels in recent years. However, conventional approaches are unable to model the underlying spatial relations between labels in multi-label images, because spatial annotations of the labels are generally not provided. In this paper, we...

chapter

Conditional Similarity Networks

Andreas Veit, Serge Belongie, Theofanis Karaletsos

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 1781 - 1789

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

What makes images similar? To measure the similarity between images, they are typically embedded in a feature-vector space, in which their distance preserve the relative dissimilarity. However, when learning such similarity embeddings the simplifying assumption is commonly made that images are only compared to one unique measure of similarity. A main reason for this is that contradicting notions of...

chapter

Soft-Margin Mixture of Regressions

Dong Huang, Longfei Han, Fernando De la Torre

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 4058 - 4066

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

Nonlinear regression is a common statistical tool to solve many computer vision problems (e.g., age estimation, pose estimation). Existing approaches to nonlinear regression fall into two main categories: (1) The universal approach provides an implicit or explicit homogeneous feature mapping (e.g., kernel ridge regression, Gaussian process regression, neural networks). These approaches may fail when...

chapter

Learning Deep Match Kernels for Image-Set Classification

Haoliang Sun, Xiantong Zhen, Yuanjie Zheng, Gongping Yang, more

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 6240 - 6249

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

Image-set classification has recently generated great popularity due to its widespread applications in computer vision. The great challenges arise from effectively and efficiently measuring the similarity between image sets with high inter-class ambiguity and huge intra-class variability. In this paper, we propose deep match kernels (DMK) to directly measure the similarity between image sets in the...

chapter

Spatio-Temporal Self-Organizing Map Deep Network for Dynamic Object Detection from Videos

Yang Du, Chunfeng Yuan, Bing Li, Weiming Hu, more

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) > 4245 - 4254

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

In dynamic object detection, it is challenging to construct an effective model to sufficiently characterize the spatial-temporal properties of the background. This paper proposes a new Spatio-Temporal Self-Organizing Map (STSOM) deep network to detect dynamic objects in complex scenarios. The proposed approach has several contributions: First, a novel STSOM shared by all pixels in a video frame is...

chapter

When Kernel Methods Meet Feature Learning: Log-Covariance Network for Action Recognition From Skeletal Data

Jacopo Cavazza, Pietro Morerio, Vittorio Murino

2017 IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) > 1251 - 1258

2017 IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)

Human action recognition from skeletal data is a hot research topic and important in many open domain applications of computer vision, thanks to recently introduced 3D sensors. In the literature, naive methods simply transfer off-the-shelf techniques from video to the skeletal representation. However, the current state-of-the-art is contended between to different paradigms: kernel-based methods and...

chapter

The human detection in images using the depth map

Dmitriy Tatarenkov, Dmitry Podolsky

2017 Systems of Signal Synchronization, Generating and Processing in Telecommunications (SINKHROINFO) > 1 - 4

2017 Systems of Signal Synchronization, Generating and Processing in Telecommunications (SINKHROINFO)

In today world the necessity for the autonomous mobile robots and vehicles is increasing. The safety autonomous moving demands the reliable and fast detection algorithms. The Histogram of Oriented Gradients (HOG) descriptors show significantly outperforms the existing feature sets for a human detection. Though the given method has a lot of type I errors. The amount of these errors can be decreased...

chapter

Action recognition using three dimension convolution and long short term memory

Yu-Cheng Liu, Jian-Jiun Ding, Yao-Jen Chang, Chien-Yao Wang, more

2017 IEEE International Conference on Consumer Electronics - Taiwan (ICCE-TW) > 83 - 84

2017 IEEE International Conference on Consumer Electronics - Taiwan (ICCE-TW)

The convolutional neural network (CNN) is more and more popular in computer vision and widely used in acoustic signal processing, image classification, and image segmentation. In this work, an architecture which is a combination of the 3-D convolutional neural network and the long short term memory (LSTM) was proposed for action recognition. It stacks the consecutive video frames, extracts spatial...

chapter

Feature extraction with convolutional neural networks for aerial image retrieval

Hakan Cevikalp, Golara Ghorban Dordinejad, Merve Elmas

2017 25th Signal Processing and Communications Applications Conference (SIU) > 1 - 4

2017 25th Signal Processing and Communications Applications Conference (SIU)

Deep learning methods have been effectively used to provide great improvement in various research fields such as machine learning, image processing and computer vision. One of the most frequently used deep learning methods in image processing is the convolutional neural networks. Compared to the traditional artificial neural networks, convolutional neural networks do not use the predefined kernels,...

INFONA - science communication portal

Search results

Learned Multi-patch Similarity

A vision based traffic light detection and recognition approach for intelligent vehicles

Maximum correntropy criterion for convex anc semi-nonnegative matrix factorization

Factorized Bilinear Models for Image Recognition

Deep Determinantal Point Process for Large-Scale Multi-label Classification

Occlusion detector using convolutional neural network for person re-identification

Understanding convolutional neural networks using a minimal model for handwritten digit recognition

Non-linear Convolution Filters for CNN-Based Learning

Scale-Adaptive Convolutions for Scene Parsing

Preprocessing of barley grain images for defect identification

Design of a real-time pedestrian detection system for autonomous vehicles

Learning Spatial Regularization with Image-Level Supervisions for Multi-label Image Classification

Conditional Similarity Networks

Soft-Margin Mixture of Regressions

Learning Deep Match Kernels for Image-Set Classification

Spatio-Temporal Self-Organizing Map Deep Network for Dynamic Object Detection from Videos

When Kernel Methods Meet Feature Learning: Log-Covariance Network for Action Recognition From Skeletal Data

The human detection in images using the depth map

Action recognition using three dimension convolution and long short term memory

Feature extraction with convolutional neural networks for aerial image retrieval

Filter options

Publication date

Content availability

Keywords

INFONA - science communication portal

Search results

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Content availability

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options