-
Vehicle tire (tyre) detection and text recognition using deep learning
This paper presents an industrial system to read text on tire sidewalls. Images of vehicle tires in motion are acquired using roadside cameras. Firstly, the tire circularity is detected using Circular Hough Transform (CHT) with dynamic radius detection. The tire…
-
Self-supervised Monocular Depth Estimation: Let’s Talk About The Weather
Current, self-supervised depth estimation architectures rely on clear and sunny weather scenes to train deep neural networks. However, in many locations, this assumption is too strong. For example in the UK (2021), 149 days consisted of rain. For these architectures…
-
Neural Caption Generation for News Images
Automatic caption generation of images has gained significant interest. It gives rise to a lot of interesting image-related applications. For example, it could help in image/video retrieval and management of vast amount of multimedia data available on the Internet. It…
-
Probabilistic visibility for multi-view stereo
We present a new formulation to multi-view stereo that treats the problem as probabilistic 3D segmentation. Previous work has used the stereo photo-consistency criterion as a detector of the boundary between the 3D scene and the surrounding empty space. Here…
-
Multi-Agent Deep Reinforcement Learning for Traffic optimization through Multiple Road Intersections using Live Camera Feed
Traffic signals provide one of the primary means to administer conflicting traffic flows. Existing signal control strategies, operating on hand-crafted rules, fail to efficiently, autonomously adapt to the changing traffic patterns. Each signal control system independently manages one intersection at…
-
Using frontier points to recover shape, reflectance and illumination
We describe a method to recover the surface reflectance and the 3D shape of a non-Lambertian object as well as illumination, from a collection of images. It is based on the so-called frontier points, which are extracted from the outlines…
-
Variational Recurrent Sequence-to-Sequence Retrieval for Stepwise Illustration
-
QuiltGAN: An Adversarially Trained, Procedural Algorithm for Texture Generation
-
How to Read Paintings: Semantic Art Understanding with Multi-Modal Retrieval
Automatic art analysis has been mostly focused on classifying artworks intondifferent artistic styles. However, understanding an artistic representationninvolves more complex processes, such as identifying the elements in the scenenor recognizing author influences. We present SemArt, a multi-modal dataset fornsemantic art…
-
Dyna-DM: Dynamic Object-aware Self-supervised Monocular Depth Maps
Self-supervised monocular depth estimation has been a subject of intense study in recent years, because of its applications in robotics and autonomous driving. Much of the recent work focuses on improving depth estimation by increasing architecture complexity. This paper shows…
