Search CORE

4 research outputs found

Bayesian segnet: Model uncertainty in deep convolutional encoder-decoder architectures for scene understanding

Author: Badrinarayanan V
Cipolla R
Kendall A
Publication venue: British Machine Vision Conference 2017, BMVC 2017
Publication date: 10/10/2016
Field of study

We present a deep learning framework for probabilistic pixel-wise semantic segmentation, which we term Bayesian SegNet. Semantic segmentation is an important tool for visual scene understanding and a meaningful measure of uncertainty is essential for decision making. Our contribution is a practical system which is able to predict pixel-wise class labels with a measure of model uncertainty. We achieve this by Monte Carlo sampling with dropout at test time to generate a posterior distribution of pixel class labels. In addition, we show that modelling uncertainty improves segmentation performance by 2-3% across a number of state of the art architectures such as SegNet, FCN and Dilation Network, with no additional parametrisation. We also observe a significant improvement in performance for smaller datasets where modelling uncertainty is more effective. We benchmark Bayesian SegNet on the indoor SUN Scene Understanding and outdoor CamVid driving scenes datasets.Toyota Corporatio

arXiv.org e-Print Archive

Crossref

Apollo (Cambridge)

Bayesian segnet: Model uncertainty in deep convolutional encoder-decoder architectures for scene understanding

Author: Badrinarayanan V
Cipolla R
Kendall A
Publication venue
Publication date: 01/07/2017
Field of study

© 2017. The copyright of this document resides with its authors. We present a deep learning framework for probabilistic pixel-wise semantic segmentation, which we term Bayesian SegNet. Semantic segmentation is an important tool for visual scene understanding and a meaningful measure of uncertainty is essential for decision making. Our contribution is a practical system which is able to predict pixel-wise class labels with a measure of model uncertainty using Bayesian deep learning. We achieve this by Monte Carlo sampling with dropout at test time to generate a posterior distribution of pixel class labels. In addition, we show that modelling uncertainty improves segmentation performance by 2-3% across a number of datasets and architectures such as SegNet, FCN, Dilation Network and DenseNet

CUED - Cambridge University Engineering Department

SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation

Author: Badrinarayanan V
Cipolla R
Kendall A
Publication venue
Publication date: 06/01/2016
Field of study

We present a novel and practical deep fully convolutional neural network architecture for semantic pixel-wise segmentation termed SegNet. This core trainable segmentation engine consists of an encoder network, a corresponding decoder network followed by a pixel-wise classification layer. The architecture of the encoder network is topologically identical to the 13 convolutional layers in the VGG16 network [1]. The role of the decoder network is to map the low resolution encoder feature maps to full input resolution feature maps for pixel-wise classification. The novelty of SegNet lies is in the manner in which the decoder upsamples its lower resolution input feature map(s). Specifically, the decoder uses pooling indices computed in the max-pooling step of the corresponding encoder to perform non-linear upsampling. This eliminates the need for learning to upsample. The upsampled maps are sparse and are then convolved with trainable filters to produce dense feature maps. We compare our proposed architecture with the widely adopted FCN [2] and also with the well known DeepLab-LargeFOV [3] , DeconvNet [4] architectures. This comparison reveals the memory versus accuracy trade-off involved in achieving good segmentation performance. SegNet was primarily motivated by scene understanding applications. Hence, it is designed to be efficient both in terms of memory and computational time during inference. It is also significantly smaller in the number of trainable parameters than other competing architectures and can be trained end-to-end using stochastic gradient descent. We also performed a controlled benchmark of SegNet and other architectures on both road scenes and SUN RGB-D indoor scene segmentation tasks. These quantitative assessments show that SegNet provides good performance with competitive inference time and most efficient inference memory-wise as compared to other architectures. We also provide a Caffe implementation of SegNet and a web demo at http://mi.eng.cam.ac.uk/projects/segnet/

CiteSeerX

CUED - Cambridge University Engineering Department

Deep learning for actinic keratosis classification

Author: Badrinarayanan V Kendall A, Cipolla R
Bookstein FL
Fagerland MW Lydersen S, Laake P
Guo Z Zhang L, Zhang D
Hames SC Sinnya S, Tan JM, et al.
Nanni L Paci M, Brahnam S, et al.
Nosaka R Fukui K
Pietik&auml
Ren S He K, Girshick RB, et al.
Rigel DS Gold LFS
Russakovsky O Deng J, Su H, et al.
Song T Li H, Meng F, et al.
Spyridonos P Gaitanis G, Likas A, et al.
Spyridonos P Gaitanis G, Likas A, et al.
Tan X Triggs W
Wassef C Rao BK
Zhu Z You X, Chen CLP, et al.
Publication venue: 'American Institute of Mathematical Sciences (AIMS)'
Publication date: 01/01/2020
Field of study

Crossref