Deep Learning for Computer Vision: Backward Propagation (UPC 2016)

•

2 gefällt mir•1,504 views

http://imatge-upc.github.io/telecombcn-2016-dlcv/ Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of big annotated data and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which had been addressed until now with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks and Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles and applications of deep learning to computer vision problems, such as image classification, object detection or text captioning.

Ingenieurwesen

Day 1 Lecture 4
Backward Propagation
Elisa Sayrol
[course site]

Learning
Purely Supervised
Typically Backpropagation + Stochastic Gradient Descent (SGD)
Good when there are lots of labeled data
Layer-wise Unsupervised + Supervised classifier
Train each layer in sequence, using regularized auto-encoders or Restricted Boltzmann
Machines (RBM)
Hold the feature extractor, on top train linear classifier on features
Good when labeled data is scarce but there are lots of unlabeled data
Layer-wise Unsupervised + Supervised Backprop
Train each layer in sequence
Backprop through the whole system
Good when learning problem is very difficult
Slide Credit: Lecun 2

From Lecture 3
L Hidden Layers
Hidden pre-activation (k>0)
Hidden activation (k=1,…L)
Output activation (k=L+1)
Figure Credit: Hugo Laroche NN course 3

Backpropagation algorithm
The output of the Network gives class scores that depens on the input
and the parameters
• Define a loss function that quantifies our unhappiness with the
scores across the training data.
• Come up with a way of efficiently finding the parameters that
minimize the loss function (optimization)
4

Probability Class given an input
(softmax)
Minimize the loss (plus some
regularization term) w.r.t. Parameters
over the whole training set.
Loss function; e.g., negative log-
likelihood (good for classification)
h2
h3
a3
a4 h4
Loss
Hidden Hidden Output
W2
W3
x a2
Input
W1
Regularization term (L2 Norm)
aka as weight decay
Figure Credit: Kevin McGuiness
Forward Pass
5

Backpropagation algorithm
• We need a way to fit the model to data: find parameters (W(k)
, b(k)
) of the
network that (locally) minimize the loss function.
• We can use stochastic gradient descent. Or better yet, mini-batch
stochastic gradient descent.
• To do this, we need to find the gradient of the loss function with respect to
all the parameters of the model (W(k)
, b(k)
)
• These can be found using the chain rule of differentiation.
• The calculations reveal that the gradient wrt. the parameters in layer k only
depends on the error from the above layer and the output from the layer
below.
• This means that the gradients for each layer can be computed iteratively,
starting at the last layer and propagating the error back through the network.
This is known as the backpropagation algorithm.
Slide Credit: Kevin McGuiness 6

1. Find the error in the top layer: 3. Backpropagate error to layer below2. Compute weight updates
h2
h3
a3
a4 h4
Loss
Hidden Hidden Output
W2
W3
x a2
Input
W1
L
Figure Credit: Kevin McGuiness
Backward Pass
7

Optimization
Stochastic Gradient Descent
Stochastic Gradient Descent with momentum
Stochastic Gradient Descent with L2 regularization
http://cs231n.github.io/optimization-1/
http://cs231n.github.io/optimization-2/
: learning rate
: weight decay
Recommended lectures:
8

Weitere ähnliche Inhalte

Was ist angesagt?

https://github.com/telecombcn-dl/dlmm-2017-dcu Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of big annotated data and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which had been addressed until now with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks and Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles and applications of deep learning to computer vision problems, such as image classification, object detection or text captioning.

Optimizing Deep Networks (D1L6 Insight@DCU Machine Learning Workshop 2017)

Universitat Politècnica de Catalunya

Joint unsupervised learning of deep representations and image clusters

Universitat Politècnica de Catalunya

D1L5 Visualization (D1L2 Insight@DCU Machine Learning Workshop 2017)

Universitat Politècnica de Catalunya

Training Deep Networks with Backprop (D1L4 Insight@DCU Machine Learning Works...

Universitat Politècnica de Catalunya

https://telecombcn-dl.github.io/dlmm-2017-dcu/ Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of big annotated data and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which had been addressed until now with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks and Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles and applications of deep learning to computer vision problems, such as image classification, object detection or image captioning.

Object Segmentation (D2L7 Insight@DCU Machine Learning Workshop 2017)

Universitat Politècnica de Catalunya

Deep Learning for Computer Vision: Unsupervised Learning (UPC 2016)

Universitat Politècnica de Catalunya

Generative Models and Adversarial Training (D2L3 Insight@DCU Machine Learning...

Universitat Politècnica de Catalunya

Deep Learning for Computer Vision: Transfer Learning and Domain Adaptation (U...

Universitat Politècnica de Catalunya

Transfer Learning (D2L4 Insight@DCU Machine Learning Workshop 2017)

Universitat Politècnica de Catalunya

https://telecombcn-dl.github.io/2017-dlcv/ Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of large-scale annotated datasets and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which were previously addressed with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks and Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles and applications of deep learning to computer vision problems, such as image classification, object detection or image captioning.

Optimization for Deep Networks (D2L1 2017 UPC Deep Learning for Computer Vision)

Universitat Politècnica de Catalunya

https://telecombcn-dl.github.io/2018-dlcv/ Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of large-scale annotated datasets and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which were previously addressed with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks and Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles and applications of deep learning to computer vision problems, such as image classification, object detection or image captioning.

Semantic Segmentation - Míriam Bellver - UPC Barcelona 2018

Universitat Politècnica de Catalunya

The first part of this dissertation focuses on an analysis of the spatial context in semantic image segmentation. First, we review how spatial context has been tackled in the literature by local features and spatial aggregation techniques. From a discussion about whether the context is beneficial or not for object recognition, we extend a Figure-Border-Ground segmentation for local feature aggregation with ground truth annotations to a more realistic scenario where object proposals techniques are used instead. Whereas the Figure and Ground regions represent the object and the surround respectively, the Border is a region around the object contour, which is found to be the region with the richest contextual information for object recognition. Furthermore, we propose a new contour-based spatial aggregation technique of the local features within the object region by a division of the region into four subregions. Both contributions have been tested on a semantic segmentation benchmark with a combination of free and non-free context local features that allows the models automatically learn whether the context is beneficial or not for each semantic category. The second part of this dissertation addresses the semantic segmentation for a set of closely-related images from an uncalibrated multiview scenario. State-of-the-art semantic segmentation algorithms fail on correctly segmenting the objects from some viewpoints when the techniques are independently applied to each viewpoint image. The lack of large annotations available for multiview segmentation do not allow to obtain a proper model that is robust to viewpoint changes. In this second part, we exploit the spatial correlation that exists between the dierent viewpoints images to obtain a more robust semantic segmentation. First, we review the state-of-the-art co-clustering, co-segmentation and video segmentation techniques that aim to segment the set of images in a generic way, i.e. without considering semantics. Then, a new architecture that considers motion information and provides a multiresolution segmentation is proposed for the co-clustering framework and outperforms state-of-the-art techniques for generic multiview segmentation. Finally, the proposed multiview segmentation is combined with the semantic segmentation results giving a method for automatic resolution selection and a coherent semantic multiview segmentation.

Visual Object Analysis using Regions and Local Features

Universitat Politècnica de Catalunya

Image Segmentation (D3L1 2017 UPC Deep Learning for Computer Vision)

Universitat Politècnica de Catalunya

https://telecombcn-dl.github.io/2017-dlai/ Deep learning technologies are at the core of the current revolution in artificial intelligence for multimedia data analysis. The convergence of large-scale annotated datasets and affordable GPU hardware has allowed the training of neural networks for data analysis tasks which were previously addressed with hand-crafted features. Architectures such as convolutional neural networks, recurrent neural networks or Q-nets for reinforcement learning have shaped a brand new scenario in signal processing. This course will cover the basic principles of deep learning from both an algorithmic and computational perspectives.

Transfer Learning and Domain Adaptation (DLAI D5L2 2017 UPC Deep Learning for...

Universitat Politècnica de Catalunya

DeconvNet, DecoupledNet, TransferNet in Image Segmentation

NamHyuk Ahn

Deep Generative Models - Kevin McGuinness - UPC Barcelona 2018

Universitat Politècnica de Catalunya

Object detection - RCNNs vs Retinanet

Rishabh Indoria

Unsupervised Learning (D2L6 2017 UPC Deep Learning for Computer Vision)

Universitat Politècnica de Catalunya

Image Classification using deep learning

Asma-AH

Image Classification with Deep Learning | DevFest + GDay, George Town, Mala...

Virot "Ta" Chiraphadhanakul

Was ist angesagt? (20)