In this video we demonstrate how to use Deep Learning Toolkit of speech recognition problem. We employ the waveform to spectrogram conversion techniques to represent 1-dimensional audio waveforms into 2-dimensional spectrograms to feed the converted dataset into convolutional neural networks for object classification.