For a given audio dataset, can we do audio classification using Spectrogram? well, let's try it out ourselves and let's use Google AutoML Vision to fail fast :D
We'll be converting our audio files into their respective spectrograms and use spectrogram as images for our classification problem.
Here is the formal definition of the Spectrogram
A Spectrogram is a visual representation of the spectrum of frequencies of a signal as it varies with time.
For this experiment, I'm going to use the following audio dataset from Kaggle
/assets/images/1085/original/6c326146-d22d-4264-919b-a1eaf62ec167.png?1409142750)