Process waveforms and spectrograms, extract acoustic features, and build speech recognition, speaker, and audio classification systems.