End-to-end Keyword Spotting using Neural Architecture Search and Quantization

David Peter, Wolfgang Roth, Franz Pernkopf

Publikation: Beitrag in Buch/Bericht/KonferenzbandBeitrag in einem KonferenzbandBegutachtung

Abstract

This paper introduces neural architecture search (NAS) for the automatic discovery of end-to-end keyword spotting (KWS) models for limited resource environments. We employ a differentiable NAS approach to optimize the structure of convolutional neural networks (CNNs) operating on raw audio waveforms. After a suitable KWS model is found with NAS, we conduct quantization of weights and activations to reduce the memory footprint. We conduct extensive experiments on the Google speech commands dataset. In particular, we compare our end-to-end models to mel-frequency cepstral coefficient (MFCC) based CNNs. For quantization, we compare fixed bit-width quantization and trained bit-width quantization. Using NAS only, we were able to obtain a highly efficient model with an accuracy of 95.55% using 75.7k parameters and 13.6M operations. Using trained bit-width quantization, the same model achieves a test accuracy of 93.76% while using on average only 2.91 bits per activation and 2.51 bits per weight
Originalspracheenglisch
Titel2022 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2022 - Proceedings
Herausgeber (Verlag)Institute of Electrical and Electronics Engineers
Seiten3423-3427
Seitenumfang5
ISBN (elektronisch)9781665405409
DOIs
PublikationsstatusVeröffentlicht - 2022
Veranstaltung47th IEEE International Conference on Acoustics, Speech and Signal Processing: ICASSP 2022 - Virtual, Online, Singapur
Dauer: 22 Mai 202227 Mai 2022

Publikationsreihe

NameICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings
Band2022-May
ISSN (Print)1520-6149

Konferenz

Konferenz47th IEEE International Conference on Acoustics, Speech and Signal Processing
KurztitelICASSP 2022
Land/GebietSingapur
OrtVirtual, Online
Zeitraum22/05/2227/05/22

ASJC Scopus subject areas

  • Software
  • Signalverarbeitung
  • Elektrotechnik und Elektronik

Fields of Expertise

  • Information, Communication & Computing

Fingerprint

Untersuchen Sie die Forschungsthemen von „End-to-end Keyword Spotting using Neural Architecture Search and Quantization“. Zusammen bilden sie einen einzigartigen Fingerprint.
  • Intelligent Systems

    Pernkopf, F.

    1/01/02 → …

    Projekt: Arbeitsgebiet

Dieses zitieren