QS4D: Quantization-aware training for efficient hardware deployment of structured state-space sequential models

Siegel, Sebastian; Fabre, Maxime; Yang, Ming-Jay; Strachan, John Paul; Bouhadjar, Younes; Neftci, Emre

doi:10.48550/ARXIV.2507.06079

Preprint

FZJ-2026-00222

QS4D: Quantization-aware training for efficient hardware deployment of structured state-space sequential models

Siegel, S. (Corresponding author)FZJ* ; Yang, M.-J.FZJ* ; Bouhadjar, Y.FZJ* ; Fabre, M.FZJ* ; Neftci, E.FZJ* ; Strachan, J. P.FZJ*

2025
arXiv

arXiv (2025) [10.48550/ARXIV.2507.06079]

This record in other databases:

Please use a persistent id in citations: doi:10.48550/ARXIV.2507.06079 doi:10.34734/FZJ-2026-00222

Abstract: Structured State Space models (SSM) have recently emerged as a new class of deep learning models, particularly well-suited for processing long sequences. Their constant memory footprint, in contrast to the linearly scaling memory demands of Transformers, makes them attractive candidates for deployment on resource-constrained edge-computing devices. While recent works have explored the effect of quantization-aware training (QAT) on SSMs, they typically do not address its implications for specialized edge hardware, for example, analog in-memory computing (AIMC) chips. In this work, we demonstrate that QAT can significantly reduce the complexity of SSMs by up to two orders of magnitude across various performance metrics. We analyze the relation between model size and numerical precision, and show that QAT enhances robustness to analog noise and enables structural pruning. Finally, we integrate these techniques to deploy SSMs on a memristive analog in-memory computing substrate and highlight the resulting benefits in terms of computational efficiency.

Keyword(s): Machine Learning (cs.LG) ; Artificial Intelligence (cs.AI) ; FOS: Computer and information sciences

Contributing Institute(s):

Research Program(s):

Appears in the scientific report 2025

Database coverage:
OpenAccess

Click to display QR Code for this record

The record appears in these collections:
Dokumenttypen > Berichte > Vorabdrucke
Institutssammlungen > PGI > PGI-15
Institutssammlungen > PGI > PGI-14
Workflowsammlungen > Öffentliche Einträge
Publikationsdatenbank
Open Access

Datensatz erzeugt am 2026-01-12, letzte Änderung am 2026-02-20

Ähnliche Datensätze

OpenAccess:

PDF

Dieses Dokument bewerten:

(Bisher nicht rezensiert)

Zum persönlichen Korb hinzufügen
Export als Author List with IDs BibTeX (UTF-8), EndNote XML, EndNote Text, RIS, MARC, Print MARC, MARCXML, DC,
Request correction
Submit fulltext

Gast :: Anmelden JuSER
		Suchen		Absenden		Personalisieren Ihre Benachrichtigungen Ihre Körbe Ihre Suchanfragen		Hilfe