Abstract
Edge devices can execute pre-trained Artificial Intelligence (AI) models optimized on large Graphical Processing Units (GPU) but often need fine-tuning for real-world data. This process, known as edge learning, is crucial for personalized learning for tasks such as speech and gesture recognition and often requires recurrent neural networks (RNNs). However, training RNNs on edge devices faces challenges due to limited resources. We propose a system for RNN training through sequence partitioning using the Forward Propagation Through Time (FPTT) training method, facilitating edge learning. Our optimized HW/SW co-design for FPTT is the first of its kind. In our work, we have implemented the complete computational process for training Long Short-Term Memory (LSTM) networks using FPTT, and we have optimized and explored the hardware architecture leveraging the Chipyard framework. Our findings indicate considerable memory savings, with only a slight increase in latency, when training small-batch size sequential MNIST (S-MNIST) data.
| Original language | English |
|---|---|
| Title of host publication | 2024 IFIP/IEEE 32nd International Conference on Very Large Scale Integration, VLSI-SoC 2024 |
| Publisher | Institute of Electrical and Electronics Engineers |
| Number of pages | 4 |
| ISBN (Electronic) | 979-8-3315-3967-2 |
| DOIs | |
| Publication status | Published - 3 Dec 2024 |
| Event | IFIP/IEEE International Conference on Very Large Scale Integration, VLSI-SoC 2024 - Morocco, Tanger, Morocco Duration: 6 Oct 2024 → 9 Oct 2024 https://vlsisoc2024.nl/ |
Conference
| Conference | IFIP/IEEE International Conference on Very Large Scale Integration, VLSI-SoC 2024 |
|---|---|
| Abbreviated title | VLSI-SoC 2024 |
| Country/Territory | Morocco |
| City | Tanger |
| Period | 6/10/24 → 9/10/24 |
| Internet address |
Funding
This work has been funded by the Dutch Organization for Scientific Research (NWO) with Grant KICH1.ST04.22.021.
| Funders | Funder number |
|---|---|
| Netherlands Organisation for Applied Scientific Research | KICH1.ST04.22.021 |
Keywords
- edge-learning
- deep learning
- HW/SW co-design
- HW/SW Co-design
- LSTM
- Edge Learning
Fingerprint
Dive into the research topics of 'A Scalable Hardware Architecture for Efficient Learning of Recurrent Neural Networks at the Edge'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver