Human brain functions, such as the internal workings of neurons, have historically inspired the design of neural networks. Recently, researchers have increased network parameters to enhance performance and emulate the brain’s complex connectivity, with models like Megatron-Turing NLG reaching 530 billion parameters. These large models, though powerful, require high-end hardware, making them impractical for resource-limited devices. Inspired by glial cells, which create, maintain, and destroy synapses based on their performance, this paper introduces a reinforcement learning agent to optimize neural network structures by adding or pruning nodes in the dense layers of a multi-layer perceptron based on specific reward functions that account for their effectiveness. Experiments on the Fashion-MNIST and CIFAR-10 datasets demonstrate that the RL agent can reduce model parameters by up to 80.95% without losing accuracy. Drawing from neuroscience, this method explores the potential to create efficient, high-performing models suitable for various hardware platforms without a loss of generality.

Glia Cell Inspired Reinforcement Learning Agent for Neural Network Optimization

Cascio, Marco
2025-01-01

Abstract

Human brain functions, such as the internal workings of neurons, have historically inspired the design of neural networks. Recently, researchers have increased network parameters to enhance performance and emulate the brain’s complex connectivity, with models like Megatron-Turing NLG reaching 530 billion parameters. These large models, though powerful, require high-end hardware, making them impractical for resource-limited devices. Inspired by glial cells, which create, maintain, and destroy synapses based on their performance, this paper introduces a reinforcement learning agent to optimize neural network structures by adding or pruning nodes in the dense layers of a multi-layer perceptron based on specific reward functions that account for their effectiveness. Experiments on the Fashion-MNIST and CIFAR-10 datasets demonstrate that the RL agent can reduce model parameters by up to 80.95% without losing accuracy. Drawing from neuroscience, this method explores the potential to create efficient, high-performing models suitable for various hardware platforms without a loss of generality.
2025
Inglese
Inglese
Lecture Notes in Computer Science (LNCS)
European Conference on Computer Vision (ECCV) 2024 Workshops
15636
183
193
11
9783031915772
Springer
Cham
SVIZZERA
Esperti anonimi
29 Settembre - 04 Ottobre, 2024
Milano, Italia
Internazionale
Cross-Dataset Evaluation, Glia Cells, Neural Network Optimization, Reinforcement Learning
5
none
Fagioli, Alessio; Cinque, Luigi; Distante, Damiano; Foresti, Gian Luca; Cascio, Marco
273
info:eu-repo/semantics/conferenceObject
4 Contributo in Atti di Convegno (Proceeding)::4.1 Contributo in Atti di convegno
File in questo prodotto:
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/20.500.14085/56061
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 0
  • ???jsp.display-item.citation.isi??? 0
social impact