In-Network Q-Learning-Based Packet Forwarding for Delay Sensitive Applications

Name: In-Network Q-Learning-Based Packet Forwarding for Delay Sensitive Applications
Keywords: low latency communications, in-network computing, in-band network telemetry, network programmability

descriptionPublicationkeyboard_double_arrow_right Article 01 May 2025Publisher:Institute of Electrical and Electronics Engineers (IEEE)Journal:IEEE Network, volume 39, pages 127-133 (issn: 0890-8044, eissn: 1558-156X,

Authors: Marco Polverini; Antonio Cianfrani; Marco Listanti; Tommaso Caiazzi; Mariano Scazzariello;

doi: 10.1109/mnet.2025.3552929

handle: 11695/147670

In-Network Q-Learning-Based Packet Forwarding for Delay Sensitive Applications

- Summary
- Subjects
- Metrics

Abstract

The use of Artificial Intelligence principles represents the next research challenge to support future network applications in the upcoming 6G era. In this work, we propose a novel approach: exploiting the principles of Reinforcement Learning (RL) and the availability of programmable switches to implement a new forwarding mechanism in the data plane of the 6G core network. More in detail, we define a Q-learning-based forwarding mechanism that acts at packet level and is able to select the minimum latency path at line rate. Our solution, referred to as Q-Learning-based Queue Length Routing in DAta Plane ((QL)2-RODAP), is fully decentralized and exploits in-band network telemetry to distribute network states among network nodes. We show that, either in random and real network topologies, our (QL)2-RODAP algorithm promptly reacts to sudden traffic bursts, and allows reducing the peak of queuing delays of about 65 − 85% with respect to other RL based approaches, thus cutting off the long tail of end-to-end latency that is critical for delay sensitive applications.

Related Organizations

Sapienza University of Rome
Italy
Roma Tre University
Italy
University of Molise
Italy
RISE Research Institutes of Sweden
Sweden

Keywords

low latency communications, in-network computing, in-band network telemetry, network programmability

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

Average

Upload OA version

Are you the author of this publication? Upload your Open Access version to Zenodo!

It’s fast and easy, just two clicks!

uploadUpload now