• español
    • English
  • Login
  • español 
    • español
    • English
  • Tipos de Publicaciones
    • bookbook partconference objectdoctoral thesisjournal articlemagazinemaster thesispatenttechnical documentationtechnical report
Ver ítem 
  •   IMDEA Networks Principal
  • Ver ítem
  •   IMDEA Networks Principal
  • Ver ítem
JavaScript is disabled for your browser. Some features of this site may not work without it.

Distributing Inference in the User Plane of Complex Network Topologies with DUNE

Compartir
Ficheros
Distributed_User_Plane_Inference_AuthorVersion.pdf (5.392Mb)
Identificadores
URI: https://hdl.handle.net/20.500.12761/2064
Metadatos
Mostrar el registro completo del ítem
Autor(es)
Bütün, Beyza; de Andrés Hernández, David; Carré, Alexis; Gucciardo, Michele; Fiore, Marco; Bütün, Beyza
Fecha
2026
Resumen
The deployment of Machine Learning (ML) models in the user plane has emerged as a promising approach to enable line-rate in-network inference, thereby reducing end-to-end latency and improving the scalability of network functions such as network telemetry. Nevertheless, integrating ML models into programmable switches remains challenging due to stringent memory and computational constraints. Prior work has predominantly focused on deploying monolithic ML models into individual programmable network devices, an approach fundamentally limited by hardware resources, particularly for complex classification tasks. In this paper, we introduce DUNE, a novel framework that, for the first time, enables distributed user-plane inference across multiple programmable network devices. DUNE employs original, fully automated techniques to (i) decompose large ML models into lightweight sub-models that retain inference accuracy while reducing resource consumption, and (ii) determine the design, sequencing, and placement of these sub-models to support efficient, distributed joint packet- and flow-level inference. We deploy P4 implementations of DUNE both in an experimental testbed with industry-grade programmable switches and in an emulation environment. We leverage the former setup to demonstrate the practical viability of the solution with real-world hardware in a simple linear topology, and the latter to evaluate DUNE’s scalability to larger and more complex network topologies coexisting with routing strategies. In both cases, we evaluate the framework using real-world traffic on two challenging classification tasks: our results show that DUNE not only reduces per-switch resource usage compared to traditional monolithic ML deployments but also improves inference accuracy by up to 7.5%.
Compartir
Ficheros
Distributed_User_Plane_Inference_AuthorVersion.pdf (5.392Mb)
Identificadores
URI: https://hdl.handle.net/20.500.12761/2064
Metadatos
Mostrar el registro completo del ítem

Listar

Todo IMDEA NetworksPor fecha de publicaciónAutoresTítulosPalabras claveTipos de contenido

Mi cuenta

Acceder

Estadísticas

Ver Estadísticas de uso

Difusión

emailContacto person Directorio wifi Eduroam rss_feed Noticias
Iniciativa IMDEA Sobre IMDEA Networks Organización Memorias anuales Transparencia
Síguenos en:
Comunidad de Madrid

UNIÓN EUROPEA

Fondo Social Europeo

UNIÓN EUROPEA

Fondo Europeo de Desarrollo Regional

UNIÓN EUROPEA

Fondos Estructurales y de Inversión Europeos

© 2021 IMDEA Networks. | Declaración de accesibilidad | Política de Privacidad | Aviso legal | Política de Cookies - Valoramos su privacidad: ¡este sitio no utiliza cookies!