NeFut Logo NeFut
Admin Login

[CS.AI] Modality-Decoupled Federated Learning for Privacy-Preserving Embodied Intelligence in 6G

Published at: 2026-09-12 22:00 Last updated: 2026-09-15 01:15
#algorithm #Machine Learning #Artificial Intelligence

6G is expected to provide the key infrastructure for large‑scale embodied intelligence, where heterogeneous robots collaborate via low‑latency connectivity, edge intelligence, and distributed sensing. Vision‑language‑action (VLA) models form a unified closed‑loop policy that integrates visual perception, language understanding, and action generation. However, training and adapting VLA models for distributed robotic agents raises challenges in privacy protection, communication efficiency, and model heterogeneity. Existing federated learning (FL) approaches overlook the intrinsic differences among vision, language, and action pathways in parameter scale, privacy exposure, update dynamics, and tolerance to compression or perturbation.

To address this, we propose FedMVLA, a modality‑decoupled FL framework tailored for privacy‑preserving embodied intelligence in 6G networks. FedMVLA comprises three mechanisms:

FedMVLA also introduces a modality‑sliced transport design that routes the precision‑critical action stream through a protected ultra‑reliable low‑latency (URLLC) slice, while other modalities use the conventional enhanced mobile broadband (eMBB) slice.

A case study is conducted on a 3GPP‑based wireless substrate, accounting for fading, co‑channel interference, and malicious jamming. Results show:

These findings demonstrate that modality‑aware aggregation, privacy allocation, and communication compression can dramatically enhance the efficiency and reliability of distributed robotic collaboration while preserving privacy and real‑time performance.

Review

Original Source: https://arxiv.org/abs/2609.09591

[h] Back to Home