HMS-SCP: Task-Oriented Multi-Scale Semantic Communication for V2X Cooperative Perception
Authors: Chun-Yeow Yeoh, Chee Keong Tan, Joanne Mun-Yee Lim, Heng-Siong Lim
First: 2026-07-03T14:02:05+00:00 · Latest: 2026-08-18T16:39:12+00:00
Comments: 15 pages, 7 figures, 6 tables, Submitted to IEEE Transactions on Vehicular Technology (TVT)
Abstract
Cooperative perception enables vehicles and infrastructure to exchange sensor data via Vehicle-to-Everything (V2X) communication, extending sensing coverage beyond occlusions and mitigating blind spots. While critical for autonomous driving and safety, practical deployments often rely on bandwidth-efficient late fusion. Recently, intermediate fusion has emerged as a promising approach for an optimal bandwidth-accuracy trade-off. However, in dense urban environments, cumulative bandwidth demands can overwhelm network capacity, potentially compromising safety-critical Cooperative Intelligent Transport Systems (C-ITS) functions. To alleviate these problems, this paper proposes Hierarchical Multi-Scale Semantic-Aware Cooperative Perception (HMS-SCP), a robust noise-resilient and bandwidth-efficient framework for task-oriented semantic communication in cooperative perception. HMS-SCP employs a spatial importance predictor to identify task-relevant grid elements at each scale, which are then directly mapped into complex-valued symbols for Joint Source-Channel Coding (JSCC). Unlike prior methods that rely on high-dimensional symbol projections for robustness, HMS-SCP exploits structural semantic redundancy across multiple scales to enhance resilience against channel noise, while maintaining an ultra-low symbol rate. This design significantly reduces bandwidth consumption and mitigates network congestion in high-density vehicular environments. Extensive evaluations on the simulated OPV2V and real-world DAIR-V2X datasets demonstrate that HMS-SCP effectively prevents performance collapse under severe Rayleigh fading and extreme compression ratio, maintaining high-confidence far-field detection with a real-time latency of below 16~ms, well within the safety-critical thresholds for dynamic V2X environments.
Summary / 总结
Cooperative perception enables vehicles and infrastructure to exchange sensor data via Vehicle-to-Everything (V2X) communication, extending sensing coverage beyond occlusions and mitigating blind spots.
LLM-Driven Large-Scale Spectrum Access
Authors: Ning Yang, Jinliang Gao, Haijun Zhang
First: 2026-04-14T02:08:49+00:00 · Latest: 2026-08-18T05:43:28+00:00
Comments: 11 pages, 2 figures, 8 tables. Submitted to IEEE Transactions on Mobile Computing (TMC)
Abstract
Efficient spectrum management in massive-scale wireless networks is increasingly challenged by explosive action spaces and the computational intractability of traditional optimization. This study proposes a LLM-Driven Large-Scale Spectrum Access (LSA) framework rooted in Group Relative Policy Optimization (GRPO). To overcome the computational intractability caused by ultra-long prompts in large-scale scenarios, we develop a hierarchical state serialization mechanism that synthesizes global environment statistics with localized critical constraints, enabling the LLM to perform high-dimensional reasoning within a bounded context window. Simulation results under strictly time-bounded inference protocols reveal that the code-driven paradigm eliminates the Supervised Fine-Tuning (SFT) cold-start bottleneck and leverages direct execution feedback to achieve superior scaling laws. The framework maintains robust spectral utility and generalization across varying network scales, yielding consistent and empirically superior performance over stochastic heuristics, and surpassing partitioned classical solvers in ultra-dense regimes under matched compute budgets. Code is available at https://github.com/Xtdzs/LLM-Driven-Large-Scale-Spectrum-Access.
Summary / 总结
Efficient spectrum management in massive-scale wireless networks is increasingly challenged by explosive action spaces and the computational intractability of traditional optimization.
An O-RAN-Assisted MARL Approach for Dynamic Sidelink and Infrastructure Selection in V2X Communications
Authors: Maria Katarine Santana Barbosa, Kelvin Lopes Dias
First: 2026-08-17T23:46:39+00:00 · Latest: 2026-08-17T23:46:39+00:00
Comments: This paper has been accepted for publication in IEEE Transactions on Vehicular Technology
Abstract
Future applications in the 6G-based Internet of Vehicles will leverage sidelink (SL) transmissions in Vehicle-to-Everything (V2X) scenarios. However, SL-based direct communication can significantly increase interference among vehicles and between vehicles and other entities of the Intelligent Transportation System. Thus, both Vehicle-to-Vehicle communications and Vulnerable Road Users (VRUs) uplink resources may be degraded or subject to starvation. Existing solutions primarily focus on improving resource allocation and pair selection. Nonetheless, they lack a comprehensive approach to tackle the communication modes and the entire network. To address these challenges, this paper leverages Open RAN to manage V2X communication and proposes a multi-agent reinforcement learning (MARL) resource-aware system. Open RAN provides control loops through a global view of the network and also an open interface-based framework for machine learning models applied to resource decision-making. Meanwhile, the MARL model aims to mitigate interference, optimize resource usage, and enhance quality of service by optimally selecting between sidelink and network transmissions. To reduce system complexity, this work employs a clustering strategy. Each agent manages a group of pairs, rather than assigning one agent to each pair. The solution supports this design by adopting a centralized training with decentralized execution approach, empowered by Open RAN. The strategy uses offline training and an off-policy approach, in which each agent stores experience for fine-tuning. Results indicate that the MARL approach reduces average loss by 21% and latency by 19% in Vehicle-only scenarios. In coexistence VRU scenarios, loss and latency drop by 18% and 30%, respectively, compared to the single-agent approach.
Summary / 总结
Future applications in the 6G-based Internet of Vehicles will leverage sidelink (SL) transmissions in Vehicle-to-Everything (V2X) scenarios.
Age of Gossip in Ring Networks With Non-Poisson Updates
Authors: Arunabh Srivastava, Sennur Ulukus
First: 2026-05-06T17:23:50+00:00 · Latest: 2026-08-17T17:35:58+00:00
Abstract
We consider a network consisting of $n$ nodes connected in a ring formation and a source that generates updates according to a renewal process and disseminates them to the ring network according to a Poisson process. The nodes in the network gossip with each other according to a push-based gossiping protocol, and disseminate version updates. Gossip between two neighbors happens at the arrivals of renewal processes with finite mean and variance. All renewal processes and Poisson processes in the network are independent but not identically distributed. We consider both uni-directional ring networks and bi-directional ring networks. We use version age of information to quantify the freshness of information at each node. Prior work has used the stochastic hybrid systems (SHS) approach or a first passage percolation (FPP) approach to analyze ring networks with edges following identical Poisson processes. In this work, we use a sample-path backtracking approach to characterize the probabilistic scaling of the version age of information of an arbitrary node in the gossip network, where each edge follows an independent but not identically distributed renewal process. We show that the version age of information of any node in the network is stochastically equivalent to $\sqrt{n}$ at any time instant after the node has received its first update from the source.
Summary / 总结
We consider a network consisting of $n$ nodes connected in a ring formation and a source that generates updates according to a renewal process and disseminates them to the ring network according to a Poisson process.
Expanding Access, Exposing Risk: A Short Study of Exposed Starlink Hosts
Authors: Omar Elamri, Isaac-Neil Zanoria, Jacob Zhi, Ben Du, Liz Izhikevich
Venue: ACM IMC Workshop of Policy-Relevant Internet Measurements and Experimentation (PRIMES), October 2025
First: 2026-08-17T17:25:10+00:00 · Latest: 2026-08-17T17:25:10+00:00
Abstract
In this very short paper, we present a measurement-driven analysis of the security characteristics of Starlink-connected hosts and uncover several concerning trends. We find that Starlink hosts are more likely to run outdated or vulnerable operating systems and network protocols than non-Starlink hosts. Regions like Latin America, Southeast Asia, and Eastern Europe show disproportionately higher risk. Our findings raise important questions for the Internet measurement and policy communities.
Summary / 总结
In this very short paper, we present a measurement-driven analysis of the security characteristics of Starlink-connected hosts and uncover several concerning trends.
LLMs for Zero-Shot Threat Detection via Structured Risk Indicators
Authors: Abdullah Alghamdi, Siamak Layeghy, Marius Portmann
First: 2026-08-17T12:47:38+00:00 · Latest: 2026-08-17T12:47:38+00:00
Abstract
We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from heterogeneous security logs. The framework models user activity as chronological timelines and incorporates retrieval-augmented generation (RAG) to provide personalised behavioural context from each user's historical activity. Rather than performing end-to-end classification directly from raw logs, it first generates structured, interpretable sets of threat-specific risk indicators, which are then classified jointly across temporal sequences to capture attack patterns spanning multiple windows.The framework is evaluated on two benchmark datasets, CERT r5.2 for insider threat detection and PicoDomain for APT detection, using four combinations of two open-weight LLMs under both retrieval and non-retrieval settings. All configurations outperform the previous state-of-the-art LLM-based framework (GABM), with the best configuration improving the F1-score by 11.40 percentage points on CERT r5.2 and 31.50 percentage points on PicoDomain. Results further show that retrieval mainly benefits weaker LLMs by generating more discriminative risk indicators, whereas stronger models achieve comparable performance without retrieved context. The most effective assignment of LLMs to the two stages depends on the dataset. These findings show that the quality of the generated risk indicators is the main driver of zero-shot cyber threat detection performance.
Summary / 总结
We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from heterogeneous security logs.
Towards the Interplanetary Internet: An IoT Perspective
Authors: Carles Gomez, Jon Crowcroft
First: 2026-08-17T09:33:20+00:00 · Latest: 2026-08-17T09:33:20+00:00
Abstract
Public administrations and private companies have announced plans to deploy networking infrastructure to support future robotic and human presence on or near space targets, such as the Moon and Mars. While using an IP protocol stack for deepspace communication had been neglected, recent events have motivated the reconsideration of IP to enable the Interplanetary Internet. This new paradigm facilitates the integration of IP-based Internet of Things (IoT) protocols for deep-space environments. This paper illustrates the similarities between deep-space and IoT scenarios, presents related IETF standardization work, and discusses opportunities and future directions for IP-based IoT protocols in the Interplanetary Internet.
Summary / 总结
Public administrations and private companies have announced plans to deploy networking infrastructure to support future robotic and human presence on or near space targets, such as the Moon and Mars.
Quantum-Safe Web Service Architecture Using Time-Based One-Time Passwords
Authors: Abel C. H. Chen
First: 2026-08-17T04:58:23+00:00 · Latest: 2026-08-17T04:58:23+00:00
Abstract
One-Time Passwords (OTPs) have become a common option for multi-factor authentication in several applications. For instance, during website login processes, OTPs are often used in conjunction with traditional text-based usernames and passwords to verify whether the access request originates from a legitimate human user rather than an automated agent. However, in scenarios involving automated connections and system-to-system interoperability, Time-Based One-Time Passwords (TOTPs) may be required to establish secure connections and access Web Services (WSs). Therefore, this study focuses on exploring the development of a quantum-safe web service architecture. The proposed approach achieves transmission security management by implementing Transport Layer Security (TLS) and HyperText Transfer Protocol Secure (HTTPS) based on Post-Quantum Cryptography (PQC). Furthermore, web service security management is realized through the construction of keyed-Hash Message Authentication Code (HMAC)-driven TOTPs. Within the experimental environment, this study evaluates and compares the computational performance of the Secure Hash Algorithm-2 (SHA-2), SHA-3, Ascon-Hash256, and SM3. The required computation time under different hardware resource conditions is analyzed for future web service deployment.
Summary / 总结
One-Time Passwords (OTPs) have become a common option for multi-factor authentication in several applications.
Scaling the Lightning Network with Practical Set Reconciliation
Authors: Xingyu Chen, Anish Sinha, David Starobinski, Ari Trachtenberg
Venue: 2026 IEEE International Conference on Blockchain and Cryptocurrency (ICBC), 2026, pp. 1-5
First: 2026-08-16T20:31:14+00:00 · Latest: 2026-08-16T20:31:14+00:00
Comments: Published in the 2026 IEEE International Conference on Blockchain and Cryptocurrency (ICBC 2026)
Abstract
The Lightning Network (LN) utilizes gossip to share network topology, channel announcements and updates, and node announcements among its local constituents. Yet, our measurements show that this flooding-based gossip reconciliation is fundamentally inefficient. We propose, instead, to use set reconciliation protocols for sharing this information, and we systematically evaluate existing approaches under realistic network conditions. We further propose ADAPTIVEIBLT, a novel adaptive IBLT (Invertible Bloom Lookup Table) protocol with a partial-decoding enhancement. By simulating reconciliation in Core-Lightning and evaluating real gossip snapshots, we demonstrate the practical benefits of reconciliation in scaling gossip reconciliation from hours down to a few minutes.
Summary / 总结
The Lightning Network (LN) utilizes gossip to share network topology, channel announcements and updates, and node announcements among its local constituents.
WiFiSpectralJam: A Large-Scale Open Wi-Fi Spectral Scan Dataset with Controlled RF Jamming
Authors: Dania Herzalla, Govind Singh, Willian T. Lunardi, Martin Andreoni
First: 2026-08-16T13:06:32+00:00 · Latest: 2026-08-16T13:06:32+00:00
Abstract
WiFiSpectralJam is a Wi-Fi spectral-scan dataset comprising 14.52 GB, 96,090 CSV files, and 522,771,130 ordered spectral observations using commodity Wi-Fi sensing hardware. Measurements were acquired with a Raspberry Pi Compute Module 4 equipped with a Qualcomm Atheros QCA9880 802.11ac network interface and the Linux ath10k spectral-scan interface. The dataset spans active and passive scan modalities across the 2.4 and 5 GHz bands and includes real-world benign background captures, benign RF-chamber floor captures, and controlled RF-jamming captures generated with a HackRF One. Jamming conditions vary by transmit power, target channel, and, in the active subset, waveform type. The release provides the raw spectral-scan records together with a file-level metadata manifest, derived spectral-summary features, validation outputs, and reproducible benchmark protocols. These resources support reuse in RF interference characterisation, jamming detection, spectrum monitoring, distribution-shift evaluation, and machine-learning studies using commodity-NIC spectral measurements. The dataset is publicly available at: https://www.kaggle.com/datasets/daniaherzalla/radio-frequency-jamming/data.
Summary / 总结
WiFiSpectralJam is a Wi-Fi spectral-scan dataset comprising 14.52 GB, 96,090 CSV files, and 522,771,130 ordered spectral observations using commodity Wi-Fi sensing hardware.
OTel: Building Domain-Specialized Telecom LLM Foundations for Intelligent Networks
Authors: Farbod Tavakkoli, Roderic Paulk, Jorden Terrazas, Kenneth Church, Mark Austin, Louis Powell, Gregory Diamos, Lina Bariah, Syed Ali Raza Zaidi, Maryam Hafeez, Ali Maatouk, Imtiaz Karim
First: 2026-08-15T22:42:34+00:00 · Latest: 2026-08-15T22:42:34+00:00
Comments: Accepted at the ACM AI Leadership Summit, Breakthrough Impact Highlights Track, 2026
Abstract
Frontier AI models have advanced rapidly, but they still struggle with telecom-specific tasks. We present Open Telco (OTel), an open telecom AI resource with derived datasets for retrieval, reranking, instruction tuning, and safety/abstention, plus 30 full-parameter post-trained baselines across embedding, reranking, and language models. The community has already engaged substantially with the resource: as of May 3, 2026, the released models have been downloaded over 16 million times, and the project has received 157+ pieces of media coverage worldwide. Building on prior open telecom datasets and benchmarks, OTel provides documented telecom data sources, held-out evaluation partitions, trained embedding models, rerankers, context-grounded LLMs, and safety/abstention data in one unified resource. OTel post-training improves performance across all three model families: embedding retrieval reaches 93.5% NDCG@10, reranking reaches 0.952 MRR@10, and language-model correctness reaches 88.2%. We release OTel as a reproducible starting point and invite the community to expand the data, improve embedding and reranking models, and build stronger context-grounded telecom LLMs.
Summary / 总结
Frontier AI models have advanced rapidly, but they still struggle with telecom-specific tasks.
Exploring the Suitability of QUIC for the Internet of Things
Authors: Carles Gomez, Nika Soltani-Tehrani, Jon Crowcroft
First: 2026-08-15T18:39:26+00:00 · Latest: 2026-08-15T18:39:26+00:00
Abstract
QUIC is an emerging transport-layer protocol that provides reliability and security. QUIC was designed to overcome issues from other protocol stacks used in the Internet, such as TCP/TLS, especially focusing on web traffic performance improvement. Therefore, QUIC was not conceived for Internet of Things (IoT) scenarios, which are characterized by significant resource constraints. However, as QUIC prominance increases, and the IoT continues to expand, QUIC may offer connectivity opportunities for IoT devices. In this paper, we explore the suitability of QUIC for IoT environments. Leveraging optional functionality, we propose, discuss, and evaluate a QUIC profile for IoT scenarios that is currently being considered for IETF standardization.
Summary / 总结
QUIC is an emerging transport-layer protocol that provides reliability and security.
ISAC in 3GPP: Evolution Toward 6G
Authors: Neeraj Varshney
First: 2026-08-15T15:28:31+00:00 · Latest: 2026-08-15T15:28:31+00:00
Comments: submitted for possible publication in IEEE Journal
Abstract
Integrated sensing and communication (ISAC) is emerging as an important direction in the Third Generation Partnership Project (3GPP) evolution toward 6G because it allows cellular networks to provide environmental awareness in addition to connectivity. This paper surveys the current 3GPP trajectory from Release~19 feasibility studies to Release~20 radio, protocol, and architecture studies, while distinguishing established requirements, ongoing study assumptions, and possible forward directions. The survey covers service requirements, sensing topologies, channel model evolution beyond 3GPP Technical Report (TR)~38.901, Radio Access Network Working Group~1 (RAN1) physical layer design, Radio Access Network Working Groups~2 and~3 (RAN2 and RAN3) system implications, and the role of sensing-assisted communication. It also synthesizes the main unresolved issues in waveform and reference signal design, multi-node coordination, sensing data reporting, service exposure, privacy, and implementation constraints. By connecting service-level motivations to physical layer, protocol, and architecture implications, the paper provides a standards-centric reading of how 3GPP may evolve toward practical 6G ISAC support.
Summary / 总结
Integrated sensing and communication (ISAC) is emerging as an important direction in the Third Generation Partnership Project (3GPP) evolution toward 6G because it allows cellular networks to provide environmental awareness in addition to connectivity.
An Asynchronous Triggered MAC Protocol for Underwater Acoustic Networks
Authors: Bingwen Huangfu, Jiani Guo, Shanshan Song, Nan Sun, Jun Liu, Miao Pan
First: 2026-08-11T06:15:38+00:00 · Latest: 2026-08-15T11:18:23+00:00
Abstract
Time Division Multiple Access (TDMA)-based Medium Access Control (MAC) protocols have proven their practicality through extensive field trials in Underwater Acoustic Networks (UANs), attributable to their hardware compatibility and ease of implementation. In conventional TDMA-based MAC designs, channel access is typically organized using synchronized, fixed-length slots to mitigate contention and coordinate transmissions. However, this paradigm imposes significant clock synchronization overhead in UANs with long and variable propagation delays and struggles to improve scheduling flexibility. Although some protocols attempt to refine this slot paradigm (adjust the slot length to improve channel reuse efficiency or scheduling frequency), they are still constrained by the trade-off between channel utilization and scheduling complexity. To this end, this paper proposes AT-MAC, an Asynchronous Triggered MAC protocol that aims to achieve efficient and fair channel access through coordinated asynchronous scheduling. AT-MAC introduces a triggered slot paradigm without time synchronization, decoupling transmission scheduling from a rigid timeline and enabling asynchronous, variable-length slots to accommodate the long and diverse propagation delays. To power this slot paradigm, AT-MAC augments conventional Multi-Agent Deep Reinforcement Learning to handle asynchronous interaction, achieving coordinated channel access under partial observations. It further devises a load-aware fairness guard mechanism to enable network-wide fairness status inference solely through local overhearing, thereby guiding adaptive scheduling correction to maintain fairness. Field-reconstructed simulations and on-board inference benchmarking demonstrate the feasibility of AT-MAC. Extensive simulation results further demonstrate its consistent performance gains across the evaluated scenarios and traffic conditions.
Summary / 总结
Time Division Multiple Access (TDMA)-based Medium Access Control (MAC) protocols have proven their practicality through extensive field trials in Underwater Acoustic Networks (UANs), attributable to their hardware compatibility and ease of implementation.
Presto: A Match-Action TCP Stack for the Terabit Era
Authors: Rajath Shashidhara, Antoine Kaufmann, Simon Peter
Venue: Presto: A Match-Action TCP Stack for the Terabit Era. In Proceedings of the ACM SIGCOMM 2026 Conference (SIGCOMM'26). Association for Computing Machinery, New York, NY, USA, 1360-1375
First: 2025-04-27T00:13:02+00:00 · Latest: 2026-08-14T15:37:38+00:00
Comments: 19 pages, 14 figures, 3 Tables, Published at ACM SIGCOMM'26
Abstract
We present Presto, the first TCP stack that delivers ASIC-class performance and energy efficiency on programmable Reconfigurable Match-Action Table (RMT) pipelines, providing flexibility while retaining standard TCP semantics and POSIX socket compatibility. The key challenge in designing Presto is reconciling TCP's complex, dependent state updates with RMT's unidirectional, lock-step execution model. To overcome this challenge, Presto introduces three novel techniques: optimistic concurrency (speculative updates validated downstream), pseudo-segment injection (circular dependency resolution without stalls), and bump-in-the-wire processing (single-pass segment handling). Together, these enable TCP retransmission, reassembly, flow, and congestion control, as a pipeline of simple match-action operations.
Our Intel Tofino 2 prototype demonstrates Presto's scalability to terabit speeds, flexibility, and robustness to network dynamics. Presto matches RDMA performance and efficiency for both RPC and streaming workloads (including NVMe-oF with SPDK), while maintaining TCP/POSIX compatibility. Presto saves up to 16 host CPU cores versus state-of-the-art kernel-bypass TCP, while achieving 5$\times$ lower 99.99p tail latency and 2$\times$ better throughput-per-watt for key-value stores. At scale, Presto drives nearly $1$ Bpps at 20 $μ$s RPC tail latency. Unlike fixed-function offloads, Presto supports transport evolution through in-data-path extensions (selective ACKs, congestion control variants, application co-design for shared logs). Finally, Presto generalizes to FPGA SmartNICs, outperforming Tonic's monolithic design by $3\times$ under equal timing.
Summary / 总结
We present Presto, the first TCP stack that delivers ASIC-class performance and energy efficiency on programmable Reconfigurable Match-Action Table (RMT) pipelines, providing flexibility while retaining standard TCP semantics and POSIX socket compatibility.
TurboRetry: Mitigating Large-Scale QUIC Handshake Floods with Off-the-Shelf DPU Offloading
Authors: Jiahao Wu, Heng Pan, Kai Lv, Zhenyu Li, Yanbiao Li, Gaogang Xie
First: 2026-08-03T14:08:04+00:00 · Latest: 2026-08-14T14:36:40+00:00
Abstract
The modern transport protocol QUIC is designed to enhance network performance and security, but it remains vulnerable to handshake flooding attacks. Such attacks exhaust CPU resources by forcing the server to perform expensive cryptographic operations via a large number of handshaking requests. QUIC provides a built-in defense mechanism, the Retry mechanism, to mitigate these attacks. However, our experiments reveal that it can still become a performance bottleneck under large-scale QUIC handshake floods due to substantial computational overhead. In this paper, we design and implement TurboRetry, a split design, that offloads the Retry mechanism onto DPUs to efficiently mitigate QUIC handshake floods. TurboRetry partitions the tasks of the Retry into two categories, and then assigns them to the DPUs and the host, respectively. To preserve QUIC semantics and reduce the coordination overhead, TurboRetry designs an extended Retry token format and an efficient cooperation scheme. In addition, TurboRetry offloads the connection authorization task to the on-path DPA to further improve both performance and security. Our evaluation shows that TurboRetry outperforms the host-side implementation by a wide margin, improving throughput by 10-20$\times$.
Summary / 总结
The modern transport protocol QUIC is designed to enhance network performance and security, but it remains vulnerable to handshake flooding attacks.
Robust Constraint-Aware Bayesian Tuning of BBRv2 for QUIC under Tactile Internet Constraints
Authors: Muhammad Hanif Lashari, Shakil Ahmed, Wafa Batayneh, Ashfaq Khokhar
First: 2026-08-14T14:00:30+00:00 · Latest: 2026-08-14T14:00:30+00:00
Abstract
Tactile Internet applications place strict require- ments on latency, jitter, loss, and responsiveness, which makes transport configuration a critical design factor. Although BBRv2 offers a model-based congestion control framework with strong throughput potential, its default behavior may not be well aligned with delay-sensitive interactive scenarios. This paper presents a robust and constraint-aware tuning framework for BBRv2 in QUIC, where parameter selection is formulated as an expensive black-box optimization problem over multiple emulated network conditions. The tuning process uses Bayesian optimization with the Tree Structured Parzen Estimator to efficiently explore a bounded parameter space under noisy experimental measure- ments. The objective is designed to preserve throughput while enforcing limits on tail latency and loss, while delay instability is evaluated separately through the jitter metric. Experimental results across low, medium, and high impairment scenarios show that the tuned configuration improves tail latency, jitter behavior, and loss performance while maintaining competitive goodput relative to standard QUIC congestion control baselines. These results support robust black-box tuning as a practical method for adapting QUIC transport behavior to tactile Internet style requirements.
Summary / 总结
Tactile Internet applications place strict require- ments on latency, jitter, loss, and responsiveness, which makes transport configuration a critical design factor.
Bridging Network Fragmentation: A Semantic-Augmented DRL Framework for UAV-aided VANETs
Authors: Gaoxiang Cao, Wenke Yuan, Huasen He, Yunpeng Hou, Xiaofeng Jiang, Shuangwu Chen, Jian Yang
First: 2026-03-19T13:15:52+00:00 · Latest: 2026-08-14T10:07:35+00:00
Comments: Revised version. This update includes substantial improvements to the methodology, ablation studies, temporal robustness analysis, and cross-city generalization experiments. Submitted to IEEE Transactions on Cognitive Communications and Networking. 15 pages, 14 figures
Abstract
Urban Vehicular Ad-Hoc Networks (VANETs) can become fragmented because buildings obstruct wireless links and vehicle mobility continuously changes the network topology. Unmanned Aerial Vehicles (UAVs) can serve as mobile relays, but Deep Reinforcement Learning (DRL)-based deployment often suffers from inefficient exploration because it lacks road-topology guidance. To address this problem, we propose Semantic-Augmented DRL (SA-DRL), which models network fragmentation over the road topology and aligns a pretrained Large Language Model (LLM) to generate a topology-dependent action prior from dynamic traffic states. The resulting Semantic-Augmented PPO (SA-PPO) algorithm combines this prior with the PPO policy through Logit Fusion, guiding exploration toward promising intersections while retaining adaptation through environmental returns. Simulations driven by real-world urban trajectories show that SA-PPO reaches the final converged reward of Vanilla PPO using only 28.6% of its training episodes. It improves the average number of vehicles in connected components and the average connected-component size by 7.9% and 8.7%, respectively, while reducing UAV energy consumption by 21.3%.
Summary / 总结
Urban Vehicular Ad-Hoc Networks (VANETs) can become fragmented because buildings obstruct wireless links and vehicle mobility continuously changes the network topology.
CipherSight: Robust Website Fingerprinting via Record-Resource Semantic Supervision under Distribution Shifts
Authors: Runhan Song, Qiqi Liu, Chuanzhou Pan, Zhenquan Ding, Youquan Xian, Chongru Fan, Lei Cui, Wei Wang, Zhiyu Hao
First: 2026-08-14T03:17:57+00:00 · Latest: 2026-08-14T03:17:57+00:00
Abstract
HTTPS website fingerprinting (WF) aims to identify visited websites from metadata observable in encrypted traffic. However, real-world deployments introduce a significant out-of-distribution (OOD) problem caused by temporal and geographic changes, while previously unseen websites are common in open-world scenarios. Existing methods primarily learn from raw TCP packet sequences and struggle to capture stable and generalizable website representations, resulting in performance degradation under practical conditions.
We propose CipherSight, a TLS-record-based hierarchical framework for robust HTTPS WF. Unlike existing approaches that rely on TCP packet sequences and are sensitive to transport-layer artifacts, CipherSight learns website representations from TLS records by jointly encoding multiple record-level attributes. It introduces a hierarchical architecture that captures both intra-flow dependencies among TLS records and inter-flow interactions across concurrent flows, enabling the model to exploit structural patterns in HTTPS traffic. Besides, to learn robust representations, CipherSight employs a masked record modeling (MRM) task to capture contextual traffic semantics and leverages fine-grained record-resource annotations as privileged supervision through structure-aware objectives and semantic distillation. Experiments show that CipherSight achieves 95.41% accuracy across more than 2,000 website classes in the closed-world setting and maintains over 90% accuracy under both temporal and geographic drift, consistently outperforming all evaluated baselines.
Summary / 总结
HTTPS website fingerprinting (WF) aims to identify visited websites from metadata observable in encrypted traffic.
InterSAGE: The Secure and Verifiable Interoperability Protocol for An Internet of Agents
Authors: Zhenhua Zou, Sheng Guo, Qiuyang Zhan, Lepeng Zhao, Shuo Li, Zhuotao Liu
First: 2026-08-13T10:00:13+00:00 · Latest: 2026-08-14T02:36:49+00:00
Comments: 35 pages, 4 figures, 7 tables. Positioning paper
Abstract
The emerging Internet of Agents enables LLM-powered agents to discover peers, invoke tools, and delegate tasks across organizational boundaries. Existing protocols increasingly define how agents exchange messages, but not how an agent proves its identity, authorization, advertised capabilities, or accountability after delegation. We present InterSAGE, a trust-native protocol suite that supplies this missing security substrate alongside, rather than in place of, communication protocols. InterSAGE comprises four layers: Persistent Identity, Discovery, Trust Negotiation, and Accountability. Its four core primitives are: (1) Agent Identity Cards that bind developer, code package, operator, and deployment context; (2) capability-aware discovery using DID-bound Verifiable Credential manifests; (3) trust negotiation combining monotonic capability attenuation with two-tier access control; and (4) kernel-mediated cryptographic audit trails that bind usage, delegation, and execution traces to agent identity without a consensus ledger. InterSAGE is designed to complement MCP, A2A, ANP, and AG-UI, allowing communication protocols to evolve independently while keeping trust semantics explicit, portable, and verifiable. We compare InterSAGE with more than 50 efforts spanning agent protocols, decentralized identity, OAuth/OIDC extensions, zero-trust governance, delegation, and audit architectures. We show that no prior architecture jointly enforces persistent identity, capability-aware discovery, trust negotiation, and accountability as a unified four-layer trust substrate for secure agent interoperability.
Summary / 总结
The emerging Internet of Agents enables LLM-powered agents to discover peers, invoke tools, and delegate tasks across organizational boundaries.
MLCC: A Congestion Control Technique to Accelerate ML Training
Authors: Anton A. Zabreyko, Sanjoli Narang, Sudarsanan Rajasekaran, Manya Ghobadi
First: 2024-02-14T21:33:18+00:00 · Latest: 2026-08-13T21:54:15+00:00
Comments: Anton A. Zabreyko, Sanjoli Narang: Equal Contribution
Abstract
We present MLCC, a novel technique to augment today's congestion control algorithms to accelerate DNN training jobs in shared GPU clusters in a fully distributed manner. At the heart of MLCC lies a straightforward principle: DNN training flows should scale their sending rate to shift other flows' communication into their compute periods, achieving interleaving. We show that integrating this principle into today's congestion control protocols is simple (requiring less than 60 lines of code for a given protocol) and enables DNN jobs to interleave within a few training iterations, thereby reducing network contention and improving job completion times. Our testbed demonstrates that MLCC accelerates the average and 99th percentile training iteration times by up to 1.9x and 2.7x respectively. Through extensive packet-level simulations, we observe a 1.35x improvement in training throughput on a 36-node, 288 GPU fat-tree topology.
Summary / 总结
We present MLCC, a novel technique to augment today's congestion control algorithms to accelerate DNN training jobs in shared GPU clusters in a fully distributed manner.
Weird Machines in Transport Layer Security
Authors: Michael Collins, Jada Cumberland, Brianne Dunn, Ross Gore, Samuel Jackson, Sachin Shetty, Jonathan Takeshita
First: 2026-08-13T18:28:01+00:00 · Latest: 2026-08-13T18:28:01+00:00
Comments: 17 pages, 3 figures, 4 tables
Abstract
Weird machines are latent computational capabilities that emerge from the composition of architectural components. Prior work has studied this phenomenon extensively in software systems, including x86 instructions, ELF metadata, and page tables, and more recently in cyber-physical systems such as industrial control networks. This paper extends weird machine theory to a new domain: the Transport Layer Security (TLS) handshake and its two dominant implementations, OpenSSL and BoringSSL.
We show that legitimate TLS primitives, including session cache entries, renegotiation logic, extension parsing, and certificate verification steps, compose into Turing-complete systems whose computation is coupled to authentication and trust decisions rather than physical actuation. We formalize this coupling, which we call trust actuation, and argue that any TLS implementation providing session storage, arithmetic on sequence counters, conditional branching on handshake state, and iteration through resumption or retry loops satisfies the conditions for arbitrary computation.
We validate this theory with two working demonstrations built on real OpenSSL code paths. The first, a sentinel system, composes standard TLS primitives into a defensive mechanism that detects anomalous handshake behavior. The second, an authentication bypass, composes the same class of primitives into an attack that defeats a cipher-strength policy check through mid-connection renegotiation, without any memory corruption or external malware. Both demonstrations run against real server and client binaries in Docker.
Summary / 总结
Weird machines are latent computational capabilities that emerge from the composition of architectural components.
A Q-learning-based QoS-aware multipath routing protocol in IoMT-based wireless body area network
Authors: Mehdi Hosseinzadeh, Roohallah Alizadehsani, Amin Beheshti, Hamid Alinejad-Roknyd, Lu Chen, Mohammad Sadegh Yousefpoor, Efat Yousefpoor, Muneera Altayeb, Thantrira Porntaveetus, Sadia Din
First: 2026-04-16T19:41:49+00:00 · Latest: 2026-08-13T05:51:10+00:00
Comments: Due to substantial changes in the contributions and responsibilities of the researchers involved in the project, the authorship of the manuscript requires revision to accurately reflect the current contributions. We therefore request withdrawal of the present version to appropriately resolve the authorship and contribution record
Abstract
The Internet of Medical Things (IoMT) enables intelligent healthcare services but faces challenges such as dynamic topology, energy constraints, and diverse QoS requirements. This paper proposes QQMR, a Q-learning-based QoS-aware multipath routing method for WBANs. QQMR classifies data into three priority levels and employs adaptive multi-level queuing and fuzzy C-means clustering to optimize routing decisions. It maintains separate learning policies for each data type and selects primary and backup paths accordingly. Experimental results demonstrate improved packet delivery ratio and significant reductions in delay, routing overhead, and energy consumption compared to existing methods.
Summary / 总结
The Internet of Medical Things (IoMT) enables intelligent healthcare services but faces challenges such as dynamic topology, energy constraints, and diverse QoS requirements.
FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting
Authors: Rentao Gu, Yihang Ding, Junjie Li, Yi Ding, Weijing Sang, Xiaoli Huo, Xin Qin, Yuefeng Ji
Venue: R. Gu, Y. Ding, J. Li, Y. Ding, W. Sang, X. Huo, X. Qin, and Y. Ji, Knowl.-Based Syst., vol.341, p.115776, 2026
First: 2026-08-12T04:09:52+00:00 · Latest: 2026-08-12T04:09:52+00:00
Abstract
Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily on textual prompts for modality alignment-introducing nontrivial computational overhead and failing to leverage the rich spectral dynamics inherent in time-series data. To enable prompt-free, frequency-aware adaptation of frozen LLMs, we propose FM-LLM (Frequency-Enhanced Mixture-of-Experts for adapting LLMs to Time Series Forecasting), an autoregressive framework grounded in constrained asymmetric coupling. A Fourier Analysis Network (FAN)-based spectral token aligner injects structured harmonic representations directly into the frozen LLM with numerical compatibility. An asymmetric Mixture-of-Experts (MoE) decoder enforces role separation: shared experts with lightweight FAN layers reconstruct the global periodic backbone, while routed experts-restricted to standard FFNs-specialize in modeling non-periodic residual dynamics. A time-frequency hybrid loss function jointly optimizes temporal accuracy and spectral consistency, mitigating error accumulation during long-horizon autoregressive rollouts. Evaluated across eleven public benchmarks, FM-LLM achieves state-of-the-art performance on 59 out of 78 evaluation metrics. Compared to the strongest autoregressive LLM-based baseline, it delivers average improvements of 5.3% in MSE and 5.6% in MAE, with maximum gains reaching 8.0% for MSE and 8.4% for MAE. FM-LLM also demonstrates robust transferability, maintaining superior performance in 10% few-shot and zero-shot forecasting scenarios.
Summary / 总结
Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting.
TrimMoE A communication aware and adaptive depth framework for distributed edge inference
Authors: Ning Li, Shuting Bai, Xin Yuan, Wenchao Xu, Song Guo, Haijun Zhang
First: 2026-08-01T10:19:02+00:00 · Latest: 2026-08-12T01:54:17+00:00
Comments: 17 pages, 11 figures
Abstract
Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission. The existing approaches mainly focus on how to reach a remote expert faster. However, in this paper, we instead consider whether a given layer, and the layers after it, need to be executed at all. To this end, a communication-aware adaptive-depth framework is proposed in this paper, termed TrimMoE, which couples layer skipping and confidence-based early exit with substitute execution and server-expert selection under a unified quality budget. Specifically, in the offline stage, TrimMoE freezes the backbone, trains the lightweight per-layer exit heads, calibrates the per-layer importance thresholds, and allocates the expert replicas by a skip/exit-aware redundancy benefit. In the online stage, a transition-aware look-ahead anticipates the token movement, so that the depth reduction targets the costliest transmissions, and besides, two feedback rules adapt the delay-quality weights and the exit threshold. Moreover, we prove that the substitution-and-skipping proxy degradation never exceeds the configured budget, and that the early exit is admitted only under a calibrated confidence gate. On a heterogeneous 10-server testbed with Switch-Base-8E, Qwen-MoE-A2.7B, and Mixtral-8x7B, TrimMoE reduces the average latency by up to 62.8%, lowers the cross-server traffic and the remote-execution ratio, and sustains high throughput under load, while keeping the task-quality degradation within a 2% bound.
Summary / 总结
Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission.
OrderMoE: An expert similarity driven distributed edge MoE inference
Authors: Xin Yuan, Ning Li, Quan Chen, Wenchao Xu, Song Guo
First: 2026-07-19T09:20:21+00:00 · Latest: 2026-08-12T01:48:31+00:00
Comments: 17 pages, 12 figures
Abstract
Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over resource-constrained and bandwidth-limited edge infrastructures. Existing distributed MoE serving methods mainly rely on exact expert placement, caching, replication, or communication scheduling, while overlooking the functional similarity among experts, which provides an opportunity to reduce cross-server token transmission. Therefore, this paper introduces a similarity-aware expert allocation and distributed deployment framework, dubbed OrderMoE, which aims to accelerate edge MoE inference while balancing inference latency, communication overhead, server workload, and inference quality. OrderMoE first constructs an expert similarity model based on router-induced logits representations and partitions experts in each MoE layer into multiple similarity groups. Then, it develops a similarity-aware expert grouping and deployment strategy to improve local similarity coverage across edge servers. Since reducing remote expert invocation and preserving exact inference quality are conflicting objectives, OrderMoE further designs a quality-aware and trajectory-aware runtime server-expert selection algorithm to decide whether a token should invoke its remote target expert or use a feasible local substitute expert. Experimental results on a real distributed edge testbed show that OrderMoE significantly reduces average latency, tail latency, cross-server traffic, and remote expert invocation ratio, while introducing only small and controllable inference quality degradation.
Summary / 总结
Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over resource-constrained and bandwidth-limited edge infrastructures.
Self-evolving network verifiers
Authors: Ioannis Protogeros, Tibor Schneider, Laurent Vanbever
First: 2026-08-11T18:45:15+00:00 · Latest: 2026-08-11T18:45:15+00:00
Comments: 8 pages, 5 figures
Abstract
Symbolic network verifiers can reason about correctness across vast spaces of routing inputs and failures, but only for the protocols and features an expert has encoded by hand. Creating and maintaining a faithful model of the control plane is both difficult and never-ending, since no written source specifies perfectly what a network does: vendor implementations deviate from the RFCs, and behaviour shifts with releases. The burden of constant upkeep ultimately keeps verification out of many networks that need it.
We argue that the model should instead evolve automatically to faithfully capture the actual network behaviour. To achieve that, we leverage the only source that specifies it unambiguously: the router software itself. In a counterexample-guided loop, a coding agent proposes extensions to the verifier's symbolic encoding, while a trusted oracle (e.g., emulated routers) supplies the ground-truth routing state. The agent iteratively refines the network model using each disagreement with the oracle.
As early evidence, a prototype of this system taught a 3,000-line SMT-based verifier three features it did not support: OSPF areas, BGP route reflection, and L3VPN over EVPN, converging autonomously on models that match the oracle, even noticing vendor-specific behaviour. Automating model growth shifts the hard problem from writing verification systems to systematically testing them; we propose a research agenda for trusting and harnessing automatically evolved verifiers.
Summary / 总结
Symbolic network verifiers can reason about correctness across vast spaces of routing inputs and failures, but only for the protocols and features an expert has encoded by hand.
Association-based Privacy Attacks in Wireless Protocols: Formal Modeling and Mitigation
Authors: Mohit Kumar Jangid, Felix Engelmann, Zhiqiang Lin
First: 2026-08-11T18:41:56+00:00 · Latest: 2026-08-11T18:41:56+00:00
Abstract
With the surge in privacy-sensitive data from sources such as social media and IoT devices, there is a pressing need for formal, automated methods to assess privacy risks within these intricate systems. This paper formally investigates root sources of pairing-based privacy threats exploited using replay/relay techniques in wireless communication. Our research harnesses condition-oblivious responses, replay-resistance, and distance bounding measures vital for protocols utilizing shared keys in allowlists for authenticated reconnections. Particularly, the paper uses formal modeling of notable wireless networks, like the Wi-Fi P2P persistent group formation and the Bluetooth Low Energy reconnection procedure, to illustrate the root causes and countermeasures. Our model rigorously validates the proposed solution against association inference attacks, along with existing formalizations of well-authentication, frame opacity, and no-desynchronization. The ensuing analysis reveals not only uncharted privacy realms in wireless communication but also identifies old and new vulnerabilities. Our proposed design changes are acknowledged by Wi-Fi Alliance and Bluetooth SIG, paving the way for future advancements in resilient, privacy-preserving wireless protocols.
Summary / 总结
With the surge in privacy-sensitive data from sources such as social media and IoT devices, there is a pressing need for formal, automated methods to assess privacy risks within these intricate systems.
Laser-Diode LiFi With Diffused-Beam Optics: System-Level Modeling and a Cross-Validated ns-3 Simulation Framework
Authors: Hussain Ahmad, Syed Muhammad Talha Gillani, Toheed Omer, Saleem Aslam
Venue: IEEE Journal of Indoor and Seamless Positioning and Navigation 2026
First: 2026-08-11T14:19:06+00:00 · Latest: 2026-08-11T14:19:06+00:00
Comments: 8 pages, 15 figures, Journal
Abstract
Laser diodes (LDs) promise an order-of-magnitude bandwidth advantage over light-emitting diodes for indoor optical wireless access, but reported prototype studies frequently leave the gap between hardware demonstrations and system-level performance unquantified. This paper develops a complete, reproducible system model of a diffused-beam LD LiFi transceiver - a 500-mW laser source beam-shaped by a holographic diffuser, an intensity-modulation/ direct-detection (IM/DD) receiver, and adaptive M-QAM signaling - and embeds it in two cross validated simulators: an open ns-3 module providing full-stack network simulation (channel, PHY, ARQ MAC, Net Device, IP/UDP/TCP) and a Python link-level engine used for Monte Carlo validation of all analytical error models. Starting from a hardware prototype that transferred data, real-time voice, and images over a 14-m line-of-sight link, we identify and close the technical gaps typical of prototype-class reports: serial-interface throughput ceilings misread as optical-link capacity, absent noise modeling, unmeasurable error floors, and unexamined beamwidth/coverage trade-offs. The framework shows that the same optical front end, freed of its 2-Mbaud UART bottleneck and driven at its 250-MHz electrical bandwidth, supports 930 Mb/s net at 14 m under a $3.8 \times 10^{-3}$ HD-FEC threshold with 16-QAM, scales to 1.86 Gb/s at 5 m with 256-QAM, and sustains on-off keying to 23.3 m; a $20^\circ$ diffuser covers a 4.2-m-radius cell of a standard room at desk height. Network simulations over the ns-3 stack yield saturation goodput within 7% of the PHY line rate and sub-0.11-ms 99th-percentile latency at 70% load. All models, code, and figures are released for reproduction.
Summary / 总结
Laser diodes (LDs) promise an order-of-magnitude bandwidth advantage over light-emitting diodes for indoor optical wireless access, but reported prototype studies frequently leave the gap between hardware demonstrations and system-level performance unquantified.
Media-over-Multipath-QUIC for Realtime Video Applications
Authors: Tanya Shreedhar, Zuji Zhou, Nitinder Mohan, Fernando Kuipers
First: 2026-08-11T09:57:19+00:00 · Latest: 2026-08-11T09:57:19+00:00
Comments: In review
Abstract
Multipath transports place a client's WiFi, cellular, and satellite networks under one connection, yet real-time video gains little from them. The scheduler that assigns packets to paths sees only bytes, so it cannot tell a keyframe that anchors a second of video from an enhancement frame whose loss costs one image. We show that the limiting factor is not a shortage of path diversity but the absence of a channel through which the application can name what the transport cannot see. Media over QUIC Transport (MoQT), already deployed on production CDNs, carries media as named Objects, and relays forward each Object's metadata without interpreting it. The knowledge the scheduler lacks therefore already flows through the subscriber's relay. We present MoMQ, a MoQT extension that turns this metadata into path decisions. Applications and relay operators install declarative rules at the edge relay that match Object metadata and express delivery preferences against labeled paths. The relay evaluates the rules mechanically, so it acts on video semantics while containing no video logic. On a live testbed spanning Starlink and WiFi, four rules cut P99.9 frame completion time from 384.7 ms under the best transport-only scheduler to 114.1 ms and reduce the required playback buffer by 61.5%. MoMQ is the only configuration that meets the 150 ms interactive latency target. Across relays in two countries and subscribers on two continents, the same rules apply unchanged and retain their advantage wherever the paths remain disjoint.
Summary / 总结
Multipath transports place a client's WiFi, cellular, and satellite networks under one connection, yet real-time video gains little from them.
Conversational Orchestration for Organic 6G
Authors: Masoud Shokrnezhad, Tarik Taleb
First: 2026-08-11T09:32:52+00:00 · Latest: 2026-08-11T09:32:52+00:00
Comments: 7 pages, 6 figures. Accepted for publication in IEEE Network Magazine
Abstract
The Organic 6G vision of a network of networks spanning an edge-cloud continuum complemented by non-terrestrial resources requires, to realize its promise, service provisioning that is simple to operate, scalable across independently administered domains, and agile under domain churn (i.e., domains dynamically joining and leaving). Despite advances in cross-domain orchestration, many proposals rely on heavy integration fabrics, multi-layer coordinators, and deep telemetry pipelines that hinder deployability and amplify coordination overhead. We propose a lightweight, decentralized conversational orchestration framework based on Large Language Model (LLM)-driven domain agents. Each domain remains autonomous: an agent observes local state via tools, reasons in a closed loop, and exchanges summaries with neighboring agents over an Agent-to-Agent (A2A) overlay aligned with data-plane coupling. Fast feasible placement is enabled by periodic, routing-like dissemination of reachability advertisements (latency, bottleneck bandwidth, and compute capacity), while safe re-optimization, scaling, and migration are handled through event-driven requests and negotiation. To meet real-time constraints, we deploy a compact reasoning model trained with verifier-based self-verification and periodically refined online via shadow updates. Simulations show manageable, near-linear control-plane overhead as domains scale and during domain joins, and robust decision quality, including recovery after objective changes. We close by outlining future research directions for principled, secure, and uncertainty-aware agentic orchestration in Organic 6G.
Summary / 总结
The Organic 6G vision of a network of networks spanning an edge-cloud continuum complemented by non-terrestrial resources requires, to realize its promise, service provisioning that is simple to operate, scalable across independently administered domains, and agile under domain churn (i.e., domains dynamically joining and leaving).
Arcalís: Accelerating Remote Procedure Calls Using a Líghtweight Near-Cache Solution
Authors: Johnson Umeike, Pongstorn Maidee, Bahar Asgari
First: 2026-02-13T04:14:42+00:00 · Latest: 2026-08-11T07:49:01+00:00
Comments: 14 pages, 26 figures
Abstract
Modern microservices increasingly depend on high-performance remote procedure calls (RPCs) to coordinate fine-grained, distributed computation. As network bandwidths continue to scale, the CPU overhead associated with RPC processing, particularly serialization, deserialization, and protocol handling, has become a critical bottleneck. This challenge is exacerbated by fast user-space networking stacks such as DPDK, which expose RPC processing as the dominant performance limiter. While prior hardware accelerators have explored NIC-attached and FPGA-based offload, these approaches remain farther from the cache hierarchy, so the frequent data accesses during RPC processing each pay an extra interconnect traversal cost that inflates RPC time. Therefore, RPC handling should occur as close as possible to the cache; however, a near-cache solution must be small, hence practical and deployable. Our key insight to enable such a solution is taking advantage of a reconfigurable accelerator that can be configured specifically for the services currently running on the CPUs. We present Arcalís, a near-cache RPC accelerator that positions a lightweight hardware engine adjacent to the last-level cache (LLC). Arcalís offloads RPC processing to dedicated microengines that operate with cache-line latency while preserving programmability. By decoupling RPC processing logic, enabling microservice-specific execution, and positioning itself near the LLC, Arcalís achieves a 1.72-4.91$\times$ end-to-end speedup compared to the CPU baseline, significantly reduces microarchitectural overhead by up to 88\%, and achieves up to a 1.62$\times$ higher throughput than prior solutions. These results highlight the potential of near-cache RPC acceleration as a practical solution for high-performance microservice deployment.
Summary / 总结
Modern microservices increasingly depend on high-performance remote procedure calls (RPCs) to coordinate fine-grained, distributed computation.
Cost-Aware Uplink MPQUIC Scheduling via Multi-Objective Bayesian Optimization
Authors: Thanh Trung Nguyen, Thanh Le, Phi Le Nguyen, Kien Nguyen
First: 2026-07-20T18:55:44+00:00 · Latest: 2026-08-11T07:46:10+00:00
Comments: The paper was submitted by a co-author without the consent and approval of the first author and other co-authors
Abstract
Multipath QUIC (MPQUIC) enables simultaneous uplink transmission over heterogeneous access networks such as Wi-Fi and LTE, improving reliability and performance. However, aggressive LTE utilization increases operational cost, creating an inherent trade-off between upload delay and cellular usage. Existing MPQUIC schedulers typically optimize a single performance objective and operate at fixed points within this trade-off space, without explicitly supporting cost-aware operation. This paper formulates uplink MPQUIC scheduling as a multi-objective optimization problem that jointly considers maximum upload completion time and total LTE usage. We propose a Bayesian Optimization-based framework that treats the MPQUIC system as a black box and systematically explores probabilistic path selection configurations to uncover Pareto-efficient operating points. Rather than committing to a predefined scheduling policy, the framework exposes a spectrum of delay--cost trade-offs without modifying protocol internals. Experiments conducted using the Mininet-WiFi emulator show that the proposed approach characterizes a wide delay--cost region and identifies configurations that achieve substantial LTE savings (up to 80%) with controlled increases in upload time. The results further indicate that, under higher contention levels, systematic multi-objective exploration provides increased flexibility compared to fixed-policy schedulers in cost-aware heterogeneous uplink deployments.
Summary / 总结
Multipath QUIC (MPQUIC) enables simultaneous uplink transmission over heterogeneous access networks such as Wi-Fi and LTE, improving reliability and performance.
ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover
Authors: Minwoo Kim, Soochang Song, Namyoon Lee, Bang Chul Jung, Yongjune Kim
First: 2026-08-11T06:37:10+00:00 · Latest: 2026-08-11T06:37:10+00:00
Abstract
Edge LLMs must preserve inference continuity when a user hands over between edge nodes, requiring key-value (KV) cache transfer to the target node. However, simultaneous handovers saturate the backhaul, preventing full cache delivery within the mobility-imposed transfer window. Rather than allocating bandwidth as if all cache entries were equally valuable, we order each user's KV cache by importance and transmit only its most informative fraction, turning token-level sparsity into communication savings. We cast the transfer as a multi-user backhaul allocation problem that maximizes average accuracy across users. Each user's partial-cache accuracy serves as its utility: a sigmoid that fits measurements on the RULER benchmark with $R^2>0.99$ across models and context lengths. Because importance ordering front-loads the high-value entries, the concave region of the accuracy curve spans nearly the entire cache. Our proposed allocator keeps served users within this region, making each per-slot allocation problem convex. The optimum is derived via a closed-form weighted water-filling solution that generalizes information-theoretic water-filling and enables online scheduling. The proposed allocator attains over 93.7% average accuracy in a 500ms transfer window, within 0.5pp of the full-cache ceiling, and reaches 98.2-99.5% of a clairvoyant upper bound.
Summary / 总结
Edge LLMs must preserve inference continuity when a user hands over between edge nodes, requiring key-value (KV) cache transfer to the target node.
Benchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxy
Authors: Aman Chauhan, Vishnu Pendyala
First: 2026-08-11T06:15:20+00:00 · Latest: 2026-08-11T06:15:20+00:00
Comments: 43 pages, 6 figures, 15 tables. Submitted to Journal of Network and Computer Applications (Elsevier)
Abstract
Static load balancers cannot mitigate a backend that is degraded rather than down: round-robin and least-connections keep routing traffic to a server returning HTTP 500s until an operator intervenes. We ask whether a Large Language Model can replace the static routing policy itself, reading HAProxy and Prometheus telemetry every 10 seconds and isolating faulty servers through guardrailed calls to the HAProxy Data Plane API. On a reproducible benchmark with a persistent structural fault built into roughly one-third of a heterogeneous fleet, we sweep 15 open-weight models across five families (0.35B to 35B total parameters; dense, mixture-of-experts, and efficient-sparse architectures), reasoning modes, fleet scales of 3 to 9 backends, and two routing algorithms, totaling 240 runs. We find a capability threshold near 3B active parameters. Below it, LLM policies are typically unreliable and sometimes worse than no policy; above it, every model, regardless of architecture, saturates near an 88% reduction in client-perceived 5xx errors over the static baseline. The threshold is approximate: Gemma 4 E2B clears it with 2B active parameters, while the dense 3B Granite 4.0 Micro does not. The availability gain has costs. Draining concentrates load onto surviving servers, inflating tail latency 2.6 to 2.8 times, and enabling reasoning multiplies token spend roughly tenfold, overrunning the control interval and degrading effectiveness. The efficient operating point is a supra-threshold model in its cheapest non-reasoning mode, wrapped inside deterministic guardrails.
Summary / 总结
Static load balancers cannot mitigate a backend that is degraded rather than down: round-robin and least-connections keep routing traffic to a server returning HTTP 500s until an operator intervenes.
AtlasRAN: Timing-Aware Evaluation of Open-source 5G Platforms for Integrated Wireless Testbeds
Authors: Ryan Barker, Tolunay Seyfi, Alireza Ebrahimi Dorcheh, Julia Boone, Fatemeh Afghah, Joseph Boccuzzi
First: 2026-03-15T23:34:49+00:00 · Latest: 2026-08-10T17:32:28+00:00
Comments: 6 pages, 4 figures, 2 tables
Abstract
Open-source fifth-generation (5G) and Open Radio Access Network (O-RAN) experiments span simulation, host operating system (host-OS) emulation, software-defined radio hardware-in-the-loop testing, Open Radio Unit fronthaul deployments, wireless digital twins, and accelerator-backed radio access network (RAN) runtimes. These environments may expose similar interfaces while preserving different timing, input/output, synchronization, buffering, transport, and observability; functional compatibility is therefore not timing fidelity.
This paper presents AtlasRAN, a claim-to-capability framework for deciding what an open-source 5G platform can credibly measure. It combines two reference paths, an execution-regime matrix, and a measurement-backed case study comparing OpenAirInterface (OAI) radio-frequency simulator (RFSim) with the Sionna Research Kit (Sionna-RK), which offloads low-density parity-check (LDPC) decoding to CUDA while retaining the surrounding OAI host-OS path. From one to six users, aggregate uplink goodput falls from 114.59 to 35.09 Mb/s for OAI and from 103.34 to 35.01 Mb/s for Sionna-RK. At six users, both RFSim real-time factors are approximately 0.55, despite estimated cumulative LDPC work of 1.09~ms per transport block for OAI and 0.37 ms for Sionna-RK. Timing-normalized service remains 60.7-63.6~Mb per emulated second across complete runs, while falling host and accelerator utilization is consistent with an under-fed decoder. A twelve-user run is retained only as failure-region evidence because end-of-test traffic summaries are absent. The practical takeaway is that integrated wireless testbeds, edge platforms, artificial-intelligence-enabled RANs, and digital twins should report timing discipline, transport path, memory movement, and observability as first-class experimental variables.
Summary / 总结
Open-source fifth-generation (5G) and Open Radio Access Network (O-RAN) experiments span simulation, host operating system (host-OS) emulation, software-defined radio hardware-in-the-loop testing, Open Radio Unit fronthaul deployments, wireless digital twins, and accelerator-backed radio access network (RAN) runtimes.
A Bird's-Eye View on Security Considerations in RFCs
Authors: Jukka Ruohonen, Qusai Ramadan
First: 2026-08-10T17:23:42+00:00 · Latest: 2026-08-10T17:23:42+00:00
Comments: Submitted
Abstract
Request for comments (RFCs) are Internet standards, memorandums, and related technical documents about core Internet protocols made via and released by the Internet Engineering Task Force (IETF). In the early 1990s each RFC was required to have a section for security considerations. The present work examines these sections. According to the empirical results, (1) over 90% of the RFCs sampled have discussed security explicitly in these sections, (2) although mandatory security requirements have only seldom-if ever-been imposed. Furthermore, (3) the RFC-to-RFC reference network specific to the security consideration sections is sparse, although a few RFCs and their security consideration sections are heavily referenced. In addition, (4) the volume of references peaked during a period from circa mid-1990s to mid-2010s. Regarding the topics discussed in the sections, (5) these do not represent general security issues, such as spoofing or eavesdropping; rather, the topics mostly reflect distinct security issues specific to distinct protocols. With the exceptions of network security in general, security specifications, and routing, (6) also the longitudinal evolution of the topics is protocol-specific. As the subject matter has not been previously examined, these empirical results fill a gap in the standardization literature.
Summary / 总结
Request for comments (RFCs) are Internet standards, memorandums, and related technical documents about core Internet protocols made via and released by the Internet Engineering Task Force (IETF).
Quantum-Classical Coexistence Network Tomography
Authors: Xuchuang Wang, Joseph C. Chapman, Aneesh Ramaswamy, Matheus Guedes de Andrade, Yu-Zhen Janice Chen, Joseph M. Lukens, Gayane Vardoyan, Don Towsley
First: 2026-08-10T09:42:05+00:00 · Latest: 2026-08-10T09:42:05+00:00
Comments: A short version of this paper is accepted to QCE'26 under the same title
Abstract
Quantum-classical coexistence networks (QCNs) share optical fiber between quantum and classical signals via wavelength-division multiplexing, offering a practical path to quantum communication over existing telecom infrastructure. However, co- and counter-propagating classical traffic introduce distinct depolarization noise, complicating channel characterization. We develop a tomography framework that infers per-link channel parameters of a QCN from end-to-end measurements alone. We first model each coexisting fiber by decomposing the signal evolution into photon loss, successful transmission, and three direction-dependent depolarization components. We then derive closed-form link-level estimators, and extend the approach to star-topology networks through a system of multiplicative equations across end-node pairs, together with a simple classical-signal-direction-switching protocol that resolves the remaining unknowns. On single-link experimental testbed data, we recover per-link depolarization probabilities accurately, with estimated process fidelities closely tracking the Bayesian-process-tomography baseline across multiple fiber lengths and wavelengths; residual gaps reflect the depolarization-only approximation. Absent a multi-link coexistence testbed, we validate the star-network estimators on emulated paths built from measured single-link channels. We further extend the framework in two directions: (i) a channel model that factorizes the coexisting fiber into a depolarizing-with-loss signal channel and a Raman-noise-injection channel on separate optical modes -- a completely-positive, trace-preserving tensor product -- whose link observables reduce exactly to our basic model; and (ii) a generalization to arbitrary topologies via a peeling algorithm (trees) and a least-squares estimator (meshes), validated by Monte-Carlo simulations on tree and cyclic-mesh networks.
Summary / 总结
Quantum-classical coexistence networks (QCNs) share optical fiber between quantum and classical signals via wavelength-division multiplexing, offering a practical path to quantum communication over existing telecom infrastructure.
Automated Synthesis of Deterministic Cross-Domain Interfaces
Authors: Konstantinos Christodoulopoulos, Antonis Selentis-Boulntadakis
First: 2026-08-10T08:55:31+00:00 · Latest: 2026-08-10T08:55:31+00:00
Abstract
Deterministic networking spans heterogeneous domains. At each boundary, two domains must agree on an assume--guarantee contract: what traffic the client may inject, and the QoS the carrier will hold for it. Composing such contracts into an end-to-end guarantee is standardized, but deriving each domain's contract is not. Today they are hand-crafted, static, and over-provisioned. The difficulty rises when traffic changes and the contract must become dynamic. We present a framework that automatically synthesizes the per-domain contract for both static and dynamic classes, along with the registration the dynamic one rests on. Two reasoning modules implemented with large language models (LLMs) drive it: an agent takes the client's traffic declaration and searches the carrier's configuration mechanisms, and a handler builds the network model from a typed disclosure of the substrate. The handler calls a network-calculus kernel for the model's hard terms, and an independent oracle---a faithful simulator, testbed, or live network---which verifies each candidate and discovers what lacks an a-priori algebraic form: when a reconfiguration is safe, and the instant to apply it. We synthesized a dynamic contract for uplink 5G fronthaul over a TDM-PON, grounded against a packet-level simulator. Across six draws from two LLM families, every synthesis produced a feasible, verified contract tight to ${\sim}1.2\,μ$s, holding a $100$-$μ$s deadline that reactive scheduling cannot meet, at up to $3.5$ times the bandwidth efficiency of static over-provisioning. Tasked instead with computing the worst-case delay directly, the LLMs were unsound in five of six attempts---evidence for the division of labor: LLMs construct the model, formal tools hold numeric authority. The same framework, unchanged, synthesized a static 5G--TSN bridge contract on a second substrate.
Summary / 总结
Deterministic networking spans heterogeneous domains.
PSP: Low-Overhead Packet-Level Load Balancing for Stale-State and Bandwidth-Asymmetric Networks
Authors: Jiaqi Liu, Chunyang Zhang, Heng Pan, Yanbiao Li
First: 2026-08-09T02:47:03+00:00 · Latest: 2026-08-09T02:47:03+00:00
Comments: 12 pages, 10 figures, 5 tables. Accepted by IEEE LCN 2026
Abstract
With the rapid growth of large language model training and generative artificial intelligence services, data center networks face severe micro-burst traffic and high concurrency. Traditional hash-based flow-level load balancing cannot sense link states, leading to hash collisions, hotspot congestion, and tail latency in multipath Clos networks. Existing packet-level schemes are constrained by stale state information, high hardware complexity, and poor adaptation to heterogeneous links.
To address these issues, this paper proposes probabilistic state-proportional (PSP) dispatching, a packet-level load balancing algorithm. Using a Band-based discrete state representation, PSP replaces global sorting with local probability mapping, reducing hardware complexity while suppressing herding and oscillations caused by stale states.
Experiments on a cycle-accurate simulator show that PSP is robust across port scales, bandwidth-limited paths, and fixed-flow interference. It outperforms join-the-shortest-queue (JSQ) scheduling and Random in loss rate, 99th-percentile buffer occupancy, and scalability, while remaining competitive with Top-k at lower hardware cost. PSP provides an effective balance among performance, stability, and overhead for artificial intelligence data centers.
Summary / 总结
With the rapid growth of large language model training and generative artificial intelligence services, data center networks face severe micro-burst traffic and high concurrency.
WirelessOpsAgent: A Benchmark and Agent Design for Action Assurance in Wireless Networks
Authors: Zijian Lu, Yiping Zuo, Hao Xu, Weicong Chen, Xin He, Jiajia Guo, Shi Jin
First: 2026-08-08T18:12:15+00:00 · Latest: 2026-08-08T18:12:15+00:00
Comments: 10 pages, 7 figures
Abstract
Large language model (LLM) agents are emerging as planners for autonomous wireless network operations. Yet a task answer that is correct at proposal time can still be unsafe at execution time if supporting telemetry is stale or inconsistent. Existing benchmarks mainly evaluate task solving from fixed observations and leave support checking at execution time untested. We introduce WirelessOptBench, a benchmark for action assurance in wireless operations. It turns wireless tasks into execution state decision episodes with controlled telemetry faults and action constraints. We further develop WirelessOpsAgent, which grounds candidate actions in current evidence and repairs recoverable support failures before execution. Across three backbone evaluations with 600 episodes each, WirelessOpsAgent achieves up to 0.983 Exact Action Accuracy. On Claude Sonnet 4.6, the Unsafe APPLY Rate decreases from 82.2% to 10.3% relative to the safest baseline. We make WirelessOptBench available at https://anonymous.4open.science/r/wirelessopsbench-artifact-D969/.
Summary / 总结
Large language model (LLM) agents are emerging as planners for autonomous wireless network operations.
MAC-Gyver: Open, Programmable, Scheduling for AI-RAN 6G Systems
Authors: Maxime Elkael, Reshma Prasad, Tamerlan Aghayev, Salvatore D'Oro, Michele Polese, Tommaso Melodia
First: 2026-07-28T17:24:53+00:00 · Latest: 2026-08-08T11:31:47+00:00
Abstract
Cellular networks are integrating Artificial Intelli- gence (AI) into radio access network control. The MAC scheduler is a promising target because it allocates a limited resource, spectrum, at every slot, under competing latency, throughput, and reliability requirements. However, most learning-based sched- ulers are evaluated only in simulation. Production schedulers are difficult to modify, and realistic stress tests require more radio hardware than most laboratories can provide. We present MAC-Gyver, an open-source framework for developing and evaluating scheduling applications that execute directly inside the OpenAirInterface scheduler. It exposes scheduler observations and controls through typed interfaces while preserving the underlying protocol and real-time execution paths. The same applications run over the air and in mac-emu, a PHY-less emulator that executes the unmodified OpenAirInterface Layer 2 stack for up to 90 users on one host at real-time slot pace, with a 3GPP-compliant channel model. To showcase the flexibility of MAC-Gyver, we evaluate two use cases. A proactive uplink scheduler predicts packet arrivals and roughly halves median round-trip latency. A frequency-selective uplink scheduler selects contiguous sub-bands from per-PRB sounding observations and is evaluated across mobility and power-limited operating points against an offline scheduling ceiling. Together, they show how the same production stack can be an AI playground that supports implementation, controlled evaluation, and over-the-air validation through complementary scheduling use cases.
Summary / 总结
Cellular networks are integrating Artificial Intelli- gence (AI) into radio access network control.
Rate-Fidelity Control for Wide-Area Quantum Links
Authors: Connor Clayton, Cory Nunn, Quinn Carmack, Wayne McKenzie, Anne Marie Richards, Xiaodi Wu, Bobby Bhattacharjee
First: 2026-08-07T12:32:13+00:00 · Latest: 2026-08-07T12:32:13+00:00
Comments: 27 pages, 12 figures, 1 table
Abstract
Quantum network links must distribute entanglement at high rates while satisfying application-specified fidelity demands. However, wide-area deployed fiber links suffer from polarization drift which destabilizes end-to-end fidelity and forces periodic compensation. Current deployments often use active stabilization with fixed control policies, and improvements generally stem from advances in quantum hardware. Meanwhile, software control remains relatively underexplored.
Here, we formulate quantum link operation as a joint control problem over tunable rate-fidelity tradeoffs and uncontrollable link drift. From this framework, we construct a link control protocol that dynamically adapts source pump power and polarization compensation to maximize entanglement distribution rate subject to a minimum fidelity constraint. We evaluate the protocol through trace-driven simulations driven by data from a 64 km deployed optical fiber. Compared with optimized static policies, our adaptive controller improves mean entanglement distribution rate by 14% over a 24 hour trace, without requiring any offline policy optimization. Our results show that software-based physical layer control can provide a practical mechanism for improving near-term quantum link performance without requiring additional quantum hardware.
Summary / 总结
Quantum network links must distribute entanglement at high rates while satisfying application-specified fidelity demands.
EvoRIC: Reinforcement Learning Fine-Tuned LLM-empowered RAN Intelligent Control Toward Autonomous O-RAN
Authors: Lingyan Bao, Jemin Lee, Tony Q. S. Quek
First: 2026-08-07T04:13:20+00:00 · Latest: 2026-08-07T04:13:20+00:00
Comments: Manuscript submitted 23 April 2026; revised 7 August 2026
Abstract
Despite recent advances in applying artificial intelligence (AI) techniques to radio access network (RAN), critical challenges remain: traditional machine learning (ML) algorithms suffer from limited generalization across varying network topologies, whereas general-purpose large language models (LLMs) face high computational demands and lack domain-specific knowledge. To address these gaps, this article introduces the evolving RAN intelligent controller (RIC) (EvoRIC) framework, a hierarchical architecture that enables continuous evolution by leveraging a non-real-time RIC (non-RT RIC) for global model updates and a near-real-time RIC (near-RT RIC) for local execution, dynamically empowering LLMs with domain-specific decision-making capabilities. Within this framework, we employ a reinforcement learning-based fine-tuning (RLFT) mechanism where an LLM operates as an actor within a proximal policy optimization (PPO) agent. By leveraging the interaction tuples collected from the wireless environment, the LLM's parameters are iteratively updated to align semantic reasoning with rigorous network performance objectives. We evaluate the generalization and efficacy of the proposed EvoRIC framework within integrated access and backhaul (IAB) networks, and finally, discuss the open challenges and future directions of the EvoRIC framework toward realizing autonomous O-RAN.
Summary / 总结
Despite recent advances in applying artificial intelligence (AI) techniques to radio access network (RAN), critical challenges remain: traditional machine learning (ML) algorithms suffer from limited generalization across varying network topologies, whereas general-purpose large language models (LLMs) face high computational demands and lack domain-specific knowledge.
A Parameter-Specific Retrieval and Knowledge-Guided Reasoning Framework for LLM-Based GPSR Optimization in FANETs
Authors: Zhipeng Lin, Bin Duo, Tong Liu, Jie Lin, Jianting Yuan, Xiaojun Yuan
First: 2026-08-07T03:26:58+00:00 · Latest: 2026-08-07T03:26:58+00:00
Abstract
Existing Greedy Perimeter Stateless Routing (GPSR)-based protocols for Flying Ad-Hoc Networks (FANETs) struggle to adapt routing parameters, such as hello interval, multi-path number, and greedy forwarding weights, under highly dynamic environments. As an emerging artificial intelligence technology, large language models (LLMs) show potential for intelligent decision-making, providing new opportunities for adaptive adjustment of GPSR parameters to improve network performance. However, applying LLMs to GPSR remains challenging due to irrelevant experience retrieval and the absence of protocol constraints. To address these issues, we propose a Parameter-Specific Multi-Index Retrieval and Knowledge-Guided Reasoning framework for adaptive GPSR optimization (PMKR-GPSR), an LLM-based framework that enables protocol-consistent routing parameter adaptation. We design a parameter-specific multi-index retrieval mechanism to provide LLMs with parameter-relevant experiences while reducing interference from irrelevant information. We further construct a knowledge-guided constraint graph to enforce that the routing parameters satisfy dependency rules and optimization constraints. Simulation results demonstrate that PMKR-GPSR achieves higher packet delivery ratio and lower end-to-end delay under high-mobility FANETs.
Summary / 总结
Existing Greedy Perimeter Stateless Routing (GPSR)-based protocols for Flying Ad-Hoc Networks (FANETs) struggle to adapt routing parameters, such as hello interval, multi-path number, and greedy forwarding weights, under highly dynamic environments.
MultiMoQ: Multi-Access Media-Over-QUIC for Robust Immersive Video Streaming
Authors: Yitong Li, Xinjiao Li, Ruonan Chai, Dirk Kutscher
Venue: ACM MM 2026
First: 2026-08-06T14:38:48+00:00 · Latest: 2026-08-06T14:38:48+00:00
Comments: 9 pages, 15 figures. Accepted for publication in the Proceedings of the 34th ACM International Conference on Multimedia (ACM MM 2026)
Abstract
Live immersive video streaming, particularly 360-degree video, is increasingly adopted in applications such as virtual events, sports broadcasting, and remote education. Existing approaches struggle to support high-bitrate immersive streaming for large numbers of concurrent users, with coarse-grained delivery limiting responsiveness and insufficient support for coordinating concurrent tile streams. Media over QUIC (MoQ) has recently emerged as a promising solution for large-scale media delivery, yet it lacks robustness under bandwidth-constrained conditions, often resulting in playback stalls. To address these challenges, we present MultiMoQ, a multi-access tile streaming framework built on MoQ that redesigns its delivery mechanism for robust high-bitrate streaming across multiple access paths while supporting flexible tile scheduling and seamless access switching. We implement a fully functional prototype of MultiMoQ and evaluate it in network emulation under heterogeneous real-world network conditions, comparing against Dynamic Adaptive Streaming over HTTP (DASH) and standard MoQ. Results show that MultiMoQ increases goodput for enhancement tiles and base video and reduces enhancement-tile tail end-to-end latency relative to DASH, while preserving audio continuity and avoiding the persistent stalls of standard MoQ. These transport gains translate into smoother viewport playback, and the ablation results further confirm the contribution of multi-access control to playback continuity.
Summary / 总结
Live immersive video streaming, particularly 360-degree video, is increasingly adopted in applications such as virtual events, sports broadcasting, and remote education.
MARS: Multipath Adaptive Reliable Service
Authors: Yitong Li, Xinjiao Li, Dirk Kutscher
First: 2026-08-06T14:38:13+00:00 · Latest: 2026-08-06T14:38:13+00:00
Comments: 15 pages, 11 figures, 2 tables. Accepted at the 34th IEEE International Conference on Network Protocols (ICNP 2026)
Abstract
Multipath transport is increasingly important for Internet/WAN services that move large data volumes across heterogeneous paths, including geo-distributed analytics, content distribution, and cloud-service pipelines. Existing solutions, however, face a practical trade-off: end-to-end transports such as MPTCP and MPQUIC are deployable but limited by endpoint-visible paths and delayed congestion feedback, while routing-or forwarder-assisted approaches often require infrastructure support or lack safe coordination across forwarding choices. This paper presents MARS, a receiver-driven, forwarder-assisted multipath transport for Internet/WAN environments. MARS combines tier-synchronized overlay path discovery with coupled consumer/forwarder congestion control, enabling it to safely expand usable forwarding opportunities and react near bottlenecks. It runs as an incrementally deployable UDP overlay at clients, servers, relays, or CDN-like nodes. We implement MARS in simulation and as a prototype, and evaluate it through large-scale simulation and Mininet emulation under different deployment scales, loss rates, and failure scenarios. The results show that MARS provides deployment-dependent benefits: with endpoint-only deployment, it remains competitive with end-to-end multipath baselines; with cooperating overlay forwarders, it exposes richer usable path diversity and reduces max p95 flow completion time by up to 81.5\% over ECMP-limited baselines. Even against path-expanded end-to-end baselines given the same path set, MARS achieves lower worst-case p95 FCT and stronger robustness under packet loss, while also recovering quickly from transient link failures. These results demonstrate that ICN-style receiver-driven forwarding can serve as a deployable overlay transport substrate for WAN multipath, providing benefits beyond purely end-to-end designs without requiring changes to IP routing.
Summary / 总结
Multipath transport is increasingly important for Internet/WAN services that move large data volumes across heterogeneous paths, including geo-distributed analytics, content distribution, and cloud-service pipelines.
BALANCE: Hybrid Autoregressive-Speculative LLM Inference in Wireless Edge Networks
Authors: Guanqiao Qu, Shuo Chen, Qian Chen, Kin K. Leung, Xianhao Chen
First: 2026-08-06T11:57:36+00:00 · Latest: 2026-08-06T11:57:36+00:00
Comments: 10 pages, 7 figures
Abstract
Edge inference is a promising paradigm to provide large language model (LLM) inference services in next-generation mobile networks. LLM inference mainly relies on two approaches: Autoregressive decoding (AD) generates output tokens sequentially, resulting in long latency; Speculative decoding (SD) accelerates inference by using a small language model (SLM) to generate multiple draft tokens for LLM verification, but incurs extra memory costs. Due to this latency-memory tradeoff, neither approach alone can efficiently serve users with heterogeneous demands under limited edge computing resources. To address this challenge, we propose a hybrid autoregressive-speculative inference (BALANCE) framework for edge LLM inference. In BALANCE, an edge server hosts both an SLM and an LLM, assigns each user to AD or SD, and performs the two modes simultaneously. To maximize the number of served users, we formulate a task throughput maximization problem to jointly determine user scheduling and computing resource allocation between AD and SD under user latency requirements and server memory constraints. Since the problem is NP-hard, we develop a polynomial-time algorithm that transforms the original problem into two sub-problems and obtains a sub-optimal solution with a constant approximation guarantee. Experiments demonstrate that BALANCE consistently outperforms conventional AD and SD and significantly improves task throughput.
Summary / 总结
Edge inference is a promising paradigm to provide large language model (LLM) inference services in next-generation mobile networks.
Statistical Verification of Medium-Access Parameterization for Power-Grid Edge Ad Hoc Sensor Networks
Authors: Haitian Wang, Xinyu Wang, Zichen Geng, Xian Zhang, Yiren Wang, Yihao Ding
First: 2026-02-05T10:08:40+00:00 · Latest: 2026-08-05T19:39:30+00:00
Comments: 6 pages, 1 figure, 2 tables. Accepted and presented at the 31st IEEE Symposium on Computers and Communications (IEEE ISCC 2026). Camera-ready version submitted; proceedings publication pending
Abstract
The widespread deployment of power grid ad hoc sensor networks based on IEEE 802.15.4 raises reliability challenges when nodes selfishly adapt CSMA/CA parameters to maximize individual performance. Such behavior degrades reliability, energy efficiency, and compliance with strict grid constraints. Existing analytical and simulation approaches often fail to rigorously evaluate configurations under asynchronous, event-driven, and resource-limited conditions. We develop a verification framework that integrates stochastic timed hybrid automata with statistical model checking (SMC) with confidence bounds to formally assess CSMA/CA parameterizations under grid workloads. By encoding node- and system-level objectives in temporal logic and automating protocol screening via large-scale statistical evaluation, the method certifies Nash equilibrium strategies that remain robust to unilateral deviations. In a substation-scale scenario, the certified equilibrium improves utility from 0.862 to 0.914 and raises the delivery ratio from 89.5% to 93.2% when compared with an aggressive tuning baseline. Against a delivery-oriented baseline, it reduces mean per-cycle energy from 152.8 mJ to 149.2 mJ while maintaining comparable delivery performance. Certified configurations satisfy latency, reliability, and energy constraints with robustness coefficients above 0.97 and utility above 0.91.
Summary / 总结
The widespread deployment of power grid ad hoc sensor networks based on IEEE 802.15.4 raises reliability challenges when nodes selfishly adapt CSMA/CA parameters to maximize individual performance.
Introducing Large Language Models into the Design Flow of Time-Sensitive Networking
Authors: Rubi Debnath, Luxi Zhao, Mohammadreza Barzegaran, Paul Pop, Sebastian Steinhorst
First: 2025-09-30T15:04:24+00:00 · Latest: 2026-08-05T11:55:20+00:00
Abstract
The growing demand for real-time, safety-critical systems has significantly increased both the adoption and complexity of Time-Sensitive Networking (TSN). Configuring an optimized TSN network is highly challenging, requiring careful planning, design, analysis, verification, validation, and deployment. Large Language Models (LLMs) have recently demonstrated strong capabilities in solving complex tasks, positioning them as promising candidates for automating end-to-end TSN deployment and management, referred to as TSN orchestration. This paper outlines the steps involved in TSN orchestration and the associated challenges. To assess the capabilities of existing LLMs, we conduct an initial proof-of-concept case study focused on TSN tasks across multiple models. Building on these insights, we propose an LLM-assisted orchestration framework. Unlike prior research on LLMs in computer networks, which has concentrated on general configuration and management, TSN-specific orchestration has not yet been investigated. We present the building blocks for automating TSN using LLMs, describe the proposed pipeline, and analyze opportunities and limitations for real-world deployment. This work provides the first roadmap toward assessing the feasibility of LLM-assisted TSN orchestration.
Summary / 总结
The growing demand for real-time, safety-critical systems has significantly increased both the adoption and complexity of Time-Sensitive Networking (TSN).