Program Details
The detailed technical program below is a draft schedule and will be updated as sessions, chairs, and rooms are finalized. All times are in U.S. Mountain Time (UTC-7).
Tuesday, August 18
| 11:00 am – 12:25 pm | |
|---|---|
| Track A | Track B |
Research Session 1: LLM Inference & Serving KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving DualPath: Accelerating Agentic LLM Inference by Harvesting Disaggregated KV-Cache Storage I/O Efficient Remote KV Cache Reuse with GPU-native Video Codec Connex: Endpoint Mobility Primitives for Dynamic LLM Serving TurboBus: Pooling PCIe Bandwidth for LLM Workloads via Scale-Up Fabrics | Research Session 2: Network Verification & Formal Methods Towards Efficient Verification of Distributed In-Network Computing Programs When static verification is not enough: revealing BGP bugs at runtime Explainable Network Verification via Localized Subspecification VeriLucid: A Verification-aware Data-plane Programming Language Elastispec: Formalizing Enterprise Firewall Management with Informal and Elastic Specifications |
| 1:50 pm – 3:15 pm | |
|---|---|
| Track A | Track B |
Research Session 3: Collective Communication Algorithms Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks OptCCL: Scalable Synthesis of Optimal Collective Communication Algorithms DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training Theseus: Runtime-Adaptive GPU Collective Communication with Hot-Swappable Schedules | Research Session 4: Cellular & RAN Systems CausalTune: Causal Learning based Automated Cellular RAN Configuration Tuning Framework SAGE: A Real-Time AI System for Reducing Latency in NextG Cellular Networks RANPilot: Making AI Functionalities Robust to Dynamic O-RAN Reconfigurations Synchronizing with the Scheduler: Dual-Loop Congestion Control for 5G Uplink on Commodity Devices Unveiling Low-Altitude 5G Performance: Linking Key Influencing Factors with UAV Flight Parameters |
Wednesday, August 19
| 8:30 am – 9:55 am | |
|---|---|
| Track A | Track B |
Research Session 5: In-Network Aggregation for ML Turbo: Efficiently Serving Long-Context Large Language Models with In-Network Aggregation HyNA: Taming Tail Latency in MoE Training with Hybrid Switch Silicon EPIC: Abstraction and Polymorphism of In-Network Collectives on Ethernet PReCCL: Performant and Resilient Collective Communication via Integrated Inband Telemetry and Workload Reallocation UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods | Research Session 6: Cloud Networking: DPUs, SmartNICs & Gateways Dorado: Scaling SmartNIC Session Tables on Commodity DDRs XFir: Accelerating New-Flow Setup on Host Servers of a Large Cloud Network FlowTurbo: From Best-Effort to Hit-Driven MegaFlow Hardware Offloading in Open vSwitch CubeTrace: Microscopic Network Tracing for Heterogeneous Cloud Gateways Rethinking Cloud Optimization: Volatility-Driven for Better Outcomes |
| 10:10 am – 11:35 am | |
|---|---|
| Track A | Track B |
Experience Session 1: AI Datacenter Networks & Communication Connecting 100K+ GPUs: Building the Communication Stack for Large-Scale LLM Training DistDPU: A Disaggregated DPU Architecture for High-Performance and Cost-Efficient AI Clouds Pegasus: A Data Center Network for Bare-Metal AI Cloud Balancing and Beyond: Communication-Centric Optimizations in Expert Parallelism Sponsored Session: Meta Talk title TBD | Research Session 7: Wireless, Backscatter & Sensing FlowForm: Scalable Passive Metasurface Network for mmWave Coverage Expansion Concurrent OFDM Backscatter with a Single Commercial Receiver Deep-Soil Acoustic Backscatter Networking for Electrical Substation Grounding Assessment Concord: Airtime-Aware Contention Control for Taming Tail Latency from Wi-Fi Frame Bursting LITE: Loss-resilient Immersive Telepresence with Multi-modal Semantics |
| 1:50 pm – 3:00 pm | |
|---|---|
| Track A | Track B |
Research Session 8: Scale-Up Fabrics: Measurement & Design FabricPerf: Measuring NIC-less Scale-Up Network through GPU Communication Kernel Profiling Efficient and Flexible Datapaths for Fine-Grained Rack-Scale Interconnects with Elastic QP Balanced Sparse Tree: A Scalable Network Topology for Large Language Models Understanding and Profiling the Accelerator Chiplet Network Using PingPoint | Experience Session 2: Content Delivery & Video Streaming Prefetching for Short Video Streaming: Experiences from a Longitudinal Evolution at Planetary Scale Adaptive Bitrate Live Streaming over HTTP-FLV: A Practical System Perspective Horizon: A Hyper-Edge Observability Engine for Live Streaming Networks CacheFlare: Optimizing Cold Content Performance in CDNs |
| 3:10 pm – 4:20 pm | |
|---|---|
| Track A | Track B |
Experience Session 3: Diagnosis & Root-Cause Analysis AIDA: Accelerating Root Cause Analysis for Multi-Vendor Device Failures with LLM-Powered Reasoning Networked Agent Memory and Causality Representation: Experiences towards Interpretable Cloud-Scale Root-Causing Detection and Localization of End-to-end Bitflip Errors in Data Centers Anytest: Localizing the Root Cause of Hardware Transport Performance Anomalies | Experience Session 4: Cloud Data Planes & Network Virtualization Rules Offload Engine (ROE): Accelerating Host SDN Policy Evaluation Spillway: Orchestrating DPU and Host into a Unified vSwitching Fabric From Nimitz to NetPila: The Evolution of Production-Scale Container Network Open the Floodgates in a Digital Twin: Experiences of Building Spillway for 100M+-User Signaling Storms in Cellular Core Network |
Thursday, August 20
| 9:00 am – 10:25 am | |
|---|---|
| Track A | Track B |
Research Session 9: Optical & Photonic Fabrics λλ: A Programming Language for Silicon Photonics Opus: Photonic Rail-Optimized Fabric in ML Datacenters Special Session 1: Best of CCR Best of CCR paper #1 Best of CCR paper #2 Additional Best of CCR selections TBD | Research Session 10: Internet Measurement & Routing OmniPath Ping: Active Network Measurement in the Era of Packet Spraying HERMES: Repurposing User-Driven Speed Tests to Monitor the Internet A Global Inference and Assessment of Large Shared IP Addresses LARS: Keeping Local Traffic Local with Latency-Aware Route Servers at IXPs Near-optimal Online Traffic Engineering |
| 10:50 am – 12:15 pm | |
|---|---|
| Track A | Track B |
Research Session 11: Scheduling for ML Training Clusters MonkeyTree: Near-Minimal Congestion for Multi-tenant Training via Migration Aegis: Contract-Bounded Online Adaptation for Networked Accelerator Clusters LEVELLER: Fair Communication Scheduling via Progress-Rate Awareness in Multi-Tenant Training Clusters GeoOrchestra: Orchestrating Heterogeneous Geo-Distributed Training with Network-Aware Scheduling Dynamic Compute and Network Orchestration for Disaggregated RL | Research Session 12: Host Networking & Packet Processing Understanding Host Network Stack Latency Don't Stall Me Now: Hiding Memory Latency in eBPF PacketExpress: Fully Exploiting Large MTUs for Internet Traffic in Private Networks Honey, I Shrunk the Headers With Flow.ZIP Simplifying Prioritization and Scheduling with P2CS |
| 3:05 pm – 4:30 pm | |
|---|---|
| Track A | Track B |
Research Session 13: Datacenter Transport: RDMA & Congestion Control STORM: Enabling Traffic Scheduling for RDMA PSN-PATH: When Multipath RDMA Meets Lossy Networks InfiniFlow: Decoupling Virtual Channel Scalability from Buffer Requirements in Lossless Datacenter Networks Odin: Rethinking Congestion Control under All-to-All Traffic CSIG: Congestion Signaling for Datacenter Transports | Research Session 14: Security, Privacy & Networked Systems Towards High-Performance Intrusion Detection with Robustness Guarantees on Programmable Switches at ISP Scale Zero-Knowledge Cloud Analytics DeepSFU: Scalable Deepfake Detection for Video Conferencing Integrating 2PC with Consensus for Fast Replication Tuning into the Web: A Low-Cost Access Solution for Developing Countries |
Friday, August 21
| 8:30 am – 9:55 am | |
|---|---|
| Track A | Track B |
Research Session 15: Programmable Switches & Data-Plane Hardware Presto: A Match-Action TCP Stack for the Terabit Era OBM: Optimal Shared Packet Buffer Management in Switches Scale-up PIFO: Interleaving Multiple Priority Queues for High Speed Programmable Scheduling Trie-Structure-Guided Compression, Allocation, and Mapping for Storage-Efficient IPv6 Lookup Pipelines Capybara: Dynamic Load Balancing with Microsecond-Scale TCP Migration | Research Session 16: Satellite & Non-Terrestrial Networks CommSAR: Enabling Bidirectional Communication in SAR Imaging Satellites via Shared Waveform Planet-Scale IoT Connectivity via LEO Satellites Dissecting the StarLink: Characterizing Queuing and Flow Dynamics in the Starlink Network Achieving Efficient Storage and Communication via Collaboration |
| 10:10 am – 11:35 am | |
|---|---|
| Track A | Track B |
Research Session 17: Congestion, Rate & Schedule Control Improving Evaluation of Heterogenous Congestion Control Algorithm Interactions Forewarned is Forearmed: A Responsive Congestion Control with Non-intrusive Uplink Dynamics Capture Queueless and Dropless Rate Control Credit-Guided Congestion Control on Wafer-Scale On-Chip Networks for Molecular Dynamics Harvest: Adaptive Photonic Switching Schedules for Collective Communication in Scale-up Domains | Research Session 18: Learning-Based & Data-Driven Network Systems Nüwa: A Generative Control Plane for AI Network Simulation EMA: Efficient Model Adaptation for Learning-based Systems RepLLM: Toward Automatically Reproducing Network Research Results Artic: AI-oriented Real-time Communication for MLLM Video Assistant Critical Path Guided Decision Making with CALLIGATOR |
| 11:45 am – 12:55 pm | |
|---|---|
| Track A | Track B |
Experience Session 5: WAN, Configuration & Verification GGN: Experiences in Designing and Deploying the Next-Generation Google Global Network Achieving Network Efficiency Through Service Collaborative Capacity Sharing and Enforcement Verifying Non-Deterministic Convergence on a Global Production WAN Evolution of AliYANG: Model-driven and LLM-assisted Network Configuration Management | Experience Session 6: Deployment Experiences: Security, Transport & Core Comprehensive Revocation Checking at Scale: the Deployment of CRLite in Mozilla Firefox Rethinking Transparent TCP Replacement: Practical Lessons from SMC-R in the Cloud Planogram: A Multi-dimensional Physical Location Planning System for DC Networks Gryphon: Scaling Hyperscale Multi-Tenant Gateways Beyond the Petabit-Era via DPU-Augmented Hierarchical Co-Offloading |