Analyzing Reverse Address Translation Overheads in Multi-GPU Scale-Up Pods
arXiv 2604.02473•d0a6710cdfe150a466aebfdea762fb9f04db503028c30d1d01ade7855762f8c0
ASTRA-simCIDERCPU-GPU interconnectDawnKV cache sharingLink MMULink TLBMetalNVLinkOmnet++TLB prefetchingTokenDanceUALinkVulkanWebGPUdispatch overheadfused pre-translationheterogeneous memory managementkernel fusionmemory-disaggregated KV storesmulti-GPUpessimistic synchronizationreverse address translationvLLMwgpu-native
Paper metadata
- arXiv ID
- 2604.02473
- Version
- Not specified by this published record
- Category
- Computer Science — Distributed, Parallel, and Cluster Computing (cs.DC)
The PDF link points to arxiv.org. Baitaphish does not expose a private stored PDF.
Evidence and limitations
- Source ID
- arxiv_cs_dc
- Record identifier
- d0a6710cdfe150a466aebfdea762fb9f04db503028c30d1d01ade7855762f8c0
- Enrichment time
- 2026-04-06T08:52:24Z
- AI-assisted enrichment
- Yes
This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.