Home Knowledge Base RDMA (Remote Direct Memory Access) and InfiniBand

RDMA (Remote Direct Memory Access) and InfiniBand are the high-performance networking technologies that enable direct memory-to-memory data transfer between machines without involving the CPU or operating system — achieving latencies under 1 microsecond and throughputs over 400 Gbps, making them essential for HPC clusters, distributed training, and low-latency storage systems.

How RDMA Works

RDMA Operations

OperationDescriptionCPU Involvement
RDMA WriteWrite to remote memoryNone on remote side
RDMA ReadRead from remote memoryNone on remote side
Send/ReceiveTwo-sided messagingBoth sides post buffers
Atomic (CAS, FetchAdd)Atomic operation on remote memoryNone on remote side

RDMA Transports

TransportFabricBandwidthLatencyDeployment
InfiniBand (IB)Dedicated IB fabricHDR: 200 Gbps, NDR: 400 Gbps< 0.6 μsHPC, AI clusters
RoCE v2Standard Ethernet25-400 Gbps1-3 μsData centers
iWARPStandard Ethernet (TCP)10-100 Gbps5-10 μsEnterprise storage

InfiniBand Generations

GenerationPer-Lane Rate4x PortYear
QDR10 Gbps40 Gbps2008
FDR14 Gbps56 Gbps2012
EDR25 Gbps100 Gbps2015
HDR50 Gbps200 Gbps2019
NDR100 Gbps400 Gbps2022
XDR200 Gbps800 Gbps2024

RDMA in Distributed ML Training

Programming RDMA

RDMA and InfiniBand are the networking foundation of modern AI supercomputers — the ability to move data between machines at hardware speed without CPU involvement is what makes it possible to train trillion-parameter models across thousands of GPUs with near-linear scaling efficiency.

rdma infinibandremote direct memory accessrdma networkingib verbsroce rdma

Explore 500+ Semiconductor & AI Topics

From EUV lithography to CUDA optimization — search the full knowledge base or chat with our AI assistant.