RunInfraby RightNow
  • CatalogNew
  • Pricing
  • Research
  • Contact
DashboardSign inGet started
Home/Glossary/NVLink

Hardware

NVLink

What it is

NVLink is a high-bandwidth connection used between supported accelerators in the same system or fabric. Collective operations can move tensor shards and partial results across these links.

Why it moves cost and latency

Tensor-parallel latency depends on repeated communication as well as local computation. Without direct links, collectives use the available host or network path, whose bandwidth and topology may differ.

What it looks like in practice

Operators map parallel groups onto connected devices and inspect the actual communication topology. They measure collective time because link presence alone does not guarantee balanced or contention-free traffic.

Related terms

  • Tensor parallelism->
  • Pipeline parallelism->
  • Tensor versus pipeline parallelism->
  • Expert parallelism->

Questions this definition answers

What does NVLink mean in inference serving?

NVLink is a high-bandwidth connection used between supported accelerators in the same system or fabric. Collective operations can move tensor shards and partial results across these links. Operators map parallel groups onto connected devices and inspect the actual communication topology. They measure collective time because link presence alone does not guarantee balanced or contention-free traffic.

Why can NVLink move cost or latency?

Tensor-parallel latency depends on repeated communication as well as local computation. Without direct links, collectives use the available host or network path, whose bandwidth and topology may differ.

If you need custom optimization for a specific model, describe what you need

Describe the model and hardware you want optimized...
ModelsAuto engineAuto GPU
End-to-end encryption
Isolated GPU infrastructure
No training on your data
SOC 2 Type II
RunInfraby RightNow

© 2026 RunInfra. All rights reserved.

System status
Pipeline BuilderModelsCost CalculatorPricingStartupsBenchmarksDocsResearchNewsContact
Backed by
YCombinator
AICPA Type II
SOC 2
NVIDIA Inception ProgramNVIDIA Inception Program
Ask AI about RunInfra
Part of RightNow
SecurityDPAAUPCookiesTermsPrivacy