AI Accelerators

NVIDIA

NVIDIA L4 24GB

Model L4

A low-power, low-profile card for adding inference and video capacity to servers that have no spare power or slot height.

Quantity

No prices online. Add what you need and we reply with an all-in price and a delivery date.

Close-up of a blue accelerator card circuit board
Representative photo of the hardware type, not the listed product.

// NVIDIA L4 24GB specifications

ArchitectureAda Lovelace
GPU memory24GB GDDR6
Memory bandwidth300GB/s
Host interfacePCIe Gen4 x16
Form factorSingle-slot, low-profile
Max power72W

Specifications are indicative and confirmed on quote. Product and brand names belong to their respective owners and are used for identification only.

Before you order: AI Accelerators

The questions we ask before quoting. Answer them in the line notes and the first reply is usually the final one.

  • Module parts (SXM, OAM) only fit a matching baseboard and chassis. Tell us the host platform, or quote the server alongside them.
  • Passive PCIe cards rely on server airflow. Check the chassis is rated for the power of the card before ordering.
  • For multi-GPU work, say whether you need NVLink bridges or a switched fabric. It changes what gets quoted.
  • Memory size usually decides the model: size it against the largest model and batch you plan to run.

More in AI Accelerators

All AI Accelerators

NVIDIA

AI Accelerators

NVIDIA H200 SXM 141GB

Hopper-generation accelerator with the largest memory pool in its family, suited to serving and fine-tuning large language models that outgrow 80GB parts.

Architecture
Hopper
GPU memory
141GB HBM3e
Memory bandwidth
4.8TB/s

NVIDIA

AI Accelerators

NVIDIA H100 SXM 80GB

The reference training accelerator for multi-GPU nodes, deployed eight to a baseboard with full NVLink bandwidth between every GPU.

Architecture
Hopper
GPU memory
80GB HBM3
Memory bandwidth
3.35TB/s

NVIDIA

AI Accelerators

NVIDIA H100 NVL 94GB

A PCIe card version of Hopper with extra memory, built for LLM inference in standard air-cooled servers and bridged in pairs over NVLink.

Architecture
Hopper
GPU memory
94GB HBM3
Memory bandwidth
3.9TB/s