- Generative AI
- LLM inference
- LLM fine-tuning and small-model training
- NVIDIA Omniverse™ Enterprise
- Rendering and 3D graphics
- Streaming and video content
Unparalleled AI and graphics performance for the data center.
Generative AI is fueling transformative change, unlocking a new frontier of opportunities for enterprises across every industry. To transform with AI, enterprises need more compute resources, greater scale, and a broad set of capabilities to meet the demands of an ever-increasing set of diverse and complex workloads.
The NVIDIA L40S GPU is the most powerful universal GPU for the data center, delivering end-to-end acceleration for the next generation of AI-enabled applications—from gen AI, LLM inference, small-model training and fine-tuning to 3D graphics, rendering, and video applications.
| Processor |
| Graphics processor family |
NVIDIA |
| Graphics processor |
L40S |
| CUDA |
Yes |
| CUDA cores |
18176 |
| Parallel processing technology support |
Not supported |
| Memory |
| Discrete graphics card memory |
48 GB |
| Graphics card memory type |
GDDR6 |
| Memory bandwidth (max) |
864 GB/s |
| Ports & interfaces |
| Interface type |
PCI Express x16 4.0 |
| DisplayPorts quantity |
4 |
| DisplayPort version |
1.4a |
| Performance |
| Peak Single Precision Matrix (FP32) performance |
91.6 TFLOPS |
| Peak Half Precision (FP16) performance |
362.05 TFLOPS |
| TV tuner integrated |
No |
| Dual Link DVI |
No |
| Design |
| Cooling type |
Passive |
| Form factor |
Full-Height/Full-Length (FH/FL) |
| Number of slots |
2 |
| Product colour |
Brown |
| Power |
| Power consumption (max) |
350 W |
| Supplementary power connectors |
1x 16-pin |
| Weight & dimensions |
| Length |
266.7 mm |
| Height |
111.8 mm |