The NVIDIA RTX™ A400, built on the NVIDIA Ampere GPU architecture, brings the power of AI and ray-tracing acceleration into the hands of more professionals. The RTX A400 ensures AI-powered workflows and stunning ray-traced visuals are delivered with unprecedented performance in a space-saving design. Plus, you can connect up to four native displays to expand your visual horizons.
With its compact form factor, the RTX A400 fits effortlessly into any workstation, providing the necessary performance and capabilities for today’s professional workflows without compromising on efficiency or workspace.

| Part Number | VCNRTXA400ATX-PB |
| CUDA Cores | 768 |
| Tensor Cores | 24 |
| RT Cores | 6 |
| Single Precision Performance1 | 2.7 TFLOPS |
| RT Core Performance1 | 5.3 TFLOPS |
| Tensor Performance1 | 21.7 TFLOPS2 |
| GPU Memory | 4 GB GDDR6 |
| Memory Interface | 64-bit |
| Memory Bandwidth | 96 GB/sec |
| System Interface | PCI Express 4.0 x83 |
| Display Connectors | 4x mDisplayPort 1.4a |
| Maximum Power Consumption | 50 W |
|
|
Free dedicated phone and email technical support(1-800-230-0130)
Dedicated NVIDIA professional products Field Application Engineers
The NVIDIA Ampere architecture is designed to tackle the world's most important scientific, industrial, and business challenges. Delivers up to 2.5X the single precision (FP32) performance of the previous generation. Improved graphics performance due to new process.
Experience up to 4X the rendering performance with second-generation RT Cores, which offer double the efficiency in rendering. Accelerated motion blur, enhanced denoising, and a twofold increase in ray-triangle intersection speeds, along with the ability to perform concurrent RT operations, dramatically elevate the realism and speed of visual projects.
Third-generation Tensor cores deliver incredible performance for AI-powered workflows, providing up to 2.5X the generative AI performance over the previous generation. The Ampere GPU architecture supports a wider range of data formats, with support for Tensor Float 32 (TF32) precision, FP32, FP16, BF16, and INT4 data formats.
Unlock faster data-transfer speeds for data-intensive tasks with support for PCIe Gen 4, enhancing communication between CPU, memory, GPU, and streamlining your workflow efficiency.
Take advantage of 4GB of GDDR6 memory and a 20% increase in memory bandwidth. This boosts application performance, especially for content creation and CAD workflows, and compute-intensive tasks, resulting in smoother operations and greater efficiency when handling demanding projects. NVIDIA RTX A400 is an ideal platform for 3D professionals and multi-display environments, where integrated graphics fall short.
Deliver faster than real-time performance for transcoding, video editing, and other encoding applications with two dedicated H.264 and HEVC encode engines and a dedicated decode engine that are independent of 3D/compute pipeline.
NVDEC is well suited for transcoding and video playback applications for real-time decoding. The following video codecs are supported for hardware-accelerated decoding: MPEG-2, VC-1, H.264 (AVCHD), H.265 (HEVC), VP8, VP9, and AV1.
NVENC can take on 4K or 8K video encoding tasks to free up the graphics engine and the CPU for other operations. The RTX A1000 provides bettering encoding quality than software-based x264 encoders.
Pixel-level preemption provides more granular control to better support time-sensitive tasks such as VR motion tracking.
Preemption at the instruction level provides finer-grain control over compute tasks to prevent long-running applications from either monopolizing system resources or timing out.
Accelerate GPU-based lossless decompression performance by up to 100x and 20x lower CPU utilization compared to traditional storage APIs using Microsoft's DirectStorage for Windows API. RTX IO moves data from the storage to the GPU in a more efficient, compressed form, improving I/O performance.
Transparently scale the desktop and applications across up to 4 GPUs and 16 displays from a single workstation while delivering full performance and image quality.
Get more Mosaic topology choices with high-resolution display devices with a 32K max desktop size.
Support up to four 5K monitors @ 60Hz per card. The RTX A1000 supports HDR color for 4K at 60Hz for 10/12b HEVC decode and up to 4K at 60Hz for 10b HEVC encode. Each DisplayPort connector can drive ultra-high resolutions of 4096 x 2160 at 120 Hz with 30-bit color.
Gain unprecedented end-user control of the desktop experience for increased productivity in single-large display or multi-display environments, especially in the current age of large, widescreen displays.
NVIDIA RTX Experience delivers a suite of productivity tools to your desktop workstation, including desktop recording, automatic alerts for the latest NVIDIA RTX Enterprise driver updates, and access gaming features. The application is available for download here.
Deep learning frameworks such as Caffe2, MXNet, CNTK, TensorFlow, and others deliver dramatically faster training times and higher multi-node training performance. GPU-accelerated libraries such as cuDNN, cuBLAS, and TensorRT deliver higher performance for both deep learning inference and high-performance computing (HPC) applications.
Natively execute standard programming languages like C/C++ and Fortran, and APIs such as OpenCL, OpenACC, and Direct Compute to accelerate techniques such as ray tracing, video and image processing, and computation fluid dynamics.
A single, seamless 49-bit virtual address space allows for the transparent migration of data between the full allocation of CPU and GPU memory.
Maximize system uptime, seamlessly manage wide-scale deployments, and remotely control graphics and display settings for efficient operations.
| Product | NVIDIA RTX™ A400 |
| Architecture | NVIDIA Ampere |
| Process Size | 8N | NVIDIA Custom Process |
| Transistors | 8.7 Billion |
| Die Size | 200 mm2 |
| CUDA Cores | 768 |
| Tensor Cores | 24 |
| RT Cores | 6 |
| Single Precision Performance1 | 2.7 TFLOPS |
| RT Core Performance1 | 5.3 TFLOPS |
| Tensor Performance1 | 21.7 TFLOPS2 |
| GPU Memory | 4 GB GDDR6 |
| Memory Interface | 64-bit |
| Memory Bandwidth | 96 GB/sec |
| Display Connectors | 4x mDisplayPort 1.4a |
| NVENC | NVDEC | 1x | 2x (+ AV1 decode) |
| System Interface | PCI Express 4.0 x83 |
| Form Factor | 2.7" H x 6.4" L Single Slot |
| Thermal Solution | Active Fan |
| Maximum Power Consumption | 50 W |
| Max Digital Resolution | Support up to four 5K monitors @ 60Hz per card |
Performance numbers may be subject to change until product availability.
Whether you have questions about our products, need assistance with your account, or require technical support, our comprehensive resources and dedicated support team are here to provide the assistance you need
