NVIDIA DGX Spark

-5 % off new 7,319 € New 7,709 €
Installment calculator 4,9%  interest from 168.34 € monthly

NVIDIA DGX Spark

NVIDIA logo

NVIDIA DGX Spark 4TB mini computer

The world's smallest personal AI supercomputer

NVIDIA DGX Spark is a groundbreaking miniature personal AI supercomputer that, for the first time ever, packs the power of an entire AI data centre into a body measuring just 150×150×50.5mm. At its core beats the NVIDIA GB10 Grace Blackwell superchip – a combination of a 20-core ARM processor and a Blackwell architecture GPU with 5th generation tensor cores, which together achieve a performance of up to 1 petaflop (PFLOP). Its key advantage is 128GB of shared unified CPU and GPU memory connected by NVLink-C2C technology, with 5× higher throughput than PCIe 5th generation. This allows you to load models with up to 200 billion parameters, such as Llama 3.1 70B or Gemma 3 27B, directly into the memory and run them locally, without relying on the cloud or an external server. Encrypted 4TB M.2 NVMe storage, a ConnectX-7 network card with 200Gb/s InfiniBand, Wi-Fi 7, Bluetooth 5.4, and four USB-C 4.0 ports (40Gb/s) ensure connectivity for every professional scenario – including linking two DGX Spark units into a mini cluster to work with models of up to 400 billion parameters. The pre-installed NVIDIA DGX OS with CUDA libraries, Docker, and the Container Toolkit enables an immediate start to development without lengthy configuration. This system is based on Ubuntu Linux with integrated Ubuntu Pro Client support for extended ESM security updates.

Nvidia GeForce RTX 40

LPDDR5x / 128GB

Operating memory
Nvidia GeForce RTX 40

Blackwell GB10

Chip architecture
NVIDIA DGX Spark

NVIDIA ConnectX

High-speed network card
NVIDIA DGX Spark

NVIDIA AI Software Stack

Complete software stack
Nvidia GeForce RTX 40

240W

Power consumption
NVIDIA DGX Spark

Key features of the NVIDIA DGX Spark 4TB mini computer

  • NVIDIA GB10 Grace Blackwell superchip with a 20-core ARM processor and a Blackwell architecture GPU with 5th generation tensor cores
  • Up to 1 PFLOP of AI inference performance for local processing of the most demanding models
  • 128GB of shared unified CPU and GPU memory connected by NVLink-C2C technology with 5× higher throughput than PCIe 5th generation
  • Loads and runs models with up to 200 billion parameters directly in memory
  • NVIDIA ConnectX-7 network card with 200Gb/s InfiniBand connects two DGX Spark systems into a mini cluster for working with models up to 405 billion parameters
  • Encrypted 4TB M.2 NVMe storage provides ample space for large datasets, models, and development projects with maximum data security
  • Pre-installed NVIDIA DGX OS with CUDA libraries, Docker, and the Container Toolkit
  • System based on Ubuntu Linux with integrated Ubuntu Pro Client support for extended ESM security updates
  • Support for PyTorch, TensorFlow, and NVIDIA NIM microservices
  • Wi-Fi 7, Bluetooth 5.4, 10Gb Ethernet, 4× USB-C 4.0 (40Gb/s), and HDMI ensure complete connectivity in a body measuring 150×150×50.5mm and weighing just 1.2kg
NVIDIA DGX Spark

Data centre performance in a miniature body

The heart of the NVIDIA DGX Spark is the GB10 Grace Blackwell superchip – a unique combination of a 20-core ARM processor (10× Cortex-X925 + 10× Cortex-A725) and a Blackwell architecture GPU with 5th generation tensor cores and 4th generation RT cores. Both parts of the superchip are connected by NVLink-C2C technology, which provides 5× higher throughput than PCIe 5th generation and lets the CPU and GPU share a single 128GB pool of unified LPDDR5X memory. The result is immense performance that consumes a maximum of 240W and fits on any desk. This shared memory, without the traditional separation of GPU VRAM and system RAM, is the key advantage of the DGX Spark over standard desktop setups. Where a GeForce RTX 5090 graphics card offers 32GB of memory, the DGX Spark makes four times that available for AI models, without needing to quantise models or split them between multiple devices.

NVIDIA DGX Spark

A personal supercomputer that fits on any desk

NVIDIA DGX Spark proves that you no longer need a server room or a rack full of hardware to run the largest AI models. With dimensions of 150×150×50.5mm and a weight of 1.2kg, the DGX Spark fits on any desk next to your laptop, yet offers performance that, just a few years ago, would have required an entire server full of graphics cards. Four USB-C 4.0 ports with speeds up to 40Gb/s, Wi-Fi 7, 10Gb Ethernet, and an HDMI output ensure you can connect all your accessories without needing an external hub or docking station. The DGX Spark is designed for developers, scientists, and AI professionals who want the full power of a local, personal AI supercomputer without relying on the cloud – right where they work.

Accelerate all AI tasks in a compact body

The power of the Grace Blackwell architecture in a body that fits on any desk – NVIDIA DGX Spark is the ideal choice for developers, researchers, and data scientists who need full-fledged AI performance without compromise. Whether you're working on large language model inference, fine-tuning pre-trained networks, or developing AI agents, the DGX Spark handles the full spectrum of AI tasks locally, quickly, and without depending on cloud infrastructure.

NVIDIA DGX Spark

Prototype and develop AI applications without limits

The complete NVIDIA AI software stack provides developers with a full-featured platform for creating AI models, AI agents, and AI-enhanced applications – all locally, without cloud dependency. Once a solution is ready for deployment or final fine-tuning, the DGX Spark enables direct and seamless migration of workloads to the NVIDIA DGX Cloud or other NVIDIA-accelerated infrastructure.

NVIDIA DGX Spark

Fine-tune AI models with up to 70 billion parameters

Use the 128GB of unified memory in the NVIDIA DGX Spark to fine-tune pre-trained models with up to 70 billion parameters, right at your workstation. Training on your own data allows you to specialise AI models for specific needs, industry data, or particular use cases – without having to send sensitive data to the cloud or pay for expensive GPU instances. The result is a model tailored precisely to your scenario, created locally, securely, and under your full control.

NVIDIA DGX Spark

NVIDIA DGX Spark 4TB delivers top-tier data science

The combination of 128GB of unified memory and 1 PFLOP of parallel throughput performance makes the NVIDIA DGX Spark the ideal workstation for demanding data analytics and machine learning projects. Large datasets, complex computational models, and intensive training and analysis processes that previously required a powerful cloud cluster or dedicated server now run right on your desk. Fast, local, and with no waiting for remote infrastructure.

NVIDIA DGX Spark

Inference for models with up to 200 billion parameters

Fifth-generation Tensor Cores with FP4 format support achieve up to 1 PFLOP of performance and, combined with 128GB of system memory, enable you to run inference for state-of-the-art AI models with up to 200 billion parameters directly on your desk. Test, verify, and deploy models like Llama, Gemma, or Qwen locally and in real time – without cloud latency, shared infrastructure, or the risk of sensitive data leaks.

NVIDIA DGX Spark

Application development for robotics, industry, and smart cities

The NVIDIA DGX Spark is an exceptional platform for developing robotics systems, smart city solutions, and computer vision applications. Pre-installed NVIDIA Isaac frameworks for robotics, Metropolis for video analysis, and Holoscan for real-time data processing let developers take full advantage of the DGX Spark's power to rapidly prototype and deploy applications – locally, without dependence on cloud infrastructure, and with full support from the NVIDIA ecosystem.

Nvidia GeForce RTX 40

Personal AI supercomputer performance

Up to 1 petaflop for inference
Nvidia GeForce RTX 40

Up to 200 billion parameters

Local inference without cloud dependency
Nvidia GeForce RTX 40

Connecting two systems

Mini cluster for models up to 405 billion parameters
Nvidia GeForce RTX 40

Ready to work straight away

Pre-installed NVIDIA AI software stack
Nvidia GeForce RTX 40

Memory without compromise

128GB of unified CPU and GPU memory
Nvidia GeForce RTX 40

Your own data

Fine-tune models with up to 70 billion parameters
Nvidia GeForce RTX 40

Compact dimensions

150×150×50.5mm, 1.2kg
Nvidia GeForce RTX 40

Secure local work

Full data control without sending to the cloud

Specifications

Use

Assembly type Mini PC

Processor

Processor model number NVIDIA GB10
Max TDP 140 W

GPU series/model

Graphics card series NVIDIA
Model graphics cards GB10

Graphics card

Hard Drive

Storage capacity (total) 4,000 GB (4 TB)
SSD Capacity 4,000 GB (4 TB)
Internal interface M.2 (PCIe 4.0 4x NVMe)

Equipment

Basic equipment Bluetooth, Wi-Fi
WiFi 802.11be
WiFi version WiFi 7

Colour and design

Colour Gold
Front panel location front
Sidewalls Opaque

PSU

PSU 240 W

Memory

Size of operational RAM 128 GB
Memory type LPDDR5x

Outputs

USB-C 4 pc(s)
Graphics HDMI, USB-C
Other RJ-45 (LAN) 10Gbps, NVIDIA ConnectX-7 SmartNIC
Optical drive Without Optical Drive

Case

Width 150 mm
Height 50.5 mm
Depth 150 mm
Weight 1.2 kg

Operating system

Operating system NVIDIA DGX OS 7

Package contents

Package includes Power adapter, Power cable

AI

Total number of TOPS 1,000
Software platform NVIDIA CUDA
Supported precision FP4

External structure

Sidewalls Opaque
Door No door

AI specifications

Primary AI use Fine-tuning
Max. model size 70–405B
Certified AI frameworks PyTorch
SmartNIC ConnectX-7

GPU configuration

Number of GPUs in configuration 1 ×
NVLink version NVLink-C2C
Code:  NRn2000
Warranty: 24 months
Links: Producer's Website: