QDNAAdvisory and architecture for LLM inference and training platforms, on-premises or hybrid

DGX Spark

AI workstation, GB10 superchip

DGX Spark
Indicative price€5,850 incl. VAT (€4,875 excl. VAT), best price listed on idealo.fr on 2 September 2026; €5,850 to €7,511 incl. VAT depending on the seller, excluding integration (idealo.fr listing, price guide)
Memory128 GB unified LPDDR5X
Compute1 PFLOP FP4 with sparsity (1,000 TOPS)
ChipGB10 Grace Blackwell, 20 Arm cores
NetworkConnectX-7 NIC at 200 Gb/s (two QSFP ports), up to 4 units in a cluster according to NVIDIA
Power draw240 W power supply, 140 W GB10 chip TDP (NVIDIA datasheet), wall outlet; 240 W is the upper-bound assumption used in the site’s calculations, ServeTheHome measuring 60 to 200 W under load
Form factor150 mm desktop

What the machine runs

Models up to 200 billion parameters quantised on a single node. With two nodes directly connected by a QSFP cable, NVIDIA claims up to 405 billion parameters (Project DIGITS announcement, January 2025); beyond that, the ConnectX-7 NICs go through an Ethernet switch, up to four units and 700 billion parameters according to the NVIDIA datasheet re-read on 2 September 2026. Recommended runtime: llama.cpp or vLLM locally.

Positioning

A sovereign workstation for a small business puts a private model on the desk, with no reliance on the cloud.

Manufacturers

NVIDIA offers the DGX Spark through its partners, including ASUS, Dell, Gigabyte, HP, Lenovo and MSI.

GPUDirect Storage

On platforms fitted with ConnectX cards, GPUDirect Storage technology establishes a direct path between NVMe or NVMe over Fabric storage and GPU memory, bypassing the CPU buffer. Compatible storage arrays, such as NetApp, VAST, DDN or WEKA, feed training and large-scale RAG at full speed.

Support

Open stack supported by QDNA and the communities. Optional NVIDIA AI Enterprise (NIM, NeMo, Triton) with a service-level agreement, licensed per GPU.

Configure this hardware

A no-commitment conversation to size your platform.

Book a call

Which models fit on DGX Spark?

The machine offers 128 GB of memory. 1 of the 7 open models in the catalogue load on it, in the format shown.

ModelBillion parametersMost precise format that fitsSizing page
Qwen 3.8 27B27FP16Qwen 3.8 27B on DGX Spark

See the full sizing matrix.

Frequently asked questions

How much does a configured dgx spark cost for an LLM?

Pricing depends on configuration and supplier. The page shows a dated public price; we provide a detailed quote after a scoping call.

Which LLM fits in a dgx spark?

Depends on the quantization format and the input context. The table on the page indicates the recommended maximum model size.

Which runtimes support the dgx spark?

vLLM, llama.cpp, Triton Inference Server, depending on the chosen framework. The model, format and GPU combinations measured by QDNA are published in the measurements section.