NVIDIA DGX Spark
GB10 Grace Blackwell Superchip
- 128 GB coherent unified memory
- 273 GB/s memory bandwidth
- Up to 1 PFLOP FP4
- CUDA and the NVIDIA AI ecosystem
A compact Arm-based system suited to large local language models, coding agents, multimodal work and retrieval-augmented generation. Its unified memory can accommodate very large quantised models when the model, context and runtime fit within the available pool.
- gpt-oss 120B
- Larger Qwen3 variants
- Multiple smaller models
- Multimodal workloads
- Coding agents
- RAG pipelines
