NVIDIA DGX Spark HiTechCloud

NVIDIA DGX Spark – Desktop Agent Computer cho AI developer

Equipped with the NVIDIA GB10 Grace Blackwell Superchip, DGX Spark brings compact AI performance for prototyping, fine-tuning, inference and building autonomous agents at the desktop.

GB10 Grace Blackwell 128 GB Unified Memory DGX OS NVIDIA AI Stack
Overview

A desktop agent computer for agentic AI, reasoning models and AI prototyping

NVIDIA DGX Spark delivers strong AI performance in a compact, power-efficient design, with the NVIDIA AI software stack and 128 GB of memory built in so developers can prototype, fine-tune and deploy new reasoning AI models.

NVIDIA NemoClaw in the NVIDIA Agent Toolkit helps build, evaluate and optimize safer autonomous agents directly from the desktop environment.

The 1 petaFLOP FP4 figure is theoretical, based on the use of sparsity.
NVIDIA NemoClaw on DGX Spark
Agentic AI Development

Build Local AI Agents Faster with NVIDIA DGX Spark

The latest DGX OS update installs NemoClaw faster, adds current open models and speeds up inference, shortening agentic AI development cycles on DGX Spark.

NVIDIA Agent Toolkit with NVIDIA OpenShell provides open-source models and software for developers to build safer autonomous agents.

Features

NVIDIA GPU, CPU, networking and AI software technologies

01

NVIDIA GB10 Superchip

Grace Blackwell architecture delivering up to 1 petaFLOP of theoretical FP4 AI performance with sparsity.

02

128 GB Unified Memory

Unified system memory lets you develop, test and run AI models of up to around 200 billion parameters.

03

NVIDIA ConnectX Networking

High-performance interconnect lets two DGX Spark systems be paired for larger models.

04

NVIDIA AI Software Stack

DGX OS, frameworks, libraries, NVIDIA NIM, and the AI software ecosystem come preconfigured for developers.

NVIDIA DGX Spark pricing

Choose a DGX Spark plan by AI agent scale and workload

Plans are available on several billing terms, from prototype to production and enterprise AI.

3261

AI Starter

NVIDIA DGX Spark

30,000,000 VND / 3 months
  • GPU1 x NVIDIA GB10 Grace Blackwell Superchip
  • RAM/VRAM128 GB DDR5X
  • Storage4 TB SSD NVMe M.2
  • NetworkWiFi 7, NVIDIA ConnectX-7 Smart NIC, 10 GbE
  • ConnectHDMI 2.1a, Bluetooth 5.3, 4× USB 4 Type-C
  • OSNVIDIA DGX OS
  • IP1 IPv4 / 1 IPv6
  • Bandwidth1 Gbps
  • SupportEmail / Ticket during business hours
  • NotesEntry-level agentic AI desktop
Sign up now
3263

AI Production

NVIDIA DGX Spark Cluster

60,000,000 VND / 3 months
  • GPU2 × NVIDIA GB10 Grace Blackwell Superchip
  • RAM/VRAM256 GB DDR5X
  • Storage8 TB SSD NVMe M.2
  • NetworkWiFi 7, NVIDIA ConnectX-7 Smart NIC, 10 GbE
  • ConnectHDMI 2.1a, Bluetooth 5.3, 4× USB 4 Type-C
  • OSNVIDIA DGX OS
  • IP1 IPv4 / 1 IPv6
  • Bandwidth7 Gbps
  • Support24/7 priority support + dedicated engineer
  • NotesMulti-model, large workloads, HA
Sign up now
3264

AI Enterprise

NVIDIA DGX Spark Cluster

120,000,000 VND / 3 months
  • GPU4 × NVIDIA GB10 Grace Blackwell Superchip
  • RAM/VRAM512 GB DDR5X
  • Storage16 TB SSD NVMe M.2
  • NetworkWiFi 7, NVIDIA ConnectX-7 Smart NIC, 10 GbE
  • ConnectHDMI 2.1a, Bluetooth 5.3, 4× USB 4 Type-C
  • OSNVIDIA DGX OS
  • IP1 IPv4 / 1 IPv6
  • Bandwidth10 Gbps
  • Support24/7 priority support + dedicated engineer
  • NotesMulti-model, large workloads, HA
Sign up now
Why DGX Spark

From AI prototype to production-grade agentic AI

DGX Spark lets developers and businesses build AI agents, fine-tune models and test inference faster on a complete NVIDIA AI platform.

Desktop AI Supercomputer

Brings Grace Blackwell performance into a compact, power-efficient form factor that fits a workspace.

Agentic AI Ready

Ready for research into reasoning AI, autonomous agents, OpenShell and NemoClaw.

Prototype-to-Production

Shorten the development cycle from idea and prototype through fine-tuning to AI workload deployment.

Flexible system connectivity

Multiple DGX Spark units can be connected for larger workloads and multi-model scenarios.

NVIDIA AI software

Take advantage of NVIDIA NIM, frameworks, libraries, drivers and an optimized DGX OS environment.

HiTechCloud provides

Advice on plan selection, AI environment, connectivity, runtime and an operating model that fits your business.

Workloads

Accelerate all agentic AI workloads

DGX Spark suits AI developers, researchers, and data scientists who need Grace Blackwell performance in a desktop-sized form factor.

Build AI agents faster
Prototyping

Build AI agents faster

Build, test, and validate AI models or agentic AI applications directly in a desktop environment.

Fine-tune AI models
Fine-Tuning

Fine-tune AI models

Fine-tune models and experiment with RAG workflows, reasoning AI and model optimization pipelines.

Local inference testing
Inference

Local inference testing

Run, evaluate and optimize inference for large models before moving to cloud or production.

High-performance data science
Data Science

High-performance data science

Accelerates data analysis, notebooks, feature engineering, and algorithm experimentation at your desk.

Edge application development
Edge AI

Edge application development

Suited to experimenting with Isaac, Metropolis, Holoscan and edge AI applications.

NVIDIA AI Enterprise DGX Spark
NVIDIA AI Enterprise

NVIDIA AI Enterprise cho DGX Spark

A set of cloud-native tools, libraries and frameworks that speeds up production AI development, with security, optimized performance and enterprise support.

DGX Spark AI Software
DGX Spark AI Software

Pre-installed AI software stack

DGX Spark ships with the NVIDIA AI software stack and supports the NVIDIA software ecosystem, so AI projects can start quickly.

01

High-performance GPU instances for a wide range of workloads

Powerful capacity, optimized to accelerate AI/ML and high-performance workloads at any scale.

02

A broad range of infrastructure options

Training, inference, or fine-tuning - HiTechCloud offers a range of GPUs to match your needs, with transparent pricing and on-demand deployment environments.

03

Built to NVIDIA reference architecture

HiTechCloud GPU instances combine NVLink/PCIe, InfiniBand (RDMA), and RAIL topology to optimize AI/HPC performance.

GPU cloud architecture

Three core connectivity layers for high-performance GPU clusters

A network and GPU fabric designed so AI workloads scale reliably, easing bandwidth bottlenecks and holding performance at scale.

NVLink and PCIe Switch for HiTechCloud GPU instances

NVLink / PCIe Switch

High-speed GPU-to-GPU connectivity within and between nodes, reducing bottlenecks during model training.

InfiniBand RDMA for distributed training

InfiniBand (RDMA)

Low-latency connectivity optimized for distributed training and reduced processing load on the host.

RAIL topology for high-performance GPU clusters

RAIL Topology

A parallel network architecture delivering higher bandwidth, redundancy, and consistent performance at any scale.

Auto scaling

GPU auto scaling and utilization optimization

From a single GPU to large clusters, HiTechCloud provisions resources ahead of demand and gets the most out of every instance.

GPU auto scaling with Kubernetes

Auto scaling with Kubernetes

Scale GPU resources automatically from a handful to thousands of GPUs, with forecasting and provisioning ahead of demand.

SSH, TCP, and HTTP connection security for GPU instances

Security for every access connection

Access over SSH, TCP, and HTTP with built-in protection layers that keep enterprise data safe and access fully controlled.

Optimize GPU utilization with MIG

Optimize GPU utilization with MIG

Partition a single GPU into multiple independent instances to run AI workloads in parallel, improving utilization and lowering infrastructure cost.

GPU instance

Transparent pricing, no hidden fees

From large-scale training to real-time inference - pay only when you actually run, on a GPU cloud platform built for every AI and high-performance workload.

Ready to launch your first GPU instance?

From sign-up to a running GPU instance in under 5 minutes - no complex setup, no resource reservations, no charges while idle. Just deploy, run, and pay only for what you use.

The HiTechCloud ecosystem

More than compute - manage, scale, and build everything on one straightforward cloud ecosystem.

AgentBase

A complete management platform for deploying and operating AI agents securely at scale on enterprise-grade infrastructure.

Explore AgentBase

AI Platform

A unified platform for training, fine-tuning, and deploying AI models at any scale.

Explore AI Platform

Vector Database

Supports fast search, real-time analytics, log and large-scale event data, and vector databases for RAG.

Explore Vector Database

Kubernetes

Managed Kubernetes for container orchestration, AI services, and GPU cloud workloads.

Explore Kubernetes
Southeast Asia

Scale with confidence across Southeast Asia

Deploy systems and applications closer to your users, reducing latency and meeting local regulatory requirements.

01

Bangkok

BKK-01

03

Ho Chi Minh

HCM-01 · HCM-02 · HCM-03

02

Ha Noi

HAN-01 · HAN-02

Southeast Asia regional map for HiTechCloud infrastructure
1,000+ businesses

A partner for your digital transformation journey

Large enterprises and fast-growing startups choose HiTechCloud for secure, high-performance AI cloud solutions that let them innovate and scale.

Have a specific requirement? HiTechCloud is ready to help.

The HiTechCloud team advises on GPU architecture, networking, security, and an operating model matched to your real workloads.

FAQ

FAQ

Quick facts before choosing the NVIDIA DGX Spark at HiTechCloud.

Who is the NVIDIA DGX Spark suited for?

DGX Spark suits developers, researchers, data scientists and businesses that want to work on agentic AI, fine-tuning, inference, RAG and data science in a desktop environment.

Can DGX Spark run large models?

With 128 GB of unified memory per system, DGX Spark is suited to experimenting with and running large AI models. Connecting two systems lets the workload scale to larger models.

What is NemoClaw used for?

NemoClaw is an open-source reference stack in the NVIDIA Agent Toolkit that supports building, evaluating and optimizing safer AI agents with security and privacy guardrails.

What billing cycles are available for DGX Spark plans?

The plans on this page are available on 3-, 6-, 12-, 24-, 36- and 48-month terms.

Should you choose AI Starter or AI Research?

AI Starter suits getting started and prototyping. AI Research suits teams that need help setting up the AI environment, higher bandwidth and more frequent use.

When should you choose AI Production or AI Enterprise?

Choose AI Production or AI Enterprise when you need multiple DGX Spark systems, larger resources, multi-model workloads, HA and 24/7 priority support.

Does HiTechCloud support deploying AI environments?

Yes. HiTechCloud advises on DGX OS, the NVIDIA AI software stack, containers, AI frameworks, NVIDIA NIM, networking and workload-specific operating models.

Desktop AI Ready

Need advice on the NVIDIA DGX Spark for AI agents and internal workloads?

HiTechCloud helps you select the DGX Spark plan, interconnect architecture, NVIDIA AI software stack and operating roadmap that fit your business.