NVIDIA Launches DGX Spark 64GB Version Desktop AI Computer, with a Starting Price of $4999
The Block
1h ago
Ai Focus
NVIDIA officially launches the 64GB memory version of the DGX Spark desktop AI computer, with a starting price of $4,999, and will be available for purchase on October 23. This product is based on the GB10 Grace - Blackwell chips, supports multi-machine cluster expansion, and is compatible with CUDA, vLLM, llama.cpp as well as several mainstream open-source large models.
Helpful
No.Help

News from IT on October 2nd: NVIDIA officially announced an update to its DGX Spark product today, introducing a 64GB memory version of the DGX Spark desktop AI computer, with a starting price of $4,999 (approximately 33,566 RMB at current exchange rates). It will be officially available for purchase on October 23rd. Manufacturers such as Acer, Dell, ASUS, Gigabyte, MSI, and H3C will also launch complete products based on this platform.

The official previously launched 128GB version has also seen its pricing updated accordingly; the 128GB FE version is now priced at $6,950 (Note from IT: The current exchange rate is approximately 46,666 RMB).

In terms of hardware, DGX Spark is based on GB10 Grace to Blackwell supercomputing chips, and relies on a unified memory architecture to connect the CPU and GPU memory spaces. The newly added 64GB memory version has a total memory of 64GB, of which about 8GB is reserved for the system, and approximately 56GB is available for model weights and KV caching. This device supports NVFP4 quantization format and can smoothly deploy mainstream open-source large models such as Gemma4 26B, Qwen3.8-27B, Meta Muse Glimmer, Nemotron3.5, and Lightning. NVIDIA states that even after the model is loaded, there is still sufficient memory remaining for KV caching to support long-context inference tasks.

One of the major highlights of this device is its native capability for multi-machine cluster expansion, featuring a built-in ConnectX-7 high-speed network card, and it comes paired with a brand-new NVIDIA Sync toolkit. According to the official statement, ordinary developers can set up a multi-node cluster without needing in-depth knowledge of network operations and maintenance. The toolkit's built-in Cluster Assistant cluster assistant can automatically complete network configuration, device verification, and SSH environment deployment, supporting up to 4 DGX Spark devices in networking.

Official tests have shown that when two 64GB DGX Spark devices are used to form a cluster to run the Qwen3.8-27B model, the performance is increased by up to 1.7 times compared to using a single device, significantly improving inference throughput. Developers can also remotely access the DGX Spark cluster from ordinary PC devices, connect to mainstream VS Code, Cursor, etc., remotely monitor the status of the entire system, and one-click start the vLLM inference container, thereby reducing the barriers to distributed AI development.

In terms of software ecosystem, DGX Spark is fully equipped with the NVIDIA CUDA acceleration AI software stack, and is deeply adapted to the vLLM and llama.cpp inference frameworks. NVIDIA claims that the product has been specially optimized for local agent workflows, achieving a maximum speed increase of 1.9 times in local agent inference. Perplexity has completed the adaptation, launching a portable computer agent optimized for DGX Spark; moreover, the Laguna S2 118B large model can also be locally deployed and run on this hardware.

At the same time, mainstream open-source community models and frameworks such as Nemotron, Gemma, Qwen, DeepSeek, Mistral, and Stability.ai have also been adapted, enabling users to run cutting-edge models out of the box. Reports indicate that hardware manufacturers and open-source communities have formed a complete developer ecosystem matrix.

Performance tests show that the Qwen3.8 27B model deployed on DGX Spark has a score that is only 10 points behind the leading closed-source models. The report indicates that desktop hardware now has the capability to run high-quality models. The unified memory architecture has also been recognized by industry developers. Perplexity founder Aravind Srinivas commented that DGX Spark can nearly fully utilize the memory capacity, with stable heat dissipation performance. The unified memory design maximizes the efficiency per watt of power consumption, making it suitable for 24/7 uninterrupted operation of local agent tasks.

However, this platform is still targeted at professional developers. With a pricing starting from $4,999 (which is approximately 33,566 yuan at current exchange rates), its target audience mainly consists of AI researchers, independent developers, and small AI startups, rather than ordinary consumer users. With the launch of the 64GB version, the entry barrier has been lowered compared to the 128GB version, allowing more teams to experience the capabilities of local clusters.

Tip
$0
Like
0
Save
0
Views 16
WalletJYS reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
Why is Bitcoin Rising Today?
On October 2, Bitcoin briefly broke through $87,000, rising by about 3% during the session. Previously, the U.S. non-farm payroll data in September fell far short of expectations, reinforcing market expectations that the Federal Reserve would remain inactive in October. At the same time, the liquidation of over $120 million in Bitcoin shorts also contributed to the upward trend.
crypto.news
·2026-10-02 23:05:08
2
New colors added to the Extreme Karma 007GT: "Valley Green" exterior and "Ink Green" interior
GeekCar announces the addition of two new exterior colors, "Valley Green," and two new interior colors, "Ink Green," for the refreshed GeekCar 007GT. From October 1 to October 31, 2026, customers who place orders for the refreshed GeekCar 007GT will enjoy additional limited-time benefits.
The Block
·2026-10-02 22:56:04
7
Saba's CEFS ETF Announces Adjustment to Allocation Rate and Frequency
Saba Opportunistically Hedged Closed - End Funds ETF ( CEFS ) announces that starting from September 25, 2026, the fund will seek to distribute profits at an annual rate equivalent to 12% of the net asset value, and the distribution frequency will be changed from monthly to quarterly. According to the revised distribution policy, the initial distribution will be $0.246 per share, to be paid to shareholders registered as of September 28 on September 30, 2026.
PR Newswire
·2026-10-02 22:33:55
11
Facebook Whistleblower Francis Haugen questions whether AI company can exercise self-restraint
Facebook Whistleblower Francis Haugen stated on Friday that artificial intelligence companies must "step up their efforts and comply" with the spirit of the self-regulation agreements reached with the White House. In an interview with CNBC Squawk Box, she questioned whether these companies would actually abide by the "spirit" of the rules or whether they would take advantage of loopholes that are not explicitly prohibited.
CNBC
·2026-10-02 22:33:54
15
The Fed Launches FRED Connector to Enhance the Ability of AI to Access Economic Data
The Federal Reserve launched FRED MCP Connector on FRED Con on October 1st, allowing AI agents to directly access the St. Louis Federal Reserve's FRED economic database. Waller stated that AI agents already account for nearly half of the website traffic on FRED, but he also warned that there may still be errors, illusions, and mislabeling of sources in the interpretation of economic data by AI.
The Cryptonomist
·2026-10-02 22:25:31
14
View More