AI PERFORMANCE
1 PetaFLOPSHigh-density AI performance for FP4 inference workloads — right on your desk.
NVIDIA DGX SPARK · CHINA COMMUNITY
FusionXpark packs the NVIDIA GB10 Grace Blackwell superchip, unleashing 1 PFLOPS of AI compute in a compact chassis — keeping 200B-parameter models running locally.
DESKTOP SUPERCOMPUTING
From model loading and context inference to multi-node scaling, DGX Spark compresses data-center-class AI architecture into a desktop form — personal AI infrastructure, not just another workstation.
See what it can doAI PERFORMANCE
1 PetaFLOPSHigh-density AI performance for FP4 inference workloads — right on your desk.
UNIFIED MEMORY
128 GBLarger models, longer contexts, and multimodal inputs run stably in a unified memory architecture.
MODEL SCALE
200BSupports 200B-parameter-class models, with the DeepSeek, Llama, Gemma, and Qwen ecosystems.
SCALE OUT
ConnectXDual-node interconnect leaves headroom for labs, team prototypes, and small clusters.
FROM MODEL TO ACTION
01 · LOCAL MODEL
Run private models, RAG, agents, and multimodal prototypes — sensitive data never leaves the office.
02 · LIVE DEMO
Show real workflows like translation, review, and knowledge-base Q&A — turning compute specs into experience.
03 · COMMUNITY
Share images, tutorials, and local deployment experience around DGX Spark, 1Panel, and MaxKB.
DUAL-NODE LOCAL CLUSTER
Two DGX Spark units interconnect over ConnectX with RoCE, forming a 2 PFLOPS / 256 GB unified-memory local compute block. Community-verified: both DeepSeek-V4-Flash and Qwen3.8-Flash-Next serve stably on this dual-node setup — data never leaves the room.
DEEPSEEK-V4-FLASH-0731
QWEN3.8-FLASH-NEXT
The official NVFP4 sparse-MLA path with DSpark speculative decoding (MTP×5) squeezes a 1M-token context into two desktop units. Long documents, repo-scale RAG, and multi-stream chat all hold up.
A full-multimodal flagship with NVFP4 expert weights + FP8 n-gram drafts, SGLang dual-node tensor parallelism, and CUDA Graph acceleration. 900K context including vision — multimodal workflows stay local.
* Performance figures come from community dual-node testing (the MiaAI-Lab project repos and the NVIDIA developer forums); real-world results vary with quantization, concurrency, and context length. Individuals and small teams can get near-data-center-class LLM serving from a chassis-sized local cluster.
POWERED BY DGX-CN.COM
Preloaded with the 1Panel ops panel and the MaxKB knowledge base — hardware, models, and business scenarios in one local solution.
A localized multilingual expert for document translation, video subtitles, and live meeting interpretation. Content is processed locally, with industry glossary support.
A 24/7 AI legal specialist. Automatically benchmarks against historical contracts via RAG, flags risky clauses, and suggests revisions.
FREQUENTLY ASKED
No. dgx-cn.com is a Chinese fan site for DGX Spark, sharing hardware information, local deployment experience, and community application solutions with developers in China.
It targets mainstream open-model ecosystems such as DeepSeek, Qwen, Meta Llama, and Google Gemma. The runnable scale depends on model precision, quantization, and context length.
Local LLM inference, RAG knowledge bases, agent development, multimodal prototypes, and AI demos that demand data privacy and low latency.
Scan the WeChat QR code at the bottom of the page and add the note “粉丝” (fan) to get the Translation Box and Contract Review Box images plus deployment discussions.
DGX SPARK CHINA COMMUNITY
Get pricing and lead times, AI Box images and deployment tutorials, and join the DGX Spark developer community in China.
COMMUNITY SPONSOR
1Panel AI Appliance: private local deployment of DeepSeek-V4-Flash with a built-in enterprise AI gateway and horizontal scaling. Field-tested at 60+ total TPS with 2 concurrent streams, GPU temps below 70°C under sustained load.
About 1Panel