MEMORY INDUSTRY INTELLIGENCE
NVIDIA DGX Spark 64GB, 개발자가 로컬 AI를 구축하고 확장할 수 있는 더 많은 방법 제공
한국어 번역·요약·분석
원문 제목: NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI
핵심 요약
NVIDIA는 이번 달 Acer, ASUS, Dell, Gigabyte, HP, MSI 등 주요 제조 파트너를 통해 64GB 통합 메모리를 갖춘 DGX Spark를 출시한다고 발표했다. 이 신규 SKU는 GB10 Grace Blackwell Superchip, DGX OS, NVIDIA AI 소프트웨어 스택을 유지하며 최대 1,000억 파라미터 모델을 온디바이스로 지원한다. 두 대의 64GB 유닛을 NVIDIA Sync Cluster Assistant로 클러스터링하면 메모리를 128GB로 풀링하고 최대 2,000억 파라미터 모델 지원, 2배 메모리 대역폭, 최대 1.7배 성능을 제공한다. NVIDIA의 Qwen 3.8 27B 테스트에서 2대 클러스터가 단일 시스템 대비 최대 1.7배 성능을 보였다. 가격은 10월 23일 금요일부터 4,999달러부터 시작한다.
메모리 산업 영향 분석
이 문서는 NVIDIA의 DGX Spark 64GB 구성 출시를 다루며, 메모리 산업에는 통합 메모리(64GB, 클러스터 시 128GB)와 고대역폭 메모리 수요 측면에서 직접적 관련성이 있다. 관련 제품은 GB10 Grace Blackwell Superchip 기반 DGX Spark이며, 고객은 개발자, 연구자, AI 애호가 및 Acer, ASUS, Dell, Gigabyte, HP, MSI 등 제조 파트너이다. 직접 효과로는 온디바이스 AI 추론 및 에이전트 실행을 위한 고용량 통합 메모리 채택 증가, 계층 이동으로는 클라우드 의존 감소 및 로컬 디바이스로의 워크로드 이동, 사용량 효과로는 최대 1,000억 파라미터 모델 지원 및 클러스터링을 통한 2,000억 파라미터 확장이 있다. 적용 범위는 개인 AI 슈퍼컴퓨팅 및 엣지 개발이다. 반대 근거로는 클라우드 AI가 여전히 대규모 모델에 우위를 가질 수 있으며, 확인할 지표로는 DGX Spark 판매량, 통합 메모리 공급업체 수혜, 클러스터링 채택률이 있다. 분석가 가설로는 메모리 산업의 고용량 통합 메모리 수요 증가 가능성이 있으나, 문서에서 메모리 공급업체나 구체적 메모리 기술은 명시되지 않아 직접 연결 근거는 제한적이다.
한국어 번역 읽기
수집된 원문 v1의 전체 본문 기준 · 6375자
로컬 AI는 토큰 단위로 더 유용해지고 있다.
AI 에이전트가 실험에서 일상 개발로 이동함에 따라, 점점 더 유능한 오픈 모델이 더 많은 장치에 맞게 축소되어 빌더들이 로컬에서 실행할 수 있는 것이 많아지고 있다.
이번 달 출시 예정인
NVIDIA DGX Spark
는 주요 제조 파트너인 Acer, ASUS, Dell, Gigabyte, HP, MSI를 통해 64GB 통합 메모리 구성으로 제공되어, 개발자, 연구자, AI 애호가에게 DGX OS와 NVIDIA AI 소프트웨어 스택이 첫날부터 바로 사용 가능한 새로운 구성을 제공한다.
새 SKU는 클라우드 의존 없이 프라이빗하게 온디바이스에서 유능한 로컬 에이전트를 실행한다. 그리고 워크로드가 커지면 두 대의 유닛을 NVIDIA Sync Cluster Assistant를 통해 추가 설정 없이 클러스터로 묶을 수 있다.
개인 AI 슈퍼컴퓨팅의 새로운 시작점
DGX Spark는 NVIDIA Grace Blackwell 컴퓨팅, 통합 메모리,
NVIDIA ConnectX-7 네트워킹
및
NVIDIA CUDA
가속 AI 소프트웨어 스택을 하나의 시스템에 결합한다. 이는 에이전트, 추론, 파인튜닝, 데이터 과학, 엣지 개발을 위한 완전한 로컬 AI 플랫폼이다.
이 컴팩트한 개인 AI 슈퍼컴퓨터는 모든 작업마다 클라우드 인스턴스에 의존하지 않고 모델과 개발자 자신의 데이터를 실험할 수 있는 장소를 제공한다.
제조 파트너를 통해서만 독점 제공되는 새로운 64GB 구성은 GB10 Grace Blackwell Superchip, DGX OS 및 전체 NVIDIA AI 소프트웨어 스택을 유지하면서 플랫폼을 접근 가능한 가격대로 유지한다. 이는 최대 1,000억 파라미터 모델과 그 위에 구축된 에이전트 애플리케이션을 완전히 온디바이스로 지원한다.
두 대의 64GB 유닛을 클러스터로 묶으면 메모리가 단순히 두 배가 되는 것이 아니다. NVIDIA의 Qwen 3.8 27B 테스트에서 두 대의 클러스터링된 64GB 시스템은 단일 시스템 대비 최대 1.7배 성능을 제공했으며, 워크로드 요구에 따라 계속 확장할 여지가 있다.
DGX Spark는 첫날부터 에이전트 개발이 가능하도록 출하된다 — NVIDIA Agent Toolkit, CUDA-X AI 라이브러리, Nemotron 오픈 모델, 그리고 Ollama, vLLM, PyTorch with CUDA와 같은 인기 런타임이 모두 기본 지원된다. 개발자는 전원을 켜고 몇 분 만에 모델을 실행할 수 있다.
Blender는 플랫폼을 지원하는 최초의 주요 크리에이터 애플리케이션 제공업체 중 하나이며, 사전 빌드된 다운로드 가능한 설치 프로그램이 곧 제공될 예정이다.
NVIDIA Sync Cluster Assistant로 확장
개발자는 오늘 프로젝트에 필요한 메모리로 시작하고, 파이프라인이 성장함에 따라 더 큰 워크로드를 위해 다중 노드 클러스터를 원활하게 확장하도록 설계된 플랫폼 위에 구축할 수 있다.
모든 DGX Spark에는 NVIDIA ConnectX-7 NIC가 기본 내장되어 있다. 또한 두 대의 유닛을 QSFP 케이블로 직접 연결하여 메모리를
128GB
로 풀링하고 모델 지원을 최대
2,000억 파라미터
까지 확장하며 두 배의 메모리 대역폭과 최대 1.7배 성능을 제공한다.
NVIDIA Sync
앱은 이 다중 노드 클러스터를 원활하게 구성한다. 클러스터 어시스턴트 기능은 연결된 유닛을 감지하고, 장치 구성을 검증하며, ConnectX-7 네트워크를 구성하여 개발자가 인프라 대신 작업에 집중할 수 있게 한다. 모든 노드는 동일한 NVIDIA 소프트웨어 스택을 실행하므로 한 대에서 두 대로 확장할 때 아무것도 재구성할 필요가 없다.
그리고 이번 달 말에 출시되는 NVIDIA Sync Model Launcher는 로컬 AI 실행을 몇 번의 버튼 클릭만큼 간단하게 만든다. 개발자는 단일 DGX Spark 시스템 또는 클러스터에서 Qwen3.8 27B를 다운로드하고 실행할 수 있으며, NVIDIA Sync가 연결된 장치 전반에 걸쳐 모델을 실행하도록 구성하고 사용자의 노트북에서 접근 가능하게 한다. 런처는 또한 OpenCode가 모델을 사용하도록 설정하여 개발자가 브라우저에서 코딩을 시작할 수 있게 한다.
DGX Spark의 개발자 사용 사례
새로운 DGX Spark 64GB 구성은 첫날부터 실용적인 작업을 지원한다. 최대 1,000억 파라미터 모델이 완전히 온디바이스로 실행되므로, 개발자와 애호가는 메모리 내에 맞는 모델에는 단일 시스템으로 시작하거나, 더 많은 메모리와 컴퓨팅이 필요한 워크로드에는 NVIDIA Sync Cluster Assistant로 여러 DGX Spark 시스템을 연결할 수 있다.
다음은 세 가지 워크플로 예시이다:
24시간 AI 에이전트 실행:
DGX Spark에서 코딩 또는 연구 에이전트를 계속 실행하여 코드 검토, 문서 분석 또는 다단계 작업을 수행할 준비를 갖춘다. 클러스터는 더 큰 모델, 더 긴 컨텍스트 윈도우 또는 동시에 작동하는 여러 에이전트를 위한 추가 용량을 제공한다.
일상 PC에서 AI 앱 구동:
노트북이나 데스크톱에서 에이전트 또는 크리에이티브 애플리케이션을 사용하면서 DGX Spark에서 언어 또는 이미지 생성 모델을 실행한다. DGX Spark가 모델 추론을 처리하여 PC를 다른 작업에 자유롭게 한다.
작업이 커지면 확장:
단일 작업이 한 유닛을 초과할 때 — 더 큰 모델, 더 긴 컨텍스트 윈도우 또는 동시 에이전트 요청 실행 — NVIDIA Sync Cluster Assistant를 통해 200 GbE 패브릭으로 연결된 두 대의 DGX Spark 64GB 시스템이 메모리를 128GB로 풀링한다. 한 유닛에서 실행된 동일한 워크플로가 소프트웨어 환경을 재구성하지 않고 두 대로 확장된다.
DGX Spark 시작하기
DGX Spark 64GB는
Acer
,
ASUS
, Dell,
Gigabyte
, HP 및
MSI
에서 10월 23일 금요일부터 4,999달러부터 시작한다.
시작하려면:
지원되는 추론 프레임워크 다운로드 — llama.cpp, Ollama, vLLM 또는 LM Studio.
워크플로에 권장되는 로컬 모델 다운로드.
두 대의 유닛으로 확장하려면 NVIDIA ConnectX-7 포트를 통해 연결하고 NVIDIA Sync Cluster Assistant를 실행 — 네트워크를 구성하고 워크로드를 자동으로 라우팅한다.
DGX Spark의 에이전틱 AI 플레이북은 build.nvidia.com의
NemoClaw
,
OpenClaw
,
Hermes Agent
및
OpenShell
페이지를 방문하라.
#ICYMI: NVIDIA 로컬 AI의 더 많은 업데이트
DGX Spark를 위한 플레이북은
build.nvidia.com/spark
에서 탐색하라. 다음 플레이북이 64GB 장치에 곧 제공될 예정이다:
vLLM으로 LLM 서빙
로컬 LLM으로 OpenClaw 실행
분산 워크로드를 위해 여러 DGX Spark 연결
NVIDIA RTX Spark
가 탑재된 새로운 Windows PC가 이번 달 Acer, ASUS, Dell, HP, Lenovo, Microsoft 및 MSI에서 출시될 예정이다.
RTX Spark 뉴스레터
에 가입하여 향후 업데이트를 받아라.
Alibaba의
Qwen-Image-2.1
은 이미지 생성과 편집을 경량 오픈 웨이트 모델로 결합한다. NVIDIA RTX GPU, DGX Spark 및 DGX Station에서 로컬로 실행되어 크리에이터에게 자체 하드웨어에서 이미지를 생성하고 개선할 수 있는 더 많은 방법을 제공한다.
X
,
Instagram
,
TikTok
및
Facebook
에서 NVIDIA RTX Spark를 팔로우하고,
NVIDIA Local AI 뉴스레터
를 구독하여 정보를 받아라. LinkedIn과 X에서 NVIDIA Workstation을 팔로우하라.
소프트웨어 제품 정보에 관한
고지
를 참조하라.
AI 에이전트가 실험에서 일상 개발로 이동함에 따라, 점점 더 유능한 오픈 모델이 더 많은 장치에 맞게 축소되어 빌더들이 로컬에서 실행할 수 있는 것이 많아지고 있다.
이번 달 출시 예정인
NVIDIA DGX Spark
는 주요 제조 파트너인 Acer, ASUS, Dell, Gigabyte, HP, MSI를 통해 64GB 통합 메모리 구성으로 제공되어, 개발자, 연구자, AI 애호가에게 DGX OS와 NVIDIA AI 소프트웨어 스택이 첫날부터 바로 사용 가능한 새로운 구성을 제공한다.
새 SKU는 클라우드 의존 없이 프라이빗하게 온디바이스에서 유능한 로컬 에이전트를 실행한다. 그리고 워크로드가 커지면 두 대의 유닛을 NVIDIA Sync Cluster Assistant를 통해 추가 설정 없이 클러스터로 묶을 수 있다.
개인 AI 슈퍼컴퓨팅의 새로운 시작점
DGX Spark는 NVIDIA Grace Blackwell 컴퓨팅, 통합 메모리,
NVIDIA ConnectX-7 네트워킹
및
NVIDIA CUDA
가속 AI 소프트웨어 스택을 하나의 시스템에 결합한다. 이는 에이전트, 추론, 파인튜닝, 데이터 과학, 엣지 개발을 위한 완전한 로컬 AI 플랫폼이다.
이 컴팩트한 개인 AI 슈퍼컴퓨터는 모든 작업마다 클라우드 인스턴스에 의존하지 않고 모델과 개발자 자신의 데이터를 실험할 수 있는 장소를 제공한다.
제조 파트너를 통해서만 독점 제공되는 새로운 64GB 구성은 GB10 Grace Blackwell Superchip, DGX OS 및 전체 NVIDIA AI 소프트웨어 스택을 유지하면서 플랫폼을 접근 가능한 가격대로 유지한다. 이는 최대 1,000억 파라미터 모델과 그 위에 구축된 에이전트 애플리케이션을 완전히 온디바이스로 지원한다.
두 대의 64GB 유닛을 클러스터로 묶으면 메모리가 단순히 두 배가 되는 것이 아니다. NVIDIA의 Qwen 3.8 27B 테스트에서 두 대의 클러스터링된 64GB 시스템은 단일 시스템 대비 최대 1.7배 성능을 제공했으며, 워크로드 요구에 따라 계속 확장할 여지가 있다.
DGX Spark는 첫날부터 에이전트 개발이 가능하도록 출하된다 — NVIDIA Agent Toolkit, CUDA-X AI 라이브러리, Nemotron 오픈 모델, 그리고 Ollama, vLLM, PyTorch with CUDA와 같은 인기 런타임이 모두 기본 지원된다. 개발자는 전원을 켜고 몇 분 만에 모델을 실행할 수 있다.
Blender는 플랫폼을 지원하는 최초의 주요 크리에이터 애플리케이션 제공업체 중 하나이며, 사전 빌드된 다운로드 가능한 설치 프로그램이 곧 제공될 예정이다.
NVIDIA Sync Cluster Assistant로 확장
개발자는 오늘 프로젝트에 필요한 메모리로 시작하고, 파이프라인이 성장함에 따라 더 큰 워크로드를 위해 다중 노드 클러스터를 원활하게 확장하도록 설계된 플랫폼 위에 구축할 수 있다.
모든 DGX Spark에는 NVIDIA ConnectX-7 NIC가 기본 내장되어 있다. 또한 두 대의 유닛을 QSFP 케이블로 직접 연결하여 메모리를
128GB
로 풀링하고 모델 지원을 최대
2,000억 파라미터
까지 확장하며 두 배의 메모리 대역폭과 최대 1.7배 성능을 제공한다.
NVIDIA Sync
앱은 이 다중 노드 클러스터를 원활하게 구성한다. 클러스터 어시스턴트 기능은 연결된 유닛을 감지하고, 장치 구성을 검증하며, ConnectX-7 네트워크를 구성하여 개발자가 인프라 대신 작업에 집중할 수 있게 한다. 모든 노드는 동일한 NVIDIA 소프트웨어 스택을 실행하므로 한 대에서 두 대로 확장할 때 아무것도 재구성할 필요가 없다.
그리고 이번 달 말에 출시되는 NVIDIA Sync Model Launcher는 로컬 AI 실행을 몇 번의 버튼 클릭만큼 간단하게 만든다. 개발자는 단일 DGX Spark 시스템 또는 클러스터에서 Qwen3.8 27B를 다운로드하고 실행할 수 있으며, NVIDIA Sync가 연결된 장치 전반에 걸쳐 모델을 실행하도록 구성하고 사용자의 노트북에서 접근 가능하게 한다. 런처는 또한 OpenCode가 모델을 사용하도록 설정하여 개발자가 브라우저에서 코딩을 시작할 수 있게 한다.
DGX Spark의 개발자 사용 사례
새로운 DGX Spark 64GB 구성은 첫날부터 실용적인 작업을 지원한다. 최대 1,000억 파라미터 모델이 완전히 온디바이스로 실행되므로, 개발자와 애호가는 메모리 내에 맞는 모델에는 단일 시스템으로 시작하거나, 더 많은 메모리와 컴퓨팅이 필요한 워크로드에는 NVIDIA Sync Cluster Assistant로 여러 DGX Spark 시스템을 연결할 수 있다.
다음은 세 가지 워크플로 예시이다:
24시간 AI 에이전트 실행:
DGX Spark에서 코딩 또는 연구 에이전트를 계속 실행하여 코드 검토, 문서 분석 또는 다단계 작업을 수행할 준비를 갖춘다. 클러스터는 더 큰 모델, 더 긴 컨텍스트 윈도우 또는 동시에 작동하는 여러 에이전트를 위한 추가 용량을 제공한다.
일상 PC에서 AI 앱 구동:
노트북이나 데스크톱에서 에이전트 또는 크리에이티브 애플리케이션을 사용하면서 DGX Spark에서 언어 또는 이미지 생성 모델을 실행한다. DGX Spark가 모델 추론을 처리하여 PC를 다른 작업에 자유롭게 한다.
작업이 커지면 확장:
단일 작업이 한 유닛을 초과할 때 — 더 큰 모델, 더 긴 컨텍스트 윈도우 또는 동시 에이전트 요청 실행 — NVIDIA Sync Cluster Assistant를 통해 200 GbE 패브릭으로 연결된 두 대의 DGX Spark 64GB 시스템이 메모리를 128GB로 풀링한다. 한 유닛에서 실행된 동일한 워크플로가 소프트웨어 환경을 재구성하지 않고 두 대로 확장된다.
DGX Spark 시작하기
DGX Spark 64GB는
Acer
,
ASUS
, Dell,
Gigabyte
, HP 및
MSI
에서 10월 23일 금요일부터 4,999달러부터 시작한다.
시작하려면:
지원되는 추론 프레임워크 다운로드 — llama.cpp, Ollama, vLLM 또는 LM Studio.
워크플로에 권장되는 로컬 모델 다운로드.
두 대의 유닛으로 확장하려면 NVIDIA ConnectX-7 포트를 통해 연결하고 NVIDIA Sync Cluster Assistant를 실행 — 네트워크를 구성하고 워크로드를 자동으로 라우팅한다.
DGX Spark의 에이전틱 AI 플레이북은 build.nvidia.com의
NemoClaw
,
OpenClaw
,
Hermes Agent
및
OpenShell
페이지를 방문하라.
#ICYMI: NVIDIA 로컬 AI의 더 많은 업데이트
DGX Spark를 위한 플레이북은
build.nvidia.com/spark
에서 탐색하라. 다음 플레이북이 64GB 장치에 곧 제공될 예정이다:
vLLM으로 LLM 서빙
로컬 LLM으로 OpenClaw 실행
분산 워크로드를 위해 여러 DGX Spark 연결
NVIDIA RTX Spark
가 탑재된 새로운 Windows PC가 이번 달 Acer, ASUS, Dell, HP, Lenovo, Microsoft 및 MSI에서 출시될 예정이다.
RTX Spark 뉴스레터
에 가입하여 향후 업데이트를 받아라.
Alibaba의
Qwen-Image-2.1
은 이미지 생성과 편집을 경량 오픈 웨이트 모델로 결합한다. NVIDIA RTX GPU, DGX Spark 및 DGX Station에서 로컬로 실행되어 크리에이터에게 자체 하드웨어에서 이미지를 생성하고 개선할 수 있는 더 많은 방법을 제공한다.
X
,
,
TikTok
및
에서 NVIDIA RTX Spark를 팔로우하고,
NVIDIA Local AI 뉴스레터
를 구독하여 정보를 받아라. LinkedIn과 X에서 NVIDIA Workstation을 팔로우하라.
소프트웨어 제품 정보에 관한
고지
를 참조하라.
브리프용 요약 초안
NVIDIA가 64GB 통합 메모리를 갖춘 DGX Spark를 10월 23일부터 4,999달러에 출시하며, 최대 1,000억 파라미터 모델을 온디바이스로 지원한다. 두 대를 클러스터링하면 128GB 메모리와 최대 2,000억 파라미터 지원, 1.7배 성능 향상을 제공한다. 이는 로컬 AI 확산과 고용량 통합 메모리 수요에 긍정적이나, 메모리 산업 직접 영향은 추가 확인이 필요하다.
원문 텍스트
원문 열기 ↗Local AI is becoming more useful by the token.
As AI agents move from experiments into everyday development, increasingly capable open models are shrinking to fit on more devices, giving builders more to run locally.
Coming this month,
NVIDIA DGX Spark
will be available with 64GB of unified memory from top manufacturer partners — Acer, ASUS, Dell, Gigabyte, HP and MSI — giving developers, researchers and AI enthusiasts a new configuration with DGX OS and the NVIDIA AI software stack ready to use from day one.
The new SKU runs capable local agents on device — privately, without cloud dependency. And when workloads grow, two units can cluster together via NVIDIA Sync Cluster Assistant without any additional setup.
A New Starting Point for Personal AI Supercomputing
DGX Spark combines NVIDIA Grace Blackwell compute, unified memory,
NVIDIA ConnectX-7 networking
and an
NVIDIA CUDA
-accelerated AI software stack in one system. It’s a complete local AI platform for agents, inference, fine-tuning, data science and edge development.
The compact, personal AI supercomputer provides a place to experiment with models and developers’ own data without turning to a cloud instance for every task.
The new 64GB configuration, available exclusively from manufacturer partners, keeps the platform at an accessible price point while retaining the GB10 Grace Blackwell Superchip, DGX OS and full NVIDIA AI software stack — same as the 128GB model. It supports up to 100-billion-parameter models and the agentic applications built on them, fully on device.
Two 64GB units clustered together don’t just double the memory. In NVIDIA’s Qwen 3.8 27B test, two clustered 64 GB systems delivered up to 1.7x performance compared with a single system, with room to keep scaling as workloads demand.
DGX Spark ships ready for agent development from day one — NVIDIA Agent Toolkit, CUDA-X AI libraries, Nemotron open models, and popular runtimes like Ollama, vLLM, and PyTorch with CUDA are all supported out of the box. Developers can go from power-on to running models in minutes.
Blender is among the first major creator application providers to support the platform, with a
prebuilt, downloadable installer coming soon
.
Scale Up With NVIDIA Sync Cluster Assistant
Developers can start with the memory their projects need today and build on a platform designed to seamlessly scale multi-node clusters for larger workloads as their pipelines grow.
Every DGX Spark ships with a built-in NVIDIA ConnectX-7 NIC right out of the box. Plus, two units can connect directly with a QSFP cable, pooling their memory to
128GB
and expanding model support to up to
200 billion parameters while delivering twice the memory bandwidth and up to 1.7x the performance.
The
NVIDIA Sync
app configures this multi-node cluster seamlessly. The cluster assistant feature detects connected units, validates device configuration and configures the ConnectX-7 network, so developers can focus on their work rather than the infrastructure. Every node runs the same NVIDIA software stack, so nothing needs to be reconfigured when scaling from one unit to two.
And coming at the end of the month, NVIDIA Sync Model Launcher makes running local AI as simple as clicking a few buttons. Developers can download and launch Qwen3.8 27B on a single DGX Spark system or a cluster, with NVIDIA Sync configuring the model to run across connected devices and making it accessible from users’ laptops. The launcher will also set up OpenCode to use the model, so developers can start coding in their browser.
Developer Use Cases on DGX Spark
The new DGX Spark 64GB configuration supports practical work from day one. With up to 100-billion-parameter models running entirely on device, developers and enthusiasts can start with a single system for models that fit within its memory, or connect multiple DGX Spark systems with NVIDIA Sync Cluster Assistant for workloads that need more memory and compute.
Here are three workflow examples:
Run an AI agent around the clock:
Keep a coding or research agent running on DGX Spark, ready to review code, analyze documents or carry out multistep tasks. A cluster provides additional capacity for larger models, longer context windows or multiple agents working at once.
Power AI apps on your everyday PC:
Run a language- or image-generation model on DGX Spark while using an agent or creative application on laptops or desktops. DGX Spark handles the model inference, freeing PCs for other work.
Scale when the work grows:
When a single task outgrows one unit — running a larger model, a longer context window or concurrent agent requests — two DGX Spark 64GB systems connected over the 200 GbE fabric via NVIDIA Sync Cluster Assistant pool their memory to 128GB. The same workflow that ran on one unit scales to two without reconfiguring the software environment.
Get Started With DGX Spark
DGX Spark 64GB is available from
Acer
,
ASUS
, Dell,
Gigabyte
, HP and
MSI
on Friday, Oct. 23, starting at $4,999.
To get started:
Download a supported inference framework — llama.cpp, Ollama, vLLM or LM Studio.
Download the recommended local model for the workflow.
To scale to two units, connect them via their NVIDIA ConnectX-7 ports and launch NVIDIA Sync Cluster Assistant — it configures the network and routes workloads automatically.
For agentic AI playbooks on DGX Spark, visit
the
NemoClaw
,
OpenClaw
,
Hermes Agent
and
OpenShell
pages on build.nvidia.com.
#ICYMI: More Updates From NVIDIA Local AI
Explore playbooks on
build.nvidia.com/spark
for DGX Spark. The following playbooks are coming soon to 64GB devices:
Serve LLMs With vLLM
Run OpenClaw With a Local LLM
Connect Multiple DGX Sparks for Distributed Workloads
New Windows PCs powered by
NVIDIA RTX Spark
are coming this month from Acer, ASUS, Dell, HP, Lenovo, Microsoft and MSI. Sign up for the
RTX Spark newsletter
to receive future updates.
Alibaba’s
Qwen-Image-2.1
brings image generation and editing together in a lightweight, open-weight model. It runs locally on NVIDIA RTX GPUs, DGX Spark and DGX Station, giving creators more ways to create and refine images on their own hardware.
Follow NVIDIA RTX Spark on
X
,
Instagram
,
TikTok
and
Facebook
— and stay informed by subscribing to the
NVIDIA Local AI newsletter
. Follow NVIDIA Workstation on
LinkedIn
and
X
.
See
notice
regarding software product information.
As AI agents move from experiments into everyday development, increasingly capable open models are shrinking to fit on more devices, giving builders more to run locally.
Coming this month,
NVIDIA DGX Spark
will be available with 64GB of unified memory from top manufacturer partners — Acer, ASUS, Dell, Gigabyte, HP and MSI — giving developers, researchers and AI enthusiasts a new configuration with DGX OS and the NVIDIA AI software stack ready to use from day one.
The new SKU runs capable local agents on device — privately, without cloud dependency. And when workloads grow, two units can cluster together via NVIDIA Sync Cluster Assistant without any additional setup.
A New Starting Point for Personal AI Supercomputing
DGX Spark combines NVIDIA Grace Blackwell compute, unified memory,
NVIDIA ConnectX-7 networking
and an
NVIDIA CUDA
-accelerated AI software stack in one system. It’s a complete local AI platform for agents, inference, fine-tuning, data science and edge development.
The compact, personal AI supercomputer provides a place to experiment with models and developers’ own data without turning to a cloud instance for every task.
The new 64GB configuration, available exclusively from manufacturer partners, keeps the platform at an accessible price point while retaining the GB10 Grace Blackwell Superchip, DGX OS and full NVIDIA AI software stack — same as the 128GB model. It supports up to 100-billion-parameter models and the agentic applications built on them, fully on device.
Two 64GB units clustered together don’t just double the memory. In NVIDIA’s Qwen 3.8 27B test, two clustered 64 GB systems delivered up to 1.7x performance compared with a single system, with room to keep scaling as workloads demand.
DGX Spark ships ready for agent development from day one — NVIDIA Agent Toolkit, CUDA-X AI libraries, Nemotron open models, and popular runtimes like Ollama, vLLM, and PyTorch with CUDA are all supported out of the box. Developers can go from power-on to running models in minutes.
Blender is among the first major creator application providers to support the platform, with a
prebuilt, downloadable installer coming soon
.
Scale Up With NVIDIA Sync Cluster Assistant
Developers can start with the memory their projects need today and build on a platform designed to seamlessly scale multi-node clusters for larger workloads as their pipelines grow.
Every DGX Spark ships with a built-in NVIDIA ConnectX-7 NIC right out of the box. Plus, two units can connect directly with a QSFP cable, pooling their memory to
128GB
and expanding model support to up to
200 billion parameters while delivering twice the memory bandwidth and up to 1.7x the performance.
The
NVIDIA Sync
app configures this multi-node cluster seamlessly. The cluster assistant feature detects connected units, validates device configuration and configures the ConnectX-7 network, so developers can focus on their work rather than the infrastructure. Every node runs the same NVIDIA software stack, so nothing needs to be reconfigured when scaling from one unit to two.
And coming at the end of the month, NVIDIA Sync Model Launcher makes running local AI as simple as clicking a few buttons. Developers can download and launch Qwen3.8 27B on a single DGX Spark system or a cluster, with NVIDIA Sync configuring the model to run across connected devices and making it accessible from users’ laptops. The launcher will also set up OpenCode to use the model, so developers can start coding in their browser.
Developer Use Cases on DGX Spark
The new DGX Spark 64GB configuration supports practical work from day one. With up to 100-billion-parameter models running entirely on device, developers and enthusiasts can start with a single system for models that fit within its memory, or connect multiple DGX Spark systems with NVIDIA Sync Cluster Assistant for workloads that need more memory and compute.
Here are three workflow examples:
Run an AI agent around the clock:
Keep a coding or research agent running on DGX Spark, ready to review code, analyze documents or carry out multistep tasks. A cluster provides additional capacity for larger models, longer context windows or multiple agents working at once.
Power AI apps on your everyday PC:
Run a language- or image-generation model on DGX Spark while using an agent or creative application on laptops or desktops. DGX Spark handles the model inference, freeing PCs for other work.
Scale when the work grows:
When a single task outgrows one unit — running a larger model, a longer context window or concurrent agent requests — two DGX Spark 64GB systems connected over the 200 GbE fabric via NVIDIA Sync Cluster Assistant pool their memory to 128GB. The same workflow that ran on one unit scales to two without reconfiguring the software environment.
Get Started With DGX Spark
DGX Spark 64GB is available from
Acer
,
ASUS
, Dell,
Gigabyte
, HP and
MSI
on Friday, Oct. 23, starting at $4,999.
To get started:
Download a supported inference framework — llama.cpp, Ollama, vLLM or LM Studio.
Download the recommended local model for the workflow.
To scale to two units, connect them via their NVIDIA ConnectX-7 ports and launch NVIDIA Sync Cluster Assistant — it configures the network and routes workloads automatically.
For agentic AI playbooks on DGX Spark, visit
the
NemoClaw
,
OpenClaw
,
Hermes Agent
and
OpenShell
pages on build.nvidia.com.
#ICYMI: More Updates From NVIDIA Local AI
Explore playbooks on
build.nvidia.com/spark
for DGX Spark. The following playbooks are coming soon to 64GB devices:
Serve LLMs With vLLM
Run OpenClaw With a Local LLM
Connect Multiple DGX Sparks for Distributed Workloads
New Windows PCs powered by
NVIDIA RTX Spark
are coming this month from Acer, ASUS, Dell, HP, Lenovo, Microsoft and MSI. Sign up for the
RTX Spark newsletter
to receive future updates.
Alibaba’s
Qwen-Image-2.1
brings image generation and editing together in a lightweight, open-weight model. It runs locally on NVIDIA RTX GPUs, DGX Spark and DGX Station, giving creators more ways to create and refine images on their own hardware.
Follow NVIDIA RTX Spark on
X
,
,
TikTok
and
— and stay informed by subscribing to the
NVIDIA Local AI newsletter
. Follow NVIDIA Workstation on
and
X
.
See
notice
regarding software product information.