메뉴
BL
The Decoder 57일 전

엔비디아 RTX 스파크, 윈도우 로컬 AI 시대 열다

IMP
9/10
핵심 요약

엔비디아가 최대 128GB 통합 메모리와 1페타플롭 FP4 AI 연산 능력을 갖춘 'RTX 스파크' 칩을 발표하며 윈도우 기기에서 대규모 AI 에이전트를 로컬로 실행할 수 있는 기반을 마련했습니다. 이는 애플 실리콘과 퀄컴 스냅드래곤에 대응하는 그레이스 블랙웰 아키텍처 기반의 ARM 윈도우 전용 칩으로, 에이전트 보안 및 격리를 위한 '오픈쉘 런타임(OpenShell Runtime)' 등 소프트웨어 스택도 함께 제공됩니다. ASUS, Dell, HP, Lenovo, Microsoft 등 주요 OEM업체의 관련 기기는 2026년 가을에 출시될 예정입니다.

번역된 본문

엔비디아, RTX 스파크를 윈도우 기기에서 로컬 AI 에이전트를 실용화하는 칩으로 발표

핵심 요약:

  • 엔비디아는 최대 128GB 통합 메모리와 1페타플롭 FP4 AI 연산 능력을 갖춘 윈도우 랩탑용 그레이스 블랙웰(Grace Blackwell) 칩인 'RTX 스파크(RTX Spark)'를 발표했으며, 이는 애플 실리콘(Apple Silicon)과 퀄컴 스냅드래곤(Qualcomm Snapdragon)과 직접적으로 경쟁합니다.
  • 이 칩은 에이전트 격리 및 개인정보 보호 컨트롤을 위한 '오픈쉘 런타임(OpenShell Runtime)'과 같은 새로운 보안 도구를 기반으로 로컬 AI 에이전트 실행을 목표로 합니다.
  • ASUS, Dell, HP, Lenovo, Microsoft Surface를 포함한 주요 OEM업체의 기기는 2026년 가을에 출시됩니다.

RTX 스파크는 AI 에이전트를 로컬에서 구동하기 위해 설계된 엔비디아의 윈도우 랩탑 첫 진출작입니다. 이 하드웨어는 이미 알려진 DGX 스파크 칩의 윈도우 버전입니다.

엔비디아는 GTC 타이베이에서 RTX 스파크를 공개했습니다. 최상위 모델은 DGX 스파크를 구동하는 것과 동일한 GB10 그레이스 블랙웰 슈퍼칩(Grace Blackwell Superchip)입니다. 차이점은 대상에 있습니다. AI 개발자를 위한 리눅스 워크스테이션 대신, RTX 스파크는 일반 소비자를 위한 윈도우 랩탑과 콤팩트 데스크톱을 타겟으로 합니다.

엔비디아는 다양한 코어 및 SM(Streaming Multiprocessor) 수와 16GB에서 128GB까지의 메모리 옵션을 갖춘 여러 변형 모델을 제공합니다. 최상위 SKU는 6,144개의 CUDA 코어와 5세대 텐서 코어를 갖춘 블랙웰 RTX GPU를 20코어 ARM 기반 그레이스 CPU와 페어링하며, 이는 NVLink-C2C를 통해 연결됩니다. 엔비디아에 따르면 미디어텍(MediaTek)이 CPU 설계를 도왔습니다. 메모리는 최대 128GB이며 CPU와 GPU가 공유합니다. 명시된 1페타플롭의 최고 성능은 희소성(sparsity)이 적용된 FP4 정밀도를 기준으로 한 것으로, 엔비디아 사양상의 이론적 최고치입니다. 엔비디아는 워크로드에 따라 GPU 성능이 지포스 RTX 5070 랩탑 GPU(GeForce RTX 5070 Laptop GPU)와 비슷하다고 밝혔습니다.

애플 실리콘 및 스냅드래곤에 대한 엔비디아의 대안 RTX 스파크는 애플이 2020년 M시리즈 칩으로 제시한 길을 따릅니다. 즉, 하나의 패키지에 ARM CPU, GPU, 메모리 컨트롤러를 탑재하고 개별 비디오 메모리(VRAM) 대신 통합 메모리 풀을 공유하는 방식입니다. 애플의 M4 Max 역시 546GB/s 대역폭으로 최대 128GB의 통합 메모리를 제공하지만, 뉴럴 엔진(Neural Engine)은 38 TOPS(INT8)에 불과합니다. 비교하자면 RTX 스파크는 약 1,000 TOPS를 주장하지만, 이 역시 희소성이 적용된 FP4 기준이므로 비교 조건은 매우 다릅니다. 그럼에도 불구하고 순수 AI 연산 능력의 격차는 상당합니다. 엔비디아의 진정한 강점은 TensorRT 및 RTX를 기본적으로 실행하는 자사의 CUDA 스택에 남아 있습니다.

퀄컴 역시 2024년 스냅드래곤 X 엘리트(Snapdragon X Elite)로 윈도우 온 ARM(Windows-on-Arm) 랩탑 시장에 진출했고, 2025년 9월 X2 엘리트(X2 Elite)를 뒤이어 출시하며 18개의 오리온(Oryon) 코어 전체에 걸쳐 성능을 80 TOPS로 끌어올렸습니다. 이러한 칩들은 수십억 개의 매개변수를 가진 모델의 로컬 추론이 아닌, 마이크로소프트의 코파일럿+(Copilot+) 기능을 중심으로 구축되었습니다. 인텔과 AMD의 전통적인 x86 플랫폼은 여전히 훨씬 작은 NPU를 장착한 별도의 CPU 및 GPU 메모리에 의존하고 있습니다.

로컬 AI 에이전트를 위한 새로운 보안 가드레일 엔비디아는 적절한 보안 도구가 존재하지 않았기 때문에 AI 에이전트가 사용자의 주요 기기에서 거의 실행되지 않았다고 주장합니다. 새로운 윈도우 구성 요소는 ID 관리, 에이전트 격리 및 정책 시행을 제공할 예정입니다. 엔비디아의 오픈쉘 런타임(OpenShell Runtime)은 또 다른 계층을 추가합니다. 이는 에이전트가 수행할 수 있는 작업을 정의하고, 개인정보 설정에 따라 로컬 또는 클라우드 모델로 요청을 라우팅하며, 클라우드 쿼리 시 개인 데이터를 마스킹합니다. 오픈소스 프로젝트인 헤르메스 에이전트(Hermes Agent)와 오픈클로(OpenClaw)는 이미 이 계층을 윈도우 앱에 통합했다고 엔비디아는 밝혔습니다.

어도비(Adobe) 또한 최신 GPU를 위해 포토샵(Photoshop)과 프리미어(Premiere)를 리빌드할 계획을 발표했습니다. 프리미어는 엔비디아 TensorRT 통합이 포함된 새로운 비디오 파이프라인을 도입합니다. 포토샵은 GPU 가속 컴포지팅을 지원하는 새로운 엔진을 탑재합니다. RTX 스파크에서 프리미어는 또한 공유 메모리 풀의 이점을 누릴 수 있습니다. 어도비는 AI, 편집 및 효과 워크플로우의 속도를 최대 2배까지 높이는 것을 목표로 한다고 밝혔습니다.

랩탑 칩과 함께 엔비디아는 윈도우용 DGX 스테이션(DGX Station)을 선보였습니다. 이는 최대 748GB의 공유 메모리와 20 페타플롭의 FP4 성능을 제공하는 GB300 그레이스 블랙웰 울트라 데스크탑 슈퍼칩(Grace Blackwell Ultra Desktop Superchip)을 기반으로 구축되었습니다. 엔비디아에 따르면 이 장치는 모델을 실행할 수 있습니다.

원문 보기
원문 보기 (영어)
Nvidia pitches RTX Spark as the chip that finally makes local AI agents practical on Windows devices Maximilian Schreiner View the LinkedIn Profile of Maximilian Schreiner Jun 1, 2026 Nvidia Key Points Nvidia announced RTX Spark, a Grace Blackwell chip for Windows laptops with up to 128 GB unified memory and 1 petaflop FP4 AI compute, competing directly with Apple Silicon and Qualcomm Snapdragon. The chip targets local AI agent execution, backed by new security tools like OpenShell Runtime for agent isolation and privacy controls. Devices from major OEMs including ASUS, Dell, HP, Lenovo, and Microsoft Surface launch fall 2026. Ask about this article… Search RTX Spark is Nvidia's first move into Windows laptops, designed to run AI agents locally. The hardware is a Windows version of the already-known DGX Spark chip. Nvidia unveiled RTX Spark at GTC Taipei. At the top end, the chip is the same GB10 Grace Blackwell Superchip that powers the DGX Spark . The difference is who it's for. Instead of a Linux workstation aimed at AI developers, RTX Spark targets Windows laptops and compact desktops for consumers. Nvidia is offering several variants with different core and SM counts, plus memory options ranging from 16 to 128 GB. The top SKU pairs a Blackwell RTX GPU with 6,144 CUDA cores and fifth-gen Tensor Cores with a 20-core Arm-based Grace CPU, linked via NVLink-C2C. MediaTek helped design the CPU, according to Nvidia. Memory tops out at 128 GB, shared between CPU and GPU. The claimed 1 petaflop peak refers to FP4 precision with sparsity, a theoretical best case per Nvidia's specs. GPU performance sits close to a GeForce RTX 5070 Laptop GPU depending on the workload, Nvidia says. Ad Nvidia's answer to Apple Silicon and Snapdragon RTX Spark follows the path Apple charted in 2020 with its M-series chips: Arm CPU, GPU, and memory controller on one package, sharing a unified memory pool instead of separate VRAM. Apple's M4 Max also offers up to 128 GB of unified memory at 546 GB/s bandwidth, but its Neural Engine tops out at 38 TOPS (INT8). RTX Spark claims roughly 1,000 TOPS by comparison, though that's FP4 with sparsity, so the conditions are very different. Still, the gap in raw AI compute is significant. Nvidia's real edge remains its CUDA stack, including TensorRT and RTX, which runs natively. Ad DEC_D_Incontent-1 Qualcomm also pushed into Windows-on-Arm laptops with the Snapdragon X Elite in 2024, then followed up in September 2025 with the X2 Elite, boosting performance to 80 TOPS across 18 Oryon cores. Those chips are built around Microsoft's Copilot+ features, not local inference with multi-billion-parameter models. Traditional x86 platforms from Intel and AMD still rely on separate CPU and GPU memory with much smaller NPUs. Local AI agents get new security guardrails Nvidia argues that AI agents rarely run on users' primary devices because the right security tools haven't existed. New Windows components are supposed to provide identity management, agent isolation, and policy enforcement. The Nvidia OpenShell Runtime adds another layer: it defines what agents are allowed to do, routes requests to local or cloud models based on privacy settings, and masks personal data in cloud queries. The open-source projects Hermes Agent and OpenClaw already integrate this layer into their Windows apps, according to Nvidia. Ad Adobe also announced plans to rebuild Photoshop and Premiere for modern GPUs. Premiere is getting a new video pipeline with Nvidia TensorRT integration. Photoshop is getting a new engine with GPU-accelerated compositing. On RTX Spark, Premiere should also benefit from the shared memory pool. Adobe says the goal is AI, editing, and effects workflows that are up to twice as fast. Alongside the laptop chip, Nvidia showed off the DGX Station for Windows . It's built on the GB300 Grace Blackwell Ultra Desktop Superchip with up to 748 GB of shared memory and 20 petaflops of FP4 performance. Nvidia says it can run models with up to a trillion parameters locally. It ships in Q4 2026. Ad DEC_D_Incontent-2 RTX Spark devices will be available starting fall 2026 from ASUS, Dell, HP, Lenovo, Microsoft Surface, and MSI. Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: RTX Spark | Nvidia (Faster Agents) | Windows