惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

月光博客
月光博客
人人都是产品经理
人人都是产品经理
博客园 - 聂微东
WordPress大学
WordPress大学
S
SegmentFault 最新的问题
博客园 - Franky
V
V2EX
Y
Y Combinator Blog
Google DeepMind News
Google DeepMind News
J
Java Code Geeks
T
The Blog of Author Tim Ferriss
罗磊的独立博客
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Jina AI
Jina AI
博客园 - 叶小钗
F
Fortinet All Blogs
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
A
About on SuperTechFans
M
MIT News - Artificial intelligence
云风的 BLOG
云风的 BLOG
Last Week in AI
Last Week in AI
D
Docker
博客园 - 【当耐特】
阮一峰的网络日志
阮一峰的网络日志

The Decoder

Google files first joint lawsuit with FBI over Chinese AI scam network, OpenAI blocks PRC influence clusters The AI industry's platform trap is starting to look a lot like Microsoft's OpenAI buys Ona to push Codex toward long-running, autonomous coding tasks Jeff Bezos' AI startup Prometheus closes $12 billion round at a $41 billion valuation Free Deezer tool lets users on any streaming service check their playlists for AI music OpenAI vs. Anthropic: A price war over API tokens is brewing Dario Amodei's new essay reads like a Cold War playbook for the AI age Claude Fable 5: Anthropic admits "wrong tradeoff" after invisibly throttling rival AI researchers Google's new open model DiffusionGemma generates text from noise instead of word by word OpenAI's IPO slips as Altman tells staff to expect a public offering "within the next year" Anthropic study shows AI needs hours, not weeks, to build exploits from security patches OpenAI wants its biggest data center yet, and Nvidia would back the bill Claude Fable 5: The first Mythos model is powerful, expensive, and heavily filtered Germany's National Security Council greenights an AI Safety Institute modeled after the UK's AISI Google's NotebookLM now runs its own cloud computer with code execution and agent-based research Anthropic releases Claude Fable 5 and Mythos 5 with major gains in coding and science Google's Gemini 3.5 Live Translate delivers real-time voice translation across 70+ languages SpaceX wants to put data centers in orbit, and Musk says it's no big deal Landmark German ruling declares Google's AI Overviews are Google's own words and makes it liable for false answers Beijing's $295 billion AI buildout would require 80 percent domestic chips, locking out US suppliers Apple Intelligence gets a second shot with help from Google and Nvidia OpenAI now says "entirely automating everything is not the future we want" OpenAI says going public is "a complicated set of tradeoffs" and is unsure about the timing Microsoft Research's Lens proves detailed captions matter more than raw scale for training efficient image generators Intel gets a second life as Google and Nvidia explore it as a TSMC backup for AI chips Most companies are flying blind on AI spending Frontier Radar #3: How agentic AI is turning tokens into a business metric Instagram AI chatbot breach may have affected over to 20,000 accounts, Meta discloses Microsoft tightens rules for conflict zones after investigation into Israel's military use of Azure Moonshot AI targets a $30 billion valuation, more than six times its late-2025 worth
Nvidia pitches RTX Spark as the chip that finally makes l...
Maximilian Schreiner · 2026-06-01 · via The Decoder

RTX Spark is Nvidia's first move into Windows laptops, designed to run AI agents locally. The hardware is a Windows version of the already-known DGX Spark chip.

Nvidia unveiled RTX Spark at GTC Taipei. At the top end, the chip is the same GB10 Grace Blackwell Superchip that powers the DGX Spark. The difference is who it's for. Instead of a Linux workstation aimed at AI developers, RTX Spark targets Windows laptops and compact desktops for consumers. Nvidia is offering several variants with different core and SM counts, plus memory options ranging from 16 to 128 GB.

The top SKU pairs a Blackwell RTX GPU with 6,144 CUDA cores and fifth-gen Tensor Cores with a 20-core Arm-based Grace CPU, linked via NVLink-C2C. MediaTek helped design the CPU, according to Nvidia. Memory tops out at 128 GB, shared between CPU and GPU. The claimed 1 petaflop peak refers to FP4 precision with sparsity, a theoretical best case per Nvidia's specs. GPU performance sits close to a GeForce RTX 5070 Laptop GPU depending on the workload, Nvidia says.

Nvidia's answer to Apple Silicon and Snapdragon

RTX Spark follows the path Apple charted in 2020 with its M-series chips: Arm CPU, GPU, and memory controller on one package, sharing a unified memory pool instead of separate VRAM. Apple's M4 Max also offers up to 128 GB of unified memory at 546 GB/s bandwidth, but its Neural Engine tops out at 38 TOPS (INT8). RTX Spark claims roughly 1,000 TOPS by comparison, though that's FP4 with sparsity, so the conditions are very different. Still, the gap in raw AI compute is significant. Nvidia's real edge remains its CUDA stack, including TensorRT and RTX, which runs natively.

Qualcomm also pushed into Windows-on-Arm laptops with the Snapdragon X Elite in 2024, then followed up in September 2025 with the X2 Elite, boosting performance to 80 TOPS across 18 Oryon cores. Those chips are built around Microsoft's Copilot+ features, not local inference with multi-billion-parameter models. Traditional x86 platforms from Intel and AMD still rely on separate CPU and GPU memory with much smaller NPUs.

Local AI agents get new security guardrails

Nvidia argues that AI agents rarely run on users' primary devices because the right security tools haven't existed. New Windows components are supposed to provide identity management, agent isolation, and policy enforcement. The Nvidia OpenShell Runtime adds another layer: it defines what agents are allowed to do, routes requests to local or cloud models based on privacy settings, and masks personal data in cloud queries. The open-source projects Hermes Agent and OpenClaw already integrate this layer into their Windows apps, according to Nvidia.

Adobe also announced plans to rebuild Photoshop and Premiere for modern GPUs. Premiere is getting a new video pipeline with Nvidia TensorRT integration. Photoshop is getting a new engine with GPU-accelerated compositing. On RTX Spark, Premiere should also benefit from the shared memory pool. Adobe says the goal is AI, editing, and effects workflows that are up to twice as fast.

Alongside the laptop chip, Nvidia showed off the DGX Station for Windows. It's built on the GB300 Grace Blackwell Ultra Desktop Superchip with up to 748 GB of shared memory and 20 petaflops of FP4 performance. Nvidia says it can run models with up to a trillion parameters locally. It ships in Q4 2026.

RTX Spark devices will be available starting fall 2026 from ASUS, Dell, HP, Lenovo, Microsoft Surface, and MSI.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Subscribe now