Microsoft and NVIDIA Unveil RTX Spark: 1 Petaflop AI Chip for Windows PCs

Microsoft and NVIDIA have unveiled RTX Spark, a new chip delivering 1 petaflop of AI performance for Windows PCs. The unified memory architecture enables local AI workloads previously requiring cloud compute, with devices launching this fall from major manufacturers.

Last Updated: September 13, 2026 Editorial Process
Editorial Process
See more of Inside AI's trusted news by adding us as a preferred source on Google.
AI neural network visualization
Published on: June 22, 2026

June 22, 2026, (Inside AI) — Microsoft and NVIDIA have revealed RTX Spark, a new system-on-chip designed to power what they call the most capable thin-and-light Windows PCs ever built. The chip was announced at NVIDIA GTC and packs 1 petaflop of AI performance, up to 6,144 Blackwell RTX cores, 20 Arm-based CPU cores, and up to 128GB of unified memory into a single silicon package.

The architecture merges GPU, NPU, and CPU functions to handle AI-native computing locally. RTX Spark targets developers, creators, and power users who need to run advanced AI workloads, graphically demanding creative applications, and agent-based tasks without cloud reliance.

Microsoft collaborated directly with NVIDIA to optimize Windows for the chip’s heterogeneous design. This includes workload profile scheduling that distributes tasks across all 20 cores more efficiently. The Microsoft Power and Thermal Framework also maximizes performance-per-watt under sustained loads.

Unified Memory Unlocks On-Device AI

The unified memory architecture marks a leap for local AI. Systems with 128GB allow developers to load large language models entirely on-device. They can run complex rendering projects and build agentic workflows that previously required cloud compute.

Microsoft confirmed that NVIDIA will bring NVIDIA OpenShell to Windows. Hermes Agent and OpenClaw will integrate new Windows security and containment primitives for safe, local agent execution. This move could reshape how developers approach sensitive AI tasks.

Read: Apple's Core AI Lets iPhones Run 70B Language Models Locally

At launch, the app ecosystem includes Blender, DaVinci Resolve, Photoshop, Premiere, CapCut, GitHub Copilot, Claude Code, Cursor, ComfyUI, and CUDA-accelerated PyTorch. Riot Games confirmed League of Legends and VALORANT will run on RTX Spark. PUBG: Battlegrounds, Alan Wake 2, and Naraka: Bladepoint also confirmed compatibility.

From Laptops to Enterprise Supercomputers

RTX Spark-powered Copilot+ PCs will launch this fall from Microsoft Surface, ASUS, Dell, HP, Lenovo, and MSI. Microsoft also outlined a longer-term roadmap scaling Windows from RTX Spark laptops up to the NVIDIA DGX Station for Windows. That system will use the GB300 Grace Blackwell Ultra Desktop Superchip, which the company frames as putting a trillion-parameter AI supercomputer on every enterprise desk.

Industry observers note the push toward local AI compute challenges the cloud-centric model. By embedding powerful AI hardware directly into client devices, Microsoft and NVIDIA could reduce latency, improve privacy, and lower operational costs for developers. However, questions remain about battery life, thermal constraints, and real-world performance of such a dense chip in thin-and-light form factors.

Read: OpenAI and Broadcom Unveil Jalapeño: Custom AI Chip for LLM Inference at Gigawatt Scale in the U.S.

The RTX Spark announcement comes as competitors like Apple and Qualcomm also advance their own Arm-based AI silicon. Microsoft’s deep integration with NVIDIA’s software stack—including CUDA and AI frameworks—may give it an edge in developer adoption. Yet, the true test will be whether the fall hardware delivers on the promised petaflop without throttling.

More from Inside AI

  • AI Policy & Regulation

    Publishers Defend Author Accused of Using AI for Prize-Winning Novel

    September 23, 2026
  • AI Policy & Regulation

    NHTSA Investigates Comma.AI After Crashes Involving Aftermarket Driver-Assistance Devices

    September 23, 2026
  • AI In Business

    Verizon to invest $70 million in AI training effort

    September 23, 2026
  • Cybersecurity AI

    OpenAI Extends Daybreak Cyber Defense Access to Ukraine

    September 23, 2026
  • AI Policy & Regulation

    Pune Deploys 10,000 Police and AI Surveillance for Ganesh Immersion

    September 23, 2026
  • AI Policy & Regulation

    UN Security Council Briefed by AI Leaders Amid Ukraine and Iran Wars

    September 23, 2026
  • AI Hardware & Infrastructure

    Geely Unveils 4-Minute EV Charging System With AI Safety, Takes On BYD

    September 23, 2026
  • Agentic AI

    DeepSeek Unveils DSec Sandbox Infrastructure for Large-Scale Agent Training

    September 23, 2026

Never Miss a Breakthrough

Join 50,000+ readers who get our daily AI intelligence briefing. No fluff, just what matters.

Join Our Newsletter Community

Subscribe

Inside AI is an independent publication covering artificial intelligence news, machine learning research, and the tools shaping the future of technology. No hype. Just what's happening in the AI world.

Topics

  • Artificial Intelligence
  • Machine Learning
  • Generative AI
  • Agentic AI
  • Vibe Coding
  • Prompt Engineering
  • AI Policy & Regulation
  • AI Hardware & Infrastructure
  • AI Tools
  • AI In Business
  • Robotics
  • Cybersecurity AI
  • AI Safety
  • AI Tools & Reviews (Coming soon)

Company

  • Editorial Standards
  • Privacy Policy
  • Terms of Service
  • Contact
  • About Us

Others

  • Press Releases
  • Features
  • Sponsored Content
  • Advertise with us
  • Newsletter

© 2026 Inside AI. All rights reserved.

Designed by Blue Flare Digital