Wrivio
Get Wrivio
5 min readBy Wrivio Team

Microsoft Build 2026: Local AI Is No Longer Copilot+ Only

For two years the pitch was simple and a little exclusionary: if you wanted on-device AI on Windows, you needed a Copilot+ PC with a fast neural processor. At Build 2026 Microsoft walked that back. The new position is that local AI features run across CPUs, GPUs, and NPUs on a broad range of Windows hardware, not only on machines with a 40-plus TOPS neural unit.

That is a genuinely good change for anyone who wants to keep their writing on their own machine, and it is worth understanding what it actually unlocks, because the marketing will overstate it and the hardware reality will not.

The NPU-Only Gate Came Down

Microsoft’s earlier framing tied a whole class of on-device features to Copilot+ certification, which required a 40-plus TOPS NPU, 16 GB of RAM, and a 256 GB SSD. Build 2026 reframed that: as reported from the event, Microsoft announced a local AI platform that runs models and agents across CPUs, GPUs, and NPUs on any Windows hardware, and said it would stop marketing a single NPU TOPS score as the line between “can” and “cannot.”

Read plainly, this admits what people running local models already knew: a capable GPU, or even a modern CPU with enough RAM, can run a small language model perfectly well for everyday tasks. The NPU was never the only door. Tying features to it was a certification decision, not a physics one.

What This Actually Enables For Writing

The practical unlock is that you no longer need to buy a specific badge of laptop to run a small model locally for drafting and rewriting. A small instruction-tuned model doing tone and clarity work does not need a data center or a 40 TOPS accelerator. It needs a few gigabytes of RAM and a processor made in the last few years.

This is the ground Wrivio already stood on. Its Local engine runs an Apache 2.0 Qwen3 model in-process, on ordinary Windows hardware, with zero network calls during a rewrite. The Build announcement is Microsoft catching up to the idea that local writing assistance should not be gated behind a hardware tier, rather than a new capability you had to wait for.

If you have wondered whether your machine qualifies, whether you need a Copilot+ PC for local AI answers it: for writing-sized models, usually not, and Build 2026 makes that official rather than merely true.

Do Not Confuse “Runs” With “Runs Fast”

The honest caveat is speed. An NPU is good at small, sustained on-device models at low power, which is exactly the writing case, but it is not fast in raw throughput. Reported figures put an 8B model at roughly 5 to 10 tokens per second on a mobile NPU versus around 100 on a desktop GPU. For a short rewrite that streams as you read, single-digit tokens per second is fine. For bulk generation it is not.

So “any hardware can run local AI” is true and useful, and it does not mean every machine gives the same experience. The tradeoffs between NPU, CPU, and GPU for local writing still hold; Build 2026 changed the permission model, not the performance curve.

Before:

Microsoft says every Windows PC can now run local AI, so hardware does not matter anymore.

After:

Microsoft opened on-device AI beyond Copilot+ PCs at Build 2026, so a wider range of machines can run local models. Speed still depends on your CPU, GPU, and RAM, so “runs” and “runs fast” are different questions.

The second version survives a follow-up from someone about to spend money on a laptop.

A Wrivio Context for summarizing a hardware announcement could say:

Rewrite this as a short, neutral summary of a hardware or platform announcement. Keep every product name, spec figure, and date exactly as written. Preserve any distinction between what a device can do and how fast it does it. Do not upgrade a “supported” claim into a “performs well” claim.

Press Ctrl+Shift+Space, paste the draft, and check the diff for any place a cautious “supported on” quietly became an enthusiastic “runs great on.”

Common Questions

Do I still need a Copilot+ PC for local AI on Windows?

Not for writing-sized models. At Build 2026 Microsoft opened on-device AI to CPUs and GPUs across a broad range of Windows hardware, rather than restricting it to Copilot+ machines with a fast NPU. A capable non-Copilot+ PC can run a small local model for drafting and rewriting.

Why did Microsoft drop the NPU-only requirement?

Reporting from the event indicates Microsoft acknowledged the NPU developer ecosystem had not matured as hoped and that a single TOPS score was a poor purchasing signal. Opening the platform to GPUs and CPUs reflects that a small model does not require a dedicated neural unit to run usefully.

Will a local model run as fast on a laptop as on a desktop GPU?

No. Small on-device models run at single-digit to low-double-digit tokens per second on mobile NPUs, versus roughly 100 on a strong desktop GPU. For short, streamed rewrites that is comfortable; for heavy generation the gap matters.

Does this change anything about how Wrivio’s Local engine works?

Not mechanically. Wrivio’s Local engine already ran a small model in-process on ordinary Windows hardware with no network calls during a rewrite. The Build announcement validates that approach at the platform level rather than requiring any change to it.

Is on-device AI more private than cloud AI?

For the text itself, yes: a rewrite that runs locally does not send your draft anywhere. That privacy comes from where the computation happens, not from the hardware badge, which is why opening local AI to more machines is a privacy win as well as a cost one.

Download Wrivio for Windows to run a local model for your everyday writing today, on the hardware you already own.