Surface RTX Spark Dev Box $5,999: price local AI first
Microsoft's Surface RTX Spark Dev Box costs $5,999 and the Surface Laptop Ultra starts at $2,599. Price local AI against your API bill before you preorder.
Microsoft put a price on local AI yesterday. The Surface RTX Spark Dev Box is $5,999, and the Surface Laptop Ultra starts at $2,599. Both run on Nvidia's RTX Spark chip, and Microsoft says both run models over 120 billion parameters on the device. The pitch is AI with no per-token bill. The real question for a small business is simpler: how many months of API spend does $6,000 buy?
What actually happened
At its October 7 event in San Francisco, Microsoft opened preorders. Per the Windows devices blog:
- Surface Laptop Ultra: from $2,599, up to 128GB of unified memory, available October 16. Business configurations are on preorder now. A trade-in offer gives up to $1,000 back on an eligible MacBook Pro through November 23 (US and Canada, Microsoft Store online).
- Surface RTX Spark Dev Box: $5,999 with 128GB of unified memory, sold only on Microsoft.com in the US, shipping in November.
- Performance claim: up to 1 petaflop of AI compute. Microsoft's own footnote says that is theoretical FP4 with sparsity, not a sustained number.
TechCrunch reports the top Laptop Ultra configuration runs $5,900, and that Microsoft says the highest-end model is already out of stock. The Dev Box ships with VS Code, GitHub Copilot CLI, WSL, and PowerShell 7.
On the software side, Microsoft's Windows Experience blog says Windows ML now supports llama.cpp, and that local-model routing through GitHub HydraFusion reaches an experimental preview later in October.
Why it matters for your business
Do the break-even math. If your team spends $300 a month on API tokens, a $5,999 box takes 20 months to pay back. That assumes the local model does the job as well, and it ignores power and your time. At $50 a month, the math does not work. Cheap cloud models like Claude Haiku 5.5 move the bar up again.
Privacy is the better reason. If you handle client files, medical notes, or contracts you can't send to a third-party API, a local box can pay off even when the token math doesn't.
128GB is shared memory. Microsoft's footnote says the GPU can address less than the total. A 120B-parameter model at low precision fits. A 120B model at full precision does not. Ask what quantization you will run before you buy.
Hardware is the easy part. You still need patching, backups, and someone who owns the box. A local model server is a server.
Key takeaways
- Surface RTX Spark Dev Box: $5,999, 128GB unified memory, US-only, ships in November
- Surface Laptop Ultra: from $2,599, up to 128GB, available October 16, business configs on preorder
- The "1 petaflop" claim is theoretical FP4 with sparsity
- Divide the price by your monthly API spend to get a payback period before you buy
- Data privacy, not token cost, is the stronger case for local AI at small-business volume
Local box or cloud API? We build AI workflows that run on either, with the model endpoint in config, so you can start on the API and move work local when the math says so. Run your numbers in our ROI calculator, or talk to us before you spend $6,000.
Sources: Microsoft Windows Blog, TechCrunch.
- #surface-rtx-spark
- #local-ai
- #nvidia
- #windows-11
- #ai-hardware
Tommy Rush — Founder, Rush Commerce
Operator turned builder. 15+ years running operations — now shipping the systems businesses run on. More
Get The Rush Report weekly — one email, zero fluff.
Keep reading
Nano Banana 2.1 halves AI image prices: check the input tokens
Google's Nano Banana 2.1 cuts Gemini API image output to $0.0336 per 1K image, but input tokens cost 3x more. How to price your product-photo pipeline.
Read itOpenDocRouter: document parsing at $0.80 per 1,000 pages
LlamaIndex's OpenDocRouter puts 11 document parsing models behind one API with published ParseBench scores and cost per 1,000 pages. Route by document, not by brand.
Read it