Nvidia's 616.00 driver and CUDA 13.4 preview let developers begin native Windows Arm64 work for RTX Spark, but the release is an ecosystem-building step—not evidence of final performance, compatibility or pricing.
Nvidia has released a native driver and CUDA preview for preparing Windows applications for RTX Spark before the PCs arrive. The milestone is narrower than a product launch: developers can now build against the platform, but Nvidia expressly bars using this preview to judge final performance or production readiness.

An illustration from Microsoft’s RTX Spark announcement shows a chip beneath a transparent laptop. Source: Windows Experience Blog.
The release is a native Windows 11 Arm64 package aimed at development systems. An inspection of the installation files found an nv_surface_woa.inf file and device IDs associated with two N1X configurations, one listing 6,144 cores and the other 5,120. A separate package inspection reported the same entries.
Those identifiers show which hardware the preview can recognize; they are not confirmation that every retail system will use either configuration. The same package also includes an unnamed desktop-device entry, but its commercial identity has not been disclosed. Reporting on the release says version 616.00 initially targets Microsoft's Surface RTX Dev Box, a developer system rather than a consumer PC.
Native support matters at the driver layer. Microsoft says in its Windows on Arm guidance that kernel-mode drivers must be built as Arm64 binaries and that an x86 or x64 setup program cannot install an Arm64 driver. Application emulation cannot substitute for that work.
CUDA Toolkit 13.4 Developer Preview adds compiler and build-system enablement for Windows Arm64 targets. Nvidia's release notes say developers can build natively for the forthcoming RTX Spark device or cross-compile Windows Arm64 CUDA applications with the x86-64 toolkit. Drivers, firmware and system images are not part of the toolkit, and new CUDA 13.4 features or newly enabled platforms require the separately released R616 driver at version 616.00 or later.
Developers can begin before RTX Spark hardware is generally available. Nvidia's porting instructions direct them to use an existing Windows on Arm machine, audit third-party dependencies, choose Arm64 or Arm64EC, build an initial version, and later validate it on supported RTX Spark hardware and software.
That process also exposes a limit to Nvidia's leverage. The application layer is not uniquely Nvidia's: Windows already supports native Arm64 software and emulates many x86 and x64 applications, and the same Microsoft guidance warns that emulation carries performance overhead. It also identifies unsupported third-party dependencies as a potential block to a fully native build. RTX Spark can bring CUDA and RTX to Windows on Arm, but it still depends on developers, Microsoft and suppliers outside Nvidia to complete the experience.

Microsoft’s official RTX Spark announcement depicts three laptops displaying different applications as part of its app-ecosystem presentation. Source: Windows Experience Blog.
Nvidia says the full RTX Spark design combines a 20-core Grace CPU co-designed with MediaTek and a Blackwell GPU with as many as 6,144 CUDA cores. The company claims up to one petaflop of AI compute and 128GB of unified memory, and says more than 100 Windows software providers and game developers are supporting the platform in its announcement. None of those maximums establishes the performance of a particular laptop or desktop configuration.
Microsoft has described substantial platform work around the chip. It says in its platform announcement that Windows uses workload profile scheduling across the 20 CPU cores, raises the amount of system memory accessible to the GPU on high-memory systems, and tunes the Prism emulator for RTX Spark's microarchitecture. Microsoft also lists native Arm versions of creative applications including Blender, DaVinci Resolve and Photoshop, while describing other software, such as MATLAB, as supported through Prism.
That mix is more informative than the partner count. Some applications are native, some are emulated, and others are still being optimized or promised. The preview gives developers a way to reduce that uncertainty, but it does not measure how much of a buyer's actual workflow will be native at launch.
Nvidia labels CUDA 13.4 pre-release software that may be incomplete or contain design flaws. Its release notes prohibit production deployment, business-critical use and benchmarking. They also say performance data from the preview is non-representative and must not be used to characterize Nvidia hardware or software.
The developer notice identifies several specific problems:
These are disclosed development-stage issues, not findings about finished retail systems. But they affect the installation, AI build and data-transfer workflows the preview is intended to test. Until Nvidia releases validated software, even favorable results would have limited value as a retail comparison.
The commercial comparison is similarly incomplete. Windows on Arm PCs have previously centered on Qualcomm's Snapdragon platform, according to the retained market context, while conventional x86 Windows PCs remain another alternative. Nvidia has not announced RTX Spark retail prices. IDC forecasts global PC shipments will fall 11.3% for full-year 2026 and global PC average selling prices will rise 18.3%, which it attributes to a persistent memory shortage expected to see no meaningful relief before the end of 2027. That forecast is not an RTX Spark cost estimate; it only describes a difficult market backdrop for a platform advertising configurations with up to 128GB of memory.
Nvidia says RTX Spark laptops and compact desktops will be available in the fall from ASUS, Dell, HP, Lenovo, Microsoft Surface and MSI, with Acer and GIGABYTE models to follow. Version 616.00 supports the development work behind that plan, but a seasonal availability claim is not a firm date.
The launch case now turns on evidence the preview cannot supply:
Until then, the 616.00 release is best understood as an ecosystem checkpoint. Nvidia has supplied a real native driver and a path for developers to find porting failures early. It has not yet shown that the resulting PCs will deliver the reliability, coverage or economics required to compete beyond that development program.
Get concise AI news and useful context from the Magica team.
Read the newsletterKimi says a 48-hour demand surge pushed its GPU capacity close to the limit, forcing a pause in new subscriptions as a separate Kimi Code plan prepares to change how coding access is bundled and rationed.
Ant Digital Technologies has expanded Agentar with 200 preconfigured job-role templates and a multi-agent management pitch, but it has not disclosed pricing, customer use or performance data—and governance is becoming a market-wide requirement rather than a distinctive feature.
Apple is reportedly testing an opt-in tool that transcribes and summarizes Genius Bar appointments. Its current safeguards are clear, but its accuracy, retention rules and uses after the pilot are not.
Moonshot AI is seeking investor approval for a Hong Kong IPO within six months while an unfinished private round could value it above $30 billion. The pitch pairs rapid reported recurring-revenue growth with Kimi K3, but neither audited financials nor enough independent model and deployment evidence is public yet.
Andy Serkis says machine learning has a narrow role in The Hunt for Gollum’s de-aging work. That boundary remains a production claim, because the actors, shots, tools, data, labor effects and likeness terms have not been disclosed.
Nebius Group revenue rose 684% to $399 million in the first quarter of 2026 while purchases of property, equipment and intangible assets reached $2.47 billion. Microsoft and Meta reduce demand risk, but options and an unsold-capacity backstop leave delivery, financing and unit economics unresolved.
Chinese-developed models have overtaken U.S. rivals in token volume on OpenRouter, where low prices and token-heavy agent workloads favor DeepSeek V4. The crossover covers a small, platform-specific slice of AI use and does not establish leadership in revenue, enterprise demand or infrastructure control.
OpenAI said it had identified the cause of elevated ChatGPT errors and was applying mitigations, but the retained status update did not disclose the technical failure, measure the impact or confirm full recovery.
Alibaba has opened hosted access to Qwen3.8-Max-Preview and says the 2.4-trillion-parameter model will be released with downloadable weights, but the license, architecture, benchmark evidence, release date and model-level credit economics remain undisclosed.
Anthropic has put Bun’s Rust port into Claude Code ahead of Bun 1.4’s general release, creating a real but tightly controlled proving ground for an AI-led migration whose total cost and broader reliability remain unsettled.
Zhipu reportedly reached $1 billion in annual recurring revenue in July, roughly four times a March estimate, but the unconfirmed run rate is not annual sales and still sits far ahead of recognized cloud revenue while margins remain thin.
An account of PNC transaction data puts household-paid generative AI near 2%, while a separate user survey finds much broader paid access when employer-funded plans count. The gap shows why card charges alone cannot settle whether consumer AI is becoming a mass subscription business.
Kimi K3 reduces attention traffic, but Moonshot recommends deploying it across at least 64 accelerators. SemiAnalysis says expert routing will more than erase the bandwidth savings; until the promised weights are deployed independently, that remains a hardware thesis rather than a measured result.
Morgan Stanley raised its Micron fiscal-2027 gross-margin estimate to 89.3%, but Micron’s results show the forecast depends chiefly on exceptional memory pricing, customer contracts and delayed supply rather than a disclosed HBM4 margin advantage.
New Mexico’s land commissioner refused to reconsider state-land crossings for a pipeline serving Project Jupiter, preserving a fuel-supply obstacle for the planned Oracle data center. But an analyst’s 2029 forecast predates that decision and remains at odds with Oracle’s first-half-2027 delivery statement.
A reported CIA mission examined whether an influential Emirati sheikh could be trusted with sensitive U.S. technology. The public export rule that followed gives G42 and Core42 a narrow, temporary exception, but it does not connect the intelligence operation to that decision or show how compliance will be tested.
Tracebit says a guardrail-triggering string in one decoy AWS secret sharply reduced five AI agents' success in a 152-run cyber range, but the company-run test did not cover uncensored models, adaptive attackers or production deployments.
SenseTime’s U1 Pro preview combines a claimed native 8K ceiling with a multi-step image-creation loop, but its August API will need to disclose dimensions, latency, pricing and repeatable results before buyers can compare the cost of a usable asset.
Alibaba Cloud has begun invite-only testing of a 64-card Zhenwu M890 supernode instance, extending its in-house chip and infrastructure stack to outside users while leaving the service's price, benchmark methodology and customer economics undisclosed.
Open Design has attracted nearly 80,000 GitHub stars with an open, model-flexible answer to Claude Design, but its million-install claim has no published methodology and the team has not disclosed the retention and revenue figures needed to judge the business.