NVIDIA’s CUDA 13.4 brings Windows on Arm support and an early look at Rubin GPUs
NVIDIA’s CUDA Toolkit 13.4 adds Windows on Arm support, previews Rubin architecture support and expands GPU-management tools for developers.

CUDA is the software layer that lets much of the modern GPU economy get work done, so a toolkit release can matter even when it lacks the glamour of a new graphics card. Version 13.4’s Windows on Arm support is particularly relevant as Snapdragon-powered PCs try to become serious developer machines rather than attractive battery-life experiments.
What is being reported
NVIDIA says CUDA Toolkit 13.4 adds Windows on Arm support and preview functional support for its Rubin architecture, identified as compute capability 107. The release also updates CUDA Multi-Process Service V3 with a scriptable command-line interface, named server instances, TOML configuration, streaming-multiprocessor partition controls and cgroup-integrated GPU-memory limits.
The toolkit includes a CUDA Compute Fabric Transport API for moving data across NVLink fabric, updates to CUDA Python and improvements in CCCL 3.4. Those are developer-facing changes: they will not make every laptop suddenly run a frontier model, but they can reduce friction for people building, testing and sharing GPU workloads.
Rubin support is preview functionality, not a promise that the next architecture is ready for ordinary buyers. NVIDIA’s announcement is the primary source for this draft, and independent technical testing was not available during preparation.
Our opinion
CUDA 13.4 is a quietly important release. Windows on Arm support gives NVIDIA a better route into a growing PC platform, while the Rubin preview tells developers where the road is heading. The real test will be compatibility and documentation when people try to use it, because nothing turns technological optimism into a support ticket faster than a missing dependency.