# CUDA 13.1 brings GPU resource partitioning to Runtime API

_Published Tuesday, October 6, 2026 at 11:05 AM EDT · AI, Infrastructure · Latest · Tier 2 — Notable_

![CUDA 13.1 brings GPU resource partitioning to Runtime API — Primary](https://developer-blogs.nvidia.com/wp-content/uploads/2026/10/image1-3.webp)

NVIDIA says CUDA 13.1 makes green contexts accessible through its Runtime API, allowing applications to assign concurrent workloads to selected GPU execution resources. The feature has been available through the Driver API since CUDA 12.4.

Applications can divide GPU compute units between workloads and provision workqueues to reduce unintended serialization. This gives developers a way to reserve resources for latency-sensitive tasks alongside background work, while retaining the existing stream-based programming model.

Dedicated resources can let critical work avoid waiting for bulk tasks to release compute units. The tradeoff is that fewer units remain available to the bulk workload; applications can adopt the feature incrementally.

## Sources

- [NVIDIA Technical Blog](https://developer.nvidia.com/blog/control-how-your-gpu-shares-work-with-green-contexts/)

---
Canonical: https://techandbusiness.org/newswire/LyibgfeGDxfQoM_T9InbU0
Published: 2026-10-06T15:05:29.773Z
Story chronology: 2026-10-06T15:00:00.000Z
Retrieved: 2026-10-06T16:59:30.308Z
Publisher: Tech & Business (techandbusiness.org)
