
EG
Elizabeth Goodman
· 1 min read
EngineeringNVIDIA Technical Blog
Control How Your GPU Shares Work with Green Contexts
GPU applications increasingly consist of multiple independent components running at the same time within a single process: a latency-sensitive operator...
GPU applications increasingly consist of multiple independent components running at the same time within a single process: a latency-sensitive operator alongside a throughput-oriented background kernel; a data preprocessing stage alongside model inference; or multiple stages of a processing workflow sharing a single GPU. Controlling how GPU resources are shared between them remains difficult.
Source
Original source
This story was published by NVIDIA Technical Blog and written by Elizabeth Goodman. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on developer.nvidia.com


