What Is GPU Time-Slicing?
A technique for sharing a single physical GPU among multiple CDE workspaces by dividing GPU time into slices allocated to different users.
Definition
A technique for sharing a single physical GPU among multiple CDE workspaces by dividing GPU time into slices allocated to different users.
GPU time-slicing enables cost-effective AI/ML development in CDEs without dedicating an entire GPU to each developer, though it trades throughput for better resource utilization.
Where This Fits
This term is covered in depth on GPU and Accelerated Computing in CDEs.
GPU-accelerated CDEs for AI model training, 3D rendering, and scientific computing. NVIDIA A100, H100, H200 workspace provisioning.
Related Terms
Back to the full CDE and agent infrastructure glossary.
