Kubernetes

KubeRay v1.7 Launches With Upgraded Features for Ray on Kubernetes
Kubernetes

KubeRay v1.7 Launches With Upgraded Features for Ray on Kubernetes

KubeRay v1.7 introduces major upgrades like History Server beta, enhanced RayJob management, and stronger security for Ray workloads on Kubernetes.

NVIDIA’s KAI Scheduler and vCluster Enable GPU Sharing in Kubernetes
Kubernetes

NVIDIA’s KAI Scheduler and vCluster Enable GPU Sharing in Kubernetes

NVIDIA introduces a solution for isolated Kubernetes clusters and efficient GPU sharing, reducing infrastructure costs for AI/ML teams.

NVIDIA Dynamo Snapshot Tackles Kubernetes AI Cold-Start Problem
Kubernetes

NVIDIA Dynamo Snapshot Tackles Kubernetes AI Cold-Start Problem

NVIDIA's Dynamo Snapshot reduces Kubernetes AI inference cold-start times, leveraging CRIU and GPU Memory Service for sub-5-second deployment speed.

NVIDIA Open-Sources Slinky to Run Slurm GPU Workloads on Kubernetes
Kubernetes

NVIDIA Open-Sources Slinky to Run Slurm GPU Workloads on Kubernetes

NVIDIA's Slinky project enables running Slurm clusters on Kubernetes, already deployed on 8,000+ GPU systems for large-scale AI training infrastructure.

NVIDIA MIG Boosts AI Infrastructure ROI by 33% Over Time-Slicing
Kubernetes

NVIDIA MIG Boosts AI Infrastructure ROI by 33% Over Time-Slicing

New NVIDIA benchmarks show Multi-Instance GPU partitioning achieves 1.00 req/s per GPU versus 0.76 for time-slicing in production AI workloads.

NVIDIA Donates GPU Resource Driver to Kubernetes Open Source Project
Kubernetes

NVIDIA Donates GPU Resource Driver to Kubernetes Open Source Project

NVIDIA transfers critical GPU allocation software to CNCF at KubeCon Europe, marking major shift toward community-governed AI infrastructure.

NVIDIA Advances AI Infrastructure With Disaggregated LLM Inference on Kubernetes
Kubernetes

NVIDIA Advances AI Infrastructure With Disaggregated LLM Inference on Kubernetes

NVIDIA details new Kubernetes deployment patterns for disaggregated LLM inference using Dynamo and Grove, promising better GPU utilization for AI workloads.

NVIDIA Launches AI Cluster Runtime to Standardize GPU Kubernetes Deployments
Kubernetes

NVIDIA Launches AI Cluster Runtime to Standardize GPU Kubernetes Deployments

NVIDIA's new open-source AI Cluster Runtime project delivers validated, reproducible Kubernetes configurations for GPU clusters, targeting H100 and Blackwell accelerators.

NVIDIA Run:ai v2.24 Tackles GPU Scheduling Fairness for AI Workloads
Kubernetes

NVIDIA Run:ai v2.24 Tackles GPU Scheduling Fairness for AI Workloads

NVIDIA's new time-based fairshare scheduling prevents GPU resource hogging in Kubernetes clusters, addressing critical bottleneck for enterprise AI deployments.

Enhancing Kubernetes AI Cluster Stability with NVSentinel
Kubernetes

Enhancing Kubernetes AI Cluster Stability with NVSentinel

NVIDIA introduces NVSentinel, an open-source tool designed to automate health monitoring and issue remediation in Kubernetes AI clusters, ensuring GPU reliability and minimizing downtime.

Trending topics