NVIDIA
NVIDIA Dynamo Enhances Streaming for Agentic Workflows
NVIDIA Dynamo introduces new tools for faster, more accurate agentic workflows, improving token streaming and tool-call handling.
NVIDIA GB200 NVL72 Redefines Rack-Scale AI with Slurm Block Scheduling
NVIDIA's GB200 NVL72 brings exascale AI to rack-scale computing, leveraging Slurm block scheduling for efficiency. A game-changer for trillion-parameter models.
NVIDIA Model Optimizer Brings FP8 Quantization to CLIP Models
NVIDIA's Model Optimizer enhances AI efficiency with FP8 quantization for CLIP models, reducing VRAM use while maintaining performance.
NVIDIA and IREN Partner on 5GW AI Infrastructure Expansion
NVIDIA (NVDA) and IREN (IREN) announce a strategic alliance to deploy up to 5GW of AI infrastructure, including a $2.1B potential stock investment.
NVIDIA Launches Real-Time NCCL Monitoring with Prometheus
NVIDIA introduces real-time NCCL Inspector with Prometheus integration, enhancing AI workload debugging and monitoring with Grafana visualization.
NVIDIA GeForce NOW Adds Gaijin Single Sign-On, Expands Library
NVIDIA GeForce NOW introduces Gaijin single sign-on for seamless gaming and adds seven titles, including Dead as Disco and PowerWash Simulator 2.
NVIDIA Expands Spectrum-X With Open MRC Protocol for AI Scale
NVIDIA's Spectrum-X Ethernet integrates the new MRC protocol, optimizing AI network performance for hyperscale data centers like OpenAI and Microsoft.
NVIDIA and Corning Partner to Boost US AI Infrastructure Manufacturing
NVIDIA (NVDA) and Corning (GLW) team up to expand US optical manufacturing for AI infrastructure, creating 3,000+ jobs and advancing AI-driven data centers.
NVIDIA and ServiceNow Expand AI Collaboration with Project Arc
NVIDIA and ServiceNow unveil Project Arc, a secure autonomous AI agent for enterprise workflows, powered by NVIDIA's tech. Here's what it means.
NVIDIA’s Agentic AI Vision: Extreme Co-Design and Vera Rubin
NVIDIA's extreme co-design platform, Vera Rubin, tackles AI agent complexity with advanced tools for scalable, cost-efficient generative AI systems.