NVIDIA
NVIDIA SDK 1.5 Cuts AI Gaming Costs with Code Agents Over Tool-Calling
NVIDIA's In-Game Inferencing SDK 1.5 introduces code agents that slash GPU inference calls by 66%, enabling richer AI NPCs without crushing frame rates.
NVIDIA Brings CUDA Tile Programming to Julia with cuTile.jl Release
NVIDIA releases cuTile.jl, enabling Julia developers to write high-performance GPU kernels using tile-based programming with near-parity Python performance.
NVIDIA GTC 2026 Set for March 16-19 as Jensen Huang Teases Full AI Stack Reveal
NVIDIA's GTC 2026 conference runs March 16-19 in San Jose with 30,000+ attendees expected. CEO Jensen Huang keynote may unveil rumored new inference chip.
NVIDIA Drops $2B on Coherent as Optical Tech Race Heats Up
NVIDIA invests $2 billion in Coherent to develop silicon photonics for AI data centers, part of a broader $4B push into optical interconnect technology.
NVIDIA Commits $4B to Optics Partners Lumentum, Coherent for AI Data Centers
NVIDIA invests $2 billion each in Lumentum and Coherent to scale silicon photonics for AI infrastructure. LITE stock jumps 12% on the news.
NVIDIA Deploys Alibaba Qwen3.5 VLM on Blackwell GPUs for AI Agent Development
NVIDIA offers free GPU-accelerated endpoints for Alibaba's 397B parameter Qwen3.5 vision-language model, enabling developers to build multimodal AI agents.
NVIDIA Run:ai Delivers 2x GPU Utilization Gains for AI Inference Workloads
NVIDIA benchmarks show Run:ai platform doubles GPU utilization while cutting latency 61x for enterprise AI deployments running NIM inference microservices.
NVIDIA NVFP4 Training Delivers 1.59x Speed Boost Without Accuracy Loss
NVIDIA's NVFP4 4-bit training format achieves 59% faster AI model training than BF16 while matching accuracy on Llama 3 8B benchmarks, per new research.
NVIDIA Partners With Akamai, Siemens to Fortify Critical Infrastructure Security
NVIDIA teams with five major cybersecurity and industrial firms to deploy AI-powered protection for operational technology systems controlling energy, manufacturing, and utilities.
NVIDIA MIG Tech Delivers 2.25x Speedups for Power-Constrained AI Workloads
NVIDIA's Multi-Instance GPU technology shows up to 2.25x performance gains for data center workloads under power limits, with implications for AI infrastructure costs.
NVIDIA Survey Shows 89% of Telcos Boosting AI Spend in 2026
NVIDIA's telecom AI survey reveals 90% of operators report revenue gains from AI adoption, with autonomous networks delivering fastest ROI.
NVIDIA Run:ai GPU Fractioning Delivers 77% Throughput at Half Allocation
NVIDIA and Nebius benchmarks show GPU fractioning achieves 86% user capacity on 0.5 GPU allocation, enabling 3x more concurrent users for mixed AI workloads.
NVIDIA cuda.compute Brings C++ GPU Performance to Python Developers
NVIDIA's new cuda.compute library topped GPU MODE benchmarks, delivering CUDA C++ performance through pure Python with 2-4x speedups over custom kernels.
NVIDIA Blackwell Delivers 4x Inference Boost for India's Sarvam AI Models
NVIDIA's hardware-software co-design achieves 4x inference speedup for Sarvam AI's 30B parameter sovereign models, showcasing Blackwell's NVFP4 capabilities.
India Deploys 20,000 NVIDIA Blackwell GPUs in $1B AI Infrastructure Push
India partners with NVIDIA to build sovereign AI infrastructure with 20,000+ Blackwell Ultra GPUs, targeting $27.7B market by 2032 under IndiaAI Mission.