GPU Computing
NVIDIA Launches nvmath-python v1.0, Boosting GPU Math in Python
NVIDIA releases nvmath-python v1.0, bridging Python and CUDA-X for high-performance math operations across CPUs, GPUs, and distributed systems.
NVIDIA Introduces CCCL Runtime to Modernize CUDA Development
NVIDIA's CCCL Runtime brings modern C++ abstractions to CUDA, enabling safer, more efficient GPU programming for developers.
HIVE Stock Drops 11% After Announcing $75M Raise for AI Data Centers
HIVE Digital plans zero-interest notes offering to fund GPU expansion as Bitcoin miners accelerate pivot toward AI infrastructure.
NVIDIA Launches ALCHEMI Toolkit for GPU-Accelerated Chemistry Simulations
NVIDIA releases ALCHEMI Toolkit enabling researchers to build custom atomistic simulation workflows with up to 33x speedups for batched molecular dynamics on GPUs.
NVIDIA nvCOMP Cuts AI Training Checkpoint Costs by $56K Monthly
New GPU compression library reduces LLM training checkpoint sizes by 25-40%, saving teams up to $222K monthly on large-scale model training infrastructure.
NVIDIA Open-Sources Slinky to Run Slurm GPU Workloads on Kubernetes
NVIDIA's Slinky project enables running Slurm clusters on Kubernetes, already deployed on 8,000+ GPU systems for large-scale AI training infrastructure.
NVIDIA Scales AlphaFold-Multimer for Proteome-Wide Protein Complex Prediction
NVIDIA researchers detail GPU-accelerated pipeline extending AlphaFold Database with large-scale protein complex predictions using H100 clusters and optimized workflows.
NVIDIA Unveils Mission Control Software for Blackwell AI Supercomputers
NVIDIA's Mission Control bridges rack-scale GPU hardware with AI workload schedulers, enabling topology-aware job placement on GB200 and GB300 NVL72 systems.
NVIDIA Nsight Tools Slash Vision AI Decode Times by 85% in New VC-6 Batch Mode
NVIDIA's optimized VC-6 batch mode achieves submillisecond 4K image decoding, delivering up to 85% faster per-image processing for AI training pipelines.
NVIDIA GH200 Hits 4.6 Microsecond Latency in Trading Benchmark
NVIDIA's Grace Hopper Superchip achieves record single-digit microsecond inference times in STAC-ML benchmark, challenging FPGA dominance in algorithmic trading.