AI Infrastructure
Ray 2.55 Adds Fault Tolerance for Large-Scale AI Model Deployments
Anyscale's Ray Serve LLM update enables DP group fault tolerance for vLLM WideEP deployments, reducing downtime risk for distributed AI inference systems.
Together AI Kernels Team Achieves 3.6x Performance Gains on NVIDIA Hardware
Together AI's kernel research team delivers major GPU optimization breakthroughs, cutting inference latency from 281ms to 77ms for enterprise AI deployments.
NVIDIA Blackwell Ultra GPUs Crush MLPerf Benchmarks with 2.7x Performance Gains
NVIDIA's Blackwell Ultra GPUs set new MLPerf Inference records with 2.7x faster DeepSeek-R1 processing, hitting 2.5 million tokens per second across 288 GPUs.
Bitfarms Becomes Keel Infrastructure, Completes Delaware Move Amid Bitcoin Exit
Former Bitcoin miner Bitfarms officially rebrands as Keel Infrastructure, completing U.S. redomiciliation as it pivots to 2.2GW AI data center business.
Oracle Brings NVIDIA B300 GPUs and xAI Grok to Government Cloud Regions
Oracle expands AI infrastructure for U.S. government customers with NVIDIA Blackwell Ultra GPUs and xAI Grok models in secure cloud regions.
Filecoin (FIL) Onchain Cloud Hits Mainnet With 49 TiB Already Stored
Filecoin (FIL) launches programmable cloud storage for AI agents with onchain proofs, automatic payments, and two-copy replication at $2.50/TiB monthly.
NVIDIA MIG Boosts AI Infrastructure ROI by 33% Over Time-Slicing
New NVIDIA benchmarks show Multi-Instance GPU partitioning achieves 1.00 req/s per GPU versus 0.76 for time-slicing in production AI workloads.
NVIDIA Claims 1 Million X Efficiency Gains Across Six GPU Generations
NVIDIA details how Vera Rubin platform delivers 10x higher inference throughput per megawatt, reshaping AI data center economics and token factory revenue models.
Ray Serve Upgrade Delivers 88% Lower Latency for AI Inference at Scale
Anyscale announces major Ray Serve optimizations with HAProxy and gRPC, achieving 11.1x throughput gains for LLM inference workloads on enterprise deployments.
NVIDIA Donates GPU Resource Driver to Kubernetes Open Source Project
NVIDIA transfers critical GPU allocation software to CNCF at KubeCon Europe, marking major shift toward community-governed AI infrastructure.