Machine Learning
Google's Gemini Models Launch Agentic Video Understanding
Google's Gemini 3.7 Flash introduces agentic video understanding, reducing costs by 66% and token usage by 88% for video analysis.
NVIDIA TensorRT Model Connect Simplifies AI Deployment
NVIDIA's TensorRT Model Connect enables AI model deployment from checkpoint to inference in two commands, bridging open models to production.
How Learning Loops Build Durable AI Moats
Learning loops connect data curation, model training, and inference to help companies differentiate AI systems and scale intelligence efficiently.
Harvey Launches Tenet, Open-Weight Legal AI Model
Harvey debuts Tenet, its first post-trained open-weight model, promising improved legal AI performance and cost-efficiency for law firms.
NVIDIA AVO Hits 100% on ARC-AGI-3 Benchmark, Redefining AI Agents
NVIDIA's AVO achieves 100% efficiency on ARC-AGI-3, showcasing a breakthrough in long-horizon autonomous agent systems.
NVIDIA Optimizes JAX LLM Training with Host Offloading
NVIDIA's host offloading for JAX LLM training boosts GPU memory efficiency, enabling larger batch sizes and faster throughput.
Together AI Showcases Nine Papers at ICML 2026 in Seoul
Together AI presents nine groundbreaking papers at ICML 2026, covering AI agents, model efficiency, and GPU optimization. Booth B714, July 6-11.
Ray Data 2.56 Enhances AI Pipelines with Zero OOM Errors
Ray Data 2.56 eliminates OOM errors, reduces memory pressure, and improves AI training/inference speeds by over 50%.
MiniMax M3 Debuts on NVIDIA: 1M Token Context, Multimodal AI
MiniMax M3, the 428B-parameter model, launches on NVIDIA infrastructure, offering long-context reasoning and multimodal workflows for enterprise AI.
NVIDIA's Nemotron 3 Ultra Redefines AI for Long-Running Agents
NVIDIA's Nemotron 3 Ultra, a 550B-parameter AI model, offers faster, cost-efficient reasoning for complex workflows, enabling long-running agents to perform better.