Machine Learning

Google's Gemini Models Launch Agentic Video Understanding
Machine Learning

Google's Gemini Models Launch Agentic Video Understanding

Google's Gemini 3.7 Flash introduces agentic video understanding, reducing costs by 66% and token usage by 88% for video analysis.

NVIDIA TensorRT Model Connect Simplifies AI Deployment
Machine Learning

NVIDIA TensorRT Model Connect Simplifies AI Deployment

NVIDIA's TensorRT Model Connect enables AI model deployment from checkpoint to inference in two commands, bridging open models to production.

How Learning Loops Build Durable AI Moats
Machine Learning

How Learning Loops Build Durable AI Moats

Learning loops connect data curation, model training, and inference to help companies differentiate AI systems and scale intelligence efficiently.

Harvey Launches Tenet, Open-Weight Legal AI Model
Machine Learning

Harvey Launches Tenet, Open-Weight Legal AI Model

Harvey debuts Tenet, its first post-trained open-weight model, promising improved legal AI performance and cost-efficiency for law firms.

NVIDIA AVO Hits 100% on ARC-AGI-3 Benchmark, Redefining AI Agents
Machine Learning

NVIDIA AVO Hits 100% on ARC-AGI-3 Benchmark, Redefining AI Agents

NVIDIA's AVO achieves 100% efficiency on ARC-AGI-3, showcasing a breakthrough in long-horizon autonomous agent systems.

NVIDIA Optimizes JAX LLM Training with Host Offloading
Machine Learning

NVIDIA Optimizes JAX LLM Training with Host Offloading

NVIDIA's host offloading for JAX LLM training boosts GPU memory efficiency, enabling larger batch sizes and faster throughput.

Together AI Showcases Nine Papers at ICML 2026 in Seoul
Machine Learning

Together AI Showcases Nine Papers at ICML 2026 in Seoul

Together AI presents nine groundbreaking papers at ICML 2026, covering AI agents, model efficiency, and GPU optimization. Booth B714, July 6-11.

Ray Data 2.56 Enhances AI Pipelines with Zero OOM Errors
Machine Learning

Ray Data 2.56 Enhances AI Pipelines with Zero OOM Errors

Ray Data 2.56 eliminates OOM errors, reduces memory pressure, and improves AI training/inference speeds by over 50%.

MiniMax M3 Debuts on NVIDIA: 1M Token Context, Multimodal AI
Machine Learning

MiniMax M3 Debuts on NVIDIA: 1M Token Context, Multimodal AI

MiniMax M3, the 428B-parameter model, launches on NVIDIA infrastructure, offering long-context reasoning and multimodal workflows for enterprise AI.

NVIDIA's Nemotron 3 Ultra Redefines AI for Long-Running Agents
Machine Learning

NVIDIA's Nemotron 3 Ultra Redefines AI for Long-Running Agents

NVIDIA's Nemotron 3 Ultra, a 550B-parameter AI model, offers faster, cost-efficient reasoning for complex workflows, enabling long-running agents to perform better.

Trending topics