AI
AMD Quark AI Agent Streamlines Model Quantization Workflows
AMD's Quark AI Agent Skills simplify AI model optimization for PyTorch and ONNX, reducing developer friction and enhancing deployment efficiency.
NVIDIA Groq 3 LPX Achieves 3,431 TPS Benchmark on Vera Rubin
NVIDIA's Groq 3 LPX sets a new standard in AI inference, delivering 3,431 tokens/second on a 100K context benchmark and redefining high-interactivity workloads.
NVIDIA Vera CPU Targets Agentic AI Efficiency
NVIDIA unveils Vera CPU, optimized for agentic AI workloads, promising 1.5x better performance in AI factory operations.
SpaceXAI Chooses NVIDIA Vera CPU for Advanced AI Expansion
SpaceXAI adopts NVIDIA Vera CPUs to scale agentic AI applications, including groundbreaking orbital AI deployments.
NVIDIA Groq 3 LPX Hits Full Production, Boosts AI Inference Speed
NVIDIA's Groq 3 LPX enters full production, delivering record-breaking AI inference speeds for latency-critical workloads. Key to agentic AI development.
NVIDIA's Spectrum-X Ethernet Redefines AI Networking
NVIDIA unveils Spectrum-X Ethernet, optimizing AI training and inference by addressing Ethernet’s limitations for GPU clusters. Here’s what it means.
NVIDIA Vera Rubin NVL72 Delivers 30x Efficiency Boost for AI
NVIDIA's Vera Rubin NVL72 offers 30x higher throughput per watt, setting a new standard for agentic AI workloads with lower token costs.
Harvey Launches Tenet, Open-Weight Legal AI Model
Harvey debuts Tenet, its first post-trained open-weight model, promising improved legal AI performance and cost-efficiency for law firms.
NVIDIA AVO Hits 100% on ARC-AGI-3 Benchmark, Redefining AI Agents
NVIDIA's AVO achieves 100% efficiency on ARC-AGI-3, showcasing a breakthrough in long-horizon autonomous agent systems.
NVIDIA Highlights Critical Role of Security in AI Agent Stacks
NVIDIA outlines where security fits in the AI agent stack, emphasizing controls for long-horizon agents as enterprises adopt agentic systems.