Nvidia
Latest news, analysis, and insights about Nvidia.
NVIDIA, Google, and Microsoft Unite on 800 VDC Power Architecture for AI Factories
As frontier AI models scale, the bottleneck is shifting from grid capacity to on-site distribution efficiency. A new open standard backed by NVIDIA, Google, and Microsoft aims to change how we power silicon.
How NVIDIA Nemotron 3.5 Lightning Solves the Complex Multi-Agent Routing Problem
NVIDIA's latest releases tackle the unsexy but critical plumbing of the agentic transition. Learn how a 30-billion-parameter MoE model and an open-source router aim to decentralize enterprise AI.
How Nvidia is shifting the crushing cost of AI data centers onto Wall Street
Nvidia is reportedly coordinating a historic $500 billion funding package with Wall Street. The move shifts the financial burden of next-generation AI data centers from Big Tech balance sheets to institutional investors.
How NVIDIA is transforming AI storage from passive repositories into active silicon paths
As autonomous AI agents scale, traditional data architectures are buckling. NVIDIA is countering this by turning passive storage into an active, silicon-accelerated compute path with its new Vera CPU.
Why NVIDIA's B300 Blackwell Ultra is the Definitive Hardware for Agentic Reasoning
NVIDIA's latest Blackwell Ultra architecture moves past raw training throughput to tackle the massive memory bottlenecks of reasoning-class AI models. Here is how the B300 changes the game for inference scale.
How Nvidia’s venture engine uses circular financing to artificially juice chip demand
A blockbuster report reveals Nvidia facilitated $750 billion in deals, reigniting fears of a financial feedback loop. Are these investments building an ecosystem or hiding an AI bubble?
How NVIDIA Nemotron 3 Embed Dominates RTEB to Redefine Agentic Retrieval
NVIDIA's new Nemotron 3 Embed has claimed the top rank on the RTEB benchmark. Here is how this new model optimizes agentic workflows and what it means for developers choosing embedding models over existing standards.
Nvidia Releases Open-Weight Model With Learned Memory Compression That Cuts Context Costs 8x
Nvidia just dropped an open-weight 8B model with a technique that compresses key-value cache by 8x. For anyone running inference at scale—or trying to squeeze longer contexts onto consumer GPUs—this matters.
Nvidia's New China Rule: Pay First for H200 AI Chips Amid Export Uncertainty
Nvidia is demanding full upfront payment from Chinese customers for its H200 AI chips—a striking signal of how export controls are reshaping the AI hardware market. The move comes as regulatory approval from both Washington and Beijing remains in limbo.
NVIDIA Unveils Rubin Platform at CES — A Six-Chip Architecture Built for Agentic AI
NVIDIA just revealed its post-Blackwell roadmap at CES. The Rubin platform combines six purpose-built chips — including the new Vera CPU and Rubin GPU — explicitly designed for agentic AI and mixture-of-experts models. This is NVIDIA betting on where AI is headed, not where it's been.
NVIDIA's 72GB Desktop GPU Makes Running Large Language Models Locally Practical
NVIDIA just shipped a desktop GPU with 72GB of VRAM. For AI developers tired of cloud latency and API costs, this Blackwell-based workstation card finally makes running large language models locally realistic.
NVIDIA Buys Slurm Creator SchedMD, Promising Open-Source Continuity for HPC Workloads
NVIDIA has acquired SchedMD, the company behind Slurm—the workload management system running on more than half of the world's top 100 supercomputers. The move extends NVIDIA's reach deeper into the AI training stack while raising questions about infrastructure consolidation.