
NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories
NVIDIA introduces its DSX Ready program to standardize and qualify power and cooling hardware for AI data centers.

NVIDIA introduces its DSX Ready program to standardize and qualify power and cooling hardware for AI data centers.

US and Chinese officials began talks on an early-warning protocol for national security threats posed by AI.

Escalating data center debt and falling token prices threaten financial instability across hyperscalers and frontier AI labs.

Donald Trump faces growing pushback from MAGA conservatives and voters over his aggressive support for AI data center expansion.

Jina AI launched jina-ocr-v1, a 3.4B MoE document parser designed to execute fast, low-cost visual parsing using speculative decoding.

PrismML has released Ternary Bonsai 2 27B, a compressed 5.9 GB open-weights model that retains 98.2% of Qwen3.8 27B's baseline performance.

NVIDIA's Vera Rubin NVL72 benchmark debuts in MLPerf Inference v6.1, yielding up to 3.7x higher throughput than its predecessor.

NVIDIA and infrastructure partners demonstrated new power-management tools designed to optimize energy efficiency in large-scale AI factories.

Hyperscalers face immense profit pressure as AI infrastructure capital expenditure targets $1.1 trillion by 2027 against modest current revenues.

Cornelis raises $205M to challenge Nvidia's networking dominance with its open-architecture Active Compute Fabric.

ECB President Christine Lagarde warns Europe must build domestic AI infrastructure to avoid strategic leverage and economic disruption by the US or China.

Perplexity released its local Portable Computer agent on Windows for NVIDIA RTX PCs with 24GB or more VRAM.

Former EPA officials warn the Trump administration's environmental deregulation for AI data centers poses widespread public health risks.

NVIDIA detailed its full compute stack spanning training, simulation, and edge hardware powering commercial scale robotaxi operations.

Inference chipmaker d-Matrix has adopted NVIDIA NVLink Fusion to connect its Raptor XPUs directly into NVIDIA's rack-scale AI infrastructure.

AWS teamed up with Qualcomm to deploy specialized AI inference chips while Qualcomm adopts AWS Bedrock for automated semiconductor design.

NVIDIA releases CUDA Rust toolchains cuda-oxide and cutile-rs to enable compile-time-safe GPU kernel development.

OpenAI Chief Scientist Jakub Pachocki expects sustained compute scaling to drive recursive self-improvement and superhuman AI reasoning.

Perplexity published technical details on its custom GPU infrastructure stack that optimizes batch and real-time embedding serving.

The shift to real-time AI inference requires rearchitecting data center storage, memory bandwidth, and networking infrastructure.

Deepseek plans a 160,000-chip Huawei Ascend-950DT inference cluster in Inner Mongolia while remaining dependent on Nvidia for core training.

Microsoft unveiled Project Zenith, a developer-centric Windows setup paired with high-memory hardware to run 30B+ parameter AI models locally.

Nvidia released PAIR, an open-source tool that distributes local AI workloads across multiple devices on a home network.

Nvidia's RTX Spark superchip is powering a new wave of high-end AI laptops designed to run agentic workloads locally.

Perplexity open-sourced Lily, a custom Rust and Metal inference engine optimized for running Qwen3.6-35B-A3B natively on Apple Silicon.

Nvidia is officially rolling out DLSS 5, using generative AI filters to upscale graphics but requiring heavy hardware and multi-frame generation.

An opinion piece argues the EU could force a global AI slowdown by using Dutch firm ASML's lithography dominance as a trade choke-point.

China's CXMT has started small-scale production of HBM3E memory chips to help domestic AI chipmakers bypass US export controls.

An OpenAI researcher warns that ultra-fast inference speeds could allow misaligned AI systems to bypass human security controls.

Z.ai launches GLM-5.3-Flash, an open-weights model matching top benchmarks at a fraction of the cost using Chinese silicon.

OpenAI claims its custom Jalapeño ASIC outperforms NVIDIA's flagship chips on key AI inference efficiency benchmarks.

Apple released new Mac mini and Mac Studio desktops featuring M6 and M5 Ultra chips optimized for local AI workflows.

NVIDIA has put its Groq 3 LPX rack-scale inference system into full production to accelerate long-context agentic AI workloads.

NVIDIA claims its Vera Rubin NVL72 platform delivers up to 30x higher throughput per megawatt for demanding agentic AI workloads.