VAST Data Brings Confidential Compute to AI with DataEnclave

VAST Data DataEnclave

VAST Data announced its new VAST DataEnclave, a confidential compute capability for AI workloads within the VAST DataEngine, which runs AI models in hardware-isolated environments built on NVIDIA Confidential Computing. The new capability decrypts model weights and enterprise data only within protected CPU and GPU memory, and only after cryptographic attestation confirms that the environment meets policies set by each asset’s owner.

VMware AI Factory: Broadcom’s Push to Operationalize Private AI

VMware AI Factory

Broadcom introduced its new VMware AI Factory at VMware Explore 2026. The offering bundles software services built on VCF to take an enterprise from bare-metal infrastructure to a running production AI model. It combines infrastructure provisioning, GPU pooling, multi-tenant model sharing, and governance tooling into a single operational layer for enterprises that want to run AI workloads on infrastructure they own and control.

Qualcomm’s Dragonfly Data Center Portfolio

Qualcomm Dragonfly

Qualcomm unveiled Dragonfly, its full-stack data center portfolio focused on AI inference, at its recent 2026 Investor Day. The portfolio brings together the Dragonfly C1000 data center CPU, the Dragonfly AI300 inference accelerator, the company’s new High Bandwidth Compute (HBC) near-memory architecture, a broad connectivity lineup, and a custom silicon practice.

Computex 2026: AI Infrastructure, the PC Wars, and the New Connectivity Arms Race

Computex 2026

Computex has historically been a client- and consumer-focused event, but those markets now take a back seat to enterprise infrastructure and the components that define an AI factory. The key messages were all AI, all the time.

The keynotes were entirely semiconductor-first, with talks by three chip CEOs: Jensen Huang of NVIDIA, Cristiano Amon of Qualcomm, Lip-Bu Tan of Intel, and Matt Murphy of Marvell. Each painted a slightly different picture about where the AI buildout is headed.

HPE ProLiant: Edge Compute for AI and Mission-Critical Workloads

Edge AI

HPE recently expanded its HPE ProLiant edge portfolio with three new platforms: the HPE ProLiant Compute EL2000 chassis, two new Gen12 servers built for the EL2000 (the EL220 and EL240), and an enhanced version of the HPE ProLiant DL145 Gen11 server, now powered by AMD EPYC 8005 series processors. The announcement also introduced an Environmental Ruggedization Option Kit applicable across the portfolio.

Nutanix .NEXT 2026: Beyond the Hypervisor

Nutanix NEXT 2026

Nutanix held its annual .NEXT user conference earlier this month in Chicago drew more than 5,000 attendees and over 100 sponsors. The event came at an inflection point for enterprise infrastructure, as the post-VMware market continues to consolidate, AI workloads move from pilot to production, and hardware supply constraints continue to complicate infrastructure planning.

MLPerf Inference 6.0: Software Gains & Broadening Competition Shake Things Up

MLPerf 6.0 Inference

MLCommons released MLPerf Inference v6.0 results, marking what the consortium describes as the most significant update to the benchmark suite to date.

The round introduced five new workloads, including a multimodal vision-language model, a text-to-video generation benchmark, and a new interactive scenario for the DeepSeek-R1 reasoning model.

Research Note: Improving Inference with NVIDIA’s ‘CMX’ Inference Context Memory Storage Platform

NVIDIA Vera Rubin

At NVIDIA Live at CES 2026, NVIDIA introduced its Inference Context Memory Storage (ICMS) platform as part of its Rubin AI infrastructure architecture. NVIDIA’s ICMS addresses KV cache scaling challenges in LLM inference workloads.

The technology targets a specific gap in existing memory hierarchies where GPU high-bandwidth memory proves too limited for growing context requirements while general-purpose network storage introduces latency and power consumption penalties that degrade inference efficiency.

SC25: Beyond Super Computing

SC25

Supercomputing 2025 delivered a clear message to enterprise IT leaders: the infrastructure conversation has fundamentally changed. The announcements from SC25 were about architectural transformation.

From rack-scale designs to quantum integration to facility-level engineering, the building blocks of large-scale AI and HPC systems are being reimagined.

NVIDIA GTC 2025: The Super Bowl of AI

NVIDIA GTC 2025 Storage

If you thought AI was already moving fast, buckle up, Jensen Huang threw more fuel on the fire. NVIDIA’s GTC 2025 keynote wasn’t just about new GPUs; it was a full-scale vision of computing’s future, one where AI isn’t just a tool — it’s the foundation of everything.

Let’s look at what Jensen talk about during his 2+ hour keynote.

Research Note: Supermicro’s New Datacenter Scale Liquid Cooling

Supermicro Liquid Cooling

Supermicro recently announced a comprehensive, end-to-end liquid cooling solution for data centers. The solution encompasses critical hardware components such as Coolant Distribution Units (CDUs), cold plates, Coolant Distribution Manifolds (CDMs), cooling towers, and integrated management software.