Close Menu
Alpha Leaders
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
What's On
Why Enterprises Should Consider Open-Weight Models

Why Enterprises Should Consider Open-Weight Models

3 September 2026
AI visionary Ray Kurzweil joins neurotech startup Subsense as product and vision advisor

AI visionary Ray Kurzweil joins neurotech startup Subsense as product and vision advisor

3 September 2026
Before You Buy More Hardware, Check What Your Cluster Is Really Using

Before You Buy More Hardware, Check What Your Cluster Is Really Using

3 September 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Alpha Leaders
newsletter
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
Alpha Leaders
Home » Storage And Memory Enable Next Generation AI At The 2025 GTC
Innovation

Storage And Memory Enable Next Generation AI At The 2025 GTC

Press RoomBy Press Room19 March 20254 Mins Read
Facebook Twitter Copy Link Pinterest LinkedIn Tumblr Email WhatsApp
Storage And Memory Enable Next Generation AI At The 2025 GTC

Jensen Huang, CEO of Nvidia gave one of this announcement-filled presentations at the 2025 GTC in San Jose. Among announcements on GPU roadmaps, low power photonics networking, compact liquid cooled GPU systems, new robotics initiatives and the wide variety of CUDA libraries available to users, he also gave some interesting insights on digital storage and memory requirements for next generation GPUs and changes in AI storage platforms.

There is also a Nvidia driven effort underway, Storage-Next that aims to improve memory integration to GPUs versus CPUs. Several storage and memory companies were also at the GTC and we also discuss announcements from Micron, Phison, Vast Data and Vdura.

Jensen presented details on the announcement of the Blackwell GPU series, the follow on the Grace Hopper products, but he also talked about a couple of generations in the second half of 2026 and in 2027 that will be called the Vera Rubin GPU system. Vera Florence Cooper Rubin was an American astronomer who did pioneering research on galaxy rotation rates. I believe that he also indicated that the GPU generation after Vera Rubin would be named after physicist Richard Feynman.

The later 2027 Vera Rubin system, the Rubin Ultra NVL576 would have HBM4e data rates of 4.5PB/s with 365TB of fast memory and 1.5PBs NVLink7 connectivity as shown below.

The Rubin system package will also be considerably larger than prior generation Blackwell packages, as shown by comparing the size of components in the two packages below. Nvidia is taking 2D chiplet technologies to new levels with these GPUs.

Jensen, towards the end of his 2+ hour talk, announced a new class of storage infrastructure using AI to enhance data access, see image below. This Nvidia AI data platform is a reference design that digital storage partners will implement that uses AI query agents to generate insights from data in near real time using Nvidia AI enterprise software, including Nvidia NIM microservices for the company’s Nemotron models and its AI-Q Blueprint for building AI query agents.

Among the companies implementing versions of this reference design are DDN, Dell Technologies, Hewlett Packard Enterprise, Hitachi Vantara, IBM, NetApp, Nutanix, Pure Storage, VAST Data and WEKA.

Storage providers can implement these agents using Nvidia Blackwell GPUs, Bluefield DPUs, Spectrum-X networking and Dynamo open-source inference library. Using this collection of technologies, they can access large-scale data quickly and process various data types, including structured, semi-structured and unstructured data from multiple sources, including text, PDF, images and video.

There is also work going on, driven by Nvidia, for a new storage architecture for GPU computing near memory for disaggregated, data-protected, managed block storage. This would use next generation NVMe with the intention of achieving high IOPs/$, better power efficiency for fine-grained accesses by supporting direct access to storage from GPUs. This effort seeks to achieve 512B IOPs/GPU in Gen6, optimized for power and tail latency. Tail latency is the slowest response time in a system latency distribution, often measured as the outliers above 95% or higher of the distribution. Tail latency can have a significant impact on overall system performance.

GPU-initiated storage versus CPU-initiated storage is said to deliver better TCO with higher IOPS, smaller space with fewer devices needed to reach IOPS goals and lower power consumption with better utilization from reduced impact from tail latencies.

In addition to Nvidia announcements and activities, some storage and memory companies made announcement around the GTC. Micron announced SOCAMM, Small Outline Compression Attached Memory Module, developed in collaboration with Nvidia, a modular high-capacity LPDDR5X DRAM technology for use with the Nvidia GB300 Grace Blackwell Ultra Superchip. Micron also said that their HBM3E products were being used in several Nvidia platforms.

SOCAMMs are said to provide over 2.5 times higher bandwidth at the same capacity when compared to RDIMMs, allowing faster access to larger training datasets and more complex models, as well as increasing throughput for inference workloads.

Phison announced an array of expanded capabilities on aiDAPTIV+, a more affordable AI training and inference solution for on-premise infrastructure. aiDAPTIV+ is being integrated into AI laptop PCs as well as on edge computing devices running the Nvidia Jetson platform.

Vast Data announced new features to its Vast Data Platform enabling enhancements to its Vast InsightEngine including vector search and retrieval, serverless triggers and functions and fine-grained access control and AI-ready security.

VDURA launched its V5000 all-flash appliance offering a parallel file system architecture that it says is engineered for AI. It combines client-side erasure coding, remote direct memory access, RDMA, acceleration and flash-memory to scale with growing GPU clusters.

GPU memory and storage requirements are growing. New semantic storage with AI query agents and storage-next project improves data access. Micron, Phison, Vast Data and VDURA make AI storage and memory announcements at the 2025 GTC.

DDN Dell GPU GTC HPE IBM Micron Nvidia phison Vast
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link

Related Articles

Why Enterprises Should Consider Open-Weight Models

Why Enterprises Should Consider Open-Weight Models

3 September 2026
Before You Buy More Hardware, Check What Your Cluster Is Really Using

Before You Buy More Hardware, Check What Your Cluster Is Really Using

3 September 2026
The AI Productivity Race Has A False Finish Line

The AI Productivity Race Has A False Finish Line

3 September 2026
The Most Expensive Technology Decisions Are The Ones That Last Years

The Most Expensive Technology Decisions Are The Ones That Last Years

3 September 2026
Build What Better AI Can’t Commoditize

Build What Better AI Can’t Commoditize

3 September 2026
The Promise Of Precision Medicine Is Becoming Practical

The Promise Of Precision Medicine Is Becoming Practical

3 September 2026
Don't Miss
Exclusive: DeFi platform Azura launches after raising .9 million from Initialized

Exclusive: DeFi platform Azura launches after raising $6.9 million from Initialized

By Press Room22 October 2024

Azura, a new platform for decentralized finance, launched on Tuesday after raising $6.9 million in…

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

27 December 2024
Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

22 October 2024
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Latest Articles
The AI Productivity Race Has A False Finish Line

The AI Productivity Race Has A False Finish Line

3 September 20262 Views
AI wants electricity now. The electric grid needs years to catch up

AI wants electricity now. The electric grid needs years to catch up

3 September 20263 Views
The Most Expensive Technology Decisions Are The Ones That Last Years

The Most Expensive Technology Decisions Are The Ones That Last Years

3 September 20262 Views
Meloni breaks Berlusconi’s record for longest-serving uninterrupted government, longest since WWII

Meloni breaks Berlusconi’s record for longest-serving uninterrupted government, longest since WWII

3 September 20262 Views

Recent Posts

  • Why Enterprises Should Consider Open-Weight Models
  • AI visionary Ray Kurzweil joins neurotech startup Subsense as product and vision advisor
  • Before You Buy More Hardware, Check What Your Cluster Is Really Using
  • Canva’s productivity push gains traction—and in Southeast Asia, it’s happening on phones
  • The AI Productivity Race Has A False Finish Line

Recent Comments

No comments to show.
About Us
About Us

Alpha Leaders is your one-stop website for the latest Entrepreneurs and Leaders news and updates, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks
Why Enterprises Should Consider Open-Weight Models

Why Enterprises Should Consider Open-Weight Models

3 September 2026
AI visionary Ray Kurzweil joins neurotech startup Subsense as product and vision advisor

AI visionary Ray Kurzweil joins neurotech startup Subsense as product and vision advisor

3 September 2026
Before You Buy More Hardware, Check What Your Cluster Is Really Using

Before You Buy More Hardware, Check What Your Cluster Is Really Using

3 September 2026
Most Popular
Canva’s productivity push gains traction—and in Southeast Asia, it’s happening on phones

Canva’s productivity push gains traction—and in Southeast Asia, it’s happening on phones

3 September 20262 Views
The AI Productivity Race Has A False Finish Line

The AI Productivity Race Has A False Finish Line

3 September 20262 Views
AI wants electricity now. The electric grid needs years to catch up

AI wants electricity now. The electric grid needs years to catch up

3 September 20263 Views

Archives

  • September 2026
  • August 2026
  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • March 2022
  • January 2021
  • March 2020
  • January 2020

Categories

  • Blog
  • Business
  • Entrepreneurs
  • Global
  • Innovation
  • Leadership
  • Living
  • Money & Finance
  • News
  • Press Release
© 2026 Alpha Leaders. All Rights Reserved.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.