---
description: DDN and Nebul validate KV cache acceleration on NVIDIA infrastructure, targeting cost per token and time to first token gains.
title: DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories
image: https://www.storagereview.com/wp-content/uploads/2026/07/Storagereview-ddn-pr-advance-ai-inference-web-logos.png
---

[![Storage Review]()](https://storagereview.com/)

---

[Facebook](https://www.facebook.com/@storagereview) [X (Twitter)](https://twitter.twitter.com/storagereview) [Instagram](https://www.instagram.com/storagereview/) [LinkedIn](https://www.linkedin.com/company/storagereview-com) [YouTube](https://www.youtube.com/user/storagereview?sub_confirmation=1) [Email](mailto:info@storagereview.com) [Spotify](https://open.spotify.com/show/1y6VnznABhHeOSMOmbDTz0) [Reddit](https://www.reddit.com/r/StorageReview/) [Discord](https://discord.gg/TwMHb4azdC) [RSS Feed](https://www.storagereview.com/rss.xml) [TikTok](https://www.tiktok.com/@storagereview)

[Facebook](https://www.facebook.com/@storagereview) [X (Twitter)](https://twitter.com/storagereview) [Instagram](https://www.instagram.com/storagereview/) [LinkedIn](https://www.linkedin.com/company/storagereview-com) [YouTube](https://www.youtube.com/user/storagereview?sub_confirmation=1) [Email](mailto:info@storagereview.com) [Spotify](https://open.spotify.com/show/1y6VnznABhHeOSMOmbDTz0) [Reddit](https://www.reddit.com/r/StorageReview/) [Discord](https://discord.gg/TwMHb4azdC) [RSS Feed](https://www.storagereview.com/rss.xml) [TikTok](https://www.tiktok.com/@storagereview)

[![StorageReview.com]()](https://www.storagereview.com)

≡ Menu

- [Home](https://www.storagereview.com/)
- [Storage Reviews](https://www.storagereview.com/review)
  - [Consumer Reviews](https://www.storagereview.com/consumer)
  - [Enterprise Reviews](https://www.storagereview.com/enterprise)
  - [Ubiquiti Reviews](https://www.storagereview.com/best/ubiquiti-reviews)
- [SR Merch](https://store.storagereview.com)
- [Leaderboards](https://www.storagereview.com/best)
  - [Best Storage Arrays](https://www.storagereview.com/best/storage-arrays)
  - [Best Enterprise SSDs](https://www.storagereview.com/best/enterprise-ssds)
  - [Best Servers](https://www.storagereview.com/best/servers)
  - [Best Desktops for Local AI](https://www.storagereview.com/best/desktops-local-ai)
  - [Best Laptops for Local AI](https://www.storagereview.com/best/laptops-local-ai)
  - [Best Local LLM Tools](https://www.storagereview.com/best/local-llm-tools)
  - [Agentic AI Hardware](https://www.storagereview.com/best/agentic-ai-hardware)
  - [Best SSDs & Hard Drives](https://www.storagereview.com/best_drives)
  - [Best Portable SSDs](https://www.storagereview.com/best/portable-ssds)
  - [Best Business Laptops](https://www.storagereview.com/best/business-laptops)
  - [Best Mobile Workstations](https://www.storagereview.com/best/mobile-workstations)
  - [Best Desktop Workstations](https://www.storagereview.com/best/desktop-workstations)
  - [Best Laptop Battery Life](https://www.storagereview.com/best/laptop-battery-life)
- [Storage Reference Guide](https://www.storagereview.com/storage-reference-guide)
- [About SR](https://www.storagereview.com/about-storagereview)
  - [StorageReview.com Sweepstakes Rules and Regulations](https://www.storagereview.com/storagereview-com-sweepstakes-rules-and-regulations)

Search

[Home](https://www.storagereview.com/) » [News](https://www.storagereview.com/news) » DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories

# DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories

by Harold Fritts on July 13, 2026

[AI](https://www.storagereview.com/enterprise/ai)  ◇  [Enterprise](https://www.storagereview.com/enterprise)

At the RAISE Summit in Paris, DDN highlighted its ongoing collaboration with Nebul, a European sovereign-hybrid cloud provider, focused on improving the efficiency of large-scale AI inference deployments. Announced last week, the effort combines Nebul’s inference platform, DDN’s Infinia data intelligence architecture, and NVIDIA accelerated computing to address a growing production AI constraint: the cost and performance implications of moving data during inference.

![DDN Nebul NVIDIA logos]()

DDN positions the work around key production metrics, including GPU utilization, token throughput, cost per token, and latency. The premise is that model training establishes the value of an AI asset, while inference determines its operational and commercial return. As agentic AI, retrieval-augmented generation, and high-concurrency inference become more common, storage and data infrastructure can directly affect accelerator utilization and response time.

![NVIDIA AI Factory graphic]()

 

The work is an ongoing proof-of-concept engagement, and DDN characterizes the results so far as promising early data. The companies report measurable improvements in time to first token with KV cache enabled and have completed RoCE-based infrastructure validation, with benchmarking continuing across larger inference sequence lengths as they identify further optimization opportunities within the Infinia platform. The project has also expanded to include collaboration with NVIDIA on benchmarking methodologies, scalability validation, and future technical publications.

The platform uses distributed KV cache services, GPU-native data movement, data orchestration, and high-performance storage architectures. KV cache acceleration is particularly relevant for inference because it preserves and rapidly retrieves previously computed attention state, reducing repeated computation and limiting data delivery delays that can leave GPUs idle.

![DDN GPU Optimization]()

Leadership from DDN, Nebul, and NVIDIA signaled a fundamental industry shift from GPU acquisition to operational efficiency, with a focus on maximizing value from deployed accelerators. DDN CEO Alex Bouzari and Nebul CEO Arnold Juffer both emphasized that while model scale was the historical priority, the current challenge lies in inference economics and making production AI commercially viable through lower token costs. Rod Evans, Vice President of Cloud Infrastructure at NVIDIA, added that as organizations move toward large-scale agentic workloads, infrastructure success is increasingly measured by GPU utilization and latency rather than raw compute capacity.

DDN argues that AI infrastructure must evolve beyond conventional storage operations and become an active participant in AI execution. The company says its platforms support more than one million GPUs globally, spanning hyperscalers, cloud builders, enterprises, governments, and research institutions.

For infrastructure teams, the relevant production metrics increasingly include:

- **GPU utilization**, measuring how effectively expensive accelerators remain active during inference.
- **Cost per token**, connecting infrastructure efficiency to the cost of model output.
- **Tokens per watt**, measuring output efficiency against power consumption.
- **Time to first token**, which affects perceived responsiveness for interactive AI applications.
- **Time to production reflects the operational effort required to move AI services from testing to** scalable deployment.

As inference becomes the dominant AI operating workload, the ability to deliver cached context and enterprise data to GPUs with low, predictable latency will be increasingly important to sustaining utilization and controlling costs.

**Engage with StorageReview**

[Newsletter](https://www.storagereview.com/storage_newsletter) |  [YouTube](https://www.youtube.com/user/StorageReview "Opens in a new window") | Podcast  [iTunes](https://podcasts.apple.com/gb/podcast/storagereview-com-storage-reviews/id1060681115 "Opens in a new window")/ [Spotify](https://open.spotify.com/show/1y6VnznABhHeOSMOmbDTz0 "Opens in a new window") |  [Instagram](https://www.instagram.com/storagereview/ "Opens in a new window") |  [Twitter](https://twitter.com/storagereview "Opens in a new window") |  [TikTok](https://www.tiktok.com/@storagereview? "Opens in a new window") |  [RSS Feed](https://www.storagereview.com/rss.xml)

![]()

### Harold Fritts

I have been in the tech industry since IBM created Selectric. My background, though, is writing. So I decided to get out of the pre-sales biz and return to my roots, doing a bit of writing but still being involved in technology.

Previous post: [DeepInfra Opens 1.7MW Toronto Data Center With 1,000+ NVIDIA B300 GPUs](https://www.storagereview.com/news/deepinfra-opens-1-7mw-toronto-data-center-with-1000-nvidia-b300-gpus)

Next post: [WhiteFiber’s Project Redwood Links Two H200 Clusters Into One 111.2 Tbps Supercluster](https://www.storagereview.com/news/whitefibers-project-redwood-links-two-h200-clusters-into-one-111-2-tbps-supercluster)

Trusted Vendors

Products and solutions from our affiliate partners:

- [  
  ![Ubiquiti Logo]()  
   Ubiquiti](https://store.ui.com/us/en?a_aid=StorageReview "Opens in a new window")
- [  
  ![Newegg Logo]()  
   Newegg](https://click.linksynergy.com/fs-bin/click?id=g5terNMDhz0&offerid=1207190.18&subid=0&type=4 "Opens in a new window")
- [  
  ![Amazon Logo]() Amazon](https://amzn.to/4fLJUEW "Opens in a new window")

Newsletter

Subscribe to the StorageReview newsletter to stay up to date on the latest news and reviews. We promise no spam!

1  

Leave this field empty if you’re human:

## Advertisement

Content Categories

[Facebook](https://www.facebook.com/@storagereview) [X](https://twitter.com/storagereview) [Instagram](https://www.instagram.com/storagereview/) [LinkedIn](https://www.linkedin.com/company/storagereview-com) [YouTube](https://www.youtube.com/user/storagereview?sub_confirmation=1) [Email](mailto:info@storagereview.com) [Spotify](https://open.spotify.com/show/1y6VnznABhHeOSMOmbDTz0) [Reddit](https://www.reddit.com/r/StorageReview/) [Discord](https://discord.gg/TwMHb4azdC) [RSS](https://www.storagereview.com/rss.xml) [TikTok](https://www.tiktok.com/@storagereview)

Copyright © 1998-2025 Flying Pig Ventures, LLC Cincinnati, Ohio. All rights reserved.

Manage your privacy

To provide the best experiences, we and our partners use technologies like cookies to store and/or access device information. Consenting to these technologies will allow us and our partners to process personal data such as browsing behavior or unique IDs on this site and show (non-) personalized ads. Not consenting or withdrawing consent, may adversely affect certain features and functions.

Click below to consent to the above or make granular choices. Your choices will be applied to this site only. You can change your settings at any time, including withdrawing your consent, by using the toggles on the Cookie Policy, or by clicking on the manage consent button at the bottom of the screen.

Functional Functional Always active

The technical storage or access is strictly necessary for the legitimate purpose of enabling the use of a specific service explicitly requested by the subscriber or user, or for the sole purpose of carrying out the transmission of a communication over an electronic communications network.

Preferences Preferences

The technical storage or access is necessary for the legitimate purpose of storing preferences that are not requested by the subscriber or user.

Statistics Statistics

The technical storage or access that is used exclusively for statistical purposes. The technical storage or access that is used exclusively for anonymous statistical purposes. Without a subpoena, voluntary compliance on the part of your Internet Service Provider, or additional records from a third party, information stored or retrieved for this purpose alone cannot usually be used to identify you.

Marketing Marketing

The technical storage or access is required to create user profiles to send advertising, or to track the user on a website or across several websites for similar marketing purposes.

Statistics

Marketing

Features

Always active

Always active

- Manage options
- Manage services
- Manage {vendor\_count} vendors
- [Read more about these purposes](https://cookiedatabase.org/tcf/purposes/)

Accept Deny Manage options Save preferences Manage options

- {title}
- {title}
- {title}

Manage your privacy

To provide the best experiences, we use technologies like cookies to store and/or access device information. Consenting to these technologies will allow us to process data such as browsing behavior or unique IDs on this site. Not consenting or withdrawing consent, may adversely affect certain features and functions.

Functional Functional Always active

The technical storage or access is strictly necessary for the legitimate purpose of enabling the use of a specific service explicitly requested by the subscriber or user, or for the sole purpose of carrying out the transmission of a communication over an electronic communications network.

Preferences Preferences

The technical storage or access is necessary for the legitimate purpose of storing preferences that are not requested by the subscriber or user.

Statistics Statistics

The technical storage or access that is used exclusively for statistical purposes. The technical storage or access that is used exclusively for anonymous statistical purposes. Without a subpoena, voluntary compliance on the part of your Internet Service Provider, or additional records from a third party, information stored or retrieved for this purpose alone cannot usually be used to identify you.

Marketing Marketing

The technical storage or access is required to create user profiles to send advertising, or to track the user on a website or across several websites for similar marketing purposes.

Statistics

Marketing

Features

Always active

Always active

- Manage options
- Manage services
- Manage {vendor\_count} vendors
- [Read more about these purposes](https://cookiedatabase.org/tcf/purposes/)

Accept Deny Manage options Save preferences Manage options

- {title}
- {title}
- {title}

Manage consent Manage consent

```json
{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories","url":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories","name":"DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories - StorageReview.com","isPartOf":{"@id":"https:\/\/www.storagereview.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories#primaryimage"},"image":{"@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories#primaryimage"},"thumbnailUrl":"https:\/\/www.storagereview.com\/wp-content\/uploads\/2026\/07\/Storagereview-ddn-pr-advance-ai-inference-web-logos.png","datePublished":"2026-07-13T16:58:40+00:00","description":"DDN and Nebul validate KV cache acceleration on NVIDIA infrastructure, targeting cost per token and time to first token gains.","breadcrumb":{"@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories#primaryimage","url":"https:\/\/www.storagereview.com\/wp-content\/uploads\/2026\/07\/Storagereview-ddn-pr-advance-ai-inference-web-logos.png","contentUrl":"https:\/\/www.storagereview.com\/wp-content\/uploads\/2026\/07\/Storagereview-ddn-pr-advance-ai-inference-web-logos.png","width":1200,"height":628},{"@type":"BreadcrumbList","@id":"https:\/\/www.storagereview.com\/news\/ddn-and-nebul-validate-kv-cache-acceleration-for-nvidia-based-ai-factories#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.storagereview.com\/"},{"@type":"ListItem","position":2,"name":"News","item":"https:\/\/www.storagereview.com\/news"},{"@type":"ListItem","position":3,"name":"DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories"}]},{"@type":"WebSite","@id":"https:\/\/www.storagereview.com\/#website","url":"https:\/\/www.storagereview.com\/","name":"StorageReview.com","description":"StorageReview.com is a leading provider of news and reviews throughout the entire IT stack - from the datacenter to the edge, and all points in between.","publisher":{"@id":"https:\/\/www.storagereview.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.storagereview.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/www.storagereview.com\/#organization","name":"StorageReview.com","url":"https:\/\/www.storagereview.com\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.storagereview.com\/#\/schema\/logo\/image\/","url":"https:\/\/www.storagereview.com\/wp-content\/uploads\/2020\/02\/Storage-Reviews-2-2.png","contentUrl":"https:\/\/www.storagereview.com\/wp-content\/uploads\/2020\/02\/Storage-Reviews-2-2.png","width":344,"height":61,"caption":"StorageReview.com"},"image":{"@id":"https:\/\/www.storagereview.com\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/x.com\/storagereview","http:\/\/youtube.com\/user\/storagereview"]}]}
```
