{"id":32663,"date":"2026-07-08T14:00:00","date_gmt":"2026-07-08T12:00:00","guid":{"rendered":"http:\/\/stocks-future.com\/?guid=0b97a8fb5b8610903df830dfddef4346"},"modified":"2026-07-08T14:00:00","modified_gmt":"2026-07-08T12:00:00","slug":"ddn-nebul-and-nvidia-collaborate-to-advance-ai-inference-economics-through-high-performance-kv-cache-acceleration-3","status":"publish","type":"post","link":"https:\/\/stocks-future.com\/?p=32663","title":{"rendered":"DDN, Nebul, and NVIDIA Collaborate to Advance AI Inference Economics Through High-Performance KV Cache Acceleration"},"content":{"rendered":"<p class=\"bwalignc\">\n<i>Collaboration Demonstrates How Data Infrastructure Is Becoming the Key Lever for Lower<\/i><b> <\/b><i>Cost-Per-Token, Faster Time-to-First-Token, and Higher AI Factory Efficiency<\/i><\/p><p>PARIS--(BUSINESS WIRE)--<a href=\"https:\/\/twitter.com\/hashtag\/AI?src=hash\" >#AI<\/a>--<a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=http%3A%2F%2Fwww.ddn.com%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=DDN&amp;index=1&amp;md5=d27745f3fe8d85bcd0d2f77639767e74\" rel=\"nofollow\" shape=\"rect\">DDN<\/a>, the global leader in AI and data intelligence solutions, today announced continued progress in its collaboration with <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fnebul.com%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=Nebul&amp;index=2&amp;md5=7437120a4d9292f493c1a91104ef9bc1\" rel=\"nofollow\" shape=\"rect\">Nebul<\/a>, a European leader in providing sovereign-hybrid cloud solutions, to optimize large-scale AI inference performance through advanced KV Cache acceleration and high-performance data infrastructure.<\/p><br\/><a href=\"https:\/\/mms.businesswire.com\/media\/20260708206542\/en\/2846483\/5\/DDN_LOGO-RGB.jpg\"><img src=\"https:\/\/mms.businesswire.com\/media\/20260708206542\/en\/2846483\/22\/DDN_LOGO-RGB.jpg\" \/><\/a><br\/><a href=\"https:\/\/mms.businesswire.com\/media\/20260708206542\/en\/2846483\/5\/DDN_LOGO-RGB.jpg\"><img src=\"https:\/\/mms.businesswire.com\/media\/20260708206542\/en\/2846483\/21\/DDN_LOGO-RGB.jpg\" \/><\/a><p>\nThe collaboration brings together Nebul's AI inference platform, DDN's Infinia data intelligence architecture, and NVIDIA accelerated computing technologies to address one of the most important challenges facing production AI environments: maximizing the economic return on AI infrastructure investments through higher GPU utilization, faster token generation, and lower cost-per-token.<\/p><p>\nAs AI adoption moves from experimentation to production, organizations are confronting a new reality: the challenge is no longer whether AI works, but whether the economics work at scale.<\/p><p>\nAcross the industry, enterprises, cloud providers, and sovereign AI initiatives are increasingly measuring success through business outcomes such as GPU utilization, cost-per-token, tokens-per-watt, and time-to-production. In this environment, data infrastructure has emerged as a critical determinant of AI profitability.<\/p><p>\n\"The AI conversation has fundamentally changed,\" said Alex Bouzari, CEO and Co-Founder at DDN. \"For years, the industry focused on acquiring GPUs. Today, the question is how efficiently those GPUs generate value. Inference has become the economic engine of AI, and reducing the cost of every token produced is now one of the most important challenges facing the industry.\"<\/p><p>\nAs part of an ongoing proof-of-concept engagement, DDN and Nebul are validating next-generation KV Cache acceleration capabilities designed to support <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.nvidia.com%2Fen-us%2Fdata-center%2Fproducts%2Fdsx%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=NVIDIA+DSX-based+AI+factory+deployments&amp;index=3&amp;md5=f1187b4350017ebce7450f23d0b942bc\" rel=\"nofollow\" shape=\"rect\">NVIDIA DSX-based AI factory deployments<\/a> by improving inference efficiency, reducing latency, and increasing utilization of AI infrastructure.<\/p><p>\n\"For years, the industry focused on building larger models. Today, the challenge is making those models economically viable in production,\" said Arnold Juffer, CEO at Nebul. \"Every organization is looking for ways to generate more value from its AI infrastructure investments. Through our collaboration with DDN and NVIDIA, we are demonstrating how KV Cache optimization and high-performance data architectures can improve inference efficiency, reduce latency, and help unlock the next phase of AI adoption.\"<\/p><p>\n\"AI infrastructure is increasingly defined by efficiency at scale,\" said Rod Evans, Vice President of Cloud Infrastructure at NVIDIA. \"As organizations deploy larger models and agentic AI workloads into production, technologies that improve GPU utilization, reduce latency, and accelerate token generation become critical. DDN continues to be an important collaborator in advancing the data and infrastructure capabilities needed to support the next generation of AI factories.\"<\/p><p>\nRecent benchmarking efforts have demonstrated promising early results, including measurable improvements in Time-to-First-Token (TTFT) performance with KV Cache enabled. The collaboration has successfully completed RoCE-based infrastructure validation and continues to expand benchmarking activities across larger inference sequence lengths while identifying additional optimization opportunities within <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.ddn.com%2Fproducts%2Finfinia%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=DDN%27s+Infinia+platform&amp;index=4&amp;md5=f29f6f9210af035df640277616145e78\" rel=\"nofollow\" shape=\"rect\">DDN's Infinia platform<\/a>.<\/p><p>\nThe project has also expanded to include collaboration with NVIDIA around benchmarking methodologies, scalability validation, and future technical publications.<\/p><p>\n<b>The Economics of AI Have Shifted<\/b><\/p><p>\nIndustry attention has historically centered on model performance and GPU availability. However, as organizations deploy AI into production, a new bottleneck has emerged: data movement and inference efficiency.<\/p><p>\nAgentic AI workloads, retrieval-augmented generation (RAG), and large-scale inference environments place unprecedented demands on infrastructure, including networking and storage. Every millisecond of latency and every percentage point of GPU idle time directly impacts profitability.<\/p><p>\n\"Training builds the asset. Inference is where it earns,\" said Bouzari. \"The next generation of AI applications will be powered by agents that perform meaningful business transactions and decisions. The economics of those systems depend on infrastructure that can deliver data at the speed AI operates.\"<\/p><p>\nDDN's Infinia platform was purpose-built to address these challenges through distributed KV Cache services, GPU-native data movement, intelligent data orchestration, and high-performance storage architectures that maximize accelerator efficiency.<\/p><p>\n<b>Accelerating the AI Factory Era<\/b><\/p><p>\nThe Nebul collaboration reinforces DDN's broader vision that AI infrastructure must evolve beyond traditional storage architectures and become an active participant in AI execution.<\/p><p>\nThe collaboration supports the broader <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.nvidia.com%2Fen-us%2Fdata-center%2Fproducts%2Fdsx%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=NVIDIA+DSX+platform&amp;index=5&amp;md5=29cb938e6ac5b5be7b3ec2000c05cb08\" rel=\"nofollow\" shape=\"rect\">NVIDIA DSX platform<\/a> approach to AI factories, where compute, networking, storage, software, and operations are designed together to improve tokens-per-watt, cost-per-token, and time-to-production.<\/p><p>\nAs organizations seek to operationalize AI at scale, DDN believes the defining metrics of success will increasingly become:<\/p><ul class=\"bwlistdisc\">\n<li>\nGPU utilization<\/li>\n<li>\nCost-per-token<\/li>\n<li>\nTokens-per-watt<\/li>\n<li>\nTime-to-first-token<\/li>\n<li>\nTime-to-production<\/li>\n<\/ul><p>\nOrganizations that optimize these metrics will achieve sustainable AI economics. Those who do not risk deploying infrastructure that remains underutilized despite significant investment.<\/p><p>\n\"We don't sell storage,\" Bouzari added. \"We help organizations maximize the economic return on every GPU, every token, and every watt. That's the foundation of the AI economy.\"<\/p><p>\nFor more information about AI inference economics, visit DDN at RAISE or visit <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.ddn.com%2Flp%2Fevents%2Fraise-2026%2Fbook-a-meeting%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=https%3A%2F%2Fwww.ddn.com%2Flp%2Fevents%2Fraise-2026%2Fbook-a-meeting%2F&amp;index=6&amp;md5=be1988ee6dd95411c6ff23023981c55d\" rel=\"nofollow\" shape=\"rect\">https:\/\/www.ddn.com\/lp\/events\/raise-2026\/book-a-meeting\/<\/a>.<\/p><p>\n<b>About DDN<\/b><\/p><p>\nDDN is the world\u2019s leading AI and data intelligence company, powering the world\u2019s most demanding AI workloads by keeping GPUs fed, efficient, and productive\u2014at massive scale\u2014so organizations can train, checkpoint, and infer faster with less footprint and power while achieving tremendous ROI from their AI investments. From hyperscalers and next-gen cloud builders to enterprises, governments, and research institutions, DDN delivers proven data intelligence at exabyte scale across millions of GPUs\u2014so customers can deploy AI with confidence, accelerate time-to-value, and realize outsized returns. Discover more at <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.ddn.com&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=ddn.com&amp;index=7&amp;md5=ccd6553881121ce2761b3d5299871f1a\" rel=\"nofollow\" shape=\"rect\">ddn.com<\/a>.<\/p><p>\n<b>Follow DDN: <\/b><a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.linkedin.com%2Fcompany%2Fddn%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=LinkedIn&amp;index=8&amp;md5=757b2b61826b74ad0cd1144c8bf7bdca\" rel=\"nofollow\" shape=\"rect\">LinkedIn<\/a>,<a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fx.com%2Fddnintelligence&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=X&amp;index=9&amp;md5=0febfd4ba00492c5d5639f57184af439\" rel=\"nofollow\" shape=\"rect\"> X<\/a>, and <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.youtube.com%2F%40DDNintelligence&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=YouTube&amp;index=10&amp;md5=9e1efd9c12888f6912adaf0ebc917403\" rel=\"nofollow\" shape=\"rect\">YouTube<\/a><\/p><p>\n<b>About Nebul<\/b><\/p><p>\nNebul offers European values on privacy and sovereignty combined with the technology, convenience and scale of hyper scale into a genuine European Sovereign-AI Cloud. Using the full range of NVIDIA technologies, new capabilities like Agentic AI and AI Coding can be enabled more easily and quickly without sharing sensitive corporate data with third parties you do not control. Nebul AI Cloud provides access to the latest AI technologies to empower European organizations to harness the power of AI. Learn more about Nebul at <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=http%3A%2F%2Fwww.nebul.com&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=nebul.com&amp;index=11&amp;md5=090fef2f324d5a54e97ed415a1f1048f\" rel=\"nofollow\" shape=\"rect\">nebul.com<\/a>.<\/p><p>\nFollow Nebul: <a  href=\"https:\/\/cts.businesswire.com\/ct\/CT?id=smartlink&amp;url=https%3A%2F%2Fwww.linkedin.com%2Fcompany%2Fnebul%2F&amp;esheet=54566196&amp;newsitemid=20260708206542&amp;lan=en-US&amp;anchor=LinkedIn&amp;index=12&amp;md5=6f55c15f8a337753f8f5d6b9c1e43c3b\" rel=\"nofollow\" shape=\"rect\">LinkedIn<\/a><\/p><br\/> <b>Contacts<\/b> <br\/><p>\n<b>DDN Media Contact:<\/b><br\/>Amanda Lee, VP, Marketing\u2014Analyst &amp; Public Relations\n<br\/><a  href=\"mailto:amlee@ddn.com\" rel=\"nofollow\" shape=\"rect\">amlee@ddn.com<\/a><\/p>","protected":false},"excerpt":{"rendered":"<p>Collaboration Demonstrates How Data Infrastructure Is Becoming the Key Lever for Lower Cost-Per-Token, Faster Time-to-First-Token, and Higher AI Factory EfficiencyPARIS&#8211;(BUSINESS WIRE)&#8211;#AI&#8211;DDN, the global leader in AI and data intelligence solution&#8230;<\/p>\n","protected":false},"author":2,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[1],"tags":[],"class_list":["post-32663","post","type-post","status-publish","format-standard","hentry","category-infos-businesswire"],"_links":{"self":[{"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/posts\/32663","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/stocks-future.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=32663"}],"version-history":[{"count":1,"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/posts\/32663\/revisions"}],"predecessor-version":[{"id":32668,"href":"https:\/\/stocks-future.com\/index.php?rest_route=\/wp\/v2\/posts\/32663\/revisions\/32668"}],"wp:attachment":[{"href":"https:\/\/stocks-future.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=32663"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/stocks-future.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=32663"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/stocks-future.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=32663"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}