logo
Casa Casos

DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories

Certificado
China Beijing Qianxing Jietong Technology Co., Ltd. Certificações
China Beijing Qianxing Jietong Technology Co., Ltd. Certificações
Revisões do cliente
A equipe de vendas da tecnologia Co. de Qianxing Jietong do Pequim, Ltd é muito profissional e paciente. Podem fornecer cotações rapidamente. A qualidade e o empacotamento dos produtos são igualmente muito bons. Nossa cooperação é muito lisa.

—— LLC do》 de Festfing DV do 《

Quando eu procurava o processador central de intel e o SSD de Toshiba urgentemente, Sandy da tecnologia Co. de Qianxing Jietong do Pequim, Ltd deu-me muita ajuda e obteve-me os produtos que eu precisei rapidamente. Eu aprecio-a realmente.

—— Kitty Yen

Sandy da tecnologia Co. de Qianxing Jietong do Pequim, Ltd é um vendedor muito cuidadoso, que possa me lembrar de erros da configuração a tempo quando eu compro um servidor. Os coordenadores são igualmente muito profissionais e podem rapidamente terminar o processo de teste.

—— Strelkin Mikhail Vladimirovich

Estamos muito satisfeitos com a nossa experiência de trabalho com a Beijing Qianxing Jietong. A qualidade do produto é excelente e a entrega é sempre pontual. A equipe de vendas é profissional, paciente e muito prestativa com todas as nossas perguntas. Agradecemos muito o seu apoio e esperamos uma parceria de longo prazo. Altamente recomendado!

—— Ahmad Navid

Qualidade: Ótima experiência com o meu fornecedor. O MikroTik RB3011 já estava usado, mas estava em muito bom estado e tudo funcionava perfeitamente.E todas as minhas preocupações foram resolvidas rapidamente.Fornecedor muito fiável, altamente recomendado.

—— Geran Colesio

Estou Chat Online Agora

DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories

July 17, 2026
At Paris’s RAISE Summit, DDN showcased its ongoing collaboration with Nebul, a European sovereign hybrid cloud provider, focused on boosting efficiency for large-scale AI inference deployments. Unveiled last week, the joint initiative unites Nebul’s inference platform, DDN’s Infinia data intelligence architecture, and NVIDIA accelerated computing to resolve a key production AI bottleneck: data movement costs and performance limitations during inference workloads.

mais recente caso da empresa sobre DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories  0

DDN frames the collaboration around critical production metrics: GPU utilization, token throughput, cost per token, and latency. While model training builds AI asset value, inference defines its operational and commercial returns. The rising adoption of agentic AI, retrieval-augmented generation (RAG), and high-concurrency inference means storage and data infrastructure directly impact accelerator efficiency and AI response speeds.

This active proof-of-concept project has yielded promising early results. The partners have recorded measurable improvements in time-to-first-token with KV cache enabled and completed validation for RoCE-based infrastructure. Ongoing benchmarking covers longer inference sequence lengths, unlocking further optimization potential for the Infinia platform. The collaboration also expands to joint NVIDIA efforts on benchmarking frameworks, scalability verification, and upcoming technical publications.

The integrated platform leverages distributed KV cache services, GPU-native data movement, intelligent data orchestration, and high-performance storage architecture. KV cache acceleration delivers notable inference gains by preserving and rapidly retrieving pre-computed attention states, cutting redundant calculations and eliminating data delivery delays that cause GPU idling.

mais recente caso da empresa sobre DDN and Nebul Validate KV Cache Acceleration for NVIDIA-Based AI Factories  1

Leaders from DDN, Nebul, and NVIDIA highlighted a major industry shift: AI infrastructure priorities are moving from raw GPU deployment to operational efficiency, maximizing returns from existing accelerator hardware. DDN CEO Alex Bouzari and Nebul CEO Arnold Juffer noted that past focus on larger model scales has given way to optimizing inference economics to make production AI commercially viable via lower per-token costs. NVIDIA Cloud Infrastructure VP Rod Evans added that large-scale agentic workloads now measure infrastructure success by GPU utilization and latency, rather than sheer compute power.

DDN emphasizes that AI infrastructure must evolve beyond basic storage functions to actively support AI execution workflows. The firm’s infrastructure platforms currently power over one million GPUs worldwide, serving hyperscalers, cloud providers, enterprises, governments, and research institutions.

Modern AI infrastructure teams now prioritize these core production metrics:
GPU utilization: Measures effective accelerator activity during inference, maximizing value of high-end GPU hardware.
Cost per token: Links infrastructure performance directly to AI model output operational costs.
Tokens per watt: Evaluates energy efficiency of AI inference output.
Time to first token: Determines interactive AI application responsiveness and user experience.
Time to production: Quantifies operational effort to migrate AI services from testing to scalable commercial deployment.

As inference becomes the dominant AI workload, delivering cached context and enterprise data to GPUs with low, stable latency will be critical to sustaining high GPU utilization and controlling long-term operational costs.

Beijing Qianxing Jietong Technology Co., Ltd.
Sandy Yang/Global Strategy Director
WhatsApp / WeChat: +86 13426366826
Email: yangyd@qianxingdata.com
Website: www.qianxingdata.com/www.storagesserver.com
Business Focus:
ICT Product Distribution/System Integration & Services/Infrastructure Solutions
With 20+ years of IT distribution experience, we partner with leading global brands to deliver reliable products and professional services.
“Using Technology to Build an Intelligent World”Your Trusted ICT Product Service Provider!

Contacto
Beijing Qianxing Jietong Technology Co., Ltd.

Pessoa de Contato: Ms. Sandy Yang

Telefone: 13426366826

Envie sua pergunta diretamente para nós (0 / 3000)