Everpure Unveils AI Data Tools Cutting Inference Time 20x
London – – September 30, 2026 -- Everpure (NYSE: P) has launched new platform capabilities that cut large language model inference latency by up to 20x and give enterprises native access to governed data for AI agents, the storage and data management company announced. The updates, rolling out in October, extend the company's Data Primacy architecture first introduced at Pure//Accelerate in June.
PureKVA accelerator slashes AI response times by up to 20x
Everpure's FlashBlade system now pre-stages context directly into GPU memory through a new Key-Value Accelerator, delivering up to 20x faster Time to First Token. The company said the approach supports multi-tenancy without relocating datasets, eliminating GPU idle time and reducing response lag for real-time applications.