GPUCache
rustfs/GPUCache
A PB-scale, ultra-low latency distributed GPU cache for AI inference. Built with Rust, NVIDIA DOCA, RDMA, and BF-4 DPUs to bridge GPU HBM and NVMe storage, eliminating the recompute tax for large language models.
14stars
Forks
0
Open issues
0
Watchers
14
Size
0.0 MB
Apache License 2.0
ai-memorydpugpugpu-cachekvkv-cachenvidiardmastx
Created: May 26, 2026
Updated: Sep 23, 2026
Last push: Sep 23, 2026