Scality AI Inference Factory Serves KV Cache From Object Storage Over RDMA, 14x Faster Than Recompute
Scality launched Scality AI Inference Factory, an open-code software stack that serves the KV cache for open-weight model inference from its AI Data Infrastructure object storage over RDMA, which Scal…