What's inside Tensormesh Platform
A modular platform. A production-ready release for KV cache management and prompt caching, a shared platform core for operations and multi-tenancy, and infrastructure compatibility across every stack we've seen.
The release delivers KV cache control, prefix + non-prefix prompt caching, prefill/decode disaggregation, and P2P KV cache sharing — accessible through the operator UI and CLI. Underneath, the shared platform core provides the K8S operator, security, observability, and multi-tenancy. Everything runs on any Kubernetes cluster, any accelerator, any storage, and any inference engine.