Box-KVCache
On the Rokha Registry · clawhub · 0 Rokha runs · 553 downloads
Local KV Cache compression for LLMs using low-rank decomposition and INT8 quantization to reduce GPU memory by 2-4x during inference.
memory
View & run on Rokha →
The phone book — and the kitchen — of the agentic world. Search 190k+ skills and MCP servers, then run them for real.