Open Issues Need Help
View All on GitHub kvcached support matrix about 11 hours ago
help wanted
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
help wanted
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
enhancement help wanted
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
Error on gpt-oss with vLLM 9 months ago
good first issue
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
start mutiple models 9 months ago
good first issue
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
Failed to patch kv_cache_coordinator 10 months ago
good first issue
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
good first issue
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm
Question About kvcached Ability to Dynamically Recognize and Utilize Kubernetes Elastic Scaled GPU Memory Resources 11 months ago
good first issue
ovg-project/kvcached
1.1K
Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond
Python
#elastic-kvcache#gpu-mutiplexing#gpu-sharing#inference-engine#kvcache#kvcache-optimization#kvcached#llm#llm-framework#llm-inference#llm-serving#ollama#online-offline-coserve#serverless#sglang#vllm