Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

1 stars 0 forks 1 watchers Python Apache License 2.0
attention cache-compression kv-cache llm-inference pytorch transformers
9 Open Issues Need Help Last updated: Sep 6, 2026

Open Issues Need Help

View All on GitHub
help wanted good first issue area: transformers

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
help wanted good first issue area: evaluation

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
enhancement help wanted area: transformers

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
help wanted launch blocker needs GPU area: evaluation area: vllm

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
help wanted launch blocker area: vllm

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
help wanted launch blocker area: evaluation

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
documentation help wanted good first issue

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
enhancement help wanted good first issue

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers
enhancement help wanted good first issue

Deterministic, prompt-protected random KV-cache eviction for Hugging Face Transformers.

Python
#attention#cache-compression#kv-cache#llm-inference#pytorch#transformers