Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
FluffyAIcode
/
LLM-KA-Cache-Compress
like
1
Sleeping
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
LLM-KA-Cache-Compress
14.7 kB
Ctrl+K
Ctrl+K
1 contributor
History:
14 commits
cryptobiosis
Add 'When to pick KakeyaLattice over HQQ / Quanto / KIVI' comparison block
b044c1d
verified
4 months ago
.gitattributes
Safe
1.52 kB
initial commit
4 months ago
Dockerfile
Safe
1.42 kB
Fix compression-ratio display + switch to repetition-safe default prompt; align with Docker SDK manifest (Dockerfile + requirements.txt + README.md)
4 months ago
README.md
Safe
3.96 kB
Add 'When to pick KakeyaLattice over HQQ / Quanto / KIVI' comparison block
4 months ago
app.py
Safe
7.34 kB
Rewrite Space subtitle: non-Gaussian / heavy-tail product pitch
4 months ago
requirements.txt
Safe
445 Bytes
Switch default model to Qwen3-0.6B (head_dim=128, GQA); bump transformers>=4.51
4 months ago