Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Parakon
/
parakon-runtime
Like
0
Follow
Parakon
3
parakon
llama.cpp
cuda
compression
License:
parakon-community-license
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
parakon-runtime
/
examples
/
model-conversion
/
scripts
/
causal
18.3 kB
Ctrl+K
Ctrl+K
1 contributor
History:
2 commits
kat0012
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 3)
7011d4a
verified
18 days ago
compare-embeddings-logits.sh
Safe
1.25 kB
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
compare-logits.py
Safe
3.23 kB
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
convert-model.sh
1.43 kB
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
modelcard.template
Safe
185 Bytes
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
run-casual-gen-embeddings-org.py
4.24 kB
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
run-converted-model-embeddings-logits.sh
Safe
636 Bytes
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 2)
18 days ago
run-converted-model.sh
Safe
822 Bytes
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 3)
18 days ago
run-org-model.py
Safe
6.49 kB
Parakon runtime: llama.cpp fork with fused 1-bit CUDA/Metal/CPU kernels (part 3)
18 days ago