Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
zrxarchit
/
inference
like
0
Sleeping
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
30426eb
inference
11.2 kB
Ctrl+K
Ctrl+K
2 contributors
History:
6 commits
0xarchit
Support CONTEXT_LENGTH env var (default 4096) for ctx-size
30426eb
3 months ago
.dockerignore
Safe
61 Bytes
Add optimized llama.cpp CPU inference backend
3 months ago
.gitattributes
Safe
1.52 kB
initial commit
3 months ago
.gitignore
Safe
8 Bytes
Add optimized llama.cpp CPU inference backend
3 months ago
Dockerfile
Safe
1.85 kB
Use Python venv in runtime to avoid PEP 668 pip install error
3 months ago
README.md
Safe
3.09 kB
Add optimized llama.cpp CPU inference backend
3 months ago
download_model.py
Safe
2.99 kB
Add optimized llama.cpp CPU inference backend
3 months ago
requirements.txt
Safe
24 Bytes
Add optimized llama.cpp CPU inference backend
3 months ago
start.sh
Safe
1.71 kB
Support CONTEXT_LENGTH env var (default 4096) for ctx-size
3 months ago