Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
FINAL-Bench
/
POCKET-35B-CPU
like
26
Running
on
CPU Upgrade
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
POCKET-35B-CPU
28.1 kB
Ctrl+K
Ctrl+K
1 contributor
History:
23 commits
SeaWolf-AI
A/B: raise max_tokens 256->512 so POCKET's reasoning + final answer complete (not cut)
336526e
verified
about 6 hours ago
.gitattributes
Safe
1.52 kB
initial commit
about 9 hours ago
Dockerfile
Safe
1.48 kB
add Tab2 live A/B vs Bonsai (sequential: Bonsai first, then POCKET); 2nd llama-server for Bonsai Q1_0; README=BONSAI vs POCKET
about 7 hours ago
README.md
Safe
1.31 kB
A/B is now the landing tab (swap order); add prism-ml/Bonsai-27B-gguf to models relation
about 6 hours ago
app.py
Safe
5.7 kB
A/B: show reasoning again (remove no_think prefill); raise A/B max_tokens 96->256 so thinking completes
about 6 hours ago
index.html
Safe
16.9 kB
A/B: raise max_tokens 256->512 so POCKET's reasoning + final answer complete (not cut)
about 6 hours ago
requirements.txt
Safe
81 Bytes
rebuild backend on upstream llama.cpp master (qwen35moe support) via llama-server; FastAPI proxies; portable CPU build
about 8 hours ago
start.sh
Safe
1.07 kB
add Tab2 live A/B vs Bonsai (sequential: Bonsai first, then POCKET); 2nd llama-server for Bonsai Q1_0; README=BONSAI vs POCKET
about 7 hours ago