Hide token stream from Output box; show complete text only at end (no flicker) 704af0d GPUburnout commited on Jun 22
Yield placeholders between progress stages so the bar actually moves 03eb13d GPUburnout commited on Jun 22
Real progress: track HF download tqdm + update bar during token streaming 375eb0e GPUburnout commited on Jun 22
Tune llama-cpp-python for HF Spaces 2-vCPU: n_ctx=512, n_threads=2, n_batch=8 fe4e741 GPUburnout commited on Jun 22
Pin llama-cpp-python to direct manylinux wheel URL (0.2.62 cp311) 7b7f3ca GPUburnout commited on Jun 22
Pin llama-cpp-python==0.2.90 with --prefer-binary for cp311 wheels 62f2ee4 GPUburnout commited on Jun 22
Switch 1B model to s2_hf loader (use HF transformers + safetensors) f4c74fc GPUburnout Claude Opus 4.6 (1M context) commited on Apr 6
Fix 1B model ID: gpuburnout-1b -> GPUburnout-1B-160K (correct HF repo name) f481d97 GPUburnout commited on Mar 30
Add GPUburnout-2B model and update subtitle to 2 billion parameters 992cbd9 GPUburnout commited on Mar 30
Rename all models to GPUburnout-* branding, default to 134M for faster first run b087aec GPUburnout commited on Mar 7