Lower ZeroGPU duration to 45s (model pre-downloaded; cuts quota per call) 7cffb3a AfkaraLP commited on 8 days ago
Surface ZeroGPU worker traceback via try/except for diagnosis 437ce51 AfkaraLP commited on 8 days ago
Preload nvidia CUDA libs (RTLD_GLOBAL) so libllama.so finds libcudart on ZeroGPU; eager model download; duration=120 4a0f59e AfkaraLP commited on 8 days ago
gradio 6 fixes: Code language=None (rust unsupported), move theme to launch() 1e7d11c AfkaraLP commited on 8 days ago
Defer llama_cpp import into @spaces.GPU path (ZeroGPU: CUDA libs only exist inside the GPU call) 173decb AfkaraLP commited on 8 days ago
RustLean FIM Space: cursor-aware Rust completion via llama-cpp-python + ZeroGPU ef83975 AfkaraLP commited on 8 days ago