deploy: Space overrides -- CUDA requirements.txt and app_space.py entry point e901ca7 atakan Claude Opus 5 commited on 4 days ago
refactor: Collapse four inference backends into one MLX path 9e637cd atakan Claude Opus 5 commited on 5 days ago
fix: Use a prebuilt CPU wheel for llama-cpp-python, not a source build 84d5d03 atakan Claude Sonnet 5 commited on 7 days ago
fix: Wire up the GGUF deployment path and fix its tokenizer mismatch e1f9681 atakan Claude Sonnet 5 commited on 7 days ago
fix: Include the actual Gradio/ZeroGPU code changes missed in the last commit b24b1e1 atakan Claude Sonnet 5 commited on 9 days ago
feat: Enforce high-speed GGUF Q4_K_M llama_cpp execution on Hugging Face Spaces CPU cb61511 atakan commited on 9 days ago
fix: Add jsonschema, scikit-learn, pyyaml, and dependencies to requirements.txt 16e844c atakan commited on 9 days ago
fix: Simplify Colab notebook, fix rank_bm25 import handling, and clean up prompts 1f68b63 atakan commited on 9 days ago