π₯ GOD MODE ACTIVATED: Subprocess monitoring, Context Guardrails, and No-MMAP Turbo Boost 5a4ab4b abedelbahnasy55 commited on Apr 17
π¦ TRAFFIC JAM HACK: Added Parallel Processing & Invisible Keep-Alive Heartbeats 97eb54a abedelbahnasy55 commited on Apr 17
π§ Expand Context Window to 8192 tokens for long documents 4bf1c84 abedelbahnasy55 commited on Apr 17
π PERFECT OPTIMIZATION: KV Cache 8-bit Quantization & Flash Attn for max CPU speed a3e247f abedelbahnasy55 commited on Apr 17
Revert to working Q5_K_XL model with optimized performance settings 5333a08 abedelbahnasy55 commited on Apr 14
Ultimate speed optimization: Q4_K_M model + tuned threads and cache 7d46968 abedelbahnasy55 commited on Apr 14
Revert to official llama.cpp (fixes mmproj gemma4v error) + keep speed optimizations 03cdc9a abedelbahnasy55 commited on Apr 14
Enable SSL support in llama-server for HTTPS image downloads 169f223 abedelbahnasy55 commited on Apr 13
Switch to manual CMake build of llama-server for stability b3ad89a abedelbahnasy55 commited on Apr 13