Abed-gemma4-chat-api / Dockerfile

Commit History

Update Dockerfile
22a11e8
verified

abedgemma commited on

Update Dockerfile
b866073
verified

abedgemma commited on

Update Dockerfile
e789f09
verified

abedgemma commited on

Update Dockerfile
51528bc
verified

abedgemma commited on

Update Dockerfile
a3577fb
verified

abedgemma commited on

πŸ”₯ GOD MODE ACTIVATED: Subprocess monitoring, Context Guardrails, and No-MMAP Turbo Boost
5a4ab4b

abedelbahnasy55 commited on

🚦 TRAFFIC JAM HACK: Added Parallel Processing & Invisible Keep-Alive Heartbeats
97eb54a

abedelbahnasy55 commited on

🧠 Expand Context Window to 8192 tokens for long documents
4bf1c84

abedelbahnasy55 commited on

πŸ”§ Fix llama-server flag: specify --flash-attn on
e554a91

abedelbahnasy55 commited on

πŸš€ PERFECT OPTIMIZATION: KV Cache 8-bit Quantization & Flash Attn for max CPU speed
a3e247f

abedelbahnasy55 commited on

πŸ”§ Remove unsupported --image-max-pixels flag
5b51321

abedelbahnasy55 commited on

πŸ”§ Fix OOM crash: reduce memory usage and image bounds
bd5d0f1

abedelbahnasy55 commited on

Final performance tuning: threads 2/4 + mlock
6f8d76c

abedelbahnasy55 commited on

Use Python script for reliable model download
a2ce8c8

abedelbahnasy55 commited on

Switch to reliable huggingface-cli for model download
c6c330f

abedelbahnasy55 commited on

Fix flash-attn syntax and revert to stable llama.cpp
c856ee5

abedelbahnasy55 commited on

Deploy nuclear performance optimizations
79b09ad

abedelbahnasy55 commited on

Revert to working Q5_K_XL model with optimized performance settings
5333a08

abedelbahnasy55 commited on

Ultimate speed optimization: Q4_K_M model + tuned threads and cache
7d46968

abedelbahnasy55 commited on

Revert to official llama.cpp (fixes mmproj gemma4v error) + keep speed optimizations
03cdc9a

abedelbahnasy55 commited on

Remove unsupported -ctxcp argument for ik_llama.cpp
022bba3

abedelbahnasy55 commited on

Apply major performance optimizations for model and API
950c767

abedelbahnasy55 commited on

Add CA certificates and enable CURL in llama.cpp build
71a1224

abedelbahnasy55 commited on

Enable SSL support in llama-server for HTTPS image downloads
169f223

abedelbahnasy55 commited on

Switch to manual CMake build of llama-server for stability
b3ad89a

abedelbahnasy55 commited on

Fix llama-server startup command
e7dff04

abedelbahnasy55 commited on

Switch to llama-cpp-python and fix build
5b8300e

abedelbahnasy55 commited on

Add all project files for Gemma 4 deployment
4225e0b

abedelbahnasy55 commited on