Initial commit: llama.cpp OpenAI-compatible server for Gemma 4 E2B 82bd3da vedgupta commited on Apr 26