Duplicated from KBaba7/llama.cpp
A newer version of the Gradio SDK is available: 6.17.3
6.17.3
A passkey retrieval task is an evaluation method used to measure a language models ability to recall information from long contexts.
See the following PRs for more info:
make -j && ./llama-passkey -m ./models/llama-7b-v2/ggml-model-f16.gguf --junk 250