How to use James040/llama-cpp-python-wheels with llama-cpp-python:
# !pip install llama-cpp-python from llama_cpp import Llama llm = Llama.from_pretrained( repo_id="James040/llama-cpp-python-wheels", filename="{{GGUF_FILE}}", )
output = llm( "Once upon a time,", max_tokens=512, echo=True ) print(output)
The community tab is the place to discuss and collaborate with the HF community!