Instructions to use cubbk/orpheus-swedish with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use cubbk/orpheus-swedish with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-generation", model="cubbk/orpheus-swedish") messages = [ {"role": "user", "content": "Who are you?"}, ] pipe(messages)# Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("cubbk/orpheus-swedish") model = AutoModelForCausalLM.from_pretrained("cubbk/orpheus-swedish", device_map="auto") messages = [ {"role": "user", "content": "Who are you?"}, ] inputs = tokenizer.apply_chat_template( messages, add_generation_prompt=True, tokenize=True, return_dict=True, return_tensors="pt", ).to(model.device) outputs = model.generate(**inputs, max_new_tokens=40) print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[-1]:])) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- vLLM
How to use cubbk/orpheus-swedish with vLLM:
Install from pip and serve model
# Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "cubbk/orpheus-swedish" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "cubbk/orpheus-swedish", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }'Use Docker
docker model run hf.co/cubbk/orpheus-swedish
- SGLang
How to use cubbk/orpheus-swedish with SGLang:
Install from pip and serve model
# Install SGLang from pip: pip install sglang # Start the SGLang server: python3 -m sglang.launch_server \ --model-path "cubbk/orpheus-swedish" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "cubbk/orpheus-swedish", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }'Use Docker images
docker run --gpus all \ --shm-size 32g \ -p 30000:30000 \ -v ~/.cache/huggingface:/root/.cache/huggingface \ --env "HF_TOKEN=<secret>" \ --ipc=host \ lmsysorg/sglang:latest \ python3 -m sglang.launch_server \ --model-path "cubbk/orpheus-swedish" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "cubbk/orpheus-swedish", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }' - Docker Model Runner
How to use cubbk/orpheus-swedish with Docker Model Runner:
docker model run hf.co/cubbk/orpheus-swedish
Commit History
test again 91d1e00
Dan commited on
don't cut at 14s 4ee53f5
Dan commited on
register pipeline c21da3a
Dan commited on
add pipeline.py correct e4f521b
Dan commited on
Upload MyPipeline 136c393 verified
Upload MyPipeline e540e72 verified
fix audio tags a814921
Dan commited on
update readme 9a2ce5d
Dan commited on
Track .wav files with Git LFS2 f220bd1
Dan commited on
make pipeline work fa4e6e4
Dan commited on