Text Generation
MLX
Safetensors
English
pretraining
from-scratch
small-language-model
post-training
silicon
Instructions to use OpenSML/OpenSML-150M with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use OpenSML/OpenSML-150M with MLX:
# Make sure mlx-lm is installed # pip install --upgrade mlx-lm # if on a CUDA device, also pip install mlx[cuda] # Generate text with mlx-lm from mlx_lm import load, generate model, tokenizer = load("OpenSML/OpenSML-150M") prompt = "Once upon a time in" text = generate(model, tokenizer, prompt=prompt, verbose=True) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- MLX LM
How to use OpenSML/OpenSML-150M with MLX LM:
Generate or start a chat session
# Install MLX LM uv tool install mlx-lm # Generate some text mlx_lm.generate --model "OpenSML/OpenSML-150M" --prompt "Once upon a time"
- Atomic Chat
Download tokenizer/report.json from OpenSML/OpenSML-150M: direct link, hf CLI and curl.
- Browser
- Download file 889 Bytes
-
https://huggingface.co/OpenSML/OpenSML-150M/resolve/main/tokenizer/report.json
- Command line
-
hf download hf://OpenSML/OpenSML-150M/tokenizer/report.json
-
curl -L -o report.json https://huggingface.co/OpenSML/OpenSML-150M/resolve/main/tokenizer/report.json
889 Bytes
| { | |
| "fixtures_passed": 9, | |
| "sources": { | |
| "cosmopedia": { | |
| "bytes": 402259, | |
| "bytes_per_token": 5.061007523716062, | |
| "documents": 103, | |
| "roundtrip_failures": 0, | |
| "seconds": 0.09366637503262609, | |
| "tokens": 79482 | |
| }, | |
| "dclm": { | |
| "bytes": 1001509, | |
| "bytes_per_token": 4.453209483494593, | |
| "documents": 165, | |
| "roundtrip_failures": 0, | |
| "seconds": 0.26016941701527685, | |
| "tokens": 224896 | |
| }, | |
| "web": { | |
| "bytes": 2205062, | |
| "bytes_per_token": 4.625467253451697, | |
| "documents": 559, | |
| "roundtrip_failures": 0, | |
| "seconds": 0.5487007499905303, | |
| "tokens": 476722 | |
| }, | |
| "wiki": { | |
| "bytes": 401775, | |
| "bytes_per_token": 4.059276397546905, | |
| "documents": 143, | |
| "roundtrip_failures": 0, | |
| "seconds": 0.10723229206632823, | |
| "tokens": 98977 | |
| } | |
| }, | |
| "vocab_size": 32000 | |
| } | |