This is allenai/Olmo-3.1-32B-Think quantized with LLM Compressor with the recipe in the "recipe.yaml" file. Not Tested
How the models perform (token efficiency, accuracy per domain, ...) and how to use them: Quantizing Olmo 3: Most Efficient and Accurate Formats
- Developed by: The Kaitchup
- License: Apache 2.0 license
How to Support My Work
Subscribe to The Kaitchup. Or you can "buy me a kofi".
- Downloads last month
- 7
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support
Model tree for Locosdeamor/Olmo-3.1-32B-Think-w8a8-smoothquant
Base model
allenai/Olmo-3-1125-32B Finetuned
allenai/Olmo-3-32B-Think-SFT Finetuned
allenai/Olmo-3-32B-Think-DPO Finetuned
allenai/Olmo-3.1-32B-Think