Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
itroot
/
Qwen3-4B-Instruct-2507-W8A8
like
0
Text Generation
Safetensors
qwen3
qwen
qwen3-4b
instruct
quantized
w8a8
llm-compressor
conversational
8-bit precision
compressed-tensors
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
Qwen3-4B-Instruct-2507-W8A8
5.21 GB
Ctrl+K
Ctrl+K
1 contributor
History:
7 commits
itroot
Update README.md
91d7191
verified
11 months ago
.gitattributes
Safe
1.57 kB
Add W8A8 quantized model files
11 months ago
README.md
2.64 kB
Update README.md
11 months ago
added_tokens.json
Safe
707 Bytes
Add W8A8 quantized model files
11 months ago
chat_template.jinja
Safe
4.04 kB
Add W8A8 quantized model files
11 months ago
config.json
2.71 kB
Add W8A8 quantized model files
11 months ago
generation_config.json
Safe
213 Bytes
Add W8A8 quantized model files
11 months ago
merges.txt
Safe
1.67 MB
Add W8A8 quantized model files
11 months ago
model-00001-of-00002.safetensors
4.41 GB
xet
Add W8A8 quantized model files
11 months ago
model-00002-of-00002.safetensors
778 MB
xet
Add W8A8 quantized model files
11 months ago
model.safetensors.index.json
Safe
54.9 kB
Add W8A8 quantized model files
11 months ago
recipe.yaml
Safe
536 Bytes
Add W8A8 quantized model files
11 months ago
special_tokens_map.json
Safe
613 Bytes
Add W8A8 quantized model files
11 months ago
tokenizer.json
Safe
11.4 MB
xet
Add W8A8 quantized model files
11 months ago
tokenizer_config.json
Safe
5.4 kB
Add W8A8 quantized model files
11 months ago
vocab.json
Safe
2.78 MB
Add W8A8 quantized model files
11 months ago