Hugging Face
Models
Datasets
Spaces
Community
Docs
Enterprise
Pricing
Log In
Sign Up
asrith05
/
deepseek_pretrain_90k
like
0
Text Generation
Transformers
Safetensors
English
Telugu
Sanskrit
deepseek_v3
multilingual
pretrained
deepseek
base-model
text-generation-inference
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Deploy
Use this model
main
deepseek_pretrain_90k
/
generation_config.json
asrith05
Upload DeepSeek pretrained multilingual model (90k steps)
1aff7ae
verified
4 months ago
raw
Copy download link
history
blame
contribute
delete
Safe
140 Bytes
{
"_from_model_config"
:
true
,
"bos_token_id"
:
0
,
"eos_token_id"
:
1
,
"transformers_version"
:
"4.51.3"
,
"use_cache"
:
false
}