Image-to-Text
Transformers
Safetensors
Turkish
lighton_ocr
image-text-to-text
ocr
document-understanding
turkish
enterprise
vision-language
werea
Instructions to use Werea-co/Werea-DocOCR-1B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Werea-co/Werea-DocOCR-1B with Transformers:
# Use a pipeline as a high-level helper # Warning: Pipeline type "image-to-text" is no longer supported in transformers v5. # You must load the model directly (see below) or downgrade to v4.x with: # 'pip install "transformers<5.0.0' from transformers import pipeline pipe = pipeline("image-to-text", model="Werea-co/Werea-DocOCR-1B")# Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("Werea-co/Werea-DocOCR-1B") model = AutoModelForMultimodalLM.from_pretrained("Werea-co/Werea-DocOCR-1B", device_map="auto") - Notebooks
- Google Colab
- Kaggle
| { | |
| "add_prefix_space": false, | |
| "backend": "tokenizers", | |
| "bos_token": null, | |
| "clean_up_tokenization_spaces": false, | |
| "eos_token": "<|im_end|>", | |
| "errors": "replace", | |
| "image_break_token": "<|vision_pad|>", | |
| "image_end_token": "<|vision_end|>", | |
| "image_token": "<|image_pad|>", | |
| "is_local": false, | |
| "local_files_only": false, | |
| "max_length": null, | |
| "model_max_length": 131072, | |
| "model_specific_special_tokens": { | |
| "image_break_token": "<|vision_pad|>", | |
| "image_end_token": "<|vision_end|>", | |
| "image_token": "<|image_pad|>" | |
| }, | |
| "pad_to_multiple_of": null, | |
| "pad_token": "<|endoftext|>", | |
| "pad_token_type_id": 0, | |
| "padding_side": "right", | |
| "processor_class": "LightOnOcrProcessor", | |
| "split_special_tokens": false, | |
| "tokenizer_class": "TokenizersBackend", | |
| "unk_token": null | |
| } | |