How to use from the
Use from the
Transformers library
# Load model directly
from transformers import AutoTokenizer, AutoModel

tokenizer = AutoTokenizer.from_pretrained("Qdrant/gte-large-onnx-Q")
model = AutoModel.from_pretrained("Qdrant/gte-large-onnx-Q", device_map="auto")
Quick Links

Quantized ONNX port of thenlper/gte-large for text classification and similarity searches.

Downloads last month
16
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support