Pass HF_TOKEN explicitly to SentenceTransformer for gated model download 7787ac2 Vasanth6 commited on Jun 25
Optimize PyTorch threading, run generation in thread, and pre-warm model at startup d9fc469 Vasanth6 commited on Jun 25
Optimize referee model to Qwen2.5-0.5B-Instruct for CPU and add warning popup 3e566d1 Vasanth6 commited on Jun 25