rag-visualizer / backend /engines /llm_client.py

Commit History

Optimize PyTorch threading, run generation in thread, and pre-warm model at startup
d9fc469

Vasanth6 commited on

Optimize referee model to Qwen2.5-0.5B-Instruct for CPU and add warning popup
3e566d1

Vasanth6 commited on

Configure for 100% self-contained CPU deployment
e4bcad4

Vasanth6 commited on

local: phase 5 arena and text highlighter
25d4f70

Vasanth6 commited on