Fix LLM timeout: return StreamingResponse before agentic loop 6ab878f azettl Claude commited on 19 days ago
Cap max_tokens at 4096 to stay within HF Inference Provider limits 7ebe9fb azettl Claude commited on 19 days ago
Fix STT audio + faster speech: dual AudioContext, bump LLM to Qwen3-32B a390020 azettl Claude commited on 19 days ago