fix: disable GLM thinking (reasoning model burns max_tokens on reasoning_content -> empty answers); bump max_tokens to 500 08aba80 balade commited on Aug 12