Respite-API / app.py

Commit History

fix(gate): accept x-ael-key alongside x-respite-key; drop dead key read
b3596ff

pi commited on

security: give the searx bridge its own gate key
64e2bf0

cazyundee commited on

security: key-gate the searx bridge
d2bc83a

cazyundee commited on

chore: remove the temporary env diagnostic
8251b07
verified

cazyundee commited on

chore: temporary names-only env diagnostic
977bce0
verified

cazyundee commited on

fix: read the shared secret per request; add a names-only env diagnostic
3fee5c9
verified

cazyundee commited on

chore: rebuild so the container picks up RESPIRE_SPACE_KEY
bb7d60b
verified

cazyundee commited on

security: require a server-held shared secret for the compute routes
13dcd95
verified

cazyundee commited on

security: overwrite Gradio's reflected CORS on /respite/*
f1ffefe
verified

cazyundee commited on

Add SearXNG userspace runtime with DDG + Google engines
42f6abb

Buffy commited on

Add Node runtime capability probe (userspace, no root)
ba9ed6e

Buffy commited on

Expose live audio generation API route
58e1c4d
verified

cazyundee commited on

Pass Qwen prompt through async job
d3842e2
verified

cazyundee commited on

Fix Qwen multimodal prompt content
024ac84
verified

cazyundee commited on

Diagnose Qwen generation termination
8f2c140
verified

cazyundee commited on

Enable Qwen reasoning by default
ca39a8e
verified

cazyundee commited on

Add CPU Qwen3.8 inference job
e8b6e3f
verified

cazyundee commited on

Normalize TinyLlama chat template inputs
9a9fa33
verified

cazyundee commited on

Fix CPU TinyLlama PyTorch initialization
4380259
verified

cazyundee commited on

Accept JSON body for TinyLlama inference
fc0acf8
verified

cazyundee commited on

Add CPU TinyLlama inference to Gradio Space
95167c1
verified

cazyundee commited on

Test RAM loaded GGUF without pipe deadlock
3956ce2
verified

cazyundee commited on

Rebuild llama.cpp with stable CPU backend
f3c6ec5
verified

cazyundee commited on

Limit CPU thread pools to cgroup CPUs
1b0a401
verified

cazyundee commited on

Prevent llama output pipe deadlock
a8a716e
verified

cazyundee commited on

Probe llama CPU process state
8ef48d6
verified

cazyundee commited on

Run loader diagnosis as background job
3fdbf36
verified

cazyundee commited on

Add CPU model loader diagnostics
a5de9d9
verified

cazyundee commited on

Run CPU benchmarks asynchronously
318e7cb
verified

cazyundee commited on

Diagnose single-thread CPU model loading
50b3226
verified

cazyundee commited on

Add tiny GGUF CPU control test
74b57c5
verified

cazyundee commited on

Use RAM loaded CPU model benchmark
6a6f8c0
verified

cazyundee commited on

Separate model loading from CPU warmup
005dc00
verified

cazyundee commited on

Add CPU model smoke diagnostics
a903372
verified

cazyundee commited on

Serialize CPU benchmark requests
75c0e18
verified

cazyundee commited on

Keep local model benchmarks CPU only
677b59e
verified

cazyundee commited on

Select llama benchmark executable
0329f82
verified

cazyundee commited on

Fix llama benchmark binary cache
6e550c2
verified

cazyundee commited on

Use llama bench for bounded CPU measurements
4c94260
verified

cazyundee commited on

Fix CPU benchmark diagnostics
ce1fcf1
verified

cazyundee commited on

Restore stable Space server launch
bc0692a
verified

cazyundee commited on

Add share=True to bypass broken proxy
15d9481
verified

cazyundee commited on

Restore working create_app monkey-patch with staticmethod
afb0c33
verified

cazyundee commited on

Background thread inject routes into running Gradio server
c7f0d89
verified

cazyundee commited on

Fix: inject routes via background thread after launch
5e3fbbf
verified

cazyundee commited on

Fix: patch App.__init__ to inject routes into the actual served app
d0e8976
verified

cazyundee commited on

Fix: monkey-patch without staticmethod wrapper, spaces before torch, ssr_mode=False
e244295
verified

cazyundee commited on

Fix route injection: try demo.app directly, fallback to monkey-patch
e3186fc
verified

cazyundee commited on

Add model filter param to benchmark endpoint
a76c80d
verified

cazyundee commited on

Fix: build llama.cpp with static linking, find binary after build
2ab0e84
verified

cazyundee commited on