AuctionRouter / backend /app /config.py

Commit History

nav: icon tabs, and a roomier bid token cap
ebf62dc

dakshtaneja Claude Opus 5 commited on

config: strip whitespace from settings, fixing the prod image search
04475cf

dakshtaneja Claude Opus 5 commited on

pipeline: Granite web-search gate, racing the hedge
567b94e

dakshtaneja Claude Opus 5 commited on

images on web-search answers, via Tavily
7c53dcc

dakshtaneja Claude Opus 5 commited on

web-need detection at t=0, shared answer formatting, friendly copy
f54215b

dakshtaneja Claude Opus 4.8 commited on

router: GPT-OSS 120B generalist + Llama 4 Scout verifier; greeting fix; input UI
739705b

dakshtaneja Claude Opus 4.8 commited on

bid_timeout 20->30s; narrow AI bubbles to 68%; round the input box
92b2beb

dakshtaneja commited on

tier-1 lineup: Gemma 4 (general), Qwen3 Coder (coding), DeepSeek V4 Flash (logic/math)
f3a0f20

dakshtaneja commited on

qwen: force low bids outside coding
9284648

dakshtaneja commited on

frontier: GPT-5 -> GPT-5.6 Terra
2848bbb

dakshtaneja commited on

tier-1: DeepSeek V4 Flash (general), Qwen (coding), Skyfall 36B V2 (creative)
dabb994

dakshtaneja commited on

general slot: Llama 3.1 70B -> Llama 4 Maverick
48c2400

dakshtaneja commited on

general slot: Hermes 4 70B -> Llama 3.1 70B Instruct
62aaf69

dakshtaneja commited on

web search: freshness-gated OpenRouter web plugin for the agents
3072e01

dakshtaneja commited on

bidder answers: go long when detail is asked; raise token caps
175c7ab

dakshtaneja commited on

general slot: Gemini 2.5 Flash Lite -> Hermes 4 70B
0423963

dakshtaneja commited on

tier-1: Gemini 2.5 Flash Lite (general), DeepSeek V4 Flash (math), Qwen3 Coder (coding)
eea68eb

dakshtaneja commited on

revert verifier: Tencent Hy3 -> GPT-OSS 120B
40fa5b5

dakshtaneja commited on

coding slot: Qwen3 Coder Flash (paid) -> Qwen3 Coder (free-first)
8c8accb

dakshtaneja commited on

tier-1 reshuffle: V4 Flash (general), Qwen3 Coder Flash (coding), Nemotron (math)
8bbe7e0

dakshtaneja commited on

tier-1 lineup: DeepSeek V4 Pro (general), V4 Flash (math), Nemotron 3 Ultra (coding)
227a6c5

dakshtaneja commited on

swap math bidder: DeepSeek chat -> DeepSeek V4 Flash
c71a6f8

dakshtaneja commited on

swap models: Nemotron 3 Ultra as general bidder, Tencent Hy3 as verifier
50ccdb2

dakshtaneja commited on

deploy hardening: access gate, rate limiting, spend guard, env CORS
98d4633

dakshtaneja commited on

eval harness + env-swappable frontier model
6d1ccea

dakshtaneja commited on

hard gate: GPT-5 reserved exclusively for hard STEM/reasoning queries
e719504

dakshtaneja commited on

treat ambiguity as normal, not escalation-worthy
efa41b3

dakshtaneja commited on

restore 16k frontier token cap; escalations are rare and hard now
32e48a1

dakshtaneja commited on

remove the soft bid timeout: wait for all bidders
d2ca1dc

dakshtaneja commited on

topic toggle takes auction priority when its model bids confidently
4780921

dakshtaneja commited on

hedged speculative draft with user topic toggle
9eeb459

dakshtaneja commited on

low-confidence bidders skip the reason field; fix truncated deepseek specialty
3a9d087

dakshtaneja commited on

adaptive frontier token budget alongside effort
8423f9a

dakshtaneja commited on

adaptive frontier reasoning effort from bid difficulty estimates
081f1a1

dakshtaneja commited on

calibrate escalation: subjective queries stay tier-1, frontier answers shorter
57a91ca

dakshtaneja commited on

streaming-first delivery: show drafts immediately, verify in parallel
bc392c6

dakshtaneja commited on

cut easy-query latency: paid bid endpoints, low-effort verifier, grace-period bidding
587a7a1

dakshtaneja commited on

speed up inference: speculative drafts in bids, early-exit bidding, shorter retry sleeps
0cbd09d

dakshtaneja commited on

checkpoint: current behavior before latency work
001f1df

dakshtaneja commited on

AuctionRouter: cost-aware multi-agent LLM router with auction-based selection and verification-gated escalation
6b3d40b

dakshtaneja Claude Fable 5 commited on