fix: contextual retrieval latency β max_chunks gate + prompt trim + model update dec430b benroshan Claude Sonnet 4.6 commited on Jun 28
fix: revert out-of-scope model default change in contextualize_chunks 0155d2c benroshan commited on Jun 27
feat: semantic chunking via SemanticChunker, toggled by config.yaml semantic_enabled b003dda benroshan commited on Jun 23
fix: reduce max_concurrent 20β3 to stay under Groq 6k TPM; smart 429 retry wait 2ed79bb benroshan commited on Jun 20
feat: contextual retrieval in production via async BackgroundTask ee2661c benroshan commited on Jun 20
feat: add contextualize_chunks() for LLM-prefixed chunk embeddings f181551 benroshan commited on Jun 20
fix: URL size guard for OOM prevention, eliminate double retrieval in chat, fix print 6c9f809 benroshan Claude Sonnet 4.6 commited on Jun 13
feat: multi-workspace backend β workspace_id routes ChromaDB collections + BM25 dict 1e5c62a benroshan Claude Sonnet 4.6 commited on Jun 13
feat: rebrand FinRAG β Prism (collection names, system prompt, UI branding) 7dac988 benroshan Claude Sonnet 4.6 commited on Jun 13