Post
I released my own benchmark sets behind emotional-memory, my library for mood-aware recall in LLM agents. The card opens with the limits: the main set is built to favour the method, and the negative results are listed up front, e.g. on LoCoMo AFT scores F1 0.168 against 0.271 for a naive RAG baseline. Use it to check where affect-conditioned retrieval helps and where it doesn't.
gianlucamazza/emotional-memory-benchmarks
gianlucamazza/emotional-memory-benchmarks