Running Agents 15 Open Japanese LLM Leaderboard 🌸 15 Explore LLM benchmark rankings and submit your model for evaluation
Running Agents 4 LLM Evaluation Framework Demo 📊 4 Benchmark LLMs on accuracy, cost, and hallucination.
Sleeping Agents 1 PHANTASM LLM Hallucination Inverter 🔮 1 Invert LLM hallucination into productive features
Sleeping Agents 1 TemporalMesh Transformer Demo 🕸 1 Visualize TemporalMesh Transformer token flow and early exits