Can you provide Agentic and Reasoning Benchmarks?

#1
by DedeProGames - opened

Hey @Sweaterdog will you benchmark GRaPE-2.5 in Agentic or Reasoning Benchmarks? Like TerminalBench 2.1, SWE Bench Pro, NL2Repo, GPQA, and MMLU-Pro?

Skinnertopia Lab for Artifical Intelligence org

Sorry, but as compute required to run benchmarks is very very high, GRaPE 2.5 models will not have benchmarks. Our efforts are currently dedicated to experimenting with GRaPE 3 models. If someone from the community benchmarks Quasar, or Helios, we will accept a PR.

Sign up or log in to comment