"Can HAIDER-Math-32B Beat Top Math Models? Community Test & Feedback Thread”

#1
by ABDUL-HASEEB-TANOLI - opened

Hi everyone,

I’ve recently released HAIDER-Math-32B (Original), a model specifically designed for advanced mathematical reasoning. Early traction has been encouraging, and I’d like to invite the community to rigorously evaluate its performance.

🔍 Focus of evaluation:

  • Mathematical reasoning (not general chat)
  • Step-by-step solution quality
  • Final answer accuracy
  • Logical consistency across multi-step problems

🧪 Recommended test areas:

  • GSM8K-style word problems
  • Algebra and equation solving
  • Multi-step arithmetic
  • Edge-case and tricky reasoning problems

📊 If you test the model, please share:

  • The questions used
  • Model outputs
  • Correct answers
  • (Optional) Comparison with other strong math models

⚠️ Notes:

  • This is the original (full-precision) model
  • Results from this version will be considered the primary benchmark reference

🎯 Goal:
To evaluate whether this model can match or surpass strong mathematical reasoning baselines through open and transparent community testing.

I will compile and summarize community findings into a structured benchmark report.

Thanks in advance — your feedback will directly contribute to improving and validating the model.

ABDUL-HASEEB-TANOLI changed discussion status to closed

Sign up or log in to comment