Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
cs-552-2026-barn
/
general_knowledge_model
like
0
Follow
BARN
4
Safetensors
qwen3
License:
mit
Model card
Files
Files and versions
xet
Community
2
Copy to bucket
new
main
general_knowledge_model
7.52 GB
Ctrl+K
Ctrl+K
2 contributors
History:
10 commits
Nahush-27
Upload folder using huggingface_hub
d7a368e
verified
about 2 months ago
.gitattributes
Safe
1.57 kB
Add patched chat template: thinking ON + system prompt baked in
2 months ago
EVAL_REPORT.md
Safe
1.24 kB
Automated MNLP evaluation report (2026-05-20) (#1)
2 months ago
README.md
Safe
21 Bytes
initial commit
3 months ago
chat_template.jinja
Safe
4.77 kB
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago
config.json
Safe
1.42 kB
v7 ck1100: GRPO from base_fmt, step 1100/4000 (27.5%), eval-v2 16k: 57.2% MMLU-Pro / 27.8% SuperGPQA / 42.5% overall
about 2 months ago
generation_config.json
Safe
239 Bytes
Upload folder using huggingface_hub
about 2 months ago
merges.txt
Safe
1.67 MB
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago
model-00001-of-00002.safetensors
Safe
3.44 GB
xet
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago
model-00002-of-00002.safetensors
Safe
622 MB
xet
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago
model.safetensors
3.44 GB
xet
Upload folder using huggingface_hub
about 2 months ago
model.safetensors.index.json
Safe
25.6 kB
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago
tokenizer.json
Safe
11.4 MB
xet
v7 ck1100: GRPO from base_fmt, step 1100/4000 (27.5%), eval-v2 16k: 57.2% MMLU-Pro / 27.8% SuperGPQA / 42.5% overall
about 2 months ago
tokenizer_config.json
Safe
693 Bytes
v7 ck1100: GRPO from base_fmt, step 1100/4000 (27.5%), eval-v2 16k: 57.2% MMLU-Pro / 27.8% SuperGPQA / 42.5% overall
about 2 months ago
vocab.json
Safe
2.78 MB
Replace grpo_gk with base_fmt: Qwen3-1.7B base + format-forcing chat template (default system prompt + per-question boxed reminder). Zero training. eval-v2 34.7% overall / 50.2% MMLU-Pro (vs grpo_gk 27.0%). Template pre-baked; not re-patched.
about 2 months ago