Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

niranjanh123
/
lfm-grpo-gemini-simplified-reasoning

Safetensors
lfm2
Model card Files Files and versions
xet
Community
lfm-grpo-gemini-simplified-reasoning
2.35 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 10 commits
niranjanh123's picture
niranjanh123
Update README.md
b262667 verified 11 days ago
  • .gitattributes
    1.52 kB
    initial commit 17 days ago
  • LICENSE
    10.6 kB
    Upload LICENSE 11 days ago
  • README.md
    65 Bytes
    Update README.md 11 days ago
  • chat_template.jinja
    916 Bytes
    Upload GRPO model gemini_no_xml with eval efficiency 3.0 17 days ago
  • config.json
    1.29 kB
    Upload GRPO model gemini_no_xml with eval efficiency 3.0 17 days ago
  • generation_config.json
    142 Bytes
    Upload GRPO model gemini_no_xml with eval efficiency 3.0 17 days ago
  • model.safetensors
    2.34 GB
    xet
    Upload GRPO model gemini_no_xml (5x5 eff: 2.75, 10x10 eff: 2.95) 16 days ago
  • tokenizer.json
    4.73 MB
    Upload GRPO model gemini_no_xml with eval efficiency 3.0 17 days ago
  • tokenizer_config.json
    1.9 kB
    Upload GRPO model gemini_no_xml with eval efficiency 3.0 17 days ago