Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

infosave
/
GLM-5.3-Flash-cmf

Text Generation
cortiq
English
Chinese
cmf
quantized
q2tp
q4tp
mixed-precision
Mixture of Experts
hybrid-attention
2-bit
4-bit precision
Model card Files Files and versions
xet
Community

Instructions to use infosave/GLM-5.3-Flash-cmf with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • cortiq

    How to use infosave/GLM-5.3-Flash-cmf with cortiq:

    # one Rust binary, no additional dependencies
    cargo install cortiq-cli   # or a prebuilt binary from github.com/infosave2007/cmf/releases
    hf download infosave/GLM-5.3-Flash-cmf --include "*.cmf" --local-dir .
    ls *.cmf                   # some repos ship more than one quantization
    cortiq run FILE.cmf --prompt "What is the capital of France?"
    cortiq serve FILE.cmf --port 8080   # OpenAI-compatible server
  • Notebooks
  • Google Colab
  • Kaggle
GLM-5.3-Flash-cmf
283 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 6 commits
infosave's picture
infosave
Document GLM-5.3-Flash Q2TP and bounded runtime
09eca95 verified about 1 hour ago
  • .gitattributes
    1.64 kB
    Add GLM-5.3-Flash Q2TP CMF about 1 hour ago
  • README.md
    7.47 kB
    Document GLM-5.3-Flash Q2TP and bounded runtime about 1 hour ago
  • glm-5.3-flash-q2tp.cmf
    116 GB
    xet
    Add GLM-5.3-Flash Q2TP CMF about 1 hour ago
  • glm-5.3-flash-q4tp.cmf
    167 GB
    xet
    Add files using upload-large-folder tool about 7 hours ago