πŸš€ deglm-55M GGUF

deglm-55M in GGUF format, available in multiple quantization levels.

Technical Specifications

  • Total Parameters: 55,591,424 (~55.6M)
  • Vocabulary Size: 8,192
  • Embedding Dimensions: 512
  • Hidden Layers: 16
  • Attention Heads: 8
  • Head Dimension: 64
  • Context Length: 512
  • Format: GGUF

Available Quantizations:

Quantization File
Q8_0 deglm-55m.q8_0.gguf
Q4_K_M deglm-55m.q4_k_m.gguf
Q4_0 deglm-55m.q4_0.gguf
Downloads last month
-
GGUF
Model size
55.6M params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Collection including RaspizdAI/deglm-55M-GGUF