Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

amd-shark
/
sdxl-quant-fp8

Model card Files Files and versions
xet
Community
1
sdxl-quant-fp8
47.4 GB
Ctrl+K
Ctrl+K
  • 4 contributors
History: 29 commits
denchic35's picture
denchic35
Create README.md
6140e3a verified almost 2 years ago
  • all_linear_sym_8_calib8
    Fix names about 2 years ago
  • all_sym_8_calib10
    MI250 QKV fused and all layers sym, FP8 attention, guidance scale 8, calib steps 10 about 2 years ago
  • brevitas
    updated quant_params with QKV fusion about 2 years ago
  • linear_conv_fp8_sdpa_fp16_eq_bl
    Create config.json almost 2 years ago
  • linear_conv_fp8_sdpa_fp16_no_eq_bl
    Create config.json almost 2 years ago
  • linear_conv_fp8_sdpa_fp8_eq_bl
    Create config.json almost 2 years ago
  • linear_conv_fp8_sdpa_fp8_no_eq_bl
    Create config.json almost 2 years ago
  • nvidia_fp8_unet
    Upload nvidia_fp8_unet/params.safetensors with huggingface_hub almost 2 years ago
  • .gitattributes
    2.08 kB
    Added models that are fully quantized with FP8. almost 2 years ago
  • README.md
    0 Bytes
    Create README.md almost 2 years ago
  • attn.py
    6.26 kB
    Added SDPA math model & test about 2 years ago
  • sdxl.json
    2.19 MB
    Upload sdxl.json with huggingface_hub about 2 years ago
  • sdxl.safetensors
    5.14 GB
    xet
    Upload sdxl.safetensors with huggingface_hub about 2 years ago
  • test_attn.py
    1.29 kB
    Added SDPA math model & test about 2 years ago