qaihm-bot commited on
Commit
84fffe5
·
verified ·
1 Parent(s): 7f4e98b

See https://github.com/qualcomm/ai-hub-models/releases/v0.59.0 for changelog.

Files changed (2) hide show
  1. README.md +15 -3
  2. release_assets.json +1 -1
README.md CHANGED
@@ -16,7 +16,7 @@ pipeline_tag: text-generation
16
  Granite 4.0 is a family of open language models from IBM designed for enterprise AI workloads including code generation, summarization, and retrieval-augmented generation.
17
 
18
  This is based on the implementation of Granite-4.0-Micro found [here](https://huggingface.co/ibm-granite/granite-4.0-h-micro).
19
- This repository contains pre-exported model files optimized for Qualcomm® devices. You can use the [Qualcomm® AI Hub Models](https://github.com/qualcomm/ai-hub-models/blob/v0.58.0/src/qai_hub_models/models/granite_4_0_micro) library to export with custom configurations. More details on model performance across various devices, can be found [here](#performance-summary).
20
 
21
  Qualcomm AI Hub Models uses [Qualcomm AI Hub Workbench](https://workbench.aihub.qualcomm.com) to compile, profile, and evaluate this model. [Sign up](https://myaccount.qualcomm.com/signup) to run these models on a hosted Qualcomm® device.
22
 
@@ -40,14 +40,14 @@ For more device-specific assets and performance metrics, visit **[Granite-4.0-Mi
40
 
41
  ### Option 2: Export with Custom Configurations
42
 
43
- Use the [Qualcomm® AI Hub Models](https://github.com/qualcomm/ai-hub-models/blob/v0.58.0/src/qai_hub_models/models/granite_4_0_micro) Python library to compile and export the model with your own:
44
  - Custom weights (e.g., fine-tuned checkpoints)
45
  - Custom input shapes
46
  - Target device and runtime configurations
47
 
48
  This option is ideal if you need to customize the model beyond the default configuration provided here.
49
 
50
- See our repository for [Granite-4.0-Micro on GitHub](https://github.com/qualcomm/ai-hub-models/blob/v0.58.0/src/qai_hub_models/models/granite_4_0_micro) for usage instructions.
51
 
52
  ## Model Details
53
 
@@ -80,6 +80,18 @@ See our repository for [Granite-4.0-Micro on GitHub](https://github.com/qualcomm
80
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 25.288631 | 0.51570025 - 16.502408
81
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 24.661293 | 0.54267315625 - 17.365541
82
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 11.122384 | 0.37017084375000003 - 11.845467000000001
 
 
 
 
 
 
 
 
 
 
 
 
83
 
84
  ## License
85
  * The license for the original implementation of Granite-4.0-Micro can be found
 
16
  Granite 4.0 is a family of open language models from IBM designed for enterprise AI workloads including code generation, summarization, and retrieval-augmented generation.
17
 
18
  This is based on the implementation of Granite-4.0-Micro found [here](https://huggingface.co/ibm-granite/granite-4.0-h-micro).
19
+ This repository contains pre-exported model files optimized for Qualcomm® devices. You can use the [Qualcomm® AI Hub Models](https://github.com/qualcomm/ai-hub-models/blob/v0.59.0/src/qai_hub_models/models/granite_4_0_micro) library to export with custom configurations. More details on model performance across various devices, can be found [here](#performance-summary).
20
 
21
  Qualcomm AI Hub Models uses [Qualcomm AI Hub Workbench](https://workbench.aihub.qualcomm.com) to compile, profile, and evaluate this model. [Sign up](https://myaccount.qualcomm.com/signup) to run these models on a hosted Qualcomm® device.
22
 
 
40
 
41
  ### Option 2: Export with Custom Configurations
42
 
43
+ Use the [Qualcomm® AI Hub Models](https://github.com/qualcomm/ai-hub-models/blob/v0.59.0/src/qai_hub_models/models/granite_4_0_micro) Python library to compile and export the model with your own:
44
  - Custom weights (e.g., fine-tuned checkpoints)
45
  - Custom input shapes
46
  - Target device and runtime configurations
47
 
48
  This option is ideal if you need to customize the model beyond the default configuration provided here.
49
 
50
+ See our repository for [Granite-4.0-Micro on GitHub](https://github.com/qualcomm/ai-hub-models/blob/v0.59.0/src/qai_hub_models/models/granite_4_0_micro) for usage instructions.
51
 
52
  ## Model Details
53
 
 
80
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 25.288631 | 0.51570025 - 16.502408
81
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 24.661293 | 0.54267315625 - 17.365541
82
  | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Snapdragon® X Elite | 4096 | 11.122384 | 0.37017084375000003 - 11.845467000000001
83
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 512 | 17.843282 | 0.69928525 - 2.797141
84
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 512 | 13.447186 | 0.6960275 - 2.78411
85
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 512 | 11.275263 | 0.34302550000000004 - 1.3721020000000002
86
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 512 | 25.130048 | 0.9481944999999999 - 3.7927779999999998
87
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 512 | 26.288118 | 0.95749725 - 3.829989
88
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 512 | 15.760148 | 0.30860325 - 1.234413
89
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 4096 | 25.288631 | 0.51570025 - 16.502408
90
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 4096 | 24.661293 | 0.54267315625 - 17.365541
91
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ IQ-X7181 | 4096 | 11.122384 | 0.37017084375000003 - 11.845467000000001
92
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 4096 | 22.639083 | 1.1536551875 - 36.916966
93
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 4096 | 23.049816 | 1.22555709375 - 39.217827
94
+ | Granite-4.0-Micro | GENIEX_LLAMACPP | q4_0 | Qualcomm® Dragonwing™ Q-8750 | 4096 | 14.515548 | 0.36028896875 - 11.529247
95
 
96
  ## License
97
  * The license for the original implementation of Granite-4.0-Micro can be found
release_assets.json CHANGED
@@ -1,5 +1,5 @@
1
  {
2
- "version": "0.58.0",
3
  "precisions": {
4
  "q4_0": {
5
  "universal_assets": {
 
1
  {
2
+ "version": "0.59.0",
3
  "precisions": {
4
  "q4_0": {
5
  "universal_assets": {