Add measured Galaxy S26 GPU-vs-CPU rows for the .litertlm bundle

#2
by mlboydaisuke - opened
LiteRT Community (FKA TFLite) org

Thanks for publishing this bundle β€” Phi-4-mini through litert-lm 0.16.0 on a Galaxy S26 GPU delegates fully and generates correctly.

This PR adds one subsection under Performance: measured Galaxy S26 (SM8850) rows for Phi-4-mini-instruct_multi-prefill-seq_q8_ekv4096.litertlm, GPU (OpenCL) against CPU (XNNPACK), conditions inline. Both backends generated a correct answer before the numbers were quoted; the GPU rows are full delegation (decode 1648/1648 ops on LITERT_CL). The section is labeled as a different runtime path from the S24 Ultra table above it, so the two are not conflated.

Happy to reshape to the house style if you prefer.

mlboydaisuke changed pull request status to open
Ready to merge
This branch is ready to get merged automatically.

Sign up or log in to comment