Commit History

Add the family-naming model as stage 4; remove the gpu/cpu runtime choice
4a6ccb0
verified

BentoUniAcc commited on

drop llama-cpp-python (unbuildable on Spaces); explain MiMo-over-Gemma choice
0e69d49
verified

BentoUniAcc commited on

runtime picker (gpu/cpu); shorter GPU grant to stretch the free ZeroGPU quota
346a1ad
verified

BentoUniAcc commited on

ZeroGPU: duration under the 300s cap, weights prefetched outside the grant
c6af343
verified

BentoUniAcc commited on

dual runtime: @spaces.GPU 4-bit NF4 path (Part B config) + llama.cpp CPU fallback
8052148
verified

BentoUniAcc commited on

MiMo-7B PDF injection detector: CPU/GGUF runtime, verbatim Part A+B logic
fd7251d
verified

BentoUniAcc commited on