onw 1.4.8: review fixes across the window, tray, chat, API and engine; uninstall; pip after the restart; memory bar and NPU load; read-only weight mappings on Windows
onw 1.4.6: ~30% faster answers on Windows - the server starts at normal priority and runs above normal while generating (below normal put it on the low-power cores)
onw 1.4.5: first compile ~3 min shorter (4/64-token DeltaNet parts go straight to one-graph NPUW), time-left estimate while compiling; README: Panther Lake compile-time note with a Lunar Lake comparison
onw 1.4.4: fix ~3 GB extra NPU memory from the parallel compile processes of 1.4.2/1.4.3 (off by default now; older NPUW caches compiled again once); a stopped first compile no longer splits the weights over two runs
onw 1.4.2: fix loading on Windows Lunar Lake (no memory-mapped blob import), parallel first compile on every NPU (off in the memory-saving mode), fix recompiling at every load on Panther Lake, merged compile-route records
onw 1.4.1: Panther Lake (NPU 5) runs - fix DEVICE_LOST at the first inference; fix compiling again at every load on Windows (read-only cache files); ~3x faster first compile on Panther Lake
onw 1.4.0: MTP draft tokens (~1.5x answers on Lunar Lake, NF4 Qwen3.5 family), 64-token prompt blocks on Lunar Lake, less host work per step, fix all-zero logit rows