Post
90
Just posted an article about my adventures with Qwen3.8-Next-Flash running on a DGX Spark. Got up to 114 tok/s across 8 concurrent streams. https://huggingface.co/blog/krisbailey/shortlist-mtp-mostly-worked-around-a-bad-kernel-ch
Join the community of Machine Learners and AI enthusiasts.
Sign Up