Popular repositories Loading
-
qwen3.8-flash-next-dgx-spark-sglang
qwen3.8-flash-next-dgx-spark-sglang PublicProduction recipe for serving Qwen3.8-Flash-Next (180B Hybrid MoE) on a single 128GB NVIDIA DGX Spark with SGLang, NVMe PLE offloading & MTP speculative decoding.
Jinja 4
-
glm-5.3-flash-dgx-spark-dflash2
glm-5.3-flash-dgx-spark-dflash2 PublicA practical recipe for serving GLM-5.3-Flash (320B MoE) on a single 128GB NVIDIA DGX Spark with DFlash2 speculative decoding.
Shell 3
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.