🎯
Focusing
Pinned Loading
-
GLM-5.3-Flash-EXL3-DGX-Spark
GLM-5.3-Flash-EXL3-DGX-Spark PublicRun zai-org/GLM-5.3-Flash as an EXL3 4bpw checkpoint on NVIDIA DGX Spark (GB10) with vLLM and cuda-exl3 — measured two-node (TP=2) and three-node (TP=3 + expert parallel) recipes, plus the kernel f…
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.