Popular repositories Loading
-
glm-5.3-flash-2x-rtx-pro-6000-blackwell
glm-5.3-flash-2x-rtx-pro-6000-blackwell PublicDeploy GLM-5.3 Flash with DFlash2 speculative decoding on dual RTX PRO 6000 Blackwell 96GB GPUs, enabling one-million-token context and 16-image prompts via an OpenAI-compatible API.
Python
-
cellfree-polygamy3184.github.io
cellfree-polygamy3184.github.io PublicRun one-million-token GLM-5.3 Flash K3 with DFlash2 speculative decoding on 2× RTX PRO 6000 Blackwell 96GB GPUs via a safety-hardened, OpenAI-compatible runtime.
HTML
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.