Qwen3.8-Flash-Next as Cogni-Brain on NVIDIA DGX Spark (GB10): HashK GPU PLE + SGLang NEXTN, 36.8 tok/s code, 100/100 tool-eval, 262K context.
-
Updated
Sep 8, 2026 - Python
Qwen3.8-Flash-Next as Cogni-Brain on NVIDIA DGX Spark (GB10): HashK GPU PLE + SGLang NEXTN, 36.8 tok/s code, 100/100 tool-eval, 262K context.
Running Cogni-Brain on DGX Spark · Nemotron-3.5-Lightning-30B-A3B-NVFP4 + DSpark (1M context)
Laguna-S-2.1-NVFP4 Cogni-Brain on DGX Spark: 262K context, ~27 tok/s, 97/100 Tool-Eval, DFlash speculative decoding.
GLM-5.3-Flash (ox-alpha) as Cogni-Brain on single NVIDIA DGX Spark (GB10): llama.cpp SM 121a, 18.7 tok/s decode, 32K context, 100/100 tool-eval
To associate your repository with the cogni-brain topic, visit your repo's landing page and select "manage topics."