Highlights
- Pro
Popular repositories Loading
-
GLM-5.2-R9-Adaptive-MTP-FULL-CUDA-4x-DGX-Spark
GLM-5.2-R9-Adaptive-MTP-FULL-CUDA-4x-DGX-Spark PublicGLM-5.2 on 4x DGX Spark with adaptive MTP K2/K4/K5, FULL CUDA graphs, DCP2, 520K context, and a downloadable ARM64 runtime image.
-
GLM-5.2-1M-4x-DGX-Spark
GLM-5.2-1M-4x-DGX-Spark PublicUnpruned GLM-5.2 (744B) at 1M context on 4x DGX Spark (GB10/sm_121a) — NVFP4 compact-KV + B12X sparse-MLA + MTP-5. Tested, stable, honest measured numbers.
Python 9
-
glm-5.2-dgx-spark-vllm027
glm-5.2-dgx-spark-vllm027 PublicMeasured vLLM 0.27 serving recipe for GLM-5.2 on 4x NVIDIA DGX Spark GB10 (TP4) — DCP1/DCP2/DCP4 options, full speed matrix for 12 booted configs, launcher, patches, and ops runbook
-
vibeclawcoder-local-llm
vibeclawcoder-local-llm PublicForked from laurentenhoor/devclaw
Multi-project dev/qa pipeline orchestration plugin for OpenClaw
TypeScript 3
-
GLM-5.2-Harness-O14-4x-DGX-Spark
GLM-5.2-Harness-O14-4x-DGX-Spark PublicO14 Fast — 250K total KV, READY; O14 Balanced — 500K target, TESTING / DO NOT DEPLOY — for GLM-5.2 on 4× DGX Spark.
If the problem persists, check the GitHub status page or contact support.