Highlights
- Pro
Popular repositories Loading
-
GLM-5.2-R9-Adaptive-MTP-FULL-CUDA-4x-DGX-Spark
GLM-5.2-R9-Adaptive-MTP-FULL-CUDA-4x-DGX-Spark PublicGLM-5.2 on 4x DGX Spark with adaptive MTP K2/K4/K5, FULL CUDA graphs, DCP2, 520K context, and a downloadable ARM64 runtime image.
-
GLM-5.2-1M-4x-DGX-Spark
GLM-5.2-1M-4x-DGX-Spark PublicUnpruned GLM-5.2 (744B) at 1M context on 4x DGX Spark (GB10/sm_121a) — NVFP4 compact-KV + B12X sparse-MLA + MTP-5. Tested, stable, honest measured numbers.
Python 8
-
vibeclawcoder-local-llm
vibeclawcoder-local-llm PublicForked from laurentenhoor/devclaw
Multi-project dev/qa pipeline orchestration plugin for OpenClaw
TypeScript 3
-
GLM-5.2-Harness-O14-4x-DGX-Spark
GLM-5.2-Harness-O14-4x-DGX-Spark PublicO14 Fast — 250K total KV, READY; O14 Balanced — 500K target, TESTING / DO NOT DEPLOY — for GLM-5.2 on 4× DGX Spark.
-
Keys-GLM-5.2-QuantTrio-655K-MTP-k5-4x-DGX-Spark
Keys-GLM-5.2-QuantTrio-655K-MTP-k5-4x-DGX-Spark PublicValidated GLM-5.2 QuantTrio native MTP k=5 recipe for a 4x DGX Spark cluster
Python 1
If the problem persists, check the GitHub status page or contact support.